Papers - Fine-tuning - Llava - DPO Collection by matlok Apr 2, 2024 - Direct Preference Optimization of Video Large Multimodal Models from Language Model Reward Paper • 2404.01258 • Published Apr 1, 2024 • 12
Direct Preference Optimization of Video Large Multimodal Models from Language Model Reward Paper • 2404.01258 • Published Apr 1, 2024 • 12
MoE Collection by rafaelpierrehf Apr 2, 2024 - Qwen/Qwen1.5-MoE-A2.7B Text Generation • 14B • Updated Apr 18, 2024 • 717k • 229
3D shape datasets Collection by KostyaZahanych Apr 2, 2024 - allenai/objaverse Updated Mar 31, 2023 • 614k • 457
Audio Collection by kevcon80 Apr 2, 2024 - pyannote/speaker-diarization-3.1 Automatic Speech Recognition • Updated May 10, 2024 • 9.09M • 3.51k
Quiet-STaR Collection by rafaelpierrehf Apr 2, 2024 - Crystalcareai/Quiet-Star-Custom Text Generation • 7B • Updated Apr 12, 2024 • 120 • 11
MMStar An elite vision-indispensable multi-modal benchmark Collection by Lin-Chen May 26, 2024 1 Lin-Chen/MMStar Viewer • Updated Apr 7, 2024 • 1.5k • 15.7k • 53 Are We on the Right Way for Evaluating Large Vision-Language Models? Paper • 2403.20330 • Published Mar 29, 2024 • 6
Are We on the Right Way for Evaluating Large Vision-Language Models? Paper • 2403.20330 • Published Mar 29, 2024 • 6
to_download Collection by klemensnoho Apr 2, 2024 - unity/inference-engine-tiny-stories Text Generation • Updated Jun 8 • 97 • 11
Image upscaler Collection by Goat26 May 14, 2025 - Runtime error Agents 113 Image Upscaler 🔷 113 Running Agents 379 Serverless ImgGen Hub ♨ 379 Highly hackable hub w/ Flux, SD 3.5, LoRAs, no GPUs required
Running Agents 379 Serverless ImgGen Hub ♨ 379 Highly hackable hub w/ Flux, SD 3.5, LoRAs, no GPUs required
llm-memorization Collection by ooo0710 Apr 2, 2024 - Localizing Paragraph Memorization in Language Models Paper • 2403.19851 • Published Mar 28, 2024 • 15
Localizing Paragraph Memorization in Language Models Paper • 2403.19851 • Published Mar 28, 2024 • 15
Papers - Fine-tuning - Llava - DPO Collection by matlok Apr 2, 2024 - Direct Preference Optimization of Video Large Multimodal Models from Language Model Reward Paper • 2404.01258 • Published Apr 1, 2024 • 12
Direct Preference Optimization of Video Large Multimodal Models from Language Model Reward Paper • 2404.01258 • Published Apr 1, 2024 • 12
Quiet-STaR Collection by rafaelpierrehf Apr 2, 2024 - Crystalcareai/Quiet-Star-Custom Text Generation • 7B • Updated Apr 12, 2024 • 120 • 11
MoE Collection by rafaelpierrehf Apr 2, 2024 - Qwen/Qwen1.5-MoE-A2.7B Text Generation • 14B • Updated Apr 18, 2024 • 717k • 229
MMStar An elite vision-indispensable multi-modal benchmark Collection by Lin-Chen May 26, 2024 1 Lin-Chen/MMStar Viewer • Updated Apr 7, 2024 • 1.5k • 15.7k • 53 Are We on the Right Way for Evaluating Large Vision-Language Models? Paper • 2403.20330 • Published Mar 29, 2024 • 6
Are We on the Right Way for Evaluating Large Vision-Language Models? Paper • 2403.20330 • Published Mar 29, 2024 • 6
3D shape datasets Collection by KostyaZahanych Apr 2, 2024 - allenai/objaverse Updated Mar 31, 2023 • 614k • 457
to_download Collection by klemensnoho Apr 2, 2024 - unity/inference-engine-tiny-stories Text Generation • Updated Jun 8 • 97 • 11
Audio Collection by kevcon80 Apr 2, 2024 - pyannote/speaker-diarization-3.1 Automatic Speech Recognition • Updated May 10, 2024 • 9.09M • 3.51k
Image upscaler Collection by Goat26 May 14, 2025 - Runtime error Agents 113 Image Upscaler 🔷 113 Running Agents 379 Serverless ImgGen Hub ♨ 379 Highly hackable hub w/ Flux, SD 3.5, LoRAs, no GPUs required
Running Agents 379 Serverless ImgGen Hub ♨ 379 Highly hackable hub w/ Flux, SD 3.5, LoRAs, no GPUs required
llm-memorization Collection by ooo0710 Apr 2, 2024 - Localizing Paragraph Memorization in Language Models Paper • 2403.19851 • Published Mar 28, 2024 • 15
Localizing Paragraph Memorization in Language Models Paper • 2403.19851 • Published Mar 28, 2024 • 15