Text-to-Speech facebook/seamless-expressive Text-to-Speech • Updated Jan 4, 2024 • 192 metavoiceio/metavoice-1B-v0.1 Text-to-Speech • Updated Apr 3, 2024 • 241 • 789
Image Processing briaai/RMBG-1.4 Image Segmentation • 44.1M • Updated Jul 6, 2025 • 405k • 2.02k Running on Zero Agents 482 LocateAnything 💬 482 Detect and label objects in images and videos
Papers - LLMs DocLLM: A layout-aware generative language model for multimodal document understanding Paper • 2401.00908 • Published Dec 31, 2023 • 192
DocLLM: A layout-aware generative language model for multimodal document understanding Paper • 2401.00908 • Published Dec 31, 2023 • 192
LLMs andersonbcdefg/synthetic_retrieval_tasks Viewer • Updated Feb 3, 2024 • 205k • 128 • 79 microsoft/phi-2 Text Generation • 3B • Updated Dec 8, 2025 • 1.35M • 3.5k TinyLlama/TinyLlama-1.1B-Chat-v1.0 Text Generation • 1B • Updated Mar 17, 2024 • 1.63M • • 1.78k cloudyu/Mixtral_34Bx2_MoE_60B Text Generation • 61B • Updated Jan 6 • 8.52k • 114
Papers - LLMs DocLLM: A layout-aware generative language model for multimodal document understanding Paper • 2401.00908 • Published Dec 31, 2023 • 192
DocLLM: A layout-aware generative language model for multimodal document understanding Paper • 2401.00908 • Published Dec 31, 2023 • 192
LLMs andersonbcdefg/synthetic_retrieval_tasks Viewer • Updated Feb 3, 2024 • 205k • 128 • 79 microsoft/phi-2 Text Generation • 3B • Updated Dec 8, 2025 • 1.35M • 3.5k TinyLlama/TinyLlama-1.1B-Chat-v1.0 Text Generation • 1B • Updated Mar 17, 2024 • 1.63M • • 1.78k cloudyu/Mixtral_34Bx2_MoE_60B Text Generation • 61B • Updated Jan 6 • 8.52k • 114
Text-to-Speech facebook/seamless-expressive Text-to-Speech • Updated Jan 4, 2024 • 192 metavoiceio/metavoice-1B-v0.1 Text-to-Speech • Updated Apr 3, 2024 • 241 • 789
Image Processing briaai/RMBG-1.4 Image Segmentation • 44.1M • Updated Jul 6, 2025 • 405k • 2.02k Running on Zero Agents 482 LocateAnything 💬 482 Detect and label objects in images and videos