Collections
Discover the best community collections!
Collections trending this week
-
VibeVoice Technical Report
Paper • 2508.19205 • Published • 180 -
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing
Paper • 2509.22186 • Published • 177 -
SenseNova-U1: Unifying Multimodal Understanding and Generation with NEO-unify Architecture
Paper • 2605.12500 • Published • 198 -
BDH-CQ: In-Context Learning with Recurrent Latent Reasoning
Paper • 2608.09888 • Published • 782
-
vpakarinen/better-human-motion-h3-lora
Text-to-Video • Updated • 325 • • 56 -
vpakarinen/natural-face-speech-h3-lora
Text-to-Video • Updated • 343 • • 26 -
vpakarinen/insta-tiktok-aesthetics-h3-lora
Text-to-Video • Updated • 218 • • 37 -
vpakarinen/asmr-trigger-audio-h3-lora
Text-to-Video • Updated • 55 •
-
Gated Recurrent Transformers: Expressive Depth through Recurrent Modulation
Paper • 2608.15062 • Published • 12 -
Amr-Hegazy/grt-large-isoflop
Text Generation • 0.3B • Updated • 274 -
Amr-Hegazy/grt-large-isoparam
Text Generation • 0.8B • Updated • 260 -
Amr-Hegazy/grt-medium-isoflop
Text Generation • 0.4B • Updated • 265
-
prithivMLmods/ImageShield-Guardrail-Pro
Viewer • Updated • 30.6k • 53 • 1 -
prithivMLmods/ImageShield-Guardrail-Multimodal-20K
Viewer • Updated • 28k • 51 • 2 -
prithivMLmods/ImageShield-Guardrail-80K
Viewer • Updated • 80k • 32 • 2 -
prithivMLmods/ImageShield-Guardrail-Realism-60K
Viewer • Updated • 60k • 40 • 2
-
VibeVoice Technical Report
Paper • 2508.19205 • Published • 180 -
MinerU2.5: A Decoupled Vision-Language Model for Efficient High-Resolution Document Parsing
Paper • 2509.22186 • Published • 177 -
SenseNova-U1: Unifying Multimodal Understanding and Generation with NEO-unify Architecture
Paper • 2605.12500 • Published • 198 -
BDH-CQ: In-Context Learning with Recurrent Latent Reasoning
Paper • 2608.09888 • Published • 782
-
vpakarinen/better-human-motion-h3-lora
Text-to-Video • Updated • 325 • • 56 -
vpakarinen/natural-face-speech-h3-lora
Text-to-Video • Updated • 343 • • 26 -
vpakarinen/insta-tiktok-aesthetics-h3-lora
Text-to-Video • Updated • 218 • • 37 -
vpakarinen/asmr-trigger-audio-h3-lora
Text-to-Video • Updated • 55 •
-
Gated Recurrent Transformers: Expressive Depth through Recurrent Modulation
Paper • 2608.15062 • Published • 12 -
Amr-Hegazy/grt-large-isoflop
Text Generation • 0.3B • Updated • 274 -
Amr-Hegazy/grt-large-isoparam
Text Generation • 0.8B • Updated • 260 -
Amr-Hegazy/grt-medium-isoflop
Text Generation • 0.4B • Updated • 265
-
prithivMLmods/ImageShield-Guardrail-Pro
Viewer • Updated • 30.6k • 53 • 1 -
prithivMLmods/ImageShield-Guardrail-Multimodal-20K
Viewer • Updated • 28k • 51 • 2 -
prithivMLmods/ImageShield-Guardrail-80K
Viewer • Updated • 80k • 32 • 2 -
prithivMLmods/ImageShield-Guardrail-Realism-60K
Viewer • Updated • 60k • 40 • 2