-
LLaDA2.0-Uni: Unifying Multimodal Understanding and Generation with Diffusion Large Language Model
Paper • 2604.20796 • Published • 244 -
inclusionAI/LLaDA2.0-Uni
Any-to-Any • 16B • Updated • 4.58k • 251 -
inclusionAI/LLaDA2.0-Uni-FP8
Any-to-Any • 16B • Updated • 3.12k • 5 -
LLaDA2.0: Scaling Up Diffusion Language Models to 100B
Paper • 2512.15745 • Published • 89
Collections
Discover the best community collections!
Collections trending this week
-
Retreatcost/KansenSakura-Conflagration-RP-12b
Text Generation • 12B • Updated • 75 • • 15 -
Retreatcost/KansenSakura-Erosion-RP-12b
Text Generation • 12B • Updated • 139 • • 37 -
Retreatcost/KansenSakura-Radiance-RP-12b
Text Generation • 12B • Updated • 9 • • 29 -
Retreatcost/KansenSakura-Eclipse-RP-12b
Text Generation • 12B • Updated • 61 • • 39
-
Qwen/Qwen3-Next-80B-A3B-Instruct
Text Generation • 81B • Updated • 286k • • 1.05k -
Qwen/Qwen3-Next-80B-A3B-Thinking
Text Generation • 81B • Updated • 46.3k • • 494 -
Qwen/Qwen3-Next-80B-A3B-Instruct-FP8
Text Generation • 81B • Updated • 160k • 90 -
Qwen/Qwen3-Next-80B-A3B-Thinking-FP8
Text Generation • 81B • Updated • 2.95k • 54
-
LLaDA2.0-Uni: Unifying Multimodal Understanding and Generation with Diffusion Large Language Model
Paper • 2604.20796 • Published • 244 -
inclusionAI/LLaDA2.0-Uni
Any-to-Any • 16B • Updated • 4.58k • 251 -
inclusionAI/LLaDA2.0-Uni-FP8
Any-to-Any • 16B • Updated • 3.12k • 5 -
LLaDA2.0: Scaling Up Diffusion Language Models to 100B
Paper • 2512.15745 • Published • 89
-
Retreatcost/KansenSakura-Conflagration-RP-12b
Text Generation • 12B • Updated • 75 • • 15 -
Retreatcost/KansenSakura-Erosion-RP-12b
Text Generation • 12B • Updated • 139 • • 37 -
Retreatcost/KansenSakura-Radiance-RP-12b
Text Generation • 12B • Updated • 9 • • 29 -
Retreatcost/KansenSakura-Eclipse-RP-12b
Text Generation • 12B • Updated • 61 • • 39
-
Qwen/Qwen3-Next-80B-A3B-Instruct
Text Generation • 81B • Updated • 286k • • 1.05k -
Qwen/Qwen3-Next-80B-A3B-Thinking
Text Generation • 81B • Updated • 46.3k • • 494 -
Qwen/Qwen3-Next-80B-A3B-Instruct-FP8
Text Generation • 81B • Updated • 160k • 90 -
Qwen/Qwen3-Next-80B-A3B-Thinking-FP8
Text Generation • 81B • Updated • 2.95k • 54