E-MoE: Enhanced Mixture-of-Experts for Non-Factorized Diffusion Language Models Paper • 2609.37533 • Published 10 days ago • 66
Decentralized Master-Mind: Joint Action Refinement through Iterative Intent Denoising in Multi-Agent Pathfinding Paper • 2609.32019 • Published 14 days ago • 62
LANTERN: Illuminating Hidden Mathematical Knowledge in Language Models Paper • 2609.32264 • Published 13 days ago • 49
Your Transformer Can Hold Two Thoughts at Once: Evidence of Linear Superposition in LLMs Paper • 2609.29845 • Published 15 days ago • 105
Gemma 4 AmbigQA — OFT Collection OFT-adapted Gemma 4 E2B and E4B ensemble members trained on AmbigQA. • 15 items • Updated Aug 7
Gemma 4 AmbigQA — OFT Collection OFT-adapted Gemma 4 E2B and E4B ensemble members trained on AmbigQA. • 15 items • Updated Aug 7
Gemma 4 AmbigQA — OFT Collection OFT-adapted Gemma 4 E2B and E4B ensemble members trained on AmbigQA. • 15 items • Updated Aug 7
Gemma 4 AmbigQA — OFT Collection OFT-adapted Gemma 4 E2B and E4B ensemble members trained on AmbigQA. • 15 items • Updated Aug 7
Gemma 4 AmbigQA — OFT Collection OFT-adapted Gemma 4 E2B and E4B ensemble members trained on AmbigQA. • 15 items • Updated Aug 7