TakalaWang/Discussion-Phi-4-multimodal-instruct-audio-dimp-reasoning Text Generation • 6B • Updated May 15, 2025 • 96 • 3
All-in-One Multilingual Scene Text Recognition with Script-aware Mixture-of-Experts Paper • 2609.24058 • Published 7 days ago • 55
Complex KDA: Understanding and Enhancing the Expressivity of Kimi Delta Attention Paper • 2609.24797 • Published 7 days ago • 11
nezahatkorkmaz/Turkish-medical-visual-question-answering-LLaVa-dataset Viewer • Updated Mar 19, 2025 • 316 • 144 • 10
JEPA-Anything: Learning Predictive Models across Different Worlds Paper • 2609.20800 • Published 11 days ago • 74
VC-Attention: Value Smoothing and Softmax Casting for Low-bit Attention Paper • 2609.15810 • Published 14 days ago • 51