Realtime-Venus: A full-duplex interaction system with asynchronous delegation Paper • 2609.13814 • Published 14 days ago • 218
Measuring Audio's Impact on Correctness: Audio-Contribution-Aware Post-Training of Large Audio Language Models Paper • 2509.21060 • Published Sep 25, 2025 • 2
OmniVChat: Synthesizing, Benchmarking, and Training for Native Audio-Visual Dialogue Paper • 2609.21465 • Published 8 days ago • 145
Penguin-VL: Exploring the Efficiency Limits of VLM with LLM-based Vision Encoders Paper • 2603.06569 • Published Mar 6 • 120