DarwinX: Evolving Agent Harnesses Through Natural Selection Paper • 2608.07545 • Published 25 days ago • 113
AI4AI at Test-Time: Strong-to-Weak Capability Transfer via Harnesses Paper • 2608.12307 • Published 13 days ago • 113
Future Optical Flow Prediction Improves Robot Control & Video Generation Paper • 2601.10781 • Published Jan 15 • 19
Active Video Perception: Iterative Evidence Seeking for Agentic Long Video Understanding Paper • 2512.05774 • Published Dec 5, 2025 • 7
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models Paper • 2507.12806 • Published Jul 17, 2025 • 21
BLIP3-o: A Family of Fully Open Unified Multimodal Models-Architecture, Training and Dataset Paper • 2505.09568 • Published May 14, 2025 • 100
Salesforce/xgen-mm-vid-phi3-mini-r-v1.5-32tokens-8frames Image-Text-to-Text • 4B • Updated Feb 3, 2025 • 98 • 4
Salesforce/xgen-mm-vid-phi3-mini-r-v1.5-128tokens-8frames Image-Text-to-Text • 4B • Updated Feb 3, 2025 • 182 • 11
xGen-MM-Vid (BLIP-3-Video): You Only Need 32 Tokens to Represent a Video Even in VLMs Paper • 2410.16267 • Published Oct 21, 2024 • 18
xGen-VideoSyn-1: High-fidelity Text-to-Video Synthesis with Compressed Representations Paper • 2408.12590 • Published Aug 22, 2024 • 35
Salesforce/xgen-mm-phi3-mini-instruct-singleimg-r-v1.5 Image-Text-to-Text • 4B • Updated Feb 3, 2025 • 69 • 15
Salesforce/xgen-mm-phi3-mini-instruct-dpo-r-v1.5 Image-Text-to-Text • 4B • Updated Feb 3, 2025 • 61 • 19
Salesforce/xgen-mm-phi3-mini-instruct-interleave-r-v1.5 Image-Text-to-Text • 4B • Updated Feb 3, 2025 • 362 • 59