VākQA: A Benchmark and Evaluation Study for Telugu Spoken Factoid Question Answering Paper • 2609.19879 • Published 6 days ago • 26
Continual Learning Mechanisms Compose for Long-Horizon Memorization Paper • 2609.06986 • Published 16 days ago • 372
FreeFlow: A Bias-free Hierarchical Transformer for Optical Flow Estimation Paper • 2609.11486 • Published 13 days ago • 34
TempCloze: Can Video-LLMs Identify the Missing Middle? Paper • 2609.01515 • Published 22 days ago • 31
Multi-Grid Post-Training for Long-Form Multi-Shot Video Generation Paper • 2609.06373 • Published 17 days ago • 17
Encoded Early, Used Late: Where Transformers Begin to Act on an Inferred Partner's Expertise Paper • 2609.07139 • Published 16 days ago • 17
LatentPress: Context Compression Beyond Text and Vision Paper • 2609.01507 • Published 22 days ago • 119
SemComp-Bench: Benchmarking Semantic Task Completion in Video Generation Paper • 2608.17426 • Published Aug 18 • 160
Can We Defend Against AI-Generated Video Attacks on Real-World Crisis Events? A Systematic Evaluation of Detectors, Generators and Social Dissemination Paper • 2608.14391 • Published Aug 14 • 285
LLMRouter: Unified Infrastructure for Developing, Evaluating, and Deploying LLM Routers Paper • 2608.06867 • Published Aug 7 • 114
OpenART: Scaling Agent Red Teaming via Open-Ended Environment Evolution Paper • 2608.00677 • Published Aug 1 • 265
The Next Screenshot Knows: Gated Hindsight Distillation for Mobile GUI Agents Paper • 2608.06065 • Published Aug 6 • 8
OSReward: Instituting Standardized Evaluation for Cross-Platform Computer-Use Reward Models Paper • 2607.28609 • Published Jul 30 • 75
GST-Bench: Can VLMs Develop Global Spatial Awareness from Video? Paper • 2608.05747 • Published Aug 6 • 47