Where Does Retrieval-Based Open-Ended Evaluation Fail? Automatic Taxonomy Induction from Long-Form Medical Answer Factuality Verification Paper • 2609.30467 • Published 13 days ago • 11
EvolveTrade: Experience-Driven Policy Refinement for Self-Evolving LLM Trading Agents Paper • 2609.17632 • Published 22 days ago • 46
Continual Learning Mechanisms Compose for Long-Horizon Memorization Paper • 2609.06986 • Published 30 days ago • 376
What LLM Trading Agents Actually Do in Production: A Six-Month, Population-Scale Record from Two Fleets Paper • 2609.05663 • Published Sep 4 • 21
unsloth/Qwen3.8-Flash-Next-GGUF Image-Text-to-Text • 177B • Updated about 11 hours ago • 1.36M • 1.13k
Beyond Retrieval: Progressive Latent Memory Evolution for Streaming Video Understanding Paper • 2609.04131 • Published Sep 3 • 31
LLaDA-Image: Building Strong Image Generators with Fully Open Training Recipes Paper • 2609.03796 • Published Sep 3 • 186
SemComp-Bench: Benchmarking Semantic Task Completion in Video Generation Paper • 2608.17426 • Published Aug 18 • 161
Can We Defend Against AI-Generated Video Attacks on Real-World Crisis Events? A Systematic Evaluation of Detectors, Generators and Social Dissemination Paper • 2608.14391 • Published Aug 14 • 287