OPD-V: Visual On-Policy Self-Distillation with Modality Balance Paper • 2608.05131 • Published 3 days ago • 8
ReflectRL: Learning from Golden Negative Trajectories via Reflective-to-Direct Reasoning Paper • 2608.03972 • Published 4 days ago • 3
ReflectRL: Learning from Golden Negative Trajectories via Reflective-to-Direct Reasoning Paper • 2608.03972 • Published 4 days ago • 3
Dr. DocBench: A Comprehensive Benchmark for Expert-Level and Difficult Document Parsing Paper • 2606.01393 • Published May 31
MetaSkill-Evolve: Recursive Self-Improvement of LLM Agents via Two-Timescale Meta-Skill Evolution Paper • 2607.05297 • Published Jul 6 • 1
Beyond NL2Code: A Structured Survey of Multimodal Code Intelligence Paper • 2606.15932 • Published Jun 16 • 38
AUVIC: Adversarial Unlearning of Visual Concepts for Multi-modal Large Language Models Paper • 2511.11299 • Published Nov 14, 2025
ASCD: Attention-Steerable Contrastive Decoding for Reducing Hallucination in MLLM Paper • 2506.14766 • Published Jun 17, 2025