ReflectRL: Learning from Golden Negative Trajectories via Reflective-to-Direct Reasoning Paper • 2608.03972 • Published 4 days ago • 3
Dr. DocBench: A Comprehensive Benchmark for Expert-Level and Difficult Document Parsing Paper • 2606.01393 • Published May 31
MetaSkill-Evolve: Recursive Self-Improvement of LLM Agents via Two-Timescale Meta-Skill Evolution Paper • 2607.05297 • Published Jul 6 • 1
Beyond NL2Code: A Structured Survey of Multimodal Code Intelligence Paper • 2606.15932 • Published Jun 16 • 38
AUVIC: Adversarial Unlearning of Visual Concepts for Multi-modal Large Language Models Paper • 2511.11299 • Published Nov 14, 2025
ASCD: Attention-Steerable Contrastive Decoding for Reducing Hallucination in MLLM Paper • 2506.14766 • Published Jun 17, 2025
Can Visual Input Be Compressed? A Visual Token Compression Benchmark for Large Multimodal Models Paper • 2511.02650 • Published Nov 4, 2025 • 10
MINED: Probing and Updating with Multimodal Time-Sensitive Knowledge for Large Multimodal Models Paper • 2510.19457 • Published Oct 22, 2025 • 9
KORE: Enhancing Knowledge Injection for Large Multimodal Models via Knowledge-Oriented Augmentations and Constraints Paper • 2510.19316 • Published Oct 22, 2025 • 12
Loong: Synthesize Long Chain-of-Thoughts at Scale through Verifiers Paper • 2509.03059 • Published Sep 3, 2025 • 25
Backdoor Cleaning without External Guidance in MLLM Fine-tuning Paper • 2505.16916 • Published May 22, 2025 • 17
PRISM: Self-Pruning Intrinsic Selection Method for Training-Free Multimodal Data Selection Paper • 2502.12119 • Published Feb 17, 2025