MCPMark: A Benchmark for Stress-Testing Realistic and Comprehensive MCP Use Paper • 2509.24002 • Published Sep 28, 2025 • 180
Harnessing Uncertainty: Entropy-Modulated Policy Gradients for Long-Horizon LLM Agents Paper • 2509.09265 • Published Sep 11, 2025 • 47
WideSearch: Benchmarking Agentic Broad Info-Seeking Paper • 2508.07999 • Published Aug 11, 2025 • 113
Towards Multi-View Consistent Style Transfer with One-Step Diffusion via Vision Conditioning Paper • 2411.10130 • Published Nov 15, 2024
NTIRE 2025 Challenge on UGC Video Enhancement: Methods and Results Paper • 2505.03007 • Published May 5, 2025
The Tenth NTIRE 2025 Efficient Super-Resolution Challenge Report Paper • 2504.10686 • Published Apr 14, 2025
DoctorAgent-RL: A Multi-Agent Collaborative Reinforcement Learning System for Multi-Turn Clinical Dialogue Paper • 2505.19630 • Published May 26, 2025 • 7
Safeguarding Vision-Language Models: Mitigating Vulnerabilities to Gaussian Noise in Perturbation-based Attacks Paper • 2504.01308 • Published Apr 2, 2025 • 14
Dynamic Relation Transformer for Contextual Text Block Detection Paper • 2401.09232 • Published Jan 17, 2024
Safeguarding Vision-Language Models: Mitigating Vulnerabilities to Gaussian Noise in Perturbation-based Attacks Paper • 2504.01308 • Published Apr 2, 2025 • 14
DeepSeek-VL2: Mixture-of-Experts Vision-Language Models for Advanced Multimodal Understanding Paper • 2412.10302 • Published Dec 13, 2024 • 24