Can AI agents conduct open-ended AI research? Early evidence from two case studies Paper • 2607.27191 • Published 2 days ago • 12
SkillRise: Agentic Reinforcement Learning for Cross-Task Skill Evolution Paper • 2607.26784 • Published 2 days ago • 19
CoRT: Counterfactual Replay for Token-Level Rubric-Guided Policy Optimization Paper • 2607.25659 • Published 3 days ago • 78
FormalML: A Benchmark for Evaluating Formal Subgoal Completion in Machine Learning Theory Paper • 2510.02335 • Published Sep 26, 2025 • 2
Running on Zero Agents Featured 452 DeepSeek OCR Demo 🆘 452 An interactive demo for the DeepSeek-OCR model.
Video-As-Prompt: Unified Semantic Control for Video Generation Paper • 2510.20888 • Published Oct 23, 2025 • 50
Document Understanding, Measurement, and Manipulation Using Category Theory Paper • 2510.21553 • Published Oct 24, 2025 • 5
Reasoning with Sampling: Your Base Model is Smarter Than You Think Paper • 2510.14901 • Published Oct 16, 2025 • 49
A Theoretical Study on Bridging Internal Probability and Self-Consistency for LLM Reasoning Paper • 2510.15444 • Published Oct 17, 2025 • 151