-
Crownelius/Opus-4.6-Reasoning-3300x
Viewer • Updated • 2.16k • 827 • 320 -
Roman1111111/claude-opus-4.6-10000x
Viewer • Updated • 9.63k • 546 • 393 -
Roman1111111/gpt-5.4-step-by-step-reasoning
Viewer • Updated • 1.5k • 153 • 64 -
Crownelius/Opus-4.6-Reasoning-2100x-formatted
Viewer • Updated • 2.16k • 269 • 59
Nikita Shulgan
NikitaShu
AI & ML interests
None yet
Organizations
None yet
Memory
Additional for LLMs
-
Part I: Tricks or Traps? A Deep Dive into RL for LLM Reasoning
Paper • 2508.08221 • Published • 50 -
Don't Overthink It: A Survey of Efficient R1-style Large Reasoning Models
Paper • 2508.02120 • Published • 20 -
Thinking with Images for Multimodal Reasoning: Foundations, Methods, and Future Frontiers
Paper • 2506.23918 • Published • 90 -
The Landscape of Agentic Reinforcement Learning for LLMs: A Survey
Paper • 2509.02547 • Published • 239
Web agents
papers
OCR
Prior?
Agents
-
Efficient Agents: Building Effective Agents While Reducing Cost
Paper • 2508.02694 • Published • 86 -
A Comprehensive Survey of Self-Evolving AI Agents: A New Paradigm Bridging Foundation Models and Lifelong Agentic Systems
Paper • 2508.07407 • Published • 100 -
Seeing, Listening, Remembering, and Reasoning: A Multimodal Agent with Long-Term Memory
Paper • 2508.09736 • Published • 58 -
Memp: Exploring Agent Procedural Memory
Paper • 2508.06433 • Published • 36
LLMs
-
GLM-4.5: Agentic, Reasoning, and Coding (ARC) Foundation Models
Paper • 2508.06471 • Published • 213 -
GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning
Paper • 2507.01006 • Published • 257 -
Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities
Paper • 2507.06261 • Published • 68 -
SmallThinker: A Family of Efficient Large Language Models Natively Trained for Local Deployment
Paper • 2507.20984 • Published • 59
Reasoning Datasets from close llms
-
Crownelius/Opus-4.6-Reasoning-3300x
Viewer • Updated • 2.16k • 827 • 320 -
Roman1111111/claude-opus-4.6-10000x
Viewer • Updated • 9.63k • 546 • 393 -
Roman1111111/gpt-5.4-step-by-step-reasoning
Viewer • Updated • 1.5k • 153 • 64 -
Crownelius/Opus-4.6-Reasoning-2100x-formatted
Viewer • Updated • 2.16k • 269 • 59
OCR
Memory
Prior?
Additional for LLMs
-
Part I: Tricks or Traps? A Deep Dive into RL for LLM Reasoning
Paper • 2508.08221 • Published • 50 -
Don't Overthink It: A Survey of Efficient R1-style Large Reasoning Models
Paper • 2508.02120 • Published • 20 -
Thinking with Images for Multimodal Reasoning: Foundations, Methods, and Future Frontiers
Paper • 2506.23918 • Published • 90 -
The Landscape of Agentic Reinforcement Learning for LLMs: A Survey
Paper • 2509.02547 • Published • 239
Agents
-
Efficient Agents: Building Effective Agents While Reducing Cost
Paper • 2508.02694 • Published • 86 -
A Comprehensive Survey of Self-Evolving AI Agents: A New Paradigm Bridging Foundation Models and Lifelong Agentic Systems
Paper • 2508.07407 • Published • 100 -
Seeing, Listening, Remembering, and Reasoning: A Multimodal Agent with Long-Term Memory
Paper • 2508.09736 • Published • 58 -
Memp: Exploring Agent Procedural Memory
Paper • 2508.06433 • Published • 36
Web agents
LLMs
-
GLM-4.5: Agentic, Reasoning, and Coding (ARC) Foundation Models
Paper • 2508.06471 • Published • 213 -
GLM-4.1V-Thinking: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning
Paper • 2507.01006 • Published • 257 -
Gemini 2.5: Pushing the Frontier with Advanced Reasoning, Multimodality, Long Context, and Next Generation Agentic Capabilities
Paper • 2507.06261 • Published • 68 -
SmallThinker: A Family of Efficient Large Language Models Natively Trained for Local Deployment
Paper • 2507.20984 • Published • 59
papers