Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
Zhicheng Cai's picture

Zhicheng Cai

Aiolus-X
6

AI & ML interests

LLM&Agentic RL

Recent Activity

authored a paper 1 day ago
Beyond Euclidean Clipping: Overcoming Exploration Collapse in LLM RL via Riemannian Isometric Policy Optimization
authored a paper 1 day ago
FLEX: Continuous Agent Evolution via Forward Learning from Experience
authored a paper 1 day ago
Enigmata: Scaling Logical Reasoning in Large Language Models with Synthetic Verifiable Puzzles
View all activity

Organizations

None yet

upvoted a paper 1 day ago

Beyond Euclidean Clipping: Overcoming Exploration Collapse in LLM RL via Riemannian Isometric Policy Optimization

Paper • 2607.10169 • Published 14 days ago • 12
upvoted a paper 7 days ago

Spectral Rewiring for Exploration, Purification, and Model Merging

Paper • 2607.03065 • Published 22 days ago • 25
upvoted a paper 10 days ago

Weak-to-Strong Generalization via Direct On-Policy Distillation

Paper • 2607.05394 • Published 17 days ago • 137
upvoted 3 papers 6 months ago

Learning to Discover at Test Time

Paper • 2601.16175 • Published Jan 22 • 45

LLM-in-Sandbox Elicits General Agentic Intelligence

Paper • 2601.16206 • Published Jan 22 • 87

FLEX: Continuous Agent Evolution via Forward Learning from Experience

Paper • 2511.06449 • Published Nov 9, 2025 • 14
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs