Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
Chenxiao Zhao's picture

Chenxiao Zhao

ChenShawn
4 5 19
SiweiWu's profile picture dark-pen's profile picture LeNgocTu's profile picture
·
  • ChenShawn

AI & ML interests

Reinforcement learning

Recent Activity

authored a paper about 7 hours ago
From Generic Correlation to Input-Specific Credit in On-Policy Self Distillation
authored a paper about 7 hours ago
Anti-Self-Distillation for Reasoning RL via Pointwise Mutual Information
upvoted a paper about 8 hours ago
VibeLifeBench: Can Your Life Agent Be Proactive and Persistent in a Living World?
View all activity

Organizations

None yet

upvoted a paper about 8 hours ago

VibeLifeBench: Can Your Life Agent Be Proactive and Persistent in a Living World?

Paper • 2608.10875 • Published 1 day ago • 11
upvoted an article 5 months ago
view article
Article

Forge: Scalable Agent RL Framework and Algorithm

MiniMax-AI
•
Feb 13
• 157
upvoted 3 papers 6 months ago

REDSearcher: A Scalable and Cost-Efficient Framework for Long-Horizon Search Agents

Paper • 2602.14234 • Published Feb 15 • 28

VESPO: Variational Sequence-Level Soft Policy Optimization for Stable Off-Policy LLM Training

Paper • 2602.10693 • Published Feb 11 • 222

DeepEyes: Incentivizing "Thinking with Images" via Reinforcement Learning

Paper • 2505.14362 • Published May 20, 2025 • 6
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs