Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
wang's picture

wang PRO

xinpeng
1 2 9
kargaranamir's profile picture lunula's profile picture iamasQ's profile picture
·

AI & ML interests

None yet

Recent Activity

upvoted a paper about 6 hours ago
OPD Before RL: Warm-Starting Rubric-Based RL with On-Policy Distillation
submitted a paper about 6 hours ago
OPD Before RL: Warm-Starting Rubric-Based RL with On-Policy Distillation
updated a dataset 10 months ago
xinpeng/big-math-hard_tiny_instruct_cheat_rm_loophole_v2_mixed_0.5
View all activity

Organizations

CIS, LMU Munich's profile picture MaiNLP's profile picture safety-by-imitation's profile picture RewardHacking's profile picture

upvoted a paper about 6 hours ago

OPD Before RL: Warm-Starting Rubric-Based RL with On-Policy Distillation

Paper • 2610.02781 • Published 6 days ago • 4
upvoted a paper over 1 year ago

Refusal Direction is Universal Across Safety-Aligned Languages

Paper • 2505.17306 • Published May 22, 2025 • 2
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs