Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up

Huang's INTelligence lab

university
Activity Feed

AI & ML interests

None defined yet.

Recent Activity

shrango  submitted a paper 1 day ago
On the Off-Policy Teacher in On-Policy Distillation
ChengsongHuang  authored a paper 9 days ago
RRSI: Regularized Recursive Self-Improvement of Agent Harnesses
ChengsongHuang  authored a paper 10 days ago
Small RL Controller, Large Language Model: RL-Guided Adaptive Sampling for Test-Time Scaling
View all activity

Papers

Process Rewards with Learned Reliability

View all Papers

Jiaxin Huang's profile picture Chengsong Huang's profile picture Langlin Huang's profile picture Jixuan Leng's profile picture Jinyuan Li's profile picture

HINT-lab 's models 32

HINT-lab/mistral-7b-hermes-dpo-v0.2

Text Generation • 7B • Updated Oct 10, 2024 • 23

HINT-lab/mistral-7b-hermes-cdpo-v0.2

Text Generation • 7B • Updated Oct 10, 2024 • 20
  • Previous
  • 1
  • 2
  • Next
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs