Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
RL ReSearch
non-profit
rlresearch
Activity Feed
Follow
66
AI & ML interests
None defined yet.
Recent Activity
hamishivi
authored
a paper
3 days ago
Learning to Solve Hard Problems in RL for LLMs by Never Giving Up
hamishivi
authored
a paper
3 days ago
Tmax: A simple recipe for terminal agents
hamishivi
authored
a paper
3 days ago
Meta-Reinforcement Learning with Self-Reflection for Agentic Search
View all activity
Papers
DR Tulu: Reinforcement Learning with Evolving Rubrics for Deep Research
View all Papers
Team members
10
rl-research
's models
7
Sort: Recently updated
rl-research/DR-Tulu-8B-results
Updated
Mar 26
•
2
rl-research/DR-Tulu-8B-Step-1900-results
Updated
Mar 26
rl-research/DR-Tulu-8B-Step-1900
Text Generation
•
8B
•
Updated
Mar 26
•
12
rl-research/DR-Tulu-No-RLER-8B
Text Generation
•
8B
•
Updated
Feb 24
•
16
rl-research/DR-Tulu-8B
Text Generation
•
8B
•
Updated
Feb 24
•
964
•
•
76
rl-research/dr-tulu-shortform-rl-400step
8B
•
Updated
Jan 2
•
5
rl-research/DR-Tulu-SFT-8B
Text Generation
•
8B
•
Updated
Nov 29, 2025
•
40
•
•
5