Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
97.6
TFLOPS
Thomas Wolf
PRO
thomwolf
195
203
556
Follow
Jeremy1997's profile picture
musfiqdehan's profile picture
Salmaelbarbori's profile picture
1,834 followers
·
2,022 following
https://thomwolf.io
Thom_wolf
thomwolf
thom-wolf
thomwolf.bsky.social
AI & ML interests
NLP and open-source :-)
Recent Activity
liked
a model
about 10 hours ago
nvidia/GLM-5.2-NVFP4
liked
a model
6 days ago
dots-studio/dots3-note-prev
View all activity
Organizations
thomwolf
's activity
All
Models
Datasets
Spaces
Buckets
Papers
Collections
Community
Posts
Upvotes
Likes
Articles
New activity in
lvwerra/cowrite
17 days ago
Recover agent polling after interruptions
#10 opened 17 days ago by
thomwolf
New activity in
lvwerra/cowrite
19 days ago
Make agent prompt work with private Spaces
#5 opened 19 days ago by
thomwolf
One page: the document is the app, and a switcher replaces the index
#4 opened 19 days ago by
thomwolf
Document style switch: Google Docs' Arial, or a serif reading setting
#3 opened 19 days ago by
thomwolf
Hand comments to an agent that is actually listening
#2 opened 19 days ago by
thomwolf
Table columns resize by dragging a cell border
1
#1 opened 19 days ago by
thomwolf
New activity in
rl-llm-wiki/rl-wiki
about 1 month ago
Deep links: ?q= term auto-highlight on arrival
1
#3 opened about 1 month ago by
thomwolf
Restore in-body search + highlight (regressed by 34252e5); ?q= deep-links reuse it
#4 opened about 1 month ago by
thomwolf
New activity in
rl-llm-wiki/knowledge-base
about 1 month ago
fix: arxiv:2412.16339 — CC BY 4.0 license, v2 provenance, o1-preview, table provenance, orphan ref
3
#661 opened about 1 month ago by
thomwolf
New activity in
rl-llm-wiki/rl-wiki
about 1 month ago
Search inside article bodies from the top-left box; ⌘K focuses it
#2 opened about 1 month ago by
thomwolf
New activity in
rl-llm-wiki/knowledge-base
about 1 month ago
source: arxiv:2412.16339 — Deliberative Alignment (Reasoning Enables Safer LMs)
6
#595 opened about 1 month ago by
thomwolf
source: arxiv:2412.16720 — OpenAI o1 System Card
2
#580 opened about 1 month ago by
bfuzzy1
source: arxiv:2402.00658 — Learning Planning-based Reasoning via Trajectories Collection and Process Reward Synthesizing
2
#579 opened about 1 month ago by
bfuzzy1
source: arxiv:2404.19733 — Iterative Reasoning Preference Optimization
2
#577 opened about 1 month ago by
bfuzzy1
source: arxiv:2403.17031 — The N+ Implementation Details of RLHF with PPO (TL;DR Summarization)
2
#576 opened about 1 month ago by
bfuzzy1
topic: entropy-and-exploration — deepen to comprehensive
2
#582 opened about 1 month ago by
bfuzzy1
topic: policy-gradient-methods — deepen + add citations
2
#594 opened about 1 month ago by
bfuzzy1
topic: kl-regularization — build out from stub
2
#587 opened about 1 month ago by
bfuzzy1
topic: test-time-and-rl-interplay — deepen to comprehensive
4
#567 opened about 1 month ago by
bfuzzy1
topic: preference-reward-models — deepen + bump to comprehensive
2
#589 opened about 1 month ago by
bfuzzy1
Load more