AI & ML interests

None defined yet.

Recent Activity

sergiopaniegoΒ  updated a dataset about 14 hours ago
agents-course/certificates
JofthomasΒ  updated a dataset about 16 hours ago
agents-course/unit4-students-scores
burtenshawΒ  updated a dataset about 16 hours ago
agents-course/certificates
View all activity

sergiopaniegoΒ 
posted an update about 10 hours ago
sergiopaniegoΒ 
posted an update 1 day ago
view post
Post
115
Repo2RLEnv just shipped TaskSmith + 50 high quality RL envs generated from HF repos πŸ”¨

TaskSmith is a specialized harness that turns a merged PR into a verified RL env

the envs come from HF repos (Transformers, TRL, PEFT, Accelerate, Diffusers), shipped as Harbor tasks you can eval or train on

> code: github.com/huggingface/Repo2RLEnv
> dataset: huggingface.co/datasets/FineEnvs/HF_ML_Tasksmith
sergiopaniegoΒ 
posted an update 2 days ago
view post
Post
3650
ThinkingBox from @microsoft is now available as an OpenEnv env (cc @tuhink πŸ€— )!

> ThinkingBox is a sandbox for testing agents on business workflows. it simulates a customer, gives the agent MCP tools over a real database, and at the end checks what changed in that database instead of trusting the agent's last message

> ThinkingBox-Bench is the benchmark built on it: 507 tasks across retail, insurance, travel, banking and consulting

> the OpenEnv env runs each task as an episode in its own isolated backend and returns a pass/fail reward from those checks

https://huggingface.co/blog/microsoft/thinkingbox