AI & ML interests

None defined yet.

Recent Activity

burtenshawย  updated a dataset about 8 hours ago
agents-course/certificates
sergiopaniegoย  updated a dataset about 8 hours ago
agents-course/final-certificates
sergiopaniegoย  updated a dataset about 8 hours ago
agents-course/course-certificates-of-excellence
View all activity

sergiopaniegoย 
posted an update 2 days ago
sergiopaniegoย 
posted an update 3 days ago
view post
Post
132
Repo2RLEnv just shipped TaskSmith + 50 high quality RL envs generated from HF repos ๐Ÿ”จ

TaskSmith is a specialized harness that turns a merged PR into a verified RL env

the envs come from HF repos (Transformers, TRL, PEFT, Accelerate, Diffusers), shipped as Harbor tasks you can eval or train on

> code: github.com/huggingface/Repo2RLEnv
> dataset: huggingface.co/datasets/FineEnvs/HF_ML_Tasksmith
sergiopaniegoย 
posted an update 4 days ago
view post
Post
3681
ThinkingBox from @microsoft is now available as an OpenEnv env (cc @tuhink ๐Ÿค— )!

> ThinkingBox is a sandbox for testing agents on business workflows. it simulates a customer, gives the agent MCP tools over a real database, and at the end checks what changed in that database instead of trusting the agent's last message

> ThinkingBox-Bench is the benchmark built on it: 507 tasks across retail, insurance, travel, banking and consulting

> the OpenEnv env runs each task as an episode in its own isolated backend and returns a pass/fail reward from those checks

https://huggingface.co/blog/microsoft/thinkingbox