Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
SeanWang0027
/
ftb-sciworld-repro
like
0
Reinforcement Learning
English
on-policy-distillation
llm-agents
scienceworld
License:
apache-2.0
Model card
Files
Files and versions
xet
Community
Copy to bucket
new
main
ftb-sciworld-repro
/
tcod
/
trinity
/
common
799 kB
Ctrl+K
Ctrl+K
1 contributor
History:
1 commit
SeanWang0027
Upload folder using huggingface_hub
8c9ba62
verified
13 days ago
models
Upload folder using huggingface_hub
13 days ago
rewards
Upload folder using huggingface_hub
13 days ago
workflows
Upload folder using huggingface_hub
13 days ago
__init__.py
Safe
0 Bytes
Upload folder using huggingface_hub
13 days ago
config.py
33.3 kB
Upload folder using huggingface_hub
13 days ago
config_validator.py
49.9 kB
Upload folder using huggingface_hub
13 days ago
constants.py
2.85 kB
Upload folder using huggingface_hub
13 days ago
experience.py
24.3 kB
Upload folder using huggingface_hub
13 days ago
verl_config.py
24 kB
Upload folder using huggingface_hub
13 days ago