Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
SeanWang0027
/
ftb-sciworld-repro
like
0
Reinforcement Learning
English
on-policy-distillation
llm-agents
scienceworld
License:
apache-2.0
Model card
Files
Files and versions
xet
Community
Copy to bucket
new
main
ftb-sciworld-repro
/
tcod
/
TCOD_examples
/
alfworld
/
alfworld_data
1.97 MB
Ctrl+K
Ctrl+K
1 contributor
History:
1 commit
SeanWang0027
Upload folder using huggingface_hub
8c9ba62
verified
13 days ago
README.md
Safe
2.86 kB
Upload folder using huggingface_hub
13 days ago
fix_game_paths.py
Safe
4.91 kB
Upload folder using huggingface_hub
13 days ago
test.jsonl
Safe
23.1 kB
Upload folder using huggingface_hub
13 days ago
test_unseen.jsonl
Safe
22.3 kB
Upload folder using huggingface_hub
13 days ago
train.jsonl
Safe
568 kB
Upload folder using huggingface_hub
13 days ago
train_expert.jsonl
Safe
1.33 MB
Upload folder using huggingface_hub
13 days ago
train_hard.jsonl
Safe
21.1 kB
Upload folder using huggingface_hub
13 days ago