Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
Twu31
/
TanyaoDojo
like
0
Reinforcement Learning
JAX
mahjong
riichi
behavior-cloning
flax
License:
mit
Model card
Files
Files and versions
xet
Community
Copy to bucket
new
main
TanyaoDojo
89 MB
Ctrl+K
Ctrl+K
1 contributor
History:
3 commits
Twu31
Update headline checkpoint to 100k-game milestone: -4.66 +/- 0.535
6b0bbaf
verified
about 4 hours ago
.gitattributes
Safe
1.52 kB
initial commit
about 12 hours ago
README.md
3.22 kB
Update headline checkpoint to 100k-game milestone: -4.66 +/- 0.535
about 4 hours ago
bc_lean_g402.pkl
pickle
Detected Pickle imports (3)
"numpy.dtype"
,
"numpy._core.multiarray._reconstruct"
,
"numpy.ndarray"
What is a pickle import?
25 MB
xet
TanyaoDojo checkpoints: BC lean/v2 champions + oracle-critic negative result
about 12 hours ago
bc_lean_w192_ep2.pkl
14 MB
xet
TanyaoDojo checkpoints: BC lean/v2 champions + oracle-critic negative result
about 12 hours ago
bc_v2_g186.pkl
25 MB
xet
TanyaoDojo checkpoints: BC lean/v2 champions + oracle-critic negative result
about 12 hours ago
rl_oracle_800m.pkl
pickle
Detected Pickle imports (3)
"numpy.dtype"
,
"numpy._core.multiarray._reconstruct"
,
"numpy.ndarray"
What is a pickle import?
25 MB
xet
TanyaoDojo checkpoints: BC lean/v2 champions + oracle-critic negative result
about 12 hours ago