Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
fasa-org
/
MiniCPM-4-8B-DashAttention
like
0
Follow
FASA
5
PyTorch
llama
arxiv:
2605.18753
License:
apache-2.0
Model card
Files
Files and versions
xet
Community
2
Copy to bucket
new
main
MiniCPM-4-8B-DashAttention
16.4 GB
Ctrl+K
Ctrl+K
2 contributors
History:
18 commits
hyx21
Update README.md
d1d53b1
verified
2 months ago
.gitattributes
Safe
1.52 kB
initial commit
3 months ago
README.md
Safe
6.96 kB
Update README.md
2 months ago
added_tokens.json
Safe
216 Bytes
Upload 7 files
3 months ago
check_act.py
Safe
1.08 kB
Add check_act.py (ModelScope adansa source, verbatim)
3 months ago
config.json
Safe
3.42 kB
Update config.json
3 months ago
config_benchmark.json
Safe
3.42 kB
Rename config.json to config_benchmark.json
3 months ago
configuration_minicpm.py
Safe
9.66 kB
Add configuration_minicpm.py (ModelScope adansa source, verbatim)
3 months ago
generation_config.json
Safe
182 Bytes
Upload 7 files
3 months ago
modeling_llama_long_infllmv2.py
Safe
102 kB
Add modeling_llama_long_infllmv2.py (ModelScope adansa source, verbatim)
3 months ago
modeling_llama_long_infllmv2_64.py
Safe
102 kB
Add modeling_llama_long_infllmv2_64.py (ModelScope adansa source, verbatim)
3 months ago
modeling_minicpm.py
Safe
103 kB
Add modeling_minicpm.py (ModelScope adansa source, verbatim)
3 months ago
pytorch_model.bin
pickle
Detected Pickle imports (3)
"torch.BFloat16Storage"
,
"torch._utils._rebuild_tensor_v2"
,
"collections.OrderedDict"
What is a pickle import?
16.4 GB
xet
Add pytorch_model.bin weights (ModelScope adansa source)
3 months ago
special_tokens_map.json
Safe
630 Bytes
Upload 7 files
3 months ago
stage1.py
Safe
10.9 kB
Add stage1.py (ModelScope adansa source, verbatim)
3 months ago
tokenizer.json
Safe
6.7 MB
Upload 7 files
3 months ago
tokenizer.model
Safe
1.18 MB
xet
Upload 7 files
3 months ago
tokenizer_config.json
Safe
2.94 kB
Upload 7 files
3 months ago
topk_sparse_attention_decode.py
Safe
12.6 kB
Add topk_sparse_attention_decode.py (ModelScope adansa source, verbatim)
3 months ago
topk_sparse_attn.py
Safe
46.5 kB
Add topk_sparse_attn.py (ModelScope adansa source, verbatim)
3 months ago
transform_score.py
Safe
4.55 kB
Add transform_score.py (ModelScope adansa source, verbatim)
3 months ago