Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
cfontes
/
qwen-dflash2-spark
like
2
Text Generation
vllm
dflash2
speculative-decoding
nvfp4
qwen3
dgx-spark
gb10
License:
mit
Model card
Files
Files and versions
xet
Community
Copy to bucket
new
main
qwen-dflash2-spark
37 kB
Ctrl+K
Ctrl+K
1 contributor
History:
5 commits
cfontes
Point HF card mirror note at GitHub canonical
2ad2d32
verified
3 days ago
bench
Qwen3.8-27B DFlash2 on 2x DGX Spark: one-stop stack (135 tok/s C1, lm_head BF16 fix)
3 days ago
docs
Add prefill benchmark (~3.4k tok/s peak) + GH canonical link
3 days ago
scripts
Qwen3.8-27B DFlash2 on 2x DGX Spark: one-stop stack (135 tok/s C1, lm_head BF16 fix)
3 days ago
.gitattributes
Safe
1.52 kB
initial commit
3 days ago
.gitignore
Safe
64 Bytes
Qwen3.8-27B DFlash2 on 2x DGX Spark: one-stop stack (135 tok/s C1, lm_head BF16 fix)
3 days ago
LICENSE
Safe
1.07 kB
Qwen3.8-27B DFlash2 on 2x DGX Spark: one-stop stack (135 tok/s C1, lm_head BF16 fix)
3 days ago
README.md
9.68 kB
Point HF card mirror note at GitHub canonical
3 days ago