Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
pocharlies
/
Qwen3.8-Flash-Next
like
0
English
sglang
nvfp4
dgx-spark
gb10
tensor-parallel
speculative-decoding
deployment-recipe
License:
apache-2.0
Model card
Files
Files and versions
xet
Community
Copy to bucket
new
main
Qwen3.8-Flash-Next
63 kB
Ctrl+K
Ctrl+K
1 contributor
History:
7 commits
pocharlies
Shipping config: 8cc / 0.90 / ~1.37M-token KV pool (autotune on)
f1b2255
verified
about 8 hours ago
docs
Translate everything to English; genericize node names and paths
1 day ago
image
v0.2.0: patches baked; measured config table, large-context concurrency, why-not-vLLM, docker-run scripts
about 14 hours ago
k8s
Parsers verified: qwen3 + qwen3_coder
about 21 hours ago
run
Shipping config: 8cc / 0.90 / ~1.37M-token KV pool (autotune on)
about 8 hours ago
.gitattributes
Safe
1.52 kB
initial commit
1 day ago
README.md
14.1 kB
Shipping config: 8cc / 0.90 / ~1.37M-token KV pool (autotune on)
about 8 hours ago