- Qwen3-Coder-Next-Spark-Agentic
- Parameters
- Honesty / launch class
- Status
- Measure track
- Upstream
- Model tree
- Model parameters (upstream Β· not this pack)
- Download weights (official only)
- Probe quickstart (not a runnable day-0 ship path)
- Serve (measure window)
- Measurement bar
- Status ledger
- Client surface
- Related (Aug 3 family)
- Family matrix
- Primary public path
- Parameters
Qwen3-Coder-Next-Spark-Agentic
Parameters
| Total | ~80B (HF count 79.67B BF16; non-embedding ~79B) |
| Active / token | ~3B |
| Context | upstream card (see Qwen model page) |
| Arch | MoE Β· Qwen3NextForCausalLM Β· 512 experts Β· 10 act Β· 1 shared |
| Default OS path | Official FP8 primary Β· GGUF alts secondary (weights not in this repo) |
| Source | Qwen/Qwen3-Coder-Next |
| This pack | runtime / probe pack only β no weights here |
Size is this table. Hub sidebar Parameters badge stays empty on purpose β this repo has no checkpoint weights (pack / not-a-model). Do not treat badge blank as βunknown size.β Numbers above are upstream / measured-pack facts.
Preview / Unverified DGX Spark deployment pack for Qwen/Qwen3-Coder-Next (~80B total / ~3B active).
This repository is a runtime pack, not a model checkpoint and not a re-quant. Harness format/routing probe 2/2 (receipt only). Not tool-reliability proof. Not ship / gate claim. Verifier β gate clearance.
Honesty / launch class
- Pack:
Qwen3-Coder-Next-Spark-Agentic - Launch class: format/routing smoke + local serve
- Smoke β headline: agent_smoke validates tool-call format/routing on a live local endpoint. It is not a long-horizon reliability score, not a public leaderboard claim, and not gate clearance by itself.
- Verifier β gate clearance: a green smoke receipt is evidence of the smoke suite only.
- diy_gguf: false β public OS path uses official/peer published artifacts only.
- Weights: not in git. Pull scripts write outside the pack tree.
- Loopback default: serve binds
127.0.0.1unlessEXPOSE_LAN=1. - No public promo / campaign CTA before measured launch windows.
- Smoke cases: 2 (see
eval/agent_smoke/). - Default port:
8001(does not steal Laguna:8000unless pack is Laguna-S). - Engine: vLLM OpenAI server + official FP8 (
scripts/serve_vllm_fp8.sh)
Status
| Public role | Preview Β· Probe-measured β no ship claim |
| Aug 3 role | Co-ship with Laguna S (same calendar) |
| Flagship measured | Laguna S remains sole flagship measured pack |
| Primary OS | Official FP8 + vLLM Β· tool parser qwen3_coder |
| GGUF alts | Official Q4_K_M (pulled) Β· optional Q5_K_M Β· llama.cpp |
| diy_gguf | false |
| agent_smoke | Format/routing probe 2/2 (2026-07-30) β not tool metric Β· not gate |
| public_promo_before_launch | false until 2026-08-03 12:00 WIB |
| Ports | Laguna :8000 Β· this pack default :8001 |
| Tool format note | Qwen3-Coder native tool-call format β not Hermes JSON |
| Affiliation | Independent Β· personal Β· not Alibaba Β· not Ainfera product |
Sibling flagship (measured): Laguna-S-2.1-Spark-Agentic
Standard (SAQS): SPARK_AGENTIC_QUANT_STANDARD.md
Launch calendar: LAUNCH_AUG3.md Β· lock: results/launch_lock.json
Measure track
Same closed harness as Laguna S after tools actually execute.
- Pull FP8 β serve vLLM on
:8001β - Point
eval/agent_smokeathttp://127.0.0.1:8001/v1β - Write receipt under
results/β β no ship claim without dated gate - Verifier β gate clearance β
Latest measure window: results/measure_window_fp8_20260730T114320+0700.json
Smoke: results/agent_smoke.json Β· Receipt index: results/RECEIPT_INDEX.json
GGUF alts (llama.cpp) remain secondary headroom/quality paths.
Upstream
| Field | Value |
|---|---|
| Base | Qwen/Qwen3-Coder-Next |
| Official FP8 (primary) | Qwen/Qwen3-Coder-Next-FP8 |
| Official GGUF (alts) | Qwen/Qwen3-Coder-Next-GGUF |
| License | Apache-2.0 (full text in LICENSE) |
| Arch | Qwen3NextForCausalLM / qwen3_next |
Model tree
Qwen/Qwen3-Coder-Next β base (official)
βββ Qwen/Qwen3-Coder-Next-FP8 β primary OS weights
βββ Qwen/Qwen3-Coder-Next-GGUF β official GGUF alts
βββ hizrianraz/Qwen3-Coder-Next-Spark-Agentic β this pack (runtime Β· probe)
Findability: YAML base_model: Qwen/Qwen3-Coder-Next is set so this pack appears under the upstream Model tree on Hub.
Hub may auto-tag the edge as finetune/derived β that is a catalog label only. This pack is not a finetune, adapter, merge, or quantized checkpoint; it hosts no weights.
No library_name: gguf. No base_model_relation: quantized. Probe β gate.
Model parameters (upstream Β· not this pack)
This repo hosts no checkpoint weights, so the HF native Parameters badge stays empty by design. Size = table above. Native Model tree is enabled via base_model for findability only.
Sizes below are copied from the upstream base β not re-counted here. Probe β gate.
| Field | Value | Source |
|---|---|---|
| Architecture | MoE Β· Qwen3NextForCausalLM / qwen3_next |
Qwen config |
| Total parameters | ~80B (HF safetensors count 79.67B BF16; non-embedding ~79B) | Qwen/Qwen3-Coder-Next |
| Activated / token | ~3B | upstream card |
| Experts | 512 total Β· 10 activated Β· 1 shared Β· expert intermediate 512 | upstream card + config |
| Pack role | runtime / probe harness only | this repo |
Download weights (official only)
# on Spark β primary OS
chmod +x scripts/*.sh
./scripts/pull_official_fp8.sh
# expected:
# ~/models/qwen3-coder-next/Qwen3-Coder-Next-FP8/*.safetensors
# GGUF alts (secondary)
./scripts/pull_official_gguf.sh
# expected:
# ~/models/qwen3-coder-next/Qwen3-Coder-Next-Q4_K_M/*.gguf
Q4_K_M shard pins (Spark pull 2026-07-29)
| Shard | SHA256 | Bytes |
|---|---|---|
...-00001-of-00004.gguf |
6bcfc9f9c37901eeb92172e2ab871224dab36a453d263bcb2547f737409534da |
15524827040 |
...-00002-of-00004.gguf |
817def0691ee9d08bf3dc4444be7aed29c9e52091e8fa9d97901ce7e7f6f01d3 |
14872168352 |
...-00003-of-00004.gguf |
23aa634d47dca9b4ca3ea249384e6f01951b24c83cdc076f37f6f43d6c99883f |
14503294496 |
...-00004-of-00004.gguf |
249c768cc5f130dc731567d6edcbdacc48e14dec9e02c5dbe2b2185d2c5bdb2b |
3510702144 |
See docs/WEIGHTS_LOCAL.md + SHA256SUMS. Prefer revision pins over floating main.
Probe quickstart (not a runnable day-0 ship path)
- Leave Laguna on
:8000. This pack defaults to:8001. - Prefer FP8 + vLLM OS path:
./scripts/serve_vllm_fp8.sh(binds host127.0.0.1unless you overrideHOST). - Smoke always writes a receipt under
results/β probe β gate clearance. - Do not set
PORT=8000withoutMEASURE_WINDOW=1. - No ship / hero CTA from this card until an explicit measured gate β tag stays
unmeasured/probe-measured.
# host loopback defaults; override only when you mean it
./scripts/serve_vllm_fp8.sh
OPENAI_BASE_URL=http://127.0.0.1:8001/v1 \
SMOKE_MODEL=local-qwen3-coder-next \
python3 eval/agent_smoke/run_smoke.py
# receipt: results/agent_smoke.json
Serve (measure window)
chmod +x scripts/*.sh
# PRIMARY OS β FP8 + vLLM (default :8001; leave Laguna on :8000)
./scripts/serve_vllm_fp8.sh
# GGUF alt β llama.cpp (also :8001; stop vLLM first)
./scripts/serve_spark.sh
Do not set PORT=8000 without MEASURE_WINDOW=1.
Measurement bar
Same integrity rules as Laguna packs:
- smoke receipts under
results/before any pass claim - verifier β gate clearance
- no βagentic modelβ marketing from unrun packs
OPENAI_BASE_URL=http://127.0.0.1:8001/v1 \
SMOKE_MODEL=local-qwen3-coder-next \
python3 eval/agent_smoke/run_smoke.py
Current measure state: probe_measured_not_gate β FP8 complete + measure-window format/routing smoke 2/2; not a tool-pass headline; ship claim still forbidden until explicit gate.
Status ledger
| Gate | State |
|---|---|
| Scaffold | done |
| Official Q4_K_M pull (Spark) | complete (digests pinned) |
| Official FP8 pull (Spark) | complete 40/40 + SHA (results/fp8_sha256_latest.json) |
| vLLM serve script | wired (scripts/serve_vllm_fp8.sh) |
| measure window | pass 2026-07-30 11:53 WIB Β· Laguna restored |
| agent_smoke | format/routing probe 2/2 Β· not tool reliability Β· not gate |
| Public promo before 08-03 12:00 WIB | forbidden |
| Classification | deployment pack Β· not a model checkpoint |
Client surface
OpenAI-compatible notes + sample client: hermes/
(Qwen native tool format β not Hermes JSON.)
Related (Aug 3 family)
- Flagship measured:
Laguna-S-2.1-Spark-AgenticΒ· 40/40 - Co-list scaffold (unmeasured):
DeepSeek-V4-Flash-REAP25-Spark-AgenticΒ· no ship claim (legacyDeepSeek-V4-Flash-Spark-Agenticβ redirect) - Laguna XS Mac β parked until Metal MoE path is trustworthy + measured
- Qwen3-Coder-30B Mac β parked until measured
Family matrix
See parent FAMILY_MATRIX_AUG3.md + results/triple_aug3_lock.json.
Updated: 2026-07-30 β founder-auth base_model YAML for Hub tree findability Β· not a checkpoint Β· no quantized relation
Primary public path
scripts/pull_official_fp8.shβ official FP8 safetensors (diy_gguf=false).scripts/serve_vllm_fp8.sh(orscripts/serve_spark.shalias) β vLLM on:8001.- Smoke/probe is not a headline claim.
Model tree for hizrianraz/Qwen3-Coder-Next-Spark-Agentic
Base model
Qwen/Qwen3-Coder-Next