Qwen3-Coder-Next-Spark-Agentic

Parameters

Total ~80B (HF count 79.67B BF16; non-embedding ~79B)
Active / token ~3B
Context upstream card (see Qwen model page)
Arch MoE Β· Qwen3NextForCausalLM Β· 512 experts Β· 10 act Β· 1 shared
Default OS path Official FP8 primary Β· GGUF alts secondary (weights not in this repo)
Source Qwen/Qwen3-Coder-Next
This pack runtime / probe pack only β€” no weights here

Size is this table. Hub sidebar Parameters badge stays empty on purpose β€” this repo has no checkpoint weights (pack / not-a-model). Do not treat badge blank as β€œunknown size.” Numbers above are upstream / measured-pack facts.

Preview / Unverified DGX Spark deployment pack for Qwen/Qwen3-Coder-Next (~80B total / ~3B active).

This repository is a runtime pack, not a model checkpoint and not a re-quant. Harness format/routing probe 2/2 (receipt only). Not tool-reliability proof. Not ship / gate claim. Verifier β‰  gate clearance.

Honesty / launch class

  • Pack: Qwen3-Coder-Next-Spark-Agentic
  • Launch class: format/routing smoke + local serve
  • Smoke β‰  headline: agent_smoke validates tool-call format/routing on a live local endpoint. It is not a long-horizon reliability score, not a public leaderboard claim, and not gate clearance by itself.
  • Verifier β‰  gate clearance: a green smoke receipt is evidence of the smoke suite only.
  • diy_gguf: false β€” public OS path uses official/peer published artifacts only.
  • Weights: not in git. Pull scripts write outside the pack tree.
  • Loopback default: serve binds 127.0.0.1 unless EXPOSE_LAN=1.
  • No public promo / campaign CTA before measured launch windows.
  • Smoke cases: 2 (see eval/agent_smoke/).
  • Default port: 8001 (does not steal Laguna :8000 unless pack is Laguna-S).
  • Engine: vLLM OpenAI server + official FP8 (scripts/serve_vllm_fp8.sh)

Status

Public role Preview Β· Probe-measured β€” no ship claim
Aug 3 role Co-ship with Laguna S (same calendar)
Flagship measured Laguna S remains sole flagship measured pack
Primary OS Official FP8 + vLLM Β· tool parser qwen3_coder
GGUF alts Official Q4_K_M (pulled) Β· optional Q5_K_M Β· llama.cpp
diy_gguf false
agent_smoke Format/routing probe 2/2 (2026-07-30) β€” not tool metric Β· not gate
public_promo_before_launch false until 2026-08-03 12:00 WIB
Ports Laguna :8000 Β· this pack default :8001
Tool format note Qwen3-Coder native tool-call format β€” not Hermes JSON
Affiliation Independent Β· personal Β· not Alibaba Β· not Ainfera product

Sibling flagship (measured): Laguna-S-2.1-Spark-Agentic

Standard (SAQS): SPARK_AGENTIC_QUANT_STANDARD.md
Launch calendar: LAUNCH_AUG3.md Β· lock: results/launch_lock.json


Measure track

Same closed harness as Laguna S after tools actually execute.

  1. Pull FP8 β†’ serve vLLM on :8001 βœ…
  2. Point eval/agent_smoke at http://127.0.0.1:8001/v1 βœ…
  3. Write receipt under results/ βœ… β€” no ship claim without dated gate
  4. Verifier β‰  gate clearance βœ…

Latest measure window: results/measure_window_fp8_20260730T114320+0700.json Smoke: results/agent_smoke.json Β· Receipt index: results/RECEIPT_INDEX.json

GGUF alts (llama.cpp) remain secondary headroom/quality paths.


Upstream

Field Value
Base Qwen/Qwen3-Coder-Next
Official FP8 (primary) Qwen/Qwen3-Coder-Next-FP8
Official GGUF (alts) Qwen/Qwen3-Coder-Next-GGUF
License Apache-2.0 (full text in LICENSE)
Arch Qwen3NextForCausalLM / qwen3_next

Model tree

Qwen/Qwen3-Coder-Next                 ← base (official)
β”œβ”€β”€ Qwen/Qwen3-Coder-Next-FP8         ← primary OS weights
β”œβ”€β”€ Qwen/Qwen3-Coder-Next-GGUF        ← official GGUF alts
└── hizrianraz/Qwen3-Coder-Next-Spark-Agentic  ← this pack (runtime Β· probe)

Findability: YAML base_model: Qwen/Qwen3-Coder-Next is set so this pack appears under the upstream Model tree on Hub. Hub may auto-tag the edge as finetune/derived β€” that is a catalog label only. This pack is not a finetune, adapter, merge, or quantized checkpoint; it hosts no weights. No library_name: gguf. No base_model_relation: quantized. Probe β‰  gate.


Model parameters (upstream Β· not this pack)

This repo hosts no checkpoint weights, so the HF native Parameters badge stays empty by design. Size = table above. Native Model tree is enabled via base_model for findability only. Sizes below are copied from the upstream base β€” not re-counted here. Probe β‰  gate.

Field Value Source
Architecture MoE Β· Qwen3NextForCausalLM / qwen3_next Qwen config
Total parameters ~80B (HF safetensors count 79.67B BF16; non-embedding ~79B) Qwen/Qwen3-Coder-Next
Activated / token ~3B upstream card
Experts 512 total Β· 10 activated Β· 1 shared Β· expert intermediate 512 upstream card + config
Pack role runtime / probe harness only this repo

Download weights (official only)

# on Spark β€” primary OS
chmod +x scripts/*.sh
./scripts/pull_official_fp8.sh
# expected:
#   ~/models/qwen3-coder-next/Qwen3-Coder-Next-FP8/*.safetensors

# GGUF alts (secondary)
./scripts/pull_official_gguf.sh
# expected:
#   ~/models/qwen3-coder-next/Qwen3-Coder-Next-Q4_K_M/*.gguf

Q4_K_M shard pins (Spark pull 2026-07-29)

Shard SHA256 Bytes
...-00001-of-00004.gguf 6bcfc9f9c37901eeb92172e2ab871224dab36a453d263bcb2547f737409534da 15524827040
...-00002-of-00004.gguf 817def0691ee9d08bf3dc4444be7aed29c9e52091e8fa9d97901ce7e7f6f01d3 14872168352
...-00003-of-00004.gguf 23aa634d47dca9b4ca3ea249384e6f01951b24c83cdc076f37f6f43d6c99883f 14503294496
...-00004-of-00004.gguf 249c768cc5f130dc731567d6edcbdacc48e14dec9e02c5dbe2b2185d2c5bdb2b 3510702144

See docs/WEIGHTS_LOCAL.md + SHA256SUMS. Prefer revision pins over floating main.



Probe quickstart (not a runnable day-0 ship path)

  1. Leave Laguna on :8000. This pack defaults to :8001.
  2. Prefer FP8 + vLLM OS path: ./scripts/serve_vllm_fp8.sh (binds host 127.0.0.1 unless you override HOST).
  3. Smoke always writes a receipt under results/ β€” probe β‰  gate clearance.
  4. Do not set PORT=8000 without MEASURE_WINDOW=1.
  5. No ship / hero CTA from this card until an explicit measured gate β€” tag stays unmeasured / probe-measured.
# host loopback defaults; override only when you mean it
./scripts/serve_vllm_fp8.sh
OPENAI_BASE_URL=http://127.0.0.1:8001/v1 \
SMOKE_MODEL=local-qwen3-coder-next \
python3 eval/agent_smoke/run_smoke.py
# receipt: results/agent_smoke.json

Serve (measure window)

chmod +x scripts/*.sh
# PRIMARY OS β€” FP8 + vLLM (default :8001; leave Laguna on :8000)
./scripts/serve_vllm_fp8.sh

# GGUF alt β€” llama.cpp (also :8001; stop vLLM first)
./scripts/serve_spark.sh

Do not set PORT=8000 without MEASURE_WINDOW=1.


Measurement bar

Same integrity rules as Laguna packs:

  • smoke receipts under results/ before any pass claim
  • verifier β‰  gate clearance
  • no β€œagentic model” marketing from unrun packs
OPENAI_BASE_URL=http://127.0.0.1:8001/v1 \
SMOKE_MODEL=local-qwen3-coder-next \
python3 eval/agent_smoke/run_smoke.py

Current measure state: probe_measured_not_gate β€” FP8 complete + measure-window format/routing smoke 2/2; not a tool-pass headline; ship claim still forbidden until explicit gate.


Status ledger

Gate State
Scaffold done
Official Q4_K_M pull (Spark) complete (digests pinned)
Official FP8 pull (Spark) complete 40/40 + SHA (results/fp8_sha256_latest.json)
vLLM serve script wired (scripts/serve_vllm_fp8.sh)
measure window pass 2026-07-30 11:53 WIB Β· Laguna restored
agent_smoke format/routing probe 2/2 Β· not tool reliability Β· not gate
Public promo before 08-03 12:00 WIB forbidden
Classification deployment pack Β· not a model checkpoint

Client surface

OpenAI-compatible notes + sample client: hermes/ (Qwen native tool format β€” not Hermes JSON.)

Related (Aug 3 family)

  • Flagship measured: Laguna-S-2.1-Spark-Agentic Β· 40/40
  • Co-list scaffold (unmeasured): DeepSeek-V4-Flash-REAP25-Spark-Agentic Β· no ship claim (legacy DeepSeek-V4-Flash-Spark-Agentic β†’ redirect)
  • Laguna XS Mac β€” parked until Metal MoE path is trustworthy + measured
  • Qwen3-Coder-30B Mac β€” parked until measured

Family matrix

See parent FAMILY_MATRIX_AUG3.md + results/triple_aug3_lock.json.

Updated: 2026-07-30 β€” founder-auth base_model YAML for Hub tree findability Β· not a checkpoint Β· no quantized relation

Primary public path

  1. scripts/pull_official_fp8.sh β€” official FP8 safetensors (diy_gguf=false).
  2. scripts/serve_vllm_fp8.sh (or scripts/serve_spark.sh alias) β€” vLLM on :8001.
  3. Smoke/probe is not a headline claim.
Downloads last month

-

Downloads are not tracked for this model. How to track
Inference Providers NEW
This model isn't deployed by any Inference Provider. πŸ™‹ Ask for provider support

Model tree for hizrianraz/Qwen3-Coder-Next-Spark-Agentic

Finetuned
(33)
this model

Collection including hizrianraz/Qwen3-Coder-Next-Spark-Agentic