FluidAudio Web β raw model weights
Model weights for fluidaudio-web: fully in-browser inference (WebGPU + WebAssembly) for FluidAudio's speech models.
These are raw weights only β no ONNX, no onnxruntime. Each engine is a hand-written forward pass (WebGPU/WASM/JS) that loads these tensors directly and, where applicable, dequantizes them in-shader. Weights are extracted from the source models and parity-verified against the originals before publishing.
Layout (one folder per engine)
| folder | model | status |
|---|---|---|
vad/ |
Silero VAD v5 (16 kHz) | β available (fp32) |
parakeet/ |
Parakeet TDT 0.6B v3 (encoder + decoder/joint) | π§ in progress (int4/palettized) |
kokoro/ |
Kokoro TTS 82M | π§ planned |
nemotron/ Β· eou/ Β· sortformer/ Β· whisper/ |
streaming ASR / EOU / diarization / multilingual ASR | π§ planned |
Format
*.binβ concatenated little-endian tensors.manifest.jsonβname -> { dims, offset, len }(offset/len in elements, not bytes).
Load: fetch the .bin, slice each tensor per the manifest. Extraction scripts and
the forward passes live in the fluidaudio-web repo (scripts/, src/engines/).
Quantized engines additionally carry per-tensor scales / palettes (documented per folder) for in-shader dequant.
Inference Providers NEW
This model isn't deployed by any Inference Provider. π Ask for provider support