avp commited on
Commit
186de58
·
verified ·
1 Parent(s): 427fe2c

Feature the LFM Space, evolved Qwen model, and Ternary Bonsai kernels

Browse files
Files changed (1) hide show
  1. README.md +3 -13
README.md CHANGED
@@ -17,19 +17,9 @@ We study how learned systems represent, adapt, and generalize, with experiments
17
 
18
  ## Explore our work
19
 
20
- ### MONARCH · Language models on your GPU
21
-
22
- How far can careful GPU programming take a small language model in a browser?
23
-
24
- MONARCH explores WebGPU and WGSL kernel optimization for **Liquid AI's LFM2.5-230M**. The public demo runs inference on your device, with editable prompts, Markdown responses, live tokens per second, and a repeatable benchmark. The release includes kernel source, build instructions, model checksums, and documented measurement conditions.
25
-
26
- **[Try the demo](https://huggingface.co/spaces/inductiveML/monarch-webgpu)** · [Read the experiment](https://inductive.ml/experiments/monarch) · [Browse the code](https://huggingface.co/spaces/inductiveML/monarch-webgpu/tree/main) · [Model files](https://huggingface.co/inductiveML/LFM2.5-230M-MONARCH)
27
-
28
- ### Staged Learned Coordinates · Teaching trees to understand geometry
29
-
30
- Can a small neural network learn a coordinate system that makes a decision tree's job easier? This experiment studies learned representations in front of XGBoost to simplify curved and rotated decision boundaries.
31
-
32
- [Read the experiment](https://inductive.ml/staged-learned-coordinates)
33
 
34
  ## How we work
35
 
 
17
 
18
  ## Explore our work
19
 
20
+ - **[LFM2.5 WebGPU Space · MONARCH](https://huggingface.co/spaces/inductiveML/monarch-webgpu)** — Run LFM2.5-230M in your browser with our WGSL kernels, generate text locally, and measure your GPU's tokens per second.
21
+ - **[Qwen3.6-35B-A3B · Evolved Mixed-Bit](https://huggingface.co/inductiveML/Qwen3.6-35B-A3B-evolved-mxbit)** — A 12.63 GB MLX quantization with per-module precision selected by evolutionary search. Runs on Apple Silicon with stock `mlx-lm`.
22
+ - **[Ternary Bonsai kernels · Ternel](https://github.com/inductiveML/ternel)** — Custom Metal kernels that execute Ternary Bonsai 27B directly from losslessly packed 1.75-bit weights on Apple Silicon. [Get the MLX model](https://huggingface.co/inductiveML/Ternary-Bonsai-27B-mlx-lossless-1.75bpw).
 
 
 
 
 
 
 
 
 
 
23
 
24
  ## How we work
25