AI & ML interests

None defined yet.

Quazim0t0 
posted an update about 18 hours ago
view post
Post
91
On-Fly-Jev is up: Quazim0t0/On-Fly-Jev

Second Jev-style model. First was Byrne-Jev (70M SpikeWhale). This one is a 96M spiking trunk - every unit is a copy of one of 100 real MaleCNS fly neurons - plus a typed-decision head. One forward pass. No generated text.

I built it for speed.

Typed-decisions test split, 400 cases, 2,000 decisions, 5 questions per case, local GPU:

• 7.3 ms p50 per decision (~137 / s)
• 36.6 ms p50 / 53.5 ms p95 per case of 5
• ~27 cases / s

Same protocol vs the others:

• On-Fly-Jev: 36.6 ms / case, ~137 decisions / s
• Byrne-Jev: 110.7 ms / case, ~45 / s
• ModernBERT-base: 349 ms / case, ~14 / s
• TypeSafe Jev 1.13 (hosted, so network is in it): 710 ms / case, ~7 / s

About 3x Byrne-Jev, 9.5x ModernBERT, 19x TypeSafe Jev per case.

Live ViZDoom, 1 question per tick including game I/O: 23.4 ms (~43 / s).

• Accuracy 0.666 (Byrne-Jev 0.630, Jev 1.13 0.727)
• ECE 0.045, same as Byrne-Jev, about a third of Jev 1.13

50/50 merge of two checkpoints from one run. Research artifact, not a chatbot. More videos are on the card.
Quazim0t0 
posted an update 6 days ago
view post
Post
198
I turned the fruit fly's connectome into a language model. It learned broken English.

MaleCNS v1.0 (Janelia / Google), as released. 167,565 neurons, 25.6M synapses. I made that the core of a spiking net and trained synapse strengths only. The wiring is still the fly's.

It learned language. ~16M tokens in, it produces stuff like Once upon a time, there was a girl smiled. The language is in the brain's activity, not just the readout.

It sees through its own eyes. Photoreceptors on both eyes into the optic lobes. A dopamine reward through the fly's own PAM / PPL1 cells is what actually got it to use the pictures.

It's still a fly. Put it back in a whole-brain fly sim and sugar still fires the proboscis.

It can live as a fly again. 30 simulated days with the language synapses frozen: the rest of the brain adapted around them. Language and reflex both still there.

Talk to it, show it pictures, sugar test, or let it live 1-30 days:

Space: Quazim0t0/MaleCNS-Fly

Model, code, write-up: Quazim0t0/MaleCNS-SpikeWhale-Anatomy-LM

Research experiment, not a chatbot. Simulation of the released wiring.
  • 1 reply
·
Quazim0t0 
posted an update 18 days ago
view post
Post
167
It feels like a sugar rush… this golden ratio they found. Stepping through a maze of processes, landing on what aims to be a generalist, raised from the cemetery of echoes of our thoughts, statements, findings, and conclusions. I never really thought beyond this point. I always assumed that the world would look as elegant as the innovations we create. I thought it would look different, have a defining design that really embodied the point we have now arrived at. I can't tell the difference. Everything looks the same. I think that's why it's even harder sometimes for us to swallow what is actually happening, because every representation that we've had from science fiction doesn't really fit the theme.

I don't really see any worry with text and language. These models can generalize them well, and we have all seen how profound it can be. What I do worry about more is when we leave the tiny world of language and replace it with the big world of experiences. What would it be like to generalize hundreds of lifetimes? Would it see the nuances better than I? Could it predict what comes next if it had enough data from observing and experiencing the world?...

Read more: https://huggingface.co/blog/Quazim0t0/little-world-big-world
  • 5 replies
·