AI & ML interests

Models and API early access previews for KoFi donators

Recent Activity

appvoidΒ  updated a Space 1 day ago
CEAMFA/README
appvoidΒ  updated a Space 3 days ago
CEAMFA/palmer
appvoidΒ  published a Space 3 days ago
CEAMFA/palmer
View all activity

appvoidΒ 
posted an update about 13 hours ago
view post
Post
59
byte-level
deep layers
diverse data
compact size
overfitting
grokking

I know the next best small model is somewhere in the intersection of these features.
appvoidΒ 
updated a Space 1 day ago
appvoidΒ 
posted an update 1 day ago
view post
Post
189
Do you prefer AI companies/individuals to release half-baked models weekly or an overpowered model a month?
appvoidΒ 
published a Space 3 days ago
appvoidΒ 
posted an update 3 days ago
view post
Post
1324
Giving free early access to the gguf for some of you today! Tell me what you think.

CEAMFA/palmer-007-preview
appvoidΒ 
posted an update 7 days ago
view post
Post
604
i love reinforcement learning
  • 1 reply
Β·
appvoidΒ 
published a Space 10 days ago
appvoidΒ 
posted an update 10 days ago
view post
Post
1408
A Small Model is All You Need. Meet palmer-006 (90M)

After 3 years of experiments, we are finally releasing our flagship tiny model: **palmer-006**.

If you are building for edge hardware, SBCs (Raspberry Pi, etc.), or low-power devices, this is for you. Inspired by Andrej Karpathy's idea of a self-contained "cognitive core," we wanted to see how much power we could pack into a sub-100M parameter footprint.

🧠 **How we "Palmerized" it:**
We believe in starting our experiments with the absolute strongest baseline possible.
1. Light fine-tuning on highly curated data
2. Model merging
3. Another light fine-tuning round
4. Adjusted Mamba for maximum token speed ⚑️

⚠️ *Note: This is a foundational language model. It has not been instruction-tuned yet!*

Also, since this needs instruction tuning next to become a chat assistantβ€”**what dataset would you recommend we use for the instruct tune?**

---
πŸ”— **Quick Links & Info:**

* **License:** Open for research, education, hobby, and modification! (For commercial use/hosted APIs, shoot an email to nosoyhackercodigo@gmail.com. *PS: Donators can claim a free commercial license!*)

* **Attribution:** Built using AI tech from the Technology Innovation Institute (TII).

Can't wait to see what you build at the edge. Let me know your prompt completions below! πŸ‘‡

appvoid/palmer-006
  • 10 replies
Β·
appvoidΒ 
posted an update 11 days ago
view post
Post
203
...for tomorrow.
appvoidΒ 
posted an update 15 days ago
view post
Post
3335
Two big projects are open sourced soon. Get ready...
  • 11 replies
Β·
appvoidΒ 
posted an update 17 days ago
view post
Post
223
If you make cool smol πŸ€–πŸ€ models (below 0.5b parameters), leave a reply and I will follow you! I'm serious, you don't need to follow me at all just share something through the replies and I (and potentially more people) will follow you (if your models are decent ofc).
  • 11 replies
Β·
appvoidΒ 
posted an update 18 days ago
view post
Post
4658
if you are a tinkerer of small language models and want to stay ahead of what small models can do, follow me!!! seriously, start following people that actually still makes small models

i've made one recently btw

also, i'm keeping an eye on AxiomicLabs leaderboard, looks like the only current alternative to check where the things are going to

though, between us, i think they should add agentic/tool use benchmarks there

anyways,


enjoy!

appvoid/a-cool-model
  • 19 replies
Β·
appvoidΒ 
posted an update 20 days ago
view post
Post
114
A huge amount of large synthetic datasets on huggingface looks surprisingly like templates, that might be one of the main reasons open models might not be as good as other models, we need more people to create smaller, human-curated datasets instead of lazily sending millions of requests to large models for us to fulfill.
  • 4 replies
Β·
appvoidΒ 
posted an update 25 days ago
view post
Post
260
small reasoning models are overrated, these little ones just doom loop a lot by default. good data will always be the moat when training or finetuning small models and latest sota models like fable 5 and gpt 5.6 are increasingly making this a lot easier to do.
  • 3 replies
Β·
appvoidΒ 
posted an update 2 months ago
view post
Post
193
yikes! i missed the small model hackathon i guess i'll have to make sota for people to notice
appvoidΒ 
posted an update 2 months ago
view post
Post
113
model | epoch | step | val_loss |
c-gpt-1| 5 | 41365 | 1.0175 |

If you don't understand what you see, this is just one the coolest proof of concepts I've ever made. I just trained a char-level, super small model (~80 million parameters) on millions of high-quality examples on a curated dataset and is slowly getting davinci-003 vibes
appvoidΒ 
posted an update 2 months ago
view post
Post
184
As an advocate for small language models I just want to say. It might not actually be the end for small models. We are just getting started! Now that we have super good models we can find creative ways to replicate the behavior at small scale!

I'll show you in a few weeks what a small model is capable of, you will surprised.
  • 1 reply
Β·