Join the conversation

Join the community of Machine Learners and AI enthusiasts.

Sign Up
Banaxi-TechΒ 
posted an update about 12 hours ago
Post
1240
Well GPT X3 takes the lead. As of right now.




We are now announcing BananaMind 3 🍌! (not ai for those emoji guys)

All models will use BGA (which is almost just NSA) and our BM3X architecture.
Its sizes will be:
BananaMind 3 Flash Lite, 3M parameters at a context of 8K context.
BananaMind 3 Lite, 10M parameters with 16K context.
BananaMind 3 Flash, 25M Parameters with 16K context.
BananaMind 3 Pro, 50M parameters with 24K context.
BananaMind 3 Ultra, 100M parameters with 32K context.
And lastly, BananaMind 3 Max with 150M parameters and 64K CONTEXT.

I can assure you BananaMind 3 Max WILL beat GPT X3 or match it, we won't release it otherwise. We hope for a 40+ INTELLIGENCE INDEX!

BananaMind 3 may also be partnered with dot labs.

We will cancel BananaMind 2.1 and BananaMind 2 Ultra.

As of the BETU SLM Leaderboard we may need to release it after October 11, im very busy right now (even though we said We will release BETU leaderboard before Oct 11 😟)


Hyped for BananaMind3!!!

Bro... I am sorry to break it down to you, but this is mathematically impossible, bcs your context itself eats up all the parameters. you end up with negative parameters (which is absolutely NOT a thing)😭
Your model is your embedding space.

Anyways... good luck with whatever you are doing, man. research is research at the end of the day.

Β·

actually a longer context does not take up more embedding params. though it sometimes can depending on the positional embeddings used; if you are using RoPE more or less context doesnt affect param count. Out of the most popular types only learned absolute embeddings will always create more params whereas relative learned embeddings will sometimes create more. most others dont.

we delaying bananamind 3 😭