Bananamind3 is getting delayed lmao
🏗️ Building on HF
AxionLab
AxionLab-official
P(doom) 10%
AI & ML interests
Owner of SupraLabs
Recent Activity
new activity about 3 hours ago
BananaMind/README:Update README.md repliedto Banaxi-Tech's post about 3 hours ago
Well GPT X3 takes the lead. As of right now.
We are now announcing BananaMind 3 🍌! (not ai for those emoji guys)
All models will use BGA (which is almost just NSA) and our BM3X architecture.
Its sizes will be:
BananaMind 3 Flash Lite, 3M parameters at a context of 8K context.
BananaMind 3 Lite, 10M parameters with 16K context.
BananaMind 3 Flash, 25M Parameters with 16K context.
BananaMind 3 Pro, 50M parameters with 24K context.
BananaMind 3 Ultra, 100M parameters with 32K context.
And lastly, BananaMind 3 Max with 150M parameters and 64K CONTEXT.
I can assure you BananaMind 3 Max WILL beat GPT X3 or match it, we won't release it otherwise. We hope for a 40+ INTELLIGENCE INDEX!
BananaMind 3 may also be partnered with dot labs.
We will cancel BananaMind 2.1 and BananaMind 2 Ultra.
As of the BETU SLM Leaderboard we may need to release it after October 11, im very busy right now (even though we said We will release BETU leaderboard before Oct 11 😟)
reacted to Banaxi-Tech's post with 😔 about 13 hours ago
We have released BGA!
And wow, It provides 256x (and 512x at the end of 1M) yes 256x LESS attention compute at 1M context window.
That means you can train a 1M context window at the compute of a ~4K context window.
Check IT OUT: https://huggingface.co/spaces/BananaMind/blog#bananamind-gate-attention
The Accuracy Should BE WAy better than DSA but untested yet.
And, now some updates on BananaMind 3:
BananaMind 3 Will start training Soon!
Sizes: 10M, 25M, 50M, 100M, 150M
And the context windows ARE INSANE: 10M, 16K context, 25M 16k context, 50M 32K context, 100M and 150M, 64K context!!!!