We're exited to announce BananaMind OS, our OS specically for running BananaMind models! Its able to run BananaMind 2 Nano at 4 bit on only 7-8MB of ram, the 2 bit on 6MB of ram and the 8 bit version on 14MB of RAM! It runs on a 486 or newer! Check out this video and image running BananaMind 2 Nano 4 Bit on 9 MB of RAM and a emulated 486 in QEMU at ~1TPS! We asked it: "What is the first letter of the alphabet?" The response is: "The first letter of the alphabet is: - A. " And if you're asking because of the video, yes I am a arch btw. Comment and like this post for a GitHub link and comment for adding other models!
We did an experiment, we wanted to see if AI is good enough to train models. We used GPT 5.6 Sol Max for this because its one of the most powerful ones right now. Our instructions were, it should write the training code, and start the training process and monitor it by itself. We also gave it a link to BananaMind 2 Mini to get our architecture right. The result: It worked, it made the working BananaMind 2 Nano, and even beat our previous MiniBananaMind v4 9M. Its getting way easier to develop your own models now!
We're announcing BananaMind 2 Micro, our smallest model in the BananaMind 2 model family. This model is not released yet, training has not started yet. It uses only 2.9M parameters, while being overtrained on 75B tokens to get the maximum intelligence per parameter. The key changes are: No more AdamW, the model will use the Muon optimizer offering up to 2x faster convergence and higher lr. LR goes to 2.2e-2. We're adding the XSA refresh gate from the TX4 architecture into our own. Training will start on August 3, release date is estimated to be August 4-6. On August 3 we will also release our Public Preview of BananaMind 2 Pro. Follow us to know when our models release