AI & ML interests

Investigating the power of Smilyai's, Orion-Flagship model

Recent Activity

Bc-AI  updated a Space about 3 hours ago
Project-Prism/UI-Preview
Bc-AI  published a Space about 3 hours ago
Project-Prism/UI-Preview
Bc-AI  updated a Space about 18 hours ago
Project-Prism/README
View all activity

Bc-AI 
published a Space about 3 hours ago
Banaxi-Tech 
posted an update about 11 hours ago
view post
Post
54
saicr
is going to have its first model launch around October 2.
We're working so hard to get the models available as soon as possible.
  • 6 replies
·
Bc-AI 
posted an update about 17 hours ago
view post
Post
54
Hello everyone! 👋

A small SmilyAI Labs update!

G1-MINI has now seen around 8B tokens during its current run, and pretraining is still going strong.

Our E1 (Efficiency-1) prototype has also reached 15B pretraining tokens. E1 has 1B total parameters while activating under 100M parameters per token. It features adaptive activation, meaning easier tokens can use less compute while harder tokens receive more.

We plan to open-source E1 ASAP! 🚀

We’re also excited to announce Project Prism, which will provide limited access to our upcoming Orion Flagship model, powered by our T2 architecture.

Note: T2 here refers to the architecture, not our T2 (Thinker-2) model.

Applications for Project Prism are available through the org page, with more details coming soon!

Finally, welcome @soyL061215 , who joined the Hugging Face org today! 🎉

Thanks to our existing members:
@smilyai-large-team @MUK-IS-GOAT @Keeby-smilyai @Bc-AI

— Bc-AI
SmilyAI Labs
  • 3 replies
·
Bc-AI 
updated a Space about 18 hours ago
Bc-AI 
published a Space about 18 hours ago
Banaxi-Tech 
posted an update 1 day ago
view post
Post
79
This day is Sol nice.
Banaxi-Tech 
posted an update 3 days ago
view post
Post
90
We have some updates to @BananaMindBot 🍌
It can now train models, ask it to train a model, and i will train it for you.
It now can also merge PRs And like models.
  • 28 replies
·
Banaxi-Tech 
posted an update 4 days ago
view post
Post
2732
We've released @BananaMindBot .

Most things you do on HuggingFace, BananaMindBot can do. Fast

Mention @BananaMindBot on a model, dataset, Space discussion, paper, blog comment, or top-level post and it'll reply there.

It's powered by North Code Mini (Qwen3.8 27B, with GPT OSS 120B as fallback).

A few things it can do:

Search for models and datasets
Look up users and orgs and see what they've published
Read model cards, configs, dataset files, blog posts, and org profiles
Answer questions about what it finds
Write and run its own code in a locked-down sandbox when it needs to verify something
Check things like a model's real parameter count from the safetensors headers instead of just repeating the model card
Remember something for later if you explicitly ask it to
Forward a message to @Banaxi-Tech
Post a daily roundup of developments in the small-language-model space

It won't execute code you give it. It can read and review that code, but anything it runs is code it wrote itself.

It also can't access private data or credentials.

Mention it somewhere.

It's going to also find this post!

(Some parts inspired by CompactBot and @CompactAI Follow them please)

  • 37 replies
·
Banaxi-Tech 
posted an update 5 days ago
Banaxi-Tech 
posted an update 8 days ago
view post
Post
2758
Hi everyone!
We've seen some people getting confused with the BananaMind Leaderboards so ill explain!

We have 2 leaderboards, THESE are NOT the same, first BananaMind/BananaMindBench-Leaderboard which is ONLY for BananaMind Base Bench 1.1. The 10/10 scores do NOT mean that the benchmark is saturated. It isnt saturated, these models score 10/10 because they are the current best models, our /10 ranking system works by taking the ELO scores and then comparing them to the scores in the same size range. So if a better model releases that gets 10/10 and the others get lower.

And we also have the BananaMind SLM leaderboard, not the BananaMindBench leaderboard which uses ARC EASY,PIQA,Hellaswag, Arithmark 3 and the BananaMind Base Bench 1.1. This is the newer and recommended version.


Hope you understand it now!


cODeQ
  • 12 replies
·
Banaxi-Tech 
posted an update 9 days ago
view post
Post
16
What is a model?
What is it?
You don't know?
Banaxi-Tech 
posted an update 10 days ago
view post
Post
60
Its Monday. Getting back to working on ACR 1.0.
  • 11 replies
·
Banaxi-Tech 
posted an update 13 days ago
view post
Post
3022
We're releasing the BananaMind SLM Leaderboard!
It offers a easier look at which models are actually good for your specific needs.
Its primary metric, Intelligence index is a composite of BananaMind Base Bench, PIQA, Hellaswag, ARC Easy and Arithmark 3.
It also allows you to see specific categories like Commonsense on a model.


Check it out at BananaMind/BananaMind-SLM-Leaderboard

  • 2 replies
·
Banaxi-Tech 
posted an update 14 days ago
view post
Post
5588
AGI has arrived.


Just gotta wait for the GLM distill.
  • 29 replies
·
Bc-AI 
posted an update 14 days ago
view post
Post
127
Hello Everyone!
I am happy to announce a few things.
1. G1-MINI
G1-MINI is now in pretraining and is training at a steady pace. Our current ETAs state completion and launch in about 15-20 days, somewhere near the end of September.
2. G1-NANO
G1-NANO is also being pretrained as we speak at a pace of over 400K tokens per second processing more than 10B tokens in 12 hours. This allows us to train extremely fast, and we will launch it somewhere around 15th September.
3. We have begun work on FrameShot, a dual image and video generation model at around 4B dense parameters. This is expected to launch around late December with no promised date.
- Bc-AI
Bc-AI 
posted an update 16 days ago
view post
Post
2636
Hello everyone!
I have 2 announcements today!
The first one is the launch of our new API platform! You can make a account and get 5 dollars free credits. No credits card needed because i have no idea how to set up a payment's thing. If you want more credits just email me at smilyai@outlook.com .
The platform currently features G1-Preview a preview of G1 and the older Mira-1-Large.
2nd announcement is we have started working on G1-MINI so expect a late October Ish launch
- Bc-AI on behalf of Smilyai-Labs
  • 2 replies
·
Banaxi-Tech 
posted an update 18 days ago
view post
Post
158
Checkout
saicr
.
Details coming.
We're switching goals.
Join or mission.
  • 4 replies
·
Bc-AI 
posted an update 19 days ago
view post
Post
131
Hey everyone,

I just wanna say sorry about the change to G1.

I know a lot of you were really looking forward to the original 20B MoE, and honestly, I was really excited about it too.

Unfortunately, the free compute credits I was using from ML Intern Explorers were removed by Hugging Face. That changed what I can realistically do with the original plan, so I've decided to move G1 over to a Qwen 3.8 27B base instead. Its not just another finetune though, I am inserting extra layers and putting it through my vigourous pipeline. Results will be open source.

I know that's probably disappointing, especially for the people who were specifically waiting for the 20B MoE. I'm genuinely sorry about that.

I really appreciate everyone who got excited about G1 in the first place. I didn't expect this change either, but I'm still really excited to see where G1 can go from here.

The original G1 codebase will stay open too. Its under my profile: Bc-AI/train-g1

Thanks for sticking with us
Thanks to our beta testers, you can become one in the beta testers organisation: @guardamarcos @Timmy6767 @MUK-IS-GOAT @smilyai-large-team @Sbui503 @Banaxi-Tech @Bc-AI @atom77777 @Harley-ml @Datdanboi25 @Fishtiks @smartdigitalnetworks @vovaRL @EmetTheGolum @juiceb0xc0de @ProCreations

— Bc on behalf of Smilyai Labs
  • 2 replies
·
Bc-AI 
posted an update 20 days ago
view post
Post
123
SmilyAI Weekly Update

Hello everyone,

I have some unfortunate news to share with everyone. My earlier estimate for the launch of G1 in late October was inaccurate. We sincerely apologise for any inconvenience this may cause, but with our current compute resources, pretraining a 20B MoE model is not realistically possible within two months.

G1-MINI will also be postponed, but not for nearly as long — only by a few extra months.

This is disappointing, as I know I was excited to launch G1, and I know many people were also watching the model and looking forward to it.

However in my view, I would rather be honest about our limitations than be overly optimistic about something we currently cannot guarantee.

G1 is not cancelled but uh it will be postponed indefinitely until I have the resources needed to train it. This could be next month, or it could take years. For now, I don't want to give another estimated launch date until I know we have the resources to actually make it happen.

In the meantime, SmilyAI will continue developing AI and experimenting with new ideas, and we will provide updates as we go.

Thank you for sticking with us and supporting SmilyAI. We will continue working towards better models in the future.

Also, if you do have the hardware, to run it aka 8xH200s or better, my codebase is fully open under my very permissive license: i-have-no-idea-just-use-this. Basically, do whatever just mention me. Bc-AI/train-g1

— Bc-AI, on behalf of SmilyAI-Labs
  • 10 replies
·