i hope an alternative appears soon in case hugging face goes south
Boning Cui
AI & ML interests
Recent Activity
Organizations
I'm officially canceling my Hugging Face Pro subscription today.
I supported this platform because it stood for true openness and neutrality. This acquisition by NVIDIA fundamentally changes that.
Here’s why I’m against this deal:
- Neutrality is dead. NVIDIA is a US-based company. This means US regulations will inevitably dictate platform policies, creating direct pressure on Chinese developers and anyone building open-weight models outside the US.
- Community over bureaucracy. NVIDIA is a massive, slow-moving corporation. This acquisition will likely drown the community in corporate processes and commercial interests. Soon, uploading a simple finetune might become a bureaucratic nightmare.
- Open vs. Proprietary. Hugging Face was built on open-source ideals. NVIDIA? They are a fiercely proprietary hardware company with a minimal track record of meaningful open-source contributions. They sell chips, not freedom.
- And to add insult to injury, NVIDIA has practically abandoned consumer RTX GPUs in 2026 to chase data center profits. Why would I pay them for "openness" when they've turned their back on the very developers who built this ecosystem?
I paid for openness. Not for a corporate takeover.
🤗 was about community.
thanks, lol. I'll keep working just at a smaller scale. i won't stop releasing models obviously, I'm known to work with what I've got
its under my profile at g1-train
@juiceb0xc0de if youve got the resources you can run the code its fully open under my uh very very permissive license literally named: i-have-no-idea-just-use-this.
lol. i might try and set up some sort of donations funding thing later on when ive got time lol
because i used to use the ml-intern-explorers credits thing but that kinda shut down
beta testers please be informed. i am very sorry about this but its how life is 🥲: @guardamarcos @Timmy6767 @MUK-IS-GOAT @smilyai-large-team @Sbui503 @Banaxi-Tech @Bc-AI @atom77777 @Harley-ml @Datdanboi25 @Fishtiks @smartdigitalnetworks @vovaRL @EmetTheGolum @juiceb0xc0de @ProCreations
Hello everyone,
I have some unfortunate news to share with everyone. My earlier estimate for the launch of G1 in late October was inaccurate. We sincerely apologise for any inconvenience this may cause, but with our current compute resources, pretraining a 20B MoE model is not realistically possible within two months.
G1-MINI will also be postponed, but not for nearly as long — only by a few extra months.
This is disappointing, as I know I was excited to launch G1, and I know many people were also watching the model and looking forward to it.
However in my view, I would rather be honest about our limitations than be overly optimistic about something we currently cannot guarantee.
G1 is not cancelled but uh it will be postponed indefinitely until I have the resources needed to train it. This could be next month, or it could take years. For now, I don't want to give another estimated launch date until I know we have the resources to actually make it happen.
In the meantime, SmilyAI will continue developing AI and experimenting with new ideas, and we will provide updates as we go.
Thank you for sticking with us and supporting SmilyAI. We will continue working towards better models in the future.
Also, if you do have the hardware, to run it aka 8xH200s or better, my codebase is fully open under my very permissive license: i-have-no-idea-just-use-this. Basically, do whatever just mention me. Bc-AI/train-g1
— Bc-AI, on behalf of SmilyAI-Labs
@ProCreations what do you think about the whole nvidia acquiring huggingface thing
lol
uh?
We're also announcing 2 new models.
All of our models we will train are:
BananaMind 2.1 Flash Lite, 10M parameters with 8M in transformer and 2M in n-gram. 50B pretraining tokens.
BananaMind 2.1 Lite with 25M parameters, 5M in n-gram and 20M in transformer. 75B pretraining tokens.
BananaMind 2.1 Flash with 50M parameters, with undecided n-gram count yet. 100B pretraining tokens.
BananaMind 2.1 Pro with 145M parameters, with undecided n-gram count yet. 150-200B pretraining tokens.
BananaMind 2.1 Coder with 149M parameters with undecided n-gram count yet.
We're now announcing BananaMind 2.1 NanoCoder, a 10M parameter model focused specifically on coding and BananaMind 2.1 MiniCoder which is a 25M parameter model focused on coding.
Follow us:
@Banaxi-Tech
@vovaRL
@DedeProGames
@ProCreations what about nvidia?
I hate NVIDIA is a valid reason lol. I dont like the sound of this either, so if they introduce any new paywalls im out entirely, other alternatives exist
I'm officially canceling my Hugging Face Pro subscription today.
I supported this platform because it stood for true openness and neutrality. This acquisition by NVIDIA fundamentally changes that.
Here’s why I’m against this deal:
- Neutrality is dead. NVIDIA is a US-based company. This means US regulations will inevitably dictate platform policies, creating direct pressure on Chinese developers and anyone building open-weight models outside the US.
- Community over bureaucracy. NVIDIA is a massive, slow-moving corporation. This acquisition will likely drown the community in corporate processes and commercial interests. Soon, uploading a simple finetune might become a bureaucratic nightmare.
- Open vs. Proprietary. Hugging Face was built on open-source ideals. NVIDIA? They are a fiercely proprietary hardware company with a minimal track record of meaningful open-source contributions. They sell chips, not freedom.
- And to add insult to injury, NVIDIA has practically abandoned consumer RTX GPUs in 2026 to chase data center profits. Why would I pay them for "openness" when they've turned their back on the very developers who built this ecosystem?
I paid for openness. Not for a corporate takeover.
🤗 was about community.
Oh sorry lol. This is my checkpoints it has all the optimisers and other parts in the pt file. I will convert to huggingface once we have something to show. Thanks anyway lol
Also G1s training takes 7 hours to schedule lol, i started it once before school and it only started 3 hours after
1. G1 series status. G1 is training nicely, and the loss is dropping nicely. The metrics are publicly available and i made a small space you can use to see the nice graphs: hugging-science/Loss-Plot-G1-Large
G1-MINI is a lot slower in converging for reasons unknown yet, but we are investigating it.
2. I have built a small chat app for open SLMs here: ml-intern-explorers/slm-arena
Feel free to add your models in a pull request!
That's all for now, early G1 versions will be available for beta testers soon. Thanks to our beta testers: @guardamarcos @Timmy6767 @MUK-IS-GOAT @smilyai-large-team @Sbui503 @Banaxi-Tech @Bc-AI @atom77777 @Harley-ml @Datdanboi25 @Fishtiks @smartdigitalnetworks @vovaRL @EmetTheGolum @juiceb0xc0de @ProCreations