Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
🌌
Literally Sedna
15.8
TFLOPS
CNWPlayer
CNWPlayer
7
2
12
Follow
webxos's profile picture
RexTRO111's profile picture
2 followers
·
6 following
AI & ML interests
None yet
Recent Activity
replied
to
ProCreations
's
post
about 10 hours ago
I have hit 300 followers, and I think this calls for a bit of a giveaway 👀 a unique one, too. I have had countless AI projects I have wanted to make but have been (brutally) blocked by compute. Now that I finally have just enough compute to sort of get around (i still don't have enough 😭) and for hitting 300 followers (tysm!) I will be funding three of the communities projects via HuggingFace jobs, giving them 150 dollars max worth of compute each. I will personally be picking the winners, I am looking for projects that genuinely hit the compute wall: great ideas, blocked by compute, just like the countless ideas I've had. To join, head over to https://giveaway.ssh.codes RULES: - Final result must be open weight or open source - Only one submission per person - Have fun!
replied
to
Felladrin
's
post
about 10 hours ago
I've open-sourced the trainer I've been using to build tiny language models from scratch, together with the 95M base model I trained with it. The trainer runs on Deno (https://deno.com, cross-platform), trains on WebGPU, and it writes GGUF directly. No Python/PyTorch. The weights live in a GGUF file from the first step to the last, so every checkpoint is already something llama.cpp can load. The model is https://huggingface.co/Felladrin/Minueza-3-95M-Base: 94.7M parameters, 1.95B tokens seen, 8192 context. And here’s the repository on GitHub: https://github.com/felladrin/gguf-trainer Here on Hugging Face, I published the optimizer state next to the weights, so you can continue the pretraining instead of starting over. Or start your own from nothing: `deno run -A cli.ts demo` trains a tiny one end to end in under a minute. And the docs are written for coding agents, so you can point your agent of choice at the GitHub repo and have it drive the whole pipeline.
liked
a model
about 20 hours ago
bloomer010/Ling-3.0-tiny-GGUF
View all activity
Organizations
None yet
CNWPlayer
's activity
All
Models
Datasets
Spaces
Buckets
Papers
Collections
Community
Posts
Upvotes
Likes
Articles
liked
a model
about 20 hours ago
bloomer010/Ling-3.0-tiny-GGUF
Text Generation
•
8B
•
Updated
1 day ago
•
51.1k
•
47
liked
a model
2 days ago
empero-ai/Qwen3.8-9B-GGUF
Text Generation
•
9B
•
Updated
3 days ago
•
38.3k
•
94
liked
2 models
6 days ago
deepseek-ai/DeepSeek-V4-Pro-0813
Text Generation
•
1.7T
•
Updated
6 days ago
•
31k
•
•
609
Qwen/Qwen3.8-27B
Image-Text-to-Text
•
28B
•
Updated
5 days ago
•
666k
•
•
11.3k
liked
3 models
7 days ago
BananaMind/BananaMind-2-Pro-Preview
Text Generation
•
0.2B
•
Updated
2 days ago
•
972
•
24
CohereLabs/North-Micro-Vision-Instruct
Image-Text-to-Text
•
2B
•
Updated
2 days ago
•
20.4k
•
120
Qwen/Qwen3.8-2.4T-A95B
Text Generation
•
2.4T
•
Updated
7 days ago
•
11.2k
•
•
1.08k
liked
2 models
8 days ago
ni-co-la-s/gemmeh-it
1B
•
Updated
9 days ago
•
23
•
1
ni-co-la-s/gemmeh-GGUF
1B
•
Updated
Jun 1
•
74
•
1
liked
a model
17 days ago
deepseek-ai/DeepSeek-V4-Flash-0731
Text Generation
•
304B
•
Updated
18 days ago
•
2.12M
•
•
3.53k
liked
a model
23 days ago
moonshotai/Kimi-K3
Image-Text-to-Text
•
2.8T
•
Updated
23 days ago
•
2.23M
•
•
10.8k
liked
a model
about 1 month ago
google/gemma-4-26B-A4B-it
Image-Text-to-Text
•
27B
•
Updated
30 days ago
•
9.73M
•
•
1.41k