holy chatgpt
alkinun
AtAndDev
AI & ML interests
decentralize
Recent Activity
liked a Space about 5 hours ago
lerobot/visualize_dataset liked a dataset about 5 hours ago
lerobot/svla_so100_pickplace updated a dataset about 6 hours ago
AtAndDev/example_datasetOrganizations
replied to SoulInPsyAbstract's post about 20 hours ago
reacted to JonnaMat's post with 🔥 1 day ago
Post
2728
🚗 The reasoning backbone quadruples from 8B to 32B , while the action expert remains at 2.3B!
👀 We took a closer look at the architectural evolution from nvidia/Alpamayo-1.5-10B to nvidia/Alpamayo2-Super .
Read the analysis here:
https://huggingface.co/blog/JonnaMat/alpamayo2-super
Our analysis explores some implications of this design choice, especially from a distillation perspective where keeping the expert compact could be key for efficient deployment. 🧠
👀 We took a closer look at the architectural evolution from nvidia/Alpamayo-1.5-10B to nvidia/Alpamayo2-Super .
Read the analysis here:
https://huggingface.co/blog/JonnaMat/alpamayo2-super
Our analysis explores some implications of this design choice, especially from a distillation perspective where keeping the expert compact could be key for efficient deployment. 🧠
replied to SoulInPsyAbstract's post 1 day ago
holy chatgpt
reacted to AxionLab-official's post with 🚀🔥 4 days ago
reacted to Enderchef's post with 🤗❤️🚀🔥 4 days ago
Post
3395
🚀 Supra2 100M is out, and multiple other SLM orgs are gaining power!
Following takes a press. Please follow:
fromziro
SupraLabs
AxiomicLabs
Following takes a press. Please follow:
reacted to appvoid's post with 🔥 5 days ago
Post
1344
Giving free early access to the gguf for some of you today! Tell me what you think.
CEAMFA/palmer-007-preview
CEAMFA/palmer-007-preview
reacted to LH-Tech-AI's post with 🔥 7 days ago
Post
2373
Announcing The Supra2 Family And Supra2-100M
Today, we are announcing a brand-new series of SupraLabs models: Supra2
This series will feature various models, including such as:
- 🐜 Supra2-Nano (0.4M) → The smallest Supra2 model.
- 🤏 Supra2-Small (1.4M) → The tiny model that runs everywhere.
- 💪 Supra2-Medium (25M) → Our medium class model in the Supra2 family. The powerful midsizer.
- 🔥 Supra2-Pro (100M): base, instruct, reasoning, code, math and more! → The most capable model yet! A real allrounder for all your everyday tasks.
- 🎨 Supra2-IMG → our generative text-to-image model
...and many more...
Current progress:
- Nano (0.4M) and Small (1.4M): in training; almost done. Baseline set.
- Medium (25M): coming soon...
- Pro (100M): in training; finishes in 66 hours - Monday, 3rd August 2026, 12:00AM
- IMG: coming soon...
You can support us with a like and follow if you want!
Don't miss our next release! Stay tuned...
Today, we are announcing a brand-new series of SupraLabs models: Supra2
This series will feature various models, including such as:
- 🐜 Supra2-Nano (0.4M) → The smallest Supra2 model.
- 🤏 Supra2-Small (1.4M) → The tiny model that runs everywhere.
- 💪 Supra2-Medium (25M) → Our medium class model in the Supra2 family. The powerful midsizer.
- 🔥 Supra2-Pro (100M): base, instruct, reasoning, code, math and more! → The most capable model yet! A real allrounder for all your everyday tasks.
- 🎨 Supra2-IMG → our generative text-to-image model
...and many more...
Current progress:
- Nano (0.4M) and Small (1.4M): in training; almost done. Baseline set.
- Medium (25M): coming soon...
- Pro (100M): in training; finishes in 66 hours - Monday, 3rd August 2026, 12:00AM
- IMG: coming soon...
You can support us with a like and follow if you want!
Don't miss our next release! Stay tuned...
reacted to AxionLab-official's post with 🔥 8 days ago
Post
2224
reacted to hypothetical's post with 🤗 9 days ago
Post
3153
The best quality public GGUFs and experimental MLX inference for open LLMs!
Collection: https://huggingface.co/collections/TheStageAI/edge-lm
Collection: https://huggingface.co/collections/TheStageAI/edge-lm
reacted to danielhanchen's post with ❤️🔥 10 days ago
Post
3974
Kimi K3 can now be run locally! ✨
The 1-bit model retains ~78.9% accuracy after we shrunk it from 1.56TB to 594GB (-62% size).
Run on a Mac Studio connected with 128GB RAM device. Kimi K3 is the strongest open model to date.
GGUF: unsloth/Kimi-K3-GGUF
Guide: https://unsloth.ai/docs/models/kimi-k3
The 1-bit model retains ~78.9% accuracy after we shrunk it from 1.56TB to 594GB (-62% size).
Run on a Mac Studio connected with 128GB RAM device. Kimi K3 is the strongest open model to date.
GGUF: unsloth/Kimi-K3-GGUF
Guide: https://unsloth.ai/docs/models/kimi-k3
reacted to ProCreations's post with 🔥🚀 12 days ago
reacted to ProCreations's post with 👍 20 days ago
Post
2390
grug-v2 here. grug v1 good but lost tool use brain. grug v2 RL fix. grug v2 more SFT to make more grug.
Try grug v2 here:
ProCreations/grug-v2-9b-gguf
ProCreations/grug-v2-9b
Demo space here:
ProCreations/grug-v2-9b-demo
Try grug v2 here:
ProCreations/grug-v2-9b-gguf
ProCreations/grug-v2-9b
Demo space here:
ProCreations/grug-v2-9b-demo
reacted to danielhanchen's post with 🤗 20 days ago
Post
5936
Gemma 4 is now faster and much more accurate! 🚀
Google made huge improvements to tool-calling and chat accuracy, reliability + speed.
To get fixes, re-download our updated GGUF, MLX, NVFP4 quants!
Unsloth quants: https://huggingface.co/collections/unsloth/gemma-4
Gemma 4 Guide: https://unsloth.ai/docs/models/gemma-4
Google made huge improvements to tool-calling and chat accuracy, reliability + speed.
To get fixes, re-download our updated GGUF, MLX, NVFP4 quants!
Unsloth quants: https://huggingface.co/collections/unsloth/gemma-4
Gemma 4 Guide: https://unsloth.ai/docs/models/gemma-4