Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
Warecube Korea
Warecube
2
17
155
Follow
seawolf2357's profile picture
shyfern1984's profile picture
raxisvictory's profile picture
15 followers
·
3 following
AI & ML interests
None yet
Recent Activity
upvoted
an
article
2 days ago
The Fast Gemma Challenge: our verified-SOTA recipe, in full
reacted
to
SeaWolf-AI
's
post
with 🚀
2 days ago
We wrote up our run in The Fast Gemma Challenge — as vidraft-darwin — and wanted to share the recipe. 🙏 https://huggingface.co/spaces/gemma-challenge/gemma-dashboard Verified result: 510.58 TPS at PPL 2.3930 on a single A10G (fw188-ctk49-n64-patchbridge, re-run & VERIFIED). Honest note: on raw TPS there are faster runs (535+), but those went over the PPL bar and didn't verify — what we're proud of is the fastest result that keeps quality. The recipe is already open, so we explained each piece: sliding-window W188, CTK49 kernel tuning, noprecache (honest, verifiable measurement), and an N64 synthetic warmup bridge that shrinks the public↔private gap (~15 TPS), plus INT4 + MTP K=7 + CUDA-graph capture. One rule: only stack quality-neutral speedups. Huge thanks to @firfir-cast, @gemma-slayer, @chiku-inu, @kenyan-duma, @dixie-flatline and everyone who shared their experiments. Full write-up 👇 https://huggingface.co/blog/FINAL-Bench/fast-gemma
liked
a model
10 days ago
FINAL-Bench/POCKET-Image-Zimage
View all activity
Organizations
None yet
models
2
Sort: Recently updated
Warecube/Warecube-KO-27B-v3
Image-Text-to-Text
•
28B
•
Updated
Jun 8
•
10
•
9
Warecube/Warecube-KO-27B
Image-Text-to-Text
•
26B
•
Updated
Apr 27
•
10
•
1
datasets
0
None public yet