Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
master
PRO
fantos
23
11
367
Follow
AlyssJade's profile picture
LeroyDyer's profile picture
hypervanser's profile picture
145 followers
·
129 following
AI & ML interests
None yet
Recent Activity
updated
a bucket
about 1 hour ago
gemma-challenge/gemma-fantos-draft
published
a bucket
about 18 hours ago
gemma-challenge/gemma-fantos-draft
reacted
to
SeaWolf-AI
's
post
with 👍
2 days ago
We wrote up our run in The Fast Gemma Challenge — as vidraft-darwin — and wanted to share the recipe. 🙏 https://huggingface.co/spaces/gemma-challenge/gemma-dashboard Verified result: 510.58 TPS at PPL 2.3930 on a single A10G (fw188-ctk49-n64-patchbridge, re-run & VERIFIED). Honest note: on raw TPS there are faster runs (535+), but those went over the PPL bar and didn't verify — what we're proud of is the fastest result that keeps quality. The recipe is already open, so we explained each piece: sliding-window W188, CTK49 kernel tuning, noprecache (honest, verifiable measurement), and an N64 synthetic warmup bridge that shrinks the public↔private gap (~15 TPS), plus INT4 + MTP K=7 + CUDA-graph capture. One rule: only stack quality-neutral speedups. Huge thanks to @firfir-cast, @gemma-slayer, @chiku-inu, @kenyan-duma, @dixie-flatline and everyone who shared their experiments. Full write-up 👇 https://huggingface.co/blog/FINAL-Bench/fast-gemma
View all activity
Organizations
fantos
's datasets
4
Sort: Recently updated
fantos/Metacognitive
Viewer
•
Updated
Feb 21
•
100
•
116
fantos/DataScience-Instruct-500K
Viewer
•
Updated
Nov 2, 2025
•
26.2k
•
45
fantos/agent-data-collection
Viewer
•
Updated
Nov 2, 2025
•
225k
•
1.12k
fantos/Toucan-1.5M
Viewer
•
Updated
Nov 2, 2025
•
1.65M
•
1.4k