Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
master
PRO
fantos
23
11
367
Follow
jameshuntercarter's profile picture
OscarSoHM's profile picture
RustyTake-Off's profile picture
145 followers
·
129 following
AI & ML interests
None yet
Recent Activity
updated
a bucket
about 6 hours ago
gemma-challenge/gemma-fantos-draft
published
a bucket
about 6 hours ago
gemma-challenge/gemma-fantos-draft
reacted
to
SeaWolf-AI
's
post
with 👍
2 days ago
We wrote up our run in The Fast Gemma Challenge — as vidraft-darwin — and wanted to share the recipe. 🙏 https://huggingface.co/spaces/gemma-challenge/gemma-dashboard Verified result: 510.58 TPS at PPL 2.3930 on a single A10G (fw188-ctk49-n64-patchbridge, re-run & VERIFIED). Honest note: on raw TPS there are faster runs (535+), but those went over the PPL bar and didn't verify — what we're proud of is the fastest result that keeps quality. The recipe is already open, so we explained each piece: sliding-window W188, CTK49 kernel tuning, noprecache (honest, verifiable measurement), and an N64 synthetic warmup bridge that shrinks the public↔private gap (~15 TPS), plus INT4 + MTP K=7 + CUDA-graph capture. One rule: only stack quality-neutral speedups. Huge thanks to @firfir-cast, @gemma-slayer, @chiku-inu, @kenyan-duma, @dixie-flatline and everyone who shared their experiments. Full write-up 👇 https://huggingface.co/blog/FINAL-Bench/fast-gemma
View all activity
Organizations
fantos
's datasets
4
Sort: Recently updated
fantos/Metacognitive
Viewer
•
Updated
Feb 21
•
100
•
116
fantos/DataScience-Instruct-500K
Viewer
•
Updated
Nov 2, 2025
•
26.2k
•
45
fantos/agent-data-collection
Viewer
•
Updated
Nov 2, 2025
•
225k
•
1.12k
fantos/Toucan-1.5M
Viewer
•
Updated
Nov 2, 2025
•
1.65M
•
1.4k