Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
🔄
tired of nightshade?
24.3
TFLOPS
Jackson Yeskie
8BitStudio
11
1
20
Follow
iseekyou's profile picture
Nitesh45433's profile picture
123bbt's profile picture
5 followers
·
8 following
8-bitStudio
AI & ML interests
Text to image, Language models
Recent Activity
new
activity
about 11 hours ago
yethss/The-Quettaset:
Line count
reacted
to
harshitkgupta
's
post
with 🚀
about 17 hours ago
Fine-tuned Qwen 2.5 (0.5B → 3B) on real coding-agent traces, 10 controlled runs, one 16GB Mac. Compared PyTorch MPS vs. Apple MLX for local LoRA SFT — and the honest answer is "it depends on what you're optimizing for": • PyTorch MPS: 2.2x–5.7x faster raw throughput, but hits a hard memory wall — can't load a 3B model in FP16 on 16GB. • Apple MLX: 4-bit QLoRA fits 3B+ models with almost flat memory scaling as context grows (+109 MB going from 1k→4k tokens). • 4-bit quantization doesn't cost you convergence — eval loss tracks closely across backends. • The bigger surprise: most of MLX's slowdown isn't the 4-bit dequant tax. Two of the 10 runs went unquantized to isolate it — dequant only explains 1.07x–1.4x of the gap. A ~4.1–4.6x framework-level gap remains either way. All 10 LoRA adapters + Trackio logs are public so the numbers are checkable, not just claimed. Full writeup: https://huggingface.co/blog/harshitkgupta/fine-tuning-coding-agents-on-mac-pytorch-mps-mlx
new
activity
about 22 hours ago
hugging-science/ripp:
Hugging Science is DEAD???
View all activity
Organizations
None yet
8BitStudio
's models
2
Sort: Recently updated
8BitStudio/Aniimage-2
Text-to-Image
•
0.4B
•
Updated
15 days ago
•
200
•
5
8BitStudio/Aniimage-1
Text-to-Image
•
0.4B
•
Updated
19 days ago
•
133
•
2