-
The Era of 1-bit LLMs: All Large Language Models are in 1.58 Bits
Paper • 2402.17764 • Published • 630 -
Atom: Low-bit Quantization for Efficient and Accurate LLM Serving
Paper • 2310.19102 • Published • 11 -
AMSP: Super-Scaling LLM Training via Advanced Model States Partitioning
Paper • 2311.00257 • Published • 10 -
BiLLM: Pushing the Limit of Post-Training Quantization for LLMs
Paper • 2402.04291 • Published • 51
Spencer Presley
swpresley
AI & ML interests
Transformers, LLMs, Neural Networks, Hierarchical Classification.
Recent Activity
liked a model about 2 months ago
InternScience/Agents-A1 liked a model about 2 months ago
cyankiwi/Devstral-Small-2-24B-Instruct-2512-AWQ-4bit liked a Space 5 months ago
webml-community/Gemma-4-WebGPUOrganizations
None yet