Collections
Discover the best community collections!
Collections trending this week
-
The Era of 1-bit LLMs: All Large Language Models are in 1.58 Bits
Paper β’ 2402.17764 β’ Published β’ 630 -
Atom: Low-bit Quantization for Efficient and Accurate LLM Serving
Paper β’ 2310.19102 β’ Published β’ 11 -
AMSP: Super-Scaling LLM Training via Advanced Model States Partitioning
Paper β’ 2311.00257 β’ Published β’ 10 -
BiLLM: Pushing the Limit of Post-Training Quantization for LLMs
Paper β’ 2402.04291 β’ Published β’ 51
-
The Era of 1-bit LLMs: All Large Language Models are in 1.58 Bits
Paper β’ 2402.17764 β’ Published β’ 630 -
Atom: Low-bit Quantization for Efficient and Accurate LLM Serving
Paper β’ 2310.19102 β’ Published β’ 11 -
AMSP: Super-Scaling LLM Training via Advanced Model States Partitioning
Paper β’ 2311.00257 β’ Published β’ 10 -
BiLLM: Pushing the Limit of Post-Training Quantization for LLMs
Paper β’ 2402.04291 β’ Published β’ 51