arxiv:2609.33586
Shangzhen Zhu
ArgonnSZ
ยท
AI & ML interests
None yet
Recent Activity
submitted a paper about 12 hours ago
Pretraining Transformers with Quantized Softmax in Attention upvoted a paper about 12 hours ago
Softmax Reparameterization for Output-Head Quantization upvoted a paper about 16 hours ago
Approximating Softmax in Pretrained LLMs: Model Sensitivity and Kernel Acceleration