arxiv:2610.07767
Daniel Wang
tuidan
ยท
AI & ML interests
Efficient AI / LLM / ML System
Recent Activity
authored a paper about 1 hour ago
SVD-LLM V2: Optimizing Singular Value Truncation for Large Language Model Compression authored a paper about 1 hour ago
QUADS: Stabilizing NVFP4 Reinforcement Learning for MoE via QUantization-error Alignment across Dual Sides authored a paper about 1 hour ago
TRACE: Rollout-Guided Quantization-Aware Training for FP4 Reinforcement Learning of MoE Language Models