-
GaLore: Memory-Efficient LLM Training by Gradient Low-Rank Projection
Paper • 2403.03507 • Published • 192 -
Let the Expert Stick to His Last: Expert-Specialized Fine-Tuning for Sparse Architectural Large Language Models
Paper • 2407.01906 • Published • 44 -
QLoRA: Efficient Finetuning of Quantized LLMs
Paper • 2305.14314 • Published • 64 -
LoRA+: Efficient Low Rank Adaptation of Large Models
Paper • 2402.12354 • Published • 8
Ruozhou He
fward
·
AI & ML interests
None yet
Recent Activity
liked a Space about 2 months ago
Victarry/PP-schedule-visualizer upvoted an article 4 months ago
Vision Language Models Explained liked a model 4 months ago
sapientinc/HRM-Text-1BOrganizations
None yet