-
GaLore: Memory-Efficient LLM Training by Gradient Low-Rank Projection
Paper • 2403.03507 • Published • 191 -
Let the Expert Stick to His Last: Expert-Specialized Fine-Tuning for Sparse Architectural Large Language Models
Paper • 2407.01906 • Published • 47 -
QLoRA: Efficient Finetuning of Quantized LLMs
Paper • 2305.14314 • Published • 63 -
LoRA+: Efficient Low Rank Adaptation of Large Models
Paper • 2402.12354 • Published • 7
Ruozhou He
fward
·
AI & ML interests
None yet
Recent Activity
liked a Space 8 days ago
Victarry/PP-schedule-visualizer upvoted an article about 2 months ago
Vision Language Models Explained liked a model 3 months ago
sapientinc/HRM-Text-1BOrganizations
None yet