Mohammad Mozaffari
ยท
AI & ML interests
Compression of Large Language Models through Sparsity, Quantization, and Low-rank Approximation (the Compression Trinity)
Recent Activity
updated a collection 3 days ago
PATCH updated a collection 3 days ago
PATCH updated a collection 3 days ago
PATCH