Mohammad Mozaffari
ยท
AI & ML interests
Compression of Large Language Models through Sparsity, Quantization, and Low-rank Approximation (the Compression Trinity)
Recent Activity
updated a collection 1 day ago
PATCH updated a collection 1 day ago
PATCH updated a collection 1 day ago
PATCH