Tag
A tweet listing 12 compression techniques for reducing LLM size while maintaining performance, including quantization, distillation, and low-rank adaptation.