โ† LLM Quantization and Compression
structured pruning channels neurons compaction โ€” LLM Quantization and Compression | SOTAAZ Blog