Agent skill · zechenzhangagi

model-pruning

Reduce LLM size and accelerate inference using pruning techniques like Wanda and SparseGPT. Use when compressing models without retraining, achieving 50% s

Reference it in any coding agent with:

@skills zechenzhangagi/model-pruning

Browse the @skills marketplace