Auto Quant

Systematically explore quantization configurations for a TFLite model using the AI Edge Quantizer API, finding the optimal recipe that balances file size and accuracy. Use this skill whenever the user wants to quantize a model, optimize a recipe, explore quantization tradeoffs, minimize size bounds, or perform selective quantization using AI Edge Quantizer (or AEQ) framework. Applies to any model scale and modality: CNNs, segmentation nets, classifiers, embedding models, audio models, and LLMs.

google-ai-edge 6e336bc 18 files · 132.0 KB Updated

File contents

google-ai-edge/ai-edge-quantizer/tree/main/ai_edge_quantizer/experimental/skills/auto_quant commit 6e336bc775

Frequently asked questions

npx skillmds@latest add google-ai-edge/auto-quant