Litert Quantization Calib

Assists the user to calibrate, merge, and statically quantize litert LLM models (such as Gemma 3) in standard open-source (OSS) environments. Use when the user wants to run LLM calibration, merge task JSON results, align KV cache parameters across models, protect sensitive layers in Float32, or run quantized inference testing. Don't use for JAX/PyTorch custom quantization configurations or non-litert models.

google-ai-edge 91965b7 8.5 KB Updated

File contents

google-ai-edge/litert-torch/tree/main/litert_torch/generative/export_hf/experimental/calib/skills/tflite_quantization_calib commit 91965b79d6

Frequently asked questions

npx skillmds@latest add google-ai-edge/litert-quantization-calib