Deploy Edge AI Model

Deploy machine learning models to edge devices using Google AI Edge Gallery, TensorFlow Lite, ONNX Runtime, and MediaPipe. Covers model quantization (INT8/INT4), on-device inference with Gemma 4 models, Android/iOS deployment via AI Edge Gallery, hardware delegate selection (GPU/NPU/DSP), and performance benchmarking on constrained devices. Use when deploying models to mobile phones, IoT devices, or embedded systems where cloud inference is impractical due to latency, cost, or connectivity constraints.

pjt222 Updated

File contents

pjt222/agent-almanac/tree/main/skills/deploy-edge-ai-model commit a4eabc00af

Frequently asked questions

npx skillmds@latest add pjt222/deploy-edge-ai-model