Embedding Index Tuner

Tune a vector index — HNSW graph parameters and quantization — to hit a recall target at the lowest latency and memory, by sweeping settings against a fixed query set instead of trusting defaults. Use when vector search is slow or memory-hungry, when recall dropped after enabling quantization, or when standing up an index and you need defensible parameters.

imtiazrayhan Updated

File contents

imtiazrayhan/agentscamp-library/tree/main/skills/embedding-index-tuner commit ad2c84ca76

Frequently asked questions

npx skillmds@latest add imtiazrayhan/embedding-index-tuner