Edge Lm Inference Eval

Evaluates the feasibility and performance trade-offs of running small generative language models on edge hardware. It probes memory constraints, inference latency, token throughput, and energy efficiency across different quantization schemes and system configurations. Use when the user wants to benchmark on None (system-level inference benchmark), or asks about evaluating this task. Reports generation_throughput.

qhjqhj00 9ff8905 3.2 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/edge-lm-inference-eval commit 9ff8905050

Frequently asked questions

npx skillmds add qhjqhj00/edge-lm-inference-eval