Compound AI Hw Sw Bench Eval

This benchmark evaluates hardware-software co-design trade-offs for compound AI applications by measuring end-to-end latency, energy consumption, and accuracy across multi-modal workflows like video QA, evolutionary code generation, and RAG. It probes how different hardware configurations and software optimizations impact system performance under varying latency targets and workload patterns. Use when the user wants to benchmark on Google FRAMES benchmark, or asks about evaluating this task. Reports accuracy.

qhjqhj00 dda31ff 3.9 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/compound-ai-hw-sw-bench-eval commit dda31ff784

Frequently asked questions

npx skillmds add qhjqhj00/compound-ai-hw-sw-bench-eval