Inference Framework Benchmark Eval

Evaluates the inference performance of four deep learning frameworks (TensorRT, ONNX Runtime, OpenVINO, TensorFlow XLA) across four CNN architectures on GPU hardware. It probes how configuration settings, graph optimizations, and batch sizes impact inference speed and resource utilization, including co-localized model ensembles. Use when the user wants to benchmark on ImageNet, or asks about evaluating this task. Reports speed.

qhjqhj00 4fcb9d1 3.7 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/inference-framework-benchmark-eval commit 4fcb9d1b5c

Frequently asked questions

npx skillmds add qhjqhj00/inference-framework-benchmark-eval