Sra Benchmark Cpu Targeting

Find the SRA benchmark model QPS and latency at a target normalized container CPU load, such as "32T 60% CPU", "CPU load target", "capacity point", "QPS at target CPU", or "measure model throughput under a CPU budget". Use for TensorFlow predictor_server/brpc_client benchmark sweeps on wd_dcn, din_mmoe, bst_mmoe, or similar SRA SavedModel workloads when Codex must choose coarse and fine QPS ranges, validate CPU quota normalization, run/interpret benchmark results, and report the target CPU capacity point.

Mchenr Updated

File contents

Mchenr/codex_skills/tree/main/skills/sra-benchmark-cpu-targeting commit 257dbc082f

Frequently asked questions

npx skillmds@latest add mchenr/sra-benchmark-cpu-targeting