Agent Skills: omni-bench-run
Run a model through the omni-bench benchmark to MEASURE it — produce a run-artifact and score it, then read the numbers. Use when asked to run, benchmark, measure, evaluate, or "test the results of" a model with omni-bench — ASR (WER/CER, RTFx) or text generation (tok/s, TTFT, prefill tok/s, prompt-cache speedup). Covers the adapter seam (Transcriber/Generator), the prepare→run→score→diff CLI flow, the offline no-download smoke task, and how to interpret each metric. NOT for publishing to a leaderboard (use omni-bench-publish) or changing the framework itself (use omni-bench).
UncategorizedID: beshkenadze/claude-skills-marketplace/omni-bench-run
47
Install this agent skill to your local
Skill Files
Browse the full folder contents for omni-bench-run.
Loading file tree…
Select a file to preview its contents.