Latest benchmark run + per-case results for a public registry skill
Anonymous-readable view of the publisher’s latest benchmark run.
Returns the SkillTestRun row + per-case SkillTestCaseResult rows
with the same Without/With/Δ comparison fields the authenticated
BenchmarkTab uses. Visitor cannot trigger a new run — they only see
what the publisher last ran. A benchmark run is per-model (one model ×
all cases × both arms), so ?model= selects which model’s run the
case table shows; the per-model summary list is benchmark_models on
the detail payload.
Authorizations
Enter your API key (e.g. dai_sk_test_key_001)
Path Parameters
Query Parameters
Show the latest run for THIS runner model (the scorecard's per-model case-table switcher). Omitted = latest run overall.
Response
Successful Response