get_benchmark_scenario_results

shallow

ai.entityenricher/enricher · Verify this server

Filter, rank and limit a scenario's per-model benchmark results. No LLM call. overall blends quality, speed and cost using organization task weights and is null if a component is missing. Status tags are independent: success does not exclude stale or stale_score results. Missing sort metrics come last in either direction. See enricher://docs/model-benchmark for interpretation.

100.0/100

1 trials · measured 1 day ago

get_benchmark_scenario_results scores 100.0/100 on Vouch's measured behaviour index, from 1 real invocation trials against ai.entityenricher/enricher, measured 6 Oct 2026 under methodology v0.2.0. Every measured component scored 100.

Component breakdown

ComponentWeightValue
Reliability35%not applicable
Schema integrity25%100.0
Failure behaviour15%not applicable
Latency15%not applicable
Concurrency10%not applicable

Tool details

Transport
remote
Credential class
gated
Input schema
not declared
Output schema
not declared
Side-effect classification
unclassified

Score history

DayScoreTierMethodology
2026-10-06100.0shallowv0.2.0

Probe evidence

ProbeOutcomes
schema_integritypass: 1

Raw request/response logs are not archived yet — the outcome counts above are drawn directly from every recorded trial.

Embed this score

Available for every tool, scored or not — not a verification perk. Always links back to this page.

Vouch score: get_benchmark_scenario_results
[![Vouch score](https://vouch.tools/api/tools/d8994f42-db98-4cd4-93a4-d0cdd7ad1745/badge.svg)](https://vouch.tools/tools/d8994f42-db98-4cd4-93a4-d0cdd7ad1745)
get_benchmark_scenario_results — Vouch