set_benchmark_reference

shallow

ai.entityenricher/enricher · Verify this server

Save the gold reference for an enrichment or schema-generation benchmark. Only set reference_verified=true after checking its values against trusted evidence or obtaining human sign-off; a generated answer alone is not verification. Schema-generation references must be GeneratedJsonSchema objects. Sample-generation scenarios reject references because they are rubric-scored. Requires owner and a plan with benchmarks; no LLM call. A verified reference enables run_benchmark. See enricher://docs/model-benchmark.

100.0/100

1 trials · measured 2 days ago

set_benchmark_reference scores 100.0/100 on Vouch's measured behaviour index, from 1 real invocation trials against ai.entityenricher/enricher, measured 6 Oct 2026 under methodology v0.2.0. Every measured component scored 100.

Component breakdown

ComponentWeightValue
Reliability35%not applicable
Schema integrity25%100.0
Failure behaviour15%not applicable
Latency15%not applicable
Concurrency10%not applicable

Tool details

Transport
remote
Credential class
gated
Input schema
not declared
Output schema
not declared
Side-effect classification
unclassified

Score history

DayScoreTierMethodology
2026-10-06100.0shallowv0.2.0

Probe evidence

ProbeOutcomes
schema_integritypass: 1

Raw request/response logs are not archived yet — the outcome counts above are drawn directly from every recorded trial.

Embed this score

Available for every tool, scored or not — not a verification perk. Always links back to this page.

Vouch score: set_benchmark_reference
[![Vouch score](https://vouch.tools/api/tools/d072dd3d-d259-4ba9-93d2-0db0026d402c/badge.svg)](https://vouch.tools/tools/d072dd3d-d259-4ba9-93d2-0db0026d402c)
set_benchmark_reference — Vouch