get_gecko_test

shallow

ai.benchgecko/benchgecko · Verify this server

Results of one of BenchGecko's own tests (own measurements, CC BY 4.0): who-are-you (Does the model know which lab made it?) world-map (How well does the model draw the world map from memory?) censorship-index (How often does the model refuse legitimate questions?) knowledge-horizon (Where does the model's knowledge of world events actually stop?) tokenizer-tax (How many more tokens does the same text cost outside English?) same-model-different-host (Do providers serving the same open model give the same quality?) model-drift-index (Do models quietly change behind the same name?)

100.0/100

1 trials · measured 1 day ago

get_gecko_test scores 100.0/100 on Vouch's measured behaviour index, from 1 real invocation trials against ai.benchgecko/benchgecko, measured 6 Oct 2026 under methodology v0.2.0. Every measured component scored 100.

Component breakdown

ComponentWeightValue
Reliability35%not applicable
Schema integrity25%100.0
Failure behaviour15%not applicable
Latency15%not applicable
Concurrency10%not applicable

Tool details

Transport
remote
Credential class
open
Input schema
not declared
Output schema
not declared
Side-effect classification
unclassified

Score history

DayScoreTierMethodology
2026-10-06100.0shallowv0.2.0

Probe evidence

ProbeOutcomes
schema_integritypass: 1

Raw request/response logs are not archived yet — the outcome counts above are drawn directly from every recorded trial.

Embed this score

Available for every tool, scored or not — not a verification perk. Always links back to this page.

Vouch score: get_gecko_test
[![Vouch score](https://vouch.tools/api/tools/2d4b0a19-5cfa-4eee-b10e-9a884575a32b/badge.svg)](https://vouch.tools/tools/2d4b0a19-5cfa-4eee-b10e-9a884575a32b)
get_gecko_test — Vouch