get_planning_benchmark

shallow

com.wealthschema/mcp · Verify this server

Get Planning Benchmark v2 - 50 free evaluation tasks for AI financial-advice systems with answer keys by construction: required figures cited to primary source documents, wrong-but-plausible stale values enumerated with reason codes (the Stale Figure Rate metric). Pass blind=true to omit the answer keys for blind evaluation. Scoreboard + protocol: https://www.wealthschema.com/benchmark ; commercial eval packs: https://www.wealthschema.com/ai-eval-sets

100.0/100

1 trials · measured 1 day ago

get_planning_benchmark scores 100.0/100 on Vouch's measured behaviour index, from 1 real invocation trials against com.wealthschema/mcp, measured 6 Oct 2026 under methodology v0.2.0. Every measured component scored 100.

Component breakdown

ComponentWeightValue
Reliability35%not applicable
Schema integrity25%100.0
Failure behaviour15%not applicable
Latency15%not applicable
Concurrency10%not applicable

Tool details

Transport
remote
Credential class
open
Input schema
not declared
Output schema
not declared
Side-effect classification
unclassified

Score history

DayScoreTierMethodology
2026-10-06100.0shallowv0.2.0

Probe evidence

ProbeOutcomes
schema_integritypass: 1

Raw request/response logs are not archived yet — the outcome counts above are drawn directly from every recorded trial.

Embed this score

Available for every tool, scored or not — not a verification perk. Always links back to this page.

Vouch score: get_planning_benchmark
[![Vouch score](https://vouch.tools/api/tools/5faa11a9-c817-4507-beb4-dfa4f6b2397d/badge.svg)](https://vouch.tools/tools/5faa11a9-c817-4507-beb4-dfa4f6b2397d)
get_planning_benchmark — Vouch