get_performance_summary

shallow

com.olympus-bets/olympus-bets-analytics · Verify this server

Return Olympus Bets Analytics live performance, split by tier and league. Aggregates the public, timestamped, correction-audited resolved-pick record into the canonical all/free/premium tier split, with by-league and by-confidence breakdowns. Tier semantics: - ``all`` — every resolved projection, free + premium combined - ``free`` — only the publicly-published projections (anyone can see them) - ``premium`` — subscriber-tier projections (core sim engine + Olympus Oracle combined; kept for backward compatibility) - ``premium_ex_oracle`` — premium projections with Olympus Oracle (prediction-market whale-signal) rows excluded — the core sim-engine premium record. Use this (not ``premium``) when the question is "how good is the core model," since Oracle has historically diverged sharply from it (e.g. core +30.16u vs oracle -18.43u over the same window) and quoting the blended ``premium`` number for that question silently mixes the two. - ``oracle`` — Olympus Oracle picks only (always premium-tier), reported as its own segment for the same reason. Honest framing: all-time and rolling regimes are both available. Core Premium and Oracle are separated so legacy or source-specific performance cannot obscure the current production system. Both are published. Args: tier: Optional tier filter. Omit to return all five segments. league: Optional league filter applied inside each requested tier. detail: ``summary`` omits breakdowns; ``full`` includes all breakdowns. window: ``all`` preserves the historical contract; rolling windows use the same canonical ledger, grading, tier, and source rules. Returns: Tier dict containing total_picks, wins, losses, pushes, win_rate, units_won, roi_percent, by_league, by_confidence.

100.0/100

1 trials · measured 8 days ago

get_performance_summary scores 100.0/100 on Vouch's measured behaviour index, from 1 real invocation trials against com.olympus-bets/olympus-bets-analytics, measured 25 Aug 2026 under methodology v0.2.0. Every measured component scored 100.

Component breakdown

ComponentWeightValue
Reliability35%not applicable
Schema integrity25%100.0
Failure behaviour15%not applicable
Latency15%not applicable
Concurrency10%not applicable

Tool details

Transport
remote
Credential class
open
Input schema
not declared
Output schema
not declared
Side-effect classification
unclassified

Score history

DayScoreTierMethodology
2026-08-25100.0shallowv0.2.0

Probe evidence

ProbeOutcomes
schema_integritypass: 1

Raw request/response logs are not archived yet — the outcome counts above are drawn directly from every recorded trial.

Embed this score

Available for every tool, scored or not — not a verification perk. Always links back to this page.

Vouch score: get_performance_summary
[![Vouch score](https://vouch.tools/api/tools/c56e4aca-ee6f-41d1-a64a-892a95df5e66/badge.svg)](https://vouch.tools/tools/c56e4aca-ee6f-41d1-a64a-892a95df5e66)
get_performance_summary — Vouch