evidence_us
shallowio.github.beepboop2025/liquilens · Verify this server
Read the US bank current-amended-vintage diagnostic: every FDIC-insured bank scored quarterly from free call-report data, replayed against all 552 receivership failures since 2008 — recall 72.8% at a 21.7-month median lead, AUC 0.854. Includes the marquee replays (SVB, Signature, First Republic, both 2026 catches), the fraud-driven miss kept in full view, and the seven named 2023-2026 misses, published the week they failed; the 2026 scoreboard reads 2 of 4 flagged a year early. Read the semantics before quoting: the flag is a budgeted watchlist (top decile per quarter), not an institution-level verdict, so report the budget beside the recall and the misses beside the hits.
1 trials · measured 8 days ago
evidence_us scores 100.0/100 on Vouch's measured behaviour index, from 1 real invocation trials against io.github.beepboop2025/liquilens, measured 25 Aug 2026 under methodology v0.2.0. Every measured component scored 100.
Component breakdown
| Component | Weight | Value |
|---|---|---|
| Reliability | 35% | not applicable |
| Schema integrity | 25% | 100.0 |
| Failure behaviour | 15% | not applicable |
| Latency | 15% | not applicable |
| Concurrency | 10% | not applicable |
Tool details
- Transport
- remote
- Credential class
- self-provisionable
- Category
- Developer infrastructure
- Input schema
- not declared
- Output schema
- not declared
- Side-effect classification
- unclassified
Score history
| Day | Score | Tier | Methodology |
|---|---|---|---|
| 2026-08-25 | 100.0 | shallow | v0.2.0 |
Probe evidence
| Probe | Outcomes |
|---|---|
| schema_integrity | pass: 1 |
Raw request/response logs are not archived yet — the outcome counts above are drawn directly from every recorded trial.
Embed this score
Available for every tool, scored or not — not a verification perk. Always links back to this page.
[](https://vouch.tools/tools/b12504e1-6b17-41e7-afa9-c9df4e62e695)