calculate_review_metrics

shallow

com.ediscoverydecoder/mcp · Verify this server

Score a coded sample when you have a full confusion matrix (true/false positives and negatives) — e.g. comparing a TAR model's calls against a reviewer's. Returns recall, precision, F1, accuracy, and in-sample elusion. Use calculate_control_set_recall if you only have relevant-found vs relevant-missed; calculate_elusion for a discard/null-set sample. Aggregate counts only; not legal advice.

100.0/100

1 trials · measured 8 days ago

calculate_review_metrics scores 100.0/100 on Vouch's measured behaviour index, from 1 real invocation trials against com.ediscoverydecoder/mcp, measured 25 Aug 2026 under methodology v0.2.0. Every measured component scored 100.

Component breakdown

ComponentWeightValue
Reliability35%not applicable
Schema integrity25%100.0
Failure behaviour15%not applicable
Latency15%not applicable
Concurrency10%not applicable

Tool details

Transport
remote
Credential class
open
Category
Content & media
Input schema
not declared
Output schema
not declared
Side-effect classification
unclassified

Score history

DayScoreTierMethodology
2026-08-25100.0shallowv0.2.0

Probe evidence

ProbeOutcomes
schema_integritypass: 1

Raw request/response logs are not archived yet — the outcome counts above are drawn directly from every recorded trial.

Embed this score

Available for every tool, scored or not — not a verification perk. Always links back to this page.

Vouch score: calculate_review_metrics
[![Vouch score](https://vouch.tools/api/tools/a7da90ed-1c75-4b74-b24d-47072db94650/badge.svg)](https://vouch.tools/tools/a7da90ed-1c75-4b74-b24d-47072db94650)
calculate_review_metrics — Vouch