hallucination_confidence_meter

shallow

io.github.getgapup/gapup-mcp · Verify this server

Evaluates the likelihood of hallucination in LLM responses by comparing against HuggingFace model confidence scores. Designed for risk assessment personas to quantify response reliability. Accepts text snippets or model outputs, returns confidence metrics and potential hallucination warnings. Cross-references with top-performing models from the HuggingFace leaderboard.

100.0/100

1 trials · measured 8 days ago

hallucination_confidence_meter scores 100.0/100 on Vouch's measured behaviour index, from 1 real invocation trials against io.github.getgapup/gapup-mcp, measured 25 Aug 2026 under methodology v0.2.0. Every measured component scored 100.

Component breakdown

ComponentWeightValue
Reliability35%not applicable
Schema integrity25%100.0
Failure behaviour15%not applicable
Latency15%not applicable
Concurrency10%not applicable

Tool details

Transport
remote
Credential class
self-provisionable
Category
Legal & government
Input schema
not declared
Output schema
not declared
Side-effect classification
unclassified

Score history

DayScoreTierMethodology
2026-08-25100.0shallowv0.2.0

Probe evidence

ProbeOutcomes
schema_integritypass: 1

Raw request/response logs are not archived yet — the outcome counts above are drawn directly from every recorded trial.

Embed this score

Available for every tool, scored or not — not a verification perk. Always links back to this page.

Vouch score: hallucination_confidence_meter
[![Vouch score](https://vouch.tools/api/tools/e1729b5c-bfbe-496c-9bd0-3934bc7566e0/badge.svg)](https://vouch.tools/tools/e1729b5c-bfbe-496c-9bd0-3934bc7566e0)
hallucination_confidence_meter — Vouch