compare_models

shallow

com.ainetcafe/ai-netcafe · Verify this server

Run one prompt across multiple LLMs in parallel and return every answer side by side with its real measured cost and latency. This answers "which model should I actually use for this kind of task?" with data instead of guesswork — useful before committing a long job to an expensive model. Example — GET https://ainetcafe.com/t/compare_models?prompt=Explain+CAP+theorem+in+1+line

100.0/100

1 trials · measured 8 days ago

compare_models scores 100.0/100 on Vouch's measured behaviour index, from 1 real invocation trials against com.ainetcafe/ai-netcafe, measured 25 Aug 2026 under methodology v0.2.0. Every measured component scored 100.

Component breakdown

ComponentWeightValue
Reliability35%not applicable
Schema integrity25%100.0
Failure behaviour15%not applicable
Latency15%not applicable
Concurrency10%not applicable

Tool details

Transport
remote + stdio
Credential class
self-provisionable
Input schema
not declared
Output schema
not declared
Side-effect classification
unclassified

Score history

DayScoreTierMethodology
2026-08-25100.0shallowv0.2.0

Probe evidence

ProbeOutcomes
schema_integritypass: 1

Raw request/response logs are not archived yet — the outcome counts above are drawn directly from every recorded trial.

Embed this score

Available for every tool, scored or not — not a verification perk. Always links back to this page.

Vouch score: compare_models
[![Vouch score](https://vouch.tools/api/tools/86e31f0a-36be-414a-a9ae-9eda0c2bdc0f/badge.svg)](https://vouch.tools/tools/86e31f0a-36be-414a-a9ae-9eda0c2bdc0f)
compare_models — Vouch