score_variations

shallow

io.github.Lulu-The-Narwhal/dali · Verify this server

Score 2–8 prompt variations for the same generator and rank them best-to-worst. Use this when you've drafted multiple versions of a prompt and want to pick the winner without burning generation credits. Returns a ranked list with per-dimension comparison so you can see exactly why one variant beats another.

100.0/100

1 trials · measured 2 days ago

score_variations scores 100.0/100 on Vouch's measured behaviour index, from 1 real invocation trials against io.github.Lulu-The-Narwhal/dali, measured 31 Aug 2026 under methodology v0.2.0. Every measured component scored 100.

Component breakdown

ComponentWeightValue
Reliability35%not applicable
Schema integrity25%100.0
Failure behaviour15%not applicable
Latency15%not applicable
Concurrency10%not applicable

Tool details

Transport
remote + stdio
Credential class
open
Category
Finance & compliance
Input schema
not declared
Output schema
not declared
Side-effect classification
unclassified

Score history

DayScoreTierMethodology
2026-08-31100.0shallowv0.2.0

Probe evidence

ProbeOutcomes
schema_integritypass: 1

Raw request/response logs are not archived yet — the outcome counts above are drawn directly from every recorded trial.

Embed this score

Available for every tool, scored or not — not a verification perk. Always links back to this page.

Vouch score: score_variations
[![Vouch score](https://vouch.tools/api/tools/c52a6bc1-bd3a-4985-ab74-907e97b131da/badge.svg)](https://vouch.tools/tools/c52a6bc1-bd3a-4985-ab74-907e97b131da)
score_variations — Vouch