start_run

shallow

md.mostlyright/datasets · Verify this server

Runs a registered recipe. THIS IS THE TOOL THAT SPENDS MONEY. Four modes: sample (a bounded slice — always start here), full (the whole thing), refresh (forward from where the last run reached), backfill (one exact window). Example: {"recipe_id": "…", "recipe_digest": "…", "mode": "sample", "max_rows": 5000}. A sample must state at least one ceiling (max_rows, max_source_bytes or window); a backfill must state window {start, end}. Returns {run_id, status, mode, version, dashboard_url}. A run whose projected spend crosses the workspace threshold comes back status "held" with projected_bytes, projected_runtime_seconds and projected_cost — show those to the user and call confirm_run only if they agree. Next: run_events to watch it, then query_run to check the rows it built.

100.0/100

1 trials · measured 27 days ago

start_run scores 100.0/100 on Vouch's measured behaviour index, from 1 real invocation trials against md.mostlyright/datasets, measured 11 Sept 2026 under methodology v0.2.0. Every measured component scored 100.

Component breakdown

ComponentWeightValue
Reliability35%not applicable
Schema integrity25%100.0
Failure behaviour15%not applicable
Latency15%not applicable
Concurrency10%not applicable

Tool details

Transport
remote
Credential class
open
Input schema
not declared
Output schema
not declared
Side-effect classification
unclassified

Score history

DayScoreTierMethodology
2026-09-11100.0shallowv0.2.0

Probe evidence

ProbeOutcomes
schema_integritypass: 1

Raw request/response logs are not archived yet — the outcome counts above are drawn directly from every recorded trial.

Embed this score

Available for every tool, scored or not — not a verification perk. Always links back to this page.

Vouch score: start_run
[![Vouch score](https://vouch.tools/api/tools/589f0098-2191-409d-b0c8-1c658f19053f/badge.svg)](https://vouch.tools/tools/589f0098-2191-409d-b0c8-1c658f19053f)
start_run — Vouch