assay_batch
shallowcom.alphaassay/mcp · Verify this server
Use this when you optimised over a grid and want to submit the WHOLE sweep honestly -- every variant a ledger trial the family budget prices, so selection bias cannot hide. A research audit, not buy/sell advice. Submit your WHOLE parameter sweep honestly -- one call, every variant a ledger trial. You searched N variants; showing only the winner is exactly the selection bias the family budget prices. This tool makes the honest path the cheap path: send up to 25 DSL spec variants (as a specs list, OR base_spec + grid of dot-paths like {"signal.left.window": [2,3,4]} -- the built-in sweep adapter expands it in canonical order), plus your candles. Every variant is recorded in your family's trial ledger and verdicted with the CUMULATIVE deflation its siblings created -- later variants see the budget the earlier ones burned. The report is demote-only by construction: survives/deflated_out counts and per-variant verdicts, never a ranking, never a best pick. Metering is per valid processed variant (journaled individually, settle-before-run); if credit runs out mid-batch you get exactly what was paid for and an explicit declined count -- no silent truncation. Data minimisation on request: sketch_opt_out=true applies to every variant -- no return sketches persisted, every variant counts IN FULL toward the family budget (no evidence, no discount; disclosed per variant). The response is code-computed and ledger-dependent: a new call can change the family budget. Byte-identical output is promised only by stored replay of the same non-empty request_id with the same canonical request. NOT a buy/sell signal. Price: per valid variant; see https://api.alphaassay.com/v1/meta/pricing (api_key required -- account setup at https://api.alphaassay.com/account).
1 trials · measured 8 days ago
assay_batch scores 100.0/100 on Vouch's measured behaviour index, from 1 real invocation trials against com.alphaassay/mcp, measured 25 Aug 2026 under methodology v0.2.0. Every measured component scored 100.
Component breakdown
| Component | Weight | Value |
|---|---|---|
| Reliability | 35% | not applicable |
| Schema integrity | 25% | 100.0 |
| Failure behaviour | 15% | not applicable |
| Latency | 15% | not applicable |
| Concurrency | 10% | not applicable |
Tool details
- Transport
- remote
- Credential class
- self-provisionable
- Input schema
- not declared
- Output schema
- not declared
- Side-effect classification
- unclassified
Score history
| Day | Score | Tier | Methodology |
|---|---|---|---|
| 2026-08-25 | 100.0 | shallow | v0.2.0 |
Probe evidence
| Probe | Outcomes |
|---|---|
| schema_integrity | pass: 1 |
Raw request/response logs are not archived yet — the outcome counts above are drawn directly from every recorded trial.
Embed this score
Available for every tool, scored or not — not a verification perk. Always links back to this page.
[](https://vouch.tools/tools/d200d2a8-cdd4-4c51-8d06-3b12a33a4730)