Agentic_RL_search_tasks

shallow

space.hf.hoyant-su-agentic-rl/agentic-rl · Verify this server

Find real ShellOps CLI benchmark tasks by case-insensitive literal substring in the complete instruction, task ID or published task type. Empty query lists all tasks. Select partition 'all', 'shellops' or 'shellops_pro'; select published split 'all', 'train_src', 'train' or 'test'. Results are ordered by partition then task ID, with explicit pagination and no relevance scoring. The train subset is not double-counted.

100.0/100

1 trials · measured 27 days ago

Agentic_RL_search_tasks scores 100.0/100 on Vouch's measured behaviour index, from 1 real invocation trials against space.hf.hoyant-su-agentic-rl/agentic-rl, measured 11 Sept 2026 under methodology v0.2.0. Every measured component scored 100.

Component breakdown

ComponentWeightValue
Reliability35%not applicable
Schema integrity25%100.0
Failure behaviour15%not applicable
Latency15%not applicable
Concurrency10%not applicable

Tool details

Transport
remote
Credential class
self-provisionable
Input schema
not declared
Output schema
not declared
Side-effect classification
unclassified

Score history

DayScoreTierMethodology
2026-09-11100.0shallowv0.2.0

Probe evidence

ProbeOutcomes
schema_integritypass: 1

Raw request/response logs are not archived yet — the outcome counts above are drawn directly from every recorded trial.

Embed this score

Available for every tool, scored or not — not a verification perk. Always links back to this page.

Vouch score: Agentic_RL_search_tasks
[![Vouch score](https://vouch.tools/api/tools/69acfb1b-782d-4b5c-915f-eeec68bb64b7/badge.svg)](https://vouch.tools/tools/69acfb1b-782d-4b5c-915f-eeec68bb64b7)
Agentic_RL_search_tasks — Vouch