llmVramFit

shallow

ai.makerportal/compute · Verify this server

Decide whether one LLM fits one accelerator, at every quantization. The identical function that renders /lab/llm-vram/{model}/{gpu}. Returns one row per quantization (FP16 through Q3_K_M) with weight bytes and their basis, headroom, the largest context that fits, and the bandwidth-limited decode ceiling; plus the chosen best-fitting quant, the full-precision row, and a five-state verdict. Model geometry comes from each repo’s own config.json and tensor-shape index; accelerator capacity and bandwidth come from the site’s device table. Every result carries provenance.canonicalUrl — the published page for these exact inputs, or the lane hub when they are off the published grid.

100.0/100

1 trials · measured 27 days ago

llmVramFit scores 100.0/100 on Vouch's measured behaviour index, from 1 real invocation trials against ai.makerportal/compute, measured 11 Sept 2026 under methodology v0.2.0. Every measured component scored 100.

Component breakdown

ComponentWeightValue
Reliability35%not applicable
Schema integrity25%100.0
Failure behaviour15%not applicable
Latency15%not applicable
Concurrency10%not applicable

Tool details

Transport
remote + stdio
Credential class
self-provisionable
Input schema
not declared
Output schema
not declared
Side-effect classification
unclassified

Score history

DayScoreTierMethodology
2026-09-11100.0shallowv0.2.0

Probe evidence

ProbeOutcomes
schema_integritypass: 1

Raw request/response logs are not archived yet — the outcome counts above are drawn directly from every recorded trial.

Embed this score

Available for every tool, scored or not — not a verification perk. Always links back to this page.

Vouch score: llmVramFit
[![Vouch score](https://vouch.tools/api/tools/5022bd84-2c72-4977-8a4f-9f0b4a7ad89e/badge.svg)](https://vouch.tools/tools/5022bd84-2c72-4977-8a4f-9f0b4a7ad89e)
llmVramFit — Vouch