/inference/hf.co-bartowski-sarvam-1-gguf-q4_k_m

hf.co/bartowski/sarvam-1-GGUF:Q4_K_M

Measured by one node on hardware we do not own, and signed by each of them. What machines here actually proved — never what the card claims.

matrix built 13 Sept 2026 · every field verified against the signature of the node that produced it, and re-verifiable by you: each signed figure carries the exact bytes the node signed (signing_payload_b64) and the signature over them, and the ed25519 public key is embedded in node_did. effective_ctx is the context length a node PROVED by recall, never the length a model card advertises
context declarednot declaredwhat the runtime reported. A node attests it was told this; nobody attests it is true.
context proved by recallnot probedthe largest size at which a machine could still find tokens planted across the whole prompt — start, middle and end, all of which had to come back.
Both numbers are needed to state a gap, and one of them is missing.
signed by this network

What machines here proved

best throughput13.47 tok/s
machines that measured it1
samples behind those numbers1
regions it was proved in1
Every line traces to a node signature.
from the public record · unsigned

What the runtime says about it

No runtime metadata was attested for this model.The only unsigned block on this page. We attest we read these from the runtime — not that they are true. Nobody signs a model card.

Proven across regions

The machine that measured it

nothing is averaged — the machine is what you are choosing
did:epn:002408…9e3c79Bengaluru · 562130 · 13 Sept 2026x86_64 · intel · 2.0 GiB pooled
ctx proved
13.47tok/s1 samples · 0s
no capability probed