2026-10-11 16:38 UTC

Vosti's deterministic LLM inference specification and verification approach gains adoption in agent harnesses for reproducible execution.

state: seedheat: lowuncertainty: mediumconvergesscott: mediumdeterministic-inference llm-verification reproducible-execution agent-harnessesVosti authorsmatt_d

What is this?

Vosti is a new arXiv paper (2609.38981, Oct 1 2026) and research prototype from QDelta that formalizes a system-level specification for deterministic LLM inference โ€” requiring bitwise-identical logits across batching, chunking, prefill/decode splits, and KV-cache reuse. The engine is implemented in Rust, verified with Verus, and uses annotated Triton kernels checked by a relational kernel verifier. The web results show the paper and GitHub repo but no evidence yet of adoption in agent harnesses; the hypothesis projects forward from the paper's relevance to reproducible agent execution.

Why it matters to Scott

Formal-methods researchers are independently building what Scott's Agent Provenance Stack assumes is missing: a machine-checked deterministic execution layer โ€” Vosti's Verus-verified Rust engine with bitwise-identical logits across batching/chunking/KV-reuse would turn his Execution Attestation receipts from records into replayable proofs. Dated-receipts angle if it pans out, but relevance is capped at medium because the supplied evidence shows only a paper and prototype from an unfamiliar group (QDelta) with zero adoption in agent harnesses yet โ€” the hypothesis's adoption claim is projection, not evidence.
ip:concept.execution-attestationip:framework.agent-provenance-stackdev:concept.deterministic-agent-control-planedev:concept.hardware-aware-local-inferenceradar:concept.formal-verificationradar:mi300x-h100-byte-identical-inferenceradar:contract-verifier-llm-gpu-kernelsradar:concept.llm-serving
queries asked of Scott's wikis
  • deterministic inference specification verification agent harness
  • reproducible execution agent systems kernel verification
  • Verus Rust verification LLM serving Triton kernels
  • bitwise identical logits batching chunking KV cache agent evaluation
  • formal specification LLM inference harness reliability

Measured heat

now 0 pts/hpeak 1 pts/hcomments 0/hpeers p16momentum: steady2 platformsage 67h
points/hour across evidence ยท reading as of 2026-10-12 02:59:37.977291+11:00 ยท deterministic, not a model opinion

How the heat travelled

10-09 00:53 (minted)โญ origin echo-reconstructedFramework for specifying and verifying deterministic LLM inference.
Vosti authors on paper (echo) ยท attributed from hn.story.50012359 ยท published time unknown
โ€”
10-08 21:11first on hacker news ยท published ยท lag ?Vosti: Specifying, Implementing, and Verifying Deterministic LLM Inference
matt_d
โ€”
10-08 21:11amplified on hacker news ๐Ÿ‘‘hn.story.50012359
matt_d
peak 1 ยท 0 comments ยท 106% of case engagement
10-08 22:33our radar first saw it ยท lag ?discovery anchor: hn.story.50012359โ€”
pace: p10 vs 1204 stories at the 48h mark (now 67h old) โ€” behind 3jsbench-llm-3d-generation-benchmark (0.5x)

Evidence (2) โ€” โญ canonical anchor

sourceobjectauthorscorecomments
๐ŸŸง hnVosti: Specifying, Implementing, and Verifying Deterministic LLM Inferencematt_d10
๐ŸŸง echo.paper โญFramework for specifying and verifying deterministic LLM inference.Vosti authorsโ€”โ€”

Interpretation history

Decision trace