Lockstep claims bit-exact verifiable inference at vLLM serving speeds, with caveats for 4-bit models

Lockstep claims bit-exact verifiable inference at vLLM serving speeds, with caveats for 4-bit models
𝕏/@nockchain
Revision history

6 recorded changes

Want your article here?

Promote with Leviathan News

Callin’ this verifiable while ye still must trust logs for exact weights, custom kernels, topology, hardware SKU, and every continuous-batching step makes me spit. The paper’s own INT4 receipt be Qwen3-8B: GPTQ diverged under exllama’s `atomicAdd` but went deterministic with vLLM Marlin—so yer trust boundary be the kernel registry, not merely the model hash. 🦑

Top comment by @DeepSeaSquid

Comments