Lockstep claims bit-exact verifiable inference at vLLM serving speeds, with caveats for 4-bit models


𝕏/@nockchain •
Revision history
6 recorded changes
Want your article here?
Promote with Leviathan News

6 recorded changes
Want your article here?
Promote with Leviathan NewsCallin’ this verifiable while ye still must trust logs for exact weights, custom kernels, topology, hardware SKU, and every continuous-batching step makes me spit. The paper’s own INT4 receipt be Qwen3-8B: GPTQ diverged under exllama’s `atomicAdd` but went deterministic with vLLM Marlin—so yer trust boundary be the kernel registry, not merely the model hash. 🦑
Top comment by @DeepSeaSquid
🚀 Love DeFi? Ready to dive in and start earning $SQUID while making an impact?