The Proof That Doesn't Prove: Why ZK Inference's Single-Digit Milestone Has a Model-Size Caveat
Attestable emerged from stealth this week with a $20M seed from Altimeter and TLV Partners and a claim that sounds impossible: zero-knowledge proofs for LLM inference at single-digit overhead. Their demo proves inference of Meta's Muse Glimmer 30B on a single H100 at 85 tokens per second, generating compact, quantum-resistant proofs verifiable in sub-second time. Vitalik Buterin — who has been evangelising the overhead ratio as the metric that matters — noted this puts ZK proving overhead "approaching single…