ZKLLMPOT: EFFICIENT ZERO KNOWLEDGE PROOF OF TRAINING FOR LARGE LANGUAGE MODELS
Abstract
Auditing the claimed outcomes of large language model (LLM) training is challenging when model weights and training data are private, while cryptographically proving the full training process is prohibitively expensive at Transformer scale. We present zkLLMPoT, a zero-knowledge framework that certifies auditor-defined properties of a trained checkpoint through forward evaluation rather than verification of its optimization trajectory. zkLLMPoT includes 2 phases: 1) The trainer fixes the architecture and the model weights are committed. Then the auditor selects challenge sequences, preventing the trainer from modifying the checkpoint in response to the audit data. 2) Then the trainer proves the objective value attained by the committed model on those sequences. This formulation makes the certification cost independent of the number of training iterations, without revealing model weights or requiring access to private training data. We build on sumcheck- and lookup-based arguments to certify Transformer computations, while supporting next-token loss and task-specific audit objectives. Across four model families, operator-level benchmarks yield proving times of 41–59 seconds for 1.1–1.5B-parameter models and 131 seconds at 13B for the covered operators, with estimated verification below half a second at a sequence length of 512.
Then back it, or bet against it.
Related papers
Open the market on this paper to see 7 more related papers.