MID selects verifier-facing evidence, such as telemetry or power traces, by minimizing conditional mutual information about a protected property given the authorized verification result, and demonstrates perfect held-out verification with zero measured leakage on three of six tasks.
Bit-Exact AI Inference Verification Without Performance Tradeoffs
1 Pith paper cite this work. Polarity classification is still indexing.
abstract
Verifying claims about AI workloads is a prerequisite for credible AI governance of covert adversaries (who comply with monitoring only when detection likelihood is high), yet the apparent non-determinism of GPU floating-point arithmetic forces auditors to accept approximate output matches. Covert adversaries can exploit unverifiable degrees of freedom in monitored computation. Attack vectors include steganography, unreported modification of inference software, and covert computation via unreported batch elements. Empirically, we analyze how modern inference engines (vLLM, HF transformers) produce deterministic but non-invariant outputs, without needing to set performance-compromising determinism flags, if the right information is available for re-computation and no atomic functions are called in the backend. We demonstrate that such bitwise-precise re-computation does not require access to identical hardware, via a software-only emulation of LLM inference across multiple NVIDIA GPU variants. Thus, accumulated rounding errors can be an auditable signature of the software and hardware setup used for inference, instead of a constraint on verifiability.
citation-role summary
citation-polarity summary
fields
cs.CR 1years
2026 1verdicts
CONDITIONAL 1roles
background 1polarities
unclear 1representative citing papers
citing papers explorer
-
Privacy-Preserving AI Verification via Minimal Information Disclosure
MID selects verifier-facing evidence, such as telemetry or power traces, by minimizing conditional mutual information about a protected property given the authorized verification result, and demonstrates perfect held-out verification with zero measured leakage on three of six tasks.