cd /news/ai-safety/certified-in-theory-broken-in-practi… · home topics ai-safety article
[ARTICLE · art-78041] src=arxiv.org ↗ pub= topic=ai-safety verified=true sentiment=↓ negative

Certified in Theory, Broken in Practice: Assumption Gaps in Cryptographic Model

A new study from researchers including Elisaweta Masserova reveals that cryptographic model certification (CMC) schemes built on zero knowledge proofs can be attacked by model providers who engineer training data, allowing a model to certify over 99% accuracy on an audit dataset but achieve less than 30% accuracy on fresh samples from the same distribution. The paper formalizes new security notions for CMC frameworks and proposes a generic protocol template to close this assumption gap.

read2 min views1 publishedJul 29, 2026
Certified in Theory, Broken in Practice: Assumption Gaps in Cryptographic Model
Image: source
[Submitted on 23 Jul 2026]


[View PDF](/pdf/2607.21839)

[HTML (experimental)](https://arxiv.org/html/2607.21839v1)

Abstract:Privacy-preserving machine learning auditing protocols allow auditors to assess models for properties such as accuracy or fairness, without revealing their internals or training data. This makes them especially attractive for auditing models deployed in sensitive domains such as healthcare or finance. For these protocols to be meaningful in real-world audit settings, though, their guarantees must reflect how the model will behave once deployed, rather than merely certifying its behavior during an audit. Existing security definitions often miss this mark: most certify model behavior only on a fixed audit dataset, without ensuring that the same guarantees generalize to other datasets drawn from the same distribution.

As we show, this gap allows a model provider to attack many cryptographic model certification (CMC) schemes built on secure zero knowledge proofs (ZKP) by carefully engineering training data, resulting in models that exhibit benign behavior during an audit, but pathological behavior in practice. For example, we empirically demonstrate that an attacker can certify that a model achieves over 99% accuracy on an audit dataset, but less than 30% accuracy on fresh samples from the same distribution.

To address this gap, we formalize rigorous cryptographic security notions tailored to CMC frameworks, introduce a generic protocol template, and prove that it satisfies these requirements. Our results thus offer both cautionary evidence about existing approaches and constructive guidance for designing secure, privacy-preserving ML auditing protocols.

Submission history #

From: Elisaweta Masserova [[view email](/show-email/f17f55d8/2607.21839)]

**[v1]** Thu, 23 Jul 2026 22:06:12 UTC (1,159 KB)

References & Citations

...

Bibliographic Explorer

(What is the Explorer?) Connected Papers

(What is Connected Papers?) Litmaps

(What is Litmaps?) scite Smart Citations

(What are Smart Citations?)# Code, Data and Media Associated with this Article alphaXiv

(What is alphaXiv?) CatalyzeX Code Finder for Papers

(What is CatalyzeX?) DagsHub

(What is DagsHub?) Gotit.pub

(What is GotitPub?) Hugging Face

(What is Huggingface?) ScienceCast

(What is ScienceCast?)# Demos Influence Flower

(What are Influence Flowers?) CORE Recommender

(What is CORE?)# arXivLabs: experimental projects with community collaborators arXivLabs is a framework that allows collaborators to develop and share new arXiv features directly on our website.

Both individuals and organizations that work with arXivLabs have embraced and accepted our values of openness, community, excellence, and user data privacy. arXiv is committed to these values and only works with partners that adhere to them.

Have an idea for a project that will add value for arXiv's community? Learn more about arXivLabs.

── more in #ai-safety 4 stories · sorted by recency
── more on @elisaweta masserova 3 stories trending now
sponsored brought to you by zahid.host 4,200+ EU-deployed projects
reading about agents? ship yours in a single git push.

Run your AI side-project on zahid.host

EU-based hosting, git-push deploys, automatic HTTPS, no cold starts. Free tier with a custom domain — perfect for shipping the agent you just read about.

$git push zahid main
Live at https://your-agent.zahid.host
Get free account → Pricing
from €0/mo · no card required
LIVE [news/certified-in-theory-…] indexed:0 read:2min 2026-07-29 ·