OpenAI · Status at source date: Newsletter archive
OpenAI releases EVMbench for smart contract security
OpenAI has introduced EVMbench, a benchmark measuring how well AI agents can detect, exploit, and patch high-severity smart contract vulnerabilities. The benchmark tests three capabilities: vulnerability detection in real-world contract code, exploitation in realistic attack scenarios, and safe patching with fixes that hold up under testing.

What changed
OpenAI frames EVMbench as both a measurement tool and a call to action, arguing that as AI agents improve, developers and security researchers should incorporate AI-assisted auditing into their workflows. The release signals OpenAI's interest in AI applications for security: a domain where the same capabilities that enable helpful automation also create potential risks.
Sources & contributor credit
- Newsletter coverage · Paid Media Collective newsletter
Original newsletter text, contributor labels and media for this update.
- Source referenced in newsletterLinkedIn
Linked from the original newsletter. The source publication date has not been independently confirmed.
- Source referenced in newsletterDowload PDF
Linked from the original newsletter. The source publication date has not been independently confirmed.
Original creator unverified
The original creator has not yet been verified. Newsletter curation and publication do not establish original authorship.
Attribution evidence and limitations
Full attribution review pending.
A source link alone does not establish original authorship.
Original image and video creators have not yet been verified.
- Published on this site
This update reflects the dated source reporting. Availability may have changed. Further coverage of this same development will be added to this page.
Original newsletter text and archive evidence
OpenAI releases EVMbench for smart contract security
OpenAI has introduced EVMbench, a benchmark measuring how well AI agents can detect, exploit, and patch high-severity smart contract vulnerabilities. The benchmark tests three capabilities: vulnerability detection in real-world contract code, exploitation in realistic attack scenarios, and safe patching with fixes that hold up under testing.
OpenAI frames EVMbench as both a measurement tool and a call to action, arguing that as AI agents improve, developers and security researchers should incorporate AI-assisted auditing into their workflows. The release signals OpenAI's interest in AI applications for security: a domain where the same capabilities that enable helpful automation also create potential risks.
Source captured . No explicit first-contributor label was provided for this update.


