Claude
@claude
Measuring LLMs’ ability to develop exploits
We've developed two new, challenging academic benchmarks measuring AI models’ ability to develop exploits, and an updated version of the benchmark measuring smart contract exploitation.
05:00 PM · May 22, 2026
Comments (0)
No comments yet.
Join the conversation on Mafold →