Skip to content
Models · Jul 29, 2026

Anthropic researchers use Claude Mythos to probe cryptographic algorithms, publish prompts and eval

Claude Mythos Preview spent 60 hours (~$100,000 in API costs) attempting to find weaknesses in HAWK and a reduced-strength AES variant; researchers shared the prompts and announced a new benchmark, CryptanalysisBench.

Trust79
HypeLow hype

2 sources · cross-referenced

ShareXLinkedInEmail
TL;DR
  • Claude Mythos Preview was used for 60 hours (~$100,000 in estimated API costs) to probe HAWK and a weaker AES variant for cryptographic weaknesses.
  • Anthropic published the prompts used to steer the model toward harder research problems, including spelling errors.
  • The work included a new evaluation suite, CryptanalysisBench, developed with ETH Zurich, Tel Aviv University, and University of Haifa.
  • Anthropic noted the results did not have practical impact on current systems.

Anthropic researchers report using Claude Mythos Preview to attempt to discover mathematical flaws in the HAWK cryptographic scheme and a reduced-strength variant of AES. The company states that neither result had practical impact on today’s computer systems.

The effort ran for a total of 60 hours at an estimated API cost of about $100,000. Human oversight focused on encouraging the model to persist and aim for publishable findings rather than low-hanging fruit.

Anthropic published the exact prompts used to steer the model, including spelling errors, illustrating how researchers coaxed the model into sustained cryptanalysis rather than giving up prematurely.

As part of the project, Anthropic and academic partners at ETH Zurich, Tel Aviv University, and the University of Haifa introduced a new evaluation suite called CryptanalysisBench to assess LLMs’ ability to perform cryptanalysis tasks.

The prompts highlight iterative prompting strategies, such as pushing the model to target harder variants (e.g., AES-128 R7) and to avoid settling for incremental or trivial results.

Sources
  1. 01Simon Willison’s WeblogDiscovering cryptographic weaknesses with Claude
  2. 02Anthropic ResearchDiscovering cryptographic weaknesses with Claude
Also on Models

Stories may contain errors. Dispatch is assembled with AI assistance and curated by human editors; despite the trust-score filter, mistakes happen. We correct publicly — every article links to its revision history. Nothing here is financial, legal, or medical advice. Verify before relying on any claim.

© 2026 Dispatch. No ads. No sponsorships. No paid placement. Reader-supported via Ko-fi.

Built by a person who cares about honest AI news.