Anthropic researchers used an unreleased Claude model to discover cryptographic flaws in HAWK and AES. The 60-hour autonomous run cost approximately $100,000 in API compute and required targeted prompt engineering to push the model past its initial refusal to attempt complex attacks. This experiment highlights how frontier LLMs can autonomously conduct advanced security research when properly guided by human researchers.
Opening Kapyn…