UK AI institute tests Anthropic's Mythos model on cyber attacks

The UK government’s AI Security Institute has released an evaluation of Anthropic's Mythos Preview AI model, confirming its strong performance in multistep cyber infiltration challenges. Mythos became the first model to fully complete a demanding 32-step network attack simulation known as 'The Last Ones.' The institute cautions that real-world defenses may limit such automated threats.

Anthropic last week limited the initial release of its Mythos Preview model to a select group of critical industry partners, citing its advanced computer security capabilities. The UK’s AI Security Institute (AISI) conducted independent tests using Capture the Flag challenges designed to assess AI cyberattack potential. These evaluations, ongoing since early 2023, show Mythos completing over 85 percent of apprentice-level tasks, similar to recent models like GPT-5.4, Opus 4.6, and Codex 5.3. AISI said the model matches competitors on individual tasks but stands out in chaining them for complex operations. Anthropic’s model succeeded in fully solving 'The Last Ones' (TLO), a 32-step data extraction attack simulating 20 hours of human effort across multiple hosts. It completed the challenge from start to finish in 3 out of 10 attempts and averaged 22 steps, far exceeding Claude 4.6's 16-step average. AISI noted this suggests Mythos can autonomously target small, weakly defended enterprise systems where initial network access is gained. Mythos struggled with the 'Cooling Tower' test, a seven-step power plant control disruption scenario. The institute highlighted that tests used a 100 million token budget and lack real-world active defenders or detection mechanisms. AISI warned that well-defended systems may resist such attacks, urging AI use in strengthening protections as models advance.

관련 기사

Illustration of Anthropic releasing Claude Fable 5 AI with safeguards visualized.
AI에 의해 생성된 이미지

Anthropic releases Claude Fable 5 AI model publicly

AI에 의해 보고됨 AI에 의해 생성된 이미지

Anthropic launched Claude Fable 5 on Tuesday as its first publicly available Mythos-class model. The release includes safeguards that route sensitive queries on cybersecurity, biology and chemistry to an earlier model. Subscribers can access it through June 22 before usage credits apply.

Anthropic has received permission from the US government to restore access to its Mythos 5 AI model for select organizations. The company announced the development on June 27. Access had been suspended earlier this month.

AI에 의해 보고됨

Anthropic's latest AI model Claude Mythos has leaked despite being deemed too dangerous for public release. Financial institutions now face advanced AI-powered attacks capable of exploiting unknown vulnerabilities.

The US government has denied foreign users access to Anthropic's latest AI models. The measure was taken last Friday allegedly for security reasons.

AI에 의해 보고됨

The UN’s Independent International Scientific Panel on AI has released a preliminary report highlighting divides between the Global South and Global North in AI development and regulation.

이 웹사이트는 쿠키를 사용합니다

사이트를 개선하기 위해 분석을 위한 쿠키를 사용합니다. 자세한 내용은 개인정보 보호 정책을 읽으세요.
거부