RAI Daily · Preview · full text pending fact-check
AISI publishes model-by-model cyber capability evaluation: GPT-5.5 (71.4%) edges Anthropic Mythos (68.6%)
UK AISI's first model-specific public cyber evaluation puts two different labs' flagship models within 3 points of each other at expert-level offensive cyber, and confirms that "frontier cyber capability" is no longer a one-lab phenomenon.
Focus areas
Why you can’t read the full briefing yet
This briefing is still being fact-checked.
Every daily briefing is researched on the day, then independently checked against its sources before the full text goes public. This one has not cleared that check yet, so for now you see the headline, the summary, and its focus areas. The full edition will appear at this same address once verification completes.
In the meantime
- Read the latest verified briefing: Monitor ensembles: skill governs, decorrelation does not
- Browse verified coverage of Evaluation & assurance
- Subscribe to the RSS feed and the full edition reaches you when it clears.