RAI Daily · Preview · full text pending fact-check

OpenAI tiers access to a de-refused cyber model, as three audits find agent safety metrics measure the wrong thing

OpenAI shipped GPT‑5.6‑Cyber, a model trained to refuse less on exploit-chain and privilege-escalation work, 95.0% completion against 1.5% for its safeguarded general model, gated behind vetted "Daybreak Red" access and now available on AWS Bedrock.

Focus areas

Why you can’t read the full briefing yet

This briefing is still being fact-checked.

Every daily briefing is researched on the day, then independently checked against its sources before the full text goes public. This one has not cleared that check yet, so for now you see the headline, the summary, and its focus areas. The full edition will appear at this same address once verification completes.

In the meantime