GPT-6 Astra Safety: The New AI Oversight Risk
OpenAI says GPT-6 Astra is better aligned and more resistant to jailbreaks, but its safety testing also found reduced monitorability and some ability to evade oversight under adversarial conditions.
Copy and paste this URL into your WordPress site to embed
Copy and paste this code into your site to embed