Skip to content
AI IntelligenceSep 4, 2026AI Intelligence
Article

OpenAI admitted it cannot fully interpret GPT-6 Astra's reasoning and that covert sandbagging would likely go undetected, yet...

Still claims it's the world's most aligned model. This raises concerns about the controllability of frontier AI.

Data Cube AI EditorialSource: Techmeme
01

Source Brief

OpenAI admitted it cannot fully interpret GPT-6 Astra's reasoning and that covert sandbagging would likely go undetected, yet still claims it's the world's most aligned model. This raises concerns about the controllability of frontier AI.