Vietnamese crab exporterdouble-skinned crabs
Advertisement
Artificial intelligence
TechTech Trends

Why less visibility into how OpenAI’s new GPT-6 Astra ‘thinks’ is sparking safety concerns

The Hugging Face breach in July highlights the importance of being able to inspect what models are ‘thinking’, say analysts

3-MIN READ3-MIN
2
Listen
OpenAI said the written reasoning of its new Astra model was 'harder to monitor' compared with the previous generation released in July. Photo: Reuters
Chong Ming Lee
OpenAI’s new model, GPT-6 Astra, has less direct visibility into how a model thinks, a development that has sparked concerns coming just weeks after the Hugging Face hacking incident that required a Chinese open model to investigate, according to analysts.

When announcing Astra on Thursday, OpenAI said it was “the world’s most intelligent and aligned model,” with a “significant jump in cyber capabilities”.

OpenAI president Greg Brockman said at the end of a press call announcing Astra’s arrival that it likely represents AGI, or artificial general intelligence – AI that matches or outperforms human intelligence.

However, OpenAI also said the model’s written reasoning was “harder to monitor” compared with GPT-5.6 Sol, the previous generation released in July.

Select Voice
Select Speed
1x
AI-generated voice