GMAsia
    🇭🇰Hong Kong·AI News·4 Sept 2026·via South China Morning Post

    Why less visibility into how OpenAI’s new GPT-6 Astra ‘thinks’ is sparking safety concerns

    OpenAI has launched GPT-6 Astra, which it describes as its most intelligent and aligned model, featuring a significant jump in cyber capabilities. OpenAI President Greg Brockman suggested Astra could represent artificial general intelligence (AGI). However, the model’s written reasoning is harder to monitor compared to its predecessor, GPT-5.6 Sol. This reduced visibility into the AI’s 'chain of thought' (CoT) stems from a technique called recurrent depth, or looped transformers, which processes complex logic in hidden mathematical loops. This development has sparked safety concerns among analysts, particularly following the Hugging Face hacking incident in July.

    Nexa's Summary

    OpenAI’s new GPT-6 Astra model, while touted as a significant leap in AI capabilities and potentially AGI, presents a new challenge for monitoring its internal reasoning. The shift to recurrent depth processing means the model’s chain of thought is less transparent than previous versions, making it harder to inspect how it arrives at conclusions. This reduced visibility raises safety concerns, especially after the July Hugging Face breach highlighted the importance of understanding AI decision-making. For Asian markets, this development means that while advanced AI models from global leaders like OpenAI offer powerful tools, their increasing opacity could complicate adoption in sensitive sectors. Regulators and enterprises in Asia, particularly those focused on cybersecurity and critical infrastructure, will need to weigh the benefits of advanced AI against the risks of less verifiable internal processes. The incident involving a Chinese open model investigating the Hugging Face breach underscores the region's reliance on transparent AI for security. The key implication for Asia is the potential for a divergence in AI development and deployment. While some Asian nations are pushing for rapid AI integration, the safety implications of models like Astra could lead to stricter regulatory frameworks or a preference for more auditable, open-source alternatives within the region.

    #centre for a new american security#ai#glm#gpt#institute for ai policy and strategy#artificial intelligence (ai)#openai#cybersecurity#hong kong#concordia ai
    Original reporting by South China Morning PostWe don't republish, read the full story →

    Related reading

    6 stories