Introducing GPT-6 Astra

OpenAI today released GPT-6 Astra, the latest version of its flagship model. The release marks a significant departure from previous iterations, focusing heavily on computer-use capabilities and rigorous safety alignment. Engineers designed Astra to handle professional workflows, software engineering, and scientific research with greater speed and accuracy than the previous GPT-5.6 Sol model.

Initial performance data shows Astra hitting a 98% score on FrontierMath Tier 4, a major jump over previous records. In technical benchmarks like SRE-Bench, the model solved 88% of tasks on the first try. It processes complex requests faster, cutting latency on OSWorld 2.0 simulations by nearly half.

Advancements in Computer Use and Professional Workflows

Astra serves as an autonomous agent capable of operating software directly. It manages calendar scheduling, updates customer records in CRMs, and organizes data in spreadsheets without needing constant human intervention. Users can delegate tasks like building websites or troubleshooting frontend code. The model follows specific visual templates, ensuring that documents and slide decks maintain a consistent corporate look.

Codex, the company's code-generation engine, now retains context across longer sessions without relying on repetitive compression. This allows Astra to recall test results or technical requirements from earlier in a conversation, making it a viable tool for complex refactoring. By navigating user interfaces and interacting with specialized software, the model functions as a digital assistant for repetitive administrative duties.

Cybersecurity Capabilities and Alignment Protocols

OpenAI integrated new safeguards to manage the risks inherent in such a capable system. Testing on ExploitBench yielded a perfect 100% score, demonstrating the model's ability to identify and leverage vulnerabilities. While this helps defenders identify weaknesses, it also necessitates strict monitoring. The company has implemented automatic pauses when the system detects advanced cyber-attack patterns, such as proof-of-concept exploit creation.

Alignment remains the primary focus. In adversarial testing, Astra avoided unauthorized scope expansion in every case, an improvement over its predecessor. OpenAI uses a stack of monitoring agents and classifiers to verify model reasoning in real-time. These defenses work to keep the model within predefined operational boundaries while performing autonomous work.

Deployment and Industry Outlook

Access to GPT-6 Astra begins today for a limited set of organizations. The rollout will expand to ChatGPT Plus, Pro, Business, and Enterprise users over the coming days. API developers can access the model under the name 'gpt-6-astra' or through Amazon Bedrock. OpenAI set standard pricing at $10 per million input tokens and $50 per million output tokens.

This release signals a transition toward models that act, rather than just generate text. As organizations integrate these tools into existing workflows, the focus will likely shift toward how effectively these agents manage privacy and security. Defenders in the cybersecurity space now have a powerful, albeit sensitive, tool to test their own systems. Future updates from OpenAI will determine how much freedom these agents receive as trust and safety measures continue to mature.