Microsoft announced the general availability of GPT-6 Astra within Microsoft Foundry on Azure. The frontier model is engineered to transition enterprise generative AI from interactive chat interfaces toward autonomous agentic workflows, providing multi-step planning, deliberate decision support, and cross-application tool execution. OpenAI calls Astra its most aligned model to date, and Microsoft says it is designed for token efficiency on complex work.
The release focuses on solving deployment friction around operational fundamentals such as identity management, secure networking, compliance, and governance. By integrating GPT-6 Astra directly into Microsoft Foundry, enterprise IT teams can configure and deploy agentic pipelines while maintaining enterprise boundary controls across their cloud infrastructure.
Multi-Step Reasoning and Application Execution
GPT-6 Astra is designed to handle open-ended objectives by evaluating options, building structured plans, weighing execution trade-offs, and dynamically incorporating new parameters as tasks proceed. The model generates structured artifacts, including technical reports, analysis spreadsheets, and slide decks, that align with pre-existing enterprise templates and business standards for human verification.
A central architectural feature is the model’s computer-use capability, which enables it to operate software interfaces on a user’s behalf. Astra can parse on-screen visual information and manipulate approved user interfaces, allowing it to navigate legacy applications and tools that lack dedicated REST APIs. Key deployment scenarios include:
Software engineering teams can use the model to reproduce reported defects, trace root causes across code repositories, generate proposed fixes, and prepare pull requests for developer review. For business intelligence, Astra automates the construction and iterative refinement of Power BI dashboards. In broader operations, the engine supports automated record updates, form processing, and end-to-end interface testing across enterprise software suites.
Governance, Security, and Containment Controls
Because computer use and autonomous agency present unique security challenges, Microsoft Foundry integrates containment safeguards around Astra. Workflows can be constructed with scoped credentials, role-based access control (RBAC), and required human-approval checkpoints before executing high-impact actions.
Foundry wraps the model with Microsoft Entra identity and access management, pervasive encryption in transit and at rest, private networking routing, automated content filtering, and comprehensive activity logging. Microsoft states that prompts and outputs within Foundry are not used to train the models.
Industry partners evaluating the platform highlighted the balance of development velocity and operational discipline. Replit CTO Luis Hector Chavez said Astra moves AI tooling “beyond code generation to active software creation.” Anirban Nandi, VP of Data and AI at Albertsons Companies, said Azure OpenAI on Microsoft Foundry gives his teams “that balance of speed and control,” with consistent security, governance, and operational controls as they scale.
Deployment Options and Availability
| Deployment | Context Length | Input | Cached Input | Cached Writes | Output |
|---|---|---|---|---|---|
| Standard Global (USD $/million tokens) | |||||
| Standard Global | Short context | $10.00 | $1.00 | $12.50 | $50.00 |
| Standard Global | Long context | $20.00 | $2.00 | $25.00 | $75.00 |
| Standard Data Zone (US) (USD $/million tokens) | |||||
| Standard Data Zone (US) | Short context | $11.00 | $1.10 | $13.75 | $55.00 |
| Standard Data Zone (US) | Long context | $22.00 | $2.20 | $27.50 | $82.50 |
| Provisioned Throughput pricing varies by deployment type. U.S. Data Zone Provisioned Throughput is priced at a 10% premium to Global Provisioned Throughput. For current rates and terms, see the Azure OpenAI pricing page. | |||||
GPT-6 Astra is available immediately across Global and U.S. Data Zone geographies with two distinct infrastructure deployment tiers. The Standard deployment tier provides consumption-based, pay-as-you-go provisioning suited for variable enterprise demand. For workloads requiring deterministic latency baselines and dedicated compute capacity, the Provisioned Throughput tier offers guaranteed model-processing throughput. Microsoft has been expanding Azure’s inference capacity with its own Maia 200 accelerators alongside NVIDIA and AMD hardware. In terms of infrastructure pricing, U.S. Data Zone Provisioned Throughput carries a 10 percent premium over standard Global Provisioned Throughput rates.




Amazon