OpenAI calls GPT-6 Astra the “world's best computer use model”; in tests, it booked DMV appointments and searched job listings faster than the average person
And I bet they would never share the likely multiple prompts they used in order for these things to happen. [embedded post]Forums:r/ChatGPT:GPT-6 Astra Is Here—and OpenAI Thinks It May Kick Off the AGI Era
Context & Ripple Effects
OpenAI’s claim builds on its 2025 ChatGPT Agent rollout, which introduced a model designed to control a computer through multi-step tasks. Astra’s initial availability to Daybreak customers turns that capability into a higher-stakes performance claim, backed by OpenAI’s largest training run to date at its Texas Stargate site.
The company is also extending a progression from GPT-5’s routed mix of efficient and reasoning models toward an agent judged on completing browser-based work. Public criticism focused on the lack of disclosed evidence and prompt details behind the cited tests, making evaluation methodology central to the claim’s credibility.
First-order effects
- Daybreak customers gain access to OpenAI’s Astra model for evaluating computer-use tasks such as navigating appointment and job-listing workflows.
- OpenAI’s claim of above-average task speed raises the importance of reproducible test conditions for enterprises deciding whether Astra can act on their behalf.
Second-order effects
- Anthropic and other developers of computer-use agents face pressure to publish comparable task-completion evidence rather than rely on broad model-quality claims.
- Services whose workflows depend on websites and forms will face greater demand from customers for interfaces that work reliably with agents, not only human users.
Third-order effects
- If task-completion benchmarks become credible and standardized, agent competition will shift from producing answers to reliably executing multi-step work across third-party software.
- The value in AI products would move toward the systems that can safely coordinate access, permissions, and verification across those workflows, not merely generate text.
The trend: AI assistants are evolving into agentic work surfaces, with competitive differentiation increasingly tied to measurable completion of computer-based tasks.