1 link tagged with all of: enterprise-ai + computer-use-agents + gpt-6-astra
Click any tag below to further narrow down your results
Links
OpenAI released GPT-6 Astra, a model designed to operate software like a human would—clicking, typing, navigating across apps—rather than requiring custom API integrations. The company claims this marks the arrival of AGI, though the benchmark comparisons are murkier than the headlines suggest.
- Astra can autonomously complete multistep workflows across browsers, spreadsheets, and desktop apps without developers building separate integrations for each tool, potentially reshaping how enterprises deploy AI.
- OpenAI reports Astra scored 98.6% on ARC-AGI-3, but this number is misleading: NVIDIA achieved 100% on the same benchmark using Claude Opus 5 with added memory and tool architecture, showing that high scores come from the complete agent system, not just the foundation model.
- The core distinction matters for AGI claims—what's actually being measured: the neural network weights alone, or the model plus memory, tools, and orchestration? OpenAI sidesteps this by arguing enterprises care about outcomes, not benchmark purity.
- Astra was trained at unprecedented scale (over 100,000 DBUs) and represents OpenAI's largest capability jump yet, with strong performance across math, coding, and reasoning benchmarks, though the company notably didn't release GDPval results measuring real-world economic work.