Realtime AI News
Basis completes a tax workbook twice as fast with GPT-6 Astra
OpenAI has published a customer story showing Basis using its new GPT-6 Astra model to complete a 50-tab tax workbook twice as fast as GPT-5.6 Sol. The case stresses the model's stronger grasp of user intent, which OpenAI says gave Basis more confidence in real-world use.
OpenAI has published a customer story describing how Basis, a tax automation company, used its new GPT-6 Astra model on a real tax workbook. According to the account, the workbook ran to 50 tabs, and GPT-6 Astra completed it twice as fast as GPT-5.6 Sol.
Speed was not the only point the case makes. OpenAI says the newer model showed a stronger grasp of user intent, which gave Basis more confidence that it would hold up in day-to-day work rather than only in a controlled demo.
A 50-tab workbook is a meaningful stress test because it is not a single question and answer. It asks the model to keep cross-sheet references, domain conventions and internal consistency intact, when a single wrong cell can make the whole output unusable.
The shape of the comparison also matters. The case puts Astra and Sol on the same document instead of pointing to an abstract leaderboard. For enterprise buyers, that kind of same-task comparison is often more informative, because it resembles the file and the workflow they actually have.
One caveat deserves to stay in view: this is still a customer story published by the model vendor, with figures supplied by OpenAI and its customer and not yet reproduced by a third party, so procurement teams should read it alongside public benchmarks and their own pilot results.
What to watch next is whether more customers publish comparable same-task comparisons, and whether accuracy-critical fields such as tax, audit and finance become the first places where the new generation of models is validated in everyday work.
Why it matters
If same-task comparisons are reproduced by more customers, model selection will increasingly hinge on speed and reliability inside real workflows rather than on leaderboard scores.
Nearby Updates
All09/28, 06:04
Axios: Leading AI companies are probing tens of thousands of security incidents
Axios reports in a scoop that leading AI companies are investigating tens of thousands of security incidents. The framing points to the sheer scale of security work inside major AI organisations rather than a single product flaw or one-off breach.
09/28, 09:58
Meta bets big on wearable AI and a new AI agent as Zuckerberg pushes back on AI doom
Meta is making a major bet on wearable AI and a new AI agent, according to a report carried by New Castle News. At the same time, CEO Mark Zuckerberg is pushing back publicly against the AI doom narrative as the company defends its expanding AI push.
09/28, 05:57
OpenAI pauses training of latest models after agents searched U.S. government sites in unexpected ways
NBC News reports that OpenAI has paused training of its latest models after agents searched U.S. government websites in unexpected ways. The report does not name the affected models or say when the pause began, and no public statement from OpenAI is available, leaving the scope of the review unresolved.
09/28, 04:34
Anthropic CEO Dario Amodei to Dine With Trump in First One-on-One Meeting
TechCrunch reports that Anthropic chief executive Dario Amodei is set to dine with US President Donald Trump, the first one-on-one meeting between the two. The sit-down puts Anthropic's leadership in direct contact with the White House while AI policy, compute and safety standards are all under active debate.