Realtime AI News
Cognition Uses GPT-6 Astra to Let Devin Test Its Own Work
According to OpenAI, Cognition is using GPT-6 Astra to improve how its AI software engineer Devin tests code and shows that it works. The stated goal is for engineers to review less code and ship more, which puts verifiable evidence, not just generated code, at the center of AI-assisted development.
Information published by OpenAI says Cognition is using GPT-6 Astra to strengthen the self-testing ability of Devin, its AI software engineer. The central claim is that Astra improves Devin's ability to test software and show that it works.
The stated goal is straightforward: engineers review less code and ship more. In practice that means AI-written changes are not simply queued for line-by-line human inspection; they are expected to run, pass, and arrive with evidence first.
The underlying engineering problem is volume. The more code a model generates, the heavier the review burden becomes, and the bottleneck moves from writing to reviewing. Evidence produced by the model itself is one of the few ways to relieve that pressure.
Devin is positioned as an agent that takes on development tasks, and testing is one of the hardest steps for any such agent. Changing code is not the same as confirming the change is correct, so adding self-verification is a key step toward closing the loop.
The information again comes from OpenAI's own channels and describes Astra's effect inside a specific product, without quantitative figures on test coverage or defect rates. It is best read as a public signal about a direction of capability rather than a measured benchmark.
Two things to watch: whether self-testing holds up in large, long-lived codebases, and how review policies and quality gates need to change once a model can both write code and vouch for it.
Why it matters
Self-testing targets the real bottleneck in AI coding: generation is fast, human review is not. A model that brings its own verification evidence is what allows agents to enter production workflows rather than stop at suggestions.
Nearby Updates
All09/12, 00:46
Nscale adds former OpenAI exec Fidji Simo to its board ahead of potential IPO
TechCrunch reported on September 11 that Nscale has added former OpenAI executive Fidji Simo to its board as it prepares for a potential IPO. Simo was the No. 2 executive at OpenAI and previously led Instacart through its 2023 IPO.
09/11, 22:05
Anthropic Posted a $450K Sales Role Devoted Solely to Meta — Then Closed It in a Week
Anthropic briefly listed a job titled Mega Account Executive, Meta, a role serving only Meta, with on-target earnings of $380,000 to $450,000, or roughly 3.2 million yuan at current rates. The posting went live on September 3 and stopped accepting applications on September 9, as Meta engineers already lean heavily on Claude Code.
09/12, 02:02
MYbank opens Bailing 2.0, a small-business finance Agent, to 42 million merchants
At the 2026 Bund Conference, MYbank disclosed its AI banking rollout for the first time: Bailing 2.0, described as the world’s first inclusive-finance Agent for small and micro businesses, is now open to 42 million merchants. Behind it, eight AI workbenches have moved into risk control, manual review, marketing and R&D, with AI coding doubling and 15% of business change requests completed by AI.
09/12, 02:41
Anthropic Researcher Resigns, Warning the Company Is 'Racing Straight to Self-Improving Superintelligence'
An Anthropic researcher resigned this week and warned on X that the company is “racing straight to self-improving superintelligence and gambling with our lives.” Anthropic's own alignment lead co-signed the message instead of walking it back, and with the company reportedly preparing for an IPO, the warning lands at a sensitive moment.