Realtime AI News
Kimi K2.8 arrives suddenly: performance close to K3, million-token context open to all
Kimi has pushed out K2.8, a version the source describes as performing close to the higher-tier K3 while opening its million-token context to every user. The release lands as Moonshot AI sprints toward a Hong Kong IPO, and it looks aimed at widening the user and developer base.
Kimi has suddenly released K2.8, according to a report from QbitAI. The new version is described as performing close to the higher-tier K3, while its million-token context is now open to all users rather than a limited set of testers or paying customers.
The numbering places K2.8 between K2 and K3, making it a mid-cycle refresh rather than a generational leap. That cadence is becoming familiar among Chinese model makers: slot a near-next-generation, lower-cost version beneath the flagship and use it to win users and gather feedback from real workloads.
Opening the million-token context to everyone is the most tangible change in this update. A longer context window lets a model read much longer documents, codebases or conversation histories in one pass, which matters most for knowledge-base Q&A, long-document summarization and code understanding.
The timing is worth noting. The report says Moonshot AI is sprinting toward a Hong Kong IPO. At a stage when large-model companies are lining up for listings, iteration speed, user scale and developer ecosystem all feed directly into the capital-markets story, so every release gets read against that backdrop.
Competitively, K2.8 bundles near-flagship quality with general availability, positioning it against the flagship and sub-flagship models other vendors are shipping in the same window. For buyers, that means another option that sits close to the top tier while lowering the barrier to entry.
For developers, the practical question remains migration cost. If K2.8 delivers performance close to K3 with a more generous context window and a lower bar to access, existing applications built on Kimi may only need small adjustments to switch over.
Two things to watch next: when K3 formally arrives and how large the real capability gap between the two turns out to be, and whether the IPO process brings a denser stream of model and product moves.
Why it matters
For Moonshot AI, K2.8 pairs near-flagship quality with an open door, expanding its user and developer base right when the IPO story needs product substance. If K3 slips, K2.8 becomes the de facto vehicle for traffic and developer migration in the meantime.
Nearby Updates
All09/12, 08:09
Perplexity Runs GPT-6 Astra Across Communications, Code, and Production Systems
A case study published by OpenAI says Perplexity uses GPT-6 Astra to write communications, change software, and monitor production systems, checking in far less often than with earlier models. The detail points to frontier models moving past single-step assistance into longer end-to-end tasks where people set goals and validate results.
09/12, 06:58
Mecka AI nears $500M valuation in Sequoia-led round as robot data demand surges
TechCrunch reports that Mecka AI, a two-year-old startup focused on robot training data, is closing in on a $500 million valuation in a new round led by Sequoia Capital. The deal is being assembled only months after the company announced its Series A, underscoring how fast capital is moving into embodied-AI data.
09/12, 04:59
Y Combinator's Garry Tan wants U.S. open-weight AI labs to 'distill' frontier models, too
Y Combinator's Garry Tan has argued publicly that U.S. open-weight AI labs should also be allowed to distill frontier models. His reasoning is that frontier models were themselves trained on public human knowledge, so access to capable AI should be treated as a form of public good.
09/12, 04:57
Twenty-five Fields Medalists sign an open letter warning AI labs are threatening mathematicians' work
Twenty-five leading mathematicians have signed an open letter arguing that AI labs are threatening their intellectual work as they race to solve famous problems. The same week, NYU professor Tristan Buckmaster accused OpenAI of pressuring him over credit for a proof, and OpenAI withdrew its sponsorship of a CalTech math event.