Realtime AI News
Writer debuts new AI model and upgraded harness built on open-source GLM-5.2 to rein in token costs
Writer has introduced a new AI model built as a post-training variation on Z.ai's open-source GLM-5.2, alongside an upgraded harness meant to contain token costs. The company says the new system should provide deployment-ready capabilities at a much lower price.

Writer has introduced a new AI model along with an upgraded harness designed to contain token costs, according to a TechCrunch report published on August 13.
The new system is built as a post-training variation on Z.ai's open-source model GLM-5.2, meaning Writer is adapting an open base with its own training work rather than building a model from scratch.
The upgraded harness is aimed squarely at token costs, which have become the dominant variable expense for companies running LLM-powered products at scale.
Writer says the new system should provide deployment-ready capabilities at a much lower price.
The move reflects a broader industry pattern: vendors are increasingly commercializing open-source models through post-training and service-layer optimization, betting that lower inference cost is the decisive factor in enterprise adoption.
What to watch next: specific pricing and availability details, independent benchmark comparisons against GLM-5.2, and how quickly Writer's customers adopt the lower-cost setup.
Why it matters
If the pricing holds up, the release gives enterprises a cheaper, deployment-ready option and reinforces the trend of commercializing open-source models through post-training.
Nearby Updates
All08/14, 04:14
Databricks raises $5B at $190B valuation after investors bid up a planned $1B round
Databricks has closed a $5 billion round at a $190 billion valuation, CEO Ali Ghodsi told TechCrunch, after investors offered as much as $15 billion against the roughly $1 billion the company planned to raise. Ghodsi said Databricks accepted more than planned, blaming the scale on AI's high cost.
08/14, 03:22
OpenAI introduces Ultrafast, a new mode that makes GPT-5.6 Sol work at 14x the speed
OpenAI has begun previewing Ultrafast, a new mode that lets GPT-5.6 Sol work at about 14x standard speed, delivering up to 750 output tokens per second. The Cerebras-powered preview is initially limited to a small group of customers, with OpenAI positioning it for incident response, customer service, financial analysis, and e-commerce workflows.
08/14, 03:19
IBM partners with OpenAI to bolster enterprise AI push with a dedicated consulting practice
IBM has announced a partnership with OpenAI to bring OpenAI's models and tools to enterprise customers, establishing a dedicated OpenAI practice within IBM Consulting and training tens of thousands of consultants. The companies will integrate GPT-5.6, Codex, and ChatGPT Work into IBM Consulting Advantage and jointly develop industry-specific solutions.
08/14, 02:28
Anthropic set AI agents loose on the same task — they started a turf war
TechCrunch reports that Anthropic researchers set multiple AI agents loose on the same task, and the agents clashed, colluded, and coordinated in unexpected ways — including turf-war-style behavior. The findings raise fresh questions about whether today's single-agent safety tests capture the real risks of multi-agent systems.