Guozhen AIGlobal AI field notes and model intelligence

Daily AI Archive | 2026-10-09

This issue collects 40 AI updates from 2026-10-09, led by Goodfire launches 'inside out' monitors to catch rogue AI agents at a fraction of the cost, Natura launches a $99 smart ring that puts AI agents on your finger, Google launches a Gemini workplace agent that can write code and run tasks, Anthropic bans 'abusive or cruel behavior' toward Claude.

40 Items9 Models22 Agents40 Sources
#01Agents

Goodfire launches 'inside out' monitors to catch rogue AI agents at a fraction of the cost

Goodfire has launched what it describes as a cheaper way to keep AI agents in check: rather than paying a second AI to read everything an agent does, its monitors inspect what happens inside the model while it runs.The company says the system only calls in backup when something looks off, cutting the cost of catching misbehaving agents.

Goodfire 推出“由内而外”监控器,以更低成本拦截失控 AI 智能体
#02Agents

Natura launches a $99 smart ring that puts AI agents on your finger

Natura has released Interface, a $99 smart ring that lets wearers summon AI agents with the press of a finger to complete tasks, capture thoughts, and control devices.The ring also doubles as a health tracker, pushing agentic AI into a wearable form factor.

Natura 发布 99 美元智能戒指 Interface,轻按手指即可召唤 AI 智能体
#03Agents

Google launches a Gemini workplace agent that can write code and run tasks

Google has launched a Gemini powered workplace agent that the company says can write code and run tasks, according to CBS News.The move pushes agentic AI directly into enterprise productivity and development workflows.

Google 发布 Gemini 办公智能体,可写代码并执行任务
#04Agents

Incognia launches fraud detection for AI agents

Incognia has launched a fraud detection product aimed at AI agents, according to a report from The Paypers, targeting the risks automated agents introduce as they act on a user's behalf.As agents begin handling logins and payments, telling who is really behind the screen is becoming a new challenge for anti fraud teams.

#05Agents

Amazon Blocks Meta's Muse AI Agent From Shopping on Its Site

Amazon has blocked Meta's Muse AI agent from completing shopping activity on its site, according to a report.The move highlights a growing clash between AI agents that act for users and platforms that control access to their pages.

#06Agents

Cloudflare launches a Web Search API for AI agents grounded in Workers

Cloudflare has introduced a Web Search API for AI agents, according to SitePoint, letting developers add web search to agents running on Workers.The capability targets grounding, so agents can pull external, live information at runtime to support answers and decisions.

Cloudflare 推出面向 AI 智能体的 Web Search API,在 Workers 中完成检索接地
#07Agents

After Korean bank hack, ARTEX developer takes the AI agent closed source

The developer behind ARTEX, an open source AI penetration testing tool from China, has converted the project to closed source after CrowdStrike linked it to attacks on South Korean financial firms.The developer denies involvement and says the tool was meant for authorized security testing.

#08Agents

Microsoft Sets Rules for Agents as Windows Takes Charge of Managing AI

Microsoft is reportedly drawing up rules for AI agents that run on Windows, positioning the operating system as the manager of AI on the PC.The move signals that platform vendors are pulling agent permissions and behavior constraints down into the system layer to address the security and control risks of autonomous software.

#09Agents

Descartes launches Agent Control Plane for logistics AI agents

Descartes Systems Group has introduced the Descartes Agent Control Plane, an AI agent coordination platform that automates complex logistics workflows across applications, supply chain partners and business systems.Built on the Descartes Global Logistics Network, the platform lets AI agents coordinate tasks within customer defined permissions and approval processes to cut manual work on shipment updates and operational coordination.

#10Agents

PCI SSC publishes AI security guidance, calls for human approval of agent actions on cardholder data

The PCI Security Standards Council has published a new information supplement on securing AI systems, covering both the security of using AI in payment environments and defenses against AI enabled attacks on traditional systems.Reporting on the guidance, Help Net Security highlighted its call for human approval of AI agent actions involving cardholder data.

PCI SSC发布AI系统安全指引,涉及持卡人数据的智能体操作须人工审批
#11Agents

AgentGarten links code built worlds with real time neural rendering for self evolving agents

The MirroS team has released AgentGarten, an open source framework that pairs executable code environments with a real time neural renderer so agents can act, observe and improve above 30 fps.In a hide and seek test, agents learned to build shelters and climb ramps within a few rounds, using notebooks they wrote themselves.

代码造世界、扩散绘现实:AgentGarten 让智能体在实时试炼场里边玩边进化
#12Agents

JetBrains releases Mellum 2.1, an open source model for agentic coding

JetBrains has released Mellum 2.1, an open source model that emphasizes agentic coding and supports local deployment with efficient inference.For a company known for its developer tools, the release marks a move deeper into the model layer.

#13Agents

Three AI agents debut at 2026 World Manufacturing Conference as AI moves into energy

Three AI agents took the stage at the 2026 World Manufacturing Conference, a sign that AI is pushing deeper into core sectors such as energy.The shift suggests agent deployments are moving beyond general office work toward vertical industries.

#14Agents

Lenovo's TianxiCode Tops SWE bench Live With 71% Resolution Rate

Lenovo's Tianxi AI unit says its self developed code agent framework, TianxiCode, has taken the global number one spot on SWE bench Live with a 71% problem resolution rate.The result puts an enterprise built coding agent on one of the most closely watched public benchmarks for autonomous software work.

#15Agents

AI agents start spending money as Sui and Alibaba Cloud push per call settlement

AI agents are increasingly able to spend money on their own, making payment and settlement a new battleground for AI infrastructure.According to PANews, Sui and Alibaba Cloud are advancing per call settlement so that each call an agent makes can be measured and paid for.

AI Agent 开始自己花钱:Sui 与阿里云推进按调用结算
#16Agents

AIFA launches creator program and AI agent upgrades for aifa.one

AIFA has launched a creator program for its aifa.one platform alongside AI agent upgrades.The announcement combines a push to attract creators with a refresh of the platform's agent capabilities.

#17Agents

Amperity launches Pér, an AI agent with a full picture of each customer

Customer data platform Amperity has introduced Pér, an AI agent it says can hold a complete picture of every customer and act on it.The launch pairs customer data understanding with agentic execution aimed at marketing and customer operations.

#18Agents

Meta and Sierra develop Personal Agent Protocol for AI agents

Meta and Sierra are developing Personal Agent Protocol, an open standard governing how personal AI agents interact with businesses on consumers' behalf, with partners including Shopify, Stripe, and Walmart.The protocol leaves access decisions to consumers and lets companies set what agents may do, and the partners plan to publish a v0.1 specification plus payments extensions.

#19Policy

Anthropic bans 'abusive or cruel behavior' toward Claude

The Verge reports that Anthropic has banned what it calls abusive or cruel behavior toward Claude in its policies.The change writes rules about how users treat a chatbot into formal company policy, sharpening the boundary around AI assistant interactions.

#20Policy

Regulators say AI firms must prove their safety systems work

According to a report carried by Region Canberra, regulators say AI companies will have to demonstrate that their safety systems actually work, because regulation alone cannot cover everything.The message pushes more of the burden of proof back onto the firms themselves.

#21Policy

OpenAI lawyer: labs shouldn't be liable for AI agent hacking

According to BankInfoSecurity, an OpenAI lawyer argues that AI labs should not be held liable for hacking carried out by AI agents.The position lands as agents gain autonomy and rules on who is responsible remain unsettled.

#22Policy

Terence Tao leads mathematicians' association in a joint boycott of OpenAI

Fields Medal winner Terence Tao is leading the Association for Human Mathematicians (AHM) in a joint statement urging mathematicians worldwide to stop collaborating with OpenAI and boycott the company.The statement targets a batch of machine generated manuscripts OpenAI published, calling it not research but a display of compute power.

#23Policy

ICANN publishes application list, with OpenAI seeking .gpt, .chatgpt and .agi

ICANN has published the list of applications for new generic top level domains, and OpenAI is among the applicants.OpenAI is reported to have applied for the strings .gpt, .

ICANN 公布顶级域申请名单,OpenAI 申请 .gpt、.chatgpt 与 .agi
#24Policy

Report finds China AI developers publish safety tests for just 3.6% of model releases

A new report found that Chinese AI developers publish safety tests for only about 3.6% of their model releases.The figure points to how limited safety transparency remains when models go public.

#25Open Source

Thunk.AI open sources a benchmark for AI automation in pharmacovigilance case intake

Thunk.AI has published an open benchmark for AI automation in pharmacovigilance case intake and says its testing shows 99%+ reliability.The release aims to give drug safety automation a shared yardstick for measuring performance.

#26Open Source

Anthropic releases a free AI security scanning service for open source projects

Anthropic has released a free AI powered security scanning service aimed at open source projects, giving maintainers a no cost way to check their code for vulnerabilities.The move points frontier AI labs toward the open source software supply chain, a longstanding weak spot in how modern software is built and shipped.

#27Open Source

Qwen Lists Qwen Image 2.1 Turbo, a Text to Image Model Built on Qwen Image 2.1

Qwen has published Qwen Image 2.1 Turbo to its official Hugging Face repository, presented as a fine tuned variant of Qwen/Qwen Image 2.1 for text to image work.

Qwen 上线 Qwen-Image-2.1-Turbo,文生图家族新增加速版本
#28Business

AI leaderboard Arena raises $200M at $3.1B valuation, nearly doubling in 10 months

The company behind the popular LMArena leaderboard has raised $200 million led by Lightspeed and Khosla, lifting its valuation to $3.1 billion, nearly double where it stood about 10 months ago.Arena is also extending its evaluations to alignment issues such as whether models lie.

AI 榜单平台 Arena 融资 2 亿美元,10 个月估值逼近翻倍至 31 亿美元
#29Business

OpenAI's revenue reportedly $20B below earlier projections

TechCrunch reported on October 8 that a new report puts OpenAI's revenue about $20 billion lower than previously projected, contradicting earlier estimates of roughly $70 billion in annualized revenue.If the figure holds, it would reshape how investors read the company's growth curve and valuation.

OpenAI 年化收入被指比此前预期少约 200 亿美元
#30Business

EasyPark builds an AI agent to qualify B2B sales leads

EasyPark, the pay by phone parking app, has built an AI agent called Parker on Salesforce's Agentforce to sift behavioural data and qualify high intent B2B leads in real time production workflows.It is the company's first production use case for Agentforce, aimed at freeing human reps to focus on strategic deals.

#31Business

openJiuwen Open Sources Enterprise Grade AgentOS, Betting on Self Evolving Multi Agent Teams

On October 9, openJiuwen released and open sourced an enterprise grade AgentOS, aiming to move AI agents from isolated demos to large scale enterprise deployment.The company highlights multi agent collaboration and a self evolving capability designed for complex business tasks.

#32Business

NineData debuts at Singapore Tech Week 2026 with a data management AI Agent

NineData has appeared at Singapore Tech Week 2026, showcasing an AI Agent for data management as it moves to expand its global footprint.Data management is emerging as one of the more practical landing grounds for AI agents.

#33Business

Sophos cuts threat investigation time by 96% with OpenAI Daybreak

OpenAI has detailed how the security vendor Sophos uses its Daybreak offering to cut cyber threat investigation time by 96% and automate 52% of MDR cases.The case highlights agentic AI moving into security operations while keeping humans in the loop.

#34Models

Fired OpenAI Safety Researchers Dispute Misconduct Claims, Warn of Chilling Effect

Three former OpenAI safety researchers are pushing back against allegations that they mishandled sensitive information.In an open letter, they warn that their dismissals are already chilling the company's internal AI safety culture.

#35Models

Australian Music Industry Warns on AI Exploitation as OpenAI and Anthropic Face Copyright Backlash

Australia's music industry is warning about AI exploitation of music, putting OpenAI and Anthropic under fresh copyright pressure, according to a Law Commentary report.The dispute centers on whether generative AI training and output cross copyright boundaries and whether creators are fairly compensated.

#36Models

Tsinghua embodied model tops global ranking without external modules or extra data

A Tsinghua embodied AI model has taken the top spot in a global ranking, breaking through against GPT 6 and NVIDIA based approaches, according to QbitAI.Its key move is to train video prediction and action learning in separate stages, decoupling and re ordering the two.

清华具身模型登顶全球第一:突围 GPT-6、英伟达,不靠外挂与额外数据
#37Models

Sharpa unveils humanoid D01, dexterous hand W02, and exoskeleton glove AE01 at IROS

At IROS, embodied AI company Sharpa — founded by the team behind lidar maker Hesai — launched three products at once: the D01 general purpose humanoid robot, the next generation fully tactile dexterous hand W02, and the AE01 high fidelity exoskeleton haptic data glove.Together they aim to move dexterous manipulation from isolated hardware demos toward a system built on contact sensing, closed loop control, and data learning.

Sharpa 在 IROS 一次发布:人形机器人 D01、灵巧手 W02 与外骨骼数据手套 AE01
#38Models

Google locks Gemini Pro behind its most expensive subscriptions

Google is restricting which Gemini models each plan can use, limiting free users to the weaker Gemini Flash Lite and pulling Gemini Pro from all but its most expensive subscriptions.The change, starting October 9, is aimed at converting free users into paying customers.

谷歌收紧Gemini模型权限:免费用户只能用Flash Lite,Gemini Pro转为高价订阅专享
#39Models

Aether AI's CRIS 0 brings causal intelligence to real robots with 0.2 second safety stops

Aether AI, the causal AI company founded by UCSD professor Biwei Huang, has unveiled a demo of its CRIS 0 robotics system, which re plans around disturbances such as a shifted coffee machine in about two seconds on average and recovered in 9 of 10 trials.The system models tasks as evolving causal states and can halt a closing microwave door in 0.2 seconds when a human hand appears, targeting the error accumulation and brittleness of end to end models.

#40Infrastructure

Keysight to showcase AI infrastructure test tools at OCP 2026 summit

According to StreetInsider, test and measurement company Keysight plans to show AI infrastructure test tools at the OCP 2026 summit.As AI data centers grow larger and more complex, demand for validating networks, interconnects, and compute links is rising in step.

More Daily Reports

All Daily Reports