Realtime AI News
Thunk.AI open-sources a benchmark for AI automation in pharmacovigilance case intake
Thunk.AI has published an open benchmark for AI automation in pharmacovigilance case intake and says its testing shows 99%+ reliability. The release aims to give drug-safety automation a shared yardstick for measuring performance.
Thunk.AI has published an open benchmark for AI automation in pharmacovigilance case intake, claiming that its testing demonstrates reliability above 99%. The move was reported by BioSpace.
Pharmacovigilance is the post-market safety monitoring of medicines. A large part of the job is taking adverse-event reports from doctors, patients and regulators, then classifying and entering them into safety systems — high-volume, repetitive work where accuracy is non-negotiable.
Thunk.AI's approach is to automate that intake with AI while opening a benchmark to the public so the performance of such systems can be measured. According to the report, the company says its own testing pushed reliability past 99%.
The value of open-sourcing the benchmark is comparability. Drug-safety automation has long lacked a public, common way to score systems, leaving vendors to report accuracy on their own terms. A shared benchmark gives competing tools a single ruler.
Still, the 99%-plus figure is a company claim, and the report offers no independent third-party verification. For pharma companies and regulators, whether such tools can go live will hinge on how they perform on real, noisy data and how they satisfy audit requirements.
What to watch next: whether other drug-safety vendors adopt or benchmark against the standard, and whether regulators use it to push toward standardized evaluation of automation in case processing.
Why it matters
Case intake is one of the heaviest compliance burdens in pharma; if the automation's reliability can be independently verified, AI could cut the manual cost of processing safety reports and reshape the drug-safety software and outsourcing market.
Nearby Updates
All10/09, 01:35
Incognia launches fraud detection for AI agents
Incognia has launched a fraud detection product aimed at AI agents, according to a report from The Paypers, targeting the risks automated agents introduce as they act on a user's behalf. As agents begin handling logins and payments, telling who is really behind the screen is becoming a new challenge for anti-fraud teams.
10/09, 01:00
Anthropic bans 'abusive or cruel behavior' toward Claude
The Verge reports that Anthropic has banned what it calls abusive or cruel behavior toward Claude in its policies. The change writes rules about how users treat a chatbot into formal company policy, sharpening the boundary around AI-assistant interactions.
10/09, 00:21
Google launches a Gemini workplace agent that can write code and run tasks
Google has launched a Gemini-powered workplace agent that the company says can write code and run tasks, according to CBS News. The move pushes agentic AI directly into enterprise productivity and development workflows.
10/09, 00:00
Natura launches a $99 smart ring that puts AI agents on your finger
Natura has released Interface, a $99 smart ring that lets wearers summon AI agents with the press of a finger to complete tasks, capture thoughts, and control devices. The ring also doubles as a health tracker, pushing agentic AI into a wearable form factor.