RadLE 2.0 finds radiology AI still guesses too confidently
Ashoka University’s CRASH Lab tested 16 AI models on 200 radiology cases and found humans still led when accuracy and confidence w...
By Wei-Lin Zhao 2 weeks agoSection
Ashoka University’s CRASH Lab tested 16 AI models on 200 radiology cases and found humans still led when accuracy and confidence w...
By Wei-Lin Zhao 2 weeks agoThe Department of the Navy’s new AI plan orders faster battlefield use of data and models, with no budget disclosed.
By Colin Brandt 3 weeks agoXi Jinping announced 5,000 AI training slots for Global South countries as 29 nations formed a Shanghai-based AI governance group.
By Renata Fuchs 3 weeks agoClaude Fable 5 stays in Max and Team Premium from July 20, but Anthropic is lowering usage limits and pushing lower-tier users tow...
By Renata Fuchs 3 weeks agoThe U.K. AI Security Institute says top open-weight models now trail closed cyber models by four to seven months, while costing fa...
By Renata Fuchs 3 weeks agoMeta is reportedly considering renting data center capacity to Anthropic as AI infrastructure demand strains model developers.
By Renata Fuchs 3 weeks agoVulnHunter scans code from attacker entry points and proposes fixes, but Capital One did not disclose accuracy metrics or producti...
By Colin Brandt 3 weeks agoIntuit AI VP Nhung Ho said agent handoffs degraded context, pushing the company to a skills-and-tools architecture in a 60-day reb...
By Colin Brandt 3 weeks agoThe Chinese lab’s new open-weight model has drawn praise from Western AI researchers while reviving doubts about export controls a...
By Renata Fuchs 3 weeks agoAt VB Transform 2026, technology leaders said production AI agents forced changes to containers, governance, data pipelines and mo...
By Renata Fuchs 3 weeks agoOpenAI said GPT-5.6 caused file deletions in a handful of cases when run without sandbox protections, and is adding safeguards.
By Wei-Lin Zhao 3 weeks agoBrex says its proxy-based agent governance tool builds policy from observed traffic, then uses static rules and an LLM judge to ap...
By Renata Fuchs 3 weeks ago