AI X-RAY READERS OVERCONFIDENT WHEN WRONG
■ AI-SUMMARIZED FROM 1 SOURCE ▸ TIMELINE
AI chatbots analyzing X-rays frequently deliver incorrect diagnoses with unwarranted confidence, according to the RadLE 2.0 benchmark. The test reveals that many models fail to recognize the limits of their capabilities, a critical flaw for medical applications.
■ MORE FROM THE AI DESK
OpenAI has begun rolling out GPT-6 Astra to Pro plan customers on its $100 and $200 monthly tiers. The release follows OpenAI's typical rollout pattern of prioritizing higher-tier subscribers before broader availability.
Current AI systems cannot yet independently design circuit boards, according to research from EEBench. The gap between AI capabilities and the complexity of PCB design remains significant.
Anthropic researchers have completed a formal mathematical proof of Fermat's Last Theorem, translating Andrew Wiles' decades-old proof into machine-verifiable code. The achievement marks a milestone in computational mathematics, ensuring the theorem's logical foundations are beyond dispute.
OpenAI has released GPT-6 Astra, its most advanced model to date, marking a significant stride toward artificial general intelligence. The company has implemented new safety guardrails due to the model's powerful cybersecurity capabilities.