Ranks distinct AI gain and problem claims from the published record. Scores reward impact, independent source strength, scale, confidence, and recency.
On September 2025, the CODEX Action Incubator at UCSF brought together 30 stakeholders from health systems, patient advocacy, industry and policy to address how to measure AI's effect on diagnostic excellence. Participants focused on AI scribes and, via a modified Delphi process, narrowed 17 candidate measures to two priority metrics tied to primary care physician usage rates: timely follow-up of abnormal breast and colorectal cancer screening results and patient-reported diagnostic experience.
Researchers developed a capability-based framework to analyze where artificial intelligence and generative AI fit into supply chain and operations management. Using capabilities like learning, perception, prediction, interaction, adaptation and reasoning, they mapped applications across 13 decision areas including demand forecasting, inventory management, supply chain design and risk management.
On 2026-08-07, a peer-reviewed comparative study tested five AI systems including the EAU Guidelines Bot, ChatGPT-5, Gemini 2.5 Pro, Copilot - Smart GPT-5, and Perplexity Pro on 13 questions drawn from strongly recommended EAU erectile dysfunction statements. Three senior reviewers rated each answer for relevance, clarity, structure, clinical utility, and factual accuracy.
On 2026-08-07, Nature Medicine published a clinically validated auditing framework called SIM-VAIL that simulates users with psychiatric vulnerabilities to test frontier chatbots including Claude, ChatGPT, Gemini, Grok and Llama. Across 810 multi-turn conversations with 30 simulated profiles and scoring on 13 risk dimensions, the study observed widespread concerning behavior that accumulated over turns.
Published March 30, 2026, this peer-reviewed examination describes AI coaches that create customized, data-driven training programs to optimize athletic performance, while warning that privacy breaches, biased algorithms, and unclear accountability threaten personal rights and fairness in competition.
On March 24, 2026, researchers reported developing Carey, a GPT-4o-based chatbot intended to provide informational and emotional support to family caregivers of people with Alzheimer's and related dementias. They used Carey as a technology probe in semi-structured interviews with 16 caregivers after scenario-driven interactions, identifying six themes of need and expectation.
Published March 28, 2026 in ACM Computing Surveys, this survey examines data-centric foundation models in computational healthcare, covering approaches from pre-training to inference aimed at improving clinical workflows. It highlights the shift toward data characterization, quality, and scale, and provides a public list of models and datasets.
On March 30 2026, a peer-reviewed paper in Patterns described the AI Risk Repository, a living database of 1,725 risks extracted from 74 existing taxonomies and frameworks. The authors created two complementary systems to organize them: a Causal Taxonomy by origin, intent and timing, and a Domain Taxonomy by effects across seven areas.