TruaceTracing the truth around AITuesday, September 15, 2026
The Index

What the evidence says.What the public feels.

Ranks distinct AI gain and problem claims from the published record. Scores reward impact, independent source strength, scale, confidence, and recency.

1,419 results
Show filters and sorting

AI gains · 788

72
GainLabor· Stable· Evidence: Moderate (1 source)

Frontier models are approaching industry experts in deliverable quality on real-world economically valuable tasks and can perform them cheaper and faster than unaided experts when paired with human oversight.

Researchers introduced GDPval, a benchmark of real-world economically valuable tasks spanning 44 occupations and the top 9 U.S. GDP sectors, built from work of experienced industry professionals. As of January 2026, they reported frontier models improving linearly and approaching expert deliverable quality, with potential to complete tasks cheaper and faster than unaided experts when paired with human oversight.

Impact 30%49
Evidence 25%95
Scale 20%60
Confidence 15%87
Recency 10%89

Updated Jul 22, 2026 · TRV-2026-0471

72
GainScience· Stable· Evidence: Moderate (1 source)

LLM-based autonomous agents leveraging vast Web knowledge show potential for human-level intelligence and enable applications across social science, natural science, and engineering.

Published March 22, 2024, this peer-reviewed survey examines the shift from agents trained with limited knowledge in isolated environments to agents built on large language models trained on vast Web knowledge. The authors propose a unified construction framework and systematically review applications and evaluation methods.

Impact 30%49
Evidence 25%95
Scale 20%60
Confidence 15%87
Recency 10%89

Updated Jul 20, 2026 · TRV-2026-0461

72
GainHealth· Stable· Evidence: Moderate (1 source)

AI systems deployed in hospitals and clinics have improved clinical decision-making, hospital operations, medical image analysis, and patient monitoring via wearables.

This peer-reviewed review from March 2024 surveys how artificial intelligence is being integrated across hospitals and clinics, covering clinical decision support, operational management, medical image analysis, and patient monitoring with AI-powered wearables, drawing on case studies of domain-specific transformation.

Impact 30%49
Evidence 25%95
Scale 20%60
Confidence 15%87
Recency 10%89

Updated Jul 20, 2026 · TRV-2026-0454

72
GainBusiness· Stable· Evidence: Moderate (1 source)

Manufacturing SMEs in Sweden structure, bundle, and leverage AI resources to transform key business and production operations and create competitive advantage.

Published April 3 2024, this peer-reviewed study investigated AI implementation in manufacturing SMEs in Sweden across packaging, plastic, and metal sectors. It found SMEs build an AI resource portfolio through acquiring and accumulating resources, bundle them into learning and governance capabilities, and leverage them in production.

Impact 30%49
Evidence 25%95
Scale 20%60
Confidence 15%87
Recency 10%89

Updated Jul 20, 2026 · TRV-2026-0444

AI problems · 631

68
ProblemHealth· Newly added· Evidence: Moderate (1 source)

Machine learning-informed toxicology analysis indicates 6PPD-quinone exposure increases IBD risk in human colon epithelial cells by downregulating NR1H4, causing lipid and cholesterol accumulation, mitochondrial dysfunction, and elevated IL-6, TNF-alpha, and IL-8.

Researchers used network toxicology, multi-model machine learning, molecular docking, and in vitro experiments in human intestinal epithelial cells to probe the tire-derived pollutant 6PPD-quinone. The workflow identified 60 overlapping 6PPD-Q-IBD targets and prioritized six core genes, with NR1H4 as a key mediator that binds strongly to 6PPD-Q.

Impact 30%49
Evidence 25%95
Scale 20%35
Confidence 15%87
Recency 10%98

Updated Sep 6, 2026 · TRV-2026-0998

68
ProblemHealth· Newly added· Evidence: Moderate (1 source)

Canadian medical students not interested in radiology reported a higher perceived impact of AI on the field and lower perceived career sustainability, alongside limited adequate radiology exposure.

By September 2026, researchers surveyed 73 Canadian medical students to test whether perceptions of radiology's procedural scope and AI integration differed by interest in radiology. Interested students rated therapeutic and diagnostic procedures as more important and reported higher perceived career sustainability, while non-interested students reported higher perceived AI impact.

Impact 30%49
Evidence 25%95
Scale 20%35
Confidence 15%87
Recency 10%98

Updated Sep 6, 2026 · TRV-2026-0997

68
ProblemHealth· Newly added· Evidence: Moderate (1 source)

Current AI models frequently produce visually plausible but anatomically inaccurate illustrations, with major errors across all models, making them unreliable for unsupervised clinical use.

On September 4, 2026, a peer-reviewed study in Academic Radiology evaluated 25 musculoskeletal imaging cases where three AI systems generated visual summaries from report text alone. Two fellowship-trained MSK radiologists rated each image for anatomical accuracy and clinical usefulness, finding Gemini 3.0 Pro most consistent at 40-42% accurate and 60-65% useful, while ChatGPT and Perplexity frequently produced plausible but inaccurate images.

Impact 30%49
Evidence 25%95
Scale 20%35
Confidence 15%87
Recency 10%98

Updated Sep 6, 2026 · TRV-2026-0996

68
ProblemHealth· Newly added· Evidence: Moderate (1 source)

Adoption of digital pathology and AI in liver disease is constrained by access and logistics barriers, quality issues, lack of guidance, and unproven real-world effectiveness and clinical safety.

A September 2026 review in The Lancet Digital Health summarizes how digital pathology, image analysis, and AI, including deep learning on high-resolution whole-slide images, are being applied to liver disease, liver cancer diagnosis, and transplantation assessment. The authors describe expanding use from long-standing research applications to increasing clinical practice access.

Impact 30%49
Evidence 25%95
Scale 20%35
Confidence 15%87
Recency 10%98

Updated Sep 6, 2026 · TRV-2026-0995

Recomputed live from the record · Sep 15, 2026, 10:54 PM