TruaceTracing the truth around AIWednesday, September 16, 2026
TRV-2026-1114Version 1 · Certified

Written 2026-09-16 06:57:00 UTC · current record

Reason for this version

Certified into the record

Canonical text (the exact bytes fingerprinted)

TRUVACE RECORD VERSION
record: TRV-2026-1114
version: 1
kind: certified
reason: Certified into the record
timestamp: 2026-09-16T06:57:00.227213Z
status: published
lens: trace
sector: science
headline: Can AI Predict Publication? Multimodal Large Language Models and the Structural Determinants of Surgical Scholarship
dek: BackgroundWhether artificial intelligence can identify publishable scientific work is untested. We evaluated whether a multimodal large language model (MLLM) could predict, from poster content alone, which abstracts at the American Association for the Surgery of Trauma (AAST) Annual Meetings reached publication, and characterized the investigator, institutional, and domain level determinants situating model performance.MethodsWe retrospectively analyzed 260 abstracts from the 2021-2022 AAST Annual Meetings. Bibl…
gain_title: GPT-4.1 scored poster images alone and predicted which AAST abstracts reached publication with 58.5% overall accuracy, rising to 74.2% in Violence, Societal, and Behavioral and 62.2% in Hemorrhage, Resuscitation, and Vascular Control.
problem_title: Model performance fell to chance in Critical Care and Outcomes and Systems, Technology, and Process Optimization, where publication depended on institutional factors absent from the poster such as multicenter scaffolding, mentorship, and senior author fluency.
trace_subject: GPT-4.1 predicting publication of AAST surgical abstracts from poster content alone
gain_reading: GPT-4.1 scored poster images alone and predicted which AAST abstracts reached publication with 58.5% overall accuracy, rising to 74.2% in Violence, Societal, and Behavioral and 62.2% in Hemorrhage, Resuscitation, and Vascular Control.
gain_evidence: GPT-4.1 predicted publication with 58.5% accuracy overall ( P = .009) | 74.2% in Violence, Societal, and Behavioral
problem_reading: Model performance fell to chance in Critical Care and Outcomes and Systems, Technology, and Process Optimization, where publication depended on institutional factors absent from the poster such as multicenter scaffolding, mentorship, and senior author fluency.
problem_evidence: falling to chance in Critical Care and Outcomes and Systems, Technology, and Process Optimization. | fell to chance where advancement depended on institutional factors absent from the poster
quick_read: Researchers tested whether GPT-4.1 could predict publication from poster images alone for 260 abstracts presented at the 2021-2022 AAST Annual Meetings. By September 2026 publication date, 142 had published, and the model achieved 58.5% accuracy overall, with higher accuracy in Violence, Societal, and Behavioral and Hemorrhage, Resuscitation, and Vascular Control, but chance-level performance in other domains.

The result matters because it shows AI can read scientific signal legible on the page but cannot read structural advantage that also determines publication, such as multicenter scaffolding and mentorship. Uncertainty remains about generalizability beyond AAST, beyond 2021-2022, and about demographic inference from public data, plus why domain dependence occurs.
limitation: Findings limited to 260 AAST abstracts from 2021-2022, with investigator demographics inferred from public data and model performance dependent on domain and on structural factors not visible on posters.
tag: Dual reading
key_points: Retrospective analysis of 260 abstracts from 2021-2022 AAST Annual Meetings with bibliographic confirmation of publication. | 142 (54.6%) reached publication at a mean of 13.4 months, with multicenter origin the only independent predictor ( P = .02). | Hispanic investigators were underrepresented among first and senior authors and absent from published Hemorrhage, Resuscitation, and Vascular Control subset. | Model used six-domain rubric averaged over 30 iterations on poster images; performance fell to chance in Critical Care and Outcomes and Systems, Technology, and Process Optimization.
rundown: The study analyzed 260 posters from the 2021-2022 American Association for the Surgery of Trauma meetings, confirming publication via bibliographic search and scoring each poster image with GPT-4.1 on a six-domain rubric averaged over 30 iterations.

Overall 142 abstracts (54.6%) published at mean 13.4 months; multivariable analysis found multicenter origin as the only independent predictor, while Hispanic investigators were underrepresented among first ( P = .03) and senior ( P = .01) authors and absent from the published Hemorrhage, Resuscitation, and Vascular Control subset.
sources:
- peer_reviewed | The American Surgeon™ | https://doi.org/10.1177/00031348261487667 | 2026-09-14
prev: 0000000000000000000000000000000000000000000000000000000000000000
sha256
4b63f26ccef6e9c6d537055fa435e3deaaf54a7937d7775469fab3f61924b8bb
previous
0000000000000000000000000000000000000000000000000000000000000000
Verify this record
How to verify without trusting this page

Fetch the canonical text of any version from /api/record/TRV-2026-1114 and hash it yourself — for example shasum -a 256 on the saved canonical field. The result must equal content_hash, and each version’s text ends with prev:followed by the prior version’s hash (version 1 chains to 64 zeros). If a single character of any version had been altered since certification, the chain would not reproduce.