Lång K, et al. Artificial intelligence-supported screen reading versus standard double reading in the Mammography Screening with Artificial Intelligence trial (MASAI): a clinical safety analysis of a randomised, controlled, non-inferiority, single-blinded, screening accuracy study. Lancet Oncol. 2023;24:936-44.(読影の作業量 44.3%減)
Gommers J, et al. Interval cancer, sensitivity, and specificity comparing AI-supported mammography screening with standard double reading without AI in the MASAI study. Lancet. 2026;407:505-14.(スウェーデン105,934人。中間期がん 1000人あたり 1.55 vs 1.76、比 0.88〈95%CI 0.65–1.18、p=0.41〉=非劣性で、有意な減少ではない。感度 80.5% vs 73.8%。特異度はどちらも98.5%。死亡は評価項目ではない)
Yao X, et al. Artificial intelligence-enabled electrocardiograms for identification of patients with low ejection fraction: a pragmatic, randomized clinical trial. Nat Med. 2021;27:815-9.(米国。AIを使った群で、心臓のポンプ機能低下の新たな診断が増えた)
Bean AM, et al. Reliability of LLMs as medical assistants for the general public: a randomized preregistered study. Nat Med. 2026;32:609-15.(AI単体では関係する病気を94.9%の場面で挙げた。一般の人がAIを使うと、対照群〈主にネット検索〉のほうが病気を挙げられた。受診先の判断は差なし)
Ayers JW, et al. Comparing physician and artificial intelligence chatbot responses to patient questions posted to a public social media forum. JAMA Intern Med. 2023;183:589-96.