{"database":"biostudies-literature","file_versions":[],"scores":null,"additional":{"omics_type":["Unknown"],"volume":["12"],"submitter":["Takahashi H"],"pubmed_abstract":["<h4>Background</h4>Generative artificial intelligence (AI) is increasingly used in medical education, including AI-based virtual patients to improve interview skills. However, how much AI-based assessment (ABA) differs from human-based assessment (HBA) remains unclear.<h4>Objective</h4>This study aimed to compare the quality of clinical interview assessments generated via an ABA (GPT-o1 Pro [ABA-o1] and GPT-5 Pro [ABA-5]) with those generated via an HBA conducted by clinical instructors in an AI-based virtual patient setting. We also examined whether AI reduced evaluation time and assessed agreement across participants with different levels of clinical experience.<h4>Methods</h4>A standardized case of leg weakness was implemented in an AI-based virtual patient. Seven participants (2 medica"],"journal":["JMIR medical education"],"pagination":["e81673"],"full_dataset_link":["https://www.ebi.ac.uk/biostudies/studies/S-EPMC12912650"],"repository":["biostudies-literature"],"pubmed_title":["AI- vs Human-Based Assessment of Medical Interview Transcripts in a Generative AI-Simulated Patient System: Cross-Sectional Validation Study."],"pmcid":["PMC12912650"],"pubmed_authors":["Aiyama Y","Kishi M","Nagai S","Naito T","Matsuura T","Tomoda Y","Kondo T","Shikino K","Shinohara T","Yamada Y","Tokushima Y","Sano F","Enomoto A","Watanabe R","Takahashi H"],"additional_accession":[]},"is_claimable":false,"name":"AI- vs Human-Based Assessment of Medical Interview Transcripts in a Generative AI-Simulated Patient System: Cross-Sectional Validation Study.","description":"<h4>Background</h4>Generative artificial intelligence (AI) is increasingly used in medical education, including AI-based virtual patients to improve interview skills. However, how much AI-based assessment (ABA) differs from human-based assessment (HBA) remains unclear.<h4>Objective</h4>This study aimed to compare the quality of clinical interview assessments generated via an ABA (GPT-o1 Pro [ABA-o1] and GPT-5 Pro [ABA-5]) with those generated via an HBA conducted by clinical instructors in an AI-based virtual patient setting. We also examined whether AI reduced evaluation time and assessed agreement across participants with different levels of clinical experience.<h4>Methods</h4>A standardized case of leg weakness was implemented in an AI-based virtual patient. Seven participants (2 medica","dates":{"release":"2026-01-01T00:00:00Z","publication":"2026 Feb","modification":"2026-07-16T00:38:30.916Z","creation":"2026-07-09T10:26:28.154Z"},"accession":"S-EPMC12912650","cross_references":{"pubmed":["41701946"],"doi":["10.2196/81673"]}}