Performance of Multimodal Large Language Models in Detection and Position Assessment of Thoracic Devices on Chest Radiographs. [PDF]
Güzel HE, Özenbaş C, Saravi B.
europepmc +1 more source
Between Help and Harm: An Evaluation Study of Mental Health Crisis Handling by Large Language Models. [PDF]
Arnaiz-Rodriguez A +7 more
europepmc +1 more source
Is one run enough? Reproducibility of flagship large language models across temperature and reasoning settings in biomedical text processing. [PDF]
Windisch P +6 more
europepmc +1 more source
Evaluation of GPT-5.2 for melanoma detection across skin tones. [PDF]
Frederickson KL, Adunyah SE, Wang Q.
europepmc +1 more source
A Comparative Analysis of AI-Language Models' MCQ Performance versus Medical Students Across Different Pediatric Topics. [PDF]
Bolgova O +4 more
europepmc +1 more source
Large language models pass a standard three-party Turing test. [PDF]
Jones CR, Bergen BK.
europepmc +1 more source
AI agents are sensitive to nudges. [PDF]
Cherep M, Maes P, Singh N.
europepmc +1 more source
Using GPT-4 to annotate the severity of all phenotypic abnormalities within the human phenotype ontology. [PDF]
Murphy KB, Schilder BM, Skene NG.
europepmc +1 more source
Automated Approaches of Text Simplification of Patient Education Materials: Scoping Review. [PDF]
Krenn C +6 more
europepmc +1 more source

