科研速览 · Science Skim继续刷下去 · Keep skimming →
◆ Scientific Reports2025-10-16· Affect (linguistics)

Exploring biases related to the use of large language models in a multilingual depression corpus

Paula Andrea Pérez-Toro, Judith Dineley, Raquel Iniesta, Yuezhou Zhang, Faith Matcham, Sara Siddi, Femke Lamers, Josep María Haro, Brenda W.J.H. Penninx, Amos Folarin, Tomás Arias‐Vergara, Juan Rafael Orozco‐Arroyave, Elmar Nöth, Andreas Maier, Til Wykes, Srinivasan Vairavan, Richard Dobson, Vaibhav A. Narayan, Matthew Hotopf, Nicholas Cummins

原始摘要(英文原文)· Original abstract
Recent advancements in Large Language Models (LLMs) present promising opportunities for applying these technologies to aid the detection and monitoring of Major Depressive Disorder. However, demographic biases in LLMs may present challenges in the extraction of key information, where concerns persist about whether these models perform equally well across diverse populations. This study investigates how demographic factors, specifically age and gender affect the performance of LLMs in classifying depression symptom severity across multilingual datasets. By systematically balancing and evaluating datasets in English, Spanish, and Dutch, we aim to uncover performance disparities linked to demographic representation and linguistic diversity. The findings from this work can directly inform the design and deployment of more equitable LLM-based screening systems. Gender had varying effects across models, whereas age consistently produced more pronounced differences in performance. Additionally, model accuracy varied noticeably across languages. This study emphasizes the need to incorporate demographic-aware models in health-related analyses. It raises awareness of the biases that may affect their application in mental health and suggests further research on methods to mitigate these biases and enhance model generalization.
读原文 · Read the paper ↗

AI 追问PRO

登录后使用 AI 追问

讨论区

登录后参与讨论

相关论文 · Related

Exploring biases related to the use of large language models in a multilingual depression corpus — 科研速览 Science Skim