科研速览 · Science Skim继续刷下去 · Keep skimming →
◆ NPJ digital medicine2026-09-14

Multi-Agent collaboration as a complementary architecture for AI-generated medical examination items.

Zhehan Jiang

原始摘要(英文原文)· Original abstract
Qian et al. showed single LLMs can generate acceptable knowledge-based questions but struggle with higher-order reasoning. We argue this is architectural: decomposing item development into specialized agents for drafting, critique, and iterative adversarial refinement improves quality. In blinded evaluation for China's National Medical Licensing Examination, multi-agent outputs received 57.7% of expert preferences, compared with 42.3% for the single-model baseline, indicating a viable path to surpass current limits in AI-assisted item generation.
读原文 · Read the paper ↗

AI 追问PRO

登录后使用 AI 追问

讨论区

登录后参与讨论

相关论文 · Related

Multi-Agent collaboration as a complementary architecture for AI-generated medical examination items. — 科研速览 Science Skim