科研速览 · Science Skim继续刷下去 · Keep skimming →
◆ Cognitive Computation2026-03-13· Agency (philosophy)

LLM Alignment should go beyond Harmlessness–Helpfulness and incorporate Human Agency

Usman Naseem, Tanmoy Chakraborty, Kai-Wei Chang, Mark Dras, Preslav Nakov, Nanyun Peng, Soujanya Poria

原始摘要(英文原文)· Original abstract
Abstract Large Language Models are transforming communication, research, and decision-making, but misalignment – when models diverge from human values, safety requirements, or user intent – poses serious risks. In this position paper, we argue that many alignment failures stem from operational choices in training and deployment. We posit that alignment should shift from static, post-training constraints toward dynamic, participatory approaches that safeguard pluralism, autonomy, and human flourishing. We outline forward-looking directions, including pluralistic evaluation, transparency, and the Flourishing–Justice–Autonomy (FJA) framework, and present a roadmap for advancing alignment research and practice.
读原文 · Read the paper ↗

AI 追问PRO

登录后使用 AI 追问

讨论区

登录后参与讨论

相关论文 · Related

LLM Alignment should go beyond Harmlessness–Helpfulness and incorporate Human Agency — 科研速览 Science Skim