科研速览 · Science Skim继续刷下去 · Keep skimming →
◇ Open MIND2026-08-01· Affect (linguistics)

Does the quality of thinking behind data selection affect the model's "character"?

Katja Gorlinski

原始摘要(英文原文)· Original abstract
Does the quality of thinking behind data selection affect a model's "character"? We train the same model on four different sets of texts, all drawn from one shared pool, and give every trained copy the same exam. The only thing that changes between the sets is how the texts were chosen: at random, by an attentive human with no procedure, and by a second person following a fixed written protocol. A fourth, deliberately flattering set checks that the measurement works at all. Forty training runs. Every criterion is fixed before the first run, and a null result is published with the same precision as a success. The full frozen document, the figures and the budget are in the attached files.
读原文 · Read the paper ↗

AI 追问PRO

登录后使用 AI 追问

讨论区

登录后参与讨论

相关论文 · Related

Does the quality of thinking behind data selection affect the model's "character"? — 科研速览 Science Skim