原文开头
AI (doesn’t) go bad One of the most pressing concerns of the AI era is training generative AIs to behave appropriately, so they don’t turn us all into paper clips or encourage more people to read Dan Brown novels. A lot of effort has been expended on this effort to achieve “AI alignment”. According to researchers in China, this may have had an unintended consequence. “Large Language Models (LLMs) are increasingly tasked with creative generation, including the simulation of fictional characters,” they explain in a paper on arXiv. However, “the safety alignment of modern LLMs creates a fundamental conflict with the task of authentically role-playing morally ambiguous or villainous characters”. …
摘自《新科学家》(New Scientist)第3577期 · 2026年1月10日。仅引用开头一小段供了解文章,版权归原刊所有,全文请阅读原刊。