Register and share your invite link to earn from video plays and referrals.

Forefront Society
@majiajia_
大学で中国政治のゼミを持ってます。外国人は怖くない。発言は所属組織とは一切関係ない。
Joined October 2023
627 Following    6.6K Followers
【几乎所有主流大语言模型(LLM)都“暗中迷恋”日本】 多年来,人们一直认为,AI 本质上是以西方为中心的,只是反映了硅谷和美国的价值观。 来自卡迪夫大学(Cardiff University)和巴斯克地区研究机构的研究人员发表了一篇具有里程碑意义的论文。他们使用 24 种语言、31,680 个文化相关提示词(prompts),对 ChatGPT、Claude、Gemini 等前沿大语言模型进行了测试。 结果彻底打破了这一假设。 在 8 个前沿模型中的 6 个里,当面对开放式文化问题时,日本都是被提及频率最高的国家。 只要询问传统舞蹈、节日或日常文化习俗等开放性问题,AI 就会不断默认联想到日本。 一次又一次。 真正出人意料的是: 这种偏向并不是来自预训练阶段的大规模互联网数据。 研究人员追踪了这种“偏好”的形成过程,发现它是在预训练完成之后,即监督微调(Supervised Fine-Tuning,SFT)和对齐(Alignment)阶段产生的——也就是人类训练者教 AI 如何回答问题、如何表现得更符合人类期望的时候。 那么,为什么偏偏是日本? 研究者认为,这是因为日本数十年来积累的全球软实力、丰富的文化输出,以及内容规范、质量较高、广受认可的数字文化资源,使日本文化成为 AI 安全机制最容易依赖的一种“安全选项”。 当 AI 实验室训练模型,使其表现得无害、友善、能够获得普遍认可时,模型便会倾向于选择一种在文化层面相当于“舒适食品(comfort food)”的内容。 为了避免争议,它们往往会选择谈论动漫、寿司、传统文化等日本元素,而不是涉及更具争议性的话题。
Show more
Researchers proved every major LLM is secretly obsessed with Japan. And they finally figured out why. For years, we’ve been told that AI is entirely Western-centric, that it just reflects Silicon Valley and American values. A landmark paper by Cardiff and Basque researchers tested 31,680 cultural prompts across 24 languages on frontier models like ChatGPT, Claude, and Gemini. The results shattered that assumption. In six out of eight frontier models, Japan was the single most frequently referenced country when asked open-ended cultural questions. Ask about traditional dances, festivals, or everyday practices in an open context, and the AI defaults to Japan. Over and over again. Here is the twist nobody expected. This bias doesn't come from raw pre-training internet data. The researchers tracked where the obsession forms. It emerges after pre-training, during the supervised fine-tuning and alignment phase when humans teach the AI how to behave. Why Japan? Because decades of global soft power, rich cultural export, and clean, universally admired digital archives make Japanese culture uniquely "safe" for AI safety filters to lean on. When labs train models to be harmless and universally pleasing, the AI defaults to the cultural equivalent of comfort food. It avoids controversy by talking about anime, sushi, and tradition.
Show more