Language over Content: Tracing Cultural Understanding in Multilingual Large Language Models

2510.16565v1 cs.CL, cs.AI, cs.LG 2025-10-22

Авторы:

Seungho Cho, Changgeon Ko, Eui Jun Hwang, Junmyeong Lee, Huije Lee, Jong C. Park

Abstract

Large language models (LLMs) are increasingly used across diverse cultural contexts, making accurate cultural understanding essential. Prior evaluations have mostly focused on output-level performance, obscuring the factors that drive differences in responses, while studies using circuit analysis have covered few languages and rarely focused on culture. In this work, we trace LLMs' internal cultural understanding mechanisms by measuring activation path overlaps when answering semantically equivalent questions under two conditions: varying the target country while fixing the question language, and varying the question language while fixing the country. We also use same-language country pairs to disentangle language from cultural aspects. Results show that internal paths overlap more for same-language, cross-country questions than for cross-language, same-country questions, indicating strong language-specific patterns. Notably, the South Korea-North Korea pair exhibits low overlap and high variability, showing that linguistic similarity does not guarantee aligned internal representation.

Ссылки и действия

Читать на arXiv Скачать PDF

Дополнительные ресурсы:

Language over Content: Tracing Cultural Understanding in Multilingual Large Language Models

Авторы:

Abstract

Ссылки и действия

Связанные статьи

LYNX: Learning Dynamic Exits for Confidence-Controlled Reasoning

To Think or Not to Think: The Hidden Cost of Meta-Training with Excessive CoT Ex...

Arbitrage: Efficient Reasoning via Advantage-Aware Speculation

Structured Document Translation via Format Reinforcement Learning

Principled RL for Diffusion LLMs Emerges from a Sequence-Level Perspective

Навигация