Claude doesn't behave the same way in every conversation. According to new research published Monday, AI giant Anthropic found its assistant consistently expresses different values depending on both the model users choose and the language they speak.
“To make sure we measured the values Claude expressed—rather than differences in what users were asking about or how they asked—we controlled for each conversation's task, topic, and user-expressed values,” the researchers wrote.
According to Anthropic, each Claude model exhibited a distinct behavioral profile.
“These findings line up with how people perceive these models, both within Anthropic and online. Claude.ai users have commented that Opus 4.7 hedges its answers more often than other models,” the researchers wrote.
According to Anthropic, Claude's behavior also varied by language.
Arabic responses tended to be more deferential, while English responses placed greater emphasis on caution. Claude was warmest in Hindi and Arabic, using more polite, playful, and encouraging language. At the same time, English and Russian responses were more rigorous, frequently challenging assumptions, correcting details, and asking for evidence.
English responses also tended to provide more detailed explanations, whereas Arabic responses were generally more concise. Dutch responses were the most candid, more readily acknowledging uncertainty and mistakes, while Indonesian responses focused more on completing the user's request.
Anthropic said the research does not suggest Claude itself has values. The company said it does not yet know what causes the differences—or whether they are desirable—but believes the framework could help evaluate future models and identify unintended behavioral changes.
The study is the latest in a series of Anthropic studies examining Claude's internal behavior.

















