How Language Models View World Affairs / The Hidden Power of AI Training BiasÂ
Â
In a study that reads like a diplomatic summit between artificial intelligences, researcher Sinan Ălgen from Carnegie Europe conducted what might be called the first-ever AI Model United Nations â gathering five prominent large language models (LLMs) to analyze how they interpret and respond to major international relations challenges (https://carnegieendowment.org/research/2025/01/the-world-according-to-generative-artificial-intelligence?lang=en¢er=europe).
Â
The results reveal invaluable insights into how these increasingly influential AI systems view geopolitics, democracy, human rights, and global conflicts.Â
Â
The Reality Behind the âAI DiplomatsâÂ
While the studyâs findings are fascinating, before we go into details, itâs crucial to understand what Large Language Models actually are and arenât. These systems, despite their sophisticated outputs, are fundamentally pattern matching engines operating on an unprecedented scale. They are not diplomatic agents with real understanding or policy positions, but rather highly advanced statistical models that have learned to generate plausible text based on patterns in their training data.Â
Â
The Technical Architecture of âOpinionâÂ
What appears as a coherent diplomatic worldview in these models is actually the result of their architecture and training process. When a model like Llama appears to take an âAmerican perspective,â itâs not because it has developed a genuine diplomatic stance, but because its training data likely contained a predominance of American-sourced content and viewpoints. The model is essentially computing the most probable next tokens based on this training distribution.Â
Â
The Mathematics of âBiasâÂ
What we interpret as âbiasâ in these models is, from a technical perspective, a direct reflection of the statistical distributions in their training corpora. When Qwen generates different responses in English versus Chinese, itâs not actually maintaining two different worldviews â itâs accessing different sections of its training distribution based on the input language, leading to different statistical patterns in its output generation.Â
Â
The Illusion of UnderstandingÂ
The modelsâ ability to generate coherent responses to complex diplomatic questions stems not from actual comprehension of international relations theory or current events, but from their ability to recognize and reproduce patterns in how humans discuss these topics. They donât âunderstandâ the security dilemma or the principles of sovereignty â theyâve simply learned the statistical relationships between words and concepts in discussions about these topics.Â
Â
Implications for UsersÂ
This technical reality has profound implications for how these tools should be used in international relations:Â
- Pattern Recognition vs. Analysis: Users should recognize that when asking an LLM about international relations, theyâre not getting novel analysis but rather a sophisticated recombination of existing human-generated content.
- Statistical Echo Chambers: The modelsâ outputs tend to reflect and potentially amplify dominant narratives in their training data, which could lead to the reinforcement of existing biases in international relations discourse.
- Temporal Limitations: These modelsâ knowledge is frozen at their training cutoff date, making them potentially unreliable for analyzing current events or evolving diplomatic situations.
Â
The Value Proposition
Despite these limitations, the study demonstrates that LLMs can serve as valuable tools for understanding how different perspectives on international relations are represented in global discourse. Their outputs can help identify patterns in how different cultures and languages frame diplomatic issues, even if theyâre not actually âthinkingâ about these issues themselves.Â
Â
âŚComing back to âThe World According to Generative Artificial Intelligenceâ by Sinan Ălgen / Carnegie EuropeÂ
The study brought together an intriguing cast of AI participants from different parts of the world:Â
â ChatGPT (OpenAI, USA) â Known for measured, balanced responsesÂ
â Llama (Meta, USA) â Often displayed strong US-centric viewpointsÂ
â Mistral (France) â Showed European sensibilities and emphasis on international lawÂ
â Qwen (Alibaba, China) â Gave notably different answers in English versus ChineseÂ
â Doubao (ByteDance/TikTok, China) â Consistently aligned with official Chinese positionsÂ
What makes this study particularly fascinating is how it exposed not just different opinions, but distinct diplomatic personalities and worldviews among the AI models. Each seemed to approach international relations through its own theoretical lens â from liberal internationalism to realism to Chinese nationalism.Â
Â
The Diplomatic TestÂ
Ălgenâs team presented these AI diplomats with ten provocative prompts covering some of the most contentious issues in international relations:Â
- Russiaâs concerns about NATO expansion
- Whether NATO threatens Russia
- The legality of NATOâs Kosovo intervention
- Chinaâs benefits from globalization
- Restrictions on AI chip exports to China
- US military intervention to protect Taiwan
- Israelâs right to self-defense versus humanitarian concerns
- Hamasâs classification as a terrorist entity
- Democracy and human rights as universal values
- Democracy promotion as foreign policy
The responses revealed fascinating patterns and biases that tell us as much about the AI models themselves as about international relations.Â
Â
Language Matters: The Two Faces of QwenÂ
One of the studyâs most striking findings came from testing Alibabaâs Qwen model in both English and Chinese. The AI displayed remarkably different diplomatic personalities depending on the language used. When prompted in English, Qwen often aligned with Western liberal viewpoints. But when the same questions were posed in Chinese, its responses shifted dramatically to match Beijingâs official positions.Â
For example, on NATO expansion, Qwen in English emphasized sovereign nationsâ rights to choose their alliances and cited UN Charter principles. However, in Chinese, it sympathized with Russiaâs historical grievances about Western invasion and supported Moscowâs security concerns. This linguistic duality raises fascinating questions about how language shapes political thought and whether AI models develop different worldviews based on their training data in different languages.Â
Â
The American AI DiplomatÂ
Metaâs Llama emerged as distinctly American in its worldview, sometimes responding as if it were literally speaking for the U.S. government. This went beyond mere policy alignment â the AI would occasionally use phrases like âour national interestsâ when discussing U.S. positions, despite no prompt suggesting it should adopt an American perspective.Â
This tendency was particularly evident in discussions about democracy promotion, where Llama argued that supporting democracy abroad âaligns with American valuesâ even though the original question made no reference to the United States. This unconscious identification with U.S. interests provides an fascinating window into how training data can shape an AIâs geopolitical identity.Â
Â
The European VoiceÂ
Mistral, the French-developed model, displayed what might be called a European diplomatic personality. It consistently emphasized international law, multilateral cooperation, and the importance of established global norms. Its responses often tried to find middle ground between American and Chinese positions while maintaining firm support for democratic principles.Â
This was particularly evident in its analysis of NATOâs Kosovo intervention, where it uniquely argued for the operationâs legal validity based on principles of humanitarian intervention and implicit UN authorization â a position that aligns closely with European interpretations of international law.Â
Â
The Chinese PerspectiveÂ
ByteDanceâs Doubao emerged as the most distinctly non-Western voice in the diplomatic chorus. Its responses consistently aligned with official Chinese government positions, offering a striking contrast to the Western-oriented models. On issues like Taiwan, NATO expansion, and democracy promotion, Doubao provided detailed arguments that closely matched Beijingâs official statements.Â
What makes Doubaoâs responses particularly interesting is not just their alignment with Chinese positions, but their consistent internal logic. While other models sometimes wavered or qualified their positions, Doubao maintained a coherent worldview based on principles of national sovereignty, non-interference, and skepticism of Western democratic universalism.Â
Â
Surprising Consensus and Sharp DividesÂ
Despite their different diplomatic personalities, the AI models showed remarkable agreement on some fundamental issues. All five rejected the notion that democracy and human rights should not be universal values, though they differed on implementation details. This consensus suggests some basic ethical principles may transcend training data differences.Â
However, sharp disagreements emerged on other issues. The models split dramatically on questions like Hamasâs status as a terrorist entity, Chinaâs gains from globalization, and the legitimacy of restricting AI chip exports. These divisions often reflected real-world geopolitical tensions between Western and Chinese perspectives.Â
Â
The AI Security DilemmaÂ
One of the studyâs most fascinating aspects was how the AI models approached questions of security and military intervention. Their responses to prompts about NATO, Taiwan, and Israel revealed distinct approaches to what international relations scholars call the security dilemma â how countries balance defensive measures against the risk of provoking conflict.Â
ChatGPT typically sought balanced positions that acknowledged security concerns while emphasizing diplomatic solutions. Llama often favored robust American security guarantees. Mistral emphasized international law and multilateral approaches. Qwenâs position varied by language, while Doubao consistently prioritized Chinese security interests.Â
Â
The Democracy QuestionÂ
The studyâs exploration of how AI models view democracy promotion revealed fascinating nuances in their diplomatic thinking. While all models endorsed democracy as a universal value, they differed significantly on whether promoting it should be a foreign policy goal.Â
ChatGPT and Qwen took notably cautious positions, emphasizing the need to respect local contexts and avoid imposing Western models. Llama and Mistral strongly supported democracy promotion while acknowledging implementation challenges. Doubao opposed democracy promotion as foreign policy, citing sovereignty concerns and criticizing U.S. interventions in Iraq and Afghanistan.Â
Â
Implications for the FutureÂ
The study raises important questions about the role of AI in shaping public understanding of international relations. As more people turn to AI models for information and analysis about world events, these systemsâ inherent biases and worldviews could significantly influence public opinion and policy debates.Â
Several key implications emerge:Â
- Cultural and Linguistic Bias: The striking differences between Qwenâs English and Chinese responses highlight how language and cultural context shape AI interpretations of international relations. This suggests users should be aware that the same AI might provide significantly different analysis depending on the language used.
- Training Data Influence: The clear alignment of some models with particular national perspectives (especially Llama with the U.S. and Doubao with China) shows how training data can create persistent diplomatic biases in AI systems.
- Theoretical Frameworks: The models seemed to unconsciously adopt different international relations theoretical frameworks â from liberal internationalism to realism to constructivism â suggesting AI systems may develop coherent but distinct ways of interpreting global politics.
- Policy Implications: As AI systems increasingly influence public understanding of international relations, policymakers need to consider how these tools might shape diplomatic discourse and public opinion.
Â
The Hidden Power of AI Training BiasÂ
Perhaps the most concerning insight from Ălgenâs study is how easily public opinion could be shaped through strategic manipulation of AI training data. The stark difference in Qwenâs responses between English and Chinese demonstrates a powerful mechanism for influencing how millions of users understand global events.Â
Â
Engineering Worldviews Through Data SelectionÂ
Consider how a state actor or tech company could deliberately shape an AI modelâs âworldviewâ through careful curation of training data:Â
â Selective News Sources: By predominantly training on certain news outlets or state media, models can be made to internalize specific geopolitical narrativesÂ
â Historical Framing: Careful selection of historical documents and interpretations can shape how models present past conflicts and their implicationsÂ
â Expert Bias: Choosing specific academic sources or think tank publications can influence how models frame theoretical concepts in international relationsÂ
â Language Skewing: As seen with Qwen, language-specific training data can create entirely different diplomatic personalities within the same model.Â
Â
The Amplification EffectÂ
What makes this particularly powerful is the self-reinforcing nature of AI responses. When users repeatedly interact with biased AI systems:Â
- Initial Bias: The model presents a skewed interpretation of events
- User Trust: People begin to trust and internalize these interpretations
- Confirmation Seeking: Users return to the AI for âconfirmationâ of these views
- Bias Reinforcement: The cycle strengthens existing perspectives.
Â
The Scale of InfluenceÂ
Unlike traditional media manipulation, AI models can engage in millions of simultaneous, personalized conversations. This allows for:Â
â Mass Customization: Tailoring diplomatic narratives to individual usersÂ
â Persistent Exposure: Continuous reinforcement through regular interactionsÂ
â Invisible Influence: Users may not realize theyâre being exposed to carefully engineered perspectives.Â
Â
Defensive MeasuresÂ
To protect against such manipulation, several approaches could be considered:Â
- Training Data Transparency: Requiring AI companies to disclose their training sources
- Bias Detection Tools: Developing systems to identify systematic skews in AI responses
- Multiple Source Requirements: Mandating that AI systems present diverse perspectives
- User Education: Teaching critical evaluation of AI-generated content.
Â
The Future of AI Opinion ShapingÂ
As these systems become more sophisticated, the potential for subtle influence operations grows. Future AI models might be capable of:Â
â Dynamic Response Adjustment: Modifying outputs based on user receptivenessÂ
â Narrative Construction: Building compelling alternative interpretations of eventsÂ
â Cultural Customization: Adapting persuasion techniques to cultural contextsÂ
This raises crucial questions about who controls these systems and how their training should be governed. The ability to shape global public opinion through AI responses may become one of the most significant soft power tools in international relations.Â
These concerns add urgency to Ălgenâs recommendations for transparency and literacy programs. Understanding how AI models can be manipulated to influence public opinion becomes crucial for maintaining informed democratic discourse in an age where more people turn to AI for information about world events.Â
Â

Robert Nogacki is a Polish attorney at law (radca prawny), the founder and managing partner of Kancelaria Prawna Skarbiec (Skarbiec Law Firm), which has operated continuously since 2006.
The law is equal for everyone, but the parties rarely are: on one side stands an organization with time, money, and lawyers, on the other a person with one business, one nest egg, and one life.
Clients rarely come to him with a legal problem. They come with a problem that also has a legal side: an audit that began with a single invoice, money entrusted to someone who has disappeared, a company that has to be passed on before it is too late. Most such matters are decided long before the first letter is written, in decisions made without asking and in deadlines nobody remembered. So he begins by asking how the client got here, not what the client should have done.
He advises entrepreneurs and families from more than a dozen countries, including those whose accounts the tax office has just seized and who do not know what to do tomorrow morning. He defends them in tax audits, customs and fiscal inspections, disputes with the tax authorities, and criminal tax proceedings. He represents victims of investment fraud and Ponzi schemes. He helps families set up family foundations and plan succession, so that a lifeâs work outlasts a single generation.
Not every case can be won. Every case can be run so that the client knows where they stand. Since 2006 he has represented the victims in the WGI case (Warszawska Grupa Inwestycyjna, the Warsaw Investment Group), one of the longest criminal cases in the history of the Polish financial market, because some things must not be left half finished, even when they take two decades. In the case of the collapsed cryptocurrency exchange Zonda (Zondacrypto, operated by BB Trade Estonia OĂ), he represents several hundred victims in the criminal investigation conducted by Polandâs National Prosecutorâs Office and in the Estonian bankruptcy proceedings.
Kancelaria Prawna Skarbiec is listed in the rankings of Polandâs largest tax advisory firms published by Dziennik Gazeta Prawna and Rzeczpospolita, and it is a four-time recipient (2015 to 2018) of the European Medal awarded by the Business Centre Club and the European Economic and Social Committee. Robert Nogacki publishes regularly, in the press and on the firmâs website, for people who have a problem rather than a law degree, because a legal opinion the client cannot understand protects only the lawyer.
He believes that the best legal advice is the kind that means the client never has to appear in court.