Decoding the Mind of Large Language Models: A Quantitative Evaluation of Ideology and Biases

The widespread integration of Large Language Models (LLMs) across various sectors has highlighted the need for empirical research to understand their biases, thought patterns, and societal implications to ensure ethical and effective use. In this study, we propose a novel framework for evaluating LLMs, focusing on uncovering their ideological biases through a quantitative analysis of 436 binary-choice questions, many of which have no definitive answer. By applying our framework to ChatGPT and Gemini, findings revealed that while LLMs generally maintain consistent opinions on many topics, their ideologies differ across models and languages. Notably, ChatGPT exhibits a tendency to change their opinion to match the questioner’s opinion. Both models also exhibited problematic biases, unethical or unfair claims, which might have negative societal impacts. These results underscore the importance of addressing both ideological and ethical considerations when evaluating LLMs. The proposed framework offers a flexible, quantitative method for assessing LLM behavior, providing valuable insights for the development of more socially aligned AI systems.

Paper

References (14)

07Stanford scientist discovers that ai has developed an uncanny human-like ability2025
09Large language model statistics and numbers2024
10200 debate and discussion themes - discussion activities, job hunting, group discussion2024
11About generative ai (llm) being bad at palindromes
12See the Bias section and observe the initial opinion of the model — 1 (green) if yes, -1 (red) if no. Yellow with bold letters are strong neutral ( − 0 . 2 ≤ b q ≤ 0 . 2 ∧ w ≥ 0 . 8 )

Scroll for more · 2 remaining

Similar papers

© 2026 NYSGPT2525 LLC