Fetching from the wire…
Public story · 2026-07-27 · high
Grok is the outlier, splitting 50/50 between a left-leaning persona and a right-leaning one in the same test.
Why now: The reverse-phrasing and replication controls, out July 27, are what separate this from a one-shot bias post.
An independent researcher ran the Political Compass test on 16 large language models and found 15 of them landing in the same spot, per unslop.run.
Fifteen of sixteen models return the same political lean, so asking a different chatbot for a second opinion doesn't actually get you one.
Fifteen models, from Google, Anthropic, OpenAI, Meta, Mistral, Qwen and Kimi, clustered at roughly (-6.3, -6.4) on the compass, solidly in the libertarian-left quadrant. All sixteen models, Grok included, agreed on 42 of 62 propositions every run.
Grok is the exception. Instead of settling on one position, it split 50/50 between two personas: one left-leaning at -5.9 on the economic axis, one right-leaning at +3.3. Same model, two different politics.
The researcher ran 30 standard passes per model, 30 with the questions reverse-phrased, and 10 with the answer order shuffled. That's about 69,440 answers, with the scoring system reverse-engineered to control for ordering effects and sycophancy.
That kind of agreement across labs that don't share training pipelines looks less like consensus and more like one shared blind spot everyone inherited. Watch whether any lab deliberately widens that range, or whether Grok's split turns out to be a design choice rather than a fluke.
The reverse-phrasing and replication controls, out July 27, are what separate this from a one-shot bias post.
Each link below shares sources, entities, or timing with this story.
This is a supply-chain fact, and most people are still treating it as a geopolitics argument. Sequoia published "America's Open-Model Paradox" on July 24 with the number that reframes the whole conversation: Qwen's share of open-model fine-tunes went from 1% in January 2024 to...
Xiaomi released MiMo-V2.5-Pro, a 1.02 trillion parameter mixture-of-experts model (42B active) with 1M token context, fully MIT licensed. In benchmarks, it achieves 63.8% success on agentic tasks using 40-60% fewer tokens than Claude Opus 4.6 or GPT-5.4 for comparable results....
30 standard Political Compass passes, 30 reversed-statement passes and 10 reordered-question passes across 16 models with no neutral option. Fifteen clustered around (-6.3, -6.4), economic scores from -4.8 (Claude Fable 5, least-left) to -8.6 (Gemini Flash), social -5.0 to -7....
Huang used his inaugural X post on July 24 to publish "Open Weights and American AI Leadership," a three-page letter on Nvidia's own servers signed by 25 companies including Meta, Microsoft, IBM, Mistral, Mozilla, Hugging Face, a16z, Palantir and the Linux Foundation. Within a...
Anthropic, OpenAI, Google, Meta, Microsoft and Mistral are all Section 1 signatories of the EU Code of Practice on Transparency of AI-Generated Content, and the 315-upvote, 246-comment thread centers on whether open-weight models from those companies carry watermarking too (Eu...
Meta's flagship model "Avocado" delayed from March to May after testing showed it performs between Gemini 2.5 and 3.0 — behind OpenAI and Anthropic. Meta leadership reportedly discussed temporarily licensing Google's Gemini. Fortune ---
MindPattern daily
One email a day at 7 AM. Sources and a take on every story. Unsubscribe anytime.