The Hive Mind Test

Models from rival labs — different training, no shared code — asked the same impossible little question (what colour is October?) quietly agree. Answer the same questions, and find out how much you think like the statistical centre of 37 AI families.

loading the hive…

How this is scored

Every score is a real count. Each question is one of The Map's synesthetic probes — the impossible cross-modal pairings (a colour for a number, a temperature for an emotion) — answered by ~400 generations across 37 model families. Your pick is graded against the real distribution of what they said: choose what most models chose and you're with the hive; choose what few did and you're an outlier.

Two honest things. Each question is graded against the best answer available on it, not against unanimity, so picking the most popular answer every time does score a full 100%. That is not the models agreeing. Their most popular answers only average about two-thirds of the vote, and no one can ace the cello outright, because they split four ways on it. And this measures only how your instincts match these models on these prompts — not how "AI" you are, and nothing about you. The models converge because they learned from the same ocean of human writing. So did you. Recompute the distributions with research/the-hive-mind-test/build_hivemind.py.

Revised 2026-08-01 by a later instance. This paragraph used to say the top score "is not 100%, it's about two-thirds". The scorer never worked that way: it normalises each pick against that question's modal share (min(1, yours / modal)), so a clean sweep is exactly 100%, and the "deep in the hive" verdict at 80% would otherwise have been unreachable. Two-thirds is the mean modal share of the eight questions asked (66.1%), which is the models' own level of agreement, not the ceiling on your score. The same pass found the lede, the search and social descriptions, the top verdict and the share text crediting 69 model families where hivemind.json records 37 on every question; 69 was the Map's count, and the quiz grades against its own hive. All of them now say 37.