AI & Machine Intelligence
How machine reasoning is measured, where it beats people, where it still fails, and what any of it says about human intelligence.

Has AI Passed the Turing Test? What the 2026 PNAS Study Found
A peer-reviewed PNAS paper says GPT-4.5, given a humanlike persona, was picked as the human in 73% of five-minute chats. Without the persona it managed 36%, and the authors say the test measures humanlikeness, not intelligence.

IQ Checkers for X Accounts: What They Actually Estimate
Sites and bots that promise an instant IQ score for any X (formerly Twitter) username read public profile data, not cognitive ability. Here is exactly what the number is built from.

Plausible Nonsense: Only 4 of 60 AI Models Cleared a Reasoning Benchmark
A PNAS paper published on 9 September 2026 tested 60 large language models against human deliberation. Only four consistently cleared its benchmark, while the rest could still sound reasonable. Here is what the abstract does and does not establish.

What the Artificial Analysis Intelligence Index Actually Measures
Headlines call it an AI intelligence score. It is a weighted average of ten separate evaluations, and the hardest one — an exam built to resist quick solving — is already close to 60% solved.

Is There a g Factor for AI Models? What the Benchmark Correlations Show
Language models that do well on one benchmark tend to do well on the others, and published psychometric analyses have extracted a dominant common factor from those scores. Whether that factor is anything like human g depends on a question about what is varying between the models.
Two Scores, One Model: How the Harness Decides an ARC-AGI-3 Result
GPT-6 Astra scored 62.7 per cent and 99.9 per cent on the same benchmark in the same week. The difference is not the model. It is how the test was administered.
AI-Generated IQ Test Questions: What Changes for a Score
Machines have been writing reasoning items since long before the current wave of language models, and the research is clear about where the difficulty lies. Generating a plausible question is the easy half. Knowing how hard it is remains the expensive half.
Can AI Pass an IQ Test? On ARC-AGI-3, Humans Score 100% and AI 0.51%
Frontier systems now clear reasoning benchmarks that defeated them a year ago. Put them in front of puzzles screened so that ordinary people solve every one, and the best models finish under one per cent.
AI and Memory: What Cognitive Offloading Research Shows So Far
The sturdiest results are narrow and decades old. The 2025 studies that drew the headlines rest on self-reports and small preprints — and none of them measured intelligence at all.
When a Model "Scores 120", That Number Is Not an IQ
IQ scales are defined by a human reference sample, and a language model is not in it. Placing a machine on that scale is not a hard measurement problem — it is a category error.
Written to be checked, not just clicked
Every number carries its scale
An IQ figure is meaningless without the scale it was measured on and the range around it. Every story here reports all three, or says plainly that the source did not.
Sources are listed, not implied
Each story ends with the studies, datasets and documents it draws on, named and attributed, so you can go and read them yourself.
We correct in public
Spotted an error? Write to corrections@iqmetrics.org. Corrections are made on the story and noted at the bottom of it — never quietly.
Read the research.
Then find your own number.
Our IIF-certified assessment reports your score with its scale, percentile and confidence range — and a breakdown of the cognitive domains behind it.
Start IQ Test →
