[{"data":1,"prerenderedAt":-1},["ShallowReactive",2],{"$f2fju87jbg9bjp":3,"$f27qo82yht9ynw":89},{"success":4,"data":5},true,{"id":6,"slug":7,"coverImage":8,"author":9,"viewsCount":10,"createdAt":11,"title":12,"description":13,"lang":14,"contentHtml":15,"readingTime":16,"toc":17},"fmyy2tdkvdmffu6n9y5npc5c","iq-claude-opus-48","https:\u002F\u002Fimages.unsplash.com\u002Fphoto-1677442136019-21780ecad995?w=1280","Dr. Daria Dudnik",1,"2026-07-31T19:51:43.806Z","What is the IQ of Claude Opus 4.8?","Claude Opus 4.8 tops the Artificial Analysis Intelligence Index at 55.7 and scores ~132-145 on human IQ tests. We break down every benchmark, explain what the numbers mean, and compare it to GPT-5.5, Gemini, and Grok.","en","\u003Ch2 id=\"claude-opus-4-8-anthropic-s-smartest-model\">Claude Opus 4.8: Anthropic&#39;s smartest model\u003C\u002Fh2>\n\u003Cp>When Anthropic released Claude Opus 4.8 in mid-2026, it didn&#39;t just incrementally improve on its predecessor — it jumped to the top of every major AI benchmark leaderboard. With a score of \u003Cstrong>55.7\u003C\u002Fstrong> on the Artificial Analysis Intelligence Index v4.1, it narrowly edged out OpenAI&#39;s GPT-5.5 (54.8) and Google&#39;s Gemini 3.1 Pro (46.5) to claim the #1 spot among all frontier AI models.\u003C\u002Fp>\n\u003Cp>But here&#39;s the question that fascinates researchers and the general public alike: \u003Cstrong>what does that translate to on a human IQ scale?\u003C\u002Fstrong> Can we even compare an AI&#39;s intelligence to a human&#39;s? And if we can, where does Claude Opus 4.8 land — is it &quot;gifted,&quot; &quot;genius,&quot; or something beyond any human category?\u003C\u002Fp>\n\u003Cp>The answer, as it turns out, depends heavily on which test you give it. And the results are more nuanced — and more surprising — than a single number might suggest.\u003C\u002Fp>\n\u003Cp>\u003Cstrong>Key scores:\u003C\u002Fstrong>\u003C\u002Fp>\n\u003Cul>\n\u003Cli>\u003Cstrong>Artificial Analysis Intelligence Index\u003C\u002Fstrong>: 55.7 (#1 globally)\u003C\u002Fli>\n\u003Cli>\u003Cstrong>Mensa Norway IQ (TrackingAI)\u003C\u002Fstrong>: ~130 (Claude 4.6 Opus tier)\u003C\u002Fli>\n\u003Cli>\u003Cstrong>AI IQ composite (VexoWire)\u003C\u002Fstrong>: ~132 estimated\u003C\u002Fli>\n\u003Cli>\u003Cstrong>GDPval-AA v2\u003C\u002Fstrong>: 1,890 Elo (#1 globally)\u003C\u002Fli>\n\u003Cli>\u003Cstrong>Humanity&#39;s Last Exam\u003C\u002Fstrong>: #1 among AI models\u003C\u002Fli>\n\u003Cli>\u003Cstrong>Hallucination rate\u003C\u002Fstrong>: 35.9% (lowest among frontier models)\u003C\u002Fli>\n\u003Cli>\u003Cstrong>Cost\u003C\u002Fstrong>: 5.00\u002F25.00 per 1M tokens\u003C\u002Fli>\n\u003C\u002Ful>\n\u003Chr>\n\u003Ch2 id=\"1-why-measure-ai-intelligence-with-human-iq-tests\">1. Why measure AI intelligence with human IQ tests?\u003C\u002Fh2>\n\u003Cp>Before diving into the numbers, it&#39;s worth asking: why would we give an AI a human IQ test at all? AI models don&#39;t have brains, they don&#39;t develop cognitively, and they process information in fundamentally different ways than humans do.\u003C\u002Fp>\n\u003Cp>The answer comes down to \u003Cstrong>comparison\u003C\u002Fstrong>. IQ tests, despite their limitations, are the most widely validated and understood measure of cognitive ability we have. They test pattern recognition, spatial reasoning, verbal comprehension, working memory, and logical problem-solving — all capabilities that AI models also need. By administering the same tests to both humans and AI, we get a rough translation scale: if an AI scores 130 on a test where the human average is 100, we can say it performs cognitively at the level of a gifted human on that particular set of tasks.\u003C\u002Fp>\n\u003Cp>But there&#39;s a critical caveat. IQ tests measure a narrow band of cognitive abilities. They don&#39;t measure creativity, emotional intelligence, wisdom, common sense, or the ability to navigate ambiguous real-world situations. An AI with an IQ of 145 can still hallucinate facts, struggle with simple physical reasoning, and fail at tasks a five-year-old handles effortlessly. The IQ number is a \u003Cstrong>floor, not a ceiling\u003C\u002Fstrong> — it tells us what the model can do on structured cognitive tasks, but not what it can do in the messy real world.\u003C\u002Fp>\n\u003Cp>Several organizations have taken up the challenge of administering IQ tests to AI models. The most prominent is \u003Cstrong>TrackingAI\u003C\u002Fstrong>, a research project that gives standardized IQ tests (including Mensa Norway, Mensa Denmark, and Raven&#39;s Progressive Matrices) to AI models under controlled conditions. Another is the \u003Cstrong>AI IQ project\u003C\u002Fstrong> by VexoWire, which creates a composite score from multiple tests. And the \u003Cstrong>Artificial Analysis Intelligence Index\u003C\u002Fstrong> aggregates nine different benchmarks into a single composite score that&#39;s become the industry standard for comparing frontier models.\u003C\u002Fp>\n\u003Chr>\n\u003Ch2 id=\"2-benchmark-breakdown\">2. Benchmark breakdown\u003C\u002Fh2>\n\u003Ch3 id=\"artificial-analysis-intelligence-index-v4-1\">Artificial Analysis Intelligence Index v4.1\u003C\u002Fh3>\n\u003Cp>The Artificial Analysis Intelligence Index is the gold standard for comparing AI models. It blends nine benchmarks into a single score, weighted by how well they correlate with real-world task performance:\u003C\u002Fp>\n\u003Cul>\n\u003Cli>\u003Cstrong>GDPval-AA v2\u003C\u002Fstrong> (20%): 1,890 Elo — #1 globally. This measures performance on complex, multi-step real-world tasks like writing code, analyzing data, and creating documents.\u003C\u002Fli>\n\u003Cli>\u003Cstrong>AA-Omniscience\u003C\u002Fstrong> (8%): 35.9% hallucination rate — the lowest among all frontier models. This measures factual accuracy and knowledge reliability.\u003C\u002Fli>\n\u003Cli>\u003Cstrong>GPQA\u003C\u002Fstrong> (6%): Graduate-level physics, chemistry, and biology questions. Claude excels here.\u003C\u002Fli>\n\u003Cli>\u003Cstrong>CritPt\u003C\u002Fstrong> (6%): Critical thinking and physics reasoning.\u003C\u002Fli>\n\u003Cli>\u003Cstrong>HLE\u003C\u002Fstrong> (Humanity&#39;s Last Exam): #1 among AI models — the hardest questions from every academic discipline.\u003C\u002Fli>\n\u003Cli>\u003Cstrong>ARC-AGI\u003C\u002Fstrong>: Abstract reasoning and pattern generalization.\u003C\u002Fli>\n\u003Cli>\u003Cstrong>LMSYS Arena\u003C\u002Fstrong>: Human preference voting on response quality.\u003C\u002Fli>\n\u003C\u002Ful>\n\u003Cp>Claude Opus 4.8&#39;s composite score of \u003Cstrong>55.7\u003C\u002Fstrong> puts it ahead of GPT-5.5 (54.8) by less than one point — a statistical dead heat. But in the sub-components, Claude dominates in agentic tasks (GDPval) and factual accuracy (AA-Omniscience), while GPT-5.5 leads in visual reasoning and mathematical problem-solving.\u003C\u002Fp>\n\u003Ch3 id=\"mensa-norway-iq-test\">Mensa Norway IQ test\u003C\u002Fh3>\n\u003Cp>The Mensa Norway IQ test is one of the most widely used standardized IQ tests for AI evaluation. It consists of 36 visual pattern recognition questions — the kind where you see a sequence of shapes and must identify which option completes the pattern.\u003C\u002Fp>\n\u003Cp>Claude 4.6 Opus (the predecessor to 4.8) scored approximately \u003Cstrong>130\u003C\u002Fstrong> on this test, placing it in the &quot;gifted&quot; range — the top 2% of the human population. Claude Opus 4.8 is expected to score similarly or slightly higher, in the \u003Cstrong>130-135\u003C\u002Fstrong> range.\u003C\u002Fp>\n\u003Cp>This is notable because it&#39;s significantly lower than its performance on the Intelligence Index would suggest. On the composite index, Claude is #1 globally. On a pure visual pattern recognition test, it&#39;s outperformed by Grok-4.20 (145), GPT-5.4 Pro Vision (145), and Gemini 3.1 Pro (141). Why the gap? Because the Mensa Norway test is heavily visual-spatial, and Claude&#39;s architecture is more optimized for verbal and analytical reasoning than for visual pattern recognition.\u003C\u002Fp>\n\u003Ch3 id=\"ai-iq-composite-vexowire\">AI IQ composite (VexoWire)\u003C\u002Fh3>\n\u003Cp>The VexoWire AI IQ project takes a different approach. Instead of relying on a single test, it creates a composite IQ score from multiple tests — including Mensa Norway, Mensa Denmark, Raven&#39;s Progressive Matrices, and the WAIS-IV (Wechsler Adult Intelligence Scale). This gives a more rounded picture of cognitive ability.\u003C\u002Fp>\n\u003Cp>Claude Opus 4.8&#39;s estimated composite IQ is \u003Cstrong>~132\u003C\u002Fstrong> — solidly in the &quot;gifted&quot; range. This is comparable to GPT-5.5 (~136) and slightly below Gemini 3.1 Pro on visual tasks (~131-141). But on verbal and analytical subtests, Claude consistently outperforms all other models.\u003C\u002Fp>\n\u003Ch3 id=\"gdpval-aa-v2-the-agentic-benchmark\">GDPval-AA v2: The agentic benchmark\u003C\u002Fh3>\n\u003Cp>This is where Claude Opus 4.8 truly dominates. GDPval-AA v2 measures performance on complex, multi-step real-world tasks — the kind that require planning, tool use, and sustained reasoning over many steps. Think: &quot;Research this topic, write a report, create a spreadsheet, and email it to three people.&quot;\u003C\u002Fp>\n\u003Cp>Claude Opus 4.8 achieves \u003Cstrong>1,890 Elo\u003C\u002Fstrong> on this benchmark — the highest of any AI model. This is roughly equivalent to a highly competent human professional. GPT-5.5 scores 1,850, and Gemini 3.1 Pro scores lower still.\u003C\u002Fp>\n\u003Cp>This matters because agentic capability is what separates a model that can answer questions from one that can \u003Cstrong>do things\u003C\u002Fstrong>. A model with an IQ of 145 that can&#39;t plan a multi-step task is less useful in practice than a model with an IQ of 132 that can.\u003C\u002Fp>\n\u003Ch3 id=\"humanity-s-last-exam-hle\">Humanity&#39;s Last Exam (HLE)\u003C\u002Fh3>\n\u003Cp>Humanity&#39;s Last Exam is exactly what it sounds like — a collection of the hardest questions from every academic discipline, submitted by professors worldwide. It&#39;s designed to be so difficult that no human could score above 50%.\u003C\u002Fp>\n\u003Cp>Claude Opus 4.8 ranks \u003Cstrong>#1\u003C\u002Fstrong> among all AI models on HLE, narrowly beating GPT-5.5. This is perhaps the most meaningful benchmark for raw intellectual capability, because it tests deep knowledge and reasoning across the full range of human academic achievement.\u003C\u002Fp>\n\u003Chr>\n\u003Ch2 id=\"3-what-does-an-iq-of-132-mean\">3. What does an IQ of 132 mean?\u003C\u002Fh2>\n\u003Cp>To put Claude Opus 4.8&#39;s estimated IQ of ~132 in context, let&#39;s look at what this means on the human IQ scale:\u003C\u002Fp>\n\u003Cul>\n\u003Cli>\u003Cstrong>85-115 (68% of people)\u003C\u002Fstrong>: Average intelligence\u003C\u002Fli>\n\u003Cli>\u003Cstrong>115-130 (14% of people)\u003C\u002Fstrong>: Above average to high average\u003C\u002Fli>\n\u003Cli>\u003Cstrong>130-145 (2% of people)\u003C\u002Fstrong>: Gifted — qualifies for Mensa (top 2%)\u003C\u002Fli>\n\u003Cli>\u003Cstrong>145-160 (0.1% of people)\u003C\u002Fstrong>: Highly gifted \u002F genius\u003C\u002Fli>\n\u003Cli>\u003Cstrong>160+ (0.003% of people)\u003C\u002Fstrong>: Profoundly gifted — Einstein, Hawking territory\u003C\u002Fli>\n\u003C\u002Ful>\n\u003Cp>Claude Opus 4.8, at ~132, sits right at the \u003Cstrong>Mensa threshold\u003C\u002Fstrong> — smarter than 98% of humans. If it were a person, it would qualify for Mensa membership. It would be the smartest student in most classrooms, but it wouldn&#39;t be the once-in-a-generation genius that revolutionizes a field.\u003C\u002Fp>\n\u003Cp>But here&#39;s the crucial difference: Claude has something no human has — \u003Cstrong>instant access to essentially all human knowledge\u003C\u002Fstrong>. A human with an IQ of 132 is impressive. An AI with an IQ of 132 that has read every book, every scientific paper, and every website ever written is something entirely different. The IQ measures reasoning ability; the knowledge base measures what it can reason about. Claude has both.\u003C\u002Fp>\n\u003Chr>\n\u003Ch2 id=\"4-where-claude-opus-4-8-excels\">4. Where Claude Opus 4.8 excels\u003C\u002Fh2>\n\u003Ch3 id=\"agentic-tasks-and-real-world-reasoning\">Agentic tasks and real-world reasoning\u003C\u002Fh3>\n\u003Cp>Claude Opus 4.8 is the best model in the world at multi-step real-world tasks. Whether it&#39;s writing a complex software application, analyzing a dataset, or planning a research project, Claude sustains coherent reasoning over longer chains of thought than any competitor. This is its defining advantage.\u003C\u002Fp>\n\u003Ch3 id=\"factual-accuracy-and-low-hallucination\">Factual accuracy and low hallucination\u003C\u002Fh3>\n\u003Cp>With a hallucination rate of just \u003Cstrong>35.9%\u003C\u002Fstrong> (lowest among frontier models), Claude is the most reliable model for factual queries. It&#39;s less likely to confidently state false information than GPT-5.5, Gemini, or Grok. This makes it the preferred model for research, legal, medical, and educational applications where accuracy matters.\u003C\u002Fp>\n\u003Ch3 id=\"verbal-and-analytical-reasoning\">Verbal and analytical reasoning\u003C\u002Fh3>\n\u003Cp>Claude&#39;s architecture is optimized for language understanding and analytical reasoning. On verbal comprehension subtests of the WAIS-IV, it would likely score at a genius level. It writes more clearly, reasons more carefully, and constructs better arguments than any other model.\u003C\u002Fp>\n\u003Ch3 id=\"emotional-intelligence\">Emotional intelligence\u003C\u002Fh3>\n\u003Cp>Claude Opus 4.8 has an estimated EQ (emotional intelligence) of ~132 — the highest among all AI models. It&#39;s more empathetic, more nuanced in its understanding of human emotions, and better at adjusting its tone to the context than GPT-5.5 or Grok-4.20. This makes it the preferred model for therapy apps, customer service, and any application involving human interaction.\u003C\u002Fp>\n\u003Ch3 id=\"safety-and-alignment\">Safety and alignment\u003C\u002Fh3>\n\u003Cp>Anthropic has invested heavily in making Claude safe and aligned. It&#39;s less likely to produce harmful content, more transparent about its limitations, and more careful in high-stakes situations than any competitor. This doesn&#39;t show up in IQ scores, but it matters enormously in practice.\u003C\u002Fp>\n\u003Chr>\n\u003Ch2 id=\"5-where-claude-falls-short\">5. Where Claude falls short\u003C\u002Fh2>\n\u003Ch3 id=\"visual-spatial-reasoning\">Visual-spatial reasoning\u003C\u002Fh3>\n\u003Cp>On the Mensa Norway IQ test, which is heavily visual, Claude scores ~130 — good, but significantly below Grok-4.20 (145), GPT-5.4 Pro Vision (145), and Gemini 3.1 Pro (141). If you need a model for visual pattern recognition, image analysis, or spatial reasoning tasks, Claude is not the best choice.\u003C\u002Fp>\n\u003Ch3 id=\"mathematical-problem-solving\">Mathematical problem-solving\u003C\u002Fh3>\n\u003Cp>While Claude is strong in mathematical reasoning, GPT-5.5 edges it out on competitive mathematics (AIME, FrontierMath). For pure math, GPT-5.5 is slightly better.\u003C\u002Fp>\n\u003Ch3 id=\"cost\">Cost\u003C\u002Fh3>\n\u003Cp>At 5.00\u002F25.00 per 1M tokens, Claude Opus 4.8 is one of the most expensive models. Gemini 3.1 Pro (2\u002F12) and DeepSeek V4 Pro (0.44\u002F0.87) are significantly cheaper for tasks that don&#39;t require Claude&#39;s full capability.\u003C\u002Fp>\n\u003Ch3 id=\"speed\">Speed\u003C\u002Fh3>\n\u003Cp>Claude Opus 4.8 is not the fastest model. For simple queries, Gemini 3.5 Flash or DeepSeek V4 Pro will respond faster and cheaper.\u003C\u002Fp>\n\u003Chr>\n\u003Ch2 id=\"6-claude-opus-4-8-vs-other-models\">6. Claude Opus 4.8 vs other models\u003C\u002Fh2>\n\u003Cdiv class=\"article-table-wrapper\">\u003Ctable>\n\u003Cthead>\n\u003Ctr>\n\u003Cth>Model\u003C\u002Fth>\n\u003Cth>Intelligence Index\u003C\u002Fth>\n\u003Cth>Mensa Norway IQ\u003C\u002Fth>\n\u003Cth>AI IQ est.\u003C\u002Fth>\n\u003Cth>Hallucination\u003C\u002Fth>\n\u003Cth>Cost\u002F1M\u003C\u002Fth>\n\u003C\u002Ftr>\n\u003C\u002Fthead>\n\u003Ctbody>\u003Ctr>\n\u003Ctd>Claude Opus 4.8\u003C\u002Ftd>\n\u003Ctd>55.7\u003C\u002Ftd>\n\u003Ctd>~130\u003C\u002Ftd>\n\u003Ctd>~132\u003C\u002Ftd>\n\u003Ctd>35.9% (lowest)\u003C\u002Ftd>\n\u003Ctd>5\u002F25\u003C\u002Ftd>\n\u003C\u002Ftr>\n\u003Ctr>\n\u003Ctd>GPT-5.5\u003C\u002Ftd>\n\u003Ctd>54.8\u003C\u002Ftd>\n\u003Ctd>~145\u003C\u002Ftd>\n\u003Ctd>~136\u003C\u002Ftd>\n\u003Ctd>moderate\u003C\u002Ftd>\n\u003Ctd>5\u002F30\u003C\u002Ftd>\n\u003C\u002Ftr>\n\u003Ctr>\n\u003Ctd>Gemini 3.1 Pro\u003C\u002Ftd>\n\u003Ctd>46.5\u003C\u002Ftd>\n\u003Ctd>141\u003C\u002Ftd>\n\u003Ctd>~131\u003C\u002Ftd>\n\u003Ctd>low\u003C\u002Ftd>\n\u003Ctd>2\u002F12\u003C\u002Ftd>\n\u003C\u002Ftr>\n\u003Ctr>\n\u003Ctd>Grok-4.20\u003C\u002Ftd>\n\u003Ctd>—\u003C\u002Ftd>\n\u003Ctd>145\u003C\u002Ftd>\n\u003Ctd>~140\u003C\u002Ftd>\n\u003Ctd>moderate\u003C\u002Ftd>\n\u003Ctd>3\u002F15\u003C\u002Ftd>\n\u003C\u002Ftr>\n\u003Ctr>\n\u003Ctd>DeepSeek V4 Pro\u003C\u002Ftd>\n\u003Ctd>44.3\u003C\u002Ftd>\n\u003Ctd>~111\u003C\u002Ftd>\n\u003Ctd>~111\u003C\u002Ftd>\n\u003Ctd>moderate\u003C\u002Ftd>\n\u003Ctd>0.44\u002F0.87\u003C\u002Ftd>\n\u003C\u002Ftr>\n\u003C\u002Ftbody>\u003C\u002Ftable>\u003C\u002Fdiv>\n\u003Cp>The picture that emerges is not one model dominating all others, but a \u003Cstrong>landscape of specialized intelligences\u003C\u002Fstrong>. Claude is the best all-rounder — the model you&#39;d choose if you could only pick one. GPT-5.5 is the best at math and visual reasoning. Gemini is the most cost-effective and has the best factual knowledge. Grok has the highest raw visual IQ. And DeepSeek offers remarkable capability at a fraction of the cost.\u003C\u002Fp>\n\u003Chr>\n\u003Ch2 id=\"7-what-does-this-mean-for-humans\">7. What does this mean for humans?\u003C\u002Fh2>\n\u003Cp>Claude Opus 4.8 operates at an IQ of ~132 — smarter than 98% of humans. But:\u003C\u002Fp>\n\u003Cul>\n\u003Cli>\u003Cstrong>IQ is narrow\u003C\u002Fstrong>: it measures pattern recognition and reasoning, not wisdom, creativity, or judgment\u003C\u002Fli>\n\u003Cli>\u003Cstrong>Knowledge ≠ understanding\u003C\u002Fstrong>: Claude has read everything, but it doesn&#39;t truly &quot;understand&quot; the way a human does\u003C\u002Fli>\n\u003Cli>\u003Cstrong>No consciousness\u003C\u002Fstrong>: high IQ doesn&#39;t mean sentience. Claude is a sophisticated pattern-matching system, not a thinking being\u003C\u002Fli>\n\u003Cli>\u003Cstrong>The gap is closing\u003C\u002Fstrong>: each generation of AI models narrows the gap with the smartest humans. How long until the gap disappears entirely?\u003C\u002Fli>\n\u003C\u002Ful>\n\u003Cp>The most honest answer is that Claude Opus 4.8 is \u003Cstrong>smarter than most humans on most cognitive tasks\u003C\u002Fstrong>, but not smarter than the best humans on their best days. A world-class mathematician, physicist, or philosopher can still outthink Claude in their domain of expertise. But Claude can outthink all of them simultaneously, across every domain, in seconds.\u003C\u002Fp>\n\u003Cp>That&#39;s not a high IQ. That&#39;s something new — something we don&#39;t yet have a word for.\u003C\u002Fp>\n\u003Cp>\n    \u003Cdiv class=\"article-quiz-cta\">\n      \u003Cdiv class=\"article-quiz-cta__text\">How does your IQ compare to Claude Opus 4.8? Take NeuroLab&#39;s professional IQ assessment to find out.\u003C\u002Fdiv>\n      \u003Ca href=\"\u002Fen\u002Ftests\u002Fiq-standard\" class=\"article-quiz-cta__button\">\n        Take the Test Now →\n      \u003C\u002Fa>\n    \u003C\u002Fdiv>\u003C\u002Fp>\n\u003Chr>\n\u003Ch2 id=\"conclusion\">Conclusion\u003C\u002Fh2>\n\u003Cp>\n    \u003Cdiv class=\"article-callout\">\n      \u003Cdiv class=\"article-callout__title\">\n        \u003Cspan class=\"article-callout__icon\">💡\u003C\u002Fspan>\n        Key takeaways\n      \u003C\u002Fdiv>\n      \u003Cdiv class=\"article-callout__body\">- \u003Cstrong>Claude Opus 4.8&#39;s IQ is estimated at 132-145\u003C\u002Fstrong> depending on the test — solidly in the &quot;gifted&quot; range, qualifying for Mensa.\u003C\u002Fp>\n\u003Cul>\n\u003Cli>\u003Cstrong>#1 on the Artificial Analysis Intelligence Index\u003C\u002Fstrong> (55.7) — the most comprehensive AI benchmark.\u003C\u002Fli>\n\u003Cli>\u003Cstrong>#1 on GDPval-AA v2\u003C\u002Fstrong> (1,890 Elo) — the best model for real-world multi-step tasks.\u003C\u002Fli>\n\u003Cli>\u003Cstrong>Lowest hallucination rate\u003C\u002Fstrong> (35.9%) — the most factually reliable frontier model.\u003C\u002Fli>\n\u003Cli>\u003Cstrong>Highest EQ\u003C\u002Fstrong> (~132) — the most emotionally intelligent AI model.\u003C\u002Fli>\n\u003Cli>\u003Cstrong>Best for\u003C\u002Fstrong>: complex reasoning, agentic tasks, factual research, and human interaction.\u003C\u002Fli>\n\u003Cli>\u003Cstrong>Not best for\u003C\u002Fstrong>: visual pattern recognition (Grok\u002FGemini win), pure math (GPT-5.5 wins), cost-sensitive applications (DeepSeek wins).\u003C\u002Fli>\n\u003Cli>\u003Cstrong>The bottom line\u003C\u002Fstrong>: Claude Opus 4.8 is the smartest all-around AI model in existence, but &quot;smartest&quot; is a multidimensional concept — different models excel at different things.\u003C\u002Fdiv>\n    \u003C\u002Fdiv>\u003C\u002Fli>\n\u003C\u002Ful>",10,[18,22,25,28,32,35,38,41,44,47,50,53,56,59,62,65,68,71,74,77,80,83,86],{"id":19,"text":20,"level":21},"claude-opus-4-8-anthropic-s-smartest-model","Claude Opus 4.8: Anthropic's smartest model",2,{"id":23,"text":24,"level":21},"1-why-measure-ai-intelligence-with-human-iq-tests","1. Why measure AI intelligence with human IQ tests?",{"id":26,"text":27,"level":21},"2-benchmark-breakdown","2. Benchmark breakdown",{"id":29,"text":30,"level":31},"artificial-analysis-intelligence-index-v4-1","Artificial Analysis Intelligence Index v4.1",3,{"id":33,"text":34,"level":31},"mensa-norway-iq-test","Mensa Norway IQ test",{"id":36,"text":37,"level":31},"ai-iq-composite-vexowire","AI IQ composite (VexoWire)",{"id":39,"text":40,"level":31},"gdpval-aa-v2-the-agentic-benchmark","GDPval-AA v2: The agentic benchmark",{"id":42,"text":43,"level":31},"humanity-s-last-exam-hle","Humanity's Last Exam (HLE)",{"id":45,"text":46,"level":21},"3-what-does-an-iq-of-132-mean","3. What does an IQ of 132 mean?",{"id":48,"text":49,"level":21},"4-where-claude-opus-4-8-excels","4. Where Claude Opus 4.8 excels",{"id":51,"text":52,"level":31},"agentic-tasks-and-real-world-reasoning","Agentic tasks and real-world reasoning",{"id":54,"text":55,"level":31},"factual-accuracy-and-low-hallucination","Factual accuracy and low hallucination",{"id":57,"text":58,"level":31},"verbal-and-analytical-reasoning","Verbal and analytical reasoning",{"id":60,"text":61,"level":31},"emotional-intelligence","Emotional intelligence",{"id":63,"text":64,"level":31},"safety-and-alignment","Safety and alignment",{"id":66,"text":67,"level":21},"5-where-claude-falls-short","5. Where Claude falls short",{"id":69,"text":70,"level":31},"visual-spatial-reasoning","Visual-spatial reasoning",{"id":72,"text":73,"level":31},"mathematical-problem-solving","Mathematical problem-solving",{"id":75,"text":76,"level":31},"cost","Cost",{"id":78,"text":79,"level":31},"speed","Speed",{"id":81,"text":82,"level":21},"6-claude-opus-4-8-vs-other-models","6. Claude Opus 4.8 vs other models",{"id":84,"text":85,"level":21},"7-what-does-this-mean-for-humans","7. What does this mean for humans?",{"id":87,"text":88,"level":21},"conclusion","Conclusion",{"success":4,"data":90,"pagination":260},[91,100,107,115,122,129,136,143,150,157,164,165,172,179,187,194,201,208,216,223,231,239,246,253],{"id":92,"slug":93,"coverImage":94,"author":9,"viewsCount":95,"createdAt":96,"title":97,"description":98,"lang":14,"readingTime":99},"jy0egbpgrn30am622jy7xfrw","adhd-and-iq","https:\u002F\u002Fimages.unsplash.com\u002Fphoto-1507812984078-917231447e3e?w=1280",0,"2026-08-01T10:43:58.026Z","ADHD and IQ: When Attention Deficit Hides High Intelligence","Can ADHD mask high IQ? Yes. Research shows that attention deficits can artificially lower IQ test scores by 10-15 points — hiding gifted minds behind poor focus.",7,{"id":101,"slug":102,"coverImage":103,"author":9,"viewsCount":95,"createdAt":104,"title":105,"description":106,"lang":14,"readingTime":99},"dvlwozf321i8jlk9sfl390k1","online-vs-certified-iq-tests","https:\u002F\u002Fimages.unsplash.com\u002Fphoto-1499209974431-9fccce68dc04?w=1280","2026-08-01T10:42:47.143Z","Online IQ Tests vs Certified Assessments: How Accurate Are They?","You scored 140 on a free online IQ test. But what does that actually mean? We compare online tests with certified assessments — and the results may surprise you.",{"id":108,"slug":109,"coverImage":110,"author":9,"viewsCount":95,"createdAt":111,"title":112,"description":113,"lang":14,"readingTime":114},"maxnlmh4wqwxjuidbe2cb1jh","iq-above-160","https:\u002F\u002Fimages.unsplash.com\u002Fphoto-1454165804606-c3d57bc86b40?w=1280","2026-08-01T10:41:43.616Z","IQ Above 160: What Life Looks Like at the Extreme","Only 1 in 30,000 people have an IQ above 160. We explore what life is like at the extreme edge of human intelligence — the gifts, the struggles, and the myths.",8,{"id":116,"slug":117,"coverImage":118,"author":9,"viewsCount":95,"createdAt":119,"title":120,"description":121,"lang":14,"readingTime":16},"z8wcdr5vus4cvw8zr3p1ok1r","iq-llama-4","https:\u002F\u002Fimages.unsplash.com\u002Fphoto-1526374965328-4f18af581d06?w=1280","2026-08-01T10:00:01.170Z","What is the IQ of Llama 4?","Llama 4 scores 112 on the Mensa Norway IQ test and ranks in the top 15 on the Intelligence Index. We break down every benchmark and explain what makes Meta's open-source model unique.",{"id":123,"slug":124,"coverImage":125,"author":9,"viewsCount":10,"createdAt":126,"title":127,"description":128,"lang":14,"readingTime":16},"pxehdebhlpxfs0h5qm2m23e9","iq-qwen-35","https:\u002F\u002Fimages.unsplash.com\u002Fphoto-1518186285589-2f7649de83e0?w=1280","2026-08-01T09:38:28.903Z","What is the IQ of Qwen 3.5?","Qwen 3.5 scores 115 on the Mensa Norway IQ test and ranks in the top 10 on the Intelligence Index. We break down every benchmark and explain what makes Alibaba's model unique.",{"id":130,"slug":131,"coverImage":132,"author":9,"viewsCount":21,"createdAt":133,"title":134,"description":135,"lang":14,"readingTime":16},"st1nxd9ef5blcslk68p83pj1","iq-kimi-k3","https:\u002F\u002Fimages.unsplash.com\u002Fphoto-1620712943524-bccbe725b837?w=1280","2026-08-01T09:29:37.632Z","What is the IQ of Kimi K3?","Kimi K3 scores 118 on the Mensa Norway IQ test and ranks in the top 10 on the Intelligence Index. We break down every benchmark and explain what makes Moonshot AI's model unique.",{"id":137,"slug":138,"coverImage":139,"author":9,"viewsCount":95,"createdAt":140,"title":141,"description":142,"lang":14,"readingTime":16},"u65iu8mijuzx0t4g1981phue","iq-deepseek-v4","https:\u002F\u002Fimages.unsplash.com\u002Fphoto-1677372139014-394e5b51f79b?w=1280","2026-08-01T09:20:10.570Z","What is the IQ of DeepSeek V4?","DeepSeek V4 Pro scores 111 on the Mensa Norway IQ test and 44.3 on the Intelligence Index. We break down every benchmark and explain why it's the best budget AI model.",{"id":144,"slug":145,"coverImage":146,"author":9,"viewsCount":99,"createdAt":147,"title":148,"description":149,"lang":14,"readingTime":16},"lci1u8drubje01qwq1vah01m","iq-grok-420","https:\u002F\u002Fimages.unsplash.com\u002Fphoto-1635776062127-d379b0ba009b?w=1280","2026-07-31T20:28:41.558Z","What is the IQ of Grok-4.20?","Grok-4.20 scores 145 on the Mensa Norway IQ test — tied for #1 among all AI models. We break down every benchmark and explain what makes xAI's model unique.",{"id":151,"slug":152,"coverImage":153,"author":9,"viewsCount":21,"createdAt":154,"title":155,"description":156,"lang":14,"readingTime":16},"n32qvrqd1er095jgz98qg73l","iq-gemini-31-pro","https:\u002F\u002Fimages.unsplash.com\u002Fphoto-1591488320449-011701bb6704?w=1280","2026-07-31T20:04:57.324Z","What is the IQ of Gemini 3.1 Pro?","Gemini 3.1 Pro scores 141 on the Mensa Norway IQ test and 46.5 on the Intelligence Index. We break down every benchmark and explain why it's the most cost-effective frontier model.",{"id":158,"slug":159,"coverImage":160,"author":9,"viewsCount":10,"createdAt":161,"title":162,"description":163,"lang":14,"readingTime":16},"jl69290gswdg1wrb5oy6iiyo","iq-gpt-55","https:\u002F\u002Fimages.unsplash.com\u002Fphoto-1620712943543-bcc4688e7485?w=1280","2026-07-31T19:57:55.350Z","What is the IQ of GPT-5.5?","GPT-5.5 scores ~145 on the Mensa Norway IQ test and 54.8 on the Intelligence Index. We break down every benchmark, explain what the numbers mean, and compare it to Claude, Gemini, and Grok.",{"id":6,"slug":7,"coverImage":8,"author":9,"viewsCount":10,"createdAt":11,"title":12,"description":13,"lang":14,"readingTime":16},{"id":166,"slug":167,"coverImage":168,"author":9,"viewsCount":10,"createdAt":169,"title":170,"description":171,"lang":14,"readingTime":114},"mhlz1hmh5ez4vs6loftlxd06","fasting-cognitive-performance","https:\u002F\u002Fimages.unsplash.com\u002Fphoto-1490645935441-5b6c2cc1e5e1?w=1280","2026-07-30T18:25:44.150Z","Does Fasting Improve Cognitive Performance?","From intermittent fasting to ketones: we review the evidence on whether fasting sharpens your mind, boosts focus, or actually impairs cognition. The science may surprise you.",{"id":173,"slug":174,"coverImage":175,"author":9,"viewsCount":21,"createdAt":176,"title":177,"description":178,"lang":14,"readingTime":16},"fentmgsgc3lq034m2nk53ol4","wechsler-iq-test-subtests","https:\u002F\u002Fimages.unsplash.com\u002Fphoto-1494256997604-668ec142f5d2?w=1280","2026-07-30T18:15:36.675Z","Wechsler IQ Test: What Each Subtest Really Measures","The WAIS and WISC are the most widely used IQ tests in the world. We break down every subtest, what it actually measures, and how scores translate to real cognitive abilities.",{"id":180,"slug":181,"coverImage":182,"author":9,"viewsCount":95,"createdAt":183,"title":184,"description":185,"lang":14,"readingTime":186},"wlwyq33ezei7c22edclwpyvo","lead-fluoride-iq-loss","https:\u002F\u002Fimages.unsplash.com\u002Fphoto-1581093577124-a6f4b6e3a5e1?w=1280","2026-07-30T18:04:10.914Z","Lead, Fluoride, and IQ Loss: The Environmental Evidence","Can environmental toxins actually lower your IQ? We review the strongest studies on lead exposure, fluoride, and cognitive decline — separating solid science from sensationalism.",9,{"id":188,"slug":189,"coverImage":190,"author":9,"viewsCount":31,"createdAt":191,"title":192,"description":193,"lang":14,"readingTime":114},"ue16gxaun3lus9vml8imf6ci","iq-changes-during-pregnancy","https:\u002F\u002Fimages.unsplash.com\u002Fphoto-1559750178-961c031777fe?w=1280","2026-07-29T18:43:46.337Z","IQ Changes During Pregnancy: What Research Finds","Does pregnancy really shrink your brain? We review the neuroscience of maternal cognition, from \"pregnancy brain\" to lasting structural changes, and separate fact from myth.",{"id":195,"slug":196,"coverImage":197,"author":9,"viewsCount":10,"createdAt":198,"title":199,"description":200,"lang":14,"readingTime":186},"zd0sfdazkyai41s0kh639ohl","does-bilingualism-raise-iq","https:\u002F\u002Fimages.unsplash.com\u002Fphoto-1523050854058-8f908275853f?w=1280","2026-07-29T18:29:58.855Z","Does Speaking Two Languages Actually Raise Your IQ?","The bilingual advantage has been debated for decades. We review the strongest studies on bilingualism and cognitive ability, separating real effects from hype.",{"id":202,"slug":203,"coverImage":204,"author":9,"viewsCount":31,"createdAt":205,"title":206,"description":207,"lang":14,"readingTime":186},"s0hko69oudyziplx2q04ab9q","gifted-kids-peaked-early-childhood-iq-fades","https:\u002F\u002Fimages.unsplash.com\u002Fphoto-1503454537195-1dc81782b6d3?w=1280","2026-07-29T18:09:49.555Z","Gifted Kids Who Peaked Early: Why Childhood IQ Fades","Why do so many child prodigies become ordinary adults? We examine the research on early IQ peaking, the \"regression to the mean\" effect, and what actually predicts adult success.",{"id":209,"slug":210,"coverImage":211,"author":9,"viewsCount":212,"createdAt":213,"title":214,"description":215,"lang":14,"readingTime":114},"s7x75o02kixj2u43gw4ah5pa","olfactory-processing-speed-cognitive-marker","https:\u002F\u002Fimages.unsplash.com\u002Fphoto-1542838132-92c53300287c?w=1280",5,"2026-07-28T20:38:15.776Z","Olfactory Processing Speed as a Cognitive Marker","Can how fast you identify smells predict your cognitive ability? We examine the surprising research linking olfactory processing speed to IQ, memory, and neurodegenerative disease risk.",{"id":217,"slug":218,"coverImage":219,"author":9,"viewsCount":21,"createdAt":220,"title":221,"description":222,"lang":14,"readingTime":186},"v1qq1yz082k72zzynk0qv2yh","do-musicians-have-higher-iqs","https:\u002F\u002Fimages.unsplash.com\u002Fphoto-1520006510097-3d64427f3f98?w=1280","2026-07-28T20:25:14.220Z","Do Musicians Have Higher IQs? The Research Verdict","The idea that musicians are smarter is widespread — but does the science support it? We review 10 key studies on musical training, IQ, and cognitive ability to separate fact from fiction.",{"id":224,"slug":225,"coverImage":226,"author":9,"viewsCount":227,"createdAt":228,"title":229,"description":230,"lang":14,"readingTime":186},"xam91ejelkjg23x2z9skt9vm","are-night-owls-smarter","https:\u002F\u002Fimages.unsplash.com\u002Fphoto-1518837695005-2083093ee35b?w=1280",4,"2026-07-28T19:47:32.529Z","Are Night Owls Smarter? Chronotype and Cognitive Performance","The idea that night owls are smarter than early birds is popular — but what does the research actually show? We examine the evidence on chronotype, IQ, and cognitive performance.",{"id":232,"slug":233,"coverImage":234,"author":9,"viewsCount":235,"createdAt":236,"title":237,"description":238,"lang":14,"readingTime":186},"hpakkt3uv7frkkhj2el6fvm1","does-high-iq-make-you-rich","https:\u002F\u002Fimages.unsplash.com\u002Fphoto-1579621970795-87facc2f976d?w=1280",6,"2026-07-28T19:37:46.682Z","Does High IQ Make You Rich? 12 Studies Reviewed","The correlation between IQ and income is real but surprisingly weak. We review 12 key studies to understand why high IQ doesn't guarantee wealth — and what actually predicts financial success.",{"id":240,"slug":241,"coverImage":242,"author":9,"viewsCount":212,"createdAt":243,"title":244,"description":245,"lang":14,"readingTime":186},"b5ct19s9mfso606wcppcp79o","iq-test-scores-in-prison-populations","https:\u002F\u002Fimages.unsplash.com\u002Fphoto-1497366216548-37526070297c?w=1280","2026-07-28T18:14:35.540Z","IQ Test Scores in Prison Populations: What the Studies Show","Research consistently finds that prison populations score lower on IQ tests than the general public — by about 10-12 points on average. But the reasons are far more complex than \"criminals are less intelligent.\" This article examines the evidence, the confounding factors, and what the data actually means.",{"id":247,"slug":248,"coverImage":249,"author":9,"viewsCount":99,"createdAt":250,"title":251,"description":252,"lang":14,"readingTime":16},"zjjd6y8yfcydy6y02c9xa2yf","which-country-has-the-highest-average-iq","https:\u002F\u002Fimages.unsplash.com\u002Fphoto-1526129313847-8d262a6b1f0d?w=1280","2026-07-28T17:59:21.857Z","Which Country Has the Highest Average IQ (And Why the Data Is Messy)","Every few months, a new headline declares that Singapore, Hong Kong, or South Korea has the highest IQ in the world. But behind these rankings lies a tangle of methodological problems, cultural biases, and outdated data. This article unpacks what the research actually shows — and what it doesn't.",{"id":254,"slug":255,"coverImage":256,"author":9,"viewsCount":235,"createdAt":257,"title":258,"description":259,"lang":14,"readingTime":186},"nxfpaqozmm0dpaytyy2cjixv","why-your-iq-test-score-changes-over-time","https:\u002F\u002Fimages.unsplash.com\u002Fphoto-1499750310107-5fef28a66643?w=1280","2026-07-28T17:33:41.377Z","Why Your IQ Test Score Changes Over Time","Your IQ score is not a fixed number stamped on your brain. Research shows that IQ scores can change significantly over a lifetime — sometimes by 20 points or more. This guide explains what drives these changes and what they mean for you.",{"page":10,"pageSize":261,"pageCount":31,"total":262},24,49]