Abstract reasoning (novel domains)
Cognition · Human leadsHumans transfer schemas across unrelated domains; LLMs still fail on out-of-distribution logic puzzles.
BRAINMATTER Index · Updated 2026-07-19
Human Intelligence and Artificial Intelligence scored side-by-side across 24 dimensions — cognition, learning, perception, embodiment, social skills, ethics, economics, and meta-cognition.
| Category | Human avg | AI avg | Leader |
|---|---|---|---|
| Cognition | 6.4 | 7.3 | AI |
| Learning | 7 | 5 | Human |
| Perception | 8.5 | 8 | Human |
| Physical | 7.7 | 5.7 | Human |
| Social | 8.7 | 5.3 | Human |
| Ethics | 7.7 | 3 | Human |
| Economics | 2 | 10 | AI |
| Meta | 9.5 | 2.5 | Human |
Humans transfer schemas across unrelated domains; LLMs still fail on out-of-distribution logic puzzles.
AI processes billions of examples; humans cap at ~10⁴ meaningful exemplars per lifetime.
Agentic AI degrades sharply past ~20-step plans without tool scaffolding.
Human WM ≈ 4±1 chunks; frontier models hold 1M+ tokens with near-perfect recall.
Humans infer counterfactuals from few observations; AI conflates correlation with cause.
Trivially decisive for AI when tools are available.
A child learns a new word from 1–3 exposures; LLMs require millions of tokens.
Neural nets suffer catastrophic forgetting; humans consolidate via sleep.
AI copies a skill to 1M instances instantly; humans need years of individual training.
Humans integrate 5+ senses with proprioception; AI catching up on vision+audio+text.
AI surpasses humans on ImageNet-style benchmarks.
Robotics still fragile in unstructured environments.
No robotic hand matches a surgeon or a violinist.
AI runs 24/7 at constant quality; humans need sleep.
LLMs pass some false-belief tasks but fail robust ToM under distraction.
AI simulates affect; humans feel and respond adaptively.
2024–25 studies show LLMs match or exceed humans in structured debate.
AI reflects training-data priors; humans reason from lived stakes.
Legal personhood applies to humans; AI is a tool with attributed liability.
Both systems encode bias; neither is neutral without intervention.
AI inference approaches $0.0001/task; human labor is orders of magnitude higher.
AI is above-average creative but rarely produces genuinely field-shifting work.
No evidence current AI systems have phenomenal self-models.
Brain runs on ~20W; GPT-scale training uses gigawatt-hours.
The HI vs AI Index is a scored comparison of Human Intelligence and Artificial Intelligence across 24 dimensions spanning cognition, learning, perception, physical embodiment, social skills, ethics, economics, and meta-cognition. Each dimension is rated 0–10 for each side with a plain-language note.
AI dominates in pattern recognition at scale, working memory capacity, arithmetic speed, skill acquisition at scale, endurance, marginal cost per task, and object recognition on labeled benchmarks.
Humans lead in abstract reasoning on novel domains, long-horizon planning, sample-efficient learning, continual learning without forgetting, embodied navigation, fine motor dexterity, theory of mind, empathy, moral judgment, accountability, creative originality, self-awareness, and energy efficiency.
Each dimension is scored 0–10 by BRAINMATTER editors from a synthesis of 2024–2026 benchmarks, peer-reviewed studies, and industry evaluations. Scores are directional, not absolute — the index is designed to reveal shape and gaps, not to declare a winner.