Does 2026 AI Exhibit Intelligence, or Can Claude Outsmart Pierre or Catherine?

Using a sequence of high-school level mathematics questions that were not available on the Internet, we compare the performance of the popular AI software Claude with that of my friends and fellow human beings Pierre and Catherine. Pierre had solid scientific training as a young man, while Catherine studied literature. All three were subjected to a simulated pre-calculus oral exam with main questions and follow-up questions. Their performances are compared and the ones with the best and worst performances are identified. The outcome is that the current version of Claude, even though it is an extremely useful tool that has probably recorded the solution to nearly all calculus questions that are available on the Internet, {\em exhibits only a very limited understanding of the subject} and {\em does not exhibit the ability to make intelligent connections} between different features of a pre-calculus mathematics problem that it has never seen before.

Authors

Institutions

Publication Details

Journal
The Mathematical Intelligencer
Published
2026-09-28
DOI
https://doi.org/10.1007/s00283-026-10570-x
Primary Topic
Computability, Logic, AI Algorithms
Type
article
Field-Weighted Citation Impact
0.00
Controls
|||
ALL TIME
JAN
FEB
MAR
APR
MAY
JUN
JUL
AUG
SEP
article

Does 2026 AI Exhibit Intelligence, or Can Claude Outsmart Pierre or Catherine?

Robert C. Dalang
The Mathematical Intelligencer
Computability, Logic, AI Algorithms
article

Does 2026 AI Exhibit Intelligence, or Can Claude Outsmart Pierre or Catherine?

Robert C. Dalang
article en

Abstract

Using a sequence of high-school level mathematics questions that were not available on the Internet, we compare the performance of the popular AI software Claude with that of my friends and fellow human beings Pierre and Catherine. Pierre had solid scientific training as a young man, while Catherine studied literature. All three were subjected to a simulated pre-calculus oral exam with main questions and follow-up questions. Their performances are compared and the ones with the best and worst performances are identified. The outcome is that the current version of Claude, even though it is an extremely useful tool that has probably recorded the solution to nearly all calculus questions that are available on the Internet, {\em exhibits only a very limited understanding of the subject} and {\em does not exhibit the ability to make intelligent connections} between different features of a pre-calculus mathematics problem that it has never seen before.

The Mathematical Intelligencer
École Polytechnique Fédérale de Lausanne (CH)
Quality Education
Openalex Percentile: Top 49%
Computability, Logic, AI Algorithms
AI Navigator

Ask Laika to Summarize, Analyze, and Connect papers live on the map.

Summarize Papers & Methodologies

Extract key findings, datasets, and comparative methods across publications.

Benchmark Rankings & Visual Analytics

Rank top research institutions, authors, funders, topics, and journals by Field-Weighted Citation Impact (FWCI) and paper volume with instant charts.

Connect Distant Disciplines

Bridge topological clusters on the map to find hidden collaborative intersections.

Does 2026 AI Exhibit Intelligence, or Can Claude Outsmart Pierre or Catherine? — Robert C. Dalang · The Mathematical Intelligencer (2026) | TGRS Research Map | TGRS