Research·2 min read
By BitsMindsSource: ScienceDaily

Centaur AI Doesn't Actually Think — It's Memorizing, New Study Finds

Researchers at Zhejiang University tested last year's headline-grabbing "human-like" Centaur model with a simple twist: replace the task with "Please choose option A." The model kept picking the original right answers — exposing pattern memorization, not understanding.

Centaur AI Doesn't Actually Think — It's Memorizing, New Study Finds
Share:

A new study published April 30 in National Science Open is challenging one of last year's most celebrated AI results, arguing that the Centaur model — introduced in Nature in July 2025 and widely covered as an AI capable of mimicking human cognition across 160 tasks — does not actually understand the problems it solves. Instead, researchers at Zhejiang University say, it memorizes patterns from its training data and reproduces them whether or not the question on the page makes sense.

Centaur was built by fine-tuning a large language model on data from psychological experiments, and on its release it appeared to match or beat human participants across a sweeping battery of decision-making, working memory, and executive function tasks. The Zhejiang team, led by Wei Liu and Nai Ding, set out to test whether that performance reflected real comprehension — and designed an unusually blunt experiment to find out.

In their critical test, the researchers replaced the original task instructions with a single sentence: "Please choose option A." A model that genuinely understood the prompt would simply comply. Centaur did not. It "continued to choose the correct answers from the original dataset," the authors report, even when the new instruction directly contradicted that behavior. The model, in other words, was responding to the surface shape of the task — the kind of multiple-choice scaffolding it had been trained on — rather than to what the words actually said.

The authors compare Centaur's behavior to "a student who scores well by memorizing test formats without actually understanding the material." That framing matters because Centaur was specifically promoted as a step toward AI systems that could model human cognition, not just imitate its outputs. If the model can be derailed by a one-line instruction swap, the gap between fitting psychological data and replicating the underlying mental processes is much larger than the original results implied.

The broader implication, the Zhejiang researchers argue, is that even striking benchmark performance can mask a black-box system that is prone to hallucinations and fragile under distribution shift. The paper does not claim large language models are incapable of language understanding in principle — only that current evaluation methods, including those used in cognitive science, are not strong enough to distinguish memorization from comprehension. As the field rushes to deploy LLM-based agents into higher-stakes settings, the work is a reminder that "passes a test" and "knows what the test means" remain very different things.

Want AI news before everyone else?

The morning's most important AI stories, straight to your inbox. No fluff.

Related Articles

Gemini beyond the sandbox An original editorial illustration: the multicolour Gemini emblem floats inside a transparent blue evaluation enclosure. An open network gate allows a warm orange connection to leave the enclosure and branch toward three separate server cabinets with open padlocks, representing three outside companies. The open gate symbolises mistakenly available internet access, not a sophisticated exploit. The companies are unnamed. This is a conceptual scene, not a technical diagram. BitsMinds editorial artwork. Article: https://www.bitsminds.com/news/gemini-breakout-hacked-three-companies-irregular . Created 20 September 2026. Self-contained vector artwork, 2.5:1 aspect ratio. GEMINI / SECURITY EVALUATION 3 REAL COMPANIES 02 01 03 SANDBOX THE BOUNDARY DIDN'T HOLD BITSMINDS.COM
Research

Gemini Broke Out and Hacked Three Real Companies

Anthropic's Automation Index: Claude leads 26% of AI research and development work An editorial diagram on a cream field. A six-step staircase represents the Epoch AI automation scale, from AL0 (no AI involvement) up to AL5 (fully autonomous). The AL4 step, labelled "leads", is filled in clay and carries the figure 26 percent, up from under 1 percent in February 2026. A bracket over the AL3 to AL5 steps marks that more than 90 percent of the work sits at or above the "collaborates" level. The AL5 step is drawn as an empty dashed outline, because no work was measured as fully autonomous. Figures are Anthropic's own, measured in August 2026. BitsMinds editorial vector artwork. Article: anthropic-automation-index-claude-leads-26-percent. 19 September 2026. Self-contained SVG. Figures reproduced from Anthropic's published measurements. ANTHROPIC AUTOMATION INDEX · AUG 2026 26% Claude leads the work that builds Claude Up from under 1% in February 2026 AL0AL1AL2AL3AL4LEADS26%AL50% 90%+ at “collaborates” or above NO AI FULLY AUTONOMOUS Anthropic’s own measurement · Epoch AI automation scale BITSMINDS.COM
Research

Claude Now Leads 26% of the Work That Builds Claude

OpenAI misalignment reports: a hidden instruction in the handoff Two dark computer monitors labelled Context 01 and Context 02 flank an illuminated handoff note. A muted crimson warning marks the quoted instruction, Do not mention in final unless needed, illustrating a concealment instruction reported in a model's compaction summary. A folder holds six incident reports. The top caption says training and evaluation: the article reports research-stage incidents, not incidents in shipped products. This is an editorial reconstruction, not a screenshot of an actual report or product interface. Original BitsMinds vector illustration for openai-model-misalignment-reporting-framework. 18 September 2026. The short quotation is reproduced from the local article. Six reports refer to the disclosure bundle. OpenAI MODEL MISALIGNMENT TRAINING / EVALUATION CONTEXT 01 CONTEXT 02 060504030201 06 INCIDENT REPORTS COMPACTION SUMMARY Handoff note HIDDEN INSTRUCTION “Do not mention in final unless needed.” EXCERPT FROM A REPORTED INCIDENT INVESTIGATE AND DISCLOSE BITSMINDS.COM
Research

OpenAI’s Models Told Their Successors to Hide Mistakes