HACKS VITAE

CHECKED · CRITICAL THINKING

Why Critical Thinking Matters More in the Age of AI
What the Research Shows, and What It Doesn't

PRICES, VERSIONS AND FACTS AS OF OCTOBER 2026

What critical thinking is, whether it can be taught, what the employer surveys say, and what the studies on AI and thinking found, checked against the papers, reports and preprints themselves, with habits drawn only from what holds up.

HOVER A GLOWING POINT · DRAG TO TURN
  • Doesn't hold up 1
  • Partly 6
  • Can't confirm 1
Published
October 1, 2026
Facts as of
October 2026
Read
14 min
points
8

BACKGROUND · GEORGES DE LA TOUR, THE PENITENT MAGDALEN, C. 1640 · THE MET, OPEN ACCESS

THE SHORT VERSION

  1. Critical thinking can be taught, modestly: two reviews of classroom studies (2008 and 2015) found an average gain of about a third of a standard deviation, and how far the skill carries from one subject to another is still argued over.
  2. The job-skill claims are looser than they sound: in the World Economic Forum's employer surveys, "critical thinking and analysis" led the skill groups in 2020 but ranked fourth among individual skills, behind analytical thinking, which also topped the 2023 and 2025 lists. We could not find a study behind the figure that 65% of children will work in jobs that don't exist yet.
  3. The AI studies say less than the headlines: two surveys show associations, not causes, and the brain-scan study is a small preprint that has not been peer reviewed. In the clearest experiment, students who used GPT-4 as a crutch in practice did worse on their own later, and a tutor built to guide rather than give answers largely avoided the drop.
  4. Checking helps, modestly: leaving a page to see what others say about it, and short lessons on manipulation tricks, both improved people's judgements in tests, but the gains are small and may not last.

What we found

Our reading of the evidence on each of the 8 points, with where it comes from. Open any row, look at the source, and make up your own mind.

PartlyCritical thinking is the No. 1 skill employers want, according to the World Economic Forum

True of a skill group in 2020; not of the ranked skills, then or since. In the Forum's October 2020 report, "critical thinking and analysis" led the skill groups employers saw rising in importance, but was fourth in its list of the top 15 individual skills for 2025. In the 2023 and 2025 reports the top core skill is analytical thinking, and we did not find the phrase "critical thinking" in the text of either report. These are surveys of what employers say they want, not measurements.

SOURCE World Economic Forum, The Future of Jobs Report 2020, 2023 and 2025.

Can't confirm65% of children starting school today will work in jobs that don't exist yet

No study behind it has been found. The Forum's 2016 report gives it as "one popular estimate", and its endnote gives a web address for Shift Happens, not a study. An author who used the figure in a 2011 book wrote in 2017 that she had traced it to an Australian website that no longer existed, and now regards it as "a pretty vague figure".

SOURCE World Economic Forum, The Future of Jobs (2016), p. 3 and endnote 1 (p. 43); Cathy N. Davidson, blog post (1 June 2017).

PartlyCritical thinking can be taught

The average gain is real but modest; how far it travels is disputed. Reviews of 117 studies (2008) and 341 effect sizes (2015) found average effects of 0.341 and 0.30 standard deviations. Whether a skill learned in one subject carries over to others depends heavily on knowing the subject, and the evidence on how to make it carry over is, in one reference work's word, "scanty".

SOURCE Abrami et al., Review of Educational Research (2008, 2015); Willingham, American Educator (2007); Stanford Encyclopedia of Philosophy, "Critical Thinking".

PartlyA Microsoft study found that AI is reducing workers' critical thinking

The workers reported it; the study did not measure their thinking or show a cause. In a survey of 319 knowledge workers, higher confidence in AI went with less critical thinking, as the workers described it. The paper's own title calls the reductions "self-reported", and its authors list self-report and a younger, tech-savvy sample among the limits.

SOURCE Lee et al., CHI '25 (Microsoft Research and Carnegie Mellon University).

Doesn't hold upAn MIT study proved that ChatGPT rots your brain

It is an unreviewed preprint, and its authors reject that wording. 54 people wrote essays with ChatGPT, a search engine or no tools while their brain activity was recorded, and the ChatGPT group showed the weakest connectivity. The authors' own FAQ asks people not to use words like "brain rot" about it. As of 1 October 2026 its arXiv record listed no journal publication.

SOURCE Kosmyna et al., arXiv:2506.08872 (v2, 31 December 2025) and the authors' FAQ; Stanković et al., arXiv:2601.00856.

PartlyUsing ChatGPT makes you worse at thinking

Shown for one pattern of use, in maths learning; not shown for thinking in general. In a field experiment with nearly 1,000 high-school students, those who had used a plain GPT-4 tool during practice got 17% lower grades on their own afterwards than students who never had it, while a version designed to give hints instead of answers largely avoided the drop. That study measured maths learning, not thinking in general, and other experiments, including one with a tutor built for a course, found gains.

SOURCE Bastani et al., PNAS (2025); Kestin et al., Scientific Reports (2025); Contractor & Reyes, arXiv:2607.08849 (preprint, 2026).

PartlyPISA 2025 shows that students who use AI do worse at school

It shows an association, not a cause. Students who said they did not use AI to draft writing outperformed those who did. Among frequent users, those who used AI to help them learn and were taught to judge the quality of its output tended to do better than those who had no such guidance. The fuller analysis is due in May 2027.

SOURCE OECD, PISA 2025 press release (8 September 2026).

PartlyPrebunking makes people immune to misinformation

It helps people recognise tricks; "immune" overstates it. On YouTube, short videos raised recognition of a manipulation technique by about 5% on average. In one study the effect was no longer significant after two months without repeat testing. Whether these lessons help people tell reliable from unreliable news is disputed: a five-study reanalysis (2023) found no gain, and a 33-experiment reanalysis (2026), two of whose five authors co-wrote the YouTube study, found a consistent one. We lean cautiously toward the larger reanalysis because it covers far more experiments.

SOURCE Roozenbeek et al., Science Advances (2022); Maertens et al., Journal of Experimental Psychology: Applied (2021); Modirrousta-Galian & Higham (2023); Simchon et al. (2026).

THE ARTICLE · 14 MIN

Every tool that does some of our thinking for us brings the same worry: are we doing less of it ourselves? This page looks at what critical thinking is, whether it can be taught, what employers say about it, and what the research on AI does and does not show. Then it turns what holds up into habits you can try.

What critical thinking means

As an educational goal, the idea goes back to the American philosopher John Dewey (1910), who more commonly called it “reflective thinking”. He defined it as “active, persistent and careful consideration of any belief or supposed form of knowledge in the light of the grounds that support it, and the further conclusions to which it tends.”

Between 1988 and 1989 a panel of 46 scholars, educators and others, brought together at the request of a committee of the American Philosophical Association, agreed a longer definition, published in 1990. Critical thinking, they wrote, is “purposeful, self-regulatory judgment which results in interpretation, analysis, evaluation, and inference, as well as explanation of the evidential, conceptual, methodological, criteriological, or contextual considerations upon which that judgment is based.” Only one of the panel declined to be listed as supporting the final document.

The philosopher Robert Ennis puts it more briefly, in his 2011 wording: “Critical thinking is reasonable and reflective thinking focused on deciding what to believe or do.”

The definitions do not fully agree. Some limit critical thinking to forming a judgment; others, in the words of the Stanford Encyclopedia of Philosophy, “allow for actions as well as beliefs as the end point”. That matters for what follows, because the studies people quote do not all measure the same thing. One of the AI studies below defines critical thinking through Bloom’s taxonomy of learning goals, and its authors say plainly that “this definition of critical thinking is not uncontested.”

Can it be taught?

On average, a little. A 2008 review by Philip Abrami and colleagues pooled 117 studies with 20,698 participants and found an average effect of 0.341 standard deviations, with results that varied a great deal (“the distribution was highly heterogeneous”). The authors concluded that improvement “cannot be a matter of implicit expectation” and that “educators must take steps to make CT objectives explicit in courses”.

A 2015 review by the same group, drawing on 341 effect sizes from experimental and quasi-experimental studies, found a similar average of 0.30 and concluded that there are effective strategies “at all educational levels and across all disciplinary areas.” Three things stood out: “the opportunity for dialogue, the exposure of students to authentic or situated problems and examples, and mentoring”.

Our reading: an effect of about a third of a standard deviation is a real but modest average gain, not a transformation. The same review found that teaching critical thinking on its own and inside subject lessons together did better than either alone, but, as the Stanford Encyclopedia notes, that difference “was not statistically significant; that is, it might have arisen by chance”, and most of the studies lacked long-term follow-up.

A 2016 review by Huber and Kuncel found that critical thinking skills and dispositions “improve substantially over a normal college experience”, but that curriculum-wide programmes to improve them “do not necessarily produce incremental long-term gains.”

The harder question is whether the skill travels from one subject to another. The cognitive scientist Daniel Willingham argued in 2007 that critical thinking “is not a set of skills that can be deployed at any time, in any context” and that it is “very much dependent on domain knowledge and practice.” He did not say it never carries over: “such transfer does occur”, he wrote, but when and why is complex. The Stanford Encyclopedia describes the evidence for one popular answer, teaching critical thinking across many subjects with explicit attention to the skills they share, as “scanty”.

Our reading: teaching helps on average, and knowing a subject well is what lets the habit work in that subject.

What employers say

The line “critical thinking is the No. 1 skill employers want” is usually credited to the World Economic Forum’s Future of Jobs reports, and the Forum’s own summary of its 2020 report said that critical thinking and problem-solving “top the list”. These are surveys of employers, what companies say they expect, not measurements of what workers can do.

  • In the October 2020 report, “critical thinking and analysis” comes first among the skill groups employers saw rising in importance; the report says it and problem-solving “have stayed at the top of the agenda”. In the same report’s list of the top 15 individual skills for 2025, it ranked fourth, behind analytical thinking and innovation, active learning, and complex problem-solving.
  • The May 2023 report, based on 803 companies, says “analytical thinking and creative thinking remain the most important skills for workers in 2023.”
  • The January 2025 report, based on over 1,000 employers surveyed in late 2024, says “analytical thinking remains the most sought-after core skill among employers, with seven out of 10 companies considering it as essential in 2025.”

We searched the text of the 2023 and 2025 reports and did not find the phrase “critical thinking” in either. So the claim fits one 2020 chart of skill groups; in the ranked lists of individual skills, a close cousin, analytical thinking, comes first.

A second statistic often travels with these reports: “65% of children entering primary school today will ultimately end up working in completely new job types that don’t yet exist.” The Forum’s 2016 report gives it as “one popular estimate”, and its endnote gives a web address for Shift Happens, not a study. We could not find a study behind it. An author who used the figure in a 2011 book wrote in 2017 that it “didn’t originate with me”, that she had traced it to an Australian website that no longer existed, and that she now regards it as “a pretty vague figure”.

AI and thinking: what the studies found

Our guide to reading a study explains why a survey, a preprint and an experiment answer different questions.

A survey of knowledge workers

Researchers at Microsoft Research and Carnegie Mellon University surveyed 319 knowledge workers, who shared 936 examples of using generative AI at work. The paper, presented at the CHI 2025 conference, found that “higher confidence in GenAI is associated with less critical thinking, while higher self-confidence is associated with more critical thinking”, and that AI shifted critical thinking “toward information verification, response integration, and task stewardship.”

The paper’s own title calls the reductions in effort “self-reported”. Its authors note that participants “occasionally conflated reduced effort in using GenAI with reduced effort in critical thinking with GenAI”, and that the sample “was biased towards younger, more technologically skilled participants”. It shows an association in what people reported, not a cause.

A survey of 666 people

A 2025 paper in the journal Societies surveyed 666 people in the UK, recruited online through social media, interviewed 50 of them, and found “a significant negative correlation between frequent AI tool usage and critical thinking abilities, mediated by increased cognitive offloading.” Critical thinking was measured partly by self-report. The author lists “reliance on self-reported measures and the potential for sample bias” among the limits, and says experiments “could offer causal evidence”. A correction in September 2025 replaced a duplicated table; the author states that the scientific conclusions are unaffected.

The brain-scan preprint

The study behind many “ChatGPT and your brain” stories is a preprint from the MIT Media Lab, first posted in June 2025. In its words, “a total of 54 participants took part in sessions 1-3, with 18 completing session 4”: they wrote essays with ChatGPT, with a search engine or with no tools while their brain activity was recorded by EEG. The authors report that “LLM users displayed the weakest connectivity.”

As of 1 October 2026 its arXiv record shows a preprint, last revised on 31 December 2025, with no journal publication listed. Other researchers have posted a constructive comment raising “the limited sample size”, reproducibility and the EEG methods. The authors’ own FAQ answers the question of whether it shows that LLMs make us “dumber” with “No!” and asks people not to use words like “brain rot”.

The strongest experiment

The clearest causal evidence among these studies comes from a field experiment with nearly 1,000 students at a high school in Turkey, published in PNAS in 2025. While students had GPT-4 during practice, their grades rose (“48% improvement in grades for GPT Base and 127% for GPT Tutor”). When access was taken away, the students who had used the plain tool did worse than those who never had access (“17% reduction in grades for GPT Base”).

The tutor version, whose instructions asked it “to avoid giving them the answer and instead guide them in a step-by-step fashion”, largely avoided the drop. The authors’ explanation is that, without guardrails, students used GPT-4 “as a ‘crutch’ during practice problem sessions”. The study measured maths learning over a few sessions, not critical thinking in general, and the authors say they “focus on short-term outcomes”.

Experiments that point the other way

When the tool is built for learning, results can be positive. In a randomised trial in a Harvard undergraduate physics course with 194 students, each taking one lesson with an AI tutor and one in class, a custom tutor built on the course’s own teaching methods let students “learn significantly more in less time” than an in-class active-learning lesson (Scientific Reports, 2025).

Two 2026 studies add detail, one a preprint and one a CHI 2026 conference paper. In the preprint, access to a general AI tool raised undergraduates’ immediate test scores by 0.27 standard deviations, and the gains persisted a week later; essay gains a week later were larger among students who used AI “to explain concepts rather than generate text”. In the conference paper, an experiment with 393 people, having an AI model from the start “improved performance under time pressure but impaired it with sufficient time”.

What 15-year-olds told PISA

The OECD’s PISA 2025 results, published on 8 September 2026, add a large survey: “Students who say they did not use AI to draft text for writing assignments outperformed those who say they did.” Among frequent users, students who used AI to help them learn and were taught to assess the quality of AI-generated information “tended to perform better than those who received no such guidance.” These are associations, not proof that AI lowered anyone’s scores. The full results of PISA’s new Learning in the Digital World test are due in May 2027.

Our reading: surveys find that people who lean on AI more report, or score, less critical thinking, but they cannot say which causes which. The clearest experiment among them found that using AI to skip the work during practice left students worse off on their own, and that a tool built to make students do the thinking largely avoided that.

Over-trusting machines is an old problem

Researchers have a name for leaning too hard on a confident machine: automation bias, “the tendency to over-rely on automation”. A 2012 systematic review screened 13,821 papers and kept 74. It found that trust and confidence, workload, task complexity and time pressure all played a part, and that training and “emphasizing user accountability” helped.

In a 2023 experiment, 27 radiologists read mammograms with suggestions from what they were told came from an AI system; in 12 of 40 cases the suggestion was deliberately wrong. Inexperienced readers rated 79.7% of mammograms correctly when the suggestion was right and 19.8% when it was wrong; very experienced readers fell from 82.3% to 45.5%. The authors concluded that radiologists at every level of experience “are prone to automation bias when being supported by an AI-based system.”

Experience softened the pull without removing it. Why AI makes things up covers the errors chatbots make.

Checking what you read: what helps, and how much

The first approach is lateral reading: leaving a page to see what other sources say about it, the core of the SIFT method. In a 2022 field study in one urban school district, high-school students who had six 50-minute lessons (271 students, compared with 228 in regular classes, in a matched rather than randomised design) “grew significantly in their ability to judge the credibility of digital content”. As our SIFT page explains, a lasting benefit measured against a comparison group has not been established.

The second is prebunking, or inoculation: short games or videos that show people a manipulation technique before they meet it. In a 2022 study in Science Advances, five short videos improved recognition of techniques such as false dichotomies across six randomised experiments with 6,464 people. In a field test on YouTube the effect was smaller: recognition rose “by about 5% on average”, measured by a single test question within 24 hours of the advert. Effects also fade. In a 2021 study of the Bad News game, the benefit lasted at least three months when people were tested at regular intervals, but without regular testing it was “no longer significant” after two months.

Whether these lessons help people tell reliable from unreliable news, and not just name a trick, is disputed. A 2023 reanalysis of five studies found that two inoculation games “did not improve discrimination” but made people answer “false” more often to everything. A 2026 reanalysis of 33 experiments with 37,025 people found that games and videos “consistently improve discrimination between reliable and unreliable news, without increasing response bias.” Two of that reanalysis’s five authors also co-wrote the 2022 YouTube study; the 2023 critique was written by two researchers who were not among the authors of either. Our logical fallacies page weighs the two and leans, cautiously, toward the larger reanalysis, because it covers far more experiments. Whether any benefit survives a realistic feed has not been shown: in a 2025 study in PNAS Nexus, mixing real posts into a simulated feed “appeared to nullify the effect” of a lesson on emotional manipulation.

How to use it

  • Doing the thinking first on practice tasks. In the maths experiment, students who used a plain AI tool to get answers during practice did worse on their own later. A 2026 conference paper found that, with enough time, people who started without the AI model did better than those who had it from the start; under time pressure the pattern reversed.
  • Asking for hints and explanations rather than answers. The tutor that guided students step by step largely avoided the drop, and in a 2026 preprint students who used AI to explain concepts showed larger gains a week later.
  • Checking AI output against a source. Automation bias catches experts as well as beginners. In PISA, frequent AI users who used it to help them learn and had been taught to assess its output tended to do better, though that is an association. Our page on spotting a fake quote or an invented source covers the checks.
  • Leaving the page to check who is behind it. Lateral reading helped students judge credibility in the classroom study; whether the habit lasts has not been established.
  • Learning the subject. Willingham’s point is that critical thinking depends on knowing the subject you are thinking about.
  • Repeating the lesson. In the Bad News study, the effect held when people were tested regularly and faded when they were not.

Related: Think Again · Cognitive biases checked

Sources

Checked October 2026. What we read: the Stanford Encyclopedia entry; the Delphi Report, Ennis’s outline and Willingham’s article in full; the four World Economic Forum reports (searched as text); the full texts of the Lee, Gerlich, Bastani, Kestin and Roozenbeek papers and the Gerlich correction; the OECD’s PISA 2025 press release; and the abstracts of the other papers and preprints, with the Kosmyna authors’ FAQ. What we could not read: the full texts of the Abrami and Huber & Kuncel reviews (so their effect sizes are taken from the abstracts), the full PISA 2025 report, the full texts of the two 2026 studies, and the BBC programme on the 65% figure. If you can show any of this wrong, with a source, we want to see it.

  • critical thinking
  • ai
  • learning
  • misinformation
  • research