When a Tool Becomes a Crutch: Why Prediction Power Can Destroy Learning

Thomas Hirschmann

Hatched by Thomas Hirschmann

Jun 06, 2026

7 min read

89%

0

The Strange Risk of Being Helped Too Well

What if the most powerful AI systems do not fail by being wrong, but by being too useful?

That sounds backwards, especially at a moment when language models are being celebrated for their ability to predict, summarize, draft, diagnose, and even tutor. A system that can forecast the next word, the next likely outcome, or the next clinical risk seems like the kind of engine we should build everything around. And in many settings, that instinct is right. But there is a hidden asymmetry: a machine can be excellent at producing answers while quietly making the human around it worse at producing them.

That tension matters because prediction is not just a technical capacity. It is a psychological environment. When a model begins to stand in for our own effort, the question is no longer simply whether it is accurate. The deeper question is whether it leaves us stronger after the interaction, or weaker.

That is the real puzzle connecting AI in education and AI in medicine. In both cases, the same basic capability, high quality prediction, can either extend human judgment or slowly replace the mental work that creates judgment in the first place.

Prediction Is Not the Same as Understanding

A useful way to think about generative AI is as a prediction engine. It guesses the next word, the likely answer, the probable diagnosis, the most plausible recommendation. In medicine, that can be extraordinary. A system that can help predict risk, triage patients, or identify patterns in records can improve decisions at scale. In learning, too, a model can predict the next step in a solution and provide immediate feedback.

But prediction has a dangerous side effect: it can make the hard parts of cognition feel optional.

Imagine learning to play chess with an engine that instantly points out the best move. At first, you improve quickly. You see more patterns, make fewer mistakes, and feel your skill rising. But if the engine never lets you sit with uncertainty, then your own search process never matures. You are not learning to think in chess. You are learning to recognize when the machine has thought for you.

The same dynamic appears in classrooms and workplaces. When a student leans on GPT-4 during practice, the model can become a cognitive crutch. The student completes more problems, but the completion is misleadingly cheap. Once the tool is removed, performance can drop below that of someone who never had the tool at all. That is not a small downside. It means the tool did not just fail to teach, it may have trained dependence.

The most dangerous technologies are often the ones that reduce friction so well that they also reduce formation.

This is the central tension: efficiency can hollow out apprenticeship. The faster the answer arrives, the less opportunity there is for the mind to build the structures that make future answers possible.


Why Speed Can Steal Skill

To understand the problem, it helps to distinguish between two kinds of value in AI assistance.

  1. Output value: the immediate quality of the result.
  2. Formative value: the extent to which the user becomes better at the task over time.

These are not the same. A system can maximize output value while minimizing formative value. In fact, the better the system is at filling gaps instantly, the more it risks starving the user of the struggle that creates competence.

This is why a tutor is not the same as an answer machine. A true tutor changes the shape of the learner's attention. It slows the learner down at the right moments, asks questions, gives hints, and creates productive friction. The goal is not merely to get the right answer today, but to ensure the learner can later produce it alone.

The danger of many generic AI interactions is that they remove precisely that productive friction. They do not merely help the learner over a hump. They replace the climbing.

A simple analogy makes this clear. Consider two people learning to drive. One has a car that constantly steers itself and warns them before every mistake. The other drives a manual car, with a patient instructor who intervenes only when necessary. The first driver may feel safer and more efficient. But the second is more likely to develop real judgment, because they are constantly learning the relationship between input, consequence, and correction.

That is why the educational finding is so important: once access to the AI is removed, the previously assisted student can perform worse than a peer who never had it. The tool has not merely accelerated learning. It has potentially altered the learner's reliance structure.

The implication is unsettling: in some contexts, help can become a hidden tax on independence.

The Same Engine Can Build or Erode Judgment

Now bring medicine into the picture. A health system model that serves as a universal prediction engine sounds almost like the perfect institutional tool. Hospitals are full of prediction problems: who is at risk, who needs attention first, which intervention is likely to work, where bottlenecks will emerge.

In that context, prediction is powerful because it can compress complexity. It can reveal patterns that no single clinician could spot from a chart alone, especially across thousands or millions of cases. Used well, it becomes a kind of exoskeleton for judgment.

But the same structure that makes it valuable can also make it risky. If a system reliably predicts outcomes, teams may start trusting the system's output more than their own reasoning. Over time, clinicians can become less practiced at noticing edge cases, questioning assumptions, or exploring alternative explanations. The model becomes not just a support, but a substitute.

This is especially risky because medicine and education both depend on deliberate practice under uncertainty. Doctors learn by wrestling with ambiguous symptoms. Students learn by struggling through unsolved problems. In both cases, the human skill comes from the process of inference, not from the final answer alone.

So the real question is not whether AI can predict well. It can. The question is whether we use prediction to amplify human inference or to atrophy it.

A system that predicts everything may still be a bad partner if it prevents the human from building the habit of prediction themselves.

The Hidden Variable: Preservation of Agency

The most useful framework here is to ask of any AI system: Does it preserve agency, or does it outsource it?

Agency is not just freedom in a philosophical sense. It is the lived capacity to notice, choose, infer, and revise. A good system increases agency by making the user more capable after the interaction ends. A bad system decreases agency by making the user more dependent during the interaction itself.

This gives us a practical test.

If an AI helps a student solve five algebra problems but leaves them unable to solve the sixth alone, it is probably optimizing output, not agency.

If a hospital model flags likely sepsis cases but clinicians still need to understand why the model is confident, where it is uncertain, and when it might fail, then the model is supporting agency.

The difference is not whether the AI is used. The difference is whether the AI is designed as a scaffold or a replacement.

A scaffold is temporary by design. It exists to be removed. A replacement is designed to remain because it occupies the function permanently. Most of the danger arises when we confuse the two.

This is why safeguards matter so much. A tutor that withholds the full answer, asks guiding questions, and adapts to the learner's current stage is not just friendlier software. It is a machine designed to protect the learner's future competence. Likewise, a medical model that explains confidence levels, flags uncertainty, and keeps clinicians in the loop is more than a useful tool. It is a system that preserves professional judgment.

The best AI systems are not the ones that eliminate effort. They are the ones that make effort more intelligent.

Designing for Long-Term Strength, Not Short-Term Convenience

If prediction engines can both help and harm, then the design challenge is clear: build systems that optimize for long-term capability, not just immediate convenience.

That means asking different questions before deployment.

Instead of asking,

Sources

← Back to Library

Hatch New Ideas with Glasp AI 🐣

Glasp AI allows you to hatch new ideas based on your curated content. Let's curate and create with Glasp AI :)

Start Hatching 🐣