What Monkeys on a Polygon Reveal About the Limits of Proof
Hatched by Dhruv
Jul 31, 2026
9 min read
2 views
82%
The Strange Boundary Between What Can Be Proved and What Can Only Be Counted
What do you do when a problem is too slippery to prove directly, but too important to ignore? That question sits at the heart of both scientific inquiry and a deceptively simple counting puzzle: if monkeys move around a polygon, how many ways can they move so that at least one collision happens?
At first glance, these ideas seem to live in different universes. One belongs to the philosophy of science, where only testable phenomena can be brought under the discipline of falsifiability. The other belongs to combinatorics, where a problem can often be solved not by watching every individual event, but by counting the complement of what you want to avoid. Yet both point to the same deeper lesson: knowledge often advances not by capturing reality all at once, but by identifying the smallest conditions under which an event becomes undeniable.
That is a powerful intellectual move. It does not ask, “Can I explain everything?” It asks, “Can I define the moment when something becomes impossible to dismiss?”
The border between the knowable and the unknowable is often not drawn by certainty, but by the ability to specify a test, a condition, or a collision.
Why Falsifiability Feels Like a Constraint, Yet Functions Like a Superpower
Falsifiability is usually presented as a limitation: only phenomena that can be tested are suitable for scientific method. But that framing misses its deeper role. Falsifiability is not merely a rule for exclusion. It is a design principle for making claims that can survive contact with the world.
A claim that cannot fail under any observation may sound strong, but in practice it is often empty. If nothing can count against it, then nothing can sharpen it either. By contrast, a claim that exposes itself to possible disproof becomes more informative, because it stakes out a territory where reality can push back. The scientific method does not just ask for evidence. It asks for risk.
This is where the monkey puzzle becomes more than a puzzle. Suppose you want to know the number of possible movements that lead to at least one collision. The direct path is messy: you could try to enumerate every possible movement and inspect whether any collision occurs. That quickly becomes unwieldy. The smarter route is often to define the boundary event precisely. What counts as a collision? What counts as no collision? Once those boundaries are crisp, counting becomes possible.
Science works similarly. To test an idea, you need to define the observation that would make the idea fail. A theory that says, “Something will happen, somehow, eventually,” is like a monkey problem with no collision rule. It cannot be evaluated because it has no operational edge. A theory that says, “Under these conditions, this outcome should occur, and if it does not, the theory is in trouble,” becomes testable. The claim has acquired geometry.
This is why the most powerful scientific statements often feel narrow before they feel profound. They draw a line in the sand. That line is not a weakness. It is the beginning of traction.
Counting What You Do Not Want: The Hidden Logic of Progress
There is a beautiful trick in many counting problems: instead of counting the desired outcomes directly, count the impossible, the forbidden, or the boring case, then subtract. In the monkey problem, if you want the number of ways at least one collision happens, it is often easier to count the ways no collision happens and subtract from the total.
That strategy has a larger intellectual meaning. Humans are often bad at describing success directly. We know, vaguely, that we want a good outcome, a strong explanation, a workable plan. But vague goals are hard to compute against. Failures, on the other hand, are often easier to identify. A collision is visible. A falsified prediction is visible. A system that breaks under pressure is visible.
This suggests a useful mental model: progress is frequently measured by narrowing the set of ways something can fail. In science, that means designing experiments that rule out explanations. In life, it means building habits, institutions, and products that reduce the chance of catastrophic breakdowns. In reasoning, it means not asking only, “What could make this true?” but also, “What would make this impossible?”
Think about airplane engineering. A plane is not considered safe because engineers can prove it will never fail. It is considered safe because engineers have defined, tested, and eliminated countless failure modes. The goal is not perfect certainty. The goal is a shrinking space of plausible collision.
Or consider a medical diagnosis. Good clinicians do not merely search for evidence that supports a favorite explanation. They also look for signs that would exclude it. The absence of a symptom can matter as much as its presence. What matters is whether the pattern survives adversarial testing.
This is the common thread between counting collisions and testing hypotheses. In both cases, the real work is not to admire the total space of possibilities. It is to understand the structure of the forbidden zone.
The Deep Tension: Human Meaning Wants Breadth, Science Demands Boundaries
Here is the tension that connects the two ideas most deeply: human experience often feels richer than what can be cleanly tested, while reliable knowledge often requires exactly that kind of narrowing.
Many important parts of life are hard to formalize. Love, grief, meaning, courage, moral responsibility, aesthetic judgment, and identity resist simple measurement. That does not make them unreal. It means they may not always belong to the domain where the scientific method, in its strictest sense, can fully operate. We can study correlates, patterns, and consequences, but the full texture may exceed objective observation.
And yet, when we do want to know something with discipline, we need boundaries. We need testable claims. We need to know what would count as evidence against us. We need a way to say, “If this happens, the idea fails.” Without that, we drift into stories that may be emotionally satisfying but intellectually untouchable.
This creates a productive split. Some domains ask for interpretation, others for falsification. The mistake is to use one mode where the other is required. A poem should not be judged like a laboratory result. A vaccine hypothesis should not be judged like a poem. But the reverse mistake is just as dangerous: treating a vague intuition as if it were a tested fact, or treating a testable claim as if interpretation alone can save it.
The monkey collision puzzle illuminates this split elegantly. If you insist on solving it by watching each monkey individually, you may drown in complexity. If you insist on turning every human question into a fully testable proposition, you may flatten what matters most. The better path is knowing when to count, when to test, and when to admit a boundary.
Not everything important is scientific, but anything scientific must be made vulnerable to failure.
That sentence is not a dismissal of human richness. It is a defense of intellectual honesty.
A Better Framework: Three Questions for Any Claim
To connect these lessons into something usable, consider a three question framework for any claim, theory, or decision.
1. What is the collision?
Every serious inquiry needs a clear event that matters. In science, this is the observation that would count against the hypothesis. In life, it might be a failure condition: a budget overrun, a health marker, a relationship pattern, a deadline missed. If you cannot name the collision, you cannot meaningfully test the system.
2. What is the no collision case?
Counterintuitively, understanding what does not happen can be more revealing than chasing what does. What would success look like in the absence of breakdown? In the monkey puzzle, counting the no collision case clarifies the whole problem. In reasoning, knowing the stable baseline helps separate signal from noise.
3. What remains beyond the test?
Some questions can be made testable. Others can only be approached indirectly. This is not failure, it is category awareness. A mature thinker knows when falsifiability is appropriate and when the question itself lives in a different register. The danger is not ignorance of every answer. The danger is confusion about which kind of answer is possible.
This framework is useful because it turns abstraction into discipline. It reminds us that clarity is not a luxury. It is the price of being able to learn from reality.
Imagine a founder launching a new product. The collision could be a drop in retention below a threshold. The no collision case could be a cohort that behaves consistently over time. The beyond test layer might include the product’s cultural meaning, which cannot be reduced to metrics alone. That founder needs all three questions, not just one.
The same applies to personal growth. If you want to change a habit, identify the collision. What exact behavior signals failure? Then define the no collision condition. What does a successful week look like? Finally, admit what cannot be fully instrumented, such as motivation, dignity, or purpose. You can measure parts of the journey, but not the whole soul of it.
Key Takeaways
- A claim becomes stronger when it can fail. If nothing could disprove it, it may be too vague to teach you anything.
- Counting the forbidden case is often the shortest path to understanding. In problems and in life, clarifying what must not happen can reveal the structure of what should happen.
- Not every meaningful question is testable, but every testable claim must be bounded. Know when to use measurement, and when to use interpretation.
- Progress often means shrinking the space of collisions. Whether in science, engineering, or personal habits, improvement comes from eliminating failure modes.
- Use the three question framework: What is the collision? What is the no collision case? What remains beyond the test?
The Real Lesson: Knowledge Is a Boundary Making Machine
We often think of knowledge as accumulation, as if wisdom were just a bigger pile of facts. But the deeper pattern is different. Knowledge is what happens when we draw a boundary around reality precise enough for reality to push back.
That is why falsifiability matters. It does not reduce truth to measurement. It gives truth a place to stand where it can be challenged. And that is why a counting problem about monkeys colliding on a polygon is more than a toy. It reveals a general strategy for making complexity tractable: define the event, define the exception, and count the space between them.
In both science and reasoning, the decisive move is not always to explain more. Sometimes it is to specify less, but with greater precision. A smaller claim can be a stronger claim if it is testable. A narrower boundary can reveal a larger structure.
So the next time a problem feels too broad, ask a different question. Not, “How do I explain everything?” But, “What would count as contact with reality?” That is where collisions become evidence, and evidence becomes knowledge.
And that may be the most useful paradox of all: we learn the most not when we try to cover every possibility, but when we make a careful claim and let the world decide whether it survives.
Sources
Hatch New Ideas with Glasp AI 🐣
Glasp AI allows you to hatch new ideas based on your curated content. Let's curate and create with Glasp AI :)
Start Hatching 🐣