When Measurement Stops at the Dashboard, Learning Never Reaches Practice
Hatched by Wai-Ling Fong
Jun 28, 2026
10 min read
2 views
78%
The real question is not whether training worked
A faculty development program can look successful in a spreadsheet and still fail in real life. Participants may rate it highly, complete every module, and say they learned a great deal, yet a year later their teaching practice may look almost unchanged. That gap is not a minor evaluation problem. It is the central puzzle of educational improvement: how do we know that learning has become practice, and practice has become impact?
This matters because most institutions still confuse evidence of participation with evidence of transformation. Attendance, satisfaction surveys, and end of workshop ratings are easy to collect, which makes them seductive. But a clean dashboard can hide a messy truth: people often leave a program inspired, then return to environments that quietly erase the change. If evaluation stops too early, it rewards the appearance of learning rather than learning that survives contact with reality.
The deeper issue is not measurement alone. It is a theory of change. We need to ask what kind of evidence actually tells us that professional development has crossed the boundary from event to habit, from idea to classroom behavior, from immediate reaction to durable effect.
The hidden limitation of shallow success
Many evaluation systems are built like elevators that only go up to the second floor. They tell us who showed up, who liked the experience, and who can repeat the material afterward. Those are useful signals, but they are not the same as impact. In faculty development, especially online or hybrid programs, the most meaningful outcomes often emerge later, in more ordinary moments: a redesigned assignment, a changed discussion prompt, a new way of giving feedback, a willingness to experiment with assessment.
That makes evaluation harder, but also more honest. A participant may initially report enthusiasm, yet the real question is whether the program altered their teaching reflexes. Did it change what they notice, what they try, what they persist with when the semester becomes crowded and stressful? A one time survey cannot answer that. A year later, qualitative responses may reveal much more, because memory has had time to do its work: some practices fade, some survive, and some evolve into something new.
Think of faculty development like planting a tree rather than launching a product. If you inspect the seed after a week, you learn almost nothing. If you check the trunk after a year, you start to understand whether roots took hold, whether the climate was favorable, and whether the tree adapted. The point is not that early feedback is useless. The point is that short term reaction is not the same as long term growth.
This is where the familiar logic of evaluation becomes too narrow. In higher education, the problem is not simply that some models are imperfect. It is that they can unintentionally train us to value what is easiest to measure over what is most educationally meaningful.
What survives a year says more than what impressed someone for an hour
There is a revealing difference between momentary approval and enduring adoption. Someone can leave a session saying, “That was insightful,” and still teach exactly the way they did before. Another person may seem less enthusiastic at first, yet quietly begin using a new discussion structure, revising their feedback practices, or changing how they design online activities. The second case matters more, but it is easier to miss.
This is why longer horizon inquiry is so powerful. Asking participants a year later what they still use, what they changed, and what they abandoned uncovers the lived afterlife of a program. It reveals not only whether learning happened, but how it was translated into the realities of course planning, workload, and institutional constraints. In other words, it shows whether a program produced adaptation, not just agreement.
That distinction is crucial. Agreement is cheap. Adaptation is expensive. Agreement can happen in a workshop. Adaptation requires confidence, time, contextual fit, and repeated use. It often depends on whether participants can make the new practice compatible with their teaching identity. A pedagogical idea that feels elegant in a session may still be too fragile for a real classroom unless it can survive friction.
The most meaningful outcome of professional development is not that people liked it, but that it became thinkable and usable in their actual work.
This reframes evaluation from an audit of satisfaction into a study of transfer. Transfer is not automatic. People do not simply absorb good teaching and export it unchanged into their courses. They reinterpret, simplify, localize, and sometimes resist it. That is not failure. It is the process by which new knowledge becomes durable.
A better model: track the life cycle of change
If we want to understand whether faculty development matters, we need a more realistic mental model than a single score. A useful framework is to think in terms of the life cycle of change. Each stage reveals different things, and each stage requires different evidence.
1. Immediate response: Was the experience engaging enough to matter?
This is the entry point. If participants are confused, alienated, or overwhelmed, nothing else follows. Reactions are not trivial, because they influence whether people will return, recommend the program, or mentally dismiss it. But reaction only tells us that the door opened, not that anyone walked through it.
2. Cognitive uptake: Did participants understand the idea?
Here we ask whether they can explain the concept in their own words. This is where quizzes, reflections, and discussion can help. Understanding is necessary, but still insufficient. A person may understand a technique perfectly and never use it.
3. Practical translation: Can the idea be adapted to local conditions?
This is the hinge point. A new teaching strategy has to meet a real classroom, a real semester, a real workload, and real students. If a program does not help people answer the question “How would this look in my setting?”, it remains abstract. Practical translation is where good pedagogy becomes usable.
4. Behavioral persistence: Does it survive repetition?
One use is not enough. The real signal appears when a practice is repeated, adjusted, and incorporated into routine. Persistence matters because it shows that the new behavior has moved from novelty to habit.
5. Cultural ripple: Does it influence others or shift norms?
The deepest impact may not be individual. A faculty member who adopts a new approach may influence colleagues, graduate assistants, course teams, or departmental expectations. Impact spreads when practices become discussable, teachable, and contagious.
This framework matters because it prevents a common mistake: treating all outcomes as if they belong to the same category. They do not. A smile after a workshop, a new syllabus move, and a department wide change in teaching culture are different phenomena. They deserve different measures, different timelines, and different interpretations.
Why the scholarship of teaching and learning needs patience
The most interesting question in teaching and learning is not only “What works?” It is also “What lasts, where, and for whom?” That question is more demanding because it treats education as a living system rather than a controlled experiment. In a living system, effects are rarely immediate, linear, or uniform.
This is especially important in faculty development because the setting itself shapes the outcome. A program offered during a crisis, such as a rapid move to online teaching, may have very different effects from one delivered in a stable semester. Participants might initially adopt tools out of necessity, then later decide which ones were temporary and which ones were worth keeping. A year later, the long tail of that experience can be more revealing than the first wave of enthusiasm.
The lesson is not that we should abandon metrics. It is that we need time sensitive metrics. Some outcomes are immediate and fine grained, such as confidence or comprehension. Others are delayed and contextual, such as sustained use or shifts in teaching identity. If evaluation ignores that timing, it misreads the nature of educational change.
There is also a democratic dimension here. Inquiry into teaching should include people at different career stages, because early career educators, mid career faculty, and experienced teachers may carry change differently. The same program can function as a launchpad for one person, a refinement tool for another, and a reminder for a third. Good evaluation honors that diversity instead of pretending there is one universal pathway to improvement.
In that sense, scholarship of teaching and learning is not just about discovering effective practices. It is about learning how practices live in the hands of different people over time. That is a much richer object of study.
From proving value to understanding transformation
The deepest shift we need is philosophical. We often ask evaluation to prove that a program was worth the investment. But in complex educational work, proof is too blunt an instrument. The more useful goal is to understand how change happens, what conditions sustain it, and what kinds of support make it more likely to endure.
That suggests a different set of questions for faculty development:
- What changed immediately after the program?
- What changed six months later?
- What remained in use after a year?
- What was adapted, and why?
- What barriers prevented use even when the idea was persuasive?
- Which elements generated spillover effects beyond the original participant?
These questions turn evaluation into a narrative of transformation rather than a verdict. They also force institutions to face an uncomfortable truth: sometimes a program is intellectually strong but institutionally unsupported. The fault is not always in the design. Sometimes the surrounding environment is too misaligned to let the learning stick.
That insight should change how programs are built. A faculty development effort should not only teach a technique. It should help participants rehearse implementation, anticipate barriers, and plan for follow up. It should create opportunities to revisit practice after the initial session, because most change needs reinforcement. And it should gather evidence over time, not just at the moment of applause.
If you want to know whether a program mattered, do not ask only what participants thought of it. Ask what it changed in their Tuesday afternoon classroom.
That question is more concrete, and more revealing, than any satisfaction score.
Key Takeaways
- Do not confuse engagement with impact. Positive reactions and high completion rates are useful, but they do not show whether practice changed.
- Measure across time, not just at the end. Immediate feedback tells you about the session; delayed feedback tells you about adoption and persistence.
- Look for translation, not imitation. The best sign of learning is not exact replication of a workshop idea, but thoughtful adaptation to local teaching conditions.
- Use mixed evidence. Pair surveys with reflections, interviews, classroom artifacts, or follow up conversations to capture both breadth and depth.
- Design for durability. Build follow up, peer support, and implementation planning into faculty development so that change has a chance to survive real work.
The real measure of learning is what remains when enthusiasm fades
Educational programs are often judged at the moment they are easiest to praise. That is a mistake. The true value of faculty development appears after the applause, after the memory of the session has softened, when participants are back inside the friction of real teaching. What remains then tells us more than any immediate rating ever could.
This is the paradox at the heart of evaluation: the closer we get to genuine impact, the further we move from easy measurement. But that is not a reason to settle for shallow indicators. It is a reason to become more patient, more longitudinal, and more interested in the afterlife of learning.
In the end, the question is not whether a workshop was good. The question is whether it changed what someone does when no one is watching, when the semester is underway, when the classroom is messy, and when good intentions must become habit. That is where educational change becomes real.
Sources
Hatch New Ideas with Glasp AI 🐣
Glasp AI allows you to hatch new ideas based on your curated content. Let's curate and create with Glasp AI :)
Start Hatching 🐣