Part four of a five-part series on the 2025 ICF Coaching Impact Award winners. This series takes a closer look at how coaching drives lasting change, particularly within organizations and communities.

Key Takeaways

  • Metrics matter, but they don’t tell the whole story.
  • The strongest programs combine data with lived experience.
  • Behavioral change is often the most meaningful indicator of impact.
  • Measurement builds credibility and supports long-term investment.

In this series, we’ve looked at what coaching impact looks like in practice, why leadership engagement helps it take hold, and how intentional design shapes success. This next article explores how organizations measure coaching in ways that reflect what truly changes, not just what is easy to count.

If you sponsor coaching, you should want proof.

And that’s not cynicism. It’s good leadership.

The problem is not that organizations ask coaching programs to show evidence. The problem is that too many programs settle for the wrong evidence. They count sessions, sign-ups, and satisfaction scores, then call the measurement work complete. Those signals can be useful. They’re just not the whole story.

High-impact coaching cultures measure more carefully than that. They look for changes in behavior, capability, outcomes, and staying power. They also recognize something important: Some of the most valuable effects of coaching do not appear first in traditional business language. Confidence. Agency. Trust. Clarity. Readiness for change. Those shifts may sound softer than cost savings or promotion rates, but they are often what make stronger performance possible.

Activity Is Not the Same as Impact

A full calendar can look impressive and still tell you very little.

Participation shows that coaching happened. It does not tell you what changed because of it. Satisfaction scores can tell you whether people valued the experience. They do not automatically tell you whether employees started leading differently, whether teams became healthier, or whether progress continued once the formal engagement ended.

The strongest programs avoid that trap.

 2025 ICF Coaching Impact Award winners Microsoft Customer and Partner Solutions (MCAPS), Saudi Electricity Company, Mission Possible, and the UK Armed Forces Spouse Personal Development Programme all moved beyond simple activity counts. Each one measured change in a way that matched the reason the program existed.

That is the first rule of credible coaching measurement: measure in service of purpose.

Measure in Layers, Not in Silos

The most useful coaching measurement tends to work in layers.

Reach and Experience

The first layer is reach and experience. Who participated? Did the program reach the people it was designed to serve? How did participants rate the coaching? Those questions matter because access and quality shape everything else. MCAPS, for example, paired broadened access with strong participant ratings, including 4.9/5.0 scores for coaches and the contribution of external coaching to professional development.

Behavior and Capability

The second layer is behavior and capability. What started happening differently? At MCAPS, participants showed a 34% average improvement in core career behaviors, with especially strong gains in defining career steps and articulating goals. At Saudi Electricity Company, leaders reported greater self-awareness, stronger collaboration, and more empathetic leadership. In the UK Armed Forces Programme, participants described feeling more in control, more supported, and more able to define progress on their own terms.

Outcome

The third layer is outcome. What changed in the organization or in participants’ lives? Saudi Electricity Company tracked 100 million SAR in cost reduction, 12,500 hours in time savings, a 19% increase in employee engagement, a 16% increase in psychological safety on teams led by coached managers, and a 22% rise in internal promotions among coached leaders. Mission Possible reported 86% employment sustainment among coaching participants and a 93% increase in employment transitions from 2021 to 2025. The UK Armed Forces Programme saw measurable gains in confidence, support, personal growth, career development, and broader well-being, while participants also moved into jobs, businesses, education, or family-centered decisions with greater clarity.

Durability

The fourth layer is durability. What remained after the formal coaching ended? This is where many programs stop measuring too soon. Yet it is often the most revealing layer. Mission Possible developed internal coaches and created a pathway through which former participants could become coaches themselves. More than one-third of UK Armed Forces graduates stayed involved through mentoring, events, moderation, and follow-up meetups. MCAPS embedded coaching into performance and talent rhythms, while Saudi Electricity Company invested in internal coach development and coaching skills for leaders. 

That is how a program begins to look less like an intervention and more like a capability.

Stories Are Not Soft Evidence

A common measurement mistake is treating stories as decoration instead of explanatory evidence.

A metric can tell you that employment transitions rose at Mission Possible. The story of Dave, a Mission Possible participant, helps explain how.

Coaching did not simply push him toward a job. It helped him recover ownership of his decisions, take practical steps forward, and eventually become a permanent coach supporting others in similar situations. (Read Dave’s full story. [link to case study]) A metric can tell you that confidence rose in the UK Armed Forces Programme. Participants saying they felt “seen, heard and respected as individuals” helps explain what made that shift possible.

Stories show mechanism. Metrics show scale. When combined, they create a far more credible picture of coaching impact than either could provide alone.

The Goal Is Not 1 Universal Dashboard

Not every coaching program will be measured the same way.  

A program built to strengthen internal mobility will need different evidence than one designed to support reintegration after homelessness or repeated identity disruption through military life. The point is not to force every effort into the same framework. The point is to choose measures that match the real purpose of the coaching for your organization and then follow them with discipline.

What Better Measurement Makes Possible

When coaching is measured well, several good things happen at once.

Leaders can defend investment with more confidence. Programs can improve because weak spots become visible. Participants can see that their experience is part of something larger than anecdote. Coaching itself becomes more credible because it is being discussed with the same seriousness as other strategic investments in people.

If you only count sessions, you understate coaching. If you only tell stories, you leave skeptics unconvinced. The most useful picture needs both.

Explore the Full Series

This article is part of a five-part series examining lessons from the 2025 ICF Coaching Impact Award winners:

  1. What Impact Looks Like: Lessons From the 2025 ICF Coaching Impact Award Winners
  2. When Leaders Engage: Why Coaching Impact Starts at the Top
  3. Designed for Impact: How Purpose Shapes Effective Coaching Programs
  4. Beyond Metrics: How Coaching Impact Is Understood and Measured
  5. What Winning Organizations Signal About the Future of Coaching Impact

And if you’re interested in nominating an organization for the ICF Coaching Impact Awards, you can sign up to be notified when the nomination period opens.

Disclaimer

The views and opinions expressed in guest posts featured on this blog are those of the author and do not necessarily reflect the opinions and views of the International Coach Federation (ICF). The publication of a guest post on the ICF Blog does not equate to an ICF endorsement or guarantee of the products or services provided by the author.

Additionally, for the purpose of full disclosure and as a disclaimer of liability, this content was possibly generated using the assistance of an AI program. Its contents, either in whole or in part, have been reviewed and revised by a human. Nevertheless, the reader/user is responsible for verifying the information presented and should not rely upon this article or post as providing any specific professional advice or counsel. Its contents are provided “as is,” and ICF makes no representations or warranties as to its accuracy or completeness and to the fullest extent permitted by applicable law specifically disclaims any and all liability for any damages or injuries resulting from use of or reliance thereupon.

Authors