Sleak
Knowledge & Onboarding

Philipp HeidekerAugust 25, 202614 min read

How Do You Verify Someone Actually Understood a Methodology?

A finished course proves attendance, not understanding. How to verify that someone can apply a methodology, and how to turn that into real evidence.

TL;DR. A finished course tells you nothing about whether someone understood a methodology. What sticks is what a person applies themselves and has to defend against questions in a live conversation: why the steps run in this order, what they would do in a case that never appeared in the material, and when the methodology stops applying. That conversation used to depend on a human being with time, which is why it was never available to everyone. Sleak runs it for each person individually through the AI Coach: first in Coaching Mode as a knowledge check in dialogue, then in Training Mode as a practised conversation, both scored against a Scorecard, the standard leadership sets in advance. This piece covers how to recognise real understanding, how to turn it into evidence that holds up, and where that evidence stops.

Key Takeaways

  • At SUXXEED, average conversation quality improved by 14.5 percent after Sleak was introduced, and objection handling scores rose by more than 42.4 percent, on a base of almost 15,000 simulated conversations (SUXXEED case study).
  • Würth started with an 80-person pilot and has since extended access to more than 5,000 employees across sales and leadership, at an average rating of 4.76 out of 5 (Würth case study).
  • At FEGA & Schmitt the same approach runs across four functions, sales, procurement, leadership and HR, now covering more than 900 employees in over 100 teams at 4.83 out of 5 (FEGA & Schmitt case study).
  • All three figures come from Training Mode, where people practise against a virtual counterpart. What they establish is the effect of a repeatable conversation. Reading them as evidence about knowledge checks would be an extrapolation.
  • Proof of understanding demonstrates judgement, not conversational skill. It is what stops practice from drilling in the wrong thing, and it is no substitute for the practice itself.

Two people on your team attended the same methodology workshop, both walked out certified. One of them hits a situation a week later that never came up in the workshop and makes the right call. The other applies the methodology to a case it was never built for and does not notice. On the course report they look identical. Having read a recipe is not the same as being able to cook once an ingredient is missing.

None of this is a failure of L&D. A course produces exactly one reliable output: a record that somebody worked through it. Whether judgement came out the other end leaves no trace in any system, so it never got measured. This piece sets out what understanding a methodology means in practice, why self-assessment is useless for exactly this question, how you recognise the difference, and how Sleak runs the check one person at a time. One thing most vendors in this space leave unsaid, so here it is first: even a passed understanding check says nothing about whether the methodology lands in a real conversation. That is a second measurement.

What does it mean to understand a methodology?

Someone understands a methodology when they can explain what each step is for, apply it to a case they have never seen, and say when it does not apply. Everything below that is knowledge of the steps, and the steps are the part a course teaches well anyway.

Four stages sit between attending a workshop and understanding a methodology. The first is attendance: the person was there and worked through the material. The second is knowing: the person can list the steps. The third is understanding: the person knows what each individual step is for and what breaks without it. The fourth is applying: in a situation that has never come up before, the person recognises what is needed now. Take the third stage away and you get the single most common failure pattern. The steps get worked through cleanly, in a situation where they achieve nothing.

Whether someone can still retrieve the content at all is a separate question, and the piece on making sure a team masters product knowledge already covers it. This one sits a floor above: not whether a person can produce the material, but whether they can decide anything with it.

Why does a finished course prove nothing about understanding?

A course completion documents that the material was worked through, and that is the only quantity a learning platform can report reliably. Everything past that gets asked about rather than tested, and the answer comes from the person you were trying to assess.

The state of common practice is sobering. Most companies judge the success of a training programme on participant feedback after the session, topped up with the line manager's impression. Where measurement does happen, it happens at programme level: satisfaction surveys, transfer questionnaires a few weeks later, indicators like error or complaint rates. Those are useful instruments for asking whether a programme moved anything overall. For the question of whether one specific person understood one specific methodology, they are silent.

Behind that sits a property of the format rather than a bad decision. Steps can be written down, distributed and ticked off. Judgement cannot be written down, because it only exists inside a concrete situation. So what got measured was always the part that leaves a trace. It also explains why the gap stays quiet for so long. No report shows it. It surfaces in a decision somebody else makes three weeks later, and by then nobody is looking for the cause in a course.

Why does almost everyone overrate their own understanding?

Because people reliably overestimate what they can explain, and the effect does not show up for plain sequences of steps at all. That makes the miscalibration a property of the knowledge type rather than a character flaw in any individual.

Rozenblit and Keil (2002) demonstrated this across twelve studies. In one of them, participants rated on a seven-point scale how well they understood a mechanism. They then had to write it out step by step and rate themselves again. After that they answered a single follow-up question that required the critical detail, and rated themselves a third time. The mean fell from 3.89 to 3.10 and then to 2.49. Independent raters who scored the same written explanations later arrived at 2.72, closer to the late self-ratings than to the early ones. The correction was genuine recalibration rather than after-the-fact modesty.

The finding that matters for people development sits in the same paper. The authors compared the effect across knowledge types. For procedures, sequences like how to file a tax return, it did not appear at all and self-ratings held steady (minus 0.17 scale points). For facts it was small (plus 0.29), for narratives similar (plus 0.35), and for explanatory knowledge about mechanisms it was large (plus 0.92). A methodology has both faces. The steps feel like a procedure, the reasoning behind them is an explanation. So self-assessment is dependable in exactly the place where it is worthless.

Two details close the obvious escape routes. In Study 6 participants were warned in advance that they would have to explain themselves. The effect got smaller and did not disappear. And among the causes the authors name is how rarely people explain anything at all, which leaves almost nobody with a sense of how good they are at it. People who explain regularly know where their gaps are. That is the whole trick.

How do you recognise real understanding?

By how the person answers three questions that appear on no test: Why in this order? What would you do in this case? And when would you drop it? Each one fails differently, and the way it fails tells you what to work on.

The first question separates knowing from understanding. Someone who has only learned that step two follows step one can recite the sequence and still not explain what step one has to deliver for step two to work at all. The second question removes the template. The workshop had examples, the real situation does not, and anyone who only learned examples reaches for the nearest one and misses. The third question is the most uncomfortable and the most important. A methodology that is right for everything is right for nothing, and anyone who does not know its limits will apply it where it does damage.

The questionWhat a good answer looks likeWhat usually comes instead
Why in this order?The person says what each step is for and what goes wrong without itThe steps are listed cleanly, the reasoning stays unsaid
What would you do in this case?The person reads the new situation and works out what it calls forThe closest example from the workshop gets retold, even though it does not fit
When would you drop it?The person names conditions under which the methodology fails, and the signals that show itThe methodology counts as good for everything, limits never come up

The table explains why two people holding the same certificate behave so differently. All three questions share one property: the person has to produce the answer. Picking between supplied options covers none of them, and a test score of 90 percent says nothing about any of the three.

What turns that into evidence that holds up?

Evidence holds up when the standard exists before the conversation, the assessment cites the transcript, and both stay reconstructible afterwards. Drop any one of the three and what you have is an impression.

The standard comes first. If nobody has written down what a solid answer to those three questions has to contain, every assessment is a matter of taste. At Sleak that standard is the Scorecard: the rubric leadership writes down before anyone gets assessed. Then comes the citation. A score on its own does not tell the person what to change, and does not tell you whether the assessment holds. A record showing exactly where an explanation ran thin tells the person what to work on the next morning.

At Sleak the Knowledge Coach runs this check. It is Sleak's Coaching Mode, the KNOW half of the AI Coach. The AI Coach is the always-available coaching instance every person talks to, and Coaching Mode (KNOW) is its dialogue half: explaining, probing, building knowledge. The company's own sources, methodology decks, playbooks and process documentation, sit in a Knowledge Repository with access controlled per team. The questions come from those sources, not from general model knowledge. The dialogue-based way methodology and product knowledge get built is described in full on the product page. Scorecards and conversation partners are versioned, with a draft and an immutable published state, so it stays visible later which version somebody was assessed against. The transcript is what persists, audio recording is off by default, and changes run through a log that is only ever appended to.

That changes what you can see as a manager. Instead of a completion rate you see who can carry the methodology in a conversation and who drops out the moment the case is new. More than 900 teams now work on the platform, with over 14,000 registered users (Sleak platform data, 08/26). An honest reading belongs with it: reconstructible means checkable, not correct. How good the evidence is depends on the Scorecard, and anyone introducing something like this spends most of the effort on the rubric rather than the tool.

Does explaining a methodology prove it lands in a real conversation?

No. Proof of understanding demonstrates judgement, and judgement is not yet conversational skill. The objection is fair: someone who can explain a methodology cleanly is not thereby holding a good conversation, and any check claiming otherwise has just moved the problem one step along.

These are two separate measurements. The first tests whether a person reaches the right decision in an unfamiliar situation. Whether they then carry it out, under time pressure, against pushback and in their own words, only shows up in the conversation itself: in practice conversations with a virtual counterpart, scored against the same Scorecard. At Sleak that half is called Training Mode (DO), and how voice-based practice conversations with virtual counterparts are built sits on that page. You want the understanding check first, because practising on a faulty foundation drills in the wrong thing.

What the classic format still does better belongs here too. For first contact with a methodology, a well-built module does the job in a fraction of the time, and a phase model belongs on the desk as a reference rather than in anybody's memory. Where a rule only calls for documented acknowledgement, a test is exactly the right instrument. Courses and documents are not the problem. They become one the moment a completion date stands in for judgement.

In practice: what changes when the course becomes a conversation

Once knowledge gets asked for and applied rather than only distributed, how people show up in conversations shifts measurably. The strongest numbers come from the practice half, which is where the effect is best documented.

At SUXXEED, employees and candidates have completed almost 15,000 simulated conversations. Average conversation quality across the platform improved by 14.5 percent, and objection handling scores rose by more than 42.4 percent. That is the point where knowing turns into doing, and it does not happen in a course. It happens in repetition with feedback. An honest reading belongs with it: those values measure Training Mode, where people practise against a virtual counterpart, and not the knowledge check itself. What they show is the effect of the repeated conversation, and the step across to understanding remains an extrapolation.

The second point is volume. Würth began with a pilot of around 80 participants and has since extended access to more than 5,000 employees across sales and leadership, at an average rating of 4.76 out of 5. This is precisely where the old format broke. A person who listens, probes and refuses to accept a thin answer was always the best check available, and it was never available to 5,000 people. The method was never the problem. Its price was.

And the standard is not tied to one discipline. At FEGA & Schmitt, an electrical wholesaler with around 1,400 employees across roughly 60 locations, Sleak runs in sales, in procurement for supplier negotiations, in leadership for feedback and development conversations, and in HR for faster onboarding. What began as a focused sales initiative now covers more than 900 employees in over 100 teams, rated 4.83 out of 5. A negotiation methodology, a leadership conversation guide and an onboarding path differ completely in content and not at all in how they get checked. In all three cases you ask the same three questions. Whoever sets that standard is in the end not L&D alone but the manager who is accountable for the outcome, and for them the entry point is the conversations they already have to hold.

FAQ

Is a certificate enough to prove someone understood a methodology?

No. A certificate establishes that a person worked through a programme, which is a statement about attendance. Whether they can explain what the individual steps are for, whether they transfer the methodology to a new case, and whether they know its limits appears on no certificate, because none of it was ever asked.

Who decides what counts as a sufficient answer?

The manager, together with the subject-matter owners, before the first conversation. That is the expensive part and also the one most efforts fail on: the material exists, but nobody ever wrote down what should come out at the end. Without that standard every assessment is one person's opinion, whoever or whatever produces it.

Can understanding be compared fairly when every conversation runs differently?

What gets compared is the standard, not the conversation. Everyone is assessed against the same version of the same rubric, and the assessment cites the transcript, so it can be reviewed afterwards. What the method does not deliver is a guarantee of correctness. It makes an assessment reconstructible, not necessarily right.

Does this work for methodologies outside sales?

Yes. The three questions only assume that a methodology has an order worth explaining, gets applied to new cases, and runs into limits somewhere. That holds for an escalation procedure in service just as much as for a negotiation approach in procurement or a feedback model in leadership. The content of the Scorecard changes, the logic of the check does not.


Related articles

More on practice, feedback and getting better at the conversations the job actually turns on.

Knowledge & Onboarding

How Do You Make Sure Your Team Actually Masters Product Knowledge?

Product knowledge exists when a person can retrieve and explain it under questioning, not when they receive a document or pass a test.

Philipp Heideker14 min read
Read article
Knowledge & Onboarding

Customer Success Onboarding with AI: How New CSMs Become Productive in 14 Days Instead of 60

AI-driven CSM onboarding cuts ramp-up from 60 to 90 days down to 14, with scorecard-based measurability per skill dimension.

Philipp Heideker11 min read
Read article
Knowledge & Onboarding

Sales Rep Onboarding with AI: How to Cut Ramp-Up Time in Half

AI-powered sales onboarding cuts time-to-productivity in half. See how the 90-day program works, with scorecards, metrics, and its real limits.

Philipp Heideker8 min read
Read article