PD That Works Build my cycle →
Evidence · For school leaders

What we measure, and what we don't.

By Yechiel · Founder, PD That Works. Yechiel builds PD That Works — the workshops, the daily AI coach, and the monthly faculty reports — and wrote every page on this site.

Most PD evidence fails in one of two directions: it measures the workshop instead of the classroom, or it measures the classroom with instruments that cannot reach it and dresses the result up as proof. This page is the inventory — every kind of evidence this product produces, labeled honestly, with the claims we will not make listed alongside.

Every number carries a type

Internally, every metric and report element is tagged with what kind of evidence it is. Those labels are now visible on the leader's dashboard too, next to the sections they describe, because a number whose provenance is invisible invites the reader to assume the strongest possible reading.

Evidence typeWhat it meansExampleStatus
Product telemetryWhat happened in the applicationA check-in was completedIn use
Teacher self-reportWhat a teacher says they tried or experienced"I used wait time today"In use
Voluntarily sharedContent a teacher explicitly chose to shareA shared classroom winIn use
Leader-enteredObservation or artifact entered by a leaderA walkthrough countLater release
Verified outcomeAn externally validated resultNot offered

The bottom row is the important one. Nobody in this category of product can currently give you externally validated outcomes from a 30-day cycle, and we would rather say so on a public page than let the word "proof" do quiet work in a sales conversation.

What gets captured, and when

Day 1 — the baseline. One goal in the teacher's own words, how often they report having used the practice in the prior five teaching days, a confidence rating on a stable five-point scale, and a sentence of classroom context. Only what will actually be used later. Without this, nothing that follows is movement.

Days 2–29 — the record. Whether the check-in happened, what the teacher reported trying, whether they adapted the practice or set it down, whether they asked for help, an optional weekly confidence rating, and anything they chose to share.

Day 30 — the close. The same behavior-frequency and confidence questions asked at baseline, so the comparison is like-for-like; whether the practice was refined, retained, replaced, or abandoned; whether the coaching was useful; whether the teacher would continue; and an optional comment.

How the report is allowed to talk

The reporting language is constrained, and the constraint is enforced in the software rather than left to whoever writes the summary.

We sayWe never say
"Teachers reported trying…""Instruction improved by…"
"Participation remained…""The initiative caused…"
"The most common implementation pattern was…""The data proves…"
"Confidence increased from X to Y among N respondents.""Teachers resisted…"
"This report does not include classroom observation or student outcomes.""Practice changed" with no evidence source attached

The right-hand column is not a list of things we find distasteful. Each one is a claim this data cannot support: an inference about a classroom nobody observed, a causal claim with no counterfactual, a judgment of teachers dressed as a finding. And low engagement is never reported as resistance — it may as easily be workload, absence, a poor fit, or a teacher quietly deciding the goal was wrong.

Where a number is withheld

Some sections of a report come back deliberately empty.

Each of these makes the report less impressive and more trustworthy, which is the trade we would make every time. A blank section says the floor held.

The honest limits, stated plainly

What this is good for

Within those limits the evidence answers the question a board actually asks, at the level it can honestly be answered in a semester: teachers engaged with follow-up at this rate, here is what they reported working on, here is how their confidence on their own goals moved, and here is what they chose to share about it. That is a real answer about the weeks after the workshop, which is more than an attendance sheet has ever offered.

Common questions

Does this measure whether teaching actually improved?

It measures what teachers report trying, how consistently they engaged, and how their confidence on their own goal moved. That is self-report and telemetry, not observation. A report claiming observed instructional improvement from this data would be overstating it, so ours does not.

Do your reports include student outcomes?

No, and each report says so in its own text. Those effects take longer than a cycle and are entangled with everything else a school does.

Can I add my own observation data?

A later release will allow one cohort-level measure a school already collects — a brief walkthrough rubric, an artifact count, an existing survey item — kept separate from private coaching data and labeled as leader-entered. It is not available today.

Why is part of my report blank?

Because the cohort was too small for that section to be shown without identifying someone. The floor is five teachers for themes and confidence trends. The report names the suppression rather than filling the space.

Read one in full

A complete Day-30 school report, so you can judge the evidence yourself rather than take this page's word for it.

$495 · One 30-day cycle · Up to 25 teachers · No auto-renewal.