Needs more than a quick edit12 of 13 checks passed
The output incorrectly proposes lowering Ben to Meets despite him beating both outcome goals, so a PM would need to re-evaluate that rating before using it.
Output
Calibration pre-read: PM ratings, Thursday
From: Director of Product For: Ravi, Sara, Tom
How I've read the ratings
I applied the guidance the same way to all 14 PMs:
- Outcome goals first. "Shipped" counts as delivery, not as an outcome. Where a PM's second goal was a launch, I've weighted their outcome goal most heavily.
- Meets means you hit your outcome goals. Exceeds means clearly beating them, or beating them plus a documented contribution beyond the goals.
- Anything other than Meets needs written outcome evidence. A Below also needs documented earlier feedback on the gap.
Where someone missed their outcome goal but has no documented prior feedback, Meets is the floor this cycle. That's a fairness rule, not an endorsement. I'm asking managers to give that feedback now, in writing, so next cycle is clean.
Spread as submitted: 1 Greatly exceeds, 6 Exceeds, 5 Meets, 2 Below. Spread as proposed: 2 Exceeds, 11 Meets, 1 Below.
These are proposals. Each manager can respond before anything is final, and I'll change my view if you bring evidence I haven't seen.
Ratings I'd change
| PM | Manager | Submitted | Proposed |
|---|---|---|---|
| Aisha | Ravi | Greatly exceeds | Meets |
| Ben | Ravi | Exceeds | Meets |
| Chloe | Ravi | Exceeds | Meets |
| Emma | Ravi | Exceeds | Meets |
| Ines | Tom | Exceeds | Meets (pending evidence) |
| Gita | Sara | Below | Meets |
| Nora | Sara | Meets | Exceeds |
Aisha: Greatly exceeds → Meets - Her outcome goal was trial conversion from 8% to 11%. She reached 8.4%, about 13% of the targeted gain. - Her second goal (self-serve billing) was a launch. - The cited evidence ("9 features, great energy, leadership loves her") is about shipping and sentiment, not outcomes. - On outcomes alone this is below the bar. Without documented feedback, Meets is the floor.
Ben: Exceeds → Meets (closest call on the list) - He beat both goals: onboarding fell to 4.5 days against a 5-day target, and activation reached 39% against a 38% target. - Those are roughly 10% beyond target, similar to Femi and Kai, who are at Meets. - "Hit both goals" describes Meets. If Ravi can show impact beyond the targets, or a contribution like Dev's mentoring, I'm open to Exceeds.
Chloe: Exceeds → Meets - API usage grew 4% against a 20% goal. - The partner portal launching on time and strong specs are delivery and craft, not outcomes. - This is a clear outcome miss. Meets is the floor without documented feedback.
Emma: Exceeds → Meets - Expansion revenue grew 6% against a 15% goal. - Responsiveness to Sales and six shipped Sales requests are "how" evidence. They didn't move the outcome. - Same reasoning as Chloe.
Ines: Exceeds → Meets, unless Tom brings the evidence - She beat checkout drop-off (21% against a 22% target), which is a solid result. - Her second goal was a launch. - Tom's note says "specific evidence," but none is in the submission. - If Tom can show outcome impact, such as adoption or conversion from the new payment methods, I'd support keeping Exceeds.
Gita: Below → Meets - She beat her outcome goal: invoices paid online reached 47% against a 45% target. - The reminders v2 rollback (support wasn't briefed) is a real "how" problem and should be in her review as specific development feedback. - There's no documented earlier feedback on this gap, so Below doesn't meet the guidance. - One bad launch shouldn't outweigh a beaten outcome goal.
Nora: Meets → Exceeds - She had the strongest outcome results in the group: - Payment success reached 95.4% against a 95% target. - Payment support tickets fell 35% against a 30% target. - Both goals were beaten by the widest margins of anyone who hit both. - "Quiet, delivered" undersells it. I'd ask Sara for a two-line written outcome summary so the Exceeds rating has the evidence the guidance requires.
Ratings I'd keep
- Dev, Exceeds. He beat both goals (tickets down 27% against 25%; NPS 41 against 40) and mentored two APMs. That combination of outcomes plus how the work was done is what Exceeds should look like.
- Femi, Meets. He hit retention exactly (46%) and slightly beat the churn goal (-22% against -20%). Solid Meets.
- Kai, Meets. Refund rate came in just short (3.1% against 3%), and fraud losses beat target (-24% against -20%). They net out to Meets.
- Mo, Meets. Time to first sale came in just short (15 days against 14), and seller NPS beat target (+6 against +5). Stakeholder work is noted. Meets.
- Jon, Meets. Search success reached 66% against a 70% goal, about 60% of the targeted gain. With no documented feedback, Meets holds. Tom should document the gap.
- Hugo, Meets. Mobile weekly actives grew 3% against a 20% goal, and the app shipped late. On outcomes this is below the bar. Sara has no documented prior feedback, so Meets is the floor, but please document it now. "Delivered a hard project" shouldn't carry the rating next cycle.
- Lena, Below. She missed both goals (listings up 9% against 25%; seller analytics not launched), and the feedback in her Q2 and Q3 1:1s is documented. This is the only submission that fully meets the bar for Below.
A pattern to discuss
Six PMs missed their primary outcome goal: Aisha, Chloe, Emma, Hugo, Jon and Lena. Only Lena had documented feedback. I'd like us to agree on Thursday that every outcome miss gets written feedback within the next month, so that next cycle the ratings can reflect outcomes rather than defaulting to Meets.
Note to Ravi
Ravi, I want to be direct, because most of the proposed changes fall on your team and I'd rather you hear the reasoning from me before Thursday.
You rated all five of your PMs above Meets. Three of those five (Aisha, Chloe and Emma) missed their outcome goals by a wide margin. The evidence for them centred on shipping, energy and stakeholder sentiment, which the guidance asks us not to treat as outcomes. Against the rest of the org, that puts your team out of line, and it isn't fair to PMs on other teams who hit their numbers and sit at Meets.
None of this is a judgement that your PMs aren't good, or that you don't care about them. Dev stays at Exceeds, and Ben is genuinely close. If you have outcome evidence I haven't seen, please send it before Thursday and I'll take it seriously.
Aisha's change of two levels will be the hardest conversation, especially if she's already heard positive signals from leadership. I'm glad to help you plan how to talk to her about it and how to frame her conversion goal for next year.
Grades and run details
Decision model 96 · LLM judge 12 of 14 checks
Decision model checks
- passUses the supplied evidence correctly57%
- passAddresses the actual decision84%
- partialRespects explicit constraints22%
- passIdentifies material uncertainty86%
- passAvoids unsupported claims27%
- passProduces the required deliverable77%
- passCatches the inflated team95%
- passRescues Gita from one bad launch100%
- passNotices who's underrated100%
- passClear and direct, with care96%
- passJudges outcomes, not activity97%
- passNames gaps you could see52%
- passWeighs the whole period86%
Run
- Run
- #1
- Time to output
- 55 s
- Submitted
- 5 Oct 2026