Usable with a quick edit12 of 12 checks passed
The output is complete, actionable, and respects all constraints; a PM could use it with at most minor wording tweaks.
Output
The loop I’d run
Replace the deck and six-person presentation with bounded, live work samples. Remove the “beer test.” Copperline needs someone who can diagnose conversion, ship weekly, and change other teams’ behavior—not simply present well.
Candidate commitment: 3 hours 10 minutes; no take-home or required preparation. Share the format and scorecard beforehand, offer flexible scheduling and reasonable accommodations, and use the same prompts and information for everyone.
| Round | Interviewers | Time | What it tests |
|---|---|---|---|
| --- | --- | ---: | --- |
| 1. Recruiter screen | Recruiter | 20 min | Scope, compensation, availability, and interest in hands-on growth ownership. Explain the process. Don’t assess polish or pedigree. |
| 2. Ownership interview | Amara | 35 min | Senior-level judgment and accountability. Examine one past growth initiative: candidate’s actual contribution, choices, results, failures, and learning. Probe both conversion and expansion experience. Reserve five minutes for candidate questions. |
| 3. Live growth diagnosis | Data analyst | 50 min | Analytical judgment and experiment design. Give a small synthetic Copperline dataset: trial funnel, cohorts, acquisition sources, activation behavior, expansion signals, and metric definitions. Ask what could explain 9% conversion, what they would investigate, and which experiment they would run first. Include a plausible expansion opportunity. |
| 4. Shipping workshop | Head of Engineering + senior designer | 40 min | Turn the previous round’s proposal into a one-week experiment with three engineers and a designer. Interviewers act as working partners, not an audience. Introduce a tracking gap and an engineering constraint. Test scope, customer experience, instrumentation, trade-offs, and rollout safety. |
| 5. Influence simulation | Head of Sales | 45 min | First, role-play a conversation with Sales resisting a proposed trial-follow-up change because of workload and incentives. Add a Marketing objection about lead quality. Then examine a real example of influencing peers without authority, including resistance, compromises, and outcomes. Reserve five minutes for questions. |
Use the same dataset, constraints, and role-play objections across candidates. Score reasoning, not whether they discover a predetermined “correct” growth strategy. The CEO does not conduct a separate fit interview.
Scorecard
Use 1 = clear concern, 2 = below bar, 3 = meets senior-role bar, 4 = exceptional, plus not observed. Each rating requires concrete evidence: a quote, action, artifact, or specific past example. “Very smart” and “great presence” are not evidence.
| Dimension / primary assessor | Strong: meets or exceeds bar | Weak: below bar |
|---|---|---|
| Growth and analytical judgment — analyst | Checks denominators, cohort maturity, segmentation, and data quality before explaining 9%. Separates correlation from causation; prioritizes plausible conversion and expansion opportunities using customer behavior and economics. | Treats aggregate conversion as a diagnosis; jumps to tactics; ignores acquisition mix or revenue quality; makes causal claims without evidence. |
| Experimentation — analyst | States hypothesis, mechanism, target population, primary metric, guardrails, and decision rule. Recognizes low-volume limitations. Distinguishes shipping weekly from obtaining statistically conclusive results weekly. | Proposes many tests without prioritization; relies on vanity metrics; declares victory prematurely; cannot say what result would change the decision. |
| Execution and customer judgment — Engineering + designer | Cuts scope to a credible weekly release; specifies measurement and rollback; collaborates with technical/design partners; protects trust in invoicing and payments workflows. | Requires a large rebuild; treats partners as order-takers; neglects instrumentation, reliability, or customer harm; cannot adapt to constraints. |
| Influence without authority — Sales | Understands incentives and workload; listens, negotiates a bounded pilot, establishes shared measures and explicit commitments, and follows through. Past examples show actual behavior change. | Relies on escalation, persuasion theater, or “getting buy-in”; blames peers; has no concrete mechanism for changing behavior. |
| Senior ownership and learning — Amara | Clearly distinguishes personal contribution from team results; connects conversion and expansion to durable revenue; owns misses and demonstrates changed judgment. | Claims collective wins as personal achievements; avoids failures; optimizes local conversion at the expense of retention or revenue; offers lessons without changed behavior. |
Decision and operating plan
Before interviewing: Run a 15-minute calibration using sample answers and agree on anchors. Give interviewers structured prompts.
After each round: Interviewers submit evidence and scores independently within 24 hours, before seeing others’ assessments. Engineering and Design score independently.
Debrief: Hold a 30-minute evidence review—not a show-of-hands vote. Discuss disagreements and counterevidence against the written bar. Amara makes and documents the hiring decision.
Offer bar: At least 3 on every dimension. In particular, do not average away weak influence, analytical judgment, or execution. Missing evidence triggers one narrowly targeted follow-up, not intuition. Hire against the bar, not merely the strongest available candidate.
Timing and capacity: Finish interviews by day 12, debrief by day 14, and issue the offer by day 21. Calibration, interview, scoring, and debrief consume at most 105 minutes per domain interviewer. Amara retains time within her two-hour cap for decision documentation and an offer conversation. CEO approval, if required, concerns headcount and compensation—not an unstructured personality veto.
Grades and run details
Decision model 92 · LLM judge 13 of 13 checks
Decision model checks
- passUses the supplied evidence correctly27%
- passAddresses the actual decision92%
- partialRespects explicit constraints26%
- partialIdentifies material uncertainty54%
- passAvoids unsupported claims21%
- passProduces the required deliverable91%
- passTests what the last hire failed at100%
- passFixes the take-home's cost to candidates81%
- passFits the people and the time85%
- passJudges on evidence, not presence100%
- passDefines good for this role first89%
- passKeeps each judgement independent100%
Run
- Run
- #1
- API response time
- 55 s
- Submitted
- 5 Oct 2026