Metrics question
A new Harvey integration is enabled by enterprise admins, but weekly usage among lawyers is low and answers are rarely reused in active matters. How would you diagnose where the breakdown is, from setup and permissions to relevance, workflow fit, and trust, and what product changes would you test first?
- Harvey
- Metrics
- Hard
Practice this question out loud. An AI interviewer asks it, follows up like a real interviewer would, and scores your answer. Type or speak.
Start a mock interview on this question · Mock interview from a job description
What this question tests
Whether you can diagnose a low-adoption enterprise product problem by isolating the specific stage where it breaks down, from setup through trust.
How to approach it
- Instrument the funnel: admin enablement, individual lawyer awareness, first use, relevance of results, and reuse in active matters, so you can see where usage drops.
- Check setup and permissions first, since admin-level enablement does not guarantee individual lawyers have working access or even know it exists.
- If access is fine, check relevance: sample actual queries and results to see if the integration surfaces genuinely useful matter data or noisy results.
- If relevance is fine, check workflow fit: whether lawyers can act on results within their existing tools, or need an extra, disruptive step.
- If workflow fit is fine, check trust: whether lawyers who tried it once did not return, pointing to a first-impression confidence problem rather than a structural one.
- Test the top hypothesis with a small, targeted change, for example a setup nudge or a relevance fix, and measure whether weekly usage moves before rolling out broader changes.
What a strong answer includes
- Structures the diagnosis as a funnel with a clear stage-by-stage elimination process, rather than guessing at one cause.
- Distinguishes setup and permissions issues, invisible to the end user, from relevance and trust issues, visible but still low-adoption.
- Names reuse in active matters, not just first use, as the real bar for legal professionals.
- Proposes testing the top hypothesis with a small change before a broad rollout, a disciplined, low-risk approach.
Common mistakes
- Jumping to a fix, like more training, without isolating which funnel stage is actually broken.
- Treating low weekly usage and low reuse as the same problem, when they may have different causes.
- No plan to validate a hypothesis with a small test before a larger rollout.
Likely follow-up questions
- How would you get visibility into individual lawyer usage without violating firm confidentiality expectations?
- What would you do if the drop-off is concentrated among one practice group?
- How would you decide the diagnosis is right before committing engineering time to a fix?
More metrics questions
- What metrics prove Harvey's value to a firm like PwC or Paul Weiss?Harvey · Metrics · Hard
- What KPI hierarchy would you use for Vault, from account-level adoption and active matters to search success, document coverage, and workflow outcomes, to measure value for law firms and enterprises? How would those metrics change your roadmap if usage is high but repeat usage in critical workflows is low?Harvey · Metrics · Hard
- What metrics would you define for Command Center across adoption, admin efficiency, and governance coverage, and how would you use those metrics to decide whether the product is actually improving Harvey’s enterprise land-and-expand motion?Harvey · Metrics · Hard
- Harvey is considering expanding its platform to additional regions. How would you decide whether multi-region expansion is the right product investment now versus later? Explain the customer signals, business considerations, and success metrics you would use, and how you would compare this investment against competing platform priorities.Harvey · Metrics · Hard
- For Vault’s search and Q&A experience, what metrics would you define for adoption, engagement, trust, and business value at the workspace and firm level? Which leading indicators would you watch in the first 90 days, and how would changes in those metrics alter your product decisions?Harvey · Metrics · Hard
- A pilot agent deployment has strong qualitative feedback from senior stakeholders, but end-user adoption is inconsistent and the customer’s data environment is messy. How would you structure the deployment plan, define the success metrics, diagnose whether the issue is workflow fit vs data readiness vs change management, and decide if this should graduate into a repeatable product for the broader vertical?Harvey · Metrics · Hard
More questions from Harvey
Learn the skill behind it
Chapters of the AI PM course that teach what this question tests.
- Chapter 9: Prove it paid off: outcomes, economics, and pricing
- Chapter 2: Data fluency: SQL, logs, and reading the truth yourself
- Chapter 14: Get the job: the AI PM interview loop