Metrics question

An AI agent for patient access and scheduling has high adoption, but resolution rate is below target because too many conversations escalate to human staff. How would you break down the funnel, determine whether the root cause is workflow design, knowledge gaps, policy boundaries, or user behavior, and prioritize the first three fixes? What leading and lagging metrics would you use to know the agent is becoming core infrastructure?

Practice this question out loud. An AI interviewer asks it, follows up like a real interviewer would, and scores your answer. Type or speak.

Start a mock interview on this question · Mock interview from a job description

What this question tests

Whether you can decompose a resolution-rate shortfall into distinct root causes, prioritize fixes with real reasoning, and define metrics that track the agent becoming trusted infrastructure, not just handling volume.

How to approach it

  1. Break the funnel by escalation reason: workflow design gaps (the agent cannot complete a valid task type at all), knowledge gaps (the agent lacks information to answer correctly), policy boundaries (the agent correctly defers because a human must handle it), and user behavior (ambiguous or out-of-scope requests).
  2. Separate true failures from appropriate escalations first, since policy-boundary escalations are working as intended and should not be counted the same as workflow or knowledge failures when prioritizing fixes.
  3. Prioritize the first three fixes by volume times fixability: a high-volume knowledge gap that is cheap to close (updating a knowledge base) likely outranks a lower-volume but harder workflow redesign.
  4. Fix knowledge gaps first if they dominate volume, since they are typically the fastest to close, then workflow design gaps, then revisit policy boundaries only if they are miscalibrated, escalating things that could safely be automated.
  5. Track leading indicators like escalation reason distribution shifting away from knowledge and workflow gaps over time, and lagging indicators like sustained resolution rate improvement plus patient or staff trust signals, such as repeat usage without complaint, to know the agent is becoming core infrastructure rather than a novelty.

What a strong answer includes

Common mistakes

Likely follow-up questions

More metrics questions

More questions from Decagon

Learn the skill behind it

Chapters of the AI PM course that teach what this question tests.

Preparing for a specific role?

Book summaries for this kind of question

Browse all 4,000+ questions in the bank