AI & Technical question

How would you measure the accuracy of Harvey's legal research outputs?

Practice this question out loud. An AI interviewer asks it, follows up like a real interviewer would, and scores your answer. Type or speak.

Start a mock interview on this question · Mock interview from a job description

What this question tests

Tests AI evaluation design specifically for legal research accuracy, where correctness has a precise, checkable standard.

How to approach it

  1. Define accuracy precisely for this domain: did the research correctly identify the controlling law, and does every citation actually support the stated proposition.
  2. Build a gold standard test set: real research questions with attorney verified correct answers and citations, ideally across multiple practice areas.
  3. Score on multiple dimensions: citation existence, citation relevance to the claim, and completeness, since research can be technically correct but miss a key case.
  4. Bring in domain experts for evaluation, since only practicing attorneys can reliably judge whether a legal proposition is actually correctly supported.
  5. Track accuracy over time and by practice area, since performance on tax law may differ meaningfully from performance on employment law.
  6. Confirm with the interviewer whether the evaluation is for internal quality tracking or for a claim made externally to law firms, since the rigor bar differs.

What a strong answer includes

Common mistakes

Likely follow-up questions

More ai & technical questions

More questions from Harvey

Learn the skill behind it

Chapters of the AI PM course that teach what this question tests.

Preparing for a specific role?

Book summaries for this kind of question

Browse all 4,000+ questions in the bank