--- title: 'Senior Product Operations Manager, Evaluation Quality at Harvey' canonical: 'https://feeny.ai/job/senior-product-operations-manager-evaluation-quality-harvey-san-francisco-gaq4t0yda3yt' type: 'job' last_seen: '2026-09-07' --- # Senior Product Operations Manager, Evaluation Quality at Harvey - **Company:** [Harvey](https://feeny.ai/companies/harvey) - **Location:** San Francisco, CA - **Employment:** full-time - **Work type:** hybrid - **Posted:** 2026-08-04 - **Last confirmed live:** 2026-09-07 - **Apply:** https://jobs.ashbyhq.com/harvey/f0e4e24a-2a7a-4fc4-8872-9a8a762fecf5 ## Job description ## WHY HARVEY At Harvey, we’re transforming how legal and professional services operate. By combining frontier agentic AI, an enterprise-grade platform, and deep domain expertise, we’re reshaping how critical knowledge work gets done for decades to come. This is a rare chance to help build a generational company at a true inflection point. We have strong product-market fit and world-class investor support. We’re scaling fast and defining a new category in real time. The work is ambitious, the bar is high, and the opportunity for growth — personal, professional, and financial — is unmatched. Our team moves fast, takes ownership, and is deeply committed to the mission — operating with intensity, staying close to our customers, and pushing each other for excellence. We live by three values: Decisiveness, Simplicity, and Job's Not Finished. We act quickly on clear judgment over perfect information, we believe simplicity is what scales, and we're never satisfied with where we are. If you want to do the best work of your career alongside people who share that drive, we'd love to build with you. At Harvey, the future of professional services is being written today — and we’re just getting started. ## ROLE OVERVIEW We’re looking for a senior operator to own the quality bar behind Harvey’s human evaluations. As we scale globally, the volume of eval work is growing 10x, but volume only matters if the output is trusted. This role makes Human Data’s signal decision-grade: rigorous, calibrated, and reproducible enough that Product, Engineering, and AI Research act on it to ship. As a member of our Evaluation Operations team, you’ll work alongside our Evaluation Operations Manager (who runs throughput and coordination) and partner closely with Applied Legal Researchers, Product, Engineering, and AI Research. You'll set the standard for what "good" looks like across eval data and methodology, and own the data analyses to make conclusions that teams will rely on, building the stakeholder trust that lets EPD act on the signal. ## WHAT YOU'LL DO - Own the quality bar for Harvey’s human evaluations: define what “good” looks like for eval methodology and data analysis, and produce decision-grade outputs for EPD - Author and maintain the evaluation guidelines, instructions, and databases that contract attorneys work from - Standardize and streamline rubric and evaluation design into repeatable templates and one documented methodology, partnering with Applied Legal Research (ALR), who supplies feature-specific legal depth - Own contract-attorney quality: onboarding, calibration training, inter-rater reliability, and the feedback loop (including benchmarks and gold references) that keeps judgment consistent across attorneys and over time - Conducting quantitative and qualitative statistical analyses, diagnosing error states, investigating root causes, and turning raw eval results into a structured, prioritized signal Product and ALR can act on - Run QA on vendor and contract-attorney deliverables against a defined bar before results inform a launch decision - Ensure the quality bar holds across jurisdictions and non-English geographies as coverage expands - Establish one standard, documented way to analyze eval results, and build lightweight operational dashboards to track rater capability and eval-program health - Support ALR in a review step that certifies an evaluation is sound before it scales to contract attorneys - Partner with ALR and Analytics to determine where human eval aligns with online signal and where it can provide expanded insights ## WHAT YOU HAVE - 6+ years in product operations, research operations, evaluation/QA operations, or quality program management - A track record of owning quality inputs (guidelines, instructions, benchmarks, QA procedures) for complex, expert-driven or human-in-the-loop work - Experience onboarding, training, and calibrating a distributed pool of expert raters, annotators, or reviewers, and running the feedback loop that improves their quality over time - Enough grounding in measurement concepts (calibration, inter-rater reliability, sampling, rubric design) to independently set up and own the quality of our evaluation loop yourself - Experience with running quantitative and qualitative data analyses, interpreting and running statistical tests on evaluation data (natively or with AI tool support), and communicating conclusions to various stakeholders - A record of scaling and streamlining evaluation quality processes under shipping pressure, with a bias toward documentation and reproducibility over one-off analysis - Ability to work deeply with domain experts (e.g., ALR / lawyers) and translate nuanced judgment into repeatable, documented standards - Strong cross-functional coordination across Product, Engineering, Research, ALR, and data providers/vendors - Clear communicator who can build credibility and trust with stakeholders - Bias to action and high ownership, from writing the guideline to auditing a vendor batch line by line ## BONUS POINTS - Experience in legal tech or working with domain experts in regulated industries - Experience owning quality across multiple markets, languages, or jurisdictions - Built calibration, inter-rater reliability, or capability-tracking systems for annotation or evaluation pipelines - Experience transitioning evaluation work in-house or otherwise improving evaluation ROI - Familiarity with LLM-as-judge / automated evaluation used alongside human eval - Early employee at a hyper-growth startup, or experience at a world-class product or platform operations org ## COMPENSATION $155,400 - $233,200 USD DEPENDING ON YOUR LOCATION, AN APPLICANT PRIVACY NOTICE MAY APPLY TO YOU. YOU CAN FIND ALL OF OUR APPLICANT PRIVACY NOTICES HERE https://harveyai.notion.site/Harvey-Candidate-Privacy-Notices-319ac3fcdd7a803bb807d5094f249922?pvs=74. ## #LI-CD1 Harvey is an equal opportunity employer and does not discriminate on the basis of race, gender, sexual orientation, gender identity/expression, national origin, disability, age, genetic information, veteran status, marital status, pregnancy or related condition, or any other basis protected by law. We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made by emailing accommodations@harvey.ai ## About Harvey ## Company Overview - **One-liner**: Harvey builds domain-specific AI software for legal and professional services, enabling law firms and in-house teams to automate research, contract analysis, due diligence, and complex workflows. - **Entity Type**: Private (backed by leading venture capital firms) - **Headquarters**: San Francisco, California, United States - **Founded**: 2022 - **Founders**: Winston Weinberg (CEO), Gabriel Pereyra (President) ## Core Business - **Primary industry**: AI software for legal and professional services (LegalTech) - **Target customers**: B2B – large law firms (AmLaw 100), in-house legal teams at Fortune 500 enterprises, mid-sized firms, and professional service networks across 60+ countries. - **Mission or purpose statement**: To enable legal and professional services teams to focus on high-value work by providing a secure, expert AI platform that handles research, drafting, and complex workflows. ## Products & Services - **Harvey Assistant**: Ask questions, analyze documents, and draft faster with domain-specific AI (SaaS). - **Harvey Vault**: Securely store, organize, and bulk-analyze legal documents (SaaS). - **Harvey Knowledge**: Research complex legal, regulatory, and tax questions across domains (SaaS). - **Harvey Agents**: Purpose-built agents that execute complex legal work end-to-end (SaaS). - **Harvey Contract Intelligence**: Surface insights, strengthen negotiations, and accelerate reviews (SaaS). - **Harvey Command Center**: Analytics, benchmarking, and agentic insights for leading AI transformation (SaaS). - **Harvey Mobile**: Mobile access to stay productive from anywhere (SaaS). - **Harvey Shared Spaces**: Collaborate with legal teams across organizations in secure shared workspaces (SaaS). - **Harvey Ecosystem**: Integration with existing tools to ground answers in trusted sources (API/platform). ## Market Standing - **Valuation/Market Cap**: Not publicly disclosed (company is private). - **Key Metric**: Annual Recurring Revenue (ARR) – reported in press but exact figure not disclosed; strong growth implied by 1,500+ customers and rapid hiring. - **Notable Investors/Partners**: Sequoia, Kleiner Perkins, GV (Google Ventures), OpenAI Startup Fund, Coatue, Andreessen Horowitz, EQT. - **Growth Signals**: - Named Time100 Most Influential Companies (2025). - LinkedIn Top 50 Startup of 2025. - CNBC Disruptor 50 List. - 1,500+ customers across 60+ countries, including AmLaw 100 firms. - Headcount: ~300 employees (as of mid-2025), with aggressive hiring (254 active job postings). - Office expansion: San Francisco, New York City, London. ## Competitive Advantages - **Domain specialization**: AI models trained specifically for legal and professional services, not generic LLMs. - **Enterprise security**: SOC 2 Type II, ISO 27001/27701/42001, GDPR, CCPA compliance – meeting law firm security requirements. - **Deep investor backing**: World-class VC network provides strategic guidance and access. - **Strong product‑market fit**: Rapid adoption by top-tier law firms and in-house teams, with high switching costs due to embedded workflows. - **First‑mover in agentic legal AI**: Purpose‑built agents execute end‑to‑end tasks, differentiating from simple document Q&A tools. ## Strategic Focus - **Customization, collaboration, and simplicity**: Helping teams embed expertise into Harvey, collaborate across firms, and reduce complexity via agentic workflows. - **Expansion beyond legal**: Long‑term ambition to transform professional services broadly (consulting, accounting, etc.). - **Product innovation**: Continued investment in agentic AI, mobile access, and ecosystem integrations. - **Global growth**: Scaling sales and engineering teams in existing hubs and expanding into new markets. ## Why Work Here - **Culture**: Fast‑paced, collaborative, ambitious. Values include **Decisiveness** (“take the square root of the weather” – avoid overanalysis), **Simplicity**, and **Job’s Not Finished** (continuous improvement). - **Work environment**: Hybrid – 3+ days per week in office for roles based in San Francisco, New York City, or London. Some fully remote positions available (clearly marked). - **Benefits**: Comprehensive health/dental/vision insurance, 401(k) match, paid parental leave (immediately eligible), annual professional development stipend, in‑office daily lunch, and more. - **Growth opportunities**: Rapidly scaling (~300 employees, 254 open roles) provides internal mobility and ownership of impactful projects. - **Engineering culture**: Strong blend of AI research (DeepMind background) and legal domain expertise. CTO Siva Gurumurthy recently joined. Emphasis on building secure, enterprise‑grade products. ## Sources 1. [harvey.ai](https://www.harvey.ai/) 2. [harvey.ai/company](https://www.harvey.ai/company) 3. [harvey.ai/blog/landing-a-job-at-harvey](https://www.harvey.ai/blog/landing-a-job-at-harvey) 4. [linkedin.com/company/harvey-ai](https://www.linkedin.com/company/harvey-ai) 5. [jobs.ashbyhq.com/harvey](https://jobs.ashbyhq.com/harvey) ## Other roles at Harvey - [Staff Software Engineer, Model Infrastructure](https://feeny.ai/job/staff-software-engineer-model-infrastructure-harvey-san-francisco-8a16y83nftsj) — San Francisco, CA - [Senior Software Engineer, Model Infrastructure](https://feeny.ai/job/senior-software-engineer-model-infrastructure-harvey-san-francisco-6tywmebr0r9h) — San Francisco, CA - [Head of Mid-Market Sales](https://feeny.ai/job/head-of-mid-market-sales-harvey-london-n918xg1r4jgb) — London, United Kingdom - [Customer Experience Manager, US](https://feeny.ai/job/customer-experience-manager-us-harvey-san-francisco-2dw5w8905rz6) — San Francisco, CA - [Transformation Associate](https://feeny.ai/job/transformation-associate-harvey-san-francisco-qm17099y4cjc) — San Francisco, CA - [Transformation Associate](https://feeny.ai/job/transformation-associate-harvey-chicago-h0b3a2pdz2zp) — Chicago, IL - [Transformation Associate](https://feeny.ai/job/transformation-associate-harvey-dallas-a2hzwnh0vb51) — Dallas, TX - [Transformation Associate](https://feeny.ai/job/transformation-associate-harvey-remote-cezdb49eyhv5) - [Transformation Associate](https://feeny.ai/job/transformation-associate-harvey-new-york-78kzq8sv2gvp) — New York, NY - [Content Marketing Manager](https://feeny.ai/job/content-marketing-manager-harvey-san-francisco-zsjew4mcankn) — San Francisco, CA