Technical Program Manager, Developer Platform

Remote San Francisco FullTime

Posted 5mo ago

Job Location

San Francisco

Remote Work Policy

Fully remote

Employment Type

FullTime

Categories

AI Infrastructure Engineer

About the job

Harvey is seeking a Technical Program Manager, Quality and Reliability to lead initiatives that enhance product quality, reliability, and operational excellence. This role is pivotal in owning Harvey's product quality by collaborating with Engineering, Product, and Operations teams. You will identify and address gaps in test coverage, release safety, incident management, and system observability, while also discovering reliability risks and driving scalable improvements to support hyper-growth.

Responsibilities

  • Own end-to-end release management, ensuring timely and high-quality product releases.
  • Introduce and enforce change safety standards to minimize regressions and customer impact.
  • Lead horizontal reliability initiatives to improve test coverage, observability, and incident response.
  • Define, measure, and report on reliability metrics, driving accountability for improvement.
  • Identify systemic gaps in release processes, testing, monitoring, and incident response, creating structured improvement plans.
  • Drive rapid triage and resolution of customer-reported issues, ensuring timely follow-up.
  • Own and improve the incident management lifecycle, including post-incident reviews and tracking corrective actions.
  • Oversee vendor reliability and SLA compliance.

Requirements

  • 5+ years of experience in technical program management or release management, preferably in SaaS or fast-moving tech companies.
  • Prior experience as a Software QA or Test Engineer, preferably with SaaS products.
  • Strong understanding of engineering workflows, including CI/CD, release cycles, and infrastructure planning.
  • Experience partnering with engineering and product leadership on quality and reliability objectives.
  • Excellent communication skills, able to simplify complex information for technical and non-technical audiences.
  • Proven track record of building scalable systems and processes.
  • Comfort with ambiguity and building structure in undefined areas.
  • Bachelor's degree in Computer Science, Engineering, or a related technical field.
  • Familiarity with incident management tooling (e.g., PagerDuty, Incident.io).
  • Familiarity with monitoring stacks (e.g., Datadog, Prometheus, Grafana).
  • Familiarity with test automation frameworks (e.g., Playwright, Cypress, Selenium).

Get new AI jobs in your inbox

A weekly digest of the newest AI engineering roles.

© 2026 AI Job Board. All rights reserved.