Senior Software Engineer, Site Reliability Engineer

$200k - $260k San Francisco FullTime

Posted 8mo ago

Job Location

San Francisco

Tech Stack

Remote Work Policy

On-site

Employment Type

FullTime

Categories

Applied AI Engineer

About the job

Harvey is transforming legal and professional services by combining frontier agentic AI, an enterprise-grade platform, and deep domain expertise. We are a fast-scaling company with strong product-market fit, seeking ambitious individuals to help build a generational company. Our team operates with intensity, ownership, and a commitment to our mission, valuing decisiveness, simplicity, and continuous improvement. As a Software Engineer on the Site Reliability team, you will ensure the reliability, scalability, and performance of our legal AI platform, owning the systems that keep our platform fast, secure, and always on. Your work will be crucial in maintaining platform resilience as we grow across 50+ regions.

Responsibilities

  • Design, implement, and manage monitoring, alerting, and infrastructure resources across 50+ global regions.
  • Lead incident management processes, including postmortems and root cause analyses.
  • Automate operational tasks and workflows to maintain high reliability and reduce manual intervention.
  • Collaborate across teams to drive reliability, security, and compliance throughout the software lifecycle.
  • Optimize infrastructure costs through strategic capacity planning and build-versus-buy decisions.

Requirements

  • 5+ years of experience in Site Reliability Engineering or similar roles supporting production environments.
  • Expertise in infrastructure as code (IaC) tools (Pulumi, Terraform, CloudFormation, etc.).
  • Deep familiarity with observability tools (Datadog, Sentry, etc.) and incident response practices (PagerDuty, IncidentIO, etc.).
  • Proficiency with cloud infrastructure platforms (Azure, GCP, AWS, etc.).
  • Strong programming skills (Python, Bash, Go, or similar languages).
  • Proven track record of diagnosing complex system problems and implementing durable solutions.
  • Solid understanding of CI/CD, Kubernetes, containerization, networking, databases, and cloud security principles.
  • Excellent problem-solving skills and meticulous attention to detail.

Get new AI jobs in your inbox

A weekly digest of the newest AI engineering roles.

© 2026 AI Job Board. All rights reserved.