Staff Software Engineer, Core Infrastructure

Remote Bengaluru FullTime

Posted 4mo ago

Job Location

Bengaluru

Tech Stack

Remote Work Policy

Fully remote

Employment Type

FullTime

Categories

AI Infrastructure Engineer

About the job

Harvey is transforming legal and professional services by combining frontier agentic AI, an enterprise-grade platform, and deep domain expertise. This is a rare opportunity to help build a generational company at an inflection point, with strong product-market fit and significant investor backing. The team moves fast, takes ownership, and is deeply committed to the mission, operating with intensity and pushing for excellence. As a Staff Software Engineer on the Core Infrastructure team, you will play a critical role in designing, building, scaling, and strengthening the infrastructure that powers every user interaction with Harvey's AI platform. You will work in an environment balanced between innovation and operational excellence, ensuring the platform remains resilient and efficient as it scales.

Responsibilities

  • Design and build scalable, fault-tolerant infrastructure systems for Harvey's AI platform across multiple cloud regions.
  • Own and evolve multi-cloud infrastructure (Azure, GCP), including Kubernetes, networking, and container management.
  • Lead technical initiatives for observability, incident response, and operational excellence.
  • Architect and optimize distributed systems for reliability, including load balancing, quota management, and failover.
  • Partner with Product Engineering and Security teams to ensure infrastructure accelerates product development.
  • Drive infrastructure-as-code practices using tools like Terraform and Pulumi.
  • Mentor engineers and raise the technical bar through code and design reviews.
  • Design and implement a next-generation model proxy architecture.
  • Build distributed rate limiting and quota management systems.
  • Architect multi-region deployment strategies meeting data residency requirements.
  • Develop comprehensive observability infrastructure with SLA monitoring and cost tracking.
  • Lead the evolution of CI/CD pipelines to improve developer velocity and production stability.

Requirements

  • 10+ years of experience in Infrastructure Engineering or Platform Engineering in a production environment.
  • Long track record building and scaling complex, large-scale distributed systems.
  • Deep proficiency with cloud infrastructure platforms (Azure preferred; GCP or AWS experience is transferable).
  • Strong fluency with Infrastructure as Code (IaC) tools like Terraform, Pulumi, or CloudFormation.
  • Solid understanding of Kubernetes, container orchestration, networking, and cloud security at scale.
  • Experience with observability tools (e.g., Datadog, Sentry) and incident response practices (e.g., PagerDuty, Incident.io).
  • Strong programming skills in Python, Go, or similar languages.
  • Excellent problem-solving skills and a commitment to operational excellence.
  • Experience building infrastructure for AI/ML workloads or high-throughput inference systems (nice to have).
  • Background with distributed rate limiting, load balancing, or quota management systems (nice to have).
  • Experience operating multi-tenant platforms with strict security and compliance requirements (nice to have).
  • Track record of leading complex cross-functional projects and delivering measurable impact (nice to have).

Get new AI jobs in your inbox

A weekly digest of the newest AI engineering roles.

© 2026 AI Job Board. All rights reserved.