Senior Technical Program Manager

Hybrid

Posted 13h ago

Job Location

Hybrid

Tech Stack

Remote Work Policy

On-site

Categories

Applied AI Engineer

About the job

Cloudflare is seeking a highly skilled, self-motivated Technical Program Manager (TPM) to lead strategic, multi-quarter initiatives within our Performance and Infrastructure organization. As a vital link between Product, Engineering, Capacity Planning, Finance, and Reliability, you will drive a unified strategy from initial conception through to successful execution. In this role, you will be deeply involved in gathering requirements, defining deliverables, and reporting on progress. You will manage stakeholder expectations, proactively identify risks and dependencies, and ensure the delivery of predictable, high-impact outcomes. The ideal candidate is process-oriented yet remains flexible and iterative. You excel at navigating ambiguous technical challenges, fostering collaboration across teams, and asking the insightful questions necessary to drive stakeholders toward measurable results. This position requires a professional capable of managing complex tradeoffs and coordinating teams across global time zones.

Responsibilities

  • Lead execution for key initiatives focusing on Infrastructure Efficiency and Engineering Reliability.
  • Partner with Product and Engineering leaders to establish goals, metrics, and decision frameworks.
  • Maintain program plans, roadmaps, dependency maps, and executive status reports.
  • Align Product, Engineering, Infrastructure, Capacity, Data, Finance, Ops, and Security teams.
  • Identify tradeoffs early, document decisions, and manage escalations to resolution.
  • Represent internal technical users and integrate their workflows into program planning.
  • Establish fleet-wide reporting for resource usage to inform capacity planning and cost-efficiency strategies.
  • Build automations and AI agents to handle routine program administration at scale.
  • Coordinate shared dependencies across Infrastructure Engineering, Capacity Planning, Traffic Engineering, and Network Infrastructure workstreams.
  • Lead end-to-end chaos engineering, from test planning and execution to results publication.
  • Coordinate readiness reviews, go/no-go decisions, and safe execution of production resilience tests.
  • Track remediation, evaluate program impact, and maintain roadmaps, dashboards, and runbooks.

Requirements

  • BS+ in Computer Science, Information Technology, Engineering, or a related field, or equivalent experience.
  • 5+ years of technical program management experience, preferably in infrastructure, reliability, performance, data platforms, distributed systems, or large-scale engineering environments.
  • Strong technical fluency with infrastructure systems, telemetry/observability, reliability practices (SLO/SLIs), capacity planning, data pipelines, traffic management.

About Cloudflare

Get new AI jobs in your inbox

A weekly digest of the newest AI engineering roles.

© 2026 AI Job Board. All rights reserved.