ML Platform Engineer

Remote Europe FullTime

Posted 1mo ago

Job Location

Europe

Tech Stack

Remote Work Policy

Fully remote

Employment Type

FullTime

Categories

AI Infrastructure Engineer

About the job

Synthesia is seeking an Engineer to join its ML Platform team. This team is responsible for building and operating the systems that enable researchers and product teams to train, serve, and deploy generative models efficiently and reliably. The role involves working on research infrastructure, production serving systems, internal tooling, and platform interfaces, with a growing focus on making these systems automation-friendly and agent-oriented. This is a hands-on individual contributor role with significant ownership, where you will help shape the evolution of the ML platform as it scales.

Responsibilities

  • Design and improve platform systems for model training, evaluation, and production serving.
  • Build infrastructure and tooling to enhance the reliability, scalability, and cost-efficiency of ML workloads.
  • Develop internal tools and workflows that are easily operated by both humans and agents.
  • Work on the architecture for deploying, serving, and operating models across research and product environments.
  • Improve the scheduling, monitoring, and debugging of workloads on GPU and cloud infrastructure.
  • Develop internal tools, abstractions, and agentic systems to reduce operational overhead for researchers and engineers.
  • Drive improvements in observability, automation, reliability, and developer experience.
  • Collaborate with researchers and product engineers to address pain points and build robust platform capabilities.
  • Contribute to technical direction and make pragmatic architectural tradeoffs for platform growth.

Requirements

  • Strong experience building or operating production systems with a focus on reliability, scalability, and maintainability.
  • A systems mindset, considering bottlenecks, failure modes, interfaces, resource usage, and long-term operability.
  • Solid hands-on experience with cloud infrastructure, Linux, and infrastructure automation.
  • Experience with Kubernetes and operating distributed workloads in production.
  • Strong coding skills, ideally in Python or similar backend/tooling languages.
  • Strong judgment regarding the leverage of automation versus human control and reliability.
  • Experience building internal platforms, developer tooling, or infrastructure abstractions for other engineers.
  • Comfort working in ambiguous environments and taking ownership of open-ended technical problems.
  • A pragmatic approach focused on solving the right problem well.

About synthesia.io

Get new AI jobs in your inbox

A weekly digest of the newest AI engineering roles.

© 2026 AI Job Board. All rights reserved.