Kubernetes Jobs

410 open roles mentioning Kubernetes

Staff Platform Engineer, Voice AI

3mo ago
Together AI

Together AI

Together AI is seeking a Staff Platform Engineer to lead the architecture of their Voice AI platform, which powers real-time voice agents at scale. This role involves setting the technical direction for how developers interact with the platform, from API primitives to autoscaling systems and multi-provider abstractions. The focus is on building robust, low-latency infrastructure for voice applications, which presents unique challenges compared to text inference, such as handling bidirectional audio streams and stateful connections. This is a foundational position on a small team, where decisions will shape the platform's architecture for years to come.

$220k - $280k

San Francisco remote
KubernetesPythonTypeScript +7 more

Solutions Architect (Inference)

3mo ago
Together AI

Together AI

As a Solutions Architect (Inference) at Together AI, you will work with customers and prospects to create business value through Generative AI applications. Solutions Architects at Together are trusted advisors to our customers that evaluate, identify and demonstrate how Together can solve their AI needs. As key contributors to our sales organization, Solution Engineers add tremendous value to the customer journey and directly impact company growth and revenue. This is an exciting opportunity for a deeply technical professional passionate about AI and customer success to make a significant impact in a fast-paced, innovative environment.

London onsite
DockerKubernetesPython +8 more

Solutions Architect

3mo ago
Together AI

Together AI

As a Solutions Architect at Together AI, you will work with customers and prospects to create business value through Generative AI applications. You will act as a trusted advisor, evaluating, identifying, and demonstrating how Together AI can solve their AI needs. This role is a key contributor to the sales organization, adding significant value to the customer journey and directly impacting company growth and revenue. It's an exciting opportunity for a deeply technical professional passionate about AI and customer success to make a significant impact in a fast-paced, innovative environment.

$180k - $260k

San Francisco remote
DockerKubernetesPython +8 more

Senior Software Engineer - Together Cloud Platform

3mo ago
Together AI

Together AI

Together AI is building the AI Acceleration Cloud, an end-to-end platform for the full generative AI model lifecycle, combining a fast LLM inference engine with state-of-the-art AI cloud infrastructure. As a Senior Backend Engineer, you will play a key role in building the next generation AI cloud platform. This platform is designed to be highly available, global, and extremely fast, virtualizing cutting-edge ML hardware and enabling practitioners with self-serve AI cloud services. It serves both internal StaaS products and external cloud customers across numerous data centers worldwide.

$160k - $230k

San Francisco remote
AWSAzureKubernetes +8 more

Senior Software Engineer Together Cloud Infrastructure

3mo ago
Together AI

Together AI

Together AI is building the AI Acceleration Cloud, an end-to-end platform for the full generative AI lifecycle, combining the fastest LLM inference engine with state-of-the-art AI cloud infrastructure. As a Senior AI Infrastructure Engineer, you will play a key role in building the next generation AI cloud platform – a highly available, global, blazing-fast cloud infrastructure that virtualizes cutting-edge ML hardware and enables state-of-the-art ML practitioners with self-serve AI cloud services. This platform serves both our internal SaaS products and our external cloud customers, spanning dozens of data centers across the world.

Amsterdam hybrid
AWSAzureKubernetes +12 more

Senior Software Engineer - Together Cloud Infrastructure

3mo ago
Together AI

Together AI

Together AI is building the AI Acceleration Cloud, an end-to-end platform for the full generative AI lifecycle, combining the fastest LLM inference engine with state-of-the-art AI cloud infrastructure. As a Senior AI Infrastructure Engineer, you will play a key role in building the next generation AI cloud platform – a highly available, global, blazing-fast cloud infrastructure that virtualizes cutting-edge ML hardware and enables state-of-the-art ML practitioners with self-serve AI cloud services. This platform serves both our internal SaaS products and our external cloud customers, spanning dozens of data centers across the world.

$160k - $230k

San Francisco remote
AWSAzureKubernetes +12 more

Senior Platform Engineer, Voice AI

3mo ago
Together AI

Together AI

Together AI is building the best inference infrastructure for voice applications, powering production-grade, real-time voice agents and applications with best-in-class latency and reliability. We are seeking a Senior Platform Engineer to take ownership of the API and infrastructure layer for voice workloads. You will develop the real-time WebSocket and HTTP APIs used by developers to deploy voice experiences, design autoscaling for latency-sensitive streaming workloads, and ensure the reliability of our multi-provider voice platform for production voice agents handling millions of calls. This is a critical, foundational role on a small, high-impact team, defining how developers interact with our voice platform as we scale.

$200k - $260k

San Francisco remote
KubernetesPythonTypeScript +8 more

Senior Backend Engineer, Inference Platform

3mo ago
Together AI

Together AI

Together AI is building the Inference Platform to bring advanced generative AI models to the world, powering multi-tenant serverless workloads and dedicated endpoints. This role offers a unique opportunity to optimize latency and fully utilize tens of thousands of GPUs, working hands-on with cutting-edge hardware. You will collaborate directly with research teams to productionize frontier models and engage with the open-source community, contributing to projects that push the boundaries of inference performance and efficiency.

$160k - $250k

San Francisco remote
KubernetesPythonTypeScript +7 more

Lead/Manager Together Cloud Infrastructure

3mo ago
Together AI

Together AI

Together AI is building the AI Acceleration Cloud, an end-to-end platform for the full generative AI lifecycle, combining the fastest LLM inference engine with state-of-the-art AI cloud infrastructure. As a Lead/Manager, you will play a key role in building the Together cloud platform engineering team in the Netherlands. This platform serves both our internal SaaS products (inference, fine-tuning) and our external cloud customers, spanning dozens of data centers across the world.

Amsterdam hybrid
AWSAzureKubernetes +9 more

Customer Support Engineer (Inference), India

3mo ago
Together AI

Together AI

As a Customer Support Engineer at a pioneering AI company, you will be the first line of defense supporting customers building training, fine-tuning, and inference solutions. You will dive deep into complex technical challenges, providing swift and effective solutions while serving as a product expert. Collaborating closely with product and sales, you will drive continuous improvement of our offerings. This is an exciting opportunity for a deeply technical professional passionate about AI and customer success to make a significant impact in a fast-paced, innovative environment.

India remote
KubernetesPythonTypeScript +9 more

Customer Support Engineer (Inference)

3mo ago
Together AI

Together AI

As a Customer Support Engineer at a pioneering AI company, you will be the primary point of contact for customers building training, fine-tuning, and inference solutions with Together AI. You will tackle complex technical challenges, provide effective solutions, and act as a product expert. Collaborating with product and sales teams, you will contribute to the continuous improvement of our offerings. This role is ideal for a technically proficient individual passionate about AI and customer success, seeking to make a significant impact in a fast-paced, innovative environment.

$160k - $230k

San Francisco, CA remote
KubernetesPythonTypeScript +9 more

AI Infrastructure Engineer

3mo ago
Together AI

Together AI

As an AI Infrastructure Engineer at Together, you will be responsible for ensuring the smooth operation of all user-facing services and production systems. This role blends pragmatic operations with software engineering, applying sound engineering principles, operational discipline, and mature automation to our operating environments and codebase. You will specialize in systems (operating systems, storage subsystems, networking), implementing best practices for availability, reliability, and scalability, with varied interests in algorithms and distributed systems. Together AI is a research-driven artificial intelligence company focused on lowering the cost of modern AI systems through co-designing software, hardware, algorithms, and models, and we invite you to join our passionate group of researchers and engineers in building the next generation of AI infrastructure.

$190k - $270k

San Francisco remote
KubernetesPythonTerraform +2 more

AI infrastructure Engineer (SRE) Amsterdam

3mo ago
Together AI

Together AI

Together AI is seeking an AI Infrastructure Engineer (SRE) to ensure the smooth operation of user-facing services and production systems. This role combines the skills of a pragmatic operator and a software engineer, applying sound engineering principles, operational discipline, and automation to our operating environments and codebase. You will specialize in systems such as operating systems, storage subsystems, and networking, while implementing best practices for availability, reliability, and scalability, with interests in algorithms and distributed systems. Join a research-driven artificial intelligence company focused on advancing AI through open and transparent systems, aiming to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models.

Amsterdam onsite
KubernetesTerraformObservability +7 more

LLM Inference Frameworks and Optimization Engineer

3mo ago
Together AI

Together AI

Together.ai is building state-of-the-art infrastructure for efficient and scalable inference of large language models (LLMs). The company's mission is to optimize inference frameworks, algorithms, and infrastructure to push the boundaries of performance, scalability, and cost-efficiency. They are seeking an Inference Frameworks and Optimization Engineer to design, develop, and optimize distributed inference engines for multimodal and language models at scale. This role will focus on low-latency, high-throughput inference, GPU/accelerator optimizations, and software-hardware co-design, ensuring efficient large-scale deployment of LLMs and vision models. This position offers a unique opportunity to shape the future of LLM inference infrastructure and ensure scalable, high-performance AI deployment across diverse applications.

$160k - $230k

San Francisco, Singapore, Amsterdam remote
KubernetesPythonC# +9 more

Machine Learning, Platform Engineer

3mo ago
Together AI

Together AI

Together AI is a research-driven artificial intelligence company focused on lowering the cost of modern AI systems. This role is part of a team dedicated to enabling custom models and dedicated inference on Together's platform. The team is responsible for building a container platform, optimizing autoscaling, minimizing cold starts, achieving the best end-to-end model performance, and providing a best-in-class developer experience with great tooling. The work often involves video or audio generation across the stack, including CUDA kernels, PyTorch optimization, inference engines, container orchestration, and queueing theory.

$160k - $250k

San Francisco remote
KubernetesPythonRust +7 more

Research Engineer, Frontier Speculative Decoding

3mo ago
Together AI

Together AI

Together AI is building the Inference Platform that powers the world's most advanced generative AI models. This role will serve as a critical bridge between cutting-edge research and real-world applications, focusing on translating internal model training research into production-ready deployments for customers. The work involves a deep commitment to data-centric development, meticulous hyperparameter tuning, and rigorous checkpoint evaluation. You will transform general-purpose models into highly performant, specialized tools by fine-tuning them on customer-specific data and internal datasets, working with dedicated GPU clusters rather than training foundation models from scratch.

$190k - $270k

San Francisco, New York City remote
KubernetesPythonFine-Tuning +6 more

Staff Engineer, Distributed Storage and HPC & AI Infrastructure

3mo ago
Together AI

Together AI

Together AI is seeking a Staff Engineer to design and deliver multi-petabyte storage systems optimized for large-scale AI training and inference workloads. You will architect high-performance parallel filesystems and object stores, integrate cutting-edge technologies, and drive significant cost optimization. The role involves building Kubernetes-native storage operators and self-service platforms for automated provisioning and multi-tenancy. You will focus on optimizing data paths, designing multi-tier caching architectures, and tuning parallel filesystems for AI applications. This is a research-driven role within a company focused on lowering the cost of modern AI systems through co-design of software, hardware, algorithms, and models.

$250k - $300k

San Francisco remote
KubernetesPythonGo +7 more

Software Engineer, Agents & Automations

3mo ago
Cohere

Cohere

Cohere is seeking a Software Engineer to join the Agents & Automations team, which builds the core platform for North, Cohere's AI workspace. This platform enables customers to create AI-powered workflows, ranging from structured automations to flexible agents that can interact with enterprise systems. The role involves building the workflow builder, execution engine, integrations, debugging tools, and feedback loops that empower users to deploy and manage AI agents and automations. This is a broad product engineering role that spans frontend, backend, and AI-powered systems, with the goal of shipping reliable software that helps customers augment or automate business workflows.

London remote FullTime
CohereKubernetesPython +3 more

Member of Technical Staff (Software Engineer, Connector Platform)

3mo ago
P

Perplexity AI

The Connector Platform team is responsible for building the data layer that enables Perplexity's agents to interact with the world's software. This team manages systems that transform hundreds of diverse integrations into a unified, reliable, and well-typed interface for agents. This platform serves as the core knowledge layer, enabling agents to discover, understand, and utilize tools effectively, grounding their reasoning in real, permissioned, and up-to-date enterprise data. By maintaining a knowledge layer above connectors, the platform ensures Computer becomes the single source of truth for institutional knowledge, providing grounded, actionable, and permissioned access to customer systems, which is a key differentiator in the market.

San Francisco hybrid FullTime
AWSKubernetesPython +2 more

Senior/Staff Software Engineer, Developer Experience

3mo ago
Abridge

Abridge

Abridge is seeking Engineers to join a new, high-priority agentic engineering team focused on building the foundational infrastructure and internal systems that power software development at the company. This role operates at the intersection of AI tooling, CI/CD, and developer experience, tackling complex problems related to CI/CD pipelines, MCP servers, and AI agents. The work will directly influence how engineers build and deploy software across Abridge, shaping the future of developer productivity and operational efficiency within a rapidly evolving landscape.

SF Office hybrid FullTime
LangGraphLangChainAWS +3 more

Get new AI jobs in your inbox

A weekly digest of the newest AI engineering roles.

© 2026 AI Job Board. All rights reserved.