Kubernetes Jobs
410 open roles mentioning Kubernetes
Senior Software Engineer, Platform Security
decagon
Decagon is seeking a founding Platform Engineer, Security to design, build, and operate the security infrastructure for its leading conversational AI platform. This is a hands-on builder role focused on creating durable, well-engineered systems, including paved paths for secure service creation, automated security tooling, and infrastructure-as-code. As the first dedicated engineer in this area, you will have a significant impact on shaping the technical direction of the infrastructure security program and investing in automation, observability, and self-service. The work is closely tied to the business, directly enabling high-value deals with security-conscious enterprises.
$200k - $400k
Deployed Engineer (Seattle)
Langchain
LangChain is seeking a Deployed Engineer to join their team, focusing on making intelligent agents ubiquitous. This role involves working on challenging applied AI problems, building production systems that real teams depend on, and directly shaping how AI agents are developed and deployed in the real world. The feedback loop is fast, the impact is visible, and the work contributes to the evolution of AI agent technology.
$165k - $380k
Deployed Engineer (Boston)
Langchain
LangChain is seeking a Deployed Engineer to join their team, focusing on making intelligent agents ubiquitous. This role involves working directly with companies to build and run AI agents in production, transforming prototypes into reliable systems. You will tackle complex applied AI challenges, contributing to systems that real teams depend on, with a fast feedback loop and visible impact. Your work will directly shape how AI agents are built and adopted in the real world, bridging the gap between engineering, product, and go-to-market strategies.
$165k - $380k
Deployed Engineer (NYC)
Langchain
LangChain is seeking a Deployed Engineer to join their team in NYC. This role focuses on building and running AI agents in production, moving beyond prototypes to create reliable systems that real teams depend on. You will work on challenging applied AI problems, with a fast feedback loop and visible impact, directly shaping the future of AI agents in the real world. The Deployed Engineering team partners closely with customer engineers throughout the entire lifecycle, from pre-sales to post-deployment, ensuring technical success and reliable operation of agents at scale using the LangChain suite.
$165k - $380k
Technical Program Manager, Platform
Scale AI
As a Technical Program Manager for the Platform team, you will partner with engineering teams to directly accelerate the development and maturity of the Scale Generative AI Platform (SGP). We are looking for a TPM who has actively built and shipped products in the past and understands how to deliver robust, scalable developer tooling and distributed systems. In this role, you will own the strategic alignment and end-to-end execution of our most critical infrastructure initiatives—from initial scoping to measurable, company-wide and customer-ready adoption. You will serve as the core communication backbone and connective tissue between platform engineering, product teams, and executive leadership. Operating in a hyper-growth, demanding AI environment, you will translate SGP’s architectural complexities into clear execution strategies, unblock engineering bottlenecks, proactively mitigate deployment risks, and ensure our foundational platforms deliver reliable, performant, and secure systems capable of global-scale deployment.
$211k - $264k
Software engineer, generative AI (UK)
Writer
WRITER is a leading enterprise AI company focused on expanding human capacity through superintelligence. Our platform enables businesses to build and deploy AI agents grounded in their data, powered by enterprise-grade LLMs. We are seeking a Software Engineer, Generative AI to build the secure, scalable foundation for our AI solutions in complex corporate environments. This role is for a well-rounded engineering generalist with a strong focus on generative AI and a whole-systems mindset for architectural design. You will own projects from proposal to deployment, shaping the future of AI and contributing to a product that is changing how the world works.
Software engineer, generative AI
Writer
WRITER is seeking a Software Engineer, Generative AI to join their team. This role is at the forefront of expanding human capacity by building the secure, scalable foundation that allows generative AI solutions to thrive in complex corporate environments. The ideal candidate is a well-rounded engineering generalist who leans heavily into generative AI and brings a whole-systems mindset to architectural design. You will own projects from proposal to deployment, shaping the future of AI and contributing to a product that's changing how the world works. This is a hybrid role based out of San Francisco, New York City, or London.
Member of Technical Staff (Software Engineer, Cloud Infrastructure)
Perplexity AI
The Cloud Infrastructure team is responsible for the foundational cloud primitives and deployment models that power Perplexity's products. This includes multi-tenant public cloud solutions as well as single-tenant and on-premises options for enterprise clients. As Perplexity expands its Computer and Enterprise products, this team will build and manage the security, isolation, and compliance layers essential for customer trust. They provide the deployment topologies, multi-region infrastructure, and core services that ensure Perplexity operates reliably and efficiently for both consumer traffic and large enterprises.
DevOps Engineer - AI Systems
Scale AI
We are seeking a skilled DevOps Engineer to manage our cloud infrastructure, specifically for AI training and inference workloads. This role will involve working with major cloud providers and utilizing containerization and infrastructure-as-code tools to ensure efficient and scalable operations.
$145k - $195k
Technical Program Manager, Platform & Infrastructure
Harvey
Harvey is transforming legal and professional services by integrating frontier agentic AI with an enterprise-grade platform. We are seeking a Staff Technical Program Manager, Platform & Infrastructure to lead critical scaling initiatives. This role involves acting as the central point of contact across Core Infrastructure, Backend Platform, and product engineering teams, as well as cross-functional departments like Security and Finance. You will be responsible for multi-quarter horizontal programs focused on cloud migrations, cost optimization, capacity planning, BYOC, and building foundational internal platform elements. The position requires deep technical engagement with senior engineers, aligning leadership on long-term plans, and driving programs from initial concept to successful delivery.
$188k - $278k
Software Engineer - Baseten Inference Stack
Baseten
Baseten is seeking a Software Engineer to join their Inference Stack team. This team builds the distributed runtime that powers large-scale LLM inference across Baseten's platform, operating at the intersection of distributed systems, model performance, infrastructure, and developer experience. The role involves working across the entire stack, from customer-facing deployment tools and feature libraries to the underlying systems for orchestrating Kubernetes deployments and routing traffic. This is an ideal opportunity for engineers who thrive on owning production systems, solving complex integration challenges, and simplifying intricate infrastructure for users.
Senior Production Engineer, Managed AI
crusoe
Crusoe is seeking a Senior Production Engineer to join their team and ensure the reliability and scalability of their AI-optimized cloud platform. This role is crucial for building and operating managed AI services at scale, focusing on delivering highly available, performant, and cost-efficient AI infrastructure for compute-intensive, latency-sensitive workloads. The ideal candidate will have a strong background in distributed systems and hands-on experience with large language models.
$170k - $205k
Backend Engineer - Studio Media Platform
sarvam
Sarvam is building the bedrock of Sovereign AI for India, developing a full-stack AI platform across research, models, infrastructure, and applications. We partner with leading enterprises and public institutions, backed by prominent venture capital firms. We are hiring a Backend Engineer to work on our Studio media platform, which includes AI dubbing, live translation, and foundational services for voice cloning, stem separation, lip sync, and music generation. You will build and maintain production services, ML pipeline libraries, and platform SDKs to enable multilingual media processing at scale for our enterprise customers and Studio users.
Staff Infrastructure Engineer
Replit
Replit is seeking a Staff Infrastructure Engineer to join their Infrastructure Engineering team. This role focuses on ensuring the reliability, scalability, and performance of Replit's platform, which serves millions of developers globally. The engineer will bridge development and operations, implementing automation and best practices to enable efficient scaling and high availability. The position involves proactively identifying and resolving reliability issues, designing robust monitoring solutions, automating operational tasks, and mentoring the engineering team on reliability principles.
Applied AI Engineer, Site Reliability Engineer - EMEA
Mistral AI
Mistral AI is seeking a founding engineer for its Applied AI Site Reliability Engineering (SRE) sub-team. This role is crucial for building and operating a framework that ensures the reliability and sustainability of Mistral's AI solutions across all customer accounts, whether hosted by Mistral or the customer. You will operate in four key modes: BUILD (designing for a fleet of platforms, proactive reliability, authoring runbooks, implementing observability), RUN (operating Tier-1 customer environments, ensuring SLO compliance, managing incidents), ENABLE (productizing deployment, security, and scaling of Applied AI solutions), and SECURE (owning security operations, leading CVE response, and implementing supply-chain integrity controls). This is a framework-first, fleet management role focused on structurally solving problems for all customers, not just individual ones. The team values people and outputs, direct feedback, low ego, and high standards in a fast-paced, unstructured environment.
Software Engineer (Backend), Enterprise
Scale AI
Scale AI is seeking a Backend Engineer to join their team and build the core infrastructure for large-scale GenAI systems. This role involves designing and implementing scalable APIs, distributed data systems, and robust deployment pipelines to ensure production-grade reliability and performance for enterprise AI products. You will be instrumental in shaping how AI systems are deployed and scaled in the real world, working at the forefront of the GenAI revolution and solving complex backend and infrastructure challenges. This is an opportunity to contribute to cutting-edge solutions that transform workflows and drive efficiency for major enterprises.
Software Engineer, Enterprise
Scale AI
Scale AI is pioneering the next era of enterprise AI, providing cutting-edge solutions that transform workflows and automate complex processes for large enterprises. The Scale Generative AI Platform (SGP) offers foundational services and APIs for seamless AI integration at production scale. This role focuses on building the core infrastructure for large-scale GenAI systems, designing scalable APIs, distributed data systems, and robust deployment pipelines to ensure production-grade reliability and performance. It's an opportunity to solve hard backend and infrastructure challenges that enable AI to work at enterprise scale and shape how AI systems are deployed and scaled in the real world.
Designated Technical Support Engineer
Glean
Glean is seeking a Designated Technical Support Engineer to join its growing startup. Glean offers a Work AI platform that enhances productivity through intelligent search, an AI assistant, and AI agents. The platform provides enterprise-grade infrastructure for managing AI across businesses, with extensive connectors and LLM flexibility. The company is recognized for its innovation in AI and is expanding globally. The Designated Technical Support Engineer will act as a key technical resource for customers, ensuring a high level of service and customer satisfaction.
$120k - $190k
Applied AI, Use-case, Software Engineer (Harness)
Mistral AI
Mistral AI is seeking strong Software Engineers with cybersecurity experience to build and productionize their cybersecurity product offering. This role involves turning cutting-edge prototypes into robust, scalable services that power both offensive and defensive security capabilities. The work will directly enable the deployment of AI-powered cybersecurity at client sites and provide the infrastructure for continuous training and improvement. This is a unique opportunity to shape how AI-powered cybersecurity operates at scale within a dynamic, collaborative team passionate about AI's potential.
Senior Cloud Infrastructure Engineer
Langfuse
Langfuse is seeking a Senior Cloud Infrastructure Engineer to ensure the reliability, performance, and cost-efficiency of their open-source LLM engineering platform. This role is crucial for maintaining Langfuse Cloud on AWS ECS Fargate and ClickHouse Cloud, as well as supporting self-hosted deployments. The engineer will own the end-to-end observability stack, automate infrastructure processes, and scale the platform to meet growing demand. This is an opportunity to work closely with the ClickHouse team and directly impact a product used by major enterprises.