Kubernetes Jobs
410 open roles mentioning Kubernetes
Staff / Principal Platform Engineer - USA
inworld
Inworld is seeking a Staff / Principal Platform Engineer to take end-to-end ownership of building, securing, and scaling our AI products. This role involves being the driving force behind our cloud infrastructure, partnering with engineers to deploy and evolve services across major cloud providers using tools like Terraform and ArgoCD. You will identify needs and drive initiatives forward, directly shaping how the company operates and innovates.
$280k - $350k
Cloud Security Engineer
Applied Intuition
Applied Intuition is seeking a highly focused Cloud Security Engineer to play a crucial role in securing our infrastructure across diverse multi-cloud environments (AWS, Azure, GCP, OCI), with a heavy emphasis on Kubernetes cluster hardening. You will establish robust guardrails, enforce Identity and Access Management policies, and maintain our Cloud Security Posture Management (CSPM) to prevent insecure deployments and ensure continuous compliance. This role involves working alongside our Corporate Security & Infrastructure team to secure our cloud footprint.
$60k - $300k
Member of Technical Staff - Platform Foundations
Reflection ai
Reflection is building a company-wide foundations platform to accelerate every engineering team by providing reliable, scalable developer infrastructure, SRE capabilities, and high-throughput data ingestion tooling. This team builds and operates the core platform layer that all engineering teams depend on, defining opinionated golden paths for cloud infrastructure, networking, and access patterns. The goal is to ensure engineers can ship quickly and safely without sacrificing reliability, security, or cost predictability, enabling the development of open foundational models.
Forward Deployed Engineer (India)
cartesia
Cartesia is seeking a Forward Deployed Engineer to join their mission of building real-time multimodal intelligence. This role involves embedding directly with enterprise customers to deliver agentic voice AI solutions into their production environments. The Forward Deployed Engineer will be responsible for maximizing customer success and revenue by transforming Cartesia's core product into deployed, high-impact solutions. This is an opportunity to work on cutting-edge AI models and experiences, with a focus on innovation and rapid execution.
Software Engineer: Platform
Rogo.ai
Rogo is transforming global finance with AI, empowering professionals at top investment banks and firms with unparalleled speed, accuracy, and insight. We are redefining financial workflows and are a rapidly growing company with proven product-market fit, backed by world-class investors. We are scaling quickly and defining a new category of enterprise AI. Our team is sharp, motivated, and deeply committed to our mission, taking ownership of complex problems and focusing on our users. If you thrive in a fast-paced environment, demand excellence, and want to help build the future of finance, we invite you to join us.
Software Engineer - GPU Networking & Distributed Systems
Baseten
Baseten is building the global operating system for distributed, heterogeneous AI hardware, powering mission-critical inference for leading AI companies. As LLM and multi-modal workloads scale, the network becomes a critical component. This role focuses on leading GPU Networking efforts, making RDMA a first-class building block and optimizing distributed inference. You will architect the software fabric that unifies thousands of GPUs, co-optimizing communication and computation for advanced AI workloads.
Solutions Architect
Modal
Modal is seeking a high-impact Solutions Architect to lead technical strategy for its most strategic enterprise accounts. In this role, you will act as the executive technical counterpart to Enterprise Account Executives, guiding complex evaluations, shaping infrastructure modernization roadmaps, and promoting multi-product adoption for AI/ML workloads. This is a strategic, consultative position that requires deep architectural knowledge, executive presence, and the ability to influence significant infrastructure decisions. You will collaborate directly with CTOs, VPs of Engineering, and ML platform leaders to reimagine how AI infrastructure is built and operated. If you excel in fast-paced technical sales environments and are passionate about shaping the infrastructure powering modern AI companies, this role is for you.
Staff Cloud Support Engineer
crusoe
As a Staff Cloud Support Engineer at Crusoe, you will be a technical authority within Crusoe Cloud, acting as a force multiplier for Customer Experience, SRE, Networking, Fleet, and Product teams. Your role extends beyond simple ticket resolution; you will design reliability guardrails, influence architectural decisions, mentor other engineers, and directly contribute to revenue protection by preventing large-scale incidents. This position requires deep expertise in Linux systems, Kubernetes, networking, and AI/ML infrastructure, applied with a strong customer focus. You should be comfortable operating in ambiguous environments, leading incident response efforts, and shaping the scalability of Crusoe's high-performance AI infrastructure globally.
$156k - $190k
Security and Compliance Manager
sierra.ai
Sierra is a leading platform for customer-facing AI agents, partnering with major global brands to transform customer service and business growth. We are primarily an in-person company based in San Francisco, with expanding offices internationally. Our culture is built on core values of Trust, Customer Obsession, Craftsmanship, Intensity, and a commitment to balancing Family. The company was co-founded by Bret Taylor, former co-CEO of Salesforce and CTO of Facebook, and Clay Bavor, who spent 18 years at Google leading initiatives like Google Labs, AR/VR, and Google Lens.
$53k - $800k
Senior Software Engineer, AI Platform
decagon
Decagon is seeking an experienced software engineer to join the AI Platform team and help build the company's AI-native developer platform. This role sits at the intersection of distributed systems, developer tooling, and applied AI, focusing on experimenting with new agentic workflows and transforming promising ideas into dependable systems for the engineering organization. The position is highly product-oriented, requiring close collaboration with engineers to identify automation opportunities and ownership of solutions from architecture through production operation.
$200k - $400k
Software Engineer, Enterprise Platform
Replit
Join our Enterprise Platform team and build the infrastructure foundations that enable the world's largest organizations to run Replit within their security and compliance boundaries. As a Software Engineer on this team, you'll design and implement the deployment flexibility, networking capabilities, authorization systems, and data controls that enterprises require, from single-tenant architectures and private connectivity to custom policy enforcement and customer-managed encryption. You'll work at the intersection of cloud infrastructure and enterprise requirements, partnering with Platform Engineering, Security, and Sales to ship capabilities that unlock adoption at demanding organizations.
Software Engineer, Growth Infrastructure
Replit
Replit is seeking an experienced Growth Infrastructure Engineer to build and maintain the technical foundation for scalable growth experiments, high-performance data pipelines, and automated systems. This role is at the intersection of growth, product, and infrastructure, requiring deep technical engineering skills combined with an understanding of experimentation and data-driven optimization. You will collaborate with product, data science, and backend teams to ensure growth initiatives run smoothly and scale efficiently across systems. This is a unique opportunity to be an early member of a new Growth team, with significant ownership and influence over technical direction and product outcomes, working on a product that has experienced massive user growth.
Member of Engineering (Pre-training / Data Engineering)
poolside
Poolside is building a world where AI drives economically valuable work and scientific progress, aiming to accelerate software development through agentic systems and frontier models. This role is a core part of the Pretraining Data team, responsible for building and scaling the Model Factory, a system for rapid training, scaling, and experimentation with foundation models. The primary mission is to architect and maintain high-performance pipelines that transform trillions of raw tokens into high-quality dataset "fuel" for models. This involves engineering ingestion, deduplication, and streaming systems for petabyte-scale data, bridging the gap between raw web crawls and GPU clusters, and directly influencing model performance through superior data modeling and distributed pipeline optimization. Collaboration with Pretraining, Posttraining, Evals, and Product teams is key to generating high-quality datasets that address missing model capabilities and downstream use cases.
Software Engineer, ML Infrastructure
cursor
The ML Infrastructure team builds large-scale compute, storage, and software infrastructure to support the company's work building the world's best agentic coding model. This role works closely with ML researchers and engineers to enable their work through improvements to our training framework, systems reliability/performance, and developer experience. We are looking for strong engineers interested in building high-performance infrastructure and the software to support it.
Senior Software Engineer, Core Infrastructure
Harvey
As a Software Engineer on the Core Infrastructure team at Harvey, you will play a critical role in designing and building new infrastructure systems while equally scaling and strengthening our existing infrastructure. This foundation powers every user interaction with Harvey, processing billions of prompt tokens and millions of daily requests across our global legal AI platform. You will work in an environment balanced between innovation and operational excellence, ensuring Harvey remains resilient and efficient as it scales products, regions, customers, and usage. Your contributions will directly impact the reliability, scalability, and security of our platform as we serve the world's leading law firms and professional service providers.
$200k - $250k
Software Engineer - Training Product
Baseten
Baseten is seeking a customer-obsessed software engineer to join their team and contribute to the development of mission-critical AI inference platforms. In this role, you will own features from conception to launch, working across the entire technology stack from API and UI down to the infrastructure layer. You will have the opportunity to fine-tune models, gain a deep understanding of user workflows, and collaborate closely with research engineers to build cutting-edge experiences that accelerate model development and address real-world pain points. If you are excited about diving deep into AI model training and building impactful products, this is the role for you.
MTS, Security
fireworks ai
Fireworks is seeking a Security Engineer to play a key role in designing, implementing, and operating security controls across AI infrastructure, AI platforms, and internal systems. This role is crucial for strengthening our security posture and supporting rapid growth, ensuring the confidentiality, integrity, and availability of data, models, and infrastructure as organizations increasingly rely on large language models and cloud-native AI services. You will be instrumental in building trust by embedding security across all layers of our technology stack.
Member of Technical Staff - Platform Engineering
Modal
AI needs a new infrastructure layer, and Modal is building it. We provide instant GPU access, sub-second container starts, and native storage, enabling customers to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. As a rapidly growing cloud infrastructure company, we are seeking to dramatically improve our platform's reliability while scaling our team and customer base. This role is ideal for individuals with deep systems thinking, a passion for reliability, and a drive to enable others to move faster at scale.
Deployed Engineer (UK)
Langchain
LangChain is seeking a Deployed Engineer to join their team, focusing on building and running AI agents in production. This role involves working on complex applied AI systems that real teams depend on, with a fast feedback loop and visible impact. You will collaborate with customer engineering teams to co-architect and co-build production AI agents, own the technical win in pre-sales by designing POCs and answering technical questions, and help customers deploy and operate agent-based applications. You will also advise customers on architecture and best practices, run technical demos and trainings, and contribute reusable patterns and code. This position offers the opportunity to directly shape how AI agents are built and adopted in the real world.
Staff Software Engineer, Developer Experience
Harvey
Why Harvey At Harvey, we’re transforming how legal and professional services operate. By combining frontier agentic AI, an enterprise-grade platform, and deep domain expertise, we’re reshaping how critical knowledge work gets done for decades to come. This is a rare chance to help build a generational company at a true inflection point. With 1500+ customers in 60+ countries, strong product-market fit, and world-class investor support, we’re scaling fast and defining a new category in real time. The work is ambitious, the bar is high, and the opportunity for growth — personal, professional, and financial — is unmatched. Our team moves fast, takes ownership, and is deeply committed to the mission — operating with intensity, staying close to our customers, and pushing each other for excellence. We live by three values: Decisiveness, Simplicity, and Job's Not Finished. We act quickly on clear judgment over perfect information, we believe simplicity is what scales, and we're never satisfied with where we are. If you want to do the best work of your career alongside people who share that drive, we'd love to build with you. At Harvey, the future of professional services is being written today — and we’re just getting started. Role Overview The Developer Experience team is what keeps Harvey’s fast growth momentum possible — the kind of fast where every week brings new products, new capabilities, and new challenges. We build the CI/CD systems, internal frameworks, and platform-level guardrails that enable all Harvey engineers to ship ambitious AI features rapidly and safely. As a Software Engineer on the Developer Experience team, you will be responsible for creating frameworks and systems that maximize the velocity and efficiency of every engineer at Harvey. You’ll work across the entire tech stack and collaborate with every team to instill reliability, automation, and simplicity into our developer workflows. You'll also have a unique opportunity to apply AI directly to the problem of developer productivity, defining the future of software development within an AI-native company. If you are passionate about building foundational platforms to enable other engineers to move fast with confidence, we’d love to hear from you. What You’ll Do - Develop and scale a world-class developer platform to accelerate Harvey's hyper growth. Boost velocity and stability through robust CI/CD systems, effective test frameworks, and reliable development environments.. - Build load testing and benchmarking infrastructure essential for evaluating and optimizing the performance of AI-native applications. - Pioneer the future of software development and site reliability engineering by integrating AI agents across the software development, deployment and maintenance lifecycle. - Collaborate with Backend Platform teams to embed testability, reliability and observability into the platform, ensuring services built on our foundation are robust, easy to test and maintain. - Work closely with engineering teams to gather feedback, evangelize best practices, and make the “paved road” a reality — empowering every Harvey engineer to move fast with confidence. - Set the strategic direction and roadmap for scaling developer experience as Harvey expands, and contribute strategically to team decision-making. - Provide strong technical leadership and mentorship, upholding a high bar for engineering excellence across the team. What You Have - 7+ years of software engineering experience, including building scalable backend systems or internal developer platforms. - Proficiency in Python (or similar languages) and deep knowledge of backend development fundamentals and distributed systems. - Hands-on experience with CI/CD systems (Builtekite, Github Actions), test frameworks, or load and performance testing. - Hands-on experience with container technologies (Docker, Kubernetes) and infrastructure as code (Pulumi, Terraform). - A track record of producing high-quality, well-tested code, consistently following software development best practices to ensure quality and reliability. - Proven technical leadership throughout the entire project lifecycle, including ideation, design, implementation, and productionization. - Experience mentoring engineers, guiding architectural decisions, and shaping culture to foster engineering excellence. - Strong problem-solving skills and a passion for improving developer experience — you enjoy creating tools or frameworks that make other engineers more productive - Excellent collaboration and communication skills, with the ability to work across teams and incorporate feedback. - Experience integrating AI agents into a developer ecosystem is a strong plus. Compensation Range $238,000 - $290,000 USD Depending on your location, an Applicant Privacy Notice may apply to you. You can find all of our Applicant Privacy Notices [here]. #LI-AN2 Harvey is an equal opportunity employer and does not discriminate on the basis of race, gender, sexual orientation, gender identity/expression, national origin, disability, age, genetic information, veteran status, marital status, pregnancy or related condition, or any other basis protected by law. We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made by emailing [email protected]
$238k - $290k