Kubernetes Jobs
410 open roles mentioning Kubernetes
Software Engineer, AI Productivity
Physical Intelligence
Physical Intelligence is seeking an AI Productivity Software Engineer to build and deploy tools that enhance AI utilization across the company. This role involves collaborating with various teams, including engineering, research, and operations, to identify opportunities for AI-driven leverage and translate them into robust internal tools and workflows. The position is within the Runtime team, focusing on AI tooling to make AI agents, assistants, integrations, and automation effective for company-wide use. The goal is to deeply understand team workflows, build tailored tools, and drive their adoption to become integral to the company's operational rhythm.
Staff Software Engineer, Identity & Authorization
Replit
Replit is seeking a Staff Software Engineer to join the Product Platform team, specifically focusing on Identity & Authorization. This role is crucial for building and operating the systems that protect critical interactions across Replit's platform, including how people, agents, sandboxes, and services prove their identity and permissions. The work is high-leverage, ensuring that clear, reliable, and easy-to-adopt identity and policy systems enable all other teams to move faster and more securely. The team values curiosity, clear thinking, and collaborative, open work, prioritizing reasoning and building skills.
IAM Security Engineer
Cloudflare
As an Identity and Access Management (IAM) Security Engineer at Cloudflare, you will be instrumental in developing, deploying, and scaling robust identity and access management solutions for both internal employees and workloads. Your primary focus will be on securing our systems, applications, and data by ensuring the integrity of user access, authentication, and authorization processes. This role involves a hands-on approach to building and managing critical IAM components, contributing to the overall security posture of the organization.
Product Security Engineer, North Security
Cohere
Cohere is seeking a Senior Product Security Engineer to join their North Security team. This is a hands-on engineering role focused on securing AI-powered products for enterprises handling sensitive data. You will work closely with product engineers to review architecture and code, perform threat modeling, conduct hands-on testing, and develop scalable security guardrails. The goal is to identify and mitigate risks specific to AI systems, such as prompt injection, unsafe tool use, and data exposure, and to translate learnings into reusable security defaults for other teams. You will also be responsible for strengthening the security capabilities of engineering teams through guidance and collaboration.
Developer Productivity
Langchain
LangChain is seeking a Software Engineer to join its Infrastructure team, focusing on enhancing developer productivity across its LangGraph Cloud/Platform and LangSmith products. This role involves close collaboration with Infrastructure, Backend, and Frontend teams to ensure confident deployment of services, APIs, and UI flows. A key aspect of this position will be pioneering quality practices specifically for LLM applications, including prompt regression testing and comprehensive evaluation suites. The ideal candidate will contribute to building and maintaining the systems that power LangChain's developer platform, emphasizing reliability, scalability, and overall developer efficiency.
$175k - $240k
Threat Intelligence Software Engineer, Cloudforce One
Cloudflare
Cloudforce One is Cloudflare’s threat operations and research team, focused on identifying and disrupting cyber threats, from criminal activity to nation-state sponsored APTs. The team collaborates with external organizations and internal Cloudflare teams to develop operational tradecraft and expand threat intelligence sources for expedited threat hunting and remediation. Leveraging data from one of the world’s largest global networks, team members analyze unique data points at scale to synthesize actionable threat intelligence and protect customers. As a Software Engineer on this team, you will manage the full software development lifecycle for systems supporting threat disruption and legal response. This involves translating complex legal and security requirements into robust, scalable, and high-performance applications that enhance internet safety and power.
Security Engineer, Offensive Security
Anthropic
Anthropic is building reliable, interpretable, and steerable AI systems to be safe and beneficial for users and society. The Security Engineering team is dedicated to safeguarding these AI systems and maintaining user trust. This involves developing critical security infrastructure, establishing secure development practices, and collaborating with research and product teams to operate as a world-class security organization.
Forward Deployed Engineer, Infrastructure Specialist (Middle East)
Cohere
Cohere is a leading enterprise AI company building cutting-edge foundation AI models and end-to-end products. We are seeking engineers passionate about Agentic AI to join our team and shape how enterprises harness AI in real-world applications. As a Forward Deployed Engineer, you will act as a bridge between our core North product and client engineering teams, solving complex problems and securely integrating AI into critical sectors. This role involves working at the forefront of AI deployment, ensuring secure and efficient integration into client environments.
$20k - $40k
HPC Storage Engineer - West Coast
Runpod
Runpod is seeking a Senior Storage Engineer to join their remote-first Infrastructure team. This role is critical to the AI Developer Cloud, focusing on the design, scaling, and reliability of Runpod's multi-region storage ecosystem, which includes network volumes, local NVMe, and S3-compatible object storage. You will be responsible for implementing and maintaining distributed storage deployments, writing code and automation to manage them, and contributing to strategic decisions about future storage systems. This hands-on position requires a deep understanding of storage internals, performance tuning, and a proactive approach to problem-solving and continuous improvement, directly impacting the performance and reliability of AI workloads for over a million developers.
$180k - $260k
Developer Support Engineer (London)
Braintrust
Braintrust is seeking Developer Support Engineers, both mid-level and senior, who are passionate about helping developers overcome technical challenges and achieve their goals. In this role, you will troubleshoot issues, identify workarounds, implement fixes, and document your findings to accelerate the progress of other developers. This position combines technical problem-solving, empathy for developers, and close collaboration with Engineering, Solutions, and Product teams. If you excel at resolving complex problems, articulating technical concepts clearly, and improving the developer experience, we encourage you to apply. This role is based remotely in London.
Full-Stack Engineer
Nabla
Nabla is seeking a Full-Stack Engineer to join one of their cross-functional squads, contributing end-to-end to specific product areas. This role involves building and improving product features across web and desktop platforms, contributing to back-end APIs and services, and collaborating closely with Product, ML, and Design teams to bring projects from conception to production. The engineering team is lean, fast-moving, and deeply technical, focused on delivering AI into clinical settings reliably and at scale. Nabla is backed by a $70M Series C funding round and is at the forefront of developing AI assistants for healthcare, aiming to restore the human connection in medicine.
Platform Support Engineer (Singapore)
Braintrust
Braintrust is seeking Platform Support Engineers at both mid and senior levels to join a small, high-ownership team. This role focuses on providing technical front-line support for infrastructure, performance, and reliability for customers who run Braintrust within their own AWS, Azure, and GCP accounts. You will work closely with Cloud Infrastructure and Engineering teams, addressing complex infrastructure challenges for critical customer deployments.
Platform Support Engineer
Braintrust
Braintrust is seeking Platform Support Engineers at both mid and senior levels to join a small, high-ownership team. This role focuses on providing technical support for customers who deploy Braintrust within their own AWS, Azure, and GCP accounts, often behind their own VPCs and compliance requirements. You will be the technical front line for infrastructure, performance, and reliability issues, working closely with Cloud Infrastructure and Engineering teams. If you enjoy solving complex infrastructure challenges with a direct customer impact, this is the role for you.
Senior Distributed Systems Engineer - Cache
Cloudflare
Cloudflare is seeking a Senior Distributed Systems Engineer to join the Cache team. This team builds and operates the high-performance reverse proxy and caching data plane at Cloudflare's edge, built on Pingora, Cloudflare's open-source Rust framework. The role involves working on backend routing, load balancing, cache storage, globally distributed purge, Tiered Cache routing, and production observability. You will design, implement, test, roll out, and operate these systems, ensuring performance, correctness, and resilience across Cloudflare's global network. This position is ideal for individuals who enjoy tackling complex problems related to latency, correctness, and reliability in large-scale distributed systems.
$194k - $266k
Member of Technical Staff, New Grad (BS/MS)
fireworks ai
This role is designed for new graduates who are eager to work on real-world AI systems in production from day one. As a Member of Technical Staff at Fireworks, you will write and deploy code that operates on one of the world's most active inference platforms, serving hundreds of models to developers and enterprises at a massive scale. You will be matched to a team based on your strengths and interests, with potential work spanning inference and performance, model training and fine-tuning, distributed systems and cloud infrastructure, developer platform and APIs, or product and full-stack application engineering. You will be supported by a senior engineer mentor and a structured ramp-up program, starting with well-scoped problems and progressing to owning features and systems independently. The company emphasizes a fast-paced environment with a high bar and a commitment to employee development.
Staff Applied AI Inference Engineer
crusoe
Crusoe is seeking a Staff Applied AI Inference Engineer to accelerate the abundance of energy and intelligence by optimizing large language models for production environments. This role involves owning the inference stack end-to-end, from profiling costs and implementing modern optimization techniques to deep dives into serving code when defaults are insufficient. The work is applied, focusing on real-world deployments with specific models, traffic patterns, latency targets, and cost constraints. You will collaborate with customer engineering teams to tailor deployments, transition workloads from proof-of-concept to fully monitored production services, and ensure engineered gains benefit end-users. This is a hands-on engineering position requiring coding, profiling, and low-level optimization, with a customer-facing component involving product and technical solutions work.
$215k - $260k
Member of Technical Staff, Systems Infrastructure (2026 PhD New Grad)
fireworks ai
Fireworks is seeking PhD graduates to join their Systems Infrastructure team. This role is designed for individuals finishing their PhD in Computer Science, Computer Engineering, Electrical Engineering, or a similar field, who are interested in applying their research to large-scale, real-world AI infrastructure. You will be responsible for designing and building the core systems that power Fireworks, including schedulers, storage systems, and networks, to ensure efficient operation of tens of thousands of accelerators and low inference latency. You will tackle complex problems such as optimizing job placement on heterogeneous hardware, high-speed data movement for model weights, and maintaining saturated datacenter networks. You will be paired with a senior engineer mentor and work on impactful projects from day one, with start dates flexible around thesis defense.
Software engineer, connectors & MCP
Writer
WRITER is seeking an exceptional software engineer to join its rapidly evolving team. In this pivotal role, you will be at the forefront of expanding human capacity by building the next generation of AI-powered solutions that transform how leading enterprises operate. You will dive deep into developing a state-of-the-art platform that leverages cutting-edge generative AI technologies, from large language models to sophisticated agentic workflows, delivering seamless, scalable, and secure applications that redefine enterprise productivity. This is an unparalleled opportunity to make a tangible impact, shaping the future of AI and contributing to a product that’s changing how the world works.
Software Engineer - Voice Model
xAI
SpaceXAI's mission is to create AI systems that can accurately understand the universe and aid humanity in its pursuit of knowledge. The Grok Voice Model team is building the world's best voice AI, delivering smooth, natural, low-latency spoken interactions that are expressive, multilingual, and reliable across devices and real-time scenarios. The team owns the full training pipeline, from massive data curation and premium audio processing to frontier speech-language pre-training and intensive post-training to push quality, speed, and stability to the limit. The goal is to make talking to AI feel like conversing with a charming, kind, and knowledgeable person, and exceptionally smart, execution-oriented engineers are sought to achieve this.
$150k - $450k
Member of Technical Staff - Multimodal Understanding
xAI
SpaceXAI is seeking a Member of Technical Staff to join their multimodal team and advance the understanding and generation of AI across image, video, audio, and text. This role involves working across the full stack, from data curation and pre-training to alignment, infrastructure, and end-to-end product experiences. You will collaborate with various teams to deliver cutting-edge multimodal reasoning, world modeling, tool use, agentic behaviors, and human-AI collaboration capabilities. The goal is to build models that can perceive, reason about, and interact with the world in real-time at an unprecedented level.
$180k - $440k