Kubernetes Jobs
204 open roles mentioning Kubernetes
Senior Software Engineer, Anti-Abuse & Security
Replit
The Anti-Abuse team at Replit is responsible for defending the platform from exploitation by detecting and shutting down various forms of abuse, including phishing, cryptomining, and LLM token farming. This role involves building advanced detection systems, heuristics, and automated responses to stay ahead of constantly adapting attackers. A unique aspect of this position is the AI-native nature of Replit's platform, offering hands-on experience in applying AI to security problems, such as building guardrails for AI-generated code, detecting prompt injection attacks, and using LLMs defensively. You will own problems end-to-end, from identifying abuse patterns to shipping scalable solutions.
Forward Deployed Engineer - AI Engineer
Reflection ai
Reflection is a research lab dedicated to making intelligence open and accessible. We are seeking a core member for our Applied AI team to lead Forward Deployed Engineering efforts with enterprise customers. This role involves translating advanced AI research into high-impact, real-world applications, owning the technical strategy and delivery of agentic systems from discovery to production launch.
Product Manager, Forge
Mistral AI
Mistral AI is seeking a talented and experienced Product Manager to define and execute the strategy for Forge, a product that empowers customers to build, fine-tune, and deploy custom AI models at scale. Forge transforms cutting-edge research into enterprise-ready capabilities by supporting model fine-tuning, reinforcement learning, and post-training workflows. This role operates at the intersection of research and product, enabling customers to train specialized models for real-world business value. You will collaborate closely with applied AI scientists and research engineers to translate frontier techniques into scalable and reliable solutions, shaping a 0-1 product with significant business impact and defining the future of how organizations train and deploy AI models.
Forward Deployed Engineer, Infrastructure Specialist (Europe/Middle East)
Cohere
Cohere is a leading security-first enterprise AI company building cutting-edge foundation AI models and end-to-end products. We are seeking engineers to join our team and contribute to the widespread adoption of AI. This role offers a unique opportunity to shape how enterprises harness the power of AI in real-world applications, acting as a bridge between our core North product and client engineering teams. You will be at the forefront of solving complex problems and securely integrating AI into critical sectors like finance, healthcare, and telecommunications, working with esteemed clients and focusing on Agentic AI.
Software Engineer, New Grad
Mistral AI
Mistral AI is seeking early-career software engineers, including new graduates or those with up to 1-2 years of experience, to join their software engineering teams. In this role, you will contribute to building and enhancing the core systems that power Mistral's products, such as AI Studio and Applications, as well as supporting operations like SRE, data, and security. You will play a key part in shaping how users and developers interact with the AI platform at scale, working closely with experienced engineers who will mentor and support your growth. This is an opportunity for enthusiastic graduates to learn, collaborate, and deliver impactful products in a fast-paced environment.
Senior Software Engineer, Site Reliability Engineer
Harvey
Harvey is transforming legal and professional services by combining frontier agentic AI, an enterprise-grade platform, and deep domain expertise. We are a fast-scaling company with strong product-market fit, seeking ambitious individuals to help build a generational company. Our team operates with intensity, ownership, and a commitment to our mission, valuing decisiveness, simplicity, and continuous improvement. As a Software Engineer on the Site Reliability team, you will ensure the reliability, scalability, and performance of our legal AI platform, owning the systems that keep our platform fast, secure, and always on. Your work will be crucial in maintaining platform resilience as we grow across 50+ regions.
$200k - $260k
Staff Software Engineer, Site Reliability Engineer
Harvey
Why Harvey At Harvey, we’re transforming how legal and professional services operate. By combining frontier agentic AI, an enterprise-grade platform, and deep domain expertise, we’re reshaping how critical knowledge work gets done for decades to come. This is a rare chance to help build a generational company at a true inflection point. With 1500+ customers in 60+ countries, strong product-market fit, and world-class investor support, we’re scaling fast and defining a new category in real time. The work is ambitious, the bar is high, and the opportunity for growth — personal, professional, and financial — is unmatched. Our team moves fast, takes ownership, and is deeply committed to the mission — operating with intensity, staying close to our customers, and pushing each other for excellence. We live by three values: Decisiveness, Simplicity, and Job's Not Finished. We act quickly on clear judgment over perfect information, we believe simplicity is what scales, and we're never satisfied with where we are. If you want to do the best work of your career alongside people who share that drive, we'd love to build with you. At Harvey, the future of professional services is being written today — and we’re just getting started. Role Overview As a Staff Software Engineer on the Site Reliability team at Harvey, you will ensure the reliability, scalability, and performance of our legal AI platform. You’ll join a high-leverage team that sits at the intersection of infrastructure and product, owning the systems that keep our platform fast, secure, and always on. From scaling across 50+ regions to automating mission-critical operations, your work will ensure that Harvey remains resilient as we grow. If you’re passionate about building robust systems and reducing complexity through automation, we’d love to work with you. This role is based in San Francisco, CA. We use an in-person work model and offer relocation assistance to new employees. What You’ll Do - Design, implement, and manage monitoring, alerting, and infrastructure resources (compute, storage, networking) across 50+ global regions - Lead incident management processes, including postmortems, root cause analyses, and driving actionable improvements - Automate operational tasks and workflows, building tools and processes for capacity planning, graceful rollouts, and safe data access to maintain high reliability and reduce manual intervention - Establish best practices for security, compliance, and reliability and collaborate across teams to drive these principles throughout the software lifecycle - Optimize infrastructure costs through strategic capacity planning and build-versus-buy decisions while maintaining system performance, reliability, and functionality - Provide technical mentorship and leadership, promoting best practices and fostering team growth What You Have - 10+ years of experience in Site Reliability Engineering or similar roles supporting production environments, with proven ability to mentor and guide technical teams - Expertise in infrastructure as code(IaC) tools (Pulumi, Terraform, CloudFormation, etc.) - Deep familiarity with observability tools (Datadog, Sentry, etc.) and incident response practices (PagerDuty, IncidentIO, etc.) - Proficiency with cloud infrastructure platforms (Azure, GCP, AWS, etc.) - Strong programming skills (Python, Bash, Go, or similar languages) - Proven track record of diagnosing complex system problems and implementing durable solutions - Solid understanding of CI/CD, Kubernetes, containerization, networking, databases, and cloud security principles - Excellent problem-solving skills, meticulous attention to detail, and a commitment to operational excellence Compensation Range $238,000 - $290,000 USD Depending on your location, an Applicant Privacy Notice may apply to you. You can find all of our Applicant Privacy Notices [here]. #LI-AN2 Harvey is an equal opportunity employer and does not discriminate on the basis of race, gender, sexual orientation, gender identity/expression, national origin, disability, age, genetic information, veteran status, marital status, pregnancy or related condition, or any other basis protected by law. We are committed to providing reasonable accommodations to applicants with disabilities, and requests can be made by emailing [email protected]
$238k - $290k
Senior ML Systems Engineer, Frameworks & Tooling
Cohere
Cohere is seeking a Senior ML Systems Engineer to join their team and build, maintain, and evolve the training framework that powers their frontier-scale language models. This role is ideal for someone passionate about large-scale training, distributed systems, and HPC infrastructure, offering the opportunity to design and maintain core components for fast, reliable, and scalable model training. You will also build tooling to connect research ideas to thousands of GPUs, working across the full stack of ML systems with significant autonomy and impact.
Applied AI Engineer, Prototyping
Mistral AI
Mistral AI is seeking an Applied AI Engineer to join their Proto Team, which acts as the technical pre-sales arm of the GTM organization. This role operates at the intersection of product and customer, translating Mistral's AI technology into real-world solutions that deliver value quickly. The engineer will build high-impact, full-stack AI solutions for global customers on short timelines, owning the end-to-end execution from scoping to deployment. This position also involves collaborating with GTM, product, and engineering teams, contributing to internal tools and product improvements, and testing new capabilities to inform product direction. The ideal candidate is a fast-moving, highly curious builder who thrives in complexity and ambiguity, possesses strong problem-solving skills, and has a hacker mindset with deep engineering instincts.
AI Scientist - Warsaw
Mistral AI
Mistral AI is a pioneering company dedicated to democratizing AI through high-performance, optimized, open-source models, products, and solutions. We aim to simplify tasks, save time, and enhance learning and creativity by integrating cutting-edge AI into daily working life. Our comprehensive platform serves both enterprise and personal needs, featuring offerings like Le Chat, La Plateforme, Mistral Code, and Mistral Compute. We are a dynamic, collaborative, and diverse team passionate about AI's potential to transform society, driving innovation from our distributed teams across France, USA, UK, Germany, and Singapore. Join us to shape the future of AI and make a meaningful impact.
AI Scientist - Zurich
Mistral AI
Mistral AI is a pioneering company dedicated to democratizing AI through high-performance, optimized, open-source, and cutting-edge models, products, and solutions. Our comprehensive AI platform serves both enterprise and personal needs, offering tools like Le Chat, La Plateforme, Mistral Code, and Mistral Compute to bring frontier intelligence to end-users. We are a dynamic, collaborative, and diverse team passionate about AI's potential to transform society, thriving in competitive environments and committed to driving innovation. Join us to be part of shaping the future of AI and making a meaningful impact.
AI Scientist - Audio
Mistral AI
Mistral is on a mission to democratize AI by producing frontier intelligence for everyone, developed in the open. We are a dynamic, collaborative team passionate about AI's potential to transform society, with diverse teams distributed globally. We develop models for enterprise and consumers, focusing on systems that change business operations and integrate into daily lives, while also releasing frontier models open-source. We are hiring experts in large language model training and distributed systems to shape the future of AI.
Software Engineer, Enterprise Agents
Mistral AI
Mistral AI is seeking a Software Engineer to design and build core systems that drive the company's growth, focusing on the Enterprise Agents area. This role involves developing internal tools, integrations, connectors, workflows, and dashboards to enhance information flow across the organization. The engineer will collaborate closely with internal stakeholders to understand business needs, translate complex workflows into scalable software, and deliver pragmatic solutions that optimize team operations. The position is at the intersection of engineering, finance, data, and operations, contributing to a dynamic and collaborative team environment.
Forward Deployed Engineer, Infrastructure Specialist (North America)
Cohere
Cohere is a leading enterprise AI company building cutting-edge foundation AI models and end-to-end products. We are seeking engineers to join our team and contribute to the widespread adoption of AI. This role offers a unique opportunity to shape how enterprises harness the power of AI in real-world applications, acting as a bridge between our core North product and client engineering teams. You will be at the forefront of solving complex problems and securely integrating AI into critical sectors.
Machine Learning Infrastructure Engineer, Model Inference
Abridge
Abridge is seeking an ML Infrastructure Engineer, Model Inference to build and optimize the core inference infrastructure powering their machine learning models. This role is crucial for enhancing the scalability, efficiency, and performance of Abridge's AI-driven healthcare solutions. The engineer will collaborate with Infrastructure and Research teams to build, deploy, optimize, and orchestrate AI models, working on a platform that transforms patient-clinician conversations into structured clinical notes in real-time.
Member of Technical Staff, Integration/RL Team (Research Engineer)
Cohere
Cohere is a leading enterprise AI company focused on building cutting-edge foundation AI models and end-to-end products for real-world business problems. The integration team specifically focuses on developing and scaling machine learning algorithms and infrastructure for LLM post-training, with an emphasis on large-scale, distributed Reinforcement Learning (RL) methods. This role is crucial for enhancing the post-training codebase by implementing new research tools, optimizing algorithms, and scaling distributed RL capabilities. We are looking for passionate individuals who are meticulous in their approach to engineering and science, contributing to both production code and research efforts.
Member of Technical Staff, Post-Training
Cohere
Cohere is seeking a Member of Technical Staff to focus on post-training of AI models. This role is crucial for advancing the state of the art in model post-training and shipping cutting-edge models to production, bridging the gap between research and practical application. You will have access to significant compute resources and a talented team to contribute to increasing model capabilities and driving customer value. The position offers a unique opportunity to contribute to both production code and research efforts, depending on your interests and organizational needs.
Applied AI Engineer, Senior/Staff Devops/SRE
Mistral AI
Mistral AI is seeking an Applied AI Engineer focused on DevOps to help customers adopt its products and solve complex technical challenges. In this role, you will apply your problem-solving abilities, creativity, and technical skills to assist organizations in leveraging AI for significant impact. You will gain unique insights and contribute to critical global industries and institutions. Applied AI Engineers at Mistral AI work in small teams, owning end-to-end execution of high-stakes projects, which may involve discussing architecture, managing large-scale data, coding custom applications, engaging with customer executives, and strategizing for the Applied Engineering team.
Senior Application Security Engineer
Abridge
Abridge is seeking a highly experienced and motivated Senior Application Security Engineer to join their growing team. This is a key technical leadership role focused on building security from the ground up at the forefront of AI in healthcare. You will shape product, infrastructure, and engineering practices by impacting the vision and execution of the secure software development lifecycle (SDLC) across the entire product portfolio. The role involves collaborating with product and engineering teams to integrate security seamlessly, automate security controls, and foster a secure-by-default culture. This position demands deep technical expertise, a builder's mindset, and strong communication skills to influence security across the organization.
Software Engineer, Backend (Paris)
Mistral AI
Mistral AI is seeking passionate and skilled software engineers to join our team as Backend Engineers. You will contribute to the development of core systems powering our main products, AI Studio, Le Chat & Mistral Code, shaping how users and developers interact with our AI platform at scale. You'll work on building reliable, high-performance backend systems and APIs that serve millions of users and developers, directly impacting the developer and user experience. We welcome engineers at all levels, from fresh graduates to senior and staff engineers, who are eager to learn, collaborate, and ship impactful products.