TensorFlow Jobs

62 open roles mentioning TensorFlow

Member of Technical Staff (AI Infrastructure Engineer)

5mo ago
P

Perplexity AI

We are seeking an AI Infrastructure Engineer to join our expanding team. In this role, you will collaborate closely with our Inference and Research teams to construct, deploy, and enhance our extensive AI training and inference clusters. Your work will involve managing Kubernetes and Slurm environments, optimizing distributed training for large language models, and developing robust orchestration systems. This position offers the opportunity to significantly impact the performance and scalability of our AI infrastructure.

London hybrid FullTime
AWSKubernetesPython +5 more

Member of Technical Staff (AI Infrastructure Engineer)

5mo ago
P

Perplexity AI

We are seeking an AI Infrastructure Engineer to join our expanding team. In this role, you will collaborate closely with our Inference and Research teams to construct, deploy, and enhance our large-scale AI training and inference clusters. Our work involves Kubernetes, Slurm, Python, C++, PyTorch, and primarily operates on AWS.

San Francisco onsite FullTime
AWSKubernetesPython +5 more

Member of Technical Staff (AI Inference Engineer)

5mo ago
P

Perplexity AI

We are seeking an engineer to join our team responsible for building and running the inference engine behind Perplexity's queries. This role involves deploying dozens of model architectures at scale, managing tight latency and cost budgets, and working with a stack including Rust, Python, CUDA, and CuTe DSL. You will contribute to supporting new models, migrating GPU kernels, developing a Rust-native serving runtime, optimizing performance, and enhancing reliability and observability.

San Francisco onsite FullTime
KubernetesPythonRust +4 more

Member of Technical Staff (AI Inference Engineer)

5mo ago
P

Perplexity AI

We are seeking an AI Inference Engineer to join our dynamic team. This role is central to Perplexity's operations, as you will build and manage the inference engine that powers every query. You will deploy a variety of model architectures at scale, focusing on meeting stringent latency and cost requirements. Our technology stack includes Rust, Python, CUDA, and the CuTe DSL.

London onsite FullTime
KubernetesPythonRust +4 more

Embedded AI Engineer – Android Automotive (On-Device Intelligence)

5mo ago
A

Applied Intuition

Applied Intuition is seeking an Embedded AI Engineer to build on-device intelligence for a next-generation Android Automotive platform. This role is responsible for the entire lifecycle of embedded ML systems, ensuring models perform predictably and safely within production environments, adhering to real-world constraints like latency, thermal limits, and functional safety. The company is a leader in powering the future of physical AI, serving industries such as automotive, defense, and construction with solutions for tools, infrastructure, operating systems, and autonomy.

$60k - $300k

Sunnyvale onsite FullTime
C#TensorFlowTransformers +2 more

Internship - Search Machine Learning Engineer

5mo ago
P

Perplexity AI

Perplexity is seeking a Search Machine Learning Engineer Intern to contribute to the development of next-generation search technologies, specifically focusing on retrieval and ranking. This internship offers a hands-on opportunity to collaborate with experienced engineers, enhance search quality, experiment with novel models, and implement features that directly influence user search and information discovery experiences. The program is designed for a 12-24 week full-time engagement, conducted in person at our London office.

$12k - $24k

London onsite FullTime
PythonRustRAG +4 more

Senior Member of Technical Staff, Safety and Security for Agents

5mo ago
Cohere

Cohere

Cohere is seeking a Senior Member of Technical Staff to join the Safety and Security for Agents team. In this role, you will significantly contribute to the development of safer, fairer, more trustworthy, and more secure Large Language Models (LLMs). Your work will focus on data generation, post-training algorithms, and evaluation methods to ensure the safety of next-generation models that interact with external resources and take actions. You will collaborate closely with machine learning teams, data annotation teams, and product and policy teams, requiring a blend of machine learning expertise, ethical AI principles, experimental design, and data management skills. This position offers significant autonomy and decision-making power within a small team, with the opportunity to shape the future of LLMs for societal benefit.

London hybrid FullTime
CoherePythonPyTorch +3 more

Machine Learning Engineer, Integrity

6mo ago
OpenAI

OpenAI

As a Machine Learning Engineer in OpenAI's Integrity team, you will work on state-of-the-art models and classifiers, experiment with new architectures and approaches, and advance our capabilities in content and user understanding. This role offers the chance to transform research breakthroughs into tangible solutions that enhance the trust and safety of our platform, with a focus on training LLMs and building ML models. You will be instrumental in designing and deploying advanced machine learning models to solve real-world problems, bringing OpenAI's research from concept to implementation and creating AI-driven applications with direct impact.

San Francisco onsite FullTime
OpenAIFine-TuningPyTorch +3 more

Data Infrastructure Engineer

7mo ago
H

Heyge

At HeyGen, we are at the forefront of developing applications powered by our cutting-edge AI research. As a Data Infrastructure Engineer, you will lead the development of fundamental data systems and infrastructure. These systems are essential for powering our innovative applications, including Avatar IV, Photo Avatar, Instant Avatar, Interactive Avatar, and Video Translation. Your role will be crucial in enhancing the efficiency and scalability of these systems, which are vital to HeyGen's success.

Los Angeles, Palo Alto, San Francisco, Toronto onsite
PythonPyTorchTensorFlow

Research Engineer, Machine Learning

7mo ago
Mistral AI

Mistral AI

Mistral AI is democratizing AI through high-performance, optimized, open-source models and solutions. We are a dynamic, collaborative team passionate about AI's potential to transform society, with a diverse workforce driving innovation. As a Research Engineer – ML track, you will build and optimize large-scale learning systems powering our open-weight models. You will work hand-in-hand with Research Scientists, either enhancing the shared training framework and data pipelines or embedding within a research squad to turn fresh ideas into scalable code.

Palo Alto remote Full-time
MistralPythonPyTorch +9 more

Internship - Search Machine Learning Engineer

8mo ago
P

Perplexity AI

Perplexity is seeking a Search Machine Learning Engineer Intern to contribute to the development of next-generation search technologies, with a specific emphasis on retrieval and ranking. Interns will collaborate with seasoned engineers to enhance search quality, explore novel models, and implement features that directly influence user search and information discovery experiences. This internship program offers a duration of 12-24 weeks, operating on a full-time or part-time basis, and requires in-person attendance at our Belgrade office.

$12k - $24k

Belgrade onsite FullTime
PythonRustRAG +4 more

Member of Technical Staff, Data Analysis and Evaluation

9mo ago
Cohere

Cohere

Cohere is seeking a Member of Technical Staff in Data Analysis and Evaluation to ensure the quality, reliability, and performance of our large language models (LLMs). This role involves designing and conducting data collection tasks, assessing dataset quality, and analyzing model robustness and generalisability. You will collaborate with researchers, engineers, and data annotators to drive data-driven decisions and enhance AI system effectiveness. The position requires expertise in statistics, experimental design, and machine learning to ensure high-quality data and reliable model performance across diverse scenarios, contributing to Cohere's mission of advancing AI.

London remote FullTime
CoherePythonPyTorch +3 more

Audio Inference Engineer, Model Efficiency

10mo ago
Cohere

Cohere

Cohere is seeking an Audio Inference Engineer focused on Model Efficiency to join a fast-growing team of researchers and engineers. The mission of this team is to build reliable machine learning systems and optimize audio inference serving efficiency using innovative techniques. As an engineer on this team, you will advance core audio model serving metrics, including latency, throughput, and quality by diving deep into systems, identifying bottlenecks, and delivering creative solutions for audio processing and streaming workloads. You will collaborate closely with both the training and serving infrastructure teams to ensure seamless integration between model development and deployment, with a special focus on real-time and streaming audio inference.

New York remote FullTime
CoherePythonC# +5 more

Applied Machine Learning Engineer

11mo ago
Cohere

Cohere

Cohere is seeking a Member of Technical Staff for their Applied ML team. In this role, you will collaborate directly with customers to understand their challenges and implement solutions leveraging Large Language Models. You will apply your problem-solving skills, creativity, and technical expertise to bridge the gap in enterprise AI adoption, delivering impactful products and disrupting key industries. This is an opportunity to join at a pivotal moment, shape the company's offerings, and contribute to cutting-edge AI development.

London hybrid FullTime
CoherePythonTensorFlow +2 more

Member of Technical Staff, Agent Code

1y ago
Cohere

Cohere

Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company co-headquartered in Toronto and San Francisco, with key offices in London, New York City, Montreal, Seoul, Germany and Paris. Join us! Why this role? Code-generating LLMs and autonomous agents are revolutionizing how software is built and tasks are automated. At Cohere, we’re pushing the boundaries of what’s possible with these technologies for enterprises, and we’re looking for a senior member for the Agent Code team. You’ll be at the forefront of research and engineering, driving the development of cutting-edge code LLMs and agent systems that can interact with the digital world to solve complex tasks with minimal human oversight. This role is hands-on and research-driven. You’ll dive into the latest literature on code LLMs and agents, experiment with frontier models, and collaborate with a team of talented engineers and researchers to build scalable, production-ready solutions. At Cohere, we blend engineering and research seamlessly—everyone contributes to both, depending on their interests and organizational needs. We provide access to world-class compute resources, data, and talent to ensure you can do your best work. Note: We have offices in London, Toronto, New York and San Francisco, but we’re also remote-friendly! This team operates primarily between ET to CET time zones, so we’re seeking candidates in locations that align with these hours for effective collaboration. As a Member of Technical Staff on the Agent Code team, you will: - Stay up-to-date with the latest research in code LLMs, agents, and related fields, implementing novel ideas into our systems. - Design and implement scalable strategies to train code models, and deploy agent frameworks for inference and sampling. You will be collaborating with the pretraining team, create SFT trajectories and work on existing and new RL algorithms - Hillclimb on existing benchmarks and design new ones that reflect the needs of our enterprise users - Lead experiments on our state-of-the-art compute infrastructure, pushing the boundaries of what’s possible with frontier LLMs. You may be a good fit if you have: - A PhD in Computer Science, Machine Learning, or a related field, with publications in top-tier venues (e.g., NeurIPS, ICML, ICLR, ACL, EMNLP). - Deep expertise in code LLMs and agent systems, with a strong understanding of the latest research and trends. We are looking for people who not only have worked with code models, but have actively contributed to their development - Hands-on experience with frontier LLMs and their applications in code generation or automation. - Strong software engineering skills, with proficiency in Python and PyTorch, TensorFlow, or similar frameworks. - Experience with distributed systems, cloud infrastructure, and scalable architectures. - A proactive, self-motivated mindset, with a passion for solving ambitious, open-ended problems. What We Offer: - The opportunity to work on cutting-edge problems at the intersection of AI, code generation, and autonomous agents. - Access to world-class compute resources, data, and a collaborative team of researchers and engineers. - A remote-friendly, flexible work environment with a focus on impact and innovation. - Competitive compensation and benefits, including equity in a fast-growing AI company. If you’re passionate about shaping the future of code LLMs and agent systems, and thrive in a dynamic, research-driven environment, we’d love to hear from you! Full-Time Employees at Cohere enjoy these Perks: - A weekly lunch stipend of $75/£75 or equivalent in your local currency for lunch. - Full health and dental benefits, including a separate budget for mental health. - RRSP matching, 401K, Pension Scheme. - 100% Parental Leave top-up for up to 6 months, for either parent. - Annual enrichment benefits: Arts & culture, fitness/wellness, quality time, and a workspace improvement credit. Education & learning stipend for conferences, courses, and coaching. - 6 weeks of paid vacation (30 working days!) - Budget for traveling to other offices if you are remote, plus an annual company offsite. How and Where We Work: - Cohere is remote-friendly. We have offices in Toronto, San Francisco, New York City, London, Paris, Montreal, and more coming soon. - For those in the office: a daily lunch program, plenty of snacks, and regular community and social events. - For those not near an office: a co-working benefit so you can work alongside others in your city. - Everyone receives a $500 home office stipend to set up your workspace properly. If any of the above doesn’t line up exactly with your experience, we still encourage you to apply. We strive to create an inclusive work environment for all; we welcome applicants from all backgrounds and are committed to providing equal opportunities. Should you require any accommodations during the recruitment process, please submit an Accommodations Request Form, and we will work together to meet your needs. We may use AI-enabled tools to screen and assess applicants against the criteria for this position. This helps our recruiters identify potentially qualified candidates, but it doesn't limit the applications our recruiters may review or consider.

London hybrid FullTime
CoherePythonPyTorch +1 more

Machine Learning Infrastructure Engineer, Model Inference

1y ago
Abridge

Abridge

Abridge is seeking an ML Infrastructure Engineer, Model Inference to build and optimize the core inference infrastructure powering their machine learning models. This role is crucial for enhancing the scalability, efficiency, and performance of Abridge's AI-driven healthcare solutions. The engineer will collaborate with Infrastructure and Research teams to build, deploy, optimize, and orchestrate AI models, working on a platform that transforms patient-clinician conversations into structured clinical notes in real-time.

SF Office hybrid FullTime
KubernetesPyTorchTensorFlow +3 more

Applied AI Engineer, Senior/Staff Devops/SRE

1y ago
Mistral AI

Mistral AI

Mistral AI is seeking an Applied AI Engineer focused on DevOps to help customers adopt its products and solve complex technical challenges. In this role, you will apply your problem-solving abilities, creativity, and technical skills to assist organizations in leveraging AI for significant impact. You will gain unique insights and contribute to critical global industries and institutions. Applied AI Engineers at Mistral AI work in small teams, owning end-to-end execution of high-stakes projects, which may involve discussing architecture, managing large-scale data, coding custom applications, engaging with customer executives, and strategizing for the Applied Engineering team.

Singapore remote
MistralAWSAzure +10 more

Applied AI Engineer, ML Infrastructure Engineer / Devops - EMEA

1y ago
Mistral AI

Mistral AI

Mistral AI is seeking an Applied AI Engineer focused on DevOps to help customers adopt our AI products and solve complex technical challenges. In this role, you will apply your problem-solving abilities, creativity, and technical skills to assist organizations in leveraging AI for significant impact. You will gain unique insights and contribute to critical global industries and institutions. Your responsibilities will resemble those of a startup CTO, working in small teams to own the end-to-end execution of high-stakes projects, which may involve discussing architecture, managing large-scale data, coding applications, engaging with customer executives, and strategizing for the Applied Engineering team.

Paris remote Full-time
MistralAWSAzure +9 more

Senior Member of Technical Staff, Multimodal AI

1y ago
Cohere

Cohere

Cohere is seeking a Senior Member of Technical Staff focused on Multimodal AI to join their cutting-edge enterprise AI company. This role involves designing and developing advanced multimodal AI systems that integrate text, speech, and vision, pushing the boundaries of what's possible in AI. You will have access to exceptional compute resources and collaborate with world-class teams to innovate and shape the future of AI. The position is ideal for individuals passionate about machine learning and its real-world applications, who enjoy optimizing large models and thrive in a fast-paced, technically challenging environment.

San Francisco remote FullTime
OpenAIMistralCohere +5 more

Open-Source Software, Machine Learning Engineer

1y ago
Mistral AI

Mistral AI

Mistral AI is democratizing AI through high-performance, optimized, open-source models, products, and solutions. We are a dynamic, collaborative team passionate about AI's potential to transform society, with teams distributed globally. We are seeking an Open-Source Software, Machine Learning Engineer to join our OSS team, which is embedded within our Science team. This role is critical in helping turn research breakthroughs into tangible solutions and improving Mistral's open-source ecosystem by open-sourcing state-of-the-art models and maintaining our publicly available libraries.

Paris remote Full-time
MistralPythonGo +10 more

Get new AI jobs in your inbox

A weekly digest of the newest AI engineering roles.

© 2026 AI Job Board. All rights reserved.