JAX Jobs
75 open roles mentioning JAX
Research, Post-Training Data
thinkingmachines
Thinking Machines Lab is seeking researchers to bridge the gap between raw AI intelligence and useful, safe, and collaborative systems. This role focuses on post-training data research, combining human insight and machine learning techniques to capture and steer model behavior based on human preferences. You will be responsible for translating research ideas into actionable data through labeling and collection campaigns, understanding data quality science, and developing metrics to measure the impact of data and training interventions. The position also involves exploring new paradigms for human-AI interaction and scalable oversight, blending research, data operations, and technical implementation to advance human-centered AI systems. This role requires both fundamental research and practical engineering, making it ideal for individuals who enjoy deep theoretical exploration and hands-on experimentation.
$350k - $475k
Research Engineer, Infrastructure, Numerics
thinkingmachines
Thinking Machines Lab is seeking an infrastructure research engineer to design and build core systems for efficient large-scale model training, with a specific focus on numerics. This role involves enhancing the numerical foundations of their distributed training stack, optimizing precision formats, kernel optimizations, and communication frameworks to ensure stable, scalable, and fast training of trillion-parameter models. The ideal candidate will bridge research and systems engineering, possessing a strong understanding of both optimization mathematics and distributed compute realities.
$350k - $475k
Research Engineer, Infrastructure, Kernels
thinkingmachines
Thinking Machines Lab is seeking an infrastructure research engineer to design, optimize, and maintain the compute foundations for large-scale language model training. This role involves developing high-performance ML kernels, enabling efficient low-precision arithmetic, and improving the distributed compute stack. You will work closely with researchers and systems architects, bridging algorithmic design with hardware efficiency, prototyping new kernel implementations, and defining numerical and parallelism strategies for scaling AI systems.
$350k - $475k
Research Internship (Winter 2027)
Cohere
Cohere is seeking a Research Intern to collaborate with researchers and tools on designing and implementing novel research ideas and shipping state-of-the-art models to production. Interns will have the opportunity to work on various teams covering base model training, retrieval augmented generation, data and evaluation, safety, and finetuning, or any research area relating to LLMs. This role offers a chance to broaden research connections while gaining deep experience in a growing AI startup.
Applied Scientist / Research Engineer, AI4Engineering - EMEA
Mistral AI
Mistral AI is seeking an Applied Scientist with deep expertise in engineering sciences to work at the forefront of AI-accelerated simulation. This role involves collaborating with industrial customers and internal research teams to build and deploy AI Physics Models, complementing our existing Large Language Models (LLMs). You will be involved in the entire process, from curating high-fidelity simulation datasets and training/evaluating models to delivering production-grade AI solutions directly to engineering teams. The target domains include computational fluid dynamics, structural mechanics, semiconductor design, multi-physics modeling, and digital twins. Working cross-functionally, you will ensure our models meet stringent engineering standards beyond just benchmark metrics.
Clinician Scientist
Abridge
Abridge is seeking Clinician Scientists to advance the development of its AI-powered clinical tools. This role requires a blend of deep clinical expertise and a background in AI, focusing on shaping AI-driven tools to ensure accuracy and quality for clinicians and patients. You will collaborate closely with engineers, researchers, product managers, and fellow clinicians to refine AI models, validate outputs, and create new functionalities that enhance documentation workflows.
Member of Engineering (Post-training)
poolside
Poolside is building a company to create Artificial General Intelligence, aiming to accelerate software development through agentic systems, coding assistants, and frontier models. This role is part of the Applied Research team, focused on transforming pre-trained Large Language Models (LLMs) into well-aligned and highly capable AI systems specifically for coding and software development. You will be involved in building data pipelines and environments for agentic use cases, researching and implementing post-training algorithms, and designing experiments to test hypotheses, with access to significant GPU resources.
Machine Learning Engineer
Skild
Skild AI is seeking a Machine Learning Engineer to design and implement cutting-edge reinforcement learning algorithms for robotic applications. This role involves conducting experiments, optimizing models for real-world robotic environments, and collaborating with robotics, research, and engineering teams. The work will directly contribute to the development of intelligent, adaptable robots capable of autonomous learning and complex task performance.
Machine Learning Engineer, Reinforcement Learning
Skild
Skild AI is building the world's first general-purpose robotic intelligence that is robust and adapts to unseen scenarios without failing. We are seeking a Machine Learning Engineer to design and implement cutting-edge reinforcement learning algorithms for robotic applications. This role involves conducting experiments, optimizing models for real-world robotic environments, and collaborating closely with our robotics, research, and engineering teams. Your work will directly contribute to the development of intelligent, adaptable robots capable of autonomous learning and complex task performance.
Research Engineer, Post-training & Deployment
Skild
Skild AI is seeking a Research Engineer to join our post-training team. In this role, you will be responsible for enhancing Skild's foundation models and deploying them onto robots in real-world scenarios. You will collaborate with customers and utilize deployment data to ensure reliable robot behavior, focusing on safety, efficiency, and robustness under operational constraints. This position bridges the gap between strong lab performance and dependable customer deployments, defining the standard for scaling autonomous systems into global infrastructure.
Member of Technical Staff (AI Inference Engineer)
Perplexity AI
We are seeking an engineer to join our team responsible for building and running the inference engine behind Perplexity's queries. This role involves deploying dozens of model architectures at scale, managing tight latency and cost budgets, and working with a stack including Rust, Python, CUDA, and CuTe DSL. You will contribute to supporting new models, migrating GPU kernels, developing a Rust-native serving runtime, optimizing performance, and enhancing reliability and observability.
Member of Technical Staff (AI Inference Engineer)
Perplexity AI
We are seeking an AI Inference Engineer to join our dynamic team. This role is central to Perplexity's operations, as you will build and manage the inference engine that powers every query. You will deploy a variety of model architectures at scale, focusing on meeting stringent latency and cost requirements. Our technology stack includes Rust, Python, CUDA, and the CuTe DSL.
Internship - Search Machine Learning Engineer
Perplexity AI
Perplexity is seeking a Search Machine Learning Engineer Intern to contribute to the development of next-generation search technologies, specifically focusing on retrieval and ranking. This internship offers a hands-on opportunity to collaborate with experienced engineers, enhance search quality, experiment with novel models, and implement features that directly influence user search and information discovery experiences. The program is designed for a 12-24 week full-time engagement, conducted in person at our London office.
$12k - $24k
Senior Member of Technical Staff, Safety and Security for Agents
Cohere
Cohere is seeking a Senior Member of Technical Staff to join the Safety and Security for Agents team. In this role, you will significantly contribute to the development of safer, fairer, more trustworthy, and more secure Large Language Models (LLMs). Your work will focus on data generation, post-training algorithms, and evaluation methods to ensure the safety of next-generation models that interact with external resources and take actions. You will collaborate closely with machine learning teams, data annotation teams, and product and policy teams, requiring a blend of machine learning expertise, ethical AI principles, experimental design, and data management skills. This position offers significant autonomy and decision-making power within a small team, with the opportunity to shape the future of LLMs for societal benefit.
Research Scientist – Controlled 3D Generation
Stability AI
We are seeking a Research Scientist passionate about 3D generation, flow matching, and diffusion models. You will help advance the frontier of controllable 3D content creation by building models that generate consistent, editable, and physically grounded 3D assets and scenes. This role involves conducting cutting-edge research, designing and implementing scalable training pipelines, and developing techniques for conditioning and control. You will analyze model behavior, collaborate with cross-disciplinary teams to translate research into production-ready systems, and publish results at top-tier venues.
Research Scientist, Multimodal Alignment, Safety, and Fairness
Google DeepMind
Google DeepMind's Frontier AI unit is seeking experienced Research Scientists to join a multimodal safety research effort. This role focuses on interdisciplinary sociotechnical modeling and requires a passion for understanding AI-society interactions, a strong awareness of AI alignment and safety, and a drive to develop novel ideas, methods, interfaces, and tools. You will contribute to advancing the state of the art in AI research and Google DeepMind's mission towards Artificial General Intelligence (AGI), with a focus on leading new breakthrough research directions in areas like AI behavior exploration, assessment, and steering, particularly for subjective and creative tasks. The work involves tackling fundamental research questions to improve alignment objectives, assess adherence to desired behaviors, and enable AI agents to monitor real-world social context and evolve system behaviors over long time-horizons. You will develop new paradigms for human+AI rating that are adaptive and context-aware, driving breakthroughs within Google DeepMind, Google products, and the broader AI alignment community.
Research Scientist, Gemini Safety
Google DeepMind
The Gemini Safety team at Google DeepMind is responsible for the safety and fairness of the latest Gemini models. As a Research Scientist/Engineer, you will apply and develop cutting-edge data and algorithmic solutions to advance these user-facing models. This is a fast-paced, highly collaborative role within a supportive team dedicated to pushing the boundaries of AI for public benefit and scientific discovery, with safety and ethics as the highest priorities.
Research Engineer, Machine Learning
Mistral AI
Mistral AI is democratizing AI through high-performance, optimized, open-source models and solutions. We are a dynamic, collaborative team passionate about AI's potential to transform society, with a diverse workforce driving innovation. As a Research Engineer – ML track, you will build and optimize large-scale learning systems powering our open-weight models. You will work hand-in-hand with Research Scientists, either enhancing the shared training framework and data pipelines or embedding within a research squad to turn fresh ideas into scalable code.
Internship - Search Machine Learning Engineer
Perplexity AI
Perplexity is seeking a Search Machine Learning Engineer Intern to contribute to the development of next-generation search technologies, with a specific emphasis on retrieval and ranking. Interns will collaborate with seasoned engineers to enhance search quality, explore novel models, and implement features that directly influence user search and information discovery experiences. This internship program offers a duration of 12-24 weeks, operating on a full-time or part-time basis, and requires in-person attendance at our Belgrade office.
$12k - $24k
Member of Technical Staff, Data Analysis and Evaluation
Cohere
Cohere is seeking a Member of Technical Staff in Data Analysis and Evaluation to ensure the quality, reliability, and performance of our large language models (LLMs). This role involves designing and conducting data collection tasks, assessing dataset quality, and analyzing model robustness and generalisability. You will collaborate with researchers, engineers, and data annotators to drive data-driven decisions and enhance AI system effectiveness. The position requires expertise in statistics, experimental design, and machine learning to ensure high-quality data and reliable model performance across diverse scenarios, contributing to Cohere's mission of advancing AI.