JAX Jobs
54 open roles mentioning JAX
Machine Learning Research Scientist, Reasoning
Scale AI
This role operates at the forefront of AI research and real-world implementation, with a strong focus on reasoning within large language models (LLMs). You will play a key role in shaping Scale’s data strategy by identifying the most effective data sources and methodologies for improving LLM reasoning. Success in this role requires a deep understanding of LLMs, planning algorithms, and novel approaches to agentic reasoning, as well as creativity in tackling challenges related to data generation, model interaction, and evaluation. You will contribute to impactful research on language model reasoning, collaborate with external researchers, and work closely with engineering teams to bring state-of-the-art advancements into scalable, real-world solutions.
$252k - $315k
Senior / Staff Machine Learning Research Scientist, Agents
Scale AI
This role is at the intersection of cutting-edge AI research and practical application, with a focus on studying the data types essential for building state-of-the-art agents, such as browser and SWE agents. The ideal candidate will explore the data landscape needed to advance intelligent, adaptable AI agents, guiding the data strategy at Scale to drive innovation. This position requires not only expertise in LLM agents and planning algorithms but also creativity in addressing novel challenges related to data, interaction, and evaluation. You will contribute to impactful research publications on agents, collaborate with customer researchers, and work alongside the engineering team to translate these advancements into real-world, scalable solutions.
$302k - $378k
Research Engineer, Multimodal Reasoning For Information Literacy
Google DeepMind
Google DeepMind's research team is dedicated to tackling complex challenges in online information quality, advancing the state of the art by developing innovative solutions to detect manipulated media and misleading narratives. The team leverages interdisciplinary work spanning provenance analysis and the creation of tools for AI-assisted information literacy, with a focus on ensuring the integrity of digital discourse and a safer online environment. This role involves researching and building multimodal reasoning systems and Vision-Language Models (VLMs) to assess the trustworthiness of media on the internet, with a passion for advancing information literacy using machine learning and computational techniques.
Research, Pre-Training Data
thinkingmachines
Thinking Machines Lab is seeking pre-training researchers to join their mission of advancing collaborative general intelligence. This role is central to developing the next generation of AI models by blending research with large-scale data engineering. You will be responsible for assembling pre-training datasets and data systems, designing and implementing methods for sourcing, curating, and analyzing data for quality and performance. The position involves working with automated pipelines and human-in-the-loop processes, contributing both scientific insights and production-grade code. It's an ideal opportunity for individuals passionate about the intersection of data, machine learning, and systems, and who are eager to shape the future of AI.
$350k - $475k
Research, Post-Training Data
thinkingmachines
Thinking Machines Lab is seeking researchers to bridge the gap between raw AI intelligence and useful, safe, and collaborative systems. This role focuses on post-training data research, combining human insight and machine learning techniques to capture and steer model behavior based on human preferences. You will be responsible for translating research ideas into actionable data through labeling and collection campaigns, understanding data quality science, and developing metrics to measure the impact of data and training interventions. The position also involves exploring new paradigms for human-AI interaction and scalable oversight, blending research, data operations, and technical implementation to advance human-centered AI systems. This role requires both fundamental research and practical engineering, making it ideal for individuals who enjoy deep theoretical exploration and hands-on experimentation.
$350k - $475k
Research Engineer, Infrastructure, Numerics
thinkingmachines
Thinking Machines Lab is seeking an infrastructure research engineer to design and build core systems for efficient large-scale model training, with a specific focus on numerics. This role involves enhancing the numerical foundations of their distributed training stack, optimizing precision formats, kernel optimizations, and communication frameworks to ensure stable, scalable, and fast training of trillion-parameter models. The ideal candidate will bridge research and systems engineering, possessing a strong understanding of both optimization mathematics and distributed compute realities.
$350k - $475k
Research Engineer, Infrastructure, Kernels
thinkingmachines
Thinking Machines Lab is seeking an infrastructure research engineer to design, optimize, and maintain the compute foundations for large-scale language model training. This role involves developing high-performance ML kernels, enabling efficient low-precision arithmetic, and improving the distributed compute stack. You will work closely with researchers and systems architects, bridging algorithmic design with hardware efficiency, prototyping new kernel implementations, and defining numerical and parallelism strategies for scaling AI systems.
$350k - $475k
Research Internship (Fall, Winter 2026)
Cohere
Cohere is seeking a Research Intern to collaborate with researchers and tools on designing and implementing novel research ideas and shipping state-of-the-art models to production. Interns will have the opportunity to work on various teams covering base model training, retrieval augmented generation, data and evaluation, safety, and finetuning, or any research area relating to LLMs. This role offers a chance to broaden research connections while gaining deep experience in a growing AI startup.
Applied Scientist / Research Engineer, AI4Engineering - EMEA
Mistral AI
Mistral AI is seeking an Applied Scientist with deep expertise in engineering sciences to work at the forefront of AI-accelerated simulation. This role involves collaborating with industrial customers and internal research teams to build and deploy AI Physics Models, complementing our existing Large Language Models (LLMs). You will be involved in the entire process, from curating high-fidelity simulation datasets and training/evaluating models to delivering production-grade AI solutions directly to engineering teams. The target domains include computational fluid dynamics, structural mechanics, semiconductor design, multi-physics modeling, and digital twins. Working cross-functionally, you will ensure our models meet stringent engineering standards beyond just benchmark metrics.
Clinician Scientist
Abridge
Abridge is seeking Clinician Scientists to advance the development of its AI-powered clinical tools. This role requires a blend of deep clinical expertise and a background in AI, focusing on shaping AI-driven tools to ensure accuracy and quality for clinicians and patients. You will collaborate closely with engineers, researchers, product managers, and fellow clinicians to refine AI models, validate outputs, and create new functionalities that enhance documentation workflows.
Senior Member of Technical Staff, Safety and Security for Agents
Cohere
Cohere is seeking a Senior Member of Technical Staff to join the Safety and Security for Agents team. In this role, you will significantly contribute to the development of safer, fairer, more trustworthy, and more secure Large Language Models (LLMs). Your work will focus on data generation, post-training algorithms, and evaluation methods to ensure the safety of next-generation models that interact with external resources and take actions. You will collaborate closely with machine learning teams, data annotation teams, and product and policy teams, requiring a blend of machine learning expertise, ethical AI principles, experimental design, and data management skills. This position offers significant autonomy and decision-making power within a small team, with the opportunity to shape the future of LLMs for societal benefit.
Research Scientist – Controlled 3D Generation
Stability AI
We are seeking a Research Scientist passionate about 3D generation, flow matching, and diffusion models. You will help advance the frontier of controllable 3D content creation by building models that generate consistent, editable, and physically grounded 3D assets and scenes. This role involves conducting cutting-edge research, designing and implementing scalable training pipelines, and developing techniques for conditioning and control. You will analyze model behavior, collaborate with cross-disciplinary teams to translate research into production-ready systems, and publish results at top-tier venues.
Research Scientist, Multimodal Alignment, Safety, and Fairness
Google DeepMind
Google DeepMind's Frontier AI unit is seeking experienced Research Scientists to join a multimodal safety research effort. This role focuses on interdisciplinary sociotechnical modeling and requires a passion for understanding AI-society interactions, a strong awareness of AI alignment and safety, and a drive to develop novel ideas, methods, interfaces, and tools. You will contribute to advancing the state of the art in AI research and Google DeepMind's mission towards Artificial General Intelligence (AGI), with a focus on leading new breakthrough research directions in areas like AI behavior exploration, assessment, and steering, particularly for subjective and creative tasks. The work involves tackling fundamental research questions to improve alignment objectives, assess adherence to desired behaviors, and enable AI agents to monitor real-world social context and evolve system behaviors over long time-horizons. You will develop new paradigms for human+AI rating that are adaptive and context-aware, driving breakthroughs within Google DeepMind, Google products, and the broader AI alignment community.
Research Scientist, Gemini Safety
Google DeepMind
The Gemini Safety team at Google DeepMind is responsible for the safety and fairness of the latest Gemini models. As a Research Scientist/Engineer, you will apply and develop cutting-edge data and algorithmic solutions to advance these user-facing models. This is a fast-paced, highly collaborative role within a supportive team dedicated to pushing the boundaries of AI for public benefit and scientific discovery, with safety and ethics as the highest priorities.
Research Engineer, Machine Learning
Mistral AI
Mistral AI is democratizing AI through high-performance, optimized, open-source models and solutions. We are a dynamic, collaborative team passionate about AI's potential to transform society, with a diverse workforce driving innovation. As a Research Engineer – ML track, you will build and optimize large-scale learning systems powering our open-weight models. You will work hand-in-hand with Research Scientists, either enhancing the shared training framework and data pipelines or embedding within a research squad to turn fresh ideas into scalable code.
Software Engineer, GPU Infrastructure (HPC)
Cohere
Cohere is seeking a Staff Software Engineer to join our internal infrastructure team, responsible for building and operating world-class infrastructure and tools for training, evaluating, and serving Cohere's foundational AI models. You will work closely with AI researchers to support their AI workload needs on cutting-edge systems, focusing on stability, scalability, and observability. This role involves building and operating superclusters across multiple clouds, directly accelerating the development of industry-leading AI models. Participation in a 24x7 on-call rotation is required and compensated.
Member of Technical Staff, Data Analysis and Evaluation
Cohere
Cohere is seeking a Member of Technical Staff in Data Analysis and Evaluation to ensure the quality, reliability, and performance of our large language models (LLMs). This role involves designing and conducting data collection tasks, assessing dataset quality, and analyzing model robustness and generalisability. You will collaborate with researchers, engineers, and data annotators to drive data-driven decisions and enhance AI system effectiveness. The position requires expertise in statistics, experimental design, and machine learning to ensure high-quality data and reliable model performance across diverse scenarios, contributing to Cohere's mission of advancing AI.
Senior ML Systems Engineer, Frameworks & Tooling
Cohere
Cohere is seeking a Senior ML Systems Engineer to join their team and build, maintain, and evolve the training framework that powers their frontier-scale language models. This role is ideal for someone passionate about large-scale training, distributed systems, and HPC infrastructure, offering the opportunity to design and maintain core components for fast, reliable, and scalable model training. You will also build tooling to connect research ideas to thousands of GPUs, working across the full stack of ML systems with significant autonomy and impact.
AI Scientist - Warsaw
Mistral AI
Mistral AI is a pioneering company dedicated to democratizing AI through high-performance, optimized, open-source models, products, and solutions. We aim to simplify tasks, save time, and enhance learning and creativity by integrating cutting-edge AI into daily working life. Our comprehensive platform serves both enterprise and personal needs, featuring offerings like Le Chat, La Plateforme, Mistral Code, and Mistral Compute. We are a dynamic, collaborative, and diverse team passionate about AI's potential to transform society, driving innovation from our distributed teams across France, USA, UK, Germany, and Singapore. Join us to shape the future of AI and make a meaningful impact.
AI Scientist - Zurich
Mistral AI
Mistral AI is a pioneering company dedicated to democratizing AI through high-performance, optimized, open-source, and cutting-edge models, products, and solutions. Our comprehensive AI platform serves both enterprise and personal needs, offering tools like Le Chat, La Plateforme, Mistral Code, and Mistral Compute to bring frontier intelligence to end-users. We are a dynamic, collaborative, and diverse team passionate about AI's potential to transform society, thriving in competitive environments and committed to driving innovation. Join us to be part of shaping the future of AI and making a meaningful impact.