JAX Jobs
54 open roles mentioning JAX
Engineering Manager, GPU Infrastructure
Cohere
Cohere is a leading enterprise AI company building cutting-edge foundation AI models and end-to-end products. The GPU Clusters team is central to Cohere's infrastructure, responsible for building and operating the superclusters that power our frontier AI models. This role involves enabling research and development at the intersection of cutting-edge hardware, distributed systems, and AI research. As an Engineering Manager, you will lead a team of highly motivated engineers passionate about GPU infrastructure and AI, fostering a culture of technical excellence and innovation in a remote-first environment. This is a unique opportunity to shape the infrastructure powering the next generation of AI.
Research Engineer, Infrastructure, Inference
thinkingmachines
Thinking Machines Lab is seeking an infrastructure research engineer to design, optimize, and scale the systems that power large AI models. The goal is to make inference faster, more cost-effective, more reliable, and more reproducible, enabling research teams to focus on advancing model capabilities. This role is crucial for ensuring that every experiment, evaluation, and deployment runs smoothly at scale, with a focus on performant and efficient model inference for both real-world applications and research acceleration.
$350k - $475k
Research Engineer, Interpretability
Anthropic
The Interpretability team at Anthropic is dedicated to understanding how large language models work, believing that a mechanistic understanding is key to making advanced AI systems safe and reliable. This role involves building and maintaining the specialized infrastructure for interpretability research, akin to performing 'neuroscience' on neural networks. The work spans the entire lifecycle of a production language model, from pretraining and inference to performance optimization, pushing the boundaries of hardware and software to address critical bottlenecks. As interpretability research matures and is applied to safety audits on frontier models, engineering and infrastructure have become crucial, making this role directly impactful on one of AI's most significant open problems.
Research Engineer, Machine Learning (Reinforcement Learning)
Anthropic
As a Research Engineer within Reinforcement Learning, you will collaborate with a diverse group of researchers and engineers to advance the capabilities and safety of large language models. This role blends research and engineering responsibilities, requiring you to both implement novel approaches and contribute to the research direction. You'll work on fundamental research in reinforcement learning, creating 'agentic' models via tool use for open-ended tasks such as computer use and autonomous software generation, improving reasoning abilities in areas such as mathematics, and developing prototypes for internal use, productivity, and evaluation.
Research Engineer, Machine Learning (Reinforcement Learning)
Anthropic
As a Research Engineer within Reinforcement Learning, you will collaborate with a diverse group of researchers and engineers to advance the capabilities and safety of large language models. This role blends research and engineering responsibilities, requiring you to both implement novel approaches and contribute to the research direction. You'll work on fundamental research in reinforcement learning, creating 'agentic' models via tool use for open-ended tasks such as computer use and autonomous software generation, improving reasoning abilities in areas such as mathematics, and developing prototypes for internal use, productivity, and evaluation.
Research Engineer, Performance RL (Reinforcement Learning)
Anthropic
About Anthropic Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems. About the RL Teams Our Reinforcement Learning teams lead Anthropic's reinforcement learning research and development, playing a critical role in advancing our AI systems. We've contributed to all Claude models, with significant impacts on the autonomy and coding capabilities of Claude Sonnet 4.6 and Opus 4.6. Our work spans several key areas: - Developing systems that enable models to use computers effectively - Advancing code generation through reinforcement learning - Pioneering fundamental RL research for large language models - Building scalable RL infrastructure and training methodologies - Enhancing model reasoning capabilities We collaborate closely with Anthropic's alignment and frontier red teams to ensure our systems are both capable and safe. We partner with the applied production training team to bring research innovations into deployed models, and are dedicated to implement our research at scale. Our Reinforcement Learning teams sit at the intersection of cutting-edge research and engineering excellence, with a deep commitment to building high-quality, scalable systems that push the boundaries of what AI can accomplish. About the Role We're hiring for the Code RL team within the RL organization. As a Research Engineer, you'll advance our models' ability to safely write correct, fast code for accelerators. You'll need to know accelerator performance well to turn it into tasks and signals models can learn from. Specifically, you will: - Invent, design and implement RL environments and evaluations. - Conduct experiments and shape our research roadmap. - Deliver your work into training runs. - Collaborate with other researchers, engineers, and performance engineering specialists across and outside Anthropic. You may be a good fit if you: - Have expertise with accelerators (CUDA, ROCm, Triton, Pallas), ML framework programming (JAX or PyTorch). - Have worked across the stack – kernels, model code, distributed systems. - Know how to balance research exploration with engineering implementation. - Are passionate about AI's potential and committed to developing safe and beneficial systems. Strong candidates may also have: - Experience with reinforcement learning. - Experience porting ML workloads between different types of accelerators. - Familiarity with LLM training methodologies. The annual compensation range for this role is listed below. For sales roles, the range provided is the role’s On Target Earnings ("OTE") range, meaning that the range includes both the sales commissions/sales bonuses target and annual base salary for the role. Annual Salary: $350,000—$850,000 USD Logistics Minimum education: Bachelor’s degree or an equivalent combination of education, training, and/or experience Required field of study: A field relevant to the role as demonstrated through coursework, training, or professional experience Minimum years of experience: Years of experience required will correlate with the internal job level requirements for the position Location-based hybrid policy: Currently, we expect all staff to be in one of our offices at least 25% of the time. However, some roles may require more time in our offices. Visa sponsorship: We do sponsor visas! However, we aren't able to successfully sponsor visas for every role and every candidate. But if we make you an offer, we will make every reasonable effort to get you a visa, and we retain an immigration lawyer to help with this. We encourage you to apply even if you do not believe you meet every single qualification. Not all strong candidates will meet every single qualification as listed. Research shows that people who identify as being from underrepresented groups are more prone to experiencing imposter syndrome and doubting the strength of their candidacy, so we urge you not to exclude yourself prematurely and to submit an application if you're interested in this work. We think AI systems like the ones we're building have enormous social and ethical implications. We think this makes representation even more important, and we strive to include a range of diverse perspectives on our team. Your safety matters to us. To protect yourself from potential scams, remember that Anthropic recruiters only contact you from @anthropic.com email addresses. In some cases, we may partner with vetted recruiting agencies who will identify themselves as working on behalf of Anthropic. Be cautious of emails from other domains. Legitimate Anthropic recruiters will never ask for money, fees, or banking information before your first day. If you're ever unsure about a communication, don't click any links—visit anthropic.com/careers directly for confirmed position openings. How we're different We believe that the highest-impact AI research will be big science. At Anthropic we work as a single cohesive team on just a few large-scale research efforts. And we value impact — advancing our long-term goals of steerable, trustworthy AI — rather than work on smaller and more specific puzzles. We view AI research as an empirical science, which has as much in common with physics and biology as with traditional efforts in computer science. We're an extremely collaborative group, and we host frequent research discussions to ensure that we are pursuing the highest-impact work at any given time. As such, we greatly value communication skills. The easiest way to understand our research directions is to read our recent research. This research continues many of the directions our team worked on prior to Anthropic, including: GPT-3, Circuit-Based Interpretability, Multimodal Neurons, Scaling Laws, AI & Compute, Concrete Problems in AI Safety, and Learning from Human Preferences. Come work with us! Anthropic is a public benefit corporation headquartered in San Francisco. We offer competitive compensation and benefits, optional equity donation matching, generous vacation and parental leave, flexible working hours, and a lovely office space in which to collaborate with colleagues. Guidance on Candidates' AI Usage: Learn about our policy for using AI in our application process.
Research Engineer, Pretraining Scaling
Anthropic
Anthropic's ML Performance and Scaling team is responsible for training the company's production pretrained models, a critical function that directly shapes the future of Anthropic and its mission to build safe, beneficial AI systems. As a Research Engineer on this team, you will ensure that frontier models train reliably, efficiently, and at scale. This role bridges the gap between research and engineering, involving work across the entire production training stack, including performance optimization, hardware debugging, experimental design, and launch coordination. During model launches, the team operates in close collaboration, addressing production issues that require immediate attention.
Research Engineer, Pretraining Scaling - London
Anthropic
Anthropic's ML Performance and Scaling team is responsible for training our production pretrained models, a critical function that directly shapes the company's future and its mission to build safe, beneficial AI systems. As a Research Engineer on this team, you will ensure our frontier models train reliably, efficiently, and at scale. This demanding, high-impact role requires deep technical expertise and a passion for large-scale ML systems, operating at the boundary between research and engineering. You will work across the entire production training stack, including performance optimization, hardware debugging, experimental design, and launch coordination, responding to critical production issues during launches.
Research Engineer, Discovery
Anthropic
As a Research Engineer on our team, you will work end-to-end across the entire model stack, identifying and addressing key infrastructure blockers on the path to scientific AGI. You should have familiarity with elements of language model training, evaluation, and inference, and be eager to quickly dive into and get up to speed in areas where you are not yet an expert. This may include performance optimization, distributed systems, VM/sandboxing/container deployment, and large-scale data pipelines. Join us in our mission to develop advanced AI systems that push the frontiers of science and benefit humanity.
Performance Engineer, GPU
Anthropic
Pioneering the next generation of AI requires breakthrough innovations in GPU performance and systems engineering. As a GPU Performance Engineer, you'll architect and implement the foundational systems that power Claude and push the frontiers of what's possible with large language models. You'll be responsible for maximizing GPU utilization and performance at unprecedented scale, developing cutting-edge optimizations that directly enable new model capabilities and dramatically improve inference efficiency. Working at the intersection of hardware and software, you'll implement state-of-the-art techniques from custom kernel development to distributed system architectures. Your work will span the entire stack—from low-level tensor core optimizations to orchestrating thousands of GPUs in perfect synchronization. Strong candidates will have a track record of delivering transformative GPU performance improvements in production ML systems and will be excited to shape the future of AI infrastructure alongside world-class researchers and engineers.
Software Engineer, Research Acceleration
thinkingmachines
Thinking Machines Lab is seeking engineers to build the libraries and tools that accelerate research. You will own internal infrastructure, including evaluation libraries, RL training libraries, and experiment tracking platforms, and build systems that compound research velocity over time. This is a collaborative role where you will work directly with researchers to identify bottlenecks and pain points. Success means researchers trust your systems to just work and find them a delight to use.
$350k - $475k
Research Infrastructure Engineer, Research Acceleration
thinkingmachines
Thinking Machines Lab is seeking engineers to build the libraries and tools that accelerate research. You will own internal infrastructure, including evaluation libraries, RL training libraries, and experiment tracking platforms, to build systems that compound research velocity over time. This is a collaborative role where you will work directly with researchers to identify bottlenecks and pain points. Success means researchers trust your systems to just work and find them a delight to use.
$350k - $475k
Applied Scientist / Research Engineer
Mistral AI
Mistral AI is seeking Applied Scientists and Research Engineers to drive innovative research and collaborate with clients on complex research projects. You will develop state-of-the-art models across different modalities such as text, image, and speech. By developing novel methods and research ideas, you will apply these models across a diverse set of use cases and domains. Working cross-functionally with both external and internal science, engineering, and product teams, you will deliver high-impact AI solutions that turn the needle.
Senior/Staff Applied Scientist/Research Engineer
Mistral AI
Mistral AI is seeking Applied Scientists and Research Engineers to drive innovative research and collaborate with clients on complex research projects. You will develop state-of-the-art models across different modalities such as text, image, and speech, applying novel methods and research ideas to diverse use cases and domains. Working cross-functionally with science, engineering, and product teams, you will deliver high-impact AI solutions. We are a dynamic, collaborative team passionate about AI and its potential to transform society, with teams distributed across France, USA, UK, Germany, and Singapore.
AI research scientist
Writer
AI research at WRITER focuses on building the scientific foundation for ambitious enterprise AI deployments. As a staff AI research scientist, you will drive a high-impact research agenda centered on large language models, agentic reasoning, and system-level capabilities essential for enterprise-scale AI. This role offers a unique opportunity to advance the field while directly contributing to products used by hundreds of thousands daily. You will work on post-training, planning, multi-step reasoning, and agentic workflows, directly shaping the future of enterprise AI performance and scalability. The role provides resources, infrastructure, and cross-functional support to pursue and implement ambitious ideas rapidly.
Research Intern, Model Shaping (Fall 2026)
Together AI
As a Research Intern in the Model Shaping team, you will work on advanced post-training methods, new techniques for efficient neural network training, and robust evaluation of foundation model capabilities. The Model Shaping team at Together AI focuses on tailoring open foundation models for downstream applications, building services for machine learning developers, and developing new methods for efficient model training and evaluation. This role offers the opportunity to contribute to cutting-edge research and potentially influence open-source projects.
Frontier Agents Intern (Fall 2026)
Together AI
The Agents team investigates how to build, align, and scale frontier AI systems capable of complex, multi-step tasks and workflows across text and speech, with a focus on agentic and scientific domains. This role sits at the intersection of agent capabilities, human-computer interaction, and infrastructure, exploring areas like post-training methods for agentic behavior and developing evaluation frameworks for open-ended tasks. As a research intern, you will tackle challenges in alignment, reliability, and scalability, potentially working on new training recipes for self-learning and long-horizon reasoning, curating datasets, studying failure modes, or building scalable agent infrastructure.
Research Engineer, Materials Science
Google DeepMind
Google DeepMind is seeking a Research Engineer to join their materials science team. This role involves accelerating the discovery of new functional materials by integrating artificial intelligence, computational simulation, and automated experimentation. You will collaborate with a diverse interdisciplinary team of domain experts, ML researchers, and engineers. The work focuses on pioneering research in various scientific domains, enabling the validation of early ideas and building infrastructure for promising research lines. You will contribute your scientific domain knowledge to the team's collective expertise.
$141k - $202k
AI engineer
Writer
WRITER is seeking an AI Engineer to join their team and shape how enterprises leverage AI. This role involves building tangible AI solutions that power the future of work for leading companies. You will directly impact the performance, scalability, and ethical alignment of cutting-edge LLMs and AI agents, enabling businesses to unlock unprecedented productivity and innovation with AI grounded in their data.
AI engineer (UK)
Writer
WRITER is a leading enterprise AI platform that empowers businesses to orchestrate AI-powered work and expand human capacity through superintelligence. We build powerful, trustworthy AI solutions that unite IT and business teams for enterprise-wide transformation. Our end-to-end platform enables companies to build and deploy AI agents grounded in their data and fueled by enterprise-grade LLMs. As an AI Engineer, you will be at the forefront of shaping how enterprises harness superintelligence by building tangible AI solutions that power the future of work. Your work will directly impact the performance, scalability, and ethical alignment of our cutting-edge LLMs and AI agents, enabling businesses to unlock unprecedented levels of productivity and innovation.