PyTorch Jobs
223 open roles mentioning PyTorch
AI Scientist - Audio
Mistral AI
Mistral is on a mission to democratize AI by producing frontier intelligence for everyone, developed in the open. We are a dynamic, collaborative team passionate about AI's potential to transform society, with diverse teams distributed globally. We develop models for enterprise and consumers, focusing on systems that change business operations and integrate into daily lives, while also releasing frontier models open-source. We are hiring experts in large language model training and distributed systems to shape the future of AI.
Member of Technical Staff, Agent Code
Cohere
Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company co-headquartered in Toronto and San Francisco, with key offices in London, New York City, Montreal, Seoul, Germany and Paris. Join us! Why this role? Code-generating LLMs and autonomous agents are revolutionizing how software is built and tasks are automated. At Cohere, we’re pushing the boundaries of what’s possible with these technologies for enterprises, and we’re looking for a senior member for the Agent Code team. You’ll be at the forefront of research and engineering, driving the development of cutting-edge code LLMs and agent systems that can interact with the digital world to solve complex tasks with minimal human oversight. This role is hands-on and research-driven. You’ll dive into the latest literature on code LLMs and agents, experiment with frontier models, and collaborate with a team of talented engineers and researchers to build scalable, production-ready solutions. At Cohere, we blend engineering and research seamlessly—everyone contributes to both, depending on their interests and organizational needs. We provide access to world-class compute resources, data, and talent to ensure you can do your best work. Note: We have offices in London, Toronto, New York and San Francisco, but we’re also remote-friendly! This team operates primarily between ET to CET time zones, so we’re seeking candidates in locations that align with these hours for effective collaboration. As a Member of Technical Staff on the Agent Code team, you will: - Stay up-to-date with the latest research in code LLMs, agents, and related fields, implementing novel ideas into our systems. - Design and implement scalable strategies to train code models, and deploy agent frameworks for inference and sampling. You will be collaborating with the pretraining team, create SFT trajectories and work on existing and new RL algorithms - Hillclimb on existing benchmarks and design new ones that reflect the needs of our enterprise users - Lead experiments on our state-of-the-art compute infrastructure, pushing the boundaries of what’s possible with frontier LLMs. You may be a good fit if you have: - A PhD in Computer Science, Machine Learning, or a related field, with publications in top-tier venues (e.g., NeurIPS, ICML, ICLR, ACL, EMNLP). - Deep expertise in code LLMs and agent systems, with a strong understanding of the latest research and trends. We are looking for people who not only have worked with code models, but have actively contributed to their development - Hands-on experience with frontier LLMs and their applications in code generation or automation. - Strong software engineering skills, with proficiency in Python and PyTorch, TensorFlow, or similar frameworks. - Experience with distributed systems, cloud infrastructure, and scalable architectures. - A proactive, self-motivated mindset, with a passion for solving ambitious, open-ended problems. What We Offer: - The opportunity to work on cutting-edge problems at the intersection of AI, code generation, and autonomous agents. - Access to world-class compute resources, data, and a collaborative team of researchers and engineers. - A remote-friendly, flexible work environment with a focus on impact and innovation. - Competitive compensation and benefits, including equity in a fast-growing AI company. If you’re passionate about shaping the future of code LLMs and agent systems, and thrive in a dynamic, research-driven environment, we’d love to hear from you! Full-Time Employees at Cohere enjoy these Perks: - A weekly lunch stipend of $75/£75 or equivalent in your local currency for lunch. - Full health and dental benefits, including a separate budget for mental health. - RRSP matching, 401K, Pension Scheme. - 100% Parental Leave top-up for up to 6 months, for either parent. - Annual enrichment benefits: Arts & culture, fitness/wellness, quality time, and a workspace improvement credit. Education & learning stipend for conferences, courses, and coaching. - 6 weeks of paid vacation (30 working days!) - Budget for traveling to other offices if you are remote, plus an annual company offsite. How and Where We Work: - Cohere is remote-friendly. We have offices in Toronto, San Francisco, New York City, London, Paris, Montreal, and more coming soon. - For those in the office: a daily lunch program, plenty of snacks, and regular community and social events. - For those not near an office: a co-working benefit so you can work alongside others in your city. - Everyone receives a $500 home office stipend to set up your workspace properly. If any of the above doesn’t line up exactly with your experience, we still encourage you to apply. We strive to create an inclusive work environment for all; we welcome applicants from all backgrounds and are committed to providing equal opportunities. Should you require any accommodations during the recruitment process, please submit an Accommodations Request Form, and we will work together to meet your needs. We may use AI-enabled tools to screen and assess applicants against the criteria for this position. This helps our recruiters identify potentially qualified candidates, but it doesn't limit the applications our recruiters may review or consider.
Machine Learning Infrastructure Engineer, Model Inference
Abridge
Abridge is seeking an ML Infrastructure Engineer, Model Inference to build and optimize the core inference infrastructure powering their machine learning models. This role is crucial for enhancing the scalability, efficiency, and performance of Abridge's AI-driven healthcare solutions. The engineer will collaborate with Infrastructure and Research teams to build, deploy, optimize, and orchestrate AI models, working on a platform that transforms patient-clinician conversations into structured clinical notes in real-time.
Member of Technical Staff, Integration/RL Team (Research Engineer)
Cohere
Cohere is a leading enterprise AI company focused on building cutting-edge foundation AI models and end-to-end products for real-world business problems. The integration team specifically focuses on developing and scaling machine learning algorithms and infrastructure for LLM post-training, with an emphasis on large-scale, distributed Reinforcement Learning (RL) methods. This role is crucial for enhancing the post-training codebase by implementing new research tools, optimizing algorithms, and scaling distributed RL capabilities. We are looking for passionate individuals who are meticulous in their approach to engineering and science, contributing to both production code and research efforts.
Software Engineer, Accelerators
OpenAI
On the Accelerators team, you will help OpenAI evaluate and bring up new compute platforms that can support large-scale AI training and inference. Your work will range from prototyping system software on new accelerators to enabling performance optimizations across our AI workloads. You’ll work across the stack, collaborating with both hardware and software aspects, focusing on kernels, sharding strategies, scaling across distributed systems, and performance modeling. You'll help adapt OpenAI's software stack to non-traditional hardware and drive efficiency improvements in core AI workloads, bridging ML algorithms with system performance, especially at scale.
Member of Technical Staff, Post-Training
Cohere
Cohere is seeking a Member of Technical Staff to focus on post-training of AI models. This role is crucial for advancing the state of the art in model post-training and shipping cutting-edge models to production, bridging the gap between research and practical application. You will have access to significant compute resources and a talented team to contribute to increasing model capabilities and driving customer value. The position offers a unique opportunity to contribute to both production code and research efforts, depending on your interests and organizational needs.
Applied AI, Forward Deployed Machine Learning Engineer- Singapore
Mistral AI
Mistral AI is seeking an Applied AI Engineer to join their team and facilitate the adoption of their products among customers. This role involves working closely with clients from the pre-sale stage through post-implementation, ensuring their solutions meet and exceed expectations. The engineer will manage daily customer relations, act as a key resource for externalizing research in production settings, and collaborate with researchers, AI engineers, and product engineers on complex customer projects. The goal is to drive the successful deployment of Mistral AI products and contribute to technological transformation across various industries.
Applied AI Engineer, Senior/Staff Devops/SRE
Mistral AI
Mistral AI is seeking an Applied AI Engineer focused on DevOps to help customers adopt its products and solve complex technical challenges. In this role, you will apply your problem-solving abilities, creativity, and technical skills to assist organizations in leveraging AI for significant impact. You will gain unique insights and contribute to critical global industries and institutions. Applied AI Engineers at Mistral AI work in small teams, owning end-to-end execution of high-stakes projects, which may involve discussing architecture, managing large-scale data, coding custom applications, engaging with customer executives, and strategizing for the Applied Engineering team.
Applied Scientist / Research Engineer - Singapore
Mistral AI
Mistral AI is seeking Applied Scientists and Research Engineers to drive innovative research and collaborate with clients on complex research projects. You will develop state-of-the-art models across different modalities such as text, image, and speech. By developing novel methods and research ideas, you will apply these models across a diverse set of use cases and domains. Working cross-functionally with both external and internal science, engineering, and product teams, you will deliver high-impact AI solutions that make a significant difference. We are a dynamic, collaborative team passionate about AI and its potential to transform society, with a diverse workforce thriving in competitive environments and committed to driving innovation.
Applied Machine Learning Engineer
fireworks ai
As an Applied Machine Learning Engineer, you will serve as a vital bridge between cutting-edge AI research and practical, real-world applications. Your work will focus on developing, fine-tuning, and operationalizing machine learning models that drive business value and enhance user experiences. This is a hands-on engineering role that combines deep technical expertise with a strong customer focus to deliver scalable AI solutions.
Applied Scientist / Research Engineer (Internship)
Mistral AI
Mistral AI is seeking Applied Scientists Interns and Research Engineers Interns to drive innovative research and collaborate with clients on complex research projects. You will develop state-of-the-art models across different modalities such as text, image, and speech. By developing novel methods and research ideas, you will apply these models across a diverse set of use cases and domains. Working cross-functionally with both external and internal science, engineering, and product teams, you will deliver high-impact AI solutions that turn the needle. This position is open for our local offices in Paris and London.
Member of Technical Staff, Performance Optimization
fireworks ai
Fireworks is seeking a Software Engineer focused on Performance Optimization to enhance the speed and efficiency of their AI infrastructure. This role involves optimizing performance across all levels of the technology stack, from low-level GPU kernels to large-scale distributed systems. The primary focus will be on maximizing the performance of demanding workloads such as large language models (LLMs), vision-language models (VLMs), and advanced video models. You will collaborate with research, infrastructure, and systems teams to identify and resolve performance bottlenecks, implement advanced optimizations, and scale AI systems for production use cases, directly influencing the speed, scalability, and cost-effectiveness of cutting-edge generative AI models.
Applied AI Engineer, ML Infrastructure Engineer / Devops - EMEA
Mistral AI
Mistral AI is seeking an Applied AI Engineer focused on DevOps to help customers adopt our AI products and solve complex technical challenges. In this role, you will apply your problem-solving abilities, creativity, and technical skills to assist organizations in leveraging AI for significant impact. You will gain unique insights and contribute to critical global industries and institutions. Your responsibilities will resemble those of a startup CTO, working in small teams to own the end-to-end execution of high-stakes projects, which may involve discussing architecture, managing large-scale data, coding applications, engaging with customer executives, and strategizing for the Applied Engineering team.
Senior/Staff Applied Scientist/Research Engineer, EMEA
Mistral AI
Mistral AI is seeking Applied Scientists and Research Engineers to drive innovative research and collaborate with clients on complex research projects. You will develop state-of-the-art models across different modalities such as text, image, and speech. By developing novel methods and research ideas, you will apply these models across a diverse set of use cases and domains. Working cross-functionally with both external and internal science, engineering, and product teams, you will deliver high-impact AI solutions that make a significant difference. We are a dynamic, collaborative team passionate about AI and its potential to transform society. Our diverse workforce thrives in competitive environments and is committed to driving innovation. Join us to be part of a pioneering company shaping the future of AI.
Senior Member of Technical Staff, Multimodal AI
Cohere
Cohere is seeking a Senior Member of Technical Staff focused on Multimodal AI to join their cutting-edge enterprise AI company. This role involves designing and developing advanced multimodal AI systems that integrate text, speech, and vision, pushing the boundaries of what's possible in AI. You will have access to exceptional compute resources and collaborate with world-class teams to innovate and shape the future of AI. The position is ideal for individuals passionate about machine learning and its real-world applications, who enjoy optimizing large models and thrive in a fast-paced, technically challenging environment.
Open-Source Software, Machine Learning Engineer
Mistral AI
Mistral AI is democratizing AI through high-performance, optimized, open-source models, products, and solutions. We are a dynamic, collaborative team passionate about AI's potential to transform society, with teams distributed globally. We are seeking an Open-Source Software, Machine Learning Engineer to join our OSS team, which is embedded within our Science team. This role is critical in helping turn research breakthroughs into tangible solutions and improving Mistral's open-source ecosystem by open-sourcing state-of-the-art models and maintaining our publicly available libraries.
Member of Technical Staff - Image / Video Generation
Black Forest Labs
We are seeking a Member of Technical Staff specializing in Image/Video Generation to join our team. You will be instrumental in training large-scale diffusion models for image and video generation, pushing the boundaries of generative AI. This role involves exploring novel approaches, rigorously testing design choices, and understanding the trade-offs between speed and quality in production settings. You will contribute to shaping our research direction by conducting experiments, analyzing results, and communicating findings to the team. This is an opportunity to work on foundational technologies used by millions of creators worldwide and contribute to the advancement of generative models.
€130k - €340k
Research Engineer, Machine Learning - Paris/London/Zurich/Warsaw
Mistral AI
Mistral AI is a pioneering company focused on democratizing AI through high-performance, optimized, open-source models and solutions. We aim to simplify tasks, save time, and enhance learning and creativity by integrating AI seamlessly into daily working life. Our comprehensive AI platform serves both enterprise and personal needs, featuring offerings like Le Chat, La Plateforme, Mistral Code, and Mistral Compute. We are a dynamic, collaborative, and diverse team passionate about AI's potential to transform society, driven by innovation and a low-ego, team-spirited culture.
Applied AI, Forward Deployed Machine Learning Engineer - EMEA
Mistral AI
Mistral AI is seeking an Applied AI Engineer to drive the adoption of its products and solve complex technical challenges for customers. The Applied AI team works directly with enterprise clients from pre-sales through implementation, deploying cutting-edge AI solutions to deliver measurable business impact. This role bridges the gap between AI research and real-world applications, ensuring solutions are robust, scalable, and aligned with customer needs and Mistral's vision. The team values output and impact over hours spent, fostering a direct, low-ego, and high-standards environment where the best ideas prevail.
AI Scientist - Palo Alto
Mistral AI
Mistral is on a mission to democratize AI by producing frontier intelligence for everyone, developed in the open. We are a dynamic, collaborative team passionate about AI and its potential to transform society, with diverse teams distributed across Europe, the USA, and Asia. We develop models for enterprise and consumers, focusing on systems that change business operations and integrate into daily lives, while also releasing frontier models open-source. We are hiring experts in the training of large language models and distributed systems to join our pioneering company shaping the future of AI.