PyTorch Jobs
109 open roles mentioning PyTorch
Software Engineer, GPU Infrastructure (HPC)
Cohere
Cohere is seeking a Staff Software Engineer to join our internal infrastructure team, responsible for building and operating world-class infrastructure and tools for training, evaluating, and serving Cohere's foundational AI models. You will work closely with AI researchers to support their AI workload needs on cutting-edge systems, focusing on stability, scalability, and observability. This role involves building and operating superclusters across multiple clouds, directly accelerating the development of industry-leading AI models. Participation in a 24x7 on-call rotation is required and compensated.
Member of Technical Staff, Data Analysis and Evaluation
Cohere
Cohere is seeking a Member of Technical Staff in Data Analysis and Evaluation to ensure the quality, reliability, and performance of our large language models (LLMs). This role involves designing and conducting data collection tasks, assessing dataset quality, and analyzing model robustness and generalisability. You will collaborate with researchers, engineers, and data annotators to drive data-driven decisions and enhance AI system effectiveness. The position requires expertise in statistics, experimental design, and machine learning to ensure high-quality data and reliable model performance across diverse scenarios, contributing to Cohere's mission of advancing AI.
Senior ML Systems Engineer, Frameworks & Tooling
Cohere
Cohere is seeking a Senior ML Systems Engineer to join their team and build, maintain, and evolve the training framework that powers their frontier-scale language models. This role is ideal for someone passionate about large-scale training, distributed systems, and HPC infrastructure, offering the opportunity to design and maintain core components for fast, reliable, and scalable model training. You will also build tooling to connect research ideas to thousands of GPUs, working across the full stack of ML systems with significant autonomy and impact.
Applied AI, Technical Lead - Forward Deployed AI Engineer
Mistral AI
Mistral AI is seeking a Technical Lead, Applied AI to drive the technical strategy, execution, and delivery of complex AI solutions for enterprise customers. In this role, you will lead project teams of Applied AI Engineers, ensuring the successful deployment of Mistral AI products and the development of high-impact, scalable AI use cases. You will act as the primary technical point of contact for strategic customers, guiding them through the entire lifecycle from pre-sales to post-implementation, while collaborating closely with research, product, and engineering teams to shape the future of our offerings. As a Technical Lead, you will bridge the gap between cutting-edge AI research and real-world enterprise applications, ensuring our solutions are robust, scalable, and aligned with both customer needs and Mistral’s technological vision.
Audio Inference Engineer, Model Efficiency
Cohere
Cohere is seeking an Audio Inference Engineer focused on Model Efficiency to join a fast-growing team of researchers and engineers. The mission of this team is to build reliable machine learning systems and optimize audio inference serving efficiency using innovative techniques. As an engineer on this team, you will advance core audio model serving metrics, including latency, throughput, and quality by diving deep into systems, identifying bottlenecks, and delivering creative solutions for audio processing and streaming workloads. You will collaborate closely with both the training and serving infrastructure teams to ensure seamless integration between model development and deployment, with a special focus on real-time and streaming audio inference.
Applied AI, Forward Deployed Machine Learning Engineer - Morocco
Mistral AI
Mistral AI is seeking an Applied AI Engineer to join its customer-facing technical organization. This role involves facilitating customer adoption of Mistral AI products, addressing complex technical challenges, and deploying cutting-edge AI solutions that deliver measurable business impact. You will bridge the gap between AI research and real-world enterprise applications, ensuring solutions are robust, scalable, and aligned with customer needs and Mistral's technological vision. The team operates with a focus on outputs, direct communication, and a low-ego, high-standards environment, embracing an unstructured setting.
AI Scientist - Robotics
Mistral AI
Mistral is on a mission to democratize AI by producing frontier intelligence for everyone, developed in the open and built by a global team. We are a dynamic, collaborative group passionate about AI's potential to transform society, with a diverse workforce distributed across Europe, the USA, and Asia. We focus on developing models for enterprise and consumers that can change business operations and integrate into daily lives, while also releasing frontier models open-source. We are hiring experts in large language model training and distributed systems to shape the future of AI.
Applied AI, Technical Lead, Forward Deployed AI Engineer - EMEA
Mistral AI
Mistral AI is seeking a Technical Lead, Applied AI to drive the technical strategy, execution, and delivery of complex AI solutions for enterprise customers. This role involves leading project teams of Applied AI Engineers to ensure successful deployment of Mistral AI products and the development of high-impact, scalable AI use cases. You will serve as the primary technical point of contact for strategic customers, guiding them through the entire lifecycle from pre-sales to post-implementation, and collaborating with research, product, and engineering teams to shape future offerings. The position bridges cutting-edge AI research with real-world enterprise applications, ensuring solutions are robust, scalable, and aligned with customer needs and Mistral's vision.
AI Scientist - Warsaw
Mistral AI
Mistral AI is a pioneering company dedicated to democratizing AI through high-performance, optimized, open-source models, products, and solutions. We aim to simplify tasks, save time, and enhance learning and creativity by integrating cutting-edge AI into daily working life. Our comprehensive platform serves both enterprise and personal needs, featuring offerings like Le Chat, La Plateforme, Mistral Code, and Mistral Compute. We are a dynamic, collaborative, and diverse team passionate about AI's potential to transform society, driving innovation from our distributed teams across France, USA, UK, Germany, and Singapore. Join us to shape the future of AI and make a meaningful impact.
AI Scientist - Zurich
Mistral AI
Mistral AI is a pioneering company dedicated to democratizing AI through high-performance, optimized, open-source, and cutting-edge models, products, and solutions. Our comprehensive AI platform serves both enterprise and personal needs, offering tools like Le Chat, La Plateforme, Mistral Code, and Mistral Compute to bring frontier intelligence to end-users. We are a dynamic, collaborative, and diverse team passionate about AI's potential to transform society, thriving in competitive environments and committed to driving innovation. Join us to be part of shaping the future of AI and making a meaningful impact.
Member of Technical Staff - Pre-Training
Reflection ai
Reflection is a research lab dedicated to making intelligence open and accessible for everyone. We build open models that empower individuals to control their intelligence and shape the future of AI. As a Member of Technical Staff - Pre-Training, you will be instrumental in researching and building solutions across algorithms, scaling laws, data processing, optimizers, and model architecture. This role involves designing and executing scientific experiments to deepen our understanding of scaling large language models and data efficiency, while also implementing state-of-the-art deep learning methods. You will have the opportunity to lead small research projects independently and contribute to larger initiatives, optimizing training infrastructure for efficient scaling and working across the entire stack from low-level optimizations to high-level model design.
AI Scientist - Audio
Mistral AI
Mistral is on a mission to democratize AI by producing frontier intelligence for everyone, developed in the open. We are a dynamic, collaborative team passionate about AI's potential to transform society, with diverse teams distributed globally. We develop models for enterprise and consumers, focusing on systems that change business operations and integrate into daily lives, while also releasing frontier models open-source. We are hiring experts in large language model training and distributed systems to shape the future of AI.
Member of Technical Staff, Agent Code
Cohere
Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company co-headquartered in Toronto and San Francisco, with key offices in London, New York City, Montreal, Seoul, Germany and Paris. Join us! Why this role? Code-generating LLMs and autonomous agents are revolutionizing how software is built and tasks are automated. At Cohere, we’re pushing the boundaries of what’s possible with these technologies for enterprises, and we’re looking for a senior member for the Agent Code team. You’ll be at the forefront of research and engineering, driving the development of cutting-edge code LLMs and agent systems that can interact with the digital world to solve complex tasks with minimal human oversight. This role is hands-on and research-driven. You’ll dive into the latest literature on code LLMs and agents, experiment with frontier models, and collaborate with a team of talented engineers and researchers to build scalable, production-ready solutions. At Cohere, we blend engineering and research seamlessly—everyone contributes to both, depending on their interests and organizational needs. We provide access to world-class compute resources, data, and talent to ensure you can do your best work. Note: We have offices in London, Toronto, New York and San Francisco, but we’re also remote-friendly! This team operates primarily between ET to CET time zones, so we’re seeking candidates in locations that align with these hours for effective collaboration. As a Member of Technical Staff on the Agent Code team, you will: - Stay up-to-date with the latest research in code LLMs, agents, and related fields, implementing novel ideas into our systems. - Design and implement scalable strategies to train code models, and deploy agent frameworks for inference and sampling. You will be collaborating with the pretraining team, create SFT trajectories and work on existing and new RL algorithms - Hillclimb on existing benchmarks and design new ones that reflect the needs of our enterprise users - Lead experiments on our state-of-the-art compute infrastructure, pushing the boundaries of what’s possible with frontier LLMs. You may be a good fit if you have: - A PhD in Computer Science, Machine Learning, or a related field, with publications in top-tier venues (e.g., NeurIPS, ICML, ICLR, ACL, EMNLP). - Deep expertise in code LLMs and agent systems, with a strong understanding of the latest research and trends. We are looking for people who not only have worked with code models, but have actively contributed to their development - Hands-on experience with frontier LLMs and their applications in code generation or automation. - Strong software engineering skills, with proficiency in Python and PyTorch, TensorFlow, or similar frameworks. - Experience with distributed systems, cloud infrastructure, and scalable architectures. - A proactive, self-motivated mindset, with a passion for solving ambitious, open-ended problems. What We Offer: - The opportunity to work on cutting-edge problems at the intersection of AI, code generation, and autonomous agents. - Access to world-class compute resources, data, and a collaborative team of researchers and engineers. - A remote-friendly, flexible work environment with a focus on impact and innovation. - Competitive compensation and benefits, including equity in a fast-growing AI company. If you’re passionate about shaping the future of code LLMs and agent systems, and thrive in a dynamic, research-driven environment, we’d love to hear from you! Full-Time Employees at Cohere enjoy these Perks: - A weekly lunch stipend of $75/£75 or equivalent in your local currency for lunch. - Full health and dental benefits, including a separate budget for mental health. - RRSP matching, 401K, Pension Scheme. - 100% Parental Leave top-up for up to 6 months, for either parent. - Annual enrichment benefits: Arts & culture, fitness/wellness, quality time, and a workspace improvement credit. Education & learning stipend for conferences, courses, and coaching. - 6 weeks of paid vacation (30 working days!) - Budget for traveling to other offices if you are remote, plus an annual company offsite. How and Where We Work: - Cohere is remote-friendly. We have offices in Toronto, San Francisco, New York City, London, Paris, Montreal, and more coming soon. - For those in the office: a daily lunch program, plenty of snacks, and regular community and social events. - For those not near an office: a co-working benefit so you can work alongside others in your city. - Everyone receives a $500 home office stipend to set up your workspace properly. If any of the above doesn’t line up exactly with your experience, we still encourage you to apply. We strive to create an inclusive work environment for all; we welcome applicants from all backgrounds and are committed to providing equal opportunities. Should you require any accommodations during the recruitment process, please submit an Accommodations Request Form, and we will work together to meet your needs. We may use AI-enabled tools to screen and assess applicants against the criteria for this position. This helps our recruiters identify potentially qualified candidates, but it doesn't limit the applications our recruiters may review or consider.
Machine Learning Infrastructure Engineer, Model Inference
Abridge
Abridge is seeking an ML Infrastructure Engineer, Model Inference to build and optimize the core inference infrastructure powering their machine learning models. This role is crucial for enhancing the scalability, efficiency, and performance of Abridge's AI-driven healthcare solutions. The engineer will collaborate with Infrastructure and Research teams to build, deploy, optimize, and orchestrate AI models, working on a platform that transforms patient-clinician conversations into structured clinical notes in real-time.
Member of Technical Staff, Integration/RL Team (Research Engineer)
Cohere
Cohere is a leading enterprise AI company focused on building cutting-edge foundation AI models and end-to-end products for real-world business problems. The integration team specifically focuses on developing and scaling machine learning algorithms and infrastructure for LLM post-training, with an emphasis on large-scale, distributed Reinforcement Learning (RL) methods. This role is crucial for enhancing the post-training codebase by implementing new research tools, optimizing algorithms, and scaling distributed RL capabilities. We are looking for passionate individuals who are meticulous in their approach to engineering and science, contributing to both production code and research efforts.
Member of Technical Staff, Post-Training
Cohere
Cohere is seeking a Member of Technical Staff to focus on post-training of AI models. This role is crucial for advancing the state of the art in model post-training and shipping cutting-edge models to production, bridging the gap between research and practical application. You will have access to significant compute resources and a talented team to contribute to increasing model capabilities and driving customer value. The position offers a unique opportunity to contribute to both production code and research efforts, depending on your interests and organizational needs.
Member of Technical Staff, Search
Cohere
Cohere is seeking talented individuals to join our Search team as a Member of Technical Staff. This role focuses on developing state-of-the-art models for information retrieval, including training embedding and reranker models. You will have the opportunity to revolutionize search experiences by building intelligent, efficient, and precise search systems. Your work will advance semantic search techniques, improve accuracy and efficiency, and involve integrating novel technologies into our search infrastructure. We are looking for someone passionate about search with a strong background in information retrieval and experience working with diverse technologies and cross-functional teams.
Applied AI, Forward Deployed Machine Learning Engineer- Singapore
Mistral AI
Mistral AI is seeking an Applied AI Engineer to join their team and facilitate the adoption of their products among customers. This role involves working closely with clients from the pre-sale stage through post-implementation, ensuring their solutions meet and exceed expectations. The engineer will manage daily customer relations, act as a key resource for externalizing research in production settings, and collaborate with researchers, AI engineers, and product engineers on complex customer projects. The goal is to drive the successful deployment of Mistral AI products and contribute to technological transformation across various industries.
Applied AI Engineer, Senior/Staff Devops/SRE
Mistral AI
Mistral AI is seeking an Applied AI Engineer focused on DevOps to help customers adopt its products and solve complex technical challenges. In this role, you will apply your problem-solving abilities, creativity, and technical skills to assist organizations in leveraging AI for significant impact. You will gain unique insights and contribute to critical global industries and institutions. Applied AI Engineers at Mistral AI work in small teams, owning end-to-end execution of high-stakes projects, which may involve discussing architecture, managing large-scale data, coding custom applications, engaging with customer executives, and strategizing for the Applied Engineering team.
Applied Scientist / Research Engineer - Singapore
Mistral AI
Mistral AI is seeking Applied Scientists and Research Engineers to drive innovative research and collaborate with clients on complex research projects. You will develop state-of-the-art models across different modalities such as text, image, and speech. By developing novel methods and research ideas, you will apply these models across a diverse set of use cases and domains. Working cross-functionally with both external and internal science, engineering, and product teams, you will deliver high-impact AI solutions that make a significant difference. We are a dynamic, collaborative team passionate about AI and its potential to transform society, with a diverse workforce thriving in competitive environments and committed to driving innovation.