Cohere Jobs

77 open roles mentioning Cohere

Forward Deployed Engineer, Agentic Platform (Singapore)

8mo ago
Cohere

Cohere

Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company co-headquartered in Toronto and San Francisco, with key offices in London, New York City, Montreal, Seoul, Germany and Paris. Join us! About North: North is Cohere's cutting-edge AI workspace platform, designed to revolutionize the way enterprises utilize AI. It offers a secure and customizable environment, allowing companies to deploy AI while maintaining control over sensitive data. North integrates seamlessly with existing workflows, providing a trusted platform that connects AI agents with workplace tools and applications. Why this role? This role offers a unique opportunity to shape how enterprises harness the power of AI in real-world applications. As a bridge between our core North product and our clients’ engineering teams, you’ll be at the forefront of solving complex problems and securely integrating AI into critical sectors such as finance, healthcare, and telecommunications. We’re looking for Software Engineers with Applied AI experience who can own the design, build, and deployment of agentic workflows powered by Large Language Models (LLMs), from early prototypes to production-grade AI agents, to deliver concrete business value in enterprise workflows. You’ll work closely with customers on real-world business problems, often building first-of-their-kind agent workflows that integrate LLMs with tools, APIs, and data sources. While our pace is startup-fast, the bar is enterprise-high: agents must be reliable, observable, safe, and auditable from day one. Note: between 20 - 40% travel anticipated. In this role, you will: - Work closely with our enterprise customers to translate high-value, ambiguous business problems into well-framed agentic workflows with clear success criteria and evaluation methodologies - Lead the design, build, and delivery of LLM-powered agents that reason, plan, and act across tools, APIs, and sensitive enterprise data sources, with enterprise-grade reliability and performance - Build and ship features for North, our AI workspace platform, working across the full product lifecycle from conceptualisation through production - Take ownership of scoping and shaping use cases end-to-end, flexing into whatever technical area the problem demands (including frontend) to drive the most effective solution - Contribute to shared frameworks and patterns that enable consistent, high-quality delivery across customers and teams - Drive clarity in ambiguous situations, build alignment, and raise engineering quality across the organization - Travel up to 20–40% to work on-site with customers and partners You may be a good fit if: - You have hands-on experience building and deploying production-grade software in Python; you write clean, testable, observable, scalable code - You've built and deployed highly performant RAG and agentic applications, including agents that plan and execute multi-step tasks using patterns like ReAct or Plan-and-Execute - You're deeply familiar with the LLM stack: frontier models, vector databases, and orchestration frameworks - You have a proven ability to build robust evaluation frameworks, moving well beyond trial and error, to measure agent accuracy, safety, and latency - You’re experienced working directly with customers and can lead technical discussions with enterprise stakeholders, translating ambiguous business needs into concrete technical specs - You have experience owning the full scope of a use case end-to-end - You thrive in fast-paced and ambiguous environments and can execute well even when priorities are shifting It's a bonus if you have: - Experience setting architectural standards for AI and agentic systems across distributed teams - Experience flexing into unfamiliar technical areas, such as frontend, when the problem calls for it - Exposure to regulated or sensitive industry environments (finance, healthcare, telecoms) - Experience with enterprise security, compliance, or auditability requirements for AI systems Full-Time Employees at Cohere enjoy these Perks: - A weekly lunch stipend of $75/£75 or equivalent in your local currency for lunch. - Full health and dental benefits, including a separate budget for mental health. - RRSP matching, 401K, Pension Scheme. - 100% Parental Leave top-up for up to 6 months, for either parent. - Annual enrichment benefits: Arts & culture, fitness/wellness, quality time, and a workspace improvement credit. Education & learning stipend for conferences, courses, and coaching. - 6 weeks of paid vacation (30 working days!) - Budget for traveling to other offices if you are remote, plus an annual company offsite. How and Where We Work: - Cohere is remote-friendly. We have offices in Toronto, San Francisco, New York City, London, Paris, Montreal, and more coming soon. - For those in the office: a daily lunch program, plenty of snacks, and regular community and social events. - For those not near an office: a co-working benefit so you can work alongside others in your city. - Everyone receives a $500 home office stipend to set up your workspace properly. If any of the above doesn’t line up exactly with your experience, we still encourage you to apply. We strive to create an inclusive work environment for all; we welcome applicants from all backgrounds and are committed to providing equal opportunities. Should you require any accommodations during the recruitment process, please submit an Accommodations Request Form, and we will work together to meet your needs. We may use AI-enabled tools to screen and assess applicants against the criteria for this position. This helps our recruiters identify potentially qualified candidates, but it doesn't limit the applications our recruiters may review or consider.

$20k - $40k

Singapore remote FullTime
CoherePythonRAG +1 more

Site Reliability Engineer, Inference Infrastructure

8mo ago
Cohere

Cohere

Cohere is seeking a Site Reliability Engineer to join the Model Serving team. This role is crucial for developing, deploying, and operating the AI platform that delivers Cohere's large language models via API endpoints. You will work closely with various teams to deploy optimized NLP models into production environments, ensuring low latency, high throughput, and high availability. The position also offers the opportunity to interact with customers and create customized deployments to meet their specific needs, contributing to the widespread adoption of AI.

Toronto remote FullTime
CohereAWSAzure +5 more

Staff Software Engineer, Inference Infrastructure

8mo ago
Cohere

Cohere

Cohere is seeking Members of Technical Staff to join the Model Serving team. This role focuses on developing, deploying, and operating the AI platform that delivers Cohere's large language models via API endpoints. You will work closely with various teams to deploy optimized NLP models into production environments, ensuring low latency, high throughput, and high availability. The position also offers the opportunity to interact with customers and build customized deployments to meet their specific needs.

San Francisco hybrid FullTime
CohereAWSAzure +5 more

Member of Technical Staff, Data Analysis and Evaluation

9mo ago
Cohere

Cohere

Cohere is seeking a Member of Technical Staff in Data Analysis and Evaluation to ensure the quality, reliability, and performance of our large language models (LLMs). This role involves designing and conducting data collection tasks, assessing dataset quality, and analyzing model robustness and generalisability. You will collaborate with researchers, engineers, and data annotators to drive data-driven decisions and enhance AI system effectiveness. The position requires expertise in statistics, experimental design, and machine learning to ensure high-quality data and reliable model performance across diverse scenarios, contributing to Cohere's mission of advancing AI.

London remote FullTime
CoherePythonPyTorch +3 more

Forward Deployed Engineer, Infrastructure Specialist (Europe)

9mo ago
Cohere

Cohere

Cohere is a leading security-first enterprise AI company building cutting-edge foundation AI models and end-to-end products. We are seeking engineers to join our team and contribute to the widespread adoption of AI. This role offers a unique opportunity to shape how enterprises harness the power of AI in real-world applications, acting as a bridge between our core North product and client engineering teams. You will be at the forefront of solving complex problems and securely integrating AI into critical sectors like finance, healthcare, and telecommunications, working with esteemed clients and focusing on Agentic AI.

$20k - $40k

United Kingdom hybrid FullTime
CohereAWSAzure +3 more

Senior ML Systems Engineer, Frameworks & Tooling

9mo ago
Cohere

Cohere

Cohere is seeking a Senior ML Systems Engineer to join their team and build, maintain, and evolve the training framework that powers their frontier-scale language models. This role is ideal for someone passionate about large-scale training, distributed systems, and HPC infrastructure, offering the opportunity to design and maintain core components for fast, reliable, and scalable model training. You will also build tooling to connect research ideas to thousands of GPUs, working across the full stack of ML systems with significant autonomy and impact.

London remote FullTime
CohereDockerKubernetes +3 more

Senior Member of Technical Staff, Synthetic Data

9mo ago
Cohere

Cohere

Cohere is seeking a Senior Machine Learning Engineer specializing in synthetic data to develop and manage the synthetic data pipeline crucial for advanced language models. This role involves end-to-end management of synthetic data, including pipeline optimization, data analysis and generation, and conducting data ablations and model evaluations. You will transform diverse web and code data using generative models to enhance token efficiency and model quality, bridging research and engineering to improve throughput and accelerator utilization. This position is key to Cohere's mission of delivering efficient and reliable language capabilities and driving innovation in natural language processing.

Toronto remote FullTime
CoherePythonNLP +1 more

Member of Technical Staff - Sovereign AI

10mo ago
Cohere

Cohere

Cohere is seeking a Member of Technical Staff focused on Sovereign AI to design, build, and scale agentic AI systems for critical use cases. This role involves researching, implementing, and experimenting with novel ideas on supercompute and data infrastructure, with a strong emphasis on executing across the full AI stack to ship products serving the public interest. The position offers a unique opportunity to work with leading researchers and contribute to both production code and cutting-edge research, leveraging extensive compute resources and talent.

Canada onsite FullTime
CoherePython

Audio Inference Engineer, Model Efficiency

10mo ago
Cohere

Cohere

Cohere is seeking an Audio Inference Engineer focused on Model Efficiency to join a fast-growing team of researchers and engineers. The mission of this team is to build reliable machine learning systems and optimize audio inference serving efficiency using innovative techniques. As an engineer on this team, you will advance core audio model serving metrics, including latency, throughput, and quality by diving deep into systems, identifying bottlenecks, and delivering creative solutions for audio processing and streaming workloads. You will collaborate closely with both the training and serving infrastructure teams to ensure seamless integration between model development and deployment, with a special focus on real-time and streaming audio inference.

New York remote FullTime
CoherePythonC# +5 more

Staff Research Engineer, Model Efficiency

10mo ago
Cohere

Cohere

Cohere is seeking a Staff Research Engineer focused on Model Efficiency to join their cutting-edge enterprise AI company. This role is crucial for pushing the boundaries of Large Language Model (LLM) inference efficiency, addressing the current bottleneck in AI system capabilities. You will be responsible for developing, prototyping, and deploying techniques that significantly improve the speed and efficiency of our foundation models in production. This is an opportunity to contribute to the widespread adoption of AI by optimizing the model execution stack, including architecture, decoding, software/hardware co-design, and performance without compromising quality.

New York remote FullTime
Cohere

Member of Technical Staff, Model Efficiency

10mo ago
Cohere

Cohere

Cohere is seeking a Member of Technical Staff focused on Model Efficiency to join a fast-growing team of researchers and engineers. This role is instrumental in pushing the boundaries of LLM inference efficiency, developing techniques to improve model execution in production for lower latency and higher throughput. You will work across the inference stack, identifying bottlenecks and developing optimizations, collaborating with modeling and systems teams to ship meaningful improvements. Opportunities exist to build expertise in advanced performance techniques like GPU/CUDA optimizations and model execution strategies for large-scale architectures.

New York remote FullTime
CoherePythonGo +3 more

Senior Research Scientist, Model Evaluation

10mo ago
Cohere

Cohere

Cohere is seeking a Senior Research Scientist focused on Model Evaluation to join their team. This role is crucial for advancing AI capabilities by developing next-generation evaluation methods and infrastructure to accurately measure LLM progress. You will be responsible for creating new evaluation benchmarks, translating model feedback into reliable evaluations, and conducting research to push the state-of-the-art in LLM evaluation techniques. The ideal candidate is passionate about rigorously measuring AI capabilities and ensuring these measurements align with desired outcomes.

Toronto hybrid FullTime
Cohere

Applied Machine Learning Engineer

11mo ago
Cohere

Cohere

Cohere is seeking a Member of Technical Staff for their Applied ML team. In this role, you will collaborate directly with customers to understand their challenges and implement solutions leveraging Large Language Models. You will apply your problem-solving skills, creativity, and technical expertise to bridge the gap in enterprise AI adoption, delivering impactful products and disrupting key industries. This is an opportunity to join at a pivotal moment, shape the company's offerings, and contribute to cutting-edge AI development.

London hybrid FullTime
CoherePythonTensorFlow +2 more

Member of Technical Staff, Agent Code

1y ago
Cohere

Cohere

Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company co-headquartered in Toronto and San Francisco, with key offices in London, New York City, Montreal, Seoul, Germany and Paris. Join us! Why this role? Code-generating LLMs and autonomous agents are revolutionizing how software is built and tasks are automated. At Cohere, we’re pushing the boundaries of what’s possible with these technologies for enterprises, and we’re looking for a senior member for the Agent Code team. You’ll be at the forefront of research and engineering, driving the development of cutting-edge code LLMs and agent systems that can interact with the digital world to solve complex tasks with minimal human oversight. This role is hands-on and research-driven. You’ll dive into the latest literature on code LLMs and agents, experiment with frontier models, and collaborate with a team of talented engineers and researchers to build scalable, production-ready solutions. At Cohere, we blend engineering and research seamlessly—everyone contributes to both, depending on their interests and organizational needs. We provide access to world-class compute resources, data, and talent to ensure you can do your best work. Note: We have offices in London, Toronto, New York and San Francisco, but we’re also remote-friendly! This team operates primarily between ET to CET time zones, so we’re seeking candidates in locations that align with these hours for effective collaboration. As a Member of Technical Staff on the Agent Code team, you will: - Stay up-to-date with the latest research in code LLMs, agents, and related fields, implementing novel ideas into our systems. - Design and implement scalable strategies to train code models, and deploy agent frameworks for inference and sampling. You will be collaborating with the pretraining team, create SFT trajectories and work on existing and new RL algorithms - Hillclimb on existing benchmarks and design new ones that reflect the needs of our enterprise users - Lead experiments on our state-of-the-art compute infrastructure, pushing the boundaries of what’s possible with frontier LLMs. You may be a good fit if you have: - A PhD in Computer Science, Machine Learning, or a related field, with publications in top-tier venues (e.g., NeurIPS, ICML, ICLR, ACL, EMNLP). - Deep expertise in code LLMs and agent systems, with a strong understanding of the latest research and trends. We are looking for people who not only have worked with code models, but have actively contributed to their development - Hands-on experience with frontier LLMs and their applications in code generation or automation. - Strong software engineering skills, with proficiency in Python and PyTorch, TensorFlow, or similar frameworks. - Experience with distributed systems, cloud infrastructure, and scalable architectures. - A proactive, self-motivated mindset, with a passion for solving ambitious, open-ended problems. What We Offer: - The opportunity to work on cutting-edge problems at the intersection of AI, code generation, and autonomous agents. - Access to world-class compute resources, data, and a collaborative team of researchers and engineers. - A remote-friendly, flexible work environment with a focus on impact and innovation. - Competitive compensation and benefits, including equity in a fast-growing AI company. If you’re passionate about shaping the future of code LLMs and agent systems, and thrive in a dynamic, research-driven environment, we’d love to hear from you! Full-Time Employees at Cohere enjoy these Perks: - A weekly lunch stipend of $75/£75 or equivalent in your local currency for lunch. - Full health and dental benefits, including a separate budget for mental health. - RRSP matching, 401K, Pension Scheme. - 100% Parental Leave top-up for up to 6 months, for either parent. - Annual enrichment benefits: Arts & culture, fitness/wellness, quality time, and a workspace improvement credit. Education & learning stipend for conferences, courses, and coaching. - 6 weeks of paid vacation (30 working days!) - Budget for traveling to other offices if you are remote, plus an annual company offsite. How and Where We Work: - Cohere is remote-friendly. We have offices in Toronto, San Francisco, New York City, London, Paris, Montreal, and more coming soon. - For those in the office: a daily lunch program, plenty of snacks, and regular community and social events. - For those not near an office: a co-working benefit so you can work alongside others in your city. - Everyone receives a $500 home office stipend to set up your workspace properly. If any of the above doesn’t line up exactly with your experience, we still encourage you to apply. We strive to create an inclusive work environment for all; we welcome applicants from all backgrounds and are committed to providing equal opportunities. Should you require any accommodations during the recruitment process, please submit an Accommodations Request Form, and we will work together to meet your needs. We may use AI-enabled tools to screen and assess applicants against the criteria for this position. This helps our recruiters identify potentially qualified candidates, but it doesn't limit the applications our recruiters may review or consider.

London hybrid FullTime
CoherePythonPyTorch +1 more

Member of Technical Staff, Integration/RL Team (Research Engineer)

1y ago
Cohere

Cohere

Cohere is a leading enterprise AI company focused on building cutting-edge foundation AI models and end-to-end products for real-world business problems. The integration team specifically focuses on developing and scaling machine learning algorithms and infrastructure for LLM post-training, with an emphasis on large-scale, distributed Reinforcement Learning (RL) methods. This role is crucial for enhancing the post-training codebase by implementing new research tools, optimizing algorithms, and scaling distributed RL capabilities. We are looking for passionate individuals who are meticulous in their approach to engineering and science, contributing to both production code and research efforts.

Paris remote FullTime
CohereKubernetesPython +3 more

Member of Technical Staff, Post-Training

1y ago
Cohere

Cohere

Cohere is seeking a Member of Technical Staff to focus on post-training of AI models. This role is crucial for advancing the state of the art in model post-training and shipping cutting-edge models to production, bridging the gap between research and practical application. You will have access to significant compute resources and a talented team to contribute to increasing model capabilities and driving customer value. The position offers a unique opportunity to contribute to both production code and research efforts, depending on your interests and organizational needs.

London hybrid FullTime
CohereKubernetesPython +3 more

Senior Member of Technical Staff, Multimodal AI

1y ago
Cohere

Cohere

Cohere is seeking a Senior Member of Technical Staff focused on Multimodal AI to join their cutting-edge enterprise AI company. This role involves designing and developing advanced multimodal AI systems that integrate text, speech, and vision, pushing the boundaries of what's possible in AI. You will have access to exceptional compute resources and collaborate with world-class teams to innovate and shape the future of AI. The position is ideal for individuals passionate about machine learning and its real-world applications, who enjoy optimizing large models and thrive in a fast-paced, technically challenging environment.

San Francisco remote FullTime
OpenAIMistralCohere +5 more

Get new AI jobs in your inbox

A weekly digest of the newest AI engineering roles.

© 2026 AI Job Board. All rights reserved.