Find your next AI Engineering role

1097+ open roles · 22+ companies hiring

1097 open positions

Legal Engineering Manager (Law Firm, Corporate)

1mo ago
Harvey

Harvey

Harvey is seeking experienced Legal Engineering Managers to lead and grow high-performing teams focused on go-to-market efforts. This role bridges legal expertise, customer engagement, and commercial strategy, helping clients understand how Harvey can revolutionize legal work. You will manage a team of Legal Engineers who collaborate with Account Executives during the sales process, offering legal credibility and product knowledge. Beyond team management, you will engage directly with strategic customers, influence Harvey's go-to-market strategy, and work with Product, Marketing, Enablement, and Engineering teams to enhance Harvey's service to the legal industry. This position requires a blend of strategic thinking and hands-on involvement, ideal for seasoned legal professionals with a corporate law background who thrive in fast-paced, ambiguous environments and excel at problem-solving with a strong sense of ownership.

$315k - $385k

New York hybrid FullTime
Go

Software Engineer, North for Finance

1mo ago
Cohere

Cohere

Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company co-headquartered in Toronto and San Francisco, with key offices in London, New York City, Montreal, Seoul, Germany and Paris. Join us! Why this role? North is Cohere's cutting-edge AI workspace platform, designed for enterprise users. It offers a secure and customizable environment, allowing companies to deploy AI while maintaining control over sensitive data. North integrates seamlessly with existing workflows, providing a trusted platform that connects AI agents with workplace tools, data and applications. The North for Finance team builds specialized AI solutions for finance teams while accelerating adoption of agentic workflows for their mission-critical operations. You'll work at the intersection of AI innovation and financial services, developing features that integrate domain-specific knowledge with enterprise-grade reliability. As an engineer on this team you will help transform complex financial workflows through intelligent automation. You'll extend core agentic platform capabilities with built-in data isolation, structured-data manipulation and collaborative governance, positioning North as the leading AI workspace for finance. As a Senior Software Engineer, you will: - Design, build, ship, and maintain customer facing workflows and agentic automations - Own features end-to-end, from technical design through implementation, testing, launch, and iteration. - Collaborate closely with Product, UX, ML, and Sales teams to validate requirements and prototype quickly - Work closely with customers in the finance domain - Develop custom interfaces that seamlessly integrate with financial data sources and existing enterprise tools - Extend core agentic platform capabilities to support finance-specific requirements and ensure reliability, security, and auditability - Develop collaborative workspaces with appropriate access controls (RBAC) and data isolation - Own technical decisions that balance speed, quality and user needs, maintaining alignment with the product roadmap - Use AI actively in your work, while staying accountable for the quality and reliability of what you ship You may be a good fit if: - You have experience in full-stack product engineering - You have built, shipped, and operated production software used by thousands of users - You have strong product judgment, pay attention to details and care deeply about the end user needs - You excel in fast-paced environments and can execute while priorities and objectives are a moving target - You bring high agency: you learn quickly and drive change without waiting for perfect instructions. - You are curious and can identify, troubleshoot and fix issues across systems boundaries - You can clearly document and track architectural proposals and technical decisions - You are excited to use AI, while staying directly accountable for the systems you build - You have integrated AI agents in your day-to-day with well structured workflows It's a bonus if you have: - Experience in financial services or regulated industries - Familiarity with agentic AI architectures and implementation patterns - Knowledge of access control implementation in complex systems If this role sounds exciting, but your experience does not match every bullet, we still encourage you to apply. We value and celebrate diversity and strive to create an inclusive work environment for all. We welcome applicants from all backgrounds and are committed to providing equal opportunities. Should you require any accommodations during the recruitment process, please submit an Accommodations Request Form, and we will work together to meet your needs. Full-Time Employees at Cohere enjoy these Perks: - A weekly lunch stipend of $75/£75 or equivalent in your local currency for lunch. - Full health and dental benefits, including a separate budget for mental health. - RRSP matching, 401K, Pension Scheme. - 100% Parental Leave top-up for up to 6 months, for either parent. - Annual enrichment benefits: Arts & culture, fitness/wellness, quality time, and a workspace improvement credit. Education & learning stipend for conferences, courses, and coaching. - 6 weeks of paid vacation (30 working days!) - Budget for traveling to other offices if you are remote, plus an annual company offsite. How and Where We Work: - Cohere is remote-friendly. We have offices in Toronto, San Francisco, New York City, London, Paris, Montreal, and more coming soon. - For those in the office: a daily lunch program, plenty of snacks, and regular community and social events. - For those not near an office: a co-working benefit so you can work alongside others in your city. - Everyone receives a $500 home office stipend to set up your workspace properly. If any of the above doesn’t line up exactly with your experience, we still encourage you to apply. We strive to create an inclusive work environment for all; we welcome applicants from all backgrounds and are committed to providing equal opportunities. Should you require any accommodations during the recruitment process, please submit an Accommodations Request Form, and we will work together to meet your needs. We may use AI-enabled tools to screen and assess applicants against the criteria for this position. This helps our recruiters identify potentially qualified candidates, but it doesn't limit the applications our recruiters may review or consider.

Europe remote FullTime
CohereAI Agents

Head of Finance Systems & Automation

1mo ago
Scale AI

Scale AI

Scale AI is seeking a Head of Financial Systems to lead the strategy, implementation, and ongoing enhancement of the company's finance technology stack. This role requires a builder-first leader who can combine enterprise architecture expertise with AI automation to streamline manual processes, strengthen financial data management, and empower the Finance team with the necessary speed and accuracy. You will serve as the primary technology partner to Finance leadership, holding ultimate accountability for the reliability, scalability, and intelligence of all systems involved in the financial close, revenue, and spend lifecycles.

$198k - $248k

San Francisco, CA onsite
PythonLLMSQL +8 more

Security Engineer, Detection & Response

1mo ago
Scale AI

Scale AI

We are seeking a Senior Security Engineer specializing in Detection and Incident Response to join our Security Engineering team. This role blends security operations with software engineering, focusing on building systems for detection, containment, and prevention rather than just incident investigation. You will be responsible for designing and implementing high-precision detections across cloud and enterprise SaaS, developing automation to speed up response times, and enhancing telemetry pipelines. Your ability to write production-quality code is as crucial as your incident triage skills. You will analyze root causes, communicate incident significance and impact, and translate findings into engineering improvements like better detections, refined schemas, and smarter automation.

$238k - $297k

New York, NY; San Francisco, CA; Seattle, WA; Washington, DC onsite
AWSAzurePython +6 more

Director of Product, North

1mo ago
Cohere

Cohere

Cohere is seeking a Director of Product for North, an agentic AI platform designed to securely deploy AI agents and automations within organizations' infrastructure. This senior leadership role involves owning the product vision, strategy, and execution for the North platform, building and leading a team of product managers, and being accountable for business outcomes. The ideal candidate will set the direction for the platform, drive adoption, retention, and net expansion, and ensure a high bar for quality, security, and time-to-value for an enterprise platform deployed within customer infrastructure.

Toronto remote FullTime
CohereAI Agents

Product Manager, Integrations

1mo ago
Cohere

Cohere

Cohere is a leading enterprise AI company building cutting-edge foundation AI models and end-to-end products. We are seeking a Product Manager for Integrations to own the end-to-end integrations surface for North, Cohere's agentic AI platform. This role involves defining the integration ecosystem strategy, prioritizing and shipping integrations, and building underlying platform capabilities. You will partner with strategic customers and collaborate with engineering and field teams to drive adoption and customer value.

Toronto remote FullTime
CohereAI Agents

Product Manager of AI Applications, Global Public Sector

1mo ago
Scale AI

Scale AI

Scale is seeking an entrepreneurial and experienced Product Manager to join its Global Public Sector team. This role focuses on developing and deploying transformative AI solutions for governments and government-backed entities outside the United States. You will leverage Scale's GenAI Platform to build cutting-edge AI applications powered by customer data, working closely with clients to understand their needs and deliver bespoke solutions. This is an opportunity to own large AI projects, lead cross-functional teams, and drive significant business impact.

Riyadh, Saudi Arabia remote
PythonGoAI +2 more

Product Manager of AI Applications, Global Public Sector

1mo ago
Scale AI

Scale AI

Scale is seeking an entrepreneurial and experienced Product Manager to join its Global Public Sector team. This role focuses on developing and deploying transformative AI solutions for governments and government-backed entities outside the United States. You will leverage Scale's GenAI Platform to build cutting-edge AI applications powered by customer data, working closely with clients to understand their needs and deliver bespoke solutions. This is an opportunity to own large AI projects, lead cross-functional teams, and drive significant business impact.

Doha, Qatar ; Dubai, UAE remote
PythonGoAI +2 more

Mid-Market Customer Success Manager

1mo ago
Harvey

Harvey

Harvey is transforming how legal and professional services operate by combining frontier agentic AI, an enterprise-grade platform, and deep domain expertise. This role offers a unique opportunity to help build a generational company at an inflection point, scaling fast and defining a new category. The team operates with intensity, takes ownership, and is deeply committed to the mission, valuing decisiveness, simplicity, and a "Job's Not Finished" mentality. If you are driven to do the best work of your career alongside like-minded individuals, Harvey is the place to be.

$125k - $145k

San Francisco onsite FullTime

Mid-Market Customer Success Manager

1mo ago
Harvey

Harvey

Harvey is transforming how legal and professional services operate by combining frontier agentic AI, an enterprise-grade platform, and deep domain expertise. This role offers a unique opportunity to help build a generational company at an inflection point, scaling fast and defining a new category. The team operates with intensity, takes ownership, and is deeply committed to the mission, valuing decisiveness, simplicity, and a "Job's Not Finished" mentality. If you are driven to do the best work of your career alongside like-minded individuals, Harvey is the place to be.

$125k - $145k

Chicago hybrid FullTime

Technical Sourcer

1mo ago
Replit

Replit

Replit is seeking a Technical Sourcer for a 6-12 month fixed-term full-time position. This role is crucial in building the talent pipeline for Replit's Product and Engineering teams. You will be responsible for identifying and engaging top-tier technical talent, particularly in Software Engineering (Product, Infra, AI/ML), Product Management, and Design. The ideal candidate will be a "hunter" with a deep understanding of technology, capable of using AI tools and creative outreach strategies to find "undiscovered" talent across various platforms like GitHub and Twitter. You will partner closely with recruiters and engineering leads, utilizing data to refine sourcing strategies and ensure a high-quality funnel.

Foster City, CA hybrid FullTime
Go

Support Operations Data Analyst

1mo ago
Harvey

Harvey

Harvey is seeking its first Support Operations Data Analyst to own the analytics function for the User Operations organization. This role is crucial for transforming how support data is managed, analyzed, and utilized to improve critical metrics. You will be responsible for building and maintaining dashboards, reports, and feedback loops that track key performance indicators such as cSAT, TTR, QA scores, and escalation rates. The ideal candidate will be adept at translating complex data into clear narratives, identifying data gaps, and ensuring the organization is equipped with the necessary instrumentation as it scales. This is a solo role requiring independence, confidence in metric framing, and the ability to operate at a fast pace.

$112k - $168k

New York hybrid FullTime
PythonSQL

Support Operations Data Analyst

1mo ago
Harvey

Harvey

Harvey is seeking its first Support Operations Data Analyst to own the analytics function for the User Operations organization. This role is crucial for transforming how support data is managed, analyzed, and utilized to improve critical metrics. You will be responsible for building and maintaining dashboards, reports, and feedback loops that track key performance indicators such as cSAT, TTR, QA scores, and escalation rates. The ideal candidate will be adept at translating complex data into clear narratives, identifying data gaps, and ensuring the organization is equipped with the necessary instrumentation as it scales. This is a solo role requiring independence, confidence in metric framing, and the ability to operate at a fast pace.

$112k - $168k

Remote remote FullTime
PythonSQL

Support Operations Data Analyst

1mo ago
Harvey

Harvey

Harvey is seeking its first Support Operations Data Analyst to own the analytics function for the User Operations organization. This role is crucial for transforming how support data is managed, analyzed, and utilized to improve critical metrics. You will be responsible for building and maintaining dashboards, reports, and feedback loops that track key performance indicators such as cSAT, TTR, QA scores, and escalation rates. The ideal candidate will be adept at translating complex data into clear narratives, identifying data gaps, and ensuring the organization is equipped with the necessary instrumentation as it scales. This is a solo role requiring independence, confidence in metric framing, and the ability to operate at a fast pace.

$112k - $168k

San Francisco hybrid FullTime
PythonSQL

GTM Systems Analyst

1mo ago
Scale AI

Scale AI

Scale is seeking a technical Business Systems Analyst to act as the product owner for the Public Sector Salesforce environment and build AI tooling for the broader organization. This role is designed to be more than a standard Salesforce administrator, focusing on leveraging AI to accelerate sales cycles and streamline operations. The analyst will be responsible for architecting the GTM systems stack, building AI-powered automations, and owning integration design across various platforms. This position requires close collaboration with GTM leadership, deal desk, finance, legal, and compliance to translate business needs into scalable systems, while also ensuring compliance with FedRAMP/IL-tier and CMMC standards.

$166k - $207k

San Francisco, CA remote
AnthropicLLMClaude +9 more

Staff Platform Engineer, Voice AI

1mo ago
Together AI

Together AI

Together AI is seeking a Staff Platform Engineer to lead the architecture of their Voice AI platform, which powers real-time voice agents at scale. This role involves setting the technical direction for how developers interact with the platform, from API primitives to autoscaling systems and multi-provider abstractions. The focus is on building robust, low-latency infrastructure for voice applications, which presents unique challenges compared to text inference, such as handling bidirectional audio streams and stateful connections. This is a foundational position on a small team, where decisions will shape the platform's architecture for years to come.

$220k - $280k

San Francisco remote
KubernetesPythonTypeScript +7 more

Staff Engineer, Product UI Platform

1mo ago
Together AI

Together AI

Together AI is seeking a highly experienced Staff Engineer to own and evolve the Product UI Platform, the architectural foundation powering full-stack feature development across their web surface. This role involves evolving the product runtime from its current monolithic architecture to a scalable, modular, and high-leverage platform that supports increasing scale and reliability. You will define and drive the technical direction of the Next.js/TypeScript/Node.js web runtime, BFF layer, and application integration patterns, collaborating with Backend and API Platform leaders to ensure cohesive architectural evolution. This is a hands-on role for an experienced engineer ready to take full ownership of a critical technical domain.

$200k - $275k

San Francisco remote
TypeScriptReactNode.js +2 more

Sr. Partnerships Manager, Model Ecosystem

1mo ago
Together AI

Together AI

As the Partnerships Manager for our Model Ecosystem, you will be the primary architect of Together AI’s model library. This high-impact, cross-functional role focuses on bringing the world’s leading proprietary and open-source models onto the Together platform. You will navigate an ever-evolving landscape to negotiate non-standard, creative deals that provide developers with the best possible building blocks for AI applications. You are a "deal-maker" who thrives in ambiguity, sitting at the intersection of Product, Finance, and Marketing to ensure our model roadmap is not only technically superior but commercially viable and market-facing.

$270k - $300k

San Francisco remote
GoFine-TuningAI +1 more

Solutions Architect (Inference)

1mo ago
Together AI

Together AI

As a Solutions Architect (Inference) at Together AI, you will work with customers and prospects to create business value through Generative AI applications. Solutions Architects at Together are trusted advisors to our customers that evaluate, identify and demonstrate how Together can solve their AI needs. As key contributors to our sales organization, Solution Engineers add tremendous value to the customer journey and directly impact company growth and revenue. This is an exciting opportunity for a deeply technical professional passionate about AI and customer success to make a significant impact in a fast-paced, innovative environment.

London onsite
DockerKubernetesPython +8 more

Solutions Architect

1mo ago
Together AI

Together AI

As a Solutions Architect at Together AI, you will work with customers and prospects to create business value through Generative AI applications. You will act as a trusted advisor, evaluating, identifying, and demonstrating how Together AI can solve their AI needs. This role is a key contributor to the sales organization, adding significant value to the customer journey and directly impacting company growth and revenue. It's an exciting opportunity for a deeply technical professional passionate about AI and customer success to make a significant impact in a fast-paced, innovative environment.

$180k - $260k

San Francisco remote
DockerKubernetesPython +8 more

Senior Software Engineer - Together Cloud Platform

1mo ago
Together AI

Together AI

Together AI is building the AI Acceleration Cloud, an end-to-end platform for the full generative AI model lifecycle, combining a fast LLM inference engine with state-of-the-art AI cloud infrastructure. As a Senior Backend Engineer, you will play a key role in building the next generation AI cloud platform. This platform is designed to be highly available, global, and extremely fast, virtualizing cutting-edge ML hardware and enabling practitioners with self-serve AI cloud services. It serves both internal StaaS products and external cloud customers across numerous data centers worldwide.

$160k - $230k

San Francisco remote
AWSAzureKubernetes +8 more

Senior Software Engineer Together Cloud Infrastructure

1mo ago
Together AI

Together AI

Together AI is building the AI Acceleration Cloud, an end-to-end platform for the full generative AI lifecycle, combining the fastest LLM inference engine with state-of-the-art AI cloud infrastructure. As a Senior AI Infrastructure Engineer, you will play a key role in building the next generation AI cloud platform – a highly available, global, blazing-fast cloud infrastructure that virtualizes cutting-edge ML hardware and enables state-of-the-art ML practitioners with self-serve AI cloud services. This platform serves both our internal SaaS products and our external cloud customers, spanning dozens of data centers across the world.

Amsterdam hybrid
AWSAzureKubernetes +12 more

Senior Software Engineer - Together Cloud Infrastructure

1mo ago
Together AI

Together AI

Together AI is building the AI Acceleration Cloud, an end-to-end platform for the full generative AI lifecycle, combining the fastest LLM inference engine with state-of-the-art AI cloud infrastructure. As a Senior AI Infrastructure Engineer, you will play a key role in building the next generation AI cloud platform – a highly available, global, blazing-fast cloud infrastructure that virtualizes cutting-edge ML hardware and enables state-of-the-art ML practitioners with self-serve AI cloud services. This platform serves both our internal SaaS products and our external cloud customers, spanning dozens of data centers across the world.

$160k - $230k

San Francisco remote
AWSAzureKubernetes +12 more

Senior Platform Engineer, Voice AI

1mo ago
Together AI

Together AI

Together AI is building the best inference infrastructure for voice applications, powering production-grade, real-time voice agents and applications with best-in-class latency and reliability. We are seeking a Senior Platform Engineer to take ownership of the API and infrastructure layer for voice workloads. You will develop the real-time WebSocket and HTTP APIs used by developers to deploy voice experiences, design autoscaling for latency-sensitive streaming workloads, and ensure the reliability of our multi-provider voice platform for production voice agents handling millions of calls. This is a critical, foundational role on a small, high-impact team, defining how developers interact with our voice platform as we scale.

$200k - $260k

San Francisco remote
KubernetesPythonTypeScript +8 more

Senior Backend Engineer, Inference Platform

1mo ago
Together AI

Together AI

Together AI is building the Inference Platform to bring advanced generative AI models to the world, powering multi-tenant serverless workloads and dedicated endpoints. This role offers a unique opportunity to optimize latency and fully utilize tens of thousands of GPUs, working hands-on with cutting-edge hardware. You will collaborate directly with research teams to productionize frontier models and engage with the open-source community, contributing to projects that push the boundaries of inference performance and efficiency.

$160k - $250k

San Francisco remote
KubernetesPythonTypeScript +7 more

Research Intern, Model Shaping (Fall 2026)

1mo ago
Together AI

Together AI

As a Research Intern in the Model Shaping team, you will work on advanced post-training methods, new techniques for efficient neural network training, and robust evaluation of foundation model capabilities. The Model Shaping team at Together AI focuses on tailoring open foundation models for downstream applications, building services for machine learning developers, and developing new methods for efficient model training and evaluation. This role offers the opportunity to contribute to cutting-edge research and potentially influence open-source projects.

San Francisco hybrid
PyTorchNLPReinforcement Learning +6 more

Lead/Manager Together Cloud Infrastructure

1mo ago
Together AI

Together AI

Together AI is building the AI Acceleration Cloud, an end-to-end platform for the full generative AI lifecycle, combining the fastest LLM inference engine with state-of-the-art AI cloud infrastructure. As a Lead/Manager, you will play a key role in building the Together cloud platform engineering team in the Netherlands. This platform serves both our internal SaaS products (inference, fine-tuning) and our external cloud customers, spanning dozens of data centers across the world.

Amsterdam hybrid
AWSAzureKubernetes +9 more

Frontier Agents Intern (Fall 2026)

1mo ago
Together AI

Together AI

The Agents team investigates how to build, align, and scale frontier AI systems capable of complex, multi-step tasks and workflows across text and speech, with a focus on agentic and scientific domains. This role sits at the intersection of agent capabilities, human-computer interaction, and infrastructure, exploring areas like post-training methods for agentic behavior and developing evaluation frameworks for open-ended tasks. As a research intern, you will tackle challenges in alignment, reliability, and scalability, potentially working on new training recipes for self-learning and long-horizon reasoning, curating datasets, studying failure modes, or building scalable agent infrastructure.

San Francisco remote
PythonPyTorchNLP +6 more

Forward Deployed Engineer (Inference & Post-Training)

1mo ago
Together AI

Together AI

As a Forward Deployed Engineer (FDE) focused on Inference & Post-Training, you will be a hands-on technical partner to strategic customers, assisting production AI teams with leveraging high-quality models and performing inference at scale. You will act as a deep-domain specialist in inference optimization, fine-tuning pipelines, and production deployment, partnering with Solutions Architects. FDEs add significant value by ensuring complex Proofs of Concept (POCs) are met, facilitating platform adoption, and guiding tailored optimization efforts, directly impacting customer success and company growth.

$270k - $300k

San Francisco remote
PythonFine-TuningRLHF +9 more

Customer Support Engineer (Inference), India

1mo ago
Together AI

Together AI

As a Customer Support Engineer at a pioneering AI company, you will be the first line of defense supporting customers building training, fine-tuning, and inference solutions. You will dive deep into complex technical challenges, providing swift and effective solutions while serving as a product expert. Collaborating closely with product and sales, you will drive continuous improvement of our offerings. This is an exciting opportunity for a deeply technical professional passionate about AI and customer success to make a significant impact in a fast-paced, innovative environment.

India remote
KubernetesPythonTypeScript +9 more

Customer Support Engineer (Inference)

1mo ago
Together AI

Together AI

As a Customer Support Engineer at a pioneering AI company, you will be the primary point of contact for customers building training, fine-tuning, and inference solutions with Together AI. You will tackle complex technical challenges, provide effective solutions, and act as a product expert. Collaborating with product and sales teams, you will contribute to the continuous improvement of our offerings. This role is ideal for a technically proficient individual passionate about AI and customer success, seeking to make a significant impact in a fast-paced, innovative environment.

$160k - $230k

San Francisco, CA remote
KubernetesPythonTypeScript +9 more

Backend Software Engineer — Data Platform & AI Data Products

1mo ago
Together AI

Together AI

You will join the Data Platform team, responsible for building the backend services and data products that power how data moves through the company. This involves creating core platform primitives like high-quality event streams, reliable access layers, and developer-friendly APIs and tools. The goal is to enable teams across the organization to self-serve their data needs and ship faster. You will contribute to backend services that derive value from company data and enhance the self-serve capabilities of the data platform, allowing product and engineering teams to easily create and operate event-driven architectures, publish/consume streams, define access models, and manage data products end-to-end. Additionally, you will work on LLM-adjacent services, including prompt categorization, enrichment, and metadata systems, transforming raw telemetry into trusted, usable products with guidance from experienced engineers.

$120k - $170k

San Francisco remote
PythonGoRust +9 more

AI Researcher, Core ML (Turbo)

1mo ago
Together AI

Together AI

The Turbo team operates at the intersection of efficient inference (algorithms, architectures, engines) and post-training/RL systems. We are responsible for building and managing the systems that power Together's API, focusing on high-performance inference and RL/post-training engines capable of operating at production scale. Our core mission is to advance the frontiers of efficient inference and RL-driven training, aiming to make models significantly faster and more cost-effective to run, while simultaneously enhancing their capabilities through RL-based post-training methods. This role involves working across the entire stack, from RL algorithms and training engines to kernels and serving systems, to develop and refine state-of-the-art models using RL pipelines. We value individuals with deep expertise in one area and a strong willingness to collaborate and grow across others.

$200k - $280k

San Francisco remote
PythonTransformersRLHF +7 more

Systems Research Engineer Intern - GPU Programming (Fall 2026)

1mo ago
Together AI

Together AI

As a Systems Research Engineer Intern specialized in GPU Programming, you will play a crucial role in developing and optimizing GPU-accelerated kernels and algorithms for ML/AI applications. You will co-design GPU kernels and model architecture with the modeling and algorithm team to enhance the performance and efficiency of our AI systems. Collaborating with the hardware and software teams, you will contribute to the co-design of efficient GPU architectures and programming models, leveraging your expertise in GPU programming and parallel computing. Your research skills will be vital in staying up-to-date with the latest advancements in GPU programming techniques, ensuring that our AI infrastructure remains at the forefront of innovation.

San Francisco remote
CUDATritonParallel Computing +4 more

Systems Research Engineer, GPU Programming

1mo ago
Together AI

Together AI

As a Systems Research Engineer specialized in GPU Programming, you will play a crucial role in developing and optimizing GPU-accelerated kernels and algorithms for ML/AI applications. You will co-design GPU kernels and model architecture to enhance the performance and efficiency of our AI systems, and contribute to the co-design of efficient GPU architectures and programming models. Your research skills will be vital in staying up-to-date with the latest advancements in GPU programming techniques, ensuring that our AI infrastructure remains at the forefront of innovation.

$160k - $230k

San Francisco remote
AIMLCUDA +5 more

Staff Machine Learning Engineer, Voice AI

1mo ago
Together AI

Together AI

Together AI is building the best inference infrastructure for voice applications, powering production-grade, real-time voice agents and applications. We are seeking a Staff ML Engineer to lead the model serving layer for voice workloads. This role involves hands-on optimization of inference engines and models like Whisper, Parakeet, Orpheus, and Kokoro, focusing on pushing latency and throughput boundaries. You will address unique challenges in voice inference, such as streaming audio and real-time latency, and shape the future of how voice models are served as the industry shifts towards end-to-end speech-to-speech systems. This is a foundational hire on a small, high-impact team.

$220k - $280k

San Francisco remote
PythonGoFine-Tuning +10 more

Staff Engineer, Distributed Storage and HPC & AI Infrastructure

1mo ago
Together AI

Together AI

Together AI is seeking a Staff Engineer to design and deliver multi-petabyte storage systems optimized for large-scale AI training and inference workloads. You will architect high-performance parallel filesystems and object stores, integrate cutting-edge technologies, and drive significant cost optimization. The role involves building Kubernetes-native storage operators and self-service platforms for automated provisioning and multi-tenancy. You will focus on optimizing data paths, designing multi-tier caching architectures, and tuning parallel filesystems for AI applications. This is a research-driven role within a company focused on lowering the cost of modern AI systems through co-design of software, hardware, algorithms, and models.

$250k - $300k

San Francisco remote
KubernetesPythonGo +7 more

Senior Machine Learning Engineer, Voice AI

1mo ago
Together AI

Together AI

Together AI is building the best inference infrastructure for voice applications, powering production-grade, real-time voice agents and applications. We are seeking a Senior ML Engineer to lead the model serving layer for voice workloads. This role involves hands-on optimization of inference engines and models like Whisper, Parakeet, Orpheus, and Kokoro to achieve frontier-level latency and throughput. You will focus on unique voice inference challenges such as streaming audio, tokenization, and real-time latency budgets, shaping how voice models are served as the industry shifts towards end-to-end speech-to-speech systems. This is a foundational hire on a small, high-impact team.

$200k - $260k

San Francisco remote
PythonGoFine-Tuning +10 more

Research Engineer, Frontier Speculative Decoding

1mo ago
Together AI

Together AI

Together AI is building the Inference Platform that powers the world's most advanced generative AI models. This role will serve as a critical bridge between cutting-edge research and real-world applications, focusing on translating internal model training research into production-ready deployments for customers. The work involves a deep commitment to data-centric development, meticulous hyperparameter tuning, and rigorous checkpoint evaluation. You will transform general-purpose models into highly performant, specialized tools by fine-tuning them on customer-specific data and internal datasets, working with dedicated GPU clusters rather than training foundation models from scratch.

$190k - $270k

San Francisco, New York City remote
KubernetesPythonFine-Tuning +6 more

Research Engineer, Core ML

1mo ago
Together AI

Together AI

This research engineering role focuses on translating new Reinforcement Learning (RL) algorithms, scheduling methods, and inference optimizations into production-grade systems that power Together's API. The Core ML team operates at the intersection of efficient inference (algorithms, architectures, engines) and post-training/RL systems, building and maintaining high-performance inference and RL engines at production scale. The goal is to significantly improve model speed, cost-efficiency, and capabilities through RL-based post-training. This position requires a blend of algorithmic understanding and systems engineering, with opportunities to work across the entire stack from RL algorithms and training engines to kernels and serving systems, ultimately driving measurable improvements in latency, throughput, cost, and model quality at scale.

$200k - $280k

San Francisco remote
PythonTransformersRLHF +8 more

Frequently asked questions

What counts as an AI Engineering job?

AI Job Board lists engineering roles that build, deploy, or operate AI/ML systems — including LLM engineering, RAG (retrieval-augmented generation), AI agents, prompt engineering, AI infrastructure and MLOps, model serving/inference, and fine-tuning. It excludes non-engineering functions like AI sales or marketing roles.

What is the difference between an AI Engineer and a Machine Learning Engineer?

An AI Engineer typically builds applications on top of existing models — LLM integration, RAG pipelines, agent orchestration, and prompt design. A Machine Learning Engineer more often trains, fine-tunes, or productionizes custom models. Many companies use the titles interchangeably, so search both when browsing.

Does AI Job Board include AI infrastructure and MLOps roles?

Yes. Titles such as AI Infrastructure Engineer, ML Platform Engineer, GPU Infrastructure Engineer, Inference Engineer, MLOps Engineer, and AI Site Reliability Engineer are all covered — these roles focus on the systems that serve and scale AI models rather than building models themselves.

Are remote AI and ML jobs available?

Yes. Filter by "Remote" on the jobs page to see fully remote AI/ML engineering roles, or browse the remote jobs feed directly.

What skills are most in demand for AI engineering roles?

The most commonly requested skills are LLM APIs (OpenAI, Anthropic, Gemini), RAG and vector databases (Pinecone, Weaviate, Qdrant), agent frameworks (LangGraph, CrewAI, AutoGen), inference/serving tools (vLLM, Ray Serve, TensorRT-LLM), and fine-tuning. Use the skill filters on the jobs page to browse by specific technology.

How fresh are the job listings?

Listings are sourced continuously from company career pages and refreshed automatically. Each job shows when it was posted, and the sitemap and API expose last-updated timestamps for every posting.

How much do AI Engineers get paid?

Compensation varies widely by role, seniority, and location. Where employers disclose a salary range, it is shown directly on the job listing — filter by salary range on the jobs page to narrow results to your target compensation.

Get new AI jobs in your inbox

A weekly digest of the newest AI engineering roles.

© 2026 AI Job Board. All rights reserved.