Find your next AI Engineering role
2506+ open roles · 71+ companies hiring
2506 open positions
Senior Software Engineer - Together Cloud Infrastructure
Together AI
Together AI is building the AI Acceleration Cloud, an end-to-end platform for the full generative AI lifecycle, combining the fastest LLM inference engine with state-of-the-art AI cloud infrastructure. As a Senior AI Infrastructure Engineer, you will play a key role in building the next generation AI cloud platform – a highly available, global, blazing-fast cloud infrastructure that virtualizes cutting-edge ML hardware and enables state-of-the-art ML practitioners with self-serve AI cloud services. This platform serves both our internal SaaS products and our external cloud customers, spanning dozens of data centers across the world.
$160k - $230k
Senior Platform Engineer, Voice AI
Together AI
Together AI is building the best inference infrastructure for voice applications, powering production-grade, real-time voice agents and applications with best-in-class latency and reliability. We are seeking a Senior Platform Engineer to take ownership of the API and infrastructure layer for voice workloads. You will develop the real-time WebSocket and HTTP APIs used by developers to deploy voice experiences, design autoscaling for latency-sensitive streaming workloads, and ensure the reliability of our multi-provider voice platform for production voice agents handling millions of calls. This is a critical, foundational role on a small, high-impact team, defining how developers interact with our voice platform as we scale.
$200k - $260k
Senior Backend Engineer, Inference Platform
Together AI
Together AI is building the Inference Platform to bring advanced generative AI models to the world, powering multi-tenant serverless workloads and dedicated endpoints. This role offers a unique opportunity to optimize latency and fully utilize tens of thousands of GPUs, working hands-on with cutting-edge hardware. You will collaborate directly with research teams to productionize frontier models and engage with the open-source community, contributing to projects that push the boundaries of inference performance and efficiency.
$160k - $250k
Research Intern, Model Shaping (Fall 2026)
Together AI
As a Research Intern in the Model Shaping team, you will work on advanced post-training methods, new techniques for efficient neural network training, and robust evaluation of foundation model capabilities. The Model Shaping team at Together AI focuses on tailoring open foundation models for downstream applications, building services for machine learning developers, and developing new methods for efficient model training and evaluation. This role offers the opportunity to contribute to cutting-edge research and potentially influence open-source projects.
Lead/Manager Together Cloud Infrastructure
Together AI
Together AI is building the AI Acceleration Cloud, an end-to-end platform for the full generative AI lifecycle, combining the fastest LLM inference engine with state-of-the-art AI cloud infrastructure. As a Lead/Manager, you will play a key role in building the Together cloud platform engineering team in the Netherlands. This platform serves both our internal SaaS products (inference, fine-tuning) and our external cloud customers, spanning dozens of data centers across the world.
Frontier Agents Intern (Fall 2026)
Together AI
The Agents team investigates how to build, align, and scale frontier AI systems capable of complex, multi-step tasks and workflows across text and speech, with a focus on agentic and scientific domains. This role sits at the intersection of agent capabilities, human-computer interaction, and infrastructure, exploring areas like post-training methods for agentic behavior and developing evaluation frameworks for open-ended tasks. As a research intern, you will tackle challenges in alignment, reliability, and scalability, potentially working on new training recipes for self-learning and long-horizon reasoning, curating datasets, studying failure modes, or building scalable agent infrastructure.
Forward Deployed Engineer (Inference & Post-Training)
Together AI
As a Forward Deployed Engineer (FDE) focused on Inference & Post-Training, you will be a hands-on technical partner to strategic customers, assisting production AI teams with leveraging high-quality models and performing inference at scale. You will act as a deep-domain specialist in inference optimization, fine-tuning pipelines, and production deployment, partnering with Solutions Architects. FDEs add significant value by ensuring complex Proofs of Concept (POCs) are met, facilitating platform adoption, and guiding tailored optimization efforts, directly impacting customer success and company growth.
$270k - $300k
Customer Support Engineer (Inference), India
Together AI
As a Customer Support Engineer at a pioneering AI company, you will be the first line of defense supporting customers building training, fine-tuning, and inference solutions. You will dive deep into complex technical challenges, providing swift and effective solutions while serving as a product expert. Collaborating closely with product and sales, you will drive continuous improvement of our offerings. This is an exciting opportunity for a deeply technical professional passionate about AI and customer success to make a significant impact in a fast-paced, innovative environment.
Customer Support Engineer (Inference)
Together AI
As a Customer Support Engineer at a pioneering AI company, you will be the primary point of contact for customers building training, fine-tuning, and inference solutions with Together AI. You will tackle complex technical challenges, provide effective solutions, and act as a product expert. Collaborating with product and sales teams, you will contribute to the continuous improvement of our offerings. This role is ideal for a technically proficient individual passionate about AI and customer success, seeking to make a significant impact in a fast-paced, innovative environment.
$160k - $230k
Backend Software Engineer — Data Platform & AI Data Products
Together AI
You will join the Data Platform team, responsible for building the backend services and data products that power how data moves through the company. This involves creating core platform primitives like high-quality event streams, reliable access layers, and developer-friendly APIs and tools. The goal is to enable teams across the organization to self-serve their data needs and ship faster. You will contribute to backend services that derive value from company data and enhance the self-serve capabilities of the data platform, allowing product and engineering teams to easily create and operate event-driven architectures, publish/consume streams, define access models, and manage data products end-to-end. Additionally, you will work on LLM-adjacent services, including prompt categorization, enrichment, and metadata systems, transforming raw telemetry into trusted, usable products with guidance from experienced engineers.
$120k - $170k
AI Researcher, Core ML (Turbo)
Together AI
The Turbo team operates at the intersection of efficient inference (algorithms, architectures, engines) and post-training/RL systems. We are responsible for building and managing the systems that power Together's API, focusing on high-performance inference and RL/post-training engines capable of operating at production scale. Our core mission is to advance the frontiers of efficient inference and RL-driven training, aiming to make models significantly faster and more cost-effective to run, while simultaneously enhancing their capabilities through RL-based post-training methods. This role involves working across the entire stack, from RL algorithms and training engines to kernels and serving systems, to develop and refine state-of-the-art models using RL pipelines. We value individuals with deep expertise in one area and a strong willingness to collaborate and grow across others.
$200k - $280k
Systems Research Engineer Intern - GPU Programming (Fall 2026)
Together AI
As a Systems Research Engineer Intern specialized in GPU Programming, you will play a crucial role in developing and optimizing GPU-accelerated kernels and algorithms for ML/AI applications. You will co-design GPU kernels and model architecture with the modeling and algorithm team to enhance the performance and efficiency of our AI systems. Collaborating with the hardware and software teams, you will contribute to the co-design of efficient GPU architectures and programming models, leveraging your expertise in GPU programming and parallel computing. Your research skills will be vital in staying up-to-date with the latest advancements in GPU programming techniques, ensuring that our AI infrastructure remains at the forefront of innovation.
Systems Research Engineer, GPU Programming
Together AI
As a Systems Research Engineer specialized in GPU Programming, you will play a crucial role in developing and optimizing GPU-accelerated kernels and algorithms for ML/AI applications. You will co-design GPU kernels and model architecture to enhance the performance and efficiency of our AI systems, and contribute to the co-design of efficient GPU architectures and programming models. Your research skills will be vital in staying up-to-date with the latest advancements in GPU programming techniques, ensuring that our AI infrastructure remains at the forefront of innovation.
$160k - $230k
Staff Machine Learning Engineer, Voice AI
Together AI
Together AI is building the best inference infrastructure for voice applications, powering production-grade, real-time voice agents and applications. We are seeking a Staff ML Engineer to lead the model serving layer for voice workloads. This role involves hands-on optimization of inference engines and models like Whisper, Parakeet, Orpheus, and Kokoro, focusing on pushing latency and throughput boundaries. You will address unique challenges in voice inference, such as streaming audio and real-time latency, and shape the future of how voice models are served as the industry shifts towards end-to-end speech-to-speech systems. This is a foundational hire on a small, high-impact team.
$220k - $280k
Staff Engineer, Distributed Storage and HPC & AI Infrastructure
Together AI
Together AI is seeking a Staff Engineer to design and deliver multi-petabyte storage systems optimized for large-scale AI training and inference workloads. You will architect high-performance parallel filesystems and object stores, integrate cutting-edge technologies, and drive significant cost optimization. The role involves building Kubernetes-native storage operators and self-service platforms for automated provisioning and multi-tenancy. You will focus on optimizing data paths, designing multi-tier caching architectures, and tuning parallel filesystems for AI applications. This is a research-driven role within a company focused on lowering the cost of modern AI systems through co-design of software, hardware, algorithms, and models.
$250k - $300k
Senior Machine Learning Engineer, Voice AI
Together AI
Together AI is building the best inference infrastructure for voice applications, powering production-grade, real-time voice agents and applications. We are seeking a Senior ML Engineer to lead the model serving layer for voice workloads. This role involves hands-on optimization of inference engines and models like Whisper, Parakeet, Orpheus, and Kokoro to achieve frontier-level latency and throughput. You will focus on unique voice inference challenges such as streaming audio, tokenization, and real-time latency budgets, shaping how voice models are served as the industry shifts towards end-to-end speech-to-speech systems. This is a foundational hire on a small, high-impact team.
$200k - $260k
Research Engineer, Frontier Speculative Decoding
Together AI
Together AI is building the Inference Platform that powers the world's most advanced generative AI models. This role will serve as a critical bridge between cutting-edge research and real-world applications, focusing on translating internal model training research into production-ready deployments for customers. The work involves a deep commitment to data-centric development, meticulous hyperparameter tuning, and rigorous checkpoint evaluation. You will transform general-purpose models into highly performant, specialized tools by fine-tuning them on customer-specific data and internal datasets, working with dedicated GPU clusters rather than training foundation models from scratch.
$190k - $270k
Research Engineer, Core ML
Together AI
This research engineering role focuses on translating new Reinforcement Learning (RL) algorithms, scheduling methods, and inference optimizations into production-grade systems that power Together's API. The Core ML team operates at the intersection of efficient inference (algorithms, architectures, engines) and post-training/RL systems, building and maintaining high-performance inference and RL engines at production scale. The goal is to significantly improve model speed, cost-efficiency, and capabilities through RL-based post-training. This position requires a blend of algorithmic understanding and systems engineering, with opportunities to work across the entire stack from RL algorithms and training engines to kernels and serving systems, ultimately driving measurable improvements in latency, throughput, cost, and model quality at scale.
$200k - $280k
Machine Learning, Platform Engineer
Together AI
Together AI is a research-driven artificial intelligence company focused on lowering the cost of modern AI systems. This role is part of a team dedicated to enabling custom models and dedicated inference on Together's platform. The team is responsible for building a container platform, optimizing autoscaling, minimizing cold starts, achieving the best end-to-end model performance, and providing a best-in-class developer experience with great tooling. The work often involves video or audio generation across the stack, including CUDA kernels, PyTorch optimization, inference engines, container orchestration, and queueing theory.
$160k - $250k
Machine Learning Engineer - Inference
Together AI
Together AI is seeking a Machine Learning Engineer to join our Inference Engine team, focusing on optimizing and enhancing the performance of our AI inference systems. This role involves working with state-of-the-art large language models and ensuring they run efficiently and effectively at scale. You will collaborate closely with AI researchers and engineers to create cutting-edge AI solutions and shape the future of AI inference.
$160k - $230k
Machine Learning Engineer
Together AI
Together AI is seeking an ML Engineer to develop systems and APIs for customer inference and fine-tuning of LLMs. The ideal candidate will have experience implementing runtime systems for large-scale AI/ML model inference, including the largest LLMs. This role involves designing and building production systems for reliability and performance at scale, partnering with cross-functional teams, and improving system efficiency and stability.
$160k - $220k
LLM Inference Frameworks and Optimization Engineer
Together AI
Together.ai is building state-of-the-art infrastructure for efficient and scalable inference of large language models (LLMs). The company's mission is to optimize inference frameworks, algorithms, and infrastructure to push the boundaries of performance, scalability, and cost-efficiency. They are seeking an Inference Frameworks and Optimization Engineer to design, develop, and optimize distributed inference engines for multimodal and language models at scale. This role will focus on low-latency, high-throughput inference, GPU/accelerator optimizations, and software-hardware co-design, ensuring efficient large-scale deployment of LLMs and vision models. This position offers a unique opportunity to shape the future of LLM inference infrastructure and ensure scalable, high-performance AI deployment across diverse applications.
$160k - $230k
AI infrastructure Engineer (SRE) Amsterdam
Together AI
Together AI is seeking an AI Infrastructure Engineer (SRE) to ensure the smooth operation of user-facing services and production systems. This role combines the skills of a pragmatic operator and a software engineer, applying sound engineering principles, operational discipline, and automation to our operating environments and codebase. You will specialize in systems such as operating systems, storage subsystems, and networking, while implementing best practices for availability, reliability, and scalability, with interests in algorithms and distributed systems. Join a research-driven artificial intelligence company focused on advancing AI through open and transparent systems, aiming to significantly lower the cost of modern AI systems by co-designing software, hardware, algorithms, and models.
AI Infrastructure Engineer
Together AI
As an AI Infrastructure Engineer at Together, you will be responsible for ensuring the smooth operation of all user-facing services and production systems. This role blends pragmatic operations with software engineering, applying sound engineering principles, operational discipline, and mature automation to our operating environments and codebase. You will specialize in systems (operating systems, storage subsystems, networking), implementing best practices for availability, reliability, and scalability, with varied interests in algorithms and distributed systems. Together AI is a research-driven artificial intelligence company focused on lowering the cost of modern AI systems through co-designing software, hardware, algorithms, and models, and we invite you to join our passionate group of researchers and engineers in building the next generation of AI infrastructure.
$190k - $270k
Head of Policy & Security Research Lab
Scale AI
Scale Labs is seeking a highly experienced, strategic, and mission-driven leader to drive the policy research priorities of the organization. This is a unique opportunity to lead a team of research scientists, policy experts, and engineers focused on foundational AI safety and security work. The role involves owning the day-to-day strategy, direction, and execution of Scale's Policy Research Lab, collaborating with internal researchers and leading labs across governments, industry, and academia. You will drive initiatives related to frameworks and benchmarks for frontier AI models, stay engaged with policy and research communities to set trends, and manage the lab's projects, partnerships, and strategy, including direct work with global AI safety institutes. A key part of the role is finalizing, coordinating, and executing the lab's recruiting plan and research agenda.
$178k - $247k
Strategic Finance Manager, Gen AI
Scale AI
Scale is seeking a high-performing Strategic Finance Manager to support its rapidly growing Generative AI (GenAI) business. This role involves partnering with leadership across Product, Operations, Growth, and Go-to-Market teams to bring financial rigor to decision-making, develop actionable insights, and build scalable financial systems. The ideal candidate will have 4-6 years of experience in a fast-paced, high-growth environment, with a blend of analytical skills, business acumen, and strong execution capabilities. You will be responsible for evolving the GenAI financial forecasting model, supporting performance management, evaluating strategic initiatives, and conducting financial analyses for new ventures. Collaboration with Accounting and Corporate Finance teams to refine financial processes and systems is also a key aspect of this position.
$176k - $221k
Finance Systems & Automations Manager
Scale AI
Scale's Finance Systems and Automation team is seeking a builder-oriented individual to design and develop integrations, automations, and AI agents that streamline workflows across Finance, Accounting, People, and Recruiting. You will collaborate with stakeholders to understand their needs and build scalable solutions, leveraging internal data infrastructure, system integration tooling, and emerging AI platforms. The ideal candidate thrives on connecting systems, automating processes, and advancing AI-assisted operations, with a comfort for ambiguity and a focus on rigorous outcome validation.
$166k - $207k
Software Engineer, Agents & Automations
Cohere
Cohere is seeking a Software Engineer to join the Agents & Automations team, which builds the core platform for North, Cohere's AI workspace. This platform enables customers to create AI-powered workflows, ranging from structured automations to flexible agents that can interact with enterprise systems. The role involves building the workflow builder, execution engine, integrations, debugging tools, and feedback loops that empower users to deploy and manage AI agents and automations. This is a broad product engineering role that spans frontend, backend, and AI-powered systems, with the goal of shipping reliable software that helps customers augment or automate business workflows.
Member of Technical Staff (Software Engineer, Connector Platform)
Perplexity AI
The Connector Platform team is responsible for building the data layer that enables Perplexity's agents to interact with the world's software. This team manages systems that transform hundreds of diverse integrations into a unified, reliable, and well-typed interface for agents. This platform serves as the core knowledge layer, enabling agents to discover, understand, and utilize tools effectively, grounding their reasoning in real, permissioned, and up-to-date enterprise data. By maintaining a knowledge layer above connectors, the platform ensures Computer becomes the single source of truth for institutional knowledge, providing grounded, actionable, and permissioned access to customer systems, which is a key differentiator in the market.
Senior/Staff Software Engineer, Developer Experience
Abridge
Abridge is seeking Engineers to join a new, high-priority agentic engineering team focused on building the foundational infrastructure and internal systems that power software development at the company. This role operates at the intersection of AI tooling, CI/CD, and developer experience, tackling complex problems related to CI/CD pipelines, MCP servers, and AI agents. The work will directly influence how engineers build and deploy software across Abridge, shaping the future of developer productivity and operational efficiency within a rapidly evolving landscape.
Legal Operations Generalist
Applied Intuition
Applied Intuition is seeking a Legal Operations Analyst to join their Legal team in Sunnyvale, California. This role serves as the primary point of contact for legal requests across the company, ensuring smooth operations and follow-through. The analyst will manage a diverse range of tasks including intake triage, signature coordination, system administration, knowledge management, global entity maintenance, and compliance support. A key aspect of this role involves identifying and piloting AI-native approaches to legal operations, leveraging and building enterprise AI to enhance day-to-day efficiency and drive adoption within the team.
$60k - $300k
Software Engineer - Capacity
Baseten
Baseten is seeking a Software Engineer to join the Internal Tooling team. In this role, you will be responsible for the internal operating system that manages Baseten's capacity, balancing supply and demand to unlock revenue. You will own a product end-to-end, working directly with Capacity, Sales, and Engineering teams to define solutions and ship software that streamlines high-stakes workflows. This position is ideal for engineers who enjoy taking full ownership of a product, possess strong product intuition, and are driven to build tools that measurably improve team effectiveness.
AI Strategist, Corporate Law
Hebbia
Hebbia is seeking an AI Strategist to join their AI Strategy Team. This role is crucial for driving the strategic deployment and measurable value creation of Hebbia's AI platform across top global financial institutions. You will shape how AI transforms the enterprise by bridging product, commercial strategy, and customer impact. The ideal candidate will possess curiosity, critical thinking, commercial edge, and executive polish, seamlessly moving between strategic conversations and hands-on execution in a fast-moving environment. If you thrive on solving complex problems, building meaningful partnerships, and helping define how AI reshapes finance, this is an exciting opportunity.
$160k - $225k
Software Engineer, Agent (Cantonese Speaking)
sierra.ai
Sierra is building a platform to enable companies to create better, more human customer experiences with AI. We are seeking a Software Engineer, Agent to design and deliver production-grade AI agents that are central, mission-critical, and drive revenue. You will own the Agent Development Life Cycle (ADLC) from pilot through deployment and iteration, building, tuning, and evolving AI agents in production environments. This role involves partnering directly with leaders at large enterprises and cutting-edge startups to understand their business challenges and build AI agents that transform their operations at scale. Your work will also influence the evolution of Sierra's core platform, surfacing unmet needs and prototyping new tools and features.
Operations Program Manager (Computer Vision), Public Sector
Scale AI
Scale's Public Sector team is rapidly expanding, and you will be instrumental in accelerating the development of AI applications for national security customers. As an Operations Program Manager, you will manage multiple projects within the Computer Vision team, collaborating cross-functionally with Delivery and Engineering teams, as well as operations managers and subject matter experts. Your role will involve ownership of the data labeling system's operations, ensuring timely and high-quality data delivery across diverse modalities and clearance levels, all aimed at supporting our Public Sector customer's AI/ML objectives. You will also contribute to developing and communicating operational improvements to the customer.
$116k - $212k
Full Stack Software Engineer, Codex
OpenAI
We’re hiring a Full Stack Software Engineer to help invent the next generation of AI-powered software development workflows. This role involves owning complete product experiences, spanning user interfaces, workflow orchestration, agent and prompt design, backend systems, and cloud infrastructure. It's a highly product-oriented position where you'll work directly on workflows developers use daily, identifying bottlenecks and rethinking how software gets built in a world where AI agents are active participants. The features you ship will influence how developers around the world write, review, test, and maintain software.
Senior Software Engineer, Frontend
Harvey
Harvey is transforming legal and professional services by combining frontier agentic AI, an enterprise-grade platform, and deep domain expertise. We are a fast-scaling company with strong product-market fit, offering a unique opportunity to build a generational company. Our team is driven, takes ownership, and is committed to our mission, valuing decisiveness, simplicity, and continuous improvement. We are looking for individuals who want to do the best work of their careers alongside like-minded, driven people.
AI Field Engineer - Enterprise
fireworks ai
Fireworks is seeking an AI Field Engineer to join their team. This role is crucial for embedding with ambitious customers and technology partners to rapidly transform complex AI challenges into production systems. You will operate at the intersection of engineering, product, and customer delivery, taking a hands-on approach to building Proofs of Concept (POCs), Minimum Viable Products (MVPs), and production integrations. Simultaneously, you will engage in executive-level discussions regarding architecture, strategy, and business outcomes. The position involves significant coding, running benchmarks, debugging production issues, and architecting deployments, alongside leading discovery conversations, aligning stakeholders, and translating customer needs into product improvements. This role requires comfort working on-site with customers to build relationships and trust in person.
Head of People Operations
Langchain
LangChain is seeking a Head of People Operations to own the People Operations function end-to-end. This role is responsible for architecting and running the essential infrastructure that employees rely on, including compliance, HRIS, benefits, leave programs, global employment, policy, and the automations that streamline these processes. Reporting to the VP of People, this individual will have significant scope and authority, with the opportunity to build a small team as the company scales. This is a hands-on, build-oriented role focused on creating compounding operational systems that automate manual tasks and allow the team to focus on strategic work.
Software Engineering Manager, AI Observability & Evals Platform (New York, NY)
Langchain
LangChain is seeking an Engineering Manager to lead the team responsible for building LangSmith, our observability and evaluation platform for LLM applications. This role involves defining the technical direction, nurturing a talented engineering team, and collaborating with product and design to deliver features that empower developers to build and deploy trustworthy AI systems. You will play a key part in shaping the future of AI observability at scale, contributing to a company that has raised $125M and is experiencing rapid growth.
$200k - $250k
Frequently asked questions
What counts as an AI Engineering job?
AI Job Board lists engineering roles that build, deploy, or operate AI/ML systems — including LLM engineering, RAG (retrieval-augmented generation), AI agents, prompt engineering, AI infrastructure and MLOps, model serving/inference, and fine-tuning. It excludes non-engineering functions like AI sales or marketing roles.
What is the difference between an AI Engineer and a Machine Learning Engineer?
An AI Engineer typically builds applications on top of existing models — LLM integration, RAG pipelines, agent orchestration, and prompt design. A Machine Learning Engineer more often trains, fine-tunes, or productionizes custom models. Many companies use the titles interchangeably, so search both when browsing.
Does AI Job Board include AI infrastructure and MLOps roles?
Yes. Titles such as AI Infrastructure Engineer, ML Platform Engineer, GPU Infrastructure Engineer, Inference Engineer, MLOps Engineer, and AI Site Reliability Engineer are all covered — these roles focus on the systems that serve and scale AI models rather than building models themselves.
Are remote AI and ML jobs available?
Yes. Filter by "Remote" on the jobs page to see fully remote AI/ML engineering roles, or browse the remote jobs feed directly.
What skills are most in demand for AI engineering roles?
The most commonly requested skills are LLM APIs (OpenAI, Anthropic, Gemini), RAG and vector databases (Pinecone, Weaviate, Qdrant), agent frameworks (LangGraph, CrewAI, AutoGen), inference/serving tools (vLLM, Ray Serve, TensorRT-LLM), and fine-tuning. Use the skill filters on the jobs page to browse by specific technology.
How fresh are the job listings?
Listings are sourced continuously from company career pages and refreshed automatically. Each job shows when it was posted, and the sitemap and API expose last-updated timestamps for every posting.
How much do AI Engineers get paid?
Compensation varies widely by role, seniority, and location. Where employers disclose a salary range, it is shown directly on the job listing — filter by salary range on the jobs page to narrow results to your target compensation.