Spark Jobs

61 open roles mentioning Spark

Software Engineer - Capacity

3mo ago
B

Baseten

Baseten is seeking a Software Engineer to join the Internal Tooling team. In this role, you will be responsible for the internal operating system that manages Baseten's capacity, balancing supply and demand to unlock revenue. You will own a product end-to-end, working directly with Capacity, Sales, and Engineering teams to define solutions and ship software that streamlines high-stakes workflows. This position is ideal for engineers who enjoy taking full ownership of a product, possess strong product intuition, and are driven to build tools that measurably improve team effectiveness.

San Francisco hybrid FullTime
AWSJavaScriptPostgreSQL +1 more

Product Manager, Developer Experience

3mo ago
B

Baseten

Baseten is seeking a Product Manager to define and build the developer experience for their AI inference platform. This is a foundational role where you will work directly with founders and engineers to shape the product function from the ground up. You will own the entire developer journey, from initial sign-up to deploying and iterating on AI models, aiming to make Baseten synonymous with great developer experience. The ideal candidate is deeply technical, customer-obsessed, and driven to ship impactful product experiences, not just strategy or documentation.

San Francisco hybrid FullTime
GoAI AgentsSpark

Software Engineer, Context Engine

3mo ago
s

sierra.ai

Sierra is building a platform to enable companies to create better, more human customer experiences with AI. We are seeking a Software Engineer to join our full-stack data team. You will be instrumental in building systems that enhance the intelligence of our AI agents with every interaction. This role involves developing real-time pipelines, analytics products, and personalization primitives to transform customer conversations into business outcomes. You will contribute to architecting and building our core data platform for low-latency, large-scale data handling, including real-time eventing, streaming and batch ETL, and our data lakehouse. Additionally, you will develop systems for long-term agent memory, reusable Customer Data Platform (CDP) primitives, and optimization loops for personalized interactions that demonstrably improve business results.

San Francisco, CA onsite FullTime
OpenAIPythonTypeScript +5 more

Software Engineer - Baseten Inference Stack

3mo ago
B

Baseten

Baseten is seeking a Software Engineer to join their Inference Stack team. This team builds the distributed runtime that powers large-scale LLM inference across Baseten's platform, operating at the intersection of distributed systems, model performance, infrastructure, and developer experience. The role involves working across the entire stack, from customer-facing deployment tools and feature libraries to the underlying systems for orchestrating Kubernetes deployments and routing traffic. This is an ideal opportunity for engineers who thrive on owning production systems, solving complex integration challenges, and simplifying intricate infrastructure for users.

San Francisco hybrid FullTime
KubernetesSpark

Staff Software Engineer, Data Platform

3mo ago
Scale AI

Scale AI

Scale is at the forefront of the AI revolution, developing data engines and technologies that power the world's leading LLMs and generative models. This role is on the Platform Engineering team, responsible for the foundational data infrastructure that supports these cutting-edge AI products. You will lead the design and development of core data storage, streaming, caching, and indexing platforms, gaining exposure to the rapidly evolving AI landscape across various industries. The work involves driving architecture, implementation, and reliability of these critical systems, collaborating with stakeholders, and mentoring junior engineers.

$252k - $315k

San Francisco, CA; New York, NY remote
KubernetesPythonFine-Tuning +10 more

ML Engineer (Data), Foundational Models

3mo ago
s

sarvam

Sarvam is building India's sovereign AI platform, focusing on research, models, infrastructure, and applications to make AI work for India. This role is critical in owning the data infrastructure that fuels our foundational models. You will be responsible for building petabyte-scale curation and filtering pipelines, designing systems for training data selection and proportioning, and ensuring data quality with research-level rigor. This is an engineering- and research-intensive position involving large-scale deduplication, quality modeling, contamination detection, mixture design, curriculum learning, attribution, and debugging.

Bengaluru onsite FullTime
PythonSpark

Engineering Manager, Forward Deployed Engineering (LLM)

4mo ago
B

Baseten

Baseten is seeking an Engineering Manager (Player & Coach) to lead a team of Forward Deployed Engineers focused on building, scaling, and optimizing LLM inference workloads for their customers. This role involves both hands-on technical ownership and managerial leadership, guiding the team in designing, deploying, and managing high-performance, low-latency AI applications on Baseten’s platform. The Forward Deployed Engineering team contributes to the core Baseten codebase, influences the feature roadmap, and executes complex customer engagements. You will also collaborate with product, infrastructure, and other customer engineering teams to ensure generative AI systems deliver best-in-class performance, reliability, and cost efficiency.

San Francisco hybrid FullTime
DockerPythonHugging Face +1 more

Software Engineer - Voice AI (Inference Runtime)

4mo ago
B

Baseten

Baseten is seeking a highly impactful individual to lead the Voice AI product area, owning the end-to-end development and implementation of their in-house inference stack for Voice AI models. This role involves partnering closely with various engineering teams to push the boundaries of Voice AI, making a significant impact on industries like productivity, customer service, and education. You will be responsible for bringing state-of-the-art open-source models into production, focusing on optimizing model serving for latency, throughput, and GPU efficiency, and building large-scale, real-time infrastructure for multi-model voice agents.

San Francisco hybrid FullTime
DockerKubernetesPython +5 more

Product Manager, Inference Platform

5mo ago
B

Baseten

Baseten is seeking a Product Manager to define and build the product function for its inference platform, which powers mission-critical AI inference for leading AI companies. This role offers a unique opportunity to shape the future of AI infrastructure, working directly with founders and top engineers. You will own the product surface responsible for making AI model inference fast, reliable, and economical at scale, including autoscaling, traffic routing, failover, and workload scaling across clusters and regions. The ideal candidate is deeply technical, customer-obsessed, and enjoys tackling foundational infrastructure challenges to make AI systems scalable and reliable.

San Francisco hybrid FullTime
KubernetesSpark

Software Engineer, Systems Generalist

5mo ago
thinkingmachines

thinkingmachines

Thinking Machines Lab is seeking generalist infrastructure and systems engineers to build the core systems powering their foundation models and support internal research and product development teams. This high-impact role involves architecting and scaling critical infrastructure across the full technical stack, solving complex distributed systems problems, and building robust, scalable platforms. You will work directly with researchers to accelerate experiments, improve infrastructure efficiency, and enable key insights across models, products, and data assets.

$350k - $475k

San Francisco onsite
OpenAIMistralKubernetes +5 more

Solutions Architect

6mo ago
B

Baseten

Baseten is seeking a Solutions Architect to partner with Sales and customers, translating business needs into technical solutions. This role involves running technical discovery, guiding deployments, and proving value for clients. It's ideal for entrepreneurial, customer-facing technical professionals who want to see how companies adopt AI at scale and enjoy working across discovery, solution design, demos, deployment scoping, and hands-on implementations. You will work closely with Sales and Engineering to help build the platform engineers use to ship AI products.

San Francisco hybrid FullTime
SparkEmbeddings

Software Engineer - GPU Networking & Distributed Systems

6mo ago
B

Baseten

Baseten is building the global operating system for distributed, heterogeneous AI hardware, powering mission-critical inference for leading AI companies. As LLM and multi-modal workloads scale, the network becomes a critical component. This role focuses on leading GPU Networking efforts, making RDMA a first-class building block and optimizing distributed inference. You will architect the software fabric that unifies thousands of GPUs, co-optimizing communication and computation for advanced AI workloads.

San Francisco hybrid FullTime
KubernetesPythonGo +3 more

Software Engineer, Data

7mo ago
H

Heyge

HeyGen is seeking a Software Engineer with data engineering responsibilities to bridge the gap between core application development and large-scale data infrastructure. You will help build the data foundational layers for our next-generation features, enabling AI models to function in real-time, building robust pipelines for multimedia, and powering engaging user experiences. This role is crucial for developing cutting-edge features like PPT-to-video converters and interactive, conversational video capabilities.

$180k - $220k

Los Angeles, Palo Alto, San Francisco onsite
AWSPythonGo +5 more

Software Engineer - Training Product

7mo ago
B

Baseten

Baseten is seeking a customer-obsessed software engineer to join their team and contribute to the development of mission-critical AI inference platforms. In this role, you will own features from conception to launch, working across the entire technology stack from API and UI down to the infrastructure layer. You will have the opportunity to fine-tune models, gain a deep understanding of user workflows, and collaborate closely with research engineers to build cutting-edge experiences that accelerate model development and address real-world pain points. If you are excited about diving deep into AI model training and building impactful products, this is the role for you.

San Francisco hybrid FullTime
KubernetesFine-TuningPyTorch +3 more

Software Engineer - Model Performance Systems

8mo ago
B

Baseten

Baseten is seeking Software Engineers to join their team in a specialized, high-impact role at the intersection of high-performance computing (HPC) and Large Language Model (LLM) engineering. This position involves not only building automated performance monitoring and diagnostic tools for next-generation AI infrastructure but also defining the roadmap, driving key technical decisions, and taking full ownership of the future direction of this work. The role offers the opportunity to shape the platform that engineers use to deploy cutting-edge AI models into production.

San Francisco hybrid FullTime
PythonC#PyTorch +3 more

Senior Member of Technical Staff, Synthetic Data

9mo ago
Cohere

Cohere

Cohere is seeking a Senior Machine Learning Engineer specializing in synthetic data to develop and manage the synthetic data pipeline crucial for advanced language models. This role involves end-to-end management of synthetic data, including pipeline optimization, data analysis and generation, and conducting data ablations and model evaluations. You will transform diverse web and code data using generative models to enhance token efficiency and model quality, bridging research and engineering to improve throughput and accelerator utilization. This position is key to Cohere's mission of delivering efficient and reliable language capabilities and driving innovation in natural language processing.

Toronto remote FullTime
CoherePythonNLP +1 more

Software Engineer - Model Products

11mo ago
B

Baseten

Baseten is seeking a Software Engineer to join their Model Performance team, focusing on the infrastructure that powers hosted API endpoints for cutting-edge open-source models. This role involves working on distributed systems, model serving, and developer experience to ensure models running on the Baseten platform are fast, reliable, and cost-efficient. You will contribute to defining how developers interact with AI models at scale, joining a high-impact team at the intersection of product, model performance, and infrastructure.

San Francisco hybrid FullTime
KubernetesSparkModel Serving

Software Engineer - Enterprise Platform

1y ago
B

Baseten

Baseten powers mission-critical inference for leading AI companies, enabling them to bring cutting-edge models into production. We are seeking an Engineer for our enterprise engineering team to build capabilities that support large organizations with their security, compliance, and procurement needs. This role involves deep product and systems work across the full stack, focusing on core building blocks like identity and access management, billing, regional isolation, and deployment options. You will design authentication and authorization systems and develop administrative experiences for enterprise IT teams.

San Francisco hybrid FullTime
KubernetesSpark

Software Engineer - GPU Kernels

1y ago
B

Baseten

Baseten is seeking a GPU Kernel Engineer to join our team at the forefront of AI acceleration. In this role, you will craft the foundational code that powers modern AI workloads, optimizing every microsecond of computation to enable breakthrough applications. You will work in a fast-paced, intellectually stimulating environment where technical excellence is paramount and your contributions directly influence production systems serving millions of users across numerous products. This role offers exceptional growth potential for engineers passionate about low-level optimization and high-impact systems work.

San Francisco hybrid FullTime
AWSC#Spark +1 more

Software Engineer, Data

1y ago
Mistral AI

Mistral AI

Mistral AI is a pioneering company focused on democratizing AI through high-performance, open-source models and solutions. We are seeking passionate and skilled software engineers to join our dynamic, collaborative team. In this role, you will design, build, and maintain our data infrastructure, ensuring data accuracy, accessibility, and security. Your contributions will be crucial in enabling our science teams to enhance AI model quality and empowering business users to make informed decisions.

Paris remote Full-time
MistralPythonSQL +8 more

Get new AI jobs in your inbox

A weekly digest of the newest AI engineering roles.

© 2026 AI Job Board. All rights reserved.