Spark Jobs

61 open roles mentioning Spark

Capacity Strategy & Operations

3d ago
B

Baseten

Baseten is seeking a Capacity Strategy & Operations professional to manage the end-to-end capacity planning process. This role involves translating customer commitments and growth forecasts into supply requirements, coordinating fulfillment across various teams and vendors, and building scalable systems. You will be responsible for modeling complex scenarios, making clear recommendations, and driving alignment in a fast-paced hardware market to ensure a predictable and reliable foundation for customers and internal engineering teams. The ideal candidate is comfortable with both analytical modeling and cross-functional execution, particularly in ambiguous and rapidly evolving environments.

San Francisco hybrid FullTime
Spark

AI Engineer

3d ago
B

Baseten

Baseten is seeking an AI Engineer to join its hyper-growth Compute organization. This role is crucial for scaling the systems and workflows that balance supply and demand across Baseten's GPU fleet. You will design, build, and deploy AI-powered workflows to automate repetitive and error-prone tasks within capacity management, enabling the team to focus on critical judgment calls. The ideal candidate will audit existing systems, identify gaps, and rapidly ship solutions, understanding when to leverage existing tools versus building custom AI-powered applications. This position requires a systems-thinking approach to ensure new workflows integrate seamlessly with broader capacity systems architecture.

San Francisco hybrid FullTime
GoClaudeSpark

Software Engineer - Voice Model

5d ago
x

xAI

SpaceXAI's mission is to create AI systems that can accurately understand the universe and aid humanity in its pursuit of knowledge. The Grok Voice Model team is building the world's best voice AI, delivering smooth, natural, low-latency spoken interactions that are expressive, multilingual, and reliable across devices and real-time scenarios. The team owns the full training pipeline, from massive data curation and premium audio processing to frontier speech-language pre-training and intensive post-training to push quality, speed, and stability to the limit. The goal is to make talking to AI feel like conversing with a charming, kind, and knowledgeable person, and exceptionally smart, execution-oriented engineers are sought to achieve this.

$150k - $450k

Palo Alto, CA onsite
KubernetesPythonFine-Tuning +5 more

Software Engineer, Data Platform

5d ago
x

xAI

The Data Platform team builds and operates the infrastructure responsible for all large-scale data transport and processing across the company. We own and manage core systems including Apache Kafka, HDFS, Spark, Flink, and Trino, enabling real-time ML pipelines, feed ranking, experimentation, analytics, and observability at petabyte scale. Our team deals with latency-critical workloads, high-throughput streaming, and distributed compute systems that require fault tolerance, performance, and absolute reliability. As a software engineer on the Data Platform team, you will design, build, and operate the distributed systems powering SpaceXAI's data movement and compute. You will take ownership of infrastructure components that process trillions of events daily, driving the scalability, performance, and reliability of the systems that power product and ML workloads across the company.

$180k - $440k

Palo Alto, CA onsite
GoRustSpark +2 more

Member of Technical Staff - Multimodal Understanding

5d ago
x

xAI

SpaceXAI is seeking a Member of Technical Staff to join their multimodal team and advance the understanding and generation of AI across image, video, audio, and text. This role involves working across the full stack, from data curation and pre-training to alignment, infrastructure, and end-to-end product experiences. You will collaborate with various teams to deliver cutting-edge multimodal reasoning, world modeling, tool use, agentic behaviors, and human-AI collaboration capabilities. The goal is to build models that can perceive, reason about, and interact with the world in real-time at an unprecedented level.

$180k - $440k

Palo Alto, CA onsite
KubernetesPythonRust +5 more

Member of Technical Staff - Imagine Model

5d ago
x

xAI

SpaceXAI is seeking a multimodal engineer for the Imagine Model Team to develop cutting-edge AI experiences beyond text, focusing on high-fidelity understanding and generation across image and video modalities, with audio integration where it enhances visual content. Responsibilities cover data curation, modeling, training, inference serving, and product integration across pretraining and post-training phases. The role involves close collaboration with product teams to advance model frontiers and deliver exceptional end-to-end user experiences.

$180k - $440k

Palo Alto, CA onsite
PythonRustC# +3 more

Senior Data Scientist

10d ago
C

Cloudflare

The Data Intelligence & Analytics organization builds the core data platform and internal products that power decision-making across the company. We design and operate large-scale data systems, own the company’s data lake, ingestion infrastructure, and platform tooling, and develop end-to-end applications that transform complex datasets into fast, reliable, business-critical products used daily by go-to-market, product, and engineering teams. Our work sits at the intersection of data platforms, distributed systems, and product development, giving engineers the opportunity to own meaningful problems across the stack and build systems that truly run the business. You will focus on building scalable, reliable AI/ML models, services and GenAI powered application backends, partnering closely with data and full-stack engineers to deliver new features and operate the pipelines and platforms behind our products.

Hybrid hybrid
LangGraphLangChainPython +5 more

Principal Software Engineer, Data

10d ago
C

Cloudflare

Cloudflare is seeking an experienced, product-minded Principal Software Engineer to join the Town Lake team. This team is responsible for making vast amounts of data accessible and valuable across the company, building a modern, agentic-first data lakehouse platform. You will play a key role in taking this fast-growing internal platform to become the foundational data infrastructure for the entire company. This is a high-impact, high-visibility role where you will lead technical architecture and drive implementation for critical services, leveraging AI deeply to accelerate development and create innovative solutions. If you are a builder who thrives on solving complex problems and shaping the future of data platforms, this opportunity is for you.

Hybrid hybrid
KubernetesPythonTypeScript +5 more

Data Infrastructure Engineer, Pre-training

10d ago
Anthropic

Anthropic

Anthropic is seeking a Staff level Engineer to join our Pre-training team, responsible for developing the next generation of large language models. In this role, you will work at the intersection of cutting-edge research and practical engineering, contributing to the development of safe, steerable, and trustworthy AI systems. Our mission is to ensure that transformative AI systems are aligned with human interests.

San Francisco, CA onsite
AnthropicPythonRust +2 more

Software Engineer - Identity & Authorization

11d ago
B

Baseten

Baseten powers mission-critical inference for leading AI companies, enabling them to bring cutting-edge models into production. We are seeking a founding engineer for our identity and authorization team within enterprise engineering. This role will own the identity and access layer of the Baseten platform, including the authorization model, credential systems, and admin experiences for enterprise IT teams. You will design and build a fine-grained authorization system from the ground up, ensuring low-latency permission checks at high request volumes, consistent behavior across the product suite, and strong security guarantees for critical workloads.

San Francisco hybrid FullTime
KubernetesPythonGo +1 more

Member of Technical Staff (Search Quality Analyst)

12d ago
P

Perplexity AI

Perplexity is seeking an experienced analyst to contribute to the development and enhancement of its core search technologies. This position operates at the nexus of data analysis and engineering, involving the creation of metrics, the construction of data pipelines, and the refinement of search and answer systems.

Belgrade hybrid FullTime
PythonSQLSpark

ML Infra Engineer, Data Systems

13d ago
P

Physical Intelligence

Physical Intelligence is seeking an ML Infra Engineer specializing in Data Systems to build and operate the data infrastructure powering large-scale robot learning. This role is crucial for bridging raw data sources and training/evaluation processes, aiming to accelerate development while ensuring performance, correctness, and reliability at scale. It's a systems-focused position at the intersection of distributed systems, storage, and machine learning infrastructure, contributing to the foundational platforms that enable large-scale learning.

San Francisco onsite FullTime
Spark

Technical Program Manager, Model Performance

15d ago
B

Baseten

Baseten is seeking a Technical Program Manager to join its Model Performance organization. This is a unique opportunity to build a program framework from scratch, establishing planning structures, execution processes, and cross-functional alignment for a fast-growing team. You will be instrumental in accelerating the productization of performance R&D, directly impacting how quickly cutting-edge AI models are brought to production. If you excel at transforming ambitious initiatives into predictable, well-governed programs, this role is for you.

San Francisco hybrid FullTime
Deep LearningSpark

Member of Technical Staff - Engineering Lead, Data Platform

18d ago
Reflection ai

Reflection ai

Reflection is building the trusted data backbone for the world's most capable open-weight AI systems. The Data Platform team builds and operates the core data systems and pipelines that power our research, training, and production environments, unifying ingestion, processing, and orchestration across the entire data lifecycle. We are seeking a front-line technical leader to build, mentor, and grow this team, guiding its technical direction while remaining hands-on. This role is for someone who has earned deep technical credibility as an engineer and now multiplies it through a team, working closely with our research team to understand the needs of high-velocity experimentation and translate them into reliable, reproducible, scalable data infrastructure.

San Francisco, CA onsite FullTime
AirflowSparkKafka

Staff Software Engineer, Data Platform

19d ago
Harvey

Harvey

Harvey is seeking a Staff Software Engineer to join their central data platform team. This role is crucial for building the foundational systems that enable all teams at Harvey to work with data confidently and independently. You will be responsible for creating frameworks, tooling, and paved paths for product engineers, data engineers, and analysts, focusing on ingestion, warehousing, transformation, and real-time processing. A key aspect of this role involves ensuring data quality, lineage, governance, and handling sensitive data like PII and multi-region residency requirements, which are critical for serving security-conscious institutions.

$231k - $340k

New York hybrid FullTime
AWSAzureKubernetes +5 more

Staff Software Engineer, Data Platform

19d ago
Harvey

Harvey

Harvey is seeking a Staff Software Engineer to join their central data platform team. This role is crucial for building the foundational systems that enable all teams at Harvey to work with data confidently and independently. You will be responsible for creating frameworks, tooling, and paved paths for product engineers, data engineers, and analysts, focusing on ingestion, warehousing, transformation, and real-time processing. A key aspect of this role involves ensuring data quality, lineage, governance, and handling sensitive data like PII and multi-region residency requirements, which are critical for serving security-conscious institutions.

$231k - $340k

San Francisco hybrid FullTime
AWSAzureKubernetes +5 more

Senior Software Engineer, Data Platform

19d ago
Harvey

Harvey

Harvey is seeking a Senior Software Engineer to join our central data platform team. This role is crucial for building the systems that empower all teams at Harvey to work with data confidently and independently. You will be responsible for creating frameworks, tooling, and paved paths for product engineers, data engineers, and analysts, focusing on their leverage and trust in our data systems. The initial focus will be on establishing a reliable foundation for ingestion and warehousing, including streaming and batch processing into Snowflake, CDC, orchestration, and schema evolution, while adhering to strict data sensitivity requirements. The role will expand to encompass transformation and compute frameworks, self-serve tooling, real-time stream processing, and robust quality, lineage, and governance layers, all while ensuring PII handling and multi-region data residency are satisfied by design.

$193k - $290k

New York hybrid FullTime
AWSAzureKubernetes +5 more

Senior Software Engineer, Data Platform

19d ago
Harvey

Harvey

Harvey is seeking a Senior Software Engineer to join our central data platform team. This role is crucial for building the systems that empower all teams at Harvey to work with data confidently and independently. You will be responsible for creating frameworks, tooling, and paved paths for product engineers, data engineers, and analysts, focusing on their leverage and trust in our data systems. The initial focus will be on establishing a reliable foundation for ingestion and warehousing, including streaming and batch processing into Snowflake, CDC, orchestration, and schema evolution, while adhering to strict data sensitivity requirements. The role will expand to encompass transformation and compute frameworks, self-serve tooling, real-time stream processing, and robust quality, lineage, and governance layers, all while ensuring PII handling and multi-region data residency are satisfied by design.

$193k - $290k

San Francisco hybrid FullTime
AWSAzureKubernetes +5 more

Senior Software Engineer, Data Infrastructure

20d ago
d

decagon

Decagon is seeking a Senior Data Infrastructure Engineer to design, build, and operate the data systems that power the company's AI products. This role involves owning critical data pipelines and storage layers end-to-end, enhancing reliability and performance, and establishing streamlined processes for engineers to confidently work with data at scale. The ideal candidate will contribute to building and maintaining the foundational infrastructure for a leading conversational AI platform.

$200k - $400k

San Francisco onsite FullTime
AWSAzureKubernetes +5 more

Research Engineer, Discovery

24d ago
Anthropic

Anthropic

As a Research Engineer on our team, you will work end-to-end across the entire model stack, identifying and addressing key infrastructure blockers on the path to scientific AGI. You should have familiarity with elements of language model training, evaluation, and inference, and be eager to quickly dive into and get up to speed in areas where you are not yet an expert. This may include performance optimization, distributed systems, VM/sandboxing/container deployment, and large-scale data pipelines. Join us in our mission to develop advanced AI systems that push the frontiers of science and benefit humanity.

San Francisco, CA onsite
AnthropicAWSDocker +5 more

Get new AI jobs in your inbox

A weekly digest of the newest AI engineering roles.

© 2026 AI Job Board. All rights reserved.