ONNX Jobs

7 open roles mentioning ONNX

Senior Machine Learning Engineer

10d ago
C

Cloudflare

You will help define how machine learning models run across Cloudflare’s global network, from frontier open LLMs and real-time voice models to customer-deployed models served on heterogeneous GPUs and next-generation accelerators. You’ll work with systems engineers, product teams, hardware partners, and AI/ML engineers to bring models into production with low latency, strong reliability, and efficient resource use. This role combines applied ML, inference optimization, evaluation, and production engineering, with a focus on benchmarking models, improving serving performance, validating quality, and building tooling that helps Cloudflare and its customers ship AI applications at Internet scale.

Hybrid hybrid
PythonRAGPyTorch +5 more

Perception & Fusion Engineer

19d ago
A

Applied Intuition

Applied Intuition is seeking a Perception & Fusion Engineer to join their team in developing, integrating, and maintaining real-time sensor software solutions for autonomous vehicles across various domains. You will design, develop, and integrate sensor fusion algorithms to interface with different platforms and support autonomy team objectives. This role involves developing autonomous sensor controllers to fuse information from state-of-the-art sensor systems like EO, IR, acoustics, radar, and RF.

$10k - $20k

Sunnyvale onsite FullTime
DockerPythonC# +3 more

AI Performance Engineer

1mo ago
A

Applied Intuition

Applied Intuition is seeking a performance engineer to specialize in making large-scale machine learning workloads fast and cost-efficient within the datacenter. This role focuses on optimizing distributed training runs across multiple nodes and high-throughput batch inference for processing vast amounts of real-world autonomy logs. The primary goal is to improve throughput, cluster efficiency, and reduce cost per unit of data processed, directly impacting the company's iteration speed. You will be responsible for identifying and resolving performance bottlenecks across the entire stack, from accelerators to ML frameworks and data infrastructure, working collaboratively with various engineering teams to achieve significant improvements in training time and processing costs.

$60k - $300k

Sunnyvale onsite FullTime
KubernetesPythonGo +4 more

Robotic Software Engineer, Perception

2mo ago
A

Applied Intuition

Applied Intuition is seeking a Perception Autonomy Engineer to develop, integrate, and maintain real-time sensor software solutions for autonomous vehicles across various domains. You will work with a team to enhance capabilities and demonstrate solutions in real-world scenarios on diverse hardware platforms. The role involves designing, developing, and integrating AI/ML sensor algorithms to interface with different platforms, ingest mission information, and support autonomy objectives. This includes fusing information from state-of-the-art sensors like EO, IR, acoustics, radar, and RF to create autonomous sensor controllers.

$10k - $20k

Sunnyvale onsite FullTime
DockerPythonC# +3 more

Backend Engineer - Studio Media Platform

3mo ago
s

sarvam

Sarvam is building the bedrock of Sovereign AI for India, developing a full-stack AI platform across research, models, infrastructure, and applications. We partner with leading enterprises and public institutions, backed by prominent venture capital firms. We are hiring a Backend Engineer to work on our Studio media platform, which includes AI dubbing, live translation, and foundational services for voice cloning, stem separation, lip sync, and music generation. You will build and maintain production services, ML pipeline libraries, and platform SDKs to enable multilingual media processing at scale for our enterprise customers and Studio users.

Bengaluru onsite FullTime
AWSAzureDocker +5 more

Software Engineer - Voice AI (Inference Runtime)

4mo ago
B

Baseten

Baseten is seeking a highly impactful individual to lead the Voice AI product area, owning the end-to-end development and implementation of their in-house inference stack for Voice AI models. This role involves partnering closely with various engineering teams to push the boundaries of Voice AI, making a significant impact on industries like productivity, customer service, and education. You will be responsible for bringing state-of-the-art open-source models into production, focusing on optimizing model serving for latency, throughput, and GPU efficiency, and building large-scale, real-time infrastructure for multi-model voice agents.

San Francisco hybrid FullTime
DockerKubernetesPython +5 more

Embedded AI Engineer – Android Automotive (On-Device Intelligence)

5mo ago
A

Applied Intuition

Applied Intuition is seeking an Embedded AI Engineer to build on-device intelligence for a next-generation Android Automotive platform. This role is responsible for the entire lifecycle of embedded ML systems, ensuring models perform predictably and safely within production environments, adhering to real-world constraints like latency, thermal limits, and functional safety. The company is a leader in powering the future of physical AI, serving industries such as automotive, defense, and construction with solutions for tools, infrastructure, operating systems, and autonomy.

$60k - $300k

Sunnyvale onsite FullTime
C#TensorFlowTransformers +2 more

Get new AI jobs in your inbox

A weekly digest of the newest AI engineering roles.

© 2026 AI Job Board. All rights reserved.