ONNX Jobs
7 open roles mentioning ONNX
Senior Machine Learning Engineer
Cloudflare
You will help define how machine learning models run across Cloudflare’s global network, from frontier open LLMs and real-time voice models to customer-deployed models served on heterogeneous GPUs and next-generation accelerators. You’ll work with systems engineers, product teams, hardware partners, and AI/ML engineers to bring models into production with low latency, strong reliability, and efficient resource use. This role combines applied ML, inference optimization, evaluation, and production engineering, with a focus on benchmarking models, improving serving performance, validating quality, and building tooling that helps Cloudflare and its customers ship AI applications at Internet scale.
Perception & Fusion Engineer
Applied Intuition
Applied Intuition is seeking a Perception & Fusion Engineer to join their team in developing, integrating, and maintaining real-time sensor software solutions for autonomous vehicles across various domains. You will design, develop, and integrate sensor fusion algorithms to interface with different platforms and support autonomy team objectives. This role involves developing autonomous sensor controllers to fuse information from state-of-the-art sensor systems like EO, IR, acoustics, radar, and RF.
$10k - $20k
AI Performance Engineer
Applied Intuition
Applied Intuition is seeking a performance engineer to specialize in making large-scale machine learning workloads fast and cost-efficient within the datacenter. This role focuses on optimizing distributed training runs across multiple nodes and high-throughput batch inference for processing vast amounts of real-world autonomy logs. The primary goal is to improve throughput, cluster efficiency, and reduce cost per unit of data processed, directly impacting the company's iteration speed. You will be responsible for identifying and resolving performance bottlenecks across the entire stack, from accelerators to ML frameworks and data infrastructure, working collaboratively with various engineering teams to achieve significant improvements in training time and processing costs.
$60k - $300k
Robotic Software Engineer, Perception
Applied Intuition
Applied Intuition is seeking a Perception Autonomy Engineer to develop, integrate, and maintain real-time sensor software solutions for autonomous vehicles across various domains. You will work with a team to enhance capabilities and demonstrate solutions in real-world scenarios on diverse hardware platforms. The role involves designing, developing, and integrating AI/ML sensor algorithms to interface with different platforms, ingest mission information, and support autonomy objectives. This includes fusing information from state-of-the-art sensors like EO, IR, acoustics, radar, and RF to create autonomous sensor controllers.
$10k - $20k
Backend Engineer - Studio Media Platform
sarvam
Sarvam is building the bedrock of Sovereign AI for India, developing a full-stack AI platform across research, models, infrastructure, and applications. We partner with leading enterprises and public institutions, backed by prominent venture capital firms. We are hiring a Backend Engineer to work on our Studio media platform, which includes AI dubbing, live translation, and foundational services for voice cloning, stem separation, lip sync, and music generation. You will build and maintain production services, ML pipeline libraries, and platform SDKs to enable multilingual media processing at scale for our enterprise customers and Studio users.
Software Engineer - Voice AI (Inference Runtime)
Baseten
Baseten is seeking a highly impactful individual to lead the Voice AI product area, owning the end-to-end development and implementation of their in-house inference stack for Voice AI models. This role involves partnering closely with various engineering teams to push the boundaries of Voice AI, making a significant impact on industries like productivity, customer service, and education. You will be responsible for bringing state-of-the-art open-source models into production, focusing on optimizing model serving for latency, throughput, and GPU efficiency, and building large-scale, real-time infrastructure for multi-model voice agents.
Embedded AI Engineer – Android Automotive (On-Device Intelligence)
Applied Intuition
Applied Intuition is seeking an Embedded AI Engineer to build on-device intelligence for a next-generation Android Automotive platform. This role is responsible for the entire lifecycle of embedded ML systems, ensuring models perform predictably and safely within production environments, adhering to real-world constraints like latency, thermal limits, and functional safety. The company is a leader in powering the future of physical AI, serving industries such as automotive, defense, and construction with solutions for tools, infrastructure, operating systems, and autonomy.
$60k - $300k