flash attention Jobs

3 open roles mentioning flash attention

Machine Learning Systems Research Engineer, Agent Post-training - Enterprise GenAI

2mo ago
Scale AI

Scale AI

Scale is seeking a Machine Learning Systems Research Engineer to join their Enterprise ML Research Lab. This role will focus on building algorithms for a next-generation Agent RL training platform, supporting large-scale training, and integrating state-of-the-art technologies to optimize ML systems. You will collaborate with other ML researchers and engineers who apply these algorithms to client use cases, including AI cybersecurity firewalls and healthtech search models. If you are passionate about shaping the future of AI, this is an exciting opportunity to contribute to cutting-edge advancements in enterprise GenAI.

$265k - $331k

San Francisco, CA; New York, NY onsite
LLMPyTorchTransformers +7 more

Tech Lead Manager- MLRE, ML Systems

2mo ago
Scale AI

Scale AI

Scale's LLM post-training platform team builds our internal distributed framework for large language model training, powering MLEs, researchers, data scientists, and operators for fast and automatic training and evaluation of LLMs. This platform also serves as the underlying training framework for the data quality evaluation pipeline. You will work closely with Scale’s ML teams and researchers to build the foundation platform which supports all our ML research and development works, optimizing it to enable next generation LLM training, inference, and data curation. If you are excited about shaping the future AI via fundamental innovations, we would love to hear from you!

$265k - $331k

San Francisco, CA; New York, NY onsite
LLMPyTorchTransformers +7 more

ML Research Engineer, ML Systems

2mo ago
Scale AI

Scale AI

Scale's ML platform (RLXF) team builds our internal distributed framework for large language model training and inference. This platform powers MLEs, researchers, data scientists, and operators for fast and automatic training and evaluation of LLMs, as well as data quality evaluation. You will work closely across Scale’s ML teams and researchers to build the foundation platform that supports all our ML research and development, optimizing it to enable the next generation of LLM training, inference, and data curation. If you are excited about shaping the future of AI via fundamental innovations, we would love to hear from you!

$190k - $237k

San Francisco, CA; Seattle, WA; New York, NY onsite
LLMPyTorchTransformers +7 more

Get new AI jobs in your inbox

A weekly digest of the newest AI engineering roles.

© 2026 AI Job Board. All rights reserved.