multimodal Jobs
3 open roles mentioning multimodal
Machine Learning Research Scientist, Post-Training
Scale AI
Scale works with leading AI labs to accelerate progress in GenAI research, focusing on optimizing data curation and evaluation to enhance LLM capabilities in text and multimodal modalities. This role involves developing novel methods to improve the alignment and generalization of large-scale generative models, collaborating with researchers and engineers on best practices in data-driven AI development, and providing technical and strategic input to foundation model labs for the next generation of AI models.
$252k - $315k
ML Research Engineer, ML Systems
Scale AI
Scale's ML platform (RLXF) team builds our internal distributed framework for large language model training and inference. This platform powers MLEs, researchers, data scientists, and operators for fast and automatic training and evaluation of LLMs, as well as data quality evaluation. You will work closely across Scale’s ML teams and researchers to build the foundation platform that supports all our ML research and development, optimizing it to enable the next generation of LLM training, inference, and data curation. If you are excited about shaping the future of AI via fundamental innovations, we would love to hear from you!
$190k - $237k
Multimodal Generative AI Researcher
Stability AI
We are seeking a Research Scientist with deep expertise in training and fine-tuning large Vision-Language and Language Models (VLMs / LLMs) for downstream multimodal tasks. You will be instrumental in advancing models that reason across vision, language, and 3D, translating research breakthroughs into scalable engineering solutions. This role involves designing and fine-tuning large-scale VLMs/LLMs and hybrid architectures for complex tasks like visual reasoning, retrieval, 3D understanding, and embodied interaction.