Computer Vision Jobs

25 open roles mentioning Computer Vision

Senior/Staff Product Engineer

11d ago
r

runway

We are seeking a curious and thoughtful Product Engineer to join our team and help build cutting-edge AI-powered video creation and collaboration tools. This role involves working in a fast-paced, high-ownership environment within a highly collaborative engineering team. You will tackle a wide variety of exciting challenges, including developing new interfaces for AI content creation, building LLM-powered features and conversational experiences, and deploying cutting-edge computer vision and generative models to production. Our mission is to build AI that simulates the world by merging art and science, believing that world models are the frontier of progress in artificial intelligence.

Remote remote FullTime
AWSTypeScriptJavaScript +2 more

Member of Technical Staff (Software Engineer, Multimodal)

17d ago
P

Perplexity AI

Perplexity is seeking builders to define the future of human-AI interaction, moving beyond text and touch into real-time voice and vision. The Multimodal team is at the forefront of this, developing the experiences and infrastructure that enable AI to understand and respond through speech and sight. As a backend engineer on this team, you will be instrumental in designing and scaling the distributed systems that handle live voice sessions, driving innovation at the intersection of voice, vision, and AI agents. This role offers the opportunity to own problems end-to-end, from initial concept to production launch, working with cutting-edge technology in a fast-paced, entrepreneurial environment.

San Francisco onsite FullTime
AWSPythonGo +2 more

Research Scientist (Generative Modeling)

20d ago
W

World-labs

World Labs is a frontier AI research and product company focused on spatial intelligence, advancing beyond large language models. Co-founded by leading researchers, the company is pioneering world models that perceive, generate, reason, and interact with virtual and physical worlds. Their flagship product, Marble, transforms various media into navigable 3D worlds, with applications in gaming, film, architecture, robotics, and immersive experiences. Backed by significant investment, World Labs is building a world-class team at the intersection of AI research and real-world deployment.

$250k - $325k

San Francisco onsite
PythonFine-TuningPyTorch +3 more

Research Engineer / Scientist (SLAM)

20d ago
W

World-labs

World Labs is a frontier AI research and product company focused on spatial intelligence, co-founded by Dr. Fei-Fei Li, Justin Johnson, and Ben Mildenhall. The company is pioneering world models that perceive, generate, reason, and interact with virtual and physical worlds, with its flagship product Marble transforming text, images, and video into navigable 3D worlds. Backed by leading investors, World Labs is assembling a world-class team at the intersection of AI research and real-world deployment.

$250k - $350k

San Francisco onsite
PythonC#Computer Vision

Research Engineer / Scientist (3D Tech Lead)

20d ago
W

World-labs

World Labs is a frontier AI research and product company focused on spatial intelligence, co-founded by Dr. Fei-Fei Li, Justin Johnson, and Ben Mildenhall. The company is developing world models that can perceive, generate, reason, and interact with virtual and physical worlds, with its flagship product Marble transforming text, images, and video into navigable 3D worlds. Backed by leading investors, World Labs is building a world-class team at the intersection of AI research and real-world deployment. We are seeking a Tech Lead for 3D Modeling & Reconstruction to define technical direction and drive execution for our core 3D modeling efforts. This hands-on leadership role requires a strong research science or engineering background with significant contributions to 3D reconstruction and/or modeling. You will combine deep technical expertise with the ability to guide a high-impact team, shaping both the research roadmap and production systems.

$250k - $350k

San Francisco onsite
PythonC#Computer Vision

Pipeline Engineer (Graphics/3D)

20d ago
W

World-labs

World Labs is a frontier AI research and product company focused on spatial intelligence, developing world models that perceive, generate, reason, and interact with virtual and physical worlds. Their flagship product, Marble, transforms various inputs into navigable 3D worlds for applications in gaming, film, architecture, robotics, and immersive digital experiences. Backed by significant investment, World Labs is building a world-class team at the intersection of AI research and real-world deployment. This role involves building a production-grade web application for 3D Gaussian Splat scene generation, editing, and publishing. It's a high-ownership, full-stack role with a backend focus, bridging R&D and frontend. You will work across graphics/ML algorithms, backend services, and frontend UI, transforming proof-of-concepts into reliable, debuggable, and delightful shipped features. The ideal candidate thrives on making complex systems work smoothly in production and continuously improving them based on user feedback.

$200k - $325k

San Francisco onsite
DockerKubernetesPython +2 more

Research Engineer, Visual Knowledge Work

24d ago
Anthropic

Anthropic

We are seeking research engineers with a strong computer vision background to enhance the visual and spatial reasoning capabilities of our state-of-the-art Claude models. This role involves research, development, and evaluation, taking a full-stack approach across pretraining, RL, and runtime techniques. You will collaborate closely with the product organization to ensure that vision improvements directly impact Claude's performance on real-world tasks and address customer challenges.

New York City, NY; San Francisco, CA; Seattle, WA onsite
AnthropicFine-TuningClaude +3 more

AI Content Engineer

1mo ago
l

llamainndex ai

We are seeking a highly technical ML engineer who can produce compelling, authentic technical content at high velocity. You will combine deep expertise in document AI with strong writing skills to build benchmarks, publish technical analyses, and establish our position as the definitive leader in document understanding. This is not a traditional DevRel or Marketing role. You will write real code, build real benchmarks, and run real experiments - then translate that work into published content at a pace far faster than academic publishing. Your output will directly drive awareness and adoption among the developers building the next generation of document-powered applications.

San Francisco hybrid FullTime
LlamaIndexPythonRAG +2 more

Product Manager

2mo ago
A

Applied Intuition

Applied Intuition is seeking product managers to play a pivotal role in the growth of its agentic platform for physical AI, Dana. This role involves ownership at both the platform level for core capabilities and deep hands-on execution within specific verticals. You will collaborate directly with customers, including top OEMs and autonomy developers, as well as internal autonomy teams across various domains. This is a high-visibility position requiring a strong sense of ownership and product direction to drive significant outcomes.

$60k - $300k

Sunnyvale onsite FullTime
GoComputer Vision

Member of Technical Staff, Applied Research

2mo ago
l

llamainndex ai

We are seeking an AI Research Engineer to join our document understanding team, bridging the gap between applied research and robust engineering. In this role, you will focus on vision-language models, document processing, data curation, synthetic data generation, benchmarking, and model training and fine-tuning. The primary objective is to enhance the accuracy, speed, and cost-effectiveness of our document AI systems in production. This position requires a passion for cutting-edge AI research coupled with a strong drive for practical product impact, involving rapid prototyping, rigorous evaluation, and the transition of promising approaches into production systems.

San Francisco hybrid FullTime
LlamaIndexPythonFine-Tuning +4 more

Computer Vision AI & ML Engineer

2mo ago
S

Skild

Skild AI is seeking a Computer Vision AI & ML Engineer to design, build, and deploy advanced perception systems for real-world robotics and automation. You will work across the full machine learning lifecycle—model development, data strategy, evaluation, and production integration—to deliver robust, high-performance vision capabilities. This role combines applied research with hands-on engineering and offers the opportunity to influence both architecture and roadmap decisions.

Bengaluru, India onsite
PythonC#PyTorch +5 more

Operations Program Manager (Computer Vision), Public Sector

3mo ago
Scale AI

Scale AI

Scale's Public Sector team is rapidly expanding, and you will be instrumental in accelerating the development of AI applications for national security customers. As an Operations Program Manager, you will manage multiple projects within the Computer Vision team, collaborating cross-functionally with Delivery and Engineering teams, as well as operations managers and subject matter experts. Your role will involve ownership of the data labeling system's operations, ensuring timely and high-quality data delivery across diverse modalities and clearance levels, all aimed at supporting our Public Sector customer's AI/ML objectives. You will also contribute to developing and communicating operational improvements to the customer.

$116k - $212k

St. Louis, MO; Washington, DC onsite
Computer VisionAIMachine Learning +2 more

Program Manager (Homeland Layered Defense), Public Sector

3mo ago
Scale AI

Scale AI

Scale's public sector business is focused on delivering cutting-edge agentic AI to orchestrate portfolio management for homeland defense. We are seeking a Program Manager (PgM) to lead and own the execution for Scale's portfolio of clients dedicated to the layered defense of the United States. This is a hands-on leadership role requiring deep technical understanding, strong customer relationship management, and the ability to drive change in fast-moving, mission-driven environments. You will be instrumental in solving complex problems by working closely with our engineering team.

$233k - $291k

Washington, DC onsite
PythonGoSQL +3 more

Technical Program Manager (Computer Vision), Public Sector

3mo ago
Scale AI

Scale AI

Scale AI is seeking a Technical Program Manager (TPM) to lead and coordinate the delivery of computer vision workflows for a national security customer. In this role, you will own or support customer account plans, address technical issues, and utilize data to align internal resources with customer needs. You will drive the creation of tools that directly benefit Scale's Public Sector customers, manage executive-level to end-user relationships, and collaborate with customers to define computer vision use cases for our engineering team. You will also lead cross-functional teams to achieve AI/ML objectives and proactively identify customer needs to ensure success.

$166k - $249k

Washington, DC hybrid
Computer VisionAI/MLData analytics +1 more

Computer Vision AI & ML Engineer

3mo ago
S

Skild

Skild AI is seeking a Computer Vision AI & ML Engineer to design, build, and deploy advanced perception systems for real-world robotics and automation. You will work across the full machine learning lifecycle—model development, data strategy, evaluation, and production integration—to deliver robust, high-performance vision capabilities. This role combines applied research with hands-on engineering and offers the opportunity to influence both architecture and roadmap decisions.

San Mateo, CA onsite
PythonC#PyTorch +5 more

Product Manager, Data Engine

3mo ago
Scale AI

Scale AI

Scale AI is seeking a technical Product Manager to lead the evolution of the Public Sector Data Engine. This role is focused on building tools to measure and improve computer vision and generative AI models critical for national security systems. You will architect the vision for ML Ops tooling, creating a foundational engine for data management, curation, model development, and evaluation. This hybrid role requires a PM who understands both the strategic 'why' and the technical 'how,' bridging the gap between data labeling and full-stack AI partnership.

$178k - $266k

San Francisco, CA; St. Louis, MO; New York, NY; Washington, DC hybrid
Computer VisionAIMachine Learning +2 more

Research Engineer, Multimodal Reasoning For Information Literacy

4mo ago
Google DeepMind

Google DeepMind

Google DeepMind's research team is dedicated to tackling complex challenges in online information quality, advancing the state of the art by developing innovative solutions to detect manipulated media and misleading narratives. The team leverages interdisciplinary work spanning provenance analysis and the creation of tools for AI-assisted information literacy, with a focus on ensuring the integrity of digital discourse and a safer online environment. This role involves researching and building multimodal reasoning systems and Vision-Language Models (VLMs) to assess the trustworthiness of media on the internet, with a passion for advancing information literacy using machine learning and computational techniques.

Mountain View, California, US onsite
PythonPrompt EngineeringPyTorch +5 more

Researcher, Vision

4mo ago
s

sarvam

Sarvam is building the bedrock of Sovereign AI for India, developing a full-stack sovereign AI platform across research, models, infrastructure, and applications. The company partners with leading enterprises and public institutions, backed by prominent venture capital firms and collaborating with India's top brands. This role involves working across the full lifecycle of vision-language model (VLM) development, including data, training, evaluation, and production. We seek researchers who are comfortable with the evolving nature of the field and can take a leading role.

Bengaluru onsite FullTime
PyTorchNLPRLHF +1 more

Research Engineer, Post-training & Deployment

5mo ago
S

Skild

Skild AI is seeking a Research Engineer to join our post-training team. In this role, you will be responsible for enhancing Skild's foundation models and deploying them onto robots in real-world scenarios. You will collaborate with customers and utilize deployment data to ensure reliable robot behavior, focusing on safety, efficiency, and robustness under operational constraints. This position bridges the gap between strong lab performance and dependable customer deployments, defining the standard for scaling autonomous systems into global infrastructure.

San Mateo, CA onsite
PythonPyTorchTensorFlow +3 more

Research Scientist, Multimodal Alignment, Safety, and Fairness

6mo ago
Google DeepMind

Google DeepMind

Google DeepMind's Frontier AI unit is seeking experienced Research Scientists to join a multimodal safety research effort. This role focuses on interdisciplinary sociotechnical modeling and requires a passion for understanding AI-society interactions, a strong awareness of AI alignment and safety, and a drive to develop novel ideas, methods, interfaces, and tools. You will contribute to advancing the state of the art in AI research and Google DeepMind's mission towards Artificial General Intelligence (AGI), with a focus on leading new breakthrough research directions in areas like AI behavior exploration, assessment, and steering, particularly for subjective and creative tasks. The work involves tackling fundamental research questions to improve alignment objectives, assess adherence to desired behaviors, and enable AI agents to monitor real-world social context and evolve system behaviors over long time-horizons. You will develop new paradigms for human+AI rating that are adaptive and context-aware, driving breakthroughs within Google DeepMind, Google products, and the broader AI alignment community.

Kirkland, Washington, US; Mountain View, California, US; New York City, New York, US onsite
PythonAI AgentsFine-Tuning +5 more

Get new AI jobs in your inbox

A weekly digest of the newest AI engineering roles.

© 2026 AI Job Board. All rights reserved.