Deep Learning Jobs
79 open roles mentioning Deep Learning
Researcher, Safety Training, National Security
OpenAI
We are seeking a researcher to train and evaluate models for U.S. government use, with a focus on national security applications. You will advance safety post-training and robustness, helping models follow nuanced policies while preserving their usefulness and capabilities. This role involves researching and implementing methods for safety training, reinforcement learning, and adversarial robustness, as well as developing evaluations, identifying model failure modes, and using findings to improve training. You will also collaborate with research, engineering, security, and policy partners to support safe, reliable deployment.
Member of Technical Staff, Research (2026 PhD New Grad)
fireworks ai
Fireworks is seeking a Member of Technical Staff on the Research team, designed for PhD candidates who want their work to reach production. This role involves pushing the boundaries of generative AI, advancing LLMs and multimodal systems through foundational research. You will enhance model efficiency, accuracy, and scalability, directly shaping high-performance AI infrastructure. You will be paired with a senior researcher as a mentor and given a real research problem from day one, collaborating with top experts in deep learning, distributed systems, and optimization. The work you contribute will be deployed by leading companies, often within weeks.
Anthropic Fellows Program, AI Safety & Security
Anthropic
Anthropic is seeking talented individuals for its Fellows Program, focusing on AI Safety. This program aims to foster AI research and engineering talent by providing funding and mentorship to promising individuals, regardless of prior experience. Fellows will engage in empirical projects using external infrastructure and open-source models, with the goal of producing public outputs like research papers. The program offers a structured 4-month full-time research period, direct mentorship from Anthropic researchers, access to shared workspaces in Berkeley or London, and connections to the AI safety community. The program is designed to encourage diverse perspectives and applications, even from those who may not meet every single qualification.
$25k - $50k
Senior Data Scientist
Cloudflare
The Data Intelligence & Analytics organization builds the core data platform and internal products that power decision-making across the company. We design and operate large-scale data systems, own the company’s data lake, ingestion infrastructure, and platform tooling, and develop end-to-end applications that transform complex datasets into fast, reliable, business-critical products used daily by go-to-market, product, and engineering teams. Our work sits at the intersection of data platforms, distributed systems, and product development, giving engineers the opportunity to own meaningful problems across the stack and build systems that truly run the business. You will focus on building scalable, reliable AI/ML models, services and GenAI powered application backends, partnering closely with data and full-stack engineers to deliver new features and operate the pipelines and platforms behind our products.
Internship - Machine Learning Research Engineer
Perplexity AI
We are seeking motivated interns to join our Machine Learning Research Engineering team in Berlin for a 12-24 week full-time, in-person program. You will have the opportunity to significantly improve search quality by developing and optimizing large-scale deep learning models. This role involves conducting cutting-edge research in representation learning and building advanced RAG pipelines.
$12k - $24k
Researcher, Alignment Interpretability
OpenAI
OpenAI is seeking a researcher passionate about understanding deep networks, with a strong background in engineering, quantitative reasoning, and the research process. You will develop and carry out a research plan in mechanistic interpretability, in close collaboration with a highly motivated team. You will play a critical role in helping OpenAI ensure future models remain safe even as they grow in capability. This will make a significant impact on our goal of building and deploying safe AGI.
Technical Program Manager, Model Performance
Baseten
Baseten is seeking a Technical Program Manager to join its Model Performance organization. This is a unique opportunity to build a program framework from scratch, establishing planning structures, execution processes, and cross-functional alignment for a fast-growing team. You will be instrumental in accelerating the productization of performance R&D, directly impacting how quickly cutting-edge AI models are brought to production. If you excel at transforming ambitious initiatives into predictable, well-governed programs, this role is for you.
Research Engineer, Production Model Post-Training
Anthropic
Anthropic's production models undergo sophisticated post-training processes to enhance their capabilities, alignment, and safety. As a Research Engineer on our Post-Training team, you'll train our base models through the complete post-training stack to deliver the production Claude models that users interact with. You'll work at the intersection of cutting-edge research and production engineering, implementing, scaling, and improving post-training techniques like Constitutional AI, RLHF, and other alignment methodologies. Your work will directly impact the quality, safety, and capabilities of our production models. Note: For this role, we conduct all interviews in Python. This role may require responding to incidents on short-notice, including on weekends.
Research Engineer, Pretraining
Anthropic
Anthropic is seeking a Research Engineer to join its Pretraining team, focusing on developing the next generation of large language models. This role operates at the intersection of cutting-edge research and practical engineering, aiming to build safe, steerable, and trustworthy AI systems. The mission is to ensure that transformative AI systems are aligned with human interests and are beneficial for society. The team is dedicated to pushing the boundaries of AI while prioritizing safety and ethics.
Corporate Development Integration Lead
Anthropic
Anthropic is seeking a Corporate Development Integration Lead to manage the critical post-signing phase of M&A activities, from diligence through operational integration. This high-visibility, cross-functional role will define and scale the integration playbook, build repeatable processes for technical and organizational integration, and work with executive leadership to ensure acquisitions deliver on their strategic thesis. The ideal candidate has deep prior integration experience at high-growth technology companies and brings a structured, hands-on approach to navigating post-deal complexity. You will be energized by turning strategic intent into operational reality, thrive working across organizational boundaries, and understand that integration is where M&A value is ultimately created or destroyed. Join us in building the infrastructure that ensures Anthropic's acquisitions create lasting value for both the mission of safe AI and the brilliant teams we bring into the organization.
Member of Technical Staff (AI Researcher)
Perplexity AI
Perplexity is seeking top-tier AI Research Scientists and Engineers to advance our AI products and capabilities, focusing on building the future of AI-powered search and agent experiences. You will contribute to SOTA experiences that handle hundreds of millions of queries and continue to scale rapidly. Depending on your interests and expertise, you can join one of three specialized teams: the Core Research Team focusing on foundational models, the Agent Products Team fine-tuning models for agent and product experiences, or the Comet Agent Team dedicated to developing and enhancing the Comet Agent product.
Research Engineer / Scientist (Robot Learning)
World-labs
World Labs is a frontier AI research and product company focused on spatial intelligence, co-founded by Dr. Fei-Fei Li, Justin Johnson, and Ben Mildenhall. The company is developing world models that can perceive, generate, reason, and interact with virtual and physical worlds, with its flagship product Marble transforming text, images, and video into navigable 3D worlds. Backed by leading investors, World Labs is building a world-class team at the intersection of AI research and real-world deployment.
$250k - $350k
Senior Applied Research Engineer - Video
synthesia.io
Synthesia is seeking an Applied Research Engineer to join their Video team and contribute to building the next generation of production-grade foundation models for human-centric video generation. This role involves working at the intersection of large-scale generative modeling, distributed systems, and production engineering, with a focus on developing and optimizing video base models for realistic, controllable, and expressive synthetic humans. This is an applied research position with direct product impact, requiring the candidate to advance training recipes, scale distributed systems, improve evaluation frameworks, and optimize inference for real-world deployment. The work will directly influence models used by tens of thousands of businesses globally.
Research Scientist, Life Sciences
Anthropic
Anthropic is seeking an exceptional Research Scientist to join its Life Sciences team. This role focuses on making Claude a superhuman life sciences research assistant, operating at the intersection of machine learning, software engineering, and biology. You will directly improve model capabilities on scientific tasks through post-training, evaluation design, and RL environment development. As a core member, you will translate deep biological domain knowledge into model training objectives, benchmarks, and agentic workflows, helping establish Anthropic as a leader in AI-accelerated biology and shaping how frontier models reason about computational biology tasks. This is a unique opportunity to shape how frontier AI models learn biology, working alongside top AI researchers on problems crucial for human health and scientific understanding.
Research Engineer, Visual Knowledge Work
Anthropic
We are seeking research engineers with a strong computer vision background to enhance the visual and spatial reasoning capabilities of our state-of-the-art Claude models. This role involves research, development, and evaluation, taking a full-stack approach across pretraining, RL, and runtime techniques. You will collaborate closely with the product organization to ensure that vision improvements directly impact Claude's performance on real-world tasks and address customer challenges.
Research Engineer/Research Scientist, Pre-training
Anthropic
Anthropic is seeking a Research Engineer to join its Pre-training team, focusing on developing the next generation of large language models. This role operates at the intersection of cutting-edge research and practical engineering, contributing to the creation of safe, steerable, and trustworthy AI systems. The team is dedicated to ensuring that transformative AI systems are aligned with human interests and societal benefit.
Research Engineer, Production Model Post-Training
Anthropic
Anthropic's production models undergo sophisticated post-training processes to enhance their capabilities, alignment, and safety. As a Research Engineer on our Post-Training team, you'll train our base models through the complete post-training stack to deliver the production Claude models that users interact with. You'll work at the intersection of cutting-edge research and production engineering, implementing, scaling, and improving post-training techniques like Constitutional AI, RLHF, and other alignment methodologies. Your work will directly impact the quality, safety, and capabilities of our production models. For this role, interviews are conducted in Python, and the position may require responding to incidents on short notice, including weekends.
Forward Deployed Engineer (Training)
Baseten
Baseten is seeking a Forward Deployed Engineer (Training) to work directly with leading AI companies, taking ownership of their technical outcomes on the Baseten platform. This role involves tackling complex challenges in serving and improving AI models at scale, spanning the entire model lifecycle from inference to post-training and the systems that connect them. You will act as a technical advisor, guiding customers from initial problem framing through to production deployment, and ensuring the quality and performance of their AI workloads through rigorous evaluation and optimization.
Software Engineer
Applied Intuition
Applied Intuition is a rapidly growing company focused on powering the future of physical AI. We are building the digital infrastructure necessary to bring intelligence to every moving machine across various industries like automotive, defense, and construction. Our solutions are trusted by leading global automakers and military organizations. We are seeking a Software Engineer to join our team and contribute to the design, implementation, and deployment of software and machine learning components, with a particular focus on behavior prediction and environmental interactions. This role involves building scalable software for inference and decision-making in dynamic environments, optimizing data pipelines, and developing robust testing frameworks to ensure system performance and model accuracy.
$60k - $300k
Research Engineer, Mid-Training
Cognition
We are an applied AI lab building end-to-end software agents, known for creating Devin, the first AI software engineer. Our team is composed of highly talented individuals with backgrounds in competitive programming and leadership roles at cutting-edge AI companies. We are tackling significant global challenges and developing AI capable of real-world reasoning. This role focuses on the critical 'mid-training' phase, bridging pre-training and post-training to refine raw model capabilities. You will be instrumental in shaping our models' fundamental abilities by owning late-stage training decisions, including data mix and quality, annealing schedules, context length extension, capability injection, and synthetic data strategies.