Researcher, Alignment Interpretability
San Francisco • FullTime
Posted 12d ago
Remote Work Policy
On-site
Employment Type
FullTime
Categories
AI Research Engineer
About the job
OpenAI is seeking a researcher passionate about understanding deep networks, with a strong background in engineering, quantitative reasoning, and the research process. You will develop and carry out a research plan in mechanistic interpretability, in close collaboration with a highly motivated team. You will play a critical role in helping OpenAI ensure future models remain safe even as they grow in capability. This will make a significant impact on our goal of building and deploying safe AGI.
Responsibilities
- Develop and publish research on techniques for understanding representations of deep networks.
- Engineer infrastructure for studying model internals at scale.
- Collaborate across teams on projects that OpenAI is uniquely suited to pursue.
- Guide research directions toward demonstrable usefulness and/or long-term scalability.
Requirements
- Ph.D. or research experience in computer science, machine learning, or a related field.
- Experience in AI safety & alignment, mechanistic interpretability, or spiritually related disciplines.
- 2+ years of research engineering experience.
- Proficiency in Python or similar languages.
- Excitement about OpenAI’s mission of ensuring AGI benefits all of humanity.
- Enthusiasm for long-term AI safety & alignment, and deep thought about technical paths to safe AGI.
- Thrive in environments involving large-scale AI systems.