Researcher, Training
San Francisco • FullTime
Posted 1y ago
Remote Work Policy
On-site
Employment Type
FullTime
Categories
AI Research Engineer
About the job
OpenAI's Training team is responsible for producing the large language models that power our research and products, bringing us closer to AGI. This involves deep research into improving current architectures, datasets, and optimization techniques, alongside long-term bets on future model efficiency and capability. As a member of the architecture team, you will push the frontier of architecture development for OpenAI's flagship models, enhancing intelligence, efficiency, and adding new capabilities. The ideal candidate has a deep understanding of LLM architectures, model inference, and a hands-on empirical approach, comfortable with creative breakthroughs, strengthening baselines, designing evaluations, debugging regressions, and tracking bottlenecks.
Responsibilities
- Design, prototype, and scale new architectures to improve model intelligence.
- Execute and analyze experiments autonomously and collaboratively.
- Study, debug, and optimize both model and computational performance.
- Contribute to training and inference infrastructure.
Requirements
- Deep understanding of LLM architectures.
- Sophisticated understanding of model inference.
- Hands-on empirical approach.
- Experience landing contributions to major LLM training runs.
- Ability to thoroughly evaluate and improve deep learning architectures in a self-directed fashion.
- Motivation for safely deploying LLMs in the real world.
- Well-versed in state-of-the-art transformer modifications for efficiency.