Overview
Research Engineer for Anthropic’s Pretraining team, building the next generation of large language models with a focus on safe, steerable, and trustworthy AI.
What you'll do
- Research and implement solutions across model architecture, algorithms, data processing, and optimizer development.
- Independently lead small research projects while collaborating on larger initiatives.
- Design, run, and analyze scientific experiments for large language models.
- Optimize and scale training infrastructure to improve efficiency and reliability.
- Develop and improve developer tooling to enhance team productivity.
- Contribute across the full stack, from low-level optimizations to high-level model design.
What you'll need
- Advanced degree (MS or PhD) in Computer Science, Machine Learning, or related field.
- Strong software engineering skills with a proven track record of building complex systems.
- Python expertise and experience with deep learning frameworks (PyTorch preferred).
- Familiarity with large-scale machine learning, especially language models.
- Ability to balance research goals with practical engineering constraints.
- Excellent communication skills and ability to work collaboratively.
Details
- Hybrid policy: expected to be in one of the offices at least 25% of the time.
Read the full description and apply on the company’s own careers page.