Research Engineer, AGI Safety and Alignment, DeepMind
Summary
Research Engineer at DeepMind working on AI safety and alignment, designing experiments, prototyping systems, and publishing findings in areas like interpretability and agent control.
The Artificial General Intelligence (AGI) Safety and Alignment Team (ASAT) aims to reduce existential and catastrophic risk from AGI and eventually Artificial Superintelligence (ASI). We research novel techniques and work with the rest of GDM and Google to apply them.
Artificial intelligence will be one of humanity’s most transformative inventions. At DeepMind, we are a pioneering AI lab with exceptional interdisciplinary teams focused on advancing AI development to solve complex global challenges and accelerate high-quality product innovation for billions of users. We use our technologies for widespread public benefit and scientific discovery, ensuring safety and ethics are always our highest priority.
US: $174000 - $253000 (USD) + 15% bonus target + equity + benefits
Learn more about benefits at Google.
- Research new alignment methods, studying alignment failures, and applying AGI-scalable alignment techniques to frontier models.
- Develop adversarially robust AGI control systems and implement them in production.
- Research interpretability techniques to understand what AI systems are thinking.
- Work with product teams to ensure that our research is correctly adopted.
Minimum qualifications:
- Bachelor's degree in Computer Science, a related Software Engineering field, or equivalent practical experience.
- 3 years of experience in software development, ML engineering, or ML research.
- Experience working with research teams.
Preferred qualifications:
- Experience conducting or contributing to applied research to improve the safety and alignment of frontier AI systems.
- Experience with training large models (e.g., supervised fine-tuning, RLHF).