About the Role
Google DeepMind is at the forefront of artificial intelligence research. We are seeking an AI Safety and Alignment Researcher to develop empirical and theoretical frameworks for ensuring advanced AI systems remain safe, robust, and aligned with human values.
Responsibilities
- Conduct fundamental and applied research into model interpretability, robustness, and scalable oversight.
- Design rigorous evaluation benchmarks to test model vulnerabilities against jailbreaks, prompt injection, and deceptive alignment.
- Partner with product and engineering teams to integrate safety guardrails into deployment pipelines.
- Collaborate with academic institutions and external research bodies on safety standards.
Requirements
- Advanced degree (Ph.D. preferred) in Computer Science, Mathematics, Physics, or Philosophy with a heavy quantitative focus.
- Strong background in machine learning safety, alignment, interpretability, or adversarial robustness.
- Proficiency in Python and deep learning frameworks (TensorFlow or PyTorch).
- Excellent communication skills with a proven track record of peer-reviewed publications.
Benefits
- Industry-leading compensation including base salary and equity/stock options.
- Comprehensive health, wellness, and family care benefits.
- Access to world-class computing infrastructure and research resources.
- Generous sabbatical and time-off policies.
#J-18808-Ljbffr