DeepMind logo

Research Engineer, AGI Safety and Alignment, DeepMind

DeepMind·London·Posted today

Google DeepMind is the London-headquartered AI lab behind AlphaGo and AlphaFold, working towards artificial general intelligence. They're hiring a Research Engineer in AGI Safety and Alignment to build software and experiments that make advanced AI systems safer, based in London. This AI engineering role sits within the UK's artificial intelligence sector.

Minimum qualifications:

  • Bachelor's degree in Computer Science, a related Software Engineering field, or equivalent practical experience.
  • 3 years of experience in software development, ML engineering, or ML research.
  • Experience working with research teams.

Preferred qualifications:

  • Experience conducting or contributing to applied research to improve the safety and alignment of frontier AI systems.
  • Experience with training large models (e.g., supervised finetuning, RLHF).

About the job

The Artificial General Intelligence (AGI) Safety and Alignment Team (ASAT) aims to reduce existential and catastrophic risk from AGI and eventually Artificial Superintelligence (ASI). We research novel techniques and work with the rest of GDM and Google to apply them. We advise executive leadership on safety.


ASAT has sub-teams specialising in making future Geminis more thoroughly aligned by finding and fixing sources of misalignment and exploring alignment techniques with better generalization. Preparing for future AGI risks by simulating them today and using interpretability techniques to understand AI and solve practical problems like model forensics or eval awareness. Building control for GDM’s agents as defense-in-depth against potential misaligned internal deployments. Researching training techniques, like debate, for aligning superhuman AI and ways to retain, improve, and measure monitorability. Researching and implementing ways to assess the ways in which a given model might be imperfectly aligned and developing and implementing tools and AI assistance that accelerates safety research. Advising executive leadership on risks posed by AI systems via the frontier safety framework based on our threat models and evaluations.

We are prioritising hires for deep alignment, alignment stress testing, language model interpretability, agent control, and amplified oversight. We are looking to grow our team with researchers and engineers. Depending on your background, we have opportunities available as both Research Scientists and Software Engineers.

Artificial intelligence will be one of humanity’s most transformative inventions. At Google DeepMind, we are a pioneering AI lab with exceptional interdisciplinary teams focused on advancing AI development to solve complex global challenges and accelerate high-quality product innovation for billions of users. We use our technologies for widespread public benefit and scientific discovery, ensuring safety and ethics are always our highest priority.

We are pushing the boundaries across multiple domains. Our global teams offer diverse learning opportunities and varied career pathways for those driven to achieve exceptional results through collective effort.

Responsibilities

  • Research new alignment methods, studying alignment failures, and applying AGI-scalable alignment techniques to frontier models.
  • Develop adversarially robust AGI control systems and implement them in production.
  • Research interpretability techniques to understand what AI systems are ‘thinking’.
  • Work with product teams to ensure that our research is correctly adopted.
DeepMind logo

About DeepMind

Google DeepMind is an AI research laboratory focused on building artificial general intelligence. Founded in London in 2010 by Demis Hassabis, Shane Legg, and Mustafa Suleyman, the company was acquired by Google in 2014 for approximately $660M. DeepMind pioneered breakthroughs including AlphaGo, the first program to defeat a world champion at Go, and AlphaFold, which solved the 50-year-old protein folding problem and has been used by over two million researchers worldwide. The lab remains headquartered in London with a large UK engineering presence and continues to push the boundaries of reinforcement learning, large language models, and scientific AI.

Stage: AcquiredLondon
View DeepMind profile →