×
Register Here to Apply for Jobs or Post Jobs. X

Research Engineer, Safety Oversight, DeepMind

Job in Mountain View, Santa Clara County, California, 94043, USA
Listing for: Google
Full Time position
Listed on 2026-07-27
Job specializations:
  • Software Development
    Data Scientist, Machine Learning/ ML Engineer, AI Engineer (Applied/Software), Data Science Manager
Salary/Wage Range or Industry Benchmark: 174000 - 253000 USD Yearly USD 174000.00 253000.00 YEAR
Job Description & How to Apply Below

Research Engineer, Safety Oversight

We aim to turn production data into intelligence on the safety of deployed AI models. Safety Oversight is a new team tasked with using large-scale production traffic and a variety of automated evaluation methods to monitor the safety and alignment of deployed models. Our work will ensure we measure the real-world efficacy of our safety stack—both of in-model safety training and out-of-model safety mitigations to ensure we are effective in our goal of deploying safe models that are used for widespread public benefit.

The Safety Oversight team sits within the GenAI safety organization and is accountable for ensuring that when a model safety issue occurs in production, or when a user is misusing our model at scale, we detect and understand it, so the safety risk can be rapidly mitigated. We will collaborate closely with teams working on safety training and evaluation for Gemini and Gen Media models.

The Generative AI (GenAI) Safety team operates in a fast-paced, highly collaborative environment. We take the possibility of tangibly dangerous model capabilities seriously as AI advances, and we believe that proactive monitoring and deployment-time oversight are critical for safe AI development.

Artificial intelligence will be one of humanity's most transformative inventions. At Deep Mind, we are a pioneering AI lab with exceptional interdisciplinary teams focused on advancing AI development to solve complex global challenges and accelerate high-quality product innovation for billions of users. We use our technologies for widespread public benefit and scientific discovery, ensuring safety and ethics are always our highest priority.

We are pushing the boundaries across multiple domains. Our global teams offer varied learning opportunities and varied career pathways for those driven to achieve exceptional results through collective effort.

Individual pay is determined by factors including job-related skills, experience, and relevant education or training. US: $174000 - $253000 (USD) + 15% bonus target + equity + benefits

Responsibilities:

  • Build classifiers and large-scale data pipelines to detect model misbehavior and misuse end-to-end.
  • Research and develop cross-context monitoring systems to detect coordinated harms, developing novel signal aggregation methods across disparate user sessions to identify large-scale attack vectors.
  • Think critically about novel methods for monitoring using model activations, actions, chains-of-thought and final answers.
  • Collaborate closely with infrastructure teams and data scientists to scale your work and regularly share results with the wider safety team.
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary