×
Register Here to Apply for Jobs or Post Jobs. X
More jobs:

Member of Technical Staff, Security Engineering

Job in Berkeley, Alameda County, California, 94709, USA
Listing for: Metr
Full Time position
Listed on 2026-09-09
Job specializations:
  • IT/Tech
    Cybersecurity
Salary/Wage Range or Industry Benchmark: 180000 - 260000 USD Yearly USD 180000.00 260000.00 YEAR
Job Description & How to Apply Below

About METR

We are a nonprofit research organization that develops scientific methods to assess AI capabilities, risks, and mitigations, with a specific focus on threats related to AI R&D automation and misalignment.

We believe it is robustly good for policymakers and civil society to have a clear understanding of risks from AI systems, and we are extremely excited to build a team of ambitious, excellent people to tackle one of the most important challenges of our time.

Security at METR is becoming its own dedicated team, and you would be one of its first hires. It is extremely important that we continue to be an organization that frontier AI labs, governments, and the public trust with sensitive model access and confidential information. As misalignment incidents become more extreme and confidential information about models and frontier AI labs becomes more valuable, we expect to be under increasingly heavy pressure.

For us, security encompasses managing endpoints and securing development environments, cloud platform security, safely sandboxing agents and evaluations, VPN and VPC networking, application code reviews, account provisioning and access control, and helping ensure we use the best practices across all of our workflows.

  • Offensive security: You would be the first person on the team with an offensive security background. You'll run targeted red-team exercises against our own systems and build automated AI red teaming.

  • High-context detection and response: You will build AI systems that can quickly triage and respond to threats, both from internal agents and external attackers.

  • Blue-team engineering: Detection engineering, telemetry pipelines, incident response, and hardening across our cloud infrastructure, endpoints, and identity systems.

  • Securing a unique attack surface: METR's evaluation infrastructure runs frontier AI agents. In the past, we've run pre-deployment model evaluations - executing untrusted, model-generated code at scale on multi-day tasks.

  • Enabling bleeding-edge research: You'll work closely with our researchers to make dangerous-capability experiments safe to run. We often face extreme reward hacking and evaluation awareness during our pre-deployment evaluations, and expect internal threats from agents to become more extreme.

  • METR handles some of the most sensitive artifacts in AI - pre-release frontier model access, confidential lab information, and transcripts with raw chain-of-thought. Labs and policymakers trust us with this because of our security posture, and keeping that trust is necessary for everything else we do.

  • As AI agents are used more aggressively by malicious actors for cyber offense operations, and METR's salience rises in the public eye, we expect to face increasingly sophisticated attacks. Strengthening security at METR can be one of the highest-leverage roles to ensure third parties continue to have access to confidential information necessary to inform the world about current risks.

  • METR is one of the first organizations to see and closely study misalignment incidents that involve models breaking out of sandboxes, attacking our infrastructure, manipulating graders, and more. We also may pursue incident investigations embedded in frontier labs, in which case internal experience with similar failures will be critical.

  • Deep security expertise: You have strong fundamentals across systems, networks, cloud, and identity.
  • Offensive security: You have experience acting like an attacker, whether through red teaming, penetration testing, or adversarial research.
  • AI/LLM engineering: You build with AI: agent pipelines, LLM-powered tooling, automated workflows, and understand current limitations of those tools.
  • AWS: You should know AWS very well, including a deep understanding of IAM policies.

We don't screen on certifications, degrees, or years of experience.

  • Detection engineering at scale: Experience with SIEM/detection pipelines, writing and tuning detections, and threat hunting.

  • Cloud and container security: AWS (especially non-trivial IAM), Kubernetes, and infrastructure-as-code environments.

  • Incident response: You've led or worked severe incidents, ideally those involving AI agents.

  • AI security research: Familiarity with prompt injection, agent containment, model supply-chain risks, or red teaming AI systems themselves.

Ideally you have experience with a good portion of these technologies:

  • AWS: cloud-native software platforms
  • EKS
  • Lambda
  • ECS
  • IAM (in-depth)
  • SQS
  • Cloud Watch
  • Security Hub & Guard Duty
  • PostgreSQL: RLS, serverless Aurora
  • Pulumi: IaC
  • Data Dog: SIEM
  • Okta: IdP
  • Google Workspace: IdP
  • Tailscale: networking
  • Crowd Strike Falcon
    : endpoint security
Our Culture

METR is a mission-driven organization. We believe our work can meaningfully shape humanity's future for the better, and we want to be the best people in the world doing this work. We have a tight-knit, collaborative research culture rooted in truth-seeking and integrity. We're fiercely committed to producing high-quality, trustworthy science. We're…

To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary