Head of AI Red Teaming
Listed on 2026-08-20
-
Software Development
Software Engineer, AI Engineer (Applied/Software)
You'll manage our growing red-team, own delivery for frontier lab customers, and build the automation that allows us to scale.
This is a player-coach role: you'll split your time between writing code, setting strategy, and working with frontier lab researchers and expert red teamers.
We're an early-stage startup, so you should expect and enjoy that your responsibilities will grow and priorities change quickly.
About usOur mission is to automate AI safety , to pave the way for a future where the vast majority of AI safety work is done by AI models.
Frontier models already solve coding problems that take humans days, but a model that can be hijacked by a malicious email or web page can't be trusted to work on its own. Before AI can do the work that matters, including AI safety research itself, models have to be robust to attack. So we build the safety and alignment evals, red-teaming programs, and RL environments that find these failures and train them out.
You'll run our prompt injection red-teaming line end to end. Frontier labs want more of this work than we can currently deliver; your job is to turn that demand into safer models.
You'll:
- Manage the red team. Line-manage our red teamers, set priorities, own the QA bar, and significantly expand the team to meet demand.
- Automate the pipeline. Use coding agents to turn manual red-teaming workflows into systems that allow for >10x more impact per person.
- Own delivery for frontier labs. Scope what each customer cares about most, keep engagements on track, and be accountable for the final results.
- Set the strategy. Own pricing, packaging, and what we build next, working with our founding engineer on the technical roadmap.
You'll work closely with our founders and the teams developing frontier models, with unusual autonomy to make consequential decisions. The work you lead shapes system cards, deployment safeguards, and how much the world can trust the most capable AI systems.
What we're looking for- People and project management
: you can hire, motivate, and run a distributed team, and keep multiple customer engagements on track at once. - Technical fluency
: you can write code and build automations, and you use LLMs and coding agents as core tools. - Ownership
: you take responsibility for outcomes rather than tasks, and push until the work is done. - Analytical judgment
: you can take a messy question and turn it into a sound decision with transparent reasoning . - Early-stage startup drive
: you enjoy a fast pace, shifting priorities, limited structure, and taking on whatever will have the biggest impact. - Cyber security experience or knowledge
- Professional software engineering experience
- Experience building evals, red-teaming programs, or model-training data
- Product management, TPM, or sales experience
- Familiarity with AI safety and the alignment research community
- Experience collaborating with frontier labs or other demanding technical customers
These criteria are a guide, not a checklist.
- Location and workspace
:
Our team works out of Constellation in Berkeley, CA, and we prefer someone who can work alongside us there. We are open to remote for the right candidate. - Compensation
: $250,000–$400,000 plus equity. More for exceptional candidates. - Health coverage
:
You'll receive a generous monthly pre-tax allowance to choose the medical, dental, and vision coverage that best fits your needs. - 401(k):
We offer a 401(k) retirement plan. - Visa sponsorship
:
We sponsor visas, although we cannot successfully sponsor every visa for every role or candidate. If we make you an offer, we will make every reasonable effort to secure the visa you need, with support from an immigration lawyer we retain to guide and coordinate the process.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).