×
Register Here to Apply for Jobs or Post Jobs. X

Intern: AI Red Teaming; Fall

Job in Sunnyvale, Santa Clara County, California, 94087, USA
Listing for: Realm Labs
Apprenticeship/Internship position
Listed on 2026-08-28
Job specializations:
  • IT/Tech
    AI Engineer (Applied/Software), Machine Learning/ ML Engineer, Data Scientist
Salary/Wage Range or Industry Benchmark: 34000 - 55000 USD Yearly USD 34000.00 55000.00 YEAR
Job Description & How to Apply Below
Position: Intern: AI Red Teaming (Fall 2026)

Intern: AI Red Teaming (Fall 2026)

Sunnyvale, CA

AI/ML

In office

Full-time

Role Overview
  • You will try to break the systems we build and the models we protect, and turn what you find into something the team can act on: a reproducible attack, an evaluation that catches it, and a written account of why it works.
  • Expect some mix of: eliciting unsafe behaviour from aligned LLMs and multi-modal models; prompt injection and tool-use abuse against agentic systems; automating attack generation and evaluation rather than hand-crafting one-off prompts; and measuring whether guardrails hold under pressure. Where Realm Labs' interpretability work gives you access to a model's internals, use it.
  • We aim for a paper or public technical report out of every internship, plus attacks that stay in our evaluation suite after you leave.
Expected Background:
Adversarial ML and Red Teaming
  • Hands-on experience attacking or stress-testing models, from any direction: jailbreaks, prompt injection, adversarial examples, data poisoning, model extraction, or evaluating safety and moderation systems.
  • Able to read a paper and implement its attack.
  • (nice to have) Offensive security background outside ML: CTFs, vulnerability research, penetration testing.
  • (nice to have) Familiarity with agentic systems and their attack surface - tool calls, retrieval, memory, multi-agent orchestration.
Expected Background: ML
  • Machine learning tools: pytorch, huggingface, transformers, datasets.
  • Applied deep learning and LLM experience.
    • Training and evaluating deep models.
    • (nice to have) fine tuning LLMs, multi-modal LLMs.
  • (nice to have) Familiarity with ML[NLP,LLM,Vision] interpretability methods, sparse autoencoders, linear probes - as a way of locating failure modes, not as an end in itself.
Expected Background:
Software Engineering
  • Development environments and tools:
    • unix
      , git
      , basic clouds usage on AWS and/or GCP
    • jupyter
  • Programming:
    • python
    • (nice to have) "programming languages well-roundedness" /li>
    • experience in statically-typed and functional languages
Compensation & Benefits
  • Market aligned compensation for interns in the bay area.
Requirements
  • Must be authorized to work in the USA or must be able to obtain CPT (Curricular Practical Training) approval from host university.
#J-18808-Ljbffr
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary