×
Register Here to Apply for Jobs or Post Jobs. X
More jobs:

Technical AI Policy Researcher, Model Behaviour - Trust and Safety

Job in San Francisco, San Francisco County, California, 94199, USA
Listing for: TikTok
Full Time position
Listed on 2026-08-15
Job specializations:
  • Research/Development
    AI Evaluation
Salary/Wage Range or Industry Benchmark: 93600 - 220400 USD Yearly USD 93600.00 220400.00 YEAR
Job Description & How to Apply Below

Technical AI Policy Researcher, Model Behaviour - Trust and Safety

Location:

Employment Type:

Regular

Job Code:

A217030

Responsibilities
  • Design and maintain multimodal GenAI policies across safety‑relevant domains, including political and ideological bias, deceptive misuse, manipulation and persuasion, and fairness.
  • Translate risk and harm models into clear behavioural specifications, evaluation criteria, grading guidance, and system‑level safeguards.
  • Define practical boundaries between beneficial uses of AI and assistance that could materially enable harm, exploitation, misuse, or unsafe outcomes.
  • Build policy artifacts that support model training, evaluation, and deployment. Partner with safety researchers, engineers, product teams, and other stakeholders to operationalize policy into scalable model behaviour and measurable safeguards.
  • Design end‑to‑end policy development to pre‑launch evaluation to post‑launch monitoring workflows across safety‑relevant domains, including golden set construction, labeling guidance, calibration, adjudication, and eval coverage analysis, to ensure policies can be reliably measured and improved.
  • Use red‑teaming results, deployment data, model failures, over‑refusals, under‑refusals, and ambiguous edge cases to improve policy and evaluation quality over time.
  • Identify emerging capability areas where frontier AI systems could create new safety, fairness or bias challenges or lower barriers to harm.
  • Monitor post‑launch model activity to identify gaps in our policy framework to capture unsafe model behaviour.
  • Champion research to strengthen the defensibility and operability of policy positions, including working with Outreach and Partnerships to incorporate external expert input into relevant policy positions.
  • Combine longer‑horizon safety research with hands‑on launch and deployment work.
  • Contribute to safety reports, policy documentation, launch reviews, and AI governance reviews on the company's approach to building AI responsibly.
  • Support regulatory teams as a subject matter expert on AI compliance related initiatives.
Minimum Qualifications
  • 5 years in Trust & Safety, AI Safety Research, AI Ethics, technical AI Governance, or equivalent experience.
  • Degree in Computer Science, Human‑Computer Interaction, Engineering, Data Science or quantitative Social Sciences.
  • Direct experience in policy development, AI evaluations, red‑teaming, or AI governance work.
  • Strong technical understanding of LLM, multimodel, or generative media model behavior, model failure modes, and safety risks.
  • Demonstrated experience working with external experts and stakeholders, including civil society, government, and academia.
  • Demonstrated success working in a fast‑paced technology company or research organization conducting AI impact, risk assessments or algorithmic audits, and/or data science or product development related experience.
  • Ability to advocate for safety amongst a wide variety of business stakeholders including Product Policy, Engineering, Public Policy, Legal, Communications, and Data Science.
Preferred Qualifications
  • Ability to explain complex technical concepts to non‑technical stakeholders.
  • Experience working with governments, frontier AI companies, or AI Safety organizations.
  • Familiarity in Python and experience building ML systems.
  • Are comfortable working across the research‑to‑deployment pipeline, from exploratory experiments to production systems.
Job Information

The base salary range for this position in the selected city is $93600 - $220400 annually.

Compensation may vary outside of this range depending on a number of factors, including a candidate’s qualifications, skills, competencies and experience, and location. Base pay is one part of the Total Package that is provided to compensate and recognize employees for their work, and this role may be eligible for additional discretionary bonuses/incentives, and restricted stock units.

Benefits

Benefits may vary depending on the nature of employment and the country work location. Employees have day one access to medical, dental, and vision insurance, a 401(k) savings plan with company match, paid parental leave, short‑term and long‑term disability…

To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary