×
Register Here to Apply for Jobs or Post Jobs. X

Cyber Digital Trust & Online Safety Manager

Remote / Online - Candidates ideally in
Surprise, Maricopa County, Arizona, 85379, USA
Listing for: PowerToFly
Remote/Work from Home position
Listed on 2026-08-03
Job specializations:
  • IT/Tech
    Cybersecurity, AI Business & Operations
Salary/Wage Range or Industry Benchmark: 134500 - 265100 USD Yearly USD 134500.00 265100.00 YEAR
Job Description & How to Apply Below

Cyber Digital Trust and Online Safety Manager

The Digital Trust & Online Protection Professional will advise clients in developing, managing, and implementing policies, procedures, and strategies to ensure a safe, compliant, and trustworthy environment for our users. This individual will scale and mature digital trust and safety processes, including content compliance, user protection, and regulatory adherence across our platforms for our clients. Working closely with cross-functional stakeholders, this role will monitor regulatory changes, manage risks, and enhance our organization's approach to content safety, user trust, and online integrity.

Work you'll do

As a Manager, Strategy, Growth, and Transformation on the Deloitte Cyber team, you will be responsible for:

  • Designing and executing testing scenarios to identify how prompts or user inputs could be manipulated to generate harmful, misleading, or misaligned generative artificial intelligence outputs.
  • Researching emerging prompt injection, jailbreak, and adversarial testing techniques to evaluate model weaknesses, bias, factual inaccuracy, and misalignment with user intent.
  • Assessing the effectiveness of content moderation systems in detecting unsafe outputs and documenting vulnerabilities, failure patterns, and potential misuse impacts.
  • Recommending improvements to moderation policies, flagging mechanisms, training data, and governance controls based on testing findings.
  • Collaborating with generative artificial intelligence development, content moderation, and cross-functional stakeholders to strengthen security, trust, safety, and responsible use outcomes.
  • Developing multimodal test content and novel prompt manipulation methods to identify failure modes across text and other model inputs.

A successful candidate would possess these skills:

  • Ability to work independently and collaborate as part of a team
  • Effective written and verbal communication skills
  • Meticulous attention to detail and quality of work product
  • Ability to build and sustain professional relationships
  • Ability to lead projects or work streams
  • Ability to manage and prioritize multiple tasks in a fast-paced and dynamic environment
  • Strong interpersonal skills and professional demeanor
  • Ability to meet deadlines
  • Ability to mentor and provide clear guidance to others
The team

Enables trust and safety of online communications and digital products, protecting users, consumers, and patients from harm. Enables clients to provide consumer confidence in knowing with whom they are dealing and ensuring the integrity of access to data.

Qualifications

Required:

  • Master's degree in Computer Science, Artificial Intelligence, Machine Learning, Data Science, Cybersecurity, Linguistics, Psychology, or a related field, or equivalent professional experience
  • 10+ years of experience in threat modeling and simulation, prompt generation and analysis, novel testing, and reporting and improvement
  • Demonstrated hands‑on experience, portfolio work, publications, or research in prompt injection, jailbreak testing, model evaluation, adversarial machine learning, multimodal artificial intelligence safety, or generative artificial intelligence vulnerability assessment
  • Ability to travel 25-50%, on average, based on the work you do and the clients and industries/sectors you serve.
  • Limited immigration sponsorship may be available.

Preferred:

  • Doctor of Philosophy (PhD) in Computer Science, Artificial Intelligence, Machine Learning, Data Science, Cybersecurity, Linguistics, Psychology, or a related field, or equivalent professional experience
  • Specialized training or certifications in generative artificial intelligence red teaming, adversarial machine learning, artificial intelligence security, cybersecurity, responsible artificial intelligence, or artificial intelligence governance
  • Experience designing and operationalizing trust and safety testing programs for large‑scale consumer platforms, including escalation workflows, issue triage, and remediation tracking
  • Experience working with product, legal, policy, and engineering stakeholders to translate risk findings into practical platform controls and governance improvements

The wage range for this…

To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary