×
Register Here to Apply for Jobs or Post Jobs. X

AI Safety Expert - Fully Remote

Remote / Online - Candidates ideally in
UAE/Dubai
Listing for: mercor
Remote/Work from Home position
Listed on 2026-10-11
Job specializations:
  • IT/Tech
    AI Evaluation, Data Annotation/ AI Labeling
Salary/Wage Range or Industry Benchmark: 81000 - 111000 AED Yearly AED 81000.00 111000.00 YEAR
Job Description & How to Apply Below
Position: AI Safety Expert - Fully Remote | Upto $22/hr

About the job Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include Benchmark , General Catalyst , Peter Thiel , Adam D'Angelo , Larry Summers , and Jack Dorsey .

Position: AI Safety Experts — English & Malayalam

Type:
Contract

Compensation: $16–$22/hour

Location:

Remote

Role Responsibilities
  • Red team conversational AI models and agents to identify jailbreaks, prompt injections, and misuse cases.
  • Generate high-quality human data by annotating failures, classifying vulnerabilities, and flagging systemic risks.
  • Apply structure by following taxonomies, benchmarks, and playbooks to ensure consistent testing.
  • Document reproducibly by producing reports, datasets, and attack cases that customers can act on.
  • Work independently and asynchronously to meet deadlines while improving AI model performance .
Qualifications Must-Have
  • Fluent Language

    Skills Required:

    English & Malayalam . Native fluency in English and Malayalam is required.
  • Strong judgment about language and content accuracy.
  • Rigorous attention to detail and ability to notice subtle errors.
  • Structured approach to work following guidelines and quality standards.
  • Clear communication skills for technical and non-technical audiences.
  • Adaptability across projects, task types, and customers.
Preferred
  • Experience in Adversarial ML : jailbreak datasets, prompt injection, RLHF/DPO attacks, model extraction.
  • Cybersecurity skills: penetration testing, exploit development, reverse engineering.
  • Understanding of socio-technical risk: harassment/disinfo probing, abuse analysis, conversational AI testing.
  • Creative probing skills: psychology, acting, writing for unconventional adversarial thinking.
To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary