AI Safety Expert; remote
Seattle, King County, Washington, 98127, USA
Listed on 2026-07-18
-
Science
AI Evaluation, Data Annotation/ AI Labeling
Job Description
Job Title:
AI Safety Expert (Temporary Project)
Location:
Seattle, WA (On-site, Remote)
mpathic is keeping humans safe in the AI era through automated tools and expert datasets rooted in psychology and powered by clinicians. The Series A start‑up is backed by Foundry.vc and Next Frontier Capital.
AboutThe Role
Mpathic is seeking AI Safety Experts for a temporary project to evaluate and improve the safety, reliability, and real‑world behavior of frontier AI systems. Priority is given to applicants who can start immediately. Ideal candidates have expertise in human behavior, communication, policy, education, healthcare, technology, trust & safety, or other domains where judgment, critical thinking, and nuanced decision‑making matter.
Responsibilities- Evaluating AI‑generated conversations, responses, and reasoning for quality, safety, and usefulness
- Rating model outputs using structured evaluation rubrics and project guidelines
- Annotating conversational data to support AI training and benchmarking
- Identifying emerging risks, behavioral patterns, and opportunities for model improvement
- Providing written feedback to help researchers and engineers improve model performance
- Maintaining strict confidentiality while working with proprietary AI systems and sensitive content
- Participating in calibration sessions and quality reviews to ensure consistent evaluations
Curious, analytical, thoughtful communicators who enjoy solving complex problems and exercising sound judgment. Comfortable evaluating nuanced situations, following detailed guidelines, and contributing to trustworthy AI development.
Basic Qualifications- Professional experience or subject‑matter expertise in psychology, behavioral science, social work, trust & safety, research, or a related discipline
- Strong written communication skills with excellent attention to detail
- Comfortable learning structured evaluation frameworks and applying them consistently
- Strong critical thinking and problem‑solving skills
- High ethical standards and sound judgment when working with sensitive or ambiguous content
- Comfortable using AI tools, Google Workspace, Slack, and other web‑based collaboration platforms
- Willingness to sign NDAs and work on confidential projects
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).