mpathic

QA Reviewer & Project Coordinator (On-Site - Seattle or Boston)

mpathic Seattle, Washington, United States · $110K–$125K/yr

Technology, Information and Internet · 11-50 employees

19 h ago
qa Mid (2-5 yrs) Full-time United States
Log in to apply, save this posting, or score it against your profile with AI.

About the role

The role involves conducting quality assurance reviews and red teaming for AI systems to ensure safety, accuracy, and policy compliance. Additionally, the coordinator will manage project timelines, track deliverables, and facilitate communication across cross-functional teams.

What they look for

Quality Assurance Project Coordination AI Safety Red Teaming Prompt Engineering Data Annotation Clinical Evaluation Behavioral Policy Development Documentation Stakeholder Communication Risk Identification Workflow Optimization Calibration Sessions Google Workspace Slack Project Management

Requirements

Candidates should have experience with large language models and possess strong organizational and documentation skills. A background in clinical work, trust and safety, or project management is preferred to effectively handle sensitive content and complex workflows.

Full description

About the Role

We're seeking an experienced QA Reviewer & Project Coordinator. This is a hands-on role that combines expert evaluation and quality assurance with day-to-day project coordination. You'll spend most of your time reviewing work, testing AI systems, and ensuring high-quality outputs while also helping keep projects organized, on schedule, and running smoothly.

Position Overview

This position is ideal for someone who enjoys both deep subject matter expertise and operational execution.

You'll serve as one of the primary reviewers for AI safety work, conducting quality reviews, identifying behavioral edge cases, testing AI models through role play and red teaming, and ensuring consistency across projects. In addition, you'll coordinate portions of active programs by tracking deliverables, supporting staffing and project workflows, communicating with stakeholders, and helping identify operational risks before they become issues.

The ideal candidate enjoys solving ambiguous problems, has exceptional attention to detail, communicates clearly, and thrives in a fast-moving startup environment.

What You'll Do

QA & AI Evaluation

  • Review AI content for accuracy, safety, empathy, and policy compliance.
  • Conduct quality assurance reviews of expert evaluations and annotations.
  • Participate in calibration sessions to ensure reviewer consistency and quality.
  • Help improve QA workflows, reviewer documentation, and operational processes.
  • Role play realistic clinical scenarios with AI systems to evaluate behavior across diverse situations.
  • Perform red teaming to identify failure modes, safety risks, and behavioral edge cases.
  • Develop and refine evaluation rubrics, behavioral taxonomies, personas, and scoring guidelines.
  • Document model inconsistencies, safety concerns, and opportunities for improvement.
  • Collaborate with researchers and engineers to improve AI behavior through structured clinical feedback.
  • Maintain strict confidentiality while working with sensitive clinical content.

Project Coordination

  • Support day-to-day execution of AI safety and evaluation projects.
  • Track project timelines, deliverables, and reviewer assignments.
  • Coordinate review queues and help balance workloads across project teams.
  • Monitor project progress and proactively identify risks, blockers, or quality issues.
  • Maintain project trackers, documentation, and reporting dashboards.
  • Assist with staffing coordination as project needs evolve.
  • Help facilitate project meetings and document action items and follow-up tasks.

Required Qualifications

  • Familiarity with ChatGPT, Claude, Gemini, or other large language models.
  • Excellent written communication and documentation skills.
  • Strong organizational skills with exceptional attention to detail.
  • Comfortable managing multiple priorities simultaneously.
  • Ability to work independently while collaborating effectively across teams.
  • Comfortable working in a fast-paced startup environment with evolving priorities.

Preferred Qualifications

  • Experience reviewing or auditing clinical work for quality.
  • Experience with AI safety, LLM evaluation, prompt engineering, or red teaming.
  • Background in trust & safety, content moderation, or behavioral policy development.
  • Experience developing evaluation rubrics, taxonomies, or annotation guidelines.
  • Familiarity with data annotation or human-in-the-loop evaluation workflows.
  • Clinical experience working with serious mental illness, crisis intervention, or complex behavioral health populations.
  • Project coordination or project management experience.
  • Experience leading calibration sessions or reviewer training.
  • Experience with Google Workspace, Slack, spreadsheets, and project management tools.

You'll Thrive Here If You...

  • Notice details that others miss.
  • Can identify quality issues before they become larger problems.
  • Enjoy both analytical review work and coordinating people and projects.
  • Communicate clearly across clinical and operational teams.
  • Balance independent problem solving with knowing when to escalate.
  • Thrive in ambiguity and rapidly changing environments.
  • Care deeply about building AI systems that are safe, trustworthy, and clinically responsible.
  • Take ownership and consistently follow through.

Additional Requirements

  • Ability to work on-site at our designated office location (Seattle or Boston).
  • Willingness to sign comprehensive confidentiality and NDA agreements.
  • Comfortable working with sensitive mental health and AI safety content.
  • Participation in recurring project meetings, calibration sessions, and operational planning.
  • Availability to support occasional high-priority project deadlines as needed, including potential evening and weekend work.

Apply Even If You Don't Check Every Box

We know great candidates might not fit every bullet on a job description. If this role speaks to you and you're excited to help improve the future of healthcare research and AI safety, we'd love to hear from you.

Similar roles