Mercor
Mercor
51 – 200 Employees
Mercor’s available profile does not include enough verified information about its products, services, customers, sector, or hiring focus to support an accurate company overview. More source material is needed before describing the business or its opportunities.

AI Safety Expert – English and Marathi (Remote Contract)

Mercor seeks an AI safety expert fluent in English and Marathi to red-team conversational AI systems. The contract role focuses on adversarial testing, vulnerability analysis, and clear reporting for human-data and frontier AI projects.

Description

  • Test conversational AI models and agents against jailbreaks, prompt injection, misuse scenarios, bias exploitation, and multi-turn manipulation
  • Create human-data assets by annotating failures, categorizing vulnerabilities, and identifying systemic risks
  • Use established taxonomies, benchmarks, and playbooks to ensure consistent evaluations
  • Deliver reproducible attack cases, datasets, and customer-facing reports
  • Assess AI responses on sensitive subjects, including bias, misinformation, and harmful behavior
  • Find weaknesses that automated evaluation methods may overlook
  • Broaden testing coverage to help prevent unexpected production issues
  • Support customers in improving the safety, robustness, and reliability of their AI systems
  • Work with leading researchers on initiatives that train and advance frontier AI models

Requirements

  • Native or fluent English and Marathi language skills
  • Sound judgment when evaluating language and content
  • Ability to determine whether AI responses are accurate, complete, and suitable, with clear explanations
  • Careful attention to nuanced errors, inconsistencies, and omissions
  • Reliable adherence to testing guidelines and quality requirements
  • Clear communication of reasoning to both technical and non-technical audiences
  • Flexibility to work across different projects, tasks, and customers
  • Ability to work as an independent contractor
  • Must not require H-1B or STEM OPT sponsorship
  • Preferred: adversarial machine learning experience with jailbreak datasets, prompt injection, RLHF/DPO attacks, or model extraction
  • Preferred: cybersecurity experience in penetration testing, exploit development, or reverse engineering
  • Preferred: socio-technical risk experience in harassment or disinformation probing, abuse analysis, or conversational AI testing
  • Preferred: psychology, acting, or writing experience that supports unconventional adversarial thinking

Benefits

  • Fully remote contract work
  • Flexible schedule with the ability to set your own working hours
  • Weekly payments through Stripe or Wise
  • Project duration may change or end early based on business needs and performance
  • Optional participation in projects involving higher-sensitivity content
  • Defined guidance and wellness resources for sensitive-content assignments
  • Reasonable accommodations available upon request
  • Referral payments of up to $90 for each successful referral

Related Jobs