AI Safety Expert — Adversarial Machine Learning
Compensation
$16 – $22 /hr
Employment type
Full-time
Work setting
On-site
Location
New York, NY
Schedule
Day shift
Posted
You'll be redirected to the employer's application page.
Job overview
The AI Safety Expert — Adversarial Machine Learning role is remote; the posting identifies the work as remote without specifying a narrower geographic scope. Pay is $16–$22 per hour. Mercor connects creative and technical talent with leading AI research labs. The expert red-teams conversational AI models and agents, identifies vulnerabilities and misuse risks, and produces structured findings that customers can act on. The work supports consistent safety testing and improvements to AI model performance.
What you'll do
- Red-team conversational AI models and agents
- identify jailbreaks, prompt injections, misuse cases, and bias exploitation
- annotate failures and classify vulnerabilities
- follow testing taxonomies, benchmarks, and playbooks
- document reproducible reports, datasets, and attack cases
- work independently and asynchronously to meet deadlines.
What we're looking for
- Skills & competencies
- ai safety expertadversarial machine learningremoteenglishgujaratired teamingprompt injectionjailbreak testingai model safetyvulnerability analysiscybersecurity$16–$22/hrmercor
- Work arrangement
- Weekend coverage required
Why this role
Requires native fluency in English and Gujarati; preferred experience includes adversarial machine learning, cybersecurity, and conversational AI testing.
About the employer
Mercor is hiring for this role. Industry: Employment Placement Agencies. Sector: 56.
Additional details
- Industry sector
- 56
- Industry
- Employment Placement Agencies
- Occupation code
- 15-2051.00
You'll be redirected to the employer's application page.
Listing ID: ec3458f6-6632-4276-8c1e-5ac1efb6eae4