Back to search

AI Safety Expert — Adversarial Machine Learning

Mercor · New York, NY · AI research lab projects

Compensation

$16 – $22 /hr

Employment type

Full-time

Work setting

On-site

Location

New York, NY

Schedule

Day shift

Posted

Apply for this job

You'll be redirected to the employer's application page.

Job overview

The AI Safety Expert — Adversarial Machine Learning role is remote; the posting identifies the work as remote without specifying a narrower geographic scope. Pay is $16–$22 per hour. Mercor connects creative and technical talent with leading AI research labs. The expert red-teams conversational AI models and agents, identifies vulnerabilities and misuse risks, and produces structured findings that customers can act on. The work supports consistent safety testing and improvements to AI model performance.

What you'll do

  • Red-team conversational AI models and agents
  • identify jailbreaks, prompt injections, misuse cases, and bias exploitation
  • annotate failures and classify vulnerabilities
  • follow testing taxonomies, benchmarks, and playbooks
  • document reproducible reports, datasets, and attack cases
  • work independently and asynchronously to meet deadlines.

What we're looking for

Skills & competencies
ai safety expert
adversarial machine learning
remote
english
gujarati
red teaming
prompt injection
jailbreak testing
ai model safety
vulnerability analysis
cybersecurity
$16–$22/hr
mercor
Work arrangement
  • Weekend coverage required

Why this role

Requires native fluency in English and Gujarati; preferred experience includes adversarial machine learning, cybersecurity, and conversational AI testing.

About the employer

Mercor is hiring for this role. Industry: Employment Placement Agencies. Sector: 56.

Additional details

Industry sector
56
Industry
Employment Placement Agencies
Occupation code
15-2051.00
Apply for this job

You'll be redirected to the employer's application page.

Browse more jobs

Listing ID: ec3458f6-6632-4276-8c1e-5ac1efb6eae4

    AI Safety Expert — Adversarial Machine Learning | CollabWORK