Adversarial AI Specialist - Fully Remote

Mercor, Inc.
New York, NY, United States
3 days ago
Apply on www.careerjet.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Part-time (≤ 32 hours)
Experience level
Starter
Compensation
$41,600.0 - $45,760.0
Working hours
Regular working hours
Languages
English, Urdu

Tech stack

Artificial Intelligence Software System Penetration Testing Cyber Security Red Team (Cyber Security) Reverse Engineering

Job description

  • Red team conversational AI models and agents. Conduct jailbreaks, prompt injections, misuse cases, and bias exploitation.
  • Generate high-quality human data. Annotate failures, classify vulnerabilities, and flag systemic risks.
  • Apply structure using taxonomies, benchmarks, and playbooks to maintain testing consistency.
  • Document reproducibly. Produce reports, datasets, and attack cases that customers can act on.
  • Work independently and asynchronously to meet deadlines while improving AI model performance., PS: Our team reviews applications daily. Please complete your AI interview and application steps to be considered for this opportunity.

Requirements

Must-Have

  • Native fluency in English and Urdu.
  • Prior red teaming experience in AI adversarial work, cybersecurity, or socio-technical probing.
  • Strong communication skills to explain risks to technical and non-technical stakeholders.

Preferred

  • Experience in Adversarial ML: jailbreak datasets, prompt injection, RLHF/DPO attacks, model extraction.
  • Background in Cybersecurity: penetration testing, exploit development, reverse engineering.
  • Expertise in socio-technical risk: harassment/disinfo probing, abuse analysis, conversational AI testing.
  • Creative probing skills in psychology, acting, or writing for unconventional adversarial thinking.

Benefits & conditions

  • $75,000-90,000 per year

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.careerjet.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:18 min

Adapting engineering interviews to evaluate critical thinking and curiosity

Alex Laubscher Alex Laubscher +3 · World Congress 2025

4:11 min

Introduction to cloud-native application developer security

Micah Silverman · World Congress 2022

46 sec

Using LLMs to reverse engineer undocumented legacy code

Michele Zuccala Michele Zuccala +4 · World Congress 2026 Europe

3:46 min

Core terminology and audiences for interpretable artificial intelligence

Karol Przystalski · LIVE

2:14 min

Exploring internal AI product initiatives and global engineering roles

Maria Apazoglou · Coffee With Developers

3:44 min

Current industry adoption and future security initiatives

Alexander Allmendinger · LIVE

Videos

See all

Related articles

See all