AI Adversarial Specialist - Fully Remote

Mercor
Berlin, Germany
3 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Compensation
€99,840.0 - €128,960.0
Working hours
Regular working hours
Languages
English, Finnish
Job source

Tech stack

Artificial Intelligence Software System Penetration Testing Cyber Security Red Team (Cyber Security)

Job description

  • Red team conversational AI models and agents to identify jailbreaks, prompt injections, misuse cases, and bias exploitation.
  • Generate high-quality human data by annotating failures, classifying vulnerabilities, and flagging systemic risks.
  • Apply structured approaches using taxonomies, benchmarks, and playbooks to ensure consistent testing.
  • Document findings reproducibly to produce reports, datasets, and attack cases that customers can act on.
  • Work independently and asynchronously to meet deadlines while improving AI model performance.

Requirements

Must-Have

  • Fluent in English and Finnish.
  • Prior experience in red teaming, AI adversarial work, cybersecurity, or socio-technical probing.
  • Ability to communicate risks clearly to both technical and non-technical stakeholders.

Preferred

  • Experience with Adversarial ML, including jailbreak datasets, prompt injection, and model extraction.
  • Background in Cybersecurity, such as penetration testing and exploit development.
  • Expertise in socio-technical risk, including harassment/disinfo probing and abuse analysis.

About the company

Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include Benchmark, General Catalyst, Peter Thiel, Adam D’Angelo, Larry Summers, and Jack Dorsey.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.adzuna.de

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

4:11 min

Introduction to cloud-native application developer security

Micah Silverman · WWC 2022

2:14 min

Exploring internal AI product initiatives and global engineering roles

Maria Apazoglou · Coffee With Developers

1:51 min

Leveraging continuous penetration testing via red teams

Reto Kaeser · LIVE

3:46 min

Core terminology and audiences for interpretable artificial intelligence

Karol Przystalski · LIVE

3:44 min

Current industry adoption and future security initiatives

Alexander Allmendinger · LIVE

4:27 min

Embracing a new perspective on mobile cyber attacks

Tom Tovar · WWC 2023

Videos

See all

Related articles

See all