Mercor · Open to all
AI Safety Red Teamer
Listed on Mercor as “AI Safety Red Teamer”
What this actually is
You complete short labeling tasks, record everyday activities, or evaluate AI outputs across general knowledge domains. Open to most backgrounds. The platform title (AI Safety Red Teamer) reflects the rate band and the expertise required, not the day-to-day work.
Can you do this on your visa?
F-2 / F-4 / F-5 / F-6: open. E-1 to E-7: needs concurrent-employment permit. D-2 / D-4 students: S-3 permit, 20 hr/week cap. D-10 / D-8: case by case.
Korean tax on USD income
First 5 years in Korea: foreign-source income only taxed if remitted into Korea. After year 5: worldwide income. Full tax guide.
Original posting from Mercor
We are seeking experienced AI Safety Red Teamers to identify vulnerabilities in frontier AI systems through adversarial testing. You will design challenging prompts, uncover model weaknesses, and evaluate AI behavior across complex, high-risk, and ambiguous ("grey-area") topics.
Responsibilities
- Design adversarial prompts to stress-test frontier AI models.
- Identify jailbreaks, unsafe behaviours, hallucinations, and policy failures.
- Evaluate model robustness across misinformation, cyber, biosecurity, fraud, political content, and other sensitive domains.
- Document vulnerabilities and contribute to safety benchmarking and red-teaming reports.
- Collaborate with AI researchers to improve model alignment, robustness, and safety.
Required Qualifications
- Bachelor's degree or higher in Computer Science, Cybersecurity, Journalism, Communications, Psychology, Biology, Chemistry, Public Policy, or a related discipline.
- 5+ years of professional experience in AI Safety, AI Red Teaming, Trust & Safety, cybersecurity, investigative journalism, life sciences, or a related field.
- Strong analytical reasoning, prompt design, and written communication skills.
- Experience designing adversarial prompts or evaluating frontier AI systems.
Preferred Qualifications
- Experience with AI Red Teaming, RLHF, SFT, AI Alignment, or Trust & Safety.
- Familiarity with jailbreak testing, prompt engineering, or adversarial evaluation methodologies.
- Expertise in one or more grey-area domains, including cyber, biosecurity, political content, misinformation, or scientific safety.
Why Join?
- Help secure and strengthen the next generation of frontier AI models.
- Work on cutting-edge adversarial testing alongside leading AI researchers and safety teams.
- Influence how AI systems respond to complex, real-world safety challenges.
Quoted from Mercor’s public listing on 2026-07-21. We don’t edit platform copy; honest framing is in the title and the “what this actually is” block above.
Advertisement
Related AI training jobs
Mercor · Open to all
A/R Follow-up Manager
$75/hr · Remote · USD
Mercor · Open to all
AI Rater Guidelines Writer (Linguist / Instructional Designer)
$45-$65/hr · Remote · USD
Mercor · Open to all
AI Safety Practitioner
$60-$70/hr · Remote · USD
Mercor · Open to all
ASIC/SoC Design & Verification Engineer
$70-$100/hr · Remote · USD
More on this platform
About Mercor
AI-interview-based talent network. One application, voice interview with their AI, then matched to projects across coding, research, and specialist work. Pay scales with track and seniority.
Mercor review: AI-interview talent network
4.1/5 on Glassdoor, fastest-growing platform in the category (+509% YoY). What the AI video interview actually asks, real pay across coding/research/medical/legal/finance tracks ($25-$200/hr), and the project-availability problem.
See all AI training jobs
Browse by category and compare across all eight platforms we cover.