AI Safety Red Teamer Expert
4 settimane fa
Greater Milan Metropolitan Area, 25, Italia
Mercor
Tempo pieno
Gratuito con email o Google
Salva questo lavoro e mantieni la tua ricerca organizzata
Crea un account gratuito per salvare lavori, creare avvisi e tornare a questa inserzione dalla tua dashboard.
Gratuito con email o Google
About The Job
Mercor
connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include
Benchmark
,
General Catalyst
,
Peter Thiel
,
Adam D'Angelo
,
Larry Summers
, and
Jack Dorsey
.
Position:
AI Safety Red Teamer
Type:
Contract
Compensation:
$70–$84/hour
Location:
Remote Role Responsibilities
- Design adversarial prompts to stress-test frontier AI models.
- Identify jailbreaks, unsafe behaviors, hallucinations, and policy failures.
- Evaluate model robustness across misinformation, cyber, biosecurity, fraud, political content, and other sensitive domains.
- Document vulnerabilities and contribute to safety benchmarking and red-teaming reports.
- Collaborate with AI researchers to improve model alignment, robustness, and safety. Qualifications Must-Have
- Bachelor's degree or higher in Computer Science, Cybersecurity, Journalism, Communications, Psychology, Biology, Chemistry, Public Policy, or a related discipline.
- 5+ years of professional experience in AI Safety, AI Red Teaming, Trust & Safety, cybersecurity, investigative journalism, life sciences, or a related field.
- Strong analytical reasoning, prompt design, and written communication skills.
- Experience designing adversarial prompts or evaluating frontier AI systems. Preferred
- Experience with AI Red Teaming, RLHF, SFT, AI Alignment, or Trust & Safety.
- Familiarity with jailbreak testing, prompt engineering, or adversarial evaluation methodologies.
- Expertise in one or more grey-area domains, including cyber, biosecurity, political content, misinformation, or scientific safety. Application Process (Takes 20–30 mins to complete)
- Upload resume
- AI interview based on your resume
- Submit form Resources & Support
- For details about the interview process and platform information, please check: https://talent.docs.mercor.com/welcome
- For any help or support, reach out to: support@mercor.com *PS: Our team reviews applications daily. Please complete your AI interview and application steps to be considered for this opportunity.
* ,
Location:
Remote Role Responsibilities
- Design adversarial prompts to stress-test frontier AI models.
- Identify jailbreaks, unsafe behaviors, hallucinations, and policy failures.
- Evaluate model robustness across misinformation, cyber, biosecurity, fraud, political content, and other sensitive domains.
- Document vulnerabilities and contribute to safety benchmarking and red-teaming reports.
- Collaborate with AI researchers to improve model alignment, robustness, and safety. Qualifications Must-Have
- Bachelor's degree or higher in Computer Science, Cybersecurity, Journalism, Communications, Psychology, Biology, Chemistry, Public Policy, or a related discipline.
- 5+ years of professional experience in AI Safety, AI Red Teaming, Trust & Safety, cybersecurity, investigative journalism, life sciences, or a related field.
- Strong analytical reasoning, prompt design, and written communication skills.
- Experience designing adversarial prompts or evaluating frontier AI systems. Preferred
- Experience with AI Red Teaming, RLHF, SFT, AI Alignment, or Trust & Safety.
- Familiarity with jailbreak testing, prompt engineering, or adversarial evaluation methodologies.
- Expertise in one or more grey-area domains, including cyber, biosecurity, political content, misinformation, or scientific safety. Application Process (Takes 20–30 mins to complete)
- Upload resume
- AI interview based on your resume
- Submit form Resources & Support
- For details about the interview process and platform information, please check: https://talent.docs.mercor.com/welcome
- For any help or support, reach out to: support@mercor.com *PS: Our team reviews applications daily. Please complete your AI interview and application steps to be considered for this opportunity.
* ,