AI Safety Specialist - Remote | Upto $84/hr
mercor
PortugalPosted 17 days agoDiscoveredMatch locked
RemoteContractor
Mercor
connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include
Benchmark
,
General Catalyst
,
Peter Thiel
,
Adam D'Angelo
,
Larry Summers
, and
Jack Dorsey
.
Position
AI Safety Red Teamer
Role Responsibilities
- Design adversarial prompts to stress-test frontier
AI models
.
- Identify jailbreaks, unsafe behaviors, hallucinations, and policy failures.
- Evaluate model robustness across misinformation, cyber, biosecurity, fraud, political content, and other sensitive domains.
- Document vulnerabilities and contribute to safety benchmarking and red-teaming reports.
- Collaborate with
AI researchers
to improve model alignment, robustness, and safety.
Must-Have
•
Bachelor's degree
or higher in
Computer Science
,
Cybersecurity
,
Journalism
,
Communications
,
Psychology
,
Biology
,
Chemistry
,
Public Policy
, or a related discipline.
•
5+ years
of professional experience in
AI Safety
,
AI Red Teaming
,
Trust & Safety
, cybersecurity, investigative journalism, life sciences, or a related field.
•
Strong analytical reasoning, prompt design, and written communication skills.
•
Experience designing adversarial prompts or evaluating frontier AI systems
.
Preferred
•
Experience with AI Red Teaming
,
RLHF
,
SFT
,
AI Alignment
, or
Trust & Safety
.
•
Familiarity with jailbreak testing, prompt engineering, or adversarial evaluation methodologies.
- Expertise in one or more grey-area domains, including cyber, biosecurity, political content, misinformation, or scientific safety.
Application Process (Takes 20–30 mins to complete)
•
Upload resume
•
AI interview based on your resume
•
Resources & Support
•
For details about the interview process and platform information, please check
•
For any help or support, reach out to
PS: Our team reviews applications daily. Please complete your AI interview and application steps to be considered for this opportunity.
Originally posted on Himalayas
Not included in the source posting: benefits.
Skills
artificial-intelligencecybersecurity
Who can apply
The employer didn't state any visa, work authorization, citizenship or clearance requirements in this posting. Confirm with the employer before applying.
Read automatically from the employer's posting text. Always confirm with the employer — requirements can change after a job is published.