Any job page, tailored and filled in with one click.A tailored resume, answers and cover letter for any job page, filled in with one click.

Download the Chrome Extension
AI AdoptionFunded CompaniesJob SimulationCertificationsRoadmapsJobsPricing
Sign In
OneRoadmap

Empowering the next generation with AI education. Custom training for colleges and enterprises.

gaurav.ghai@oneroadmap.in
Delhi NCR, India

Platform

  • AI Roadmaps
  • Free Certifications
  • Learning Resources
  • Pricing

Training

  • AI Adoption Workshops
  • Expert Sessions
  • Upcoming Events
  • Workshop Gallery

Company

  • Blog
  • Contact

Legal

  • Privacy
  • Terms
  • Refunds
  • Delete your data

© 2026 OneRoadmap

Operated by Ghai Technologies, India · International operations through One Roadmap Marketing Management, Dubai, UAE

Built for your next chapter.

Open roles

AI / ML · AI-Safety

AI Safety Expert - Adversarial ML

mercor

SeniorRemote · USContractor$16–22Listed 15h ago
Apply now
Experience Senior · 6+ yrs

About the role

structured by ORI

About the job Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include Benchmark , General Catalyst , Peter Thiel , Adam D'Angelo , Larry Summers , and Jack Dorsey .

What you will do

  • Red team conversational AI models and agents. Conduct jailbreaks, prompt injections, misuse cases, and bias exploitation.
  • Generate high-quality human data. Annotate failures, classify vulnerabilities, and flag systemic risks.
  • Apply structure. Follow taxonomies, benchmarks, and playbooks to ensure consistent testing.
  • Document reproducibly. Produce reports, datasets, and attack cases that customers can act on.
  • Work independently and asynchronously . Thrive in flexible hours while improving AI model performance .

What they are looking for

  • Native fluency in English and Odia .
  • Strong judgment about language and content.
  • Rigorous attention to detail and consistency.
  • Structured approach to guidelines and quality standards.
  • Clear communication with technical and non-technical audiences.
  • Adaptability across projects, task types, and customers.

Nice to have

  • Experience in Adversarial ML : jailbreak datasets, prompt injection, RLHF/DPO attacks, model extraction.
  • Background in Cybersecurity : penetration testing, exploit development, reverse engineering.
  • Knowledge of socio-technical risk: harassment/disinfo probing, abuse analysis, conversational AI testing.
  • Creative probing skills: psychology, acting, writing for unconventional adversarial thinking.
Adversarial MLCybersecuritypenetration testingexploit developmentreverse engineeringRLHFDPO
Full posting text

About the job Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include Benchmark , General Catalyst , Peter Thiel , Adam D'Angelo , Larry Summers , and Jack Dorsey . Position: AI Safety Experts — English & Odia Type: Contract Compensation: $16–$22/hour Location: Remote Role Responsibilities Red team conversational AI models and agents. Conduct jailbreaks, prompt injections, misuse cases, and bias exploitation. Generate high-quality human data. Annotate failures, classify vulnerabilities, and flag systemic risks. Apply structure. Follow taxonomies, benchmarks, and playbooks to ensure consistent testing. Document reproducibly. Produce reports, datasets, and attack cases that customers can act on. Work independently and asynchronously . Thrive in flexible hours while improving AI model performance . Qualifications Must-Have Native fluency in English and Odia . Strong judgment about language and content. Rigorous attention to detail and consistency. Structured approach to guidelines and quality standards. Clear communication with technical and non-technical audiences. Adaptability across projects, task types, and customers. Preferred Experience in Adversarial ML : jailbreak datasets, prompt injection, RLHF/DPO attacks, model extraction. Background in Cybersecurity : penetration testing, exploit development, reverse engineering. Knowledge of socio-technical risk: harassment/disinfo probing, abuse analysis, conversational AI testing. Creative probing skills: psychology, acting, writing for unconventional adversarial thinking. Application Process (Takes 20–30 mins to complete) Upload resume AI interview based on your resume Submit form Resources & Support For details about the interview process and platform information, please check: For any help or support, reach out to: PS: Our team reviews applications daily. Please complete your AI interview and application steps to be considered for this opportunity. Originally posted on Himalayas

AI-SafetyMachine-Learning-TestingAI-Red-TeamAI-EvaluationAdversarial-MLAdversarial-ML-SpecialistAdversarial-ML-EngineerAdversarial-ML-ResearcherAI-Adversarial-SpecialistAI-Safety-ResearcherAI-Safety-Evaluator
Apply on company site

Meet Ori - your career agent on WhatsApp

Find jobs, get your roadmap, check if you're ready for a role and prepare applications - in chat, any language.

Ask Ori about this role
Checking your fit…

Opportunity details

Deadline
Closing in 60d · 21 Nov

As stated by the source. Anything not shown was not stated.

More at mercor

Product Manager II - AI

Razorpay · Product Management

Senior · 2+ yrsFunded

On-site · Bengaluru

AI / ML

3h ago
View role
T
Founding Fullstack Developer

TripSuite · Full-Stack-Developer

3+ yrs

Remote · US

SOFTWARE

Full Time · 5h ago
View role
G
Technical Writer / Quality Assurance Specialist - Remote - in USD

goPro Consultancy Group ltd. · Technical-Writer

3+ yrs

Remote · US

Full Time · 5h ago
View role
Senior DevOps Engineer

WebEngage · DevOps-Engineer

Senior · 1+ yrs

Remote · India

SOFTWARE

Full Time · 5h ago
View role
Apply