AI Safety Red Teamer
AI Safety Red Teamer
- Pay
- $70 – $84/hr
- Location
- Remote — Worldwide
- Engagement
- Contractor · full time
Earns 25 points on this device — once per role per day
Applications are handled by Mercor on their own site. Dealuxe is not the employer and does not screen applicants.
The Evolution of Artificial Intelligence Safety: Why Red Teaming is the Most Critical Discipline in Tech Today
As large language models (LLMs) and generative artificial intelligence systems rapidly evolve, their capability scope has expanded far beyond simple text generation or basic coding assistance. Today’s frontier AI architectures touch critical pillars of modern society—from handling sensitive biomedical research and writing executable cybersecurity code to summarizing complex geopolitical narratives. With extraordinary capability comes unprecedented responsibility. Ensuring these models remain safe, aligned, ethical, and resilient against malicious exploitation is no longer optional; it is the absolute foundation of the modern tech economy.
This massive paradigm shift has created an unprecedented surge in demand for elite human evaluators known as AI Safety Red Teamers. Platforms like Mercor are bridging the gap between top-tier technical talent and leading AI research laboratories by facilitating specialized, high-paying remote contracts. Offering an exceptional compensation rate of $70 to $84 per hour, the AI Safety Red Teamer role allows seasoned cybersecurity professionals, investigators, researchers, and domain experts to stress-test frontier models, uncover deep-seated vulnerabilities, and shape the future of global AI safety—all on a fully flexible schedule.
Comprehensive Job Description & Core Responsibilities
For professionals looking to transition their analytical, investigative, or technical expertise into the booming field of AI alignment and safety, understanding the core responsibilities of this position is essential. An AI Safety Red Teamer acts as an adversarial tester, deliberately probing the boundaries of advanced machine learning systems to find where safety guardrails break down.
Key Responsibilities Include:
- Adversarial Prompt Design: Engineer sophisticated, multi-layered adversarial prompts designed to stress-test frontier AI models and bypass superficial alignment filters.
- Vulnerability Identification: Uncover critical system jailbreaks, unsafe behaviors, unexpected hallucinations, and policy failures across complex prompts.
- Domain-Specific Robustness Evaluation: Evaluate model behavior across high-risk, sensitive, and ambiguous ("grey-area") domains, including cybersecurity exploits, biosecurity risks, sophisticated fraud, political propaganda, and misinformation.
- Documentation & Reporting: Systematically document discovered vulnerabilities, formulate comprehensive safety benchmarks, and contribute detailed red-teaming reports for core research teams.
- Cross-Functional Collaboration: Work directly alongside elite AI researchers and safety scientists to improve model alignment, robustness, and general safety guardrails.
Candidate Qualifications & Professional Requirements
Because AI Safety Red Teamers operate at the absolute frontier of technology, the qualification criteria reflect a need for rigorous academic training and deep professional background.
- Educational Background: Bachelor’s degree or higher in Computer Science, Cybersecurity, Journalism, Communications, Psychology, Biology, Chemistry, Public Policy, or a related discipline.
- Professional Experience: 5+ years of professional experience working in AI Safety, AI Red Teaming, Trust & Safety, cybersecurity, investigative journalism, life sciences, or a related complex field.
- Core Competencies: Exceptional analytical reasoning, sophisticated prompt design capabilities, and crystal-clear written communication skills.
- Testing Proficiency: Demonstrated experience designing adversarial prompts or systematically evaluating frontier AI systems.
Preferred Qualifications: Hands-on experience with RLHF (Reinforcement Learning from Human Feedback), SFT (Supervised Fine-Tuning), AI alignment methodologies, jailbreak testing, and specialized expertise in grey-area domains like political content or scientific safety.
Contract Terms, Flexibility, and Seamless Global Payouts
Engaging as an independent contractor through Mercor delivers unique professional advantages that traditional full-time employment simply cannot match:
- Complete Location Independence: Conduct your red-teaming projects on your own schedule from anywhere within eligible regions worldwide.
- Dynamic Scheduling: Balance high-impact adversarial testing seamlessly with consulting engagements, research projects, or personal commitments.
- Reliable Weekly Remuneration: Receive prompt, dependable weekly compensation through trusted financial infrastructure like Stripe and Wise based entirely on your rendered services.
Why This Role Defines the Cutting Edge of Your Career
If you have spent your career dissecting complex security vulnerabilities, investigating nuanced narratives, or analyzing scientific data, your skill set is immensely powerful in the age of generative AI. Traditional tech roles often involve repetitive maintenance; conversely, AI red teaming puts you on the front lines of technological evolution. You aren't just observing the AI revolution—you are actively securing it.
Furthermore, commanding up to $84 per hour while collaborating directly with world-class artificial intelligence laboratories positions you at the pinnacle of modern intellectual compensation. As AI systems become more autonomous, experts with a proven background in safety red teaming will remain the most sought-after authorities in the tech ecosystem.
Secure Your Opportunity Today
Specialized AI safety contracts represent some of the most competitive and sought-after positions in the modern digital economy, and early applicant pools fill up rapidly. If you possess the analytical rigor and professional background to challenge frontier AI systems, take control of your career trajectory and earnings by submitting your application today.
What the work is
- Adversarial Prompt Design: Engineer sophisticated, multi-layered adversarial prompts designed to stress-test frontier AI models and bypass superficial alignment filters.
- Vulnerability Identification: Uncover critical system jailbreaks, unsafe behaviors, unexpected hallucinations, and policy failures across complex prompts.
- Domain-Specific Robustness Evaluation: Evaluate model behavior across high-risk, sensitive, and ambiguous ("grey-area") domains, including cybersecurity exploits, biosecurity risks, sophisticated fraud, political propaganda, and misinformation.
- Documentation & Reporting: Systematically document discovered vulnerabilities, formulate comprehensive safety benchmarks, and contribute detailed red-teaming reports for core research teams.
- Cross-Functional Collaboration: Work directly alongside elite AI researchers and safety scientists to improve model alignment, robustness, and general safety guardrails.
What they ask for
- Educational Background: Bachelor’s degree or higher in Computer Science, Cybersecurity, Journalism, Communications, Psychology, Biology, Chemistry, Public Policy, or a related discipline.
- Professional Experience: 5+ years of professional experience working in AI Safety, AI Red Teaming, Trust & Safety, cybersecurity, investigative journalism, life sciences, or a related complex field.
- Core Competencies: Exceptional analytical reasoning, sophisticated prompt design capabilities, and crystal-clear written communication skills.
- Testing Proficiency: Demonstrated experience designing adversarial prompts or systematically evaluating frontier AI systems.
Ready to apply for AI Safety Red Teamer?
Mercor states $70 – $84/hr for this role. The application is on their site and takes a few minutes.
Earns 25 points on this device — once per role per day
Dealuxe is not the employer, does not set the pay or the hiring terms, and cannot guarantee a role is still open. If you complete a purchase or form, we may earn a small commission at no extra cost to you.
Following an offer here banks 10 points on this device — once per page, within the 500 points a day anything on the site can earn.Ad Disclosure: the application link is a referral link.
While you job-hunt, save on the brands you already use
Browse 2,287 vetted brands with live commission offers and exclusive deals — every one open to everyone, no sign-in needed.
Explore all brand categories →Scroll through the piece and stay a moment. Reading pays 5 points and sharing pays 50. Following the apply link pays 25, and buying coins pays back 25 points a dollar.
Get stories in your inbox
New brand drops, deal breakdowns and the best of the Journal, straight from The Storefront Blog. Free forever — and subscribing pays you 25 points.
Drop your email in the box above, then bank the bonus. A paid plan pays 75 — three times the free tier — plus every paid-only post.
Copies the link with your caption. Grab a username to bank points across devices.
Boost this listing
See what's trendingTrade the points you've earned to push this up Trending and the homepage, where more readers will find it.
Similar roles
Board Game Reasoning Expert
Turing
Discover how your mastery of game mechanics, strategic reasoning, and complex rule systems can train next-generation artificial intelligence models.
- Location
- Remote — Worldwide
- Posted
Autodesk Fusion CAM Expert
micro1
Explore how world-class CNC programmers, machinists, and CAD/CAM specialists are directing the evolution of artificial intelligence in advanced manufacturing while leveraging elite remote contracts.
- Location
- Remote — Worldwide
- Pay
- $80 – $120/hr
- Posted
AI Safety Practitioner
Mercor
As generative artificial intelligence systems and frontier large language models (LLMs) evolve at a breathtaking pace, their integration into everyday workflow, business operations, and public communication grows deeper. However, raw…
- Location
- Remote — Worldwide
- Pay
- $60 – $70/hr
- Posted
Product / Artifacts Experts
Mercor
Mercor is hiring Product / Artifacts professionals with product management or product development experience. In this role, you will review, assess, and provide structured feedback on your domain-specific documents, ensuring quality,…
- Location
- Remote — Europe
- Pay
- $80 – $160/hr
- Posted
Art Domain Expert
Turing
Discover how senior art historians, curators, and humanities researchers are shaping the next generation of frontier AI models through rigorous evaluation and prompt engineering.
- Location
- Remote — Worldwide
- Posted
Image/Video Annotator
Turing
Explore how remote specialists are training next-generation AI models to perceive visual context, motion, and human emotion through advanced annotation workflows.
- Location
- Remote — Worldwide
- Posted
More ai training & safety listings
- Lifestyle Expert
- Sports Domain Expert
- Customer Support Assistant$21 – $45/hr
- CFD Engineer$50 – $150/hr
- Shopify Specialist$60 – $100/hr
- Management Analyst
- Training & Development Specialist$40 – $75/hr
- Ambulance Dispatcher & Emergency Communications Specialist$30 – $55/hr
Be the first to comment
Loading comments…