AI Safety Experts — English & Finnish
AI Safety Experts — English & Finnish
- Pay
- $48 – $62/hr
- Location
- Remote — Finland
- The language requirement points to Finland, though the posting does not restrict applicants by country.
- Engagement
- Contractor
- Languages
- Finnish
Earns 25 points on this device — once per role per day
Applications are handled by Mercor on their own site. Dealuxe is not the employer and does not screen applicants.
Location: Remote | Type: Hourly Contract | Requirements: Native fluency in English & Finnish required.
The Frontier of AI Safety: Why Human Red Teaming Matters
Artificial intelligence models are scaling at a breathtaking pace, transforming everything from software development to creative arts and complex data analysis. However, with massive scale comes unprecedented vulnerability. As foundational models become more autonomous and integrated into core societal infrastructure, ensuring their safety, reliability, and alignment with human values is no longer optional—it is critical.
Enter human red teaming. While automated testing and algorithmic benchmarks can catch surface-level errors, they often fail to predict nuanced, sophisticated, or adversarial edge cases. That is why leading AI labs partner with platforms like Mercor to recruit specialized domain experts. By simulating adversarial attacks, probing for hidden biases, and rigorously stress-testing conversational agents, human experts build the crucial defenses that keep next-generation AI safe for public and enterprise deployment.
Role Overview & Key Responsibilities
The AI Safety Experts (English & Finnish) role at Mercor is designed for bilingual linguistic and technical professionals who love pushing systems to their absolute limits. In this remote, flexible contract position, you will operate as a critical line of defense for cutting-edge AI models.
Your day-to-day responsibilities will involve:
- Adversarial Red Teaming: Actively probe conversational AI models and agents using jailbreaks, prompt injections, misuse cases, multi-turn manipulation, and bias exploitation.
- Human Data Generation: Annotate failures, classify vulnerabilities, and systematically flag socio-technical risks in both English and Finnish.
- Structured Frameworks: Follow rigorous taxonomies, benchmarks, and testing playbooks to ensure consistent and reproducible safety evaluations.
- Comprehensive Documentation: Produce clear, reproducible reports, attack cases, and datasets that engineering teams can immediately act on.
Who Should Apply? Ideal Candidate Profile
Mercor is looking for sharp, curious, and adversarial thinkers who can think outside the box. Successful applicants typically bring a blend of linguistic fluency, technical intuition, and rigorous analytical discipline.
Core Qualifications:
- Native-level fluency in both English and Finnish.
- Prior experience in red teaming, cybersecurity, adversarial machine learning, or socio-technical probing.
- An innate curiosity and drive to discover system vulnerabilities.
- Exceptional communication skills to clearly articulate technical risks to diverse stakeholders.
Bonus Specialties That Stand Out:
- Adversarial ML: Familiarity with jailbreak datasets, RLHF/DPO attacks, and model extraction.
- Cybersecurity: Background in penetration testing, exploit development, or reverse engineering.
- Socio-technical Risk: Experience with abuse analysis, misinformation probing, or conversational AI evaluation.
Contract Terms, Flexibility, and Compensation
Offering a competitive hourly wage between $48 and $62 per hour, this contract position combines high earning potential with complete remote flexibility. You can manage your workload around your personal schedule, working independently as an international or domestic contractor.
Furthermore, payments are handled seamlessly on a weekly basis via trusted financial rails like Stripe or Wise. Because Mercor streamlines the recruitment process through smart matching and a singular structured interview process, qualified candidates can transition into projects rapidly without unnecessary administrative hurdles.
How to Complete Your Application Successfully
To secure a placement in competitive roles like this, submitting a complete and thorough application is vital. Ensure your resume or profile clearly highlights your linguistic background in English and Finnish, along with any relevant experience in testing, evaluation, or technical probing. Take your time during the application steps to answer all prompts thoughtfully, as automated screening and talent teams look closely at attention to detail.
Ready to Shape the Future of AI Safety?
Join top-tier researchers and engineers by lending your human expertise to frontier AI models. Apply today to lock in your hourly rate of $48–$62/hr.
Complete Your Mercor Application Now →Disclaimer: Job availability, exact compensation tiers, and project durations are subject to change based on client requirements and applicant qualifications.
What the work is
- Adversarial Red Teaming: Actively probe conversational AI models and agents using jailbreaks, prompt injections, misuse cases, multi-turn manipulation, and bias exploitation.
- Human Data Generation: Annotate failures, classify vulnerabilities, and systematically flag socio-technical risks in both English and Finnish.
- Structured Frameworks: Follow rigorous taxonomies, benchmarks, and testing playbooks to ensure consistent and reproducible safety evaluations.
- Comprehensive Documentation: Produce clear, reproducible reports, attack cases, and datasets that engineering teams can immediately act on.
What they ask for
- Native-level fluency in both English and Finnish.
- Prior experience in red teaming, cybersecurity, adversarial machine learning, or socio-technical probing.
- An innate curiosity and drive to discover system vulnerabilities.
- Exceptional communication skills to clearly articulate technical risks to diverse stakeholders.
- Adversarial ML: Familiarity with jailbreak datasets, RLHF/DPO attacks, and model extraction.
- Cybersecurity: Background in penetration testing, exploit development, or reverse engineering.
- Socio-technical Risk: Experience with abuse analysis, misinformation probing, or conversational AI evaluation.
Ready to apply for AI Safety Experts — English & Finnish?
Mercor states $48 – $62/hr for this role. The application is on their site and takes a few minutes.
Earns 25 points on this device — once per role per day
Dealuxe is not the employer, does not set the pay or the hiring terms, and cannot guarantee a role is still open. If you complete a purchase or form, we may earn a small commission at no extra cost to you.
Following an offer here banks 10 points on this device — once per page, within the 500 points a day anything on the site can earn.Ad Disclosure: the application link is a referral link.
While you job-hunt, save on the brands you already use
Browse 2,287 vetted brands with live commission offers and exclusive deals — every one open to everyone, no sign-in needed.
Explore all brand categories →Scroll through the piece and stay a moment. Reading pays 5 points and sharing pays 50. Following the apply link pays 25, and buying coins pays back 25 points a dollar.
Get stories in your inbox
New brand drops, deal breakdowns and the best of the Journal, straight from The Storefront Blog. Free forever — and subscribing pays you 25 points.
Drop your email in the box above, then bank the bonus. A paid plan pays 75 — three times the free tier — plus every paid-only post.
Copies the link with your caption. Grab a username to bank points across devices.
Boost this listing
See what's trendingTrade the points you've earned to push this up Trending and the homepage, where more readers will find it.
Similar roles
CFD Engineer
micro1
Monetize your computational fluid dynamics expertise by training next-generation artificial intelligence models. High-tier remote contracting with elite compensation.
- Location
- Remote — Worldwide
- Pay
- $50 – $150/hr
- Posted
Training & Development Specialist
micro1
Monetize your instructional design and educational program expertise by shaping next-generation artificial intelligence reasoning benchmarks. Flexible remote contracting with premium compensation.
- Location
- Remote — Global
- Pay
- $40 – $75/hr
- Posted
Supply Chain Manager
micro1
Elevate your logistics career by training next-generation artificial intelligence models. Explore high-paying remote contracts designed for seasoned supply chain professionals worldwide.
- Location
- Remote — Worldwide
- Pay
- $55 – $100/hr
- Posted
Civil Engineer
micro1
Leverage your infrastructure expertise to shape frontier artificial intelligence models while enjoying lucrative remote contracting opportunities. / Aproveche su experiencia en infraestructura para dar forma a modelos de inteligencia…
- Location
- Remote — Global
- Pay
- $50 – $100/hr
- Posted
Commercial Real Estate Manager
micro1
Explore a high-compensating remote contract role designed for senior commercial real estate professionals. Read on in English and Spanish side by side.
- Location
- Remote — Europe
- Pay
- $40 – $65/hr
- Posted
Generalist AI Trainer & Evaluator
micro1
Shape next-generation artificial intelligence models through critical thinking, rigorous annotation, and high-level analytical evaluation. Available in English & Español.
- Location
- Remote — Global
- Pay
- $50 – $90/hr
- Posted
Be the first to comment
Loading comments…