AI Safety Expert - Red Team
Mercor
Job description
About the role
We are looking for an AI Safety Expert to join our Red Team on a contract basis. You will work remotely to probe conversational AI models and agents, uncovering jailbreaks, prompt injections, misuse scenarios, and bias exploitation.
Key responsibilities
- Red‑team conversational AI systems to identify and document jailbreaks, prompt injections, misuse cases, and bias exploitation.
- Generate high‑quality human data by annotating model failures, classifying vulnerabilities, and flagging systemic risks.
- Apply structured testing approaches using taxonomies, benchmarks, and playbooks to ensure consistent and repeatable assessments.
- Document findings reproducibly, producing reports, datasets, and actionable attack cases for customers.
- Work independently and asynchronously, meeting deadlines while contributing to improvements in AI model performance.
Required profile
- Fluent in English and Finnish.
- Prior experience in red‑team operations, AI adversarial research, cybersecurity, or socio‑technical probing.
- Ability to communicate risks clearly to both technical and non‑technical stakeholders.
Required skills
- Adversarial Machine Learning (jailbreak datasets, prompt injection, model extraction).
- Penetration testing and exploit development.
- Socio‑technical risk analysis, including harassment, disinformation probing, and abuse analysis.
Questions fréquentes
Why are you reporting this job?
Explore further
Salaries, guides and searches in Suomi.
Salaries by job title
Apply in 30 seconds
Enter your email to apply. An account will be created automatically.
By continuing, you accept our terms of use.
Already have an account? Login
Une question sur cette offre ?
Posez-la ici : vous recevrez le récapitulatif de l'offre par e-mail, tout de suite.
Published 1 kuukausi sitten
Expires 4 viikon päästä
18 views · 0 interested
Boost your chances
Upload your CV — we will match you with relevant openings.
Analyzing your CV...
Mercor