07 Sep
|
Mercor
|
Australia
About the job Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include Benchmark , General Catalyst , Peter Thiel , Adam D'Angelo , Larry Summers , and Jack Dorsey . Position: AI Safety Experts — English & Indonesian Type: Contract Compensation: $17–$25/hour Location: Remote Role Responsibilities - Red team conversational AI models and agents. Focus on jailbreaks, prompt injections, misuse cases, and bias exploitation. - Generate high-quality human data. Annotate failures, classify vulnerabilities, and flag systemic risks. - Apply structure using taxonomies, benchmarks, and playbooks to maintain consistent testing. - Document reproducibly. Produce reports, datasets, and attack cases that customers can act on.
Qualifications Must-Have - Fluent in English and Indonesian . - Prior red teaming experience in AI adversarial work , cybersecurity , or socio-technical probing.
- Ability to explain risks clearly to technical and non-technical stakeholders. Preferred - Experience in Adversarial ML : jailbreak datasets, prompt injection, RLHF/DPO attacks, model extraction. - Background in Cybersecurity : penetration testing, exploit development, reverse engineering. - Knowledge of socio-technical risk: harassment/disinfo probing, abuse analysis, conversational AI testing. Application Process (Takes 20–30 mins to complete) - Upload resume - AI interview based on your resume - Submit form Resources & Support - For details about the interview process and platform information, please check: - For any help or support, reach out to: PS: Our team reviews applications daily. Please complete your AI interview and application steps to be considered for this prospect.
📌 AI Red Team Specialist - Remote | Upto $25/hr (Australia)
🏢 Mercor
📍 Australia