Mercor is assembling a panel of radiological safety experts to red-team frontier AI models. The goal is to test whether a model can correctly judge the misuse potential of a technical request while answering legitimate questions fully while refusing genuinely dangerous ones.
You will write prompts across three levels: benign, dual-use, and adversarial, evaluate responses against a defined policy standard, and craft reference answers with technical reasoning.