Mercor is hiring PhD and Master's scientists to author AI evaluation tasks (Sci Code)Mercor is partnering with leading AI labs on a recent benchmark for scientific computing.
You will author original, executable research problems that today's frontier models cannot solve.Domains — depth required in at least two subdomains (with a coding focus)Biology — ecology, biochemistry, geneticsWhat you'll doSource your own material: a published paper, a Kaggle dataset, an open-source repository, or a scenario you designWrite scientific prompts based on the inputBuild the grading criteria that define a correct answerCalibrate against frontier models — a task ships only when strong models fail it more often than they succeedRequiredPhD in biology, biological sciences, biochemistry, genetics, ecology, or a closely related fieldDemonstrated depth in at least two of the following subdomains: ecology, biochemistry,
geneticsWorking proficiency in Python, R, or another relevant programming language for scientific computingComfortable with Git/GitHub and running code in Docker — authoring runs through a pull-request workflow with automated quality checksPreferredPublications in peer-reviewed journalsPrior scientific software or research engineering experienceEngagementDuration: 6 weeksCommitment: part-time, 20+ hours per weekStart date: immediateProcessUpload your resume and application formA 25-minute conversational interview covering your background, experience, and motivationsFollow up within a few days with next steps and onboardingApply today and put your research expertise to work building the next generation of scientific AI.
#J-*****-Ljbffr
📌 Biology Phd: Scicode Ai Benchmark Engineer (Sydney)
🏢 Mercor
📍 Sydney