← All jobs
Mercor Verified
remote · task-based

Chemicals Safety for Redteaming

$65–$75/hr
Share
Life, Physical, and Social Science task-based remote
Posted Sep 9, 2026

About the work

Mercor is assembling a panel of chemistry domain experts to red-team frontier AI models. The goal is to test whether a model can correctly judge the misuse potential of a technical request; answering legitimate questions fully while refusing genuinely dangerous ones.

You will:

  • Write challenging single-turn prompts in your domain, labelled across three levels: benign, dual-use, and adversarial.

  • Evaluate the responses that come back against a defined policy standard, and judge whether each was handled correctly.

  • Write the reference answer — what a correct response looks like, and the technical reasoning for why.

The hard part is calibration, not difficulty. A benign prompt that a model wrongly refuses is as valuable as an adversarial one it wrongly answers. That judgment is why we need domain experts rather than generalists — in chemistry the line sits between mechanism and procedure, and only someone who has run the reaction knows exactly where it falls. Background we're looking for

We're looking for professionals with hands-on experience in synthesis, analytical chemistry, or chemical safety. Prior AI evaluation or red-teaming experience is preferred. Relevant backgrounds include:

  • Chemical Defense Researcher (USAMRICD, DoD, CDC) — toxicological agents and medical countermeasures.

  • Synthetic / Process Chemist — route design, reaction engineering and scale-up.

  • Analytical Chemist — detection and identification methods for hazardous compounds.

  • Industrial Hygienist / Process Safety Engineer — exposure assessment, toxic release and process hazard analysis.

  • Forensic / Toxicological Chemist — controlled substances, precursor identification and casework.

What this role demands beyond technical depth

This is writing-intensive work. Every judgment you make needs a written rationale that a non-specialist can follow. Prior technical writing, published research, or expert witness experience is a strong signal — please include a sample or link.

Prior AI red-teaming or model evaluation experience is a plus but not required.

You will also be reading and writing about misuse scenarios in your field for sustained periods. We brief experts on this in advance, and you can pause or step away at any point without penalty.

Before you apply

Your work here will not involve, and must not draw on, classified or export-controlled information, or anything covered by an NDA or prepublication review obligation. If you hold such obligations you may still be a good fit — tell us in your application and we will scope the work accordingly.

Pay range
$65–$75/hr
40h/wk
Share
Earn $250 for each successful referral on this role.
Similar roles

You might also like

Mercor Verified New
remote · hourly
Central Bank Expert — Non-G7 Policy Communication
Life, Physical, and Social Science
Posted Sep 14, 2026
$150–$250/hr 20h/wk
Mercor Verified New
remote · hourly
Political Expert — Elections Forecasting & Political Risk
Life, Physical, and Social Science
Posted Sep 14, 2026
$150–$250/hr 20h/wk
micro1 Verified New
remote · hourly
Bilingual Mental Health Expert
Life, Physical, and Social Science
Posted Sep 14, 2026
$100–$200/hr
Mercor Verified New
remote · task-based
Forensic / Toxicological Chemist
Life, Physical, and Social Science
Posted Sep 14, 2026
$65–$75/hr