AI models can explain chemistry well and still misjudge a real experiment or dataset. This role puts your experience to work finding out where. You turn problems from your own work into realistic tasks, run them through frontier models, and judge the answers the way you would review a colleague's report. Some projects involve writing code. You set your own schedule, and the work rewards careful writing as much as expertise.
KEY RESPONSIBILITIES Build realistic tasks out of problems from your own chemistry work, along with the material needed to solve them Give those tasks to frontier AI models and judge the results against the standard of your profession Weigh paired model answers against each other and explain which one holds up Write grading rubrics for your tasks and set out your reasoning in writing Review and improve tasks written by other experts IDEAL QUALIFICATIONS Degree in chemistry, chemical engineering, or a related field, completed or in progress Hands-on lab, process, or analytical work Full professional or native-level written English Based in the United States, United Kingdom, Ireland, Canada, or Australia NICE TO HAVE Programming experience, for example in Python Have used Claude or ChatGPT in your professional work WHAT SUCCESS LOOKS LIKE You clear the skills assessment on your own work, with no AI assistance Your task proposals come from real problems in your own practice, and a frontier model genuinely fails them Your rubrics are specific enough that another expert would grade the same response the same way You act on reviewer feedback without needing to be walked through it You keep at least 10 hours a week going, week after week CONTRACT & PAYMENT TERMS
$40 to $125 USD an hour. The rate varies with your expertise, education, and experience. Minimum 10 hours a week with no weekly maximum; full-time hours are available.
Independent contractor, fully remote. Payment runs through PayPal, so you need a PayPal account. As an independent contractor, you handle your own taxes.
Getting on: a skills assessment that takes most people two to four hours. It is unpaid, and you get one attempt, so take your time and do your best work. Work produced with AI tools is rejected.
After you pass: onboarding and a short training project, then you propose a task from your own practice, build it, and run it against frontier models. Each stage is reviewed before you move on, and you are paid for the work you complete.