HumanitApp
  • Home
  • Roles
  • Talent networks
  • Weekly
  • Role Match
  • Tools
  • Saved
Submit CVGet weekly role alerts
← Back to all roles

Verified partner opportunity

Chemicals Safety for Redteaming

$65 - $75 / per-task

MercorRemote - location not specifiedTask-based

About the work

Is this a fit?

  • You can show recent, specific work involving AI evaluation, Clinical reasoning, Research.
  • The listed commitment of 40 hours/week fits your schedule.
  • You can communicate your reasoning clearly in the application language.

Apply on Mercor

Complete your application with Mercor. No HumanitApp account or fee.

What happens next

  1. Continue to the official listing and enter your application details; have your resume ready.
  2. Complete the role-specific screen or assessment requested by the partner.
  3. Confirm availability, location eligibility and work authorization if requested. The partner determines the exact process.

Last verified 2026-09-26

Exact listing verified. HumanitApp is independent from Mercor and may receive a referral fee; the partner controls assessment and hiring.

Read our Mercor review
Ready to continue?Chemicals Safety for Redteaming

Official role description

About the work

Mercor is assembling a panel of chemistry domain experts to red-team frontier AI models. The goal is to test whether a model can correctly judge the misuse potential of a technical request; answering legitimate questions fully while refusing genuinely dangerous ones.

You will: • Write challenging single-turn prompts in your domain, labelled across three levels: benign, dual-use, and adversarial. • Evaluate the responses that come back against a defined policy standard, and judge whether each was handled correctly. • Write the reference answer

what a correct response looks like, and the technical reasoning for why.

The hard part is calibration, not difficulty. A benign prompt that a model wrongly refuses is as valuable as an adversarial one it wrongly answers. That judgment is why we need domain experts rather than generalists

in chemistry the line sits between mechanism and procedure, and only someone who has run the reaction knows exactly where it falls. Background we're looking for

We're looking for professionals with hands-on experience in synthesis, analytical chemistry, or chemical safety. Prior AI evaluation or red-teaming experience is preferred. Relevant backgrounds include: • Chemical Defense Researcher (USAMRICD, DoD, CDC)

toxicological agents and medical countermeasures. • Synthetic / Process Chemist

route design, reaction engineering and scale-up. • Analytical Chemist

detection and identification methods for hazardous compounds. • Industrial Hygienist / Process Safety Engineer

exposure assessment, toxic release and process hazard analysis. • Forensic / Toxicological Chemist

controlled substances, precursor identification and casework.

What this role demands beyond technical depth

This is writing-intensive work. Every judgment you make needs a written rationale that a non-specialist can follow. Prior technical writing, published research, or expert witness experience is a strong signal

please include a sample or link.

Prior AI red-teaming or model evaluation experience is a plus but not required.

You will also be reading and writing about misuse scenarios in your field for sustained periods. We brief experts on this in advance, and you can pause or step away at any point without penalty.

Before you apply

Your work here will not involve, and must not draw on, classified or export-controlled information, or anything covered by an NDA or prepublication review obligation. If you hold such obligations you may still be a good fit

tell us in your application and we will scope the work accordingly.

Relevant skills

AI evaluationClinical reasoningResearch

Who this role may fit

This opportunity may suit professionals with relevant experience in AI evaluation, Clinical reasoning, Research. Review the official description and requirements before applying.

Compensation context

The listing states $65 - $75 / per-task. Confirm the final rate, workload, and payment terms during the official application process.

Qualification checklist

  • You can show recent, specific work involving AI evaluation, Clinical reasoning, Research.
  • The listed commitment of 40 hours/week fits your schedule.
  • You can communicate your reasoning clearly in the application language.
  • You are comfortable with project availability and hours varying over time.

Not the right fit?

Use your CV to find relevant roles on HumanitApp. This is separate from a partner application.

Match my CV →

New roles, once a week.

Optional role alerts from HumanitApp. Subscribe only if you want weekly emails.

Get weekly role alerts →

Before you apply

Referral and application questions

Can I start immediately through an Instant Work Offer?

This listing is not a promise of immediate work. Mercor may send separate offers to prequalified candidates. Read how Instant Work Offers work.

How should I compare the listed pay?

The listing states $65 - $75 / per-task. Confirm the final rate, workload, and payment terms during the official application process.

Does Apply use a referral link?

Yes. The button opens the exact verified Mercorlisting using the referral URL published for this role.

Is HumanitApp the employer?

No. HumanitApp independently curates the opportunity. The partner platform manages applications and hiring decisions.

Must I share my details here?

No. Choose the primary Apply action to continue directly without giving HumanitApp your name or email address.

Continue exploring

Related verified roles

View category
Policy & Safety

Radiologicals Safety for Redteaming

$65 - $75 / per-task

Policy & Safety

Local Political Expert — Elections Forecasting & Political Risk

$150 - $250 / hour

Policy & Safety

Central Bank Expert — Non-G7 Policy Communication

$150 - $250 / hour

Policy & Safety

Political Expert — Elections Forecasting & Political Risk

$150 - $250 / hour