Role directory

$100-$149/hr jobs

96 active, referral-verified opportunities.

Finance / Remote

Investment Banking Expert

Mercor is recruiting U.S./UK/Canada/Europe/Australia-based Investment Banking Experts for a research project with a leading foundational model AI lab. — You are a good fit if you: • Have at least 2 years of experience working at top firms in investment banking and experience in at least one of the following • Financial Modeling • Pitch Decks • Investment/Analysis Summaries and Memos • Company/Industry Analysis — Here are more details about the role: • You must be able to commit at least 10 hours per week for this role • This is a minimum four week engagement, with potential for significant extension or rotation to similar, future projects • Successful contributions increase the odds that you are selected on future projects with Mercor • This role will pay between $100-$130/hour with potential for increases for top performers

$100 - $130 / hourOpen / Referral verified
Finance / Remote

Patient Financial Clearance Leader

Mercor is working with a leading AI research lab to improve the capabilities of next-generation AI systems. We are seeking experienced Patient Financial Counselling and Financial Clearance leaders to support the evaluation of AI tools designed to enhance financial clearance operations and patient financial assistance workflows. Your expertise in charity care, self-pay collections, financial screening, and patient financial advocacy will help shape AI systems that improve patient access while optimising revenue capture. — Responsibilities • Lead patient financial counselling and financial clearance operations including charity care screening, financial assistance applications, and self-pay resolution. • Evaluate AI-generated financial counselling recommendations, eligibility screening outputs, and patient communication drafts for accuracy and appropriateness. • Screen patients for Medicaid eligibility, charity care qualification, and financial assistance program enrollment. • Counsel patients on financial obligations, payment plan options, and available assistance programs. • Coordinate with social work, case management, and billing teams to address complex patient financial situations. • Develop and maintain SOPs for financial clearance workflows, including pre-service financial screening and point-of-service collections. • Monitor KPIs including charity care conversion rates, financial assistance enrollment, and point-of-service collection performance. • Ensure compliance with regulatory requirements for financial assistance programs (501(r) regulations, EMTALA). • Annotate AI outputs and provide structured feedback to support AI training datasets. — Requirements • 5+ years of experience in patient financial counselling, financial clearance, or self-pay revenue cycle management, with at least 2 years in a leadership role. • Deep knowledge of charity care programs, financial assistance eligibility, Medicaid screening, and self-pay collections. • Familiarity with 501(r) regulatory requirements and hospital financial assistance policy compliance. • Experience with point-of-service collections, payment plan administration, and patient financial advocacy. • Proficiency with EHR platforms (Epic, Cerner) and financial assistance management tools. • Exceptional written and verbal English communication skills. • High attention to detail with the ability to evaluate financial documentation and AI-generated outputs. — Preferred Qualifications • Certified Revenue Cycle Representative (CRCR) or similar revenue cycle certification. • Experience with presumptive eligibility screening tools and Medicaid enrolment facilitation. • Background in hospital, health system, or federally qualified health centre (FQHC) settings. • Familiarity with AI tools and comfort evaluating AI-generated patient financial content. • Experience developing patient financial education materials and staff training programs. — Why Join? • Contribute to the development of frontier AI systems in healthcare. • Collaborate with a world-class AI research organisation. • Gain exposure to cutting-edge AI workflows in healthcare revenue cycle management. • Opportunity to work on high-impact projects shaping the future of healthcare AI.

$135 / hourOpen / Referral verified
Finance / Remote

Netherlands Domestic Tax Specialist

We are seeking experienced Netherlands Domestic Tax Specialists to support the development and evaluation of AI systems focused on Dutch taxation. You will leverage your expertise in Dutch domestic tax law to answer complex tax questions, review AI-generated responses, identify inaccuracies, and provide detailed feedback to improve model performance. This is an excellent opportunity for tax professionals who enjoy solving complex tax issues while contributing to the next generation of AI-powered tax solutions. — What You'll Do • Answer complex questions related to Dutch domestic tax law using relevant legislation, case law, and official guidance. • Review AI-generated tax analyses for legal accuracy, completeness, and clarity. • Correct incorrect or incomplete responses and provide structured explanations. • Evaluate reasoning across personal and corporate tax scenarios. • Flag ambiguous, conflicting, or insufficient legal guidance where applicable. • Consistently apply project rubrics and quality standards. • Collaborate with calibration and feedback sessions to improve evaluation consistency. — Qualifications • 5+ years of experience advising on Dutch domestic taxation. • Experience in a Big Four firm, tax advisory firm, law firm, corporate tax department, or the Dutch Tax Administration (Belastingdienst) preferred. • Strong knowledge of Dutch Income Tax, Corporate Tax, VAT, and related domestic legislation. • Ability to interpret Dutch tax statutes, regulations, and official guidance. Dutch tax professionals routinely advise on compliance, planning, disputes, and representation before the tax authorities. • Excellent analytical and written communication skills. • Comfortable reviewing detailed legal and tax reasoning. • Native or professional fluency in Dutch; strong English proficiency is preferred. — Preferred Qualifications • Registered Dutch tax adviser (e.g., NOB, RB, or equivalent professional qualification). • Master's degree in Tax Law, Fiscal Economics, Accounting, or a related discipline. • Experience handling complex tax controversies or tax planning matters. • Prior experience reviewing legal or tax content, conducting quality assurance, or contributing to AI/LLM evaluation projects. — Why Join? • Apply your Dutch tax expertise to cutting-edge AI systems. • Work on intellectually challenging tax scenarios across multiple domains. • Collaborate with a global network of legal and tax professionals.

$100 - $130 / hourOpen / Referral verified
Medical / Remote (United States)

Medicare Advantage Members (Devoted Health) – Insight Study

Mercor is conducting a paid research study in collaboration with a leading AI research lab focused on improving healthcare and member experiences. We are seeking current or former Devoted Health Medicare Advantage members/Age 64+ to participate in a short online survey about their experience with their health plan. Participants will share perspectives on plan enrollment, benefits, member support, and day-to-day healthcare experiences. Insights gathered will help inform the development of AI tools designed to improve how health plans serve their members. — Responsibilities — Participants will be asked to: • Complete a structured ~20-minute online survey • Provide a mix of multiple-choice answers and short voice-recorded responses (approximately 10 voice responses) • Share perspectives on their experience as a Medicare Advantage plan member • Submit all responses within the given timeline — Requirements • Based in the United States • Age 64+ / Medicare-eligible • Currently or previously enrolled in a Devoted Health Medicare Advantage plan • Comfortable providing voice-recorded responses • Access to a microphone and a quiet environment • Able to independently complete a ~20-minute online survey — Engagement Details • Format: Online survey (voice-recorded responses + multiple-choice questions) • Duration: Approximately 20 minutes • Compensation: One-time payment upon successful verification of the submission • Location: Remote (United States) — Why Participate • Contribute to research shaping the next generation of healthcare AI tools • Share your real-world experience as a Medicare Advantage plan member

$120 / hourOpen / Referral verified
STEM / Remote

Research Physics Expert

Role Overview — We are seeking expert physics researchers to author and verify golden reference solutions for the CritPt benchmark (arXiv:2509.26574v3) — a frontier research-level physics benchmark. Participants will solve CritPt research-level problems end-to-end, audit solutions from other experts, or adjudicate between parallel solution attempts, producing 100%-human-verified reference data used to evaluate large language models on frontier physics reasoning. — Physics Subdomains Covered — High Energy Physics & Mathematical Physics, Biophysics & Statistical Physics, Condensed Matter & AMO, Gravitation / Cosmology / Astrophysics, Quantum Information, Optical Properties of Materials, Magnetic Materials, Measurements in QM. — Key Responsibilities • Solve research-level physics challenges end-to-end with verifiable derivations, code, and peer-reviewed references • Decompose challenges into standalone checkpoint sub-problems that require genuine physical reasoning • Author Python answer templates with auto-grading functions for symbolic or numerical answers • Audit submitted solutions for correctness, scope, and method soundness; deliver actionable feedback across iterations • Adjudicate between parallel solver attempts and decide which solution becomes the golden reference • Document chain-of-thought reasoning, error tolerances, equivalent symbolic forms, and verification test cases — Ideal Qualifications • Solver: PhD or postdoc in the relevant subfield (senior PhD student minimum) • Auditor: Postdoc or junior professor in the relevant subfield (PhD minimum) • Adjudicator: Full professor or industry research PI in the relevant subfield (senior postdoc or junior professor minimum) • Hands-on familiarity with at least two canonical methods of the target subfield, demonstrable through publications (broader coverage strongly preferred) • 3–5 representative publications (arXiv ID or DOI), ideally within the last ~5 years and in the target subfield • Working proficiency with LaTeX, Python, Jupyter, and SymPy • Strong written English (B2/C1/C2 minimum; native or near-native preferred) — More About the Opportunity • Expected commitment: ~10 hours/week, sustained across an 8–10 week window per task pool • Pay range: $80–$135 per hour, based on role and demonstrated expertise • Asynchronous work

$80 - $135 / hourOpen / Referral verified
Finance / Remote

Domain Expert – Banking (Oracle FLEXCUBE Research)

Domain Expert – Banking (Oracle FLEXCUBE Research) — To qualify you must use Oracle FLEXCUBE regularly — weekly or more — as a working part of your job, with 2+ years of professional banking experience. — FLEXCUBE knowledge we require — _Credit & exposure_ • Calculating current exposure against the bank's credit standard as at a given business date; arrears and headroom calculation • Distinguishing a validation-only posting from a booked transaction against a contract • Reading limits/facility structures; recognizing when an item needs committee sign-off vs. branch-level approval — _Customer & relationship records_ • Customer Information File (CIF) structure and relationship hierarchies • Recording contact/interaction history accurately, including on partial or handed-off information • Setting and interpreting review dates and follow-up flags on a customer file — _Contracts & core banking operations_ • Contract inquiry and lifecycle screens (origination, drawdown, repayment, release) • Branch lending repayment validation vs. contract booking • Distinguishing what's confirmed vs. assumed when picking up a file mid-process — _Reference & product data_ • Product parameter tables, GL mapping, and how product configuration constrains what a transaction can do • Settlement/correspondent references where relevant to a contract — _Regulatory & credit standards_ • Applying the bank's own credit standard/policy to determine current risk position • Version/currency awareness — which policy or limit applies as of the current business date — _Audit & handoff_ • Recording validations, conclusions, and open items so the next person in the workflow can see exactly what was decided and what's still outstanding • MIS / reporting screens relevant to exposure and customer position — What you'll do • Confirm the customer/contract position reflects what actually happened, not what's assumed, especially on partial handoffs • Work out current exposure under the bank's credit standard as at the current business date, and record it with a review date • Distinguish validation-only actions from bookings that affect a contract • Judge completeness: is the full position understood, is anything left open that shouldn't be, is the record ready for the next reviewer or for committee • Verify recorded outcomes are traceable — auditable, not just noted informally — Requirements • Frequent Oracle FLEXCUBE use — weekly or more — in your current or recent role • 2+ years of professional experience in banking operations, credit, or a related core-banking function • Reflexive familiarity with credit standards, exposure calculation, and validation-vs-booking distinctions • Bachelor's degree in Finance, Accounting, Business, or a related field; a banking/credit certification (e.g., Chartered Banker – CIOB, Certified Credit Professional) or equivalent hands-on core-banking experience

$75 - $120 / hourOpen / Referral verified
Code / Remote

LLM Research Scientist (Pre-training & Computer Vision & Adversarial Robustness)

We're looking for experienced machine learning researchers with hands-on experience training and improving deep learning models end-to-end, across vision and language. You'll work on well-scoped empirical open-ended ML research problems. — Responsibilities • Train image classifiers and generative image models from scratch, and fine-tune open-weight language models. • Get the most out of limited data, compute, and model-size budgets. • Make models robust — to adversarial inputs and to adversarial conversations. • Compress models to meet hard size and latency constraints without sacrificing accuracy. • Diagnose and resolve training issues. — Requirements — We are looking for candidates with strong expertise in one or more of the following areas: — Adversarial Robustness — Experience with: • Adversarial training of image classifiers (e.g. PGD-based training, TRADES). • Evaluating robust accuracy under standard threat models (e.g. L∞ attacks, AutoAttack) and avoiding gradient-masking pitfalls. • Managing the robustness–accuracy trade-off and robust overfitting. — Efficient Computer Vision — Experience with: • Training image classifiers end-to-end, especially for fine-grained recognition (many visually similar classes, few examples per class). • Model compression: quantization, pruning, and knowledge distillation from large teachers into small students. • Deploying models under hard size or latency budgets (on-device, edge, or embedded settings). — Generative Image Modeling — Experience with: • Training image generative models from scratch: diffusion models, GANs, VAEs, or flow-based models. • Iterating against sample-quality metrics such as FID. • Training-efficiency tricks that produce good generators quickly and at small parameter counts. — LLM Post-Training & Behavioral Robustness — Hands-on experience with one or more of: • Supervised fine-tuning and preference optimisation (DPO, RLHF, RLAIF) of open-weight language models, including building your own datasets via synthetic generation, noisy or weak supervision, and rejection sampling. • Shaping conversational behaviour over multiple turns: resistance to persuasion and sycophancy, calibrated confidence, and knowing when to accept corrections. • Alignment-style fine-tuning that changes a specific behaviour while preserving general capability. — Multilingual Pre-training — Experience with: • Training multilingual or low-resource-language models from scratch. • Tokenizer design across scripts and typologically diverse languages. • Balancing highly unequal per-language data (sampling temperatures, cross-lingual transfer) in data-constrained regimes. — Additional Areas of Interest — Experience in any of the following is a plus: • Scaling laws and training-efficiency research. • Curriculum learning and data ordering. • Model evaluation: benchmark construction, contamination control, statistically sound comparisons. • Uncertainty estimation and model calibration. • Data augmentation and synthetic data for robustness. — General Qualifications • 3+ years of machine learning research experience (PhD research counts toward this requirement). • Strong experience with PyTorch, JAX, TensorFlow, or similar ML frameworks. • Degree from a top-100 university, experience at a FAANG or comparable AI company, or an equivalent research track record through publications or impactful open-source contributions. — Why Join • Work on cutting-edge machine learning research. • Collaborate with leading AI researchers on challenging, high-impact projects. • Flexible, project-based work with competitive compensation.

$100 - $120 / hourOpen / Referral verified
Language / Australia (remote)

In-House Counsel (Australia)

Mercor is partnering with a leading legal AI company to benchmark how well AI systems answer real questions of Australian law. We are hiring experienced in-house counsel to serve as expert reviewers — running the same prompts through two AI platforms, comparing the answers, and scoring them against a standardized rubric. — What you'll do • Run the same set of prompts across two AI (LLM) platforms • Compare the outputs from both platforms side by side • Score each response against a standardized rubric we provide • Submit concise written feedback for every evaluation — Requirements • 5+ years as in-house legal counsel practising in Australia • Strong command of Australian law, especially commercial and contract matters • Sharp attention to detail and precise written English • Able to work independently to a rubric and to deadlines — Logistics • Initial commitment: ~10 hours, with strong potential for more • Confidentiality: all work is covered by NDA

$100 - $120 / hourOpen / Referral verified
Language / Remote

Pricing / ROI / revenue economics Evaluator

About the role — We are hiring expert Evaluators in Pricing / ROI / revenue economics to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. — This is a remote, hourly engagement. — Requirements (must have) — 1. 5+ years of relevant professional experience in Pricing / ROI / revenue economics. 2. Native or professional fluency in English. 3. Highly proficient in Microsoft Office and Google Workspace, especially Slides (Google Slides / PowerPoint). — Preferred (nice to have) • Advanced degree (Master's or higher) from a reputable institution. — What you'll do • Evaluate AI-generated artifacts against domain-specific quality rubrics. • Identify factual, aesthetic, and presentation errors. • Provide clear, structured written feedback.

$80 - $120 / hourOpen / Referral verified
Business / Remote

Product management / roadmap / PRD Evaluator

About the role — We are hiring expert Evaluators in Product management / roadmap / PRD to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. — This is a remote, hourly engagement. — Requirements (must have) — 1. 5+ years of relevant professional experience in Product management / roadmap / PRD. 2. Native or professional fluency in English. 3. Highly proficient in Microsoft Office and Google Workspace, especially Slides (Google Slides / PowerPoint). — Preferred (nice to have) • Advanced degree (Master's or higher) from a reputable institution. — What you'll do • Evaluate AI-generated artifacts against domain-specific quality rubrics. • Identify factual, aesthetic, and presentation errors. • Provide clear, structured written feedback.

$80 - $120 / hourOpen / Referral verified
Business / Remote

Public-sector procurement / RFI response Evaluator

About the role — We are hiring expert Evaluators in Public-sector procurement / RFI response to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. — This is a remote, hourly engagement. — Requirements (must have) — 1. 5+ years of relevant professional experience in Public-sector procurement / RFI response. 2. Native or professional fluency in English. 3. Highly proficient in Microsoft Office and Google Workspace, especially Slides (Google Slides / PowerPoint). — Preferred (nice to have) • Advanced degree (Master's or higher) from a reputable institution. — What you'll do • Evaluate AI-generated artifacts against domain-specific quality rubrics. • Identify factual, aesthetic, and presentation errors. • Provide clear, structured written feedback.

$80 - $120 / hourOpen / Referral verified
Language / Remote

User/customer research and feedback synthesis Evaluator

About the role — We are hiring expert Evaluators in User/customer research and feedback synthesis to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. — This is a remote, hourly engagement. — Requirements (must have) — 1. 5+ years of relevant professional experience in User/customer research and feedback synthesis. 2. Native or professional fluency in English. 3. Highly proficient in Microsoft Office and Google Workspace, especially Slides (Google Slides / PowerPoint). — Preferred (nice to have) • Advanced degree (Master's or higher) from a reputable institution. — What you'll do • Evaluate AI-generated artifacts against domain-specific quality rubrics. • Identify factual, aesthetic, and presentation errors. • Provide clear, structured written feedback.

$80 - $120 / hourOpen / Referral verified
Language / Remote

BI dashboards / performance reporting Evaluator

About the role — We are hiring expert Evaluators in BI dashboards / performance reporting to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. — This is a remote, hourly engagement. — Requirements (must have) — 1. 5+ years of relevant professional experience in BI dashboards / performance reporting. 2. Native or professional fluency in English. 3. Highly proficient in Microsoft Office and Google Workspace, especially Slides (Google Slides / PowerPoint). — Preferred (nice to have) • Advanced degree (Master's or higher) from a reputable institution. — What you'll do • Evaluate AI-generated artifacts against domain-specific quality rubrics. • Identify factual, aesthetic, and presentation errors. • Provide clear, structured written feedback.

$80 - $120 / hourOpen / Referral verified
Language / Remote

Nonprofit / philanthropy / community programs Evaluator

About the role — We are hiring expert Evaluators in Nonprofit / philanthropy / community programs to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. — This is a remote, hourly engagement. — Requirements (must have) — 1. 5+ years of relevant professional experience in Nonprofit / philanthropy / community programs. 2. Native or professional fluency in English. 3. Highly proficient in Microsoft Office and Google Workspace, especially Slides (Google Slides / PowerPoint). — Preferred (nice to have) • Advanced degree (Master's or higher) from a reputable institution. — What you'll do • Evaluate AI-generated artifacts against domain-specific quality rubrics. • Identify factual, aesthetic, and presentation errors. • Provide clear, structured written feedback.

$80 - $120 / hourOpen / Referral verified
Legal / Australia (remote)

Litigation Lawyer (Australia)

Mercor is partnering with a leading legal AI company to benchmark how well AI systems answer real questions of Australian law. We are hiring experienced litigation & disputes lawyers to serve as expert reviewers — running the same prompts through two AI platforms, comparing the answers, and scoring them against a standardized rubric. — What you'll do • Run the same set of prompts across two AI (LLM) platforms • Compare the outputs from both platforms side by side • Score each response against a standardized rubric we provide • Submit concise written feedback for every evaluation — Requirements • Primarily law-firm / chambers experience — at least 2 years (3+ preferred) at a law firm or barristers' chambers in litigation, disputes, or contentious practice. If you are currently in-house, that is fine only if you moved in-house within the last 2 years and your prior experience was at a firm/chambers. • Qualified to practise in Australia, with a strong command of Australian litigation and procedural law • Native or fluent English, with precise written communication • Able to work independently to a rubric and to deadlines — Logistics • Initial commitment: ~10 hours, with strong potential for more • Confidentiality: all work is covered by NDA

$100 - $120 / hourOpen / Referral verified
Legal / Australia (remote)

Corporate/M&A Lawyer (Australia)

Mercor is partnering with a leading legal AI company to benchmark how well AI systems answer real questions of Australian law. We are hiring experienced corporate/M&A lawyers to serve as expert reviewers — running the same prompts through two AI platforms, comparing the answers, and scoring them against a standardized rubric. — What you'll do • Run the same set of prompts across two AI (LLM) platforms • Compare the outputs from both platforms side by side • Score each response against a standardized rubric we provide • Submit concise written feedback for every evaluation — Requirements • Primarily law-firm experience — at least 2 years (3+ preferred) at a law firm in corporate/M&A or transactional practice. If you are currently in-house, that is fine only if you moved in-house within the last 2 years and your prior experience was at a law firm. • Qualified to practise in Australia, with a strong command of Australian corporate and commercial law • Native or fluent English, with precise written communication • Able to work independently to a rubric and to deadlines — Logistics • Initial commitment: ~10 hours, with strong potential for more • Confidentiality: all work is covered by NDA

$100 - $120 / hourOpen / Referral verified
Legal / Bay Area, CA

Senior Counsel — Legal Domain Expert

Join a leading AI lab's cutting-edge GenAI team to be at the core of the AI revolution, where your expertise fuels the development of the most advanced AI models. — 1. Overview — We are hiring an experienced counsel-level lawyer to work directly with a leading AI lab's research and program management teams, improving how frontier AI models reason about real legal work. — This listing is for practitioners at counsel or senior associate level, with roughly 8 to 15 years of post-qualification practice. If you are earlier in your career, or if you are at partner or General Counsel level, please apply to the corresponding Legal Domain Expert listing instead — we run separate listings by seniority so the rate matches the experience. — Your legal expertise is the substance of this role. You will review the quality of legal knowledge work tasks, write the instruction specs and golden solutions that define what "correct" looks like, and build the benchmarks that show whether the model is genuinely improving. We are looking for a practising specialist rather than a generalist. — This is a full-time W-2 employment position with Cincinnatus LLC, with the opportunity to be placed at a leading AI lab as part of their extended workforce. You will be provisioned with client-issued accounts and equipment, and will work inside the client's own tools alongside their research teams. — Location: This is a hybrid role based in the Bay Area, California. You must live in the Bay Area and work on-site with the client's team multiple days each week, when required. This is not a remote role. If you do not currently live in the Bay Area, you must be willing to relocate there at your own cost before the engagement starts — relocation assistance is not provided. — 2. Key Responsibilities • Data QA and reviews: Vet the quality of legal knowledge work tasks and model outputs — spotting missing behaviors, thin reasoning, and answers that read well but would not survive professional scrutiny. • Instruction specs and golden datasets: Write high-quality instruction specs, produce golden solutions to legal problems, and define new legal tasks that reflect how the work is actually done in practice. • Benchmarks and domain depth: Design challenging legal tasks and evaluation sets, and help build legal-specific skills and tools together with the research team. • Calibration: Work with client researchers and specialists in adjacent fields to keep standards consistent, translating tacit legal judgment into explicit, teachable criteria. — 3. Core Qualifications • Education: Juris Doctor (JD) from an accredited law school; a highly ranked school is strongly preferred. • Experience: 8 to 15 years of substantive post-qualification legal practice at a reputable institution — an established law firm, a corporate legal department, a regulatory body, or a court. Internships and clerkships alone do not count. • Seniority: Currently at or has reached counsel, senior associate, senior in-house counsel, or Assistant General Counsel level, with real ownership of matters. • Domain depth: Genuine specialization in at least one substantive practice area, for example corporate and transactional, litigation and dispute resolution, regulatory and compliance, intellectual property, employment and labor, or tax. • Licensure: Admission to at least one U.S. state bar, in good standing. • AI fluency: Hands-on working use of large language models in your professional work, and the judgment to tell a well-reasoned answer from a plausible-sounding wrong one. • Availability: Able to commit reliably to 40 hours per week for an initial engagement of 6 months. • Location: Living in the Bay Area, California, and able to work on-site with the client's team multiple days each week, when required. Candidates not currently based in the Bay Area must be willing to relocate there at their own cost; relocation assistance is not provided. • Excellent written communication, and the ability to give precise, well-structured written feedback. — About Cincinnatus LLC — Cincinnatus LLC is an enterprise staffing company that partners with leading technology companies to source and employ highly skilled professionals for contingent and contract-based opportunities. Cincinnatus serves as the employer of record for these engagements, providing W-2 employment, payroll, benefits, and compliance, while placing employees directly within client teams to work on high-impact initiatives. — Roles hired through Cincinnatus are not project-based or freelance engagements. They are structured, role-based positions that typically involve part-time or full-time commitments, close collaboration with a client's internal teams, and integration into standard enterprise workflows. — Cincinnatus is a legal entity separate from Mercor. While opportunities may be discovered through Mercor's platform, employment, onboarding, payroll, and benefits for these roles are administered by Cincinnatus LLC. — Equal Employment Opportunity — Cincinnatus is proud to be an Equal Employment Opportunity employer. We do not discriminate based upon race, religion, color, national origin, sex (including pregnancy, childbirth, reproductive health decisions, or related medical conditions), sexual orientation, gender identity, gender expression, age, status as a protected veteran, status as an individual with a disability, genetic information, political views or activity, or any other legally protected characteristic. — Cincinnatus is committed to providing reasonable accommodations for qualified individuals with disabilities and disabled veterans throughout the job application process.

$85 - $120 / hourOpen / Referral verified
Business / Remote

Operations / inventory / capacity planning Evaluator

About the role — We are hiring expert Evaluators in Operations / inventory / capacity planning to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. — This is a remote, hourly engagement. — Requirements (must have) — 1. 5+ years of relevant professional experience in Operations / inventory / capacity planning. 2. Native or professional fluency in English. 3. Highly proficient in Microsoft Office and Google Workspace, especially Slides (Google Slides / PowerPoint). — Preferred (nice to have) • Advanced degree (Master's or higher) from a reputable institution. — What you'll do • Evaluate AI-generated artifacts against domain-specific quality rubrics. • Identify factual, aesthetic, and presentation errors. • Provide clear, structured written feedback.

$80 - $120 / hourOpen / Referral verified
STEM / Remote

Biology / environmental science Evaluator

About the role — We are hiring expert Evaluators in Biology / environmental science to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. — This is a remote, hourly engagement. — Requirements (must have) — 1. 5+ years of relevant professional experience in Biology / environmental science. 2. Native or professional fluency in English. 3. Highly proficient in Microsoft Office and Google Workspace, especially Slides (Google Slides / PowerPoint). — Preferred (nice to have) • Advanced degree (Master's or higher) from a reputable institution. — What you'll do • Evaluate AI-generated artifacts against domain-specific quality rubrics. • Identify factual, aesthetic, and presentation errors. • Provide clear, structured written feedback.

$80 - $120 / hourOpen / Referral verified
Business / Remote

Incident management / reliability / SRE Evaluator

About the role — We are hiring expert Evaluators in Incident management / reliability / SRE to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. — This is a remote, hourly engagement. — Requirements (must have) — 1. 5+ years of relevant professional experience in Incident management / reliability / SRE. 2. Native or professional fluency in English. 3. Highly proficient in Microsoft Office and Google Workspace, especially Slides (Google Slides / PowerPoint). — Preferred (nice to have) • Advanced degree (Master's or higher) from a reputable institution. — What you'll do • Evaluate AI-generated artifacts against domain-specific quality rubrics. • Identify factual, aesthetic, and presentation errors. • Provide clear, structured written feedback.

$80 - $120 / hourOpen / Referral verified
Medical / Remote

Healthcare / clinical Evaluator

About the role — We are hiring expert Evaluators in Healthcare / clinical to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. — This is a remote, hourly engagement. — Requirements (must have) — 1. 5+ years of relevant professional experience in Healthcare / clinical. 2. Native or professional fluency in English. 3. Highly proficient in Microsoft Office and Google Workspace, especially Slides (Google Slides / PowerPoint). — Preferred (nice to have) • Advanced degree (Master's or higher) from a reputable institution. — What you'll do • Evaluate AI-generated artifacts against domain-specific quality rubrics. • Identify factual, aesthetic, and presentation errors. • Provide clear, structured written feedback.

$80 - $120 / hourOpen / Referral verified
Medical / Remote

Applied Health & Medicine Benchmark Specialist

Role Overview — We are seeking expert medical and health science professionals to author and review high-quality academic assessment content for an AI research initiative. You will write and verify rigorous multiple-choice questions across core health and medicine domains, evaluate solution quality, and help establish gold-standard benchmarks used to advance AI capabilities. — You will be assigned one of two task types: • Question Authoring — Create original, challenging multiple-choice questions in your area of medical expertise, rate their difficulty, and submit them for review. • Question Verification — Review pre-written questions for accuracy, clarity, and rigor. Edit where needed, rate difficulty, and document any changes made. — Health & Medicine Domains Covered — Clinical Medicine & Surgery, Medical Imaging & Diagnostics, Pharmacovigilance, Healthcare Management & Economics, Rehabilitation and Allied Health. — Key Responsibilities • Author original health and medicine questions that test deep conceptual understanding, not surface-level recall • Ensure questions are unambiguous, self-contained, and precisely defined — all necessary information must be in the problem statement • Rate each question's difficulty: Medium (intro undergraduate), Hard (advanced undergraduate), or Expert (post-graduate and above) • Provide 1 correct answer and 9 plausible but subtly incorrect alternatives that challenge expert-level solvers • Write step-by-step Chain-of-Thought solutions with clear, concise reasoning in markdown format • Supply 1–5 academic references per question from reputable sources (peer-reviewed journals, clinical guidelines) • For verification tasks: flag issues with clarity, completeness, precision, or solvability and justify any edits made — Ideal Qualifications • MD, DO, PhD, or doctoral candidate in Medicine, Biomedical Sciences, Public Health, or a closely related field • Master's degree considered for candidates with exceptional depth in a specific subdomain • Strong command of graduate-level medical knowledge, clinical reasoning, and biomedical research methodology • Board certification, clinical experience, or research publications in health fields is a strong plus • Excellent written English and ability to express complex ideas clearly and concisely — More About the Opportunity • Expected commitment: 10+ hours/week • Asynchronous, fully remote work

$94 - $119 / hourOpen / Referral verified
Finance / Remote

Tax Accountant / Specialist

Role Overview — Mercor is collaborating with a leading AI lab to engage experienced tax professionals. You'll translate real tax compliance, provision, and planning work into structured, high-quality training data that teaches AI to reason about how tax accountants actually work. — Focus Areas — Tax — across corporate, individual, partnership / pass-through, SALT, and international sub-specialties. — Key Responsibilities • Design realistic scenarios from your tax work — return preparation (1040 / 1120 / 1120-S / 1065), tax provision (ASC 740), responses to notices / IDRs and controversy support, tax planning and estimated payments • Review and compare AI-generated tax outputs for technical accuracy and defensible positions • Provide clear written feedback that improves how AI performs tax tasks • Collaborate asynchronously with the research team — Ideal Qualifications • CPA or EA (Enrolled Agent) • A clear tax sub-specialty — corporate, individual, partnership, SALT, or international / cross-border • Public accounting (firm) or in-house tax experience • Bachelor's degree in Accounting, Finance, or a related field • Strong written communication and attention to detail — Application Process • Submit a resume or a short summary of your tax experience • Complete a short form on your practice area, specialties, and certifications • Selected applicants may complete a brief sample task

$80 - $120 / hourOpen / Referral verified
Finance / Remote

Corporate / Controllership Accountant

Role Overview — Mercor is collaborating with a leading AI lab to engage experienced corporate / controllership accountants. You'll translate the day-to-day of running the books — financial reporting, the month-end close, and core accounting operations — into structured, high-quality training data that teaches AI to reason about real accounting work. — Focus Areas — Financial accounting & reporting · bookkeeping / client accounting services (CAS) · accounting operations (AP / AR / payroll). — Key Responsibilities • Design realistic scenarios from your controllership work — financial statement preparation, journal entries, account & bank reconciliations, month-end / period-end close, fixed-asset & depreciation schedules, accruals and prepaids • Review and compare AI-generated accounting outputs for accuracy, GAAP compliance, and sound professional judgment • Provide clear written feedback that improves how AI performs close and reporting tasks • Collaborate asynchronously with the research team — Ideal Qualifications • CPA, or 3–5+ years of corporate/in-house or firm accounting experience (no niche certification required) • Hands-on ownership of the general ledger and the period-end close • Bachelor's degree in Accounting, Finance, or a related field • Comfortable with common accounting tools (QuickBooks, NetSuite, SAP, Oracle, Excel) • Strong written communication and attention to detail — Application Process • Submit a resume or a short summary of your accounting experience • Complete a short form on your practice area, specialties, and certifications • Selected applicants may complete a brief sample task

$80 - $120 / hourOpen / Referral verified
Finance / Remote

Accounting Expert

Role Overview — Mercor is collaborating with a leading AI lab to engage experienced accounting professionals across all areas of practice — including audit, tax, financial reporting, bookkeeping, controllership, forensic, and accounting systems. Contributors help build AI systems that reason about real accounting work by translating everyday accounting workflows, judgments, and decision-making into structured, high-quality training data. — Key Responsibilities • Design realistic accounting scenarios and tasks drawn from your day-to-day work (e.g., financial statement preparation, reconciliations, journal entries, audit procedures, tax filings, month-end close, internal controls) • Review and compare AI-generated accounting outputs for accuracy, standards compliance (GAAP/IFRS), and sound professional judgment • Create structured examples that reflect how accountants actually reason through problems • Provide clear written feedback that improves how AI performs accounting tasks • Collaborate asynchronously with the research team — Ideal Qualifications • 3+ years of professional experience in accounting, audit, tax, bookkeeping, or finance operations (public accounting, corporate/in-house, advisory, or a firm) • CPA, CA, ACCA, CMA, EA, or an equivalent professional credential required • Bachelor's degree in Accounting, Finance, or a related field • Comfortable with common accounting tools (e.g., QuickBooks, NetSuite, SAP, Oracle, Excel) • Strong written communication and attention to detail — More About the Opportunity • Open to all accounting specialties — contribute where your expertise is strongest • Work spans task design, evaluation, and structured feedback on AI accounting outputs • Strong contributors advance into reviewer, lead, and domain-expert roles — Application Process • Submit a resume or a short summary of your accounting experience • Complete a short form on your practice area, specialties, and certifications • Selected applicants may complete a brief sample task • Follow-up typically provided within a few days

$80 - $120 / hourOpen / Referral verified
Math / Remote

Data analysis / quantitative readouts Evaluator

About the role — We are hiring expert Evaluators in Data analysis / quantitative readouts to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. — This is a remote, hourly engagement. — Requirements (must have) — 1. 5+ years of relevant professional experience in Data analysis / quantitative readouts. 2. Native or professional fluency in English. 3. Highly proficient in Microsoft Office and Google Workspace, especially Slides (Google Slides / PowerPoint). — Preferred (nice to have) • Advanced degree (Master's or higher) from a reputable institution. — What you'll do • Evaluate AI-generated artifacts against domain-specific quality rubrics. • Identify factual, aesthetic, and presentation errors. • Provide clear, structured written feedback.

$80 - $120 / hourOpen / Referral verified
Business / Remote

Healthcare operations Evaluator

About the role — We are hiring expert Evaluators in Healthcare operations to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. — This is a remote, hourly engagement. — Requirements (must have) — 1. 5+ years of relevant professional experience in Healthcare operations. 2. Native or professional fluency in English. 3. Highly proficient in Microsoft Office and Google Workspace, especially Slides (Google Slides / PowerPoint). — Preferred (nice to have) • Advanced degree (Master's or higher) from a reputable institution. — What you'll do • Evaluate AI-generated artifacts against domain-specific quality rubrics. • Identify factual, aesthetic, and presentation errors. • Provide clear, structured written feedback.

$80 - $120 / hourOpen / Referral verified
Code / United States Remote

Senior Software Engineer, Full Stack (Python, Java, Rust, C#, C++)

Join a leading AI lab's cutting-edge GenAI team to be at the core of the AI revolution, where your expertise fuels the development of the most advanced Large Language Models. — 1. Overview — Join a leading AI lab's cutting-edge GenAI team and help build foundational AI models from the ground up. We're seeking talented Senior Full-Stack Software Engineers with deep, hands-on production expertise across modern application stacks — Python, Java, Rust, C#, C++, and TypeScript — to build, integrate, and stress-test real software on top of frontier models before they ship. Engineers work directly with the AI lab's engineering manager on fast-moving projects and are expected to onboard quickly and contribute working code early. This is a full-time commitment of 40 hours per week. — This is a W-2 employment position with Cincinnatus LLC, with the opportunity to be placed at a leading AI Lab as part of their extended workforce. — 2. Key Responsibilities • Build and ship full-stack applications, services, and internal tools — in Python, Java, Rust, C#, C++, or TypeScript — that exercise frontier model capabilities end to end. • Integrate pre-release model APIs into working software, including the surrounding scaffolding: tool interfaces, evaluation harnesses, and telemetry. • Diagnose and document model and integration failure modes surfaced while building, translating engineering observations into clear written feedback for the research team. • Prototype quickly against shifting requirements, delivering working increments with limited up-front specification. • Collaborate with the engineering manager and other engineers to maintain consistency in code quality, architecture, and technical documentation. — 3. Core Qualifications • 6+ years of dedicated professional experience building and shipping production software at a recognized, top-tier organization, including ownership of systems end to end. • Deep production experience in at least one of Python, Java, Rust, C#, or C++, plus demonstrated delivery in a second language across a different ecosystem — this cohort is intentionally staffed across multiple language ecosystems. • Hands-on ownership across the stack: backend service and API design, a modern front-end framework (React or equivalent), data modelling, and cloud deployment and operations. • Demonstrated ability to onboard onto an unfamiliar codebase and deliver working code quickly with minimal ramp-up. • Demonstrable career progression. • Ability to engage reliably for at least 40 hours/week during weekdays. • Strong written communication skills and the ability to explain complex technical decisions clearly. — About Cincinnatus LLC: Cincinnatus LLC is an enterprise staffing company that partners with leading technology companies to source and employ highly skilled professionals for contingent and contract-based opportunities. Cincinnatus serves as the employer of record for these engagements, providing W-2 employment, payroll, benefits, and compliance, while placing employees directly within client teams to work on high-impact initiatives. — Equal Employment Opportunity: Cincinnatus is proud to be an Equal Employment Opportunity employer. We do not discriminate based upon race, religion, color, national origin, sex (including pregnancy, childbirth, reproductive health decisions, or related medical conditions), sexual orientation, gender identity, gender expression, age, status as a protected veteran, status as an individual with a disability, genetic information, political views or activity, or any other legally protected characteristic.

$90 - $110 / hourOpen / Referral verified
Code / United States Remote

MLOps Engineer (JAX, PyTorch, Pallas/Triton)

Join a leading AI lab's cutting-edge GenAI team to be at the core of the AI revolution, where your expertise fuels the development of the most advanced Large Language Models. — 1. Overview — Join a leading AI lab's cutting-edge GenAI team and help build foundational AI models from the ground up. We're seeking talented MLOps Engineers with deep, hands-on expertise in modern ML frameworks — specifically JAX, PyTorch, and kernel-level programming (Pallas/Triton). This role involves AI model training and evaluation work, including writing and assessing MLOps tasks and solutions to generate high-quality training data for frontier AI systems. — This is a W-2 employment position with Cincinnatus LLC, with the opportunity to be placed at a leading AI Lab as part of their extended workforce. This is a 40-hour full-time engagement, with no conflicts/no other engagements. — 2. Key Responsibilities • Guide research and engineering teams to close knowledge gaps and improve AI model performance in MLOps, training infrastructure, and ML framework-level topics. • Design challenging, domain-relevant tasks, and write accurate and well-structured solutions to MLOps and ML systems problems. • Evaluate MLOps tasks and solutions and provide clear, written technical feedback. • Develop guidelines and detailed rubrics/evaluation frameworks to assess training pipeline design, distributed systems reasoning, and kernel-level optimization across tasks. • Collaborate with other subject matter experts to ensure consistency and accuracy in training data. — 3. Core Qualifications • 2+ years of dedicated professional experience in ML infrastructure, MLOps, or ML systems engineering at a recognized, top-tier organization. • Hands-on production experience with JAX and/or PyTorch at scale. • Experience writing or optimizing custom GPU kernels using Pallas (JAX) or Triton. • Demonstrable career progression. • Ability to engage reliably for at least 40 hours/week during weekdays. • Strong written communication skills and the ability to explain complex technical decisions clearly. — About Cincinnatus LLC: Cincinnatus LLC is an enterprise staffing company that partners with leading technology companies to source and employ highly skilled professionals for contingent and contract-based opportunities. Cincinnatus serves as the employer of record for these engagements, providing W-2 employment, payroll, benefits, and compliance, while placing employees directly within client teams to work on high-impact initiatives. — Equal Employment Opportunity: Cincinnatus is proud to be an Equal Employment Opportunity employer. We do not discriminate based upon race, religion, color, national origin, sex (including pregnancy, childbirth, reproductive health decisions, or related medical conditions), sexual orientation, gender identity, gender expression, age, status as a protected veteran, status as an individual with a disability, genetic information, political views or activity, or any other legally protected characteristic.

$70 - $110 / hourOpen / Referral verified
Code / United States Remote

Performance Engineer (C++, Python, Rust)

Join a leading AI lab's cutting-edge GenAI team to be at the core of the AI revolution, where your expertise fuels the development of the most advanced Large Language Models. — 1. Overview — We're seeking talented Performance Engineers with deep expertise in low-level systems optimization — specifically C++, Python, and Rust — to bring hands-on technical excellence and elevate the quality of our AI training and inference infrastructure data. This role involves AI model training and evaluation work, including writing and assessing performance-engineering tasks and solutions to generate high-quality training data for frontier AI systems. — This is a W-2 employment position with Cincinnatus LLC, with the opportunity to be placed at a leading AI Lab as part of their extended workforce. This is a 40-hour full-time engagement, with no conflicts/no other engagements. — 2. Key Responsibilities • Guide research and engineering teams to close knowledge gaps and improve AI model performance in systems-level optimization, compiler engineering, and runtime performance topics. • Design challenging, domain-relevant tasks across multiple specializations, and write accurate and well-structured solutions to performance engineering problems. • Evaluate performance engineering tasks and solutions and provide clear, written technical feedback. • Develop guidelines and detailed rubrics/evaluation frameworks to assess systems design quality across AI workloads. • Collaborate with other subject matter experts to ensure consistency and accuracy in training data. — 3. Core Qualifications • 2+ years of dedicated professional experience in performance engineering, systems programming, or low-level optimization. • Deep hands-on expertise in at least one of the following: C++, Python, or Rust — with working familiarity across the others being a strong plus. • Demonstrable track record of measurable performance improvements on production systems (e.g., latency reduction, throughput gains, memory footprint optimization). • Demonstrable career progression. • Ability to engage reliably for at least 40 hours/week during weekdays. • Strong written communication skills and the ability to explain complex technical decisions clearly. — About Cincinnatus LLC: Cincinnatus LLC is an enterprise staffing company that partners with leading technology companies to source and employ highly skilled professionals for contingent and contract-based opportunities. Cincinnatus serves as the employer of record for these engagements, providing W-2 employment, payroll, benefits, and compliance, while placing employees directly within client teams to work on high-impact initiatives. — Equal Employment Opportunity: Cincinnatus is proud to be an Equal Employment Opportunity employer. We do not discriminate based upon race, religion, color, national origin, sex (including pregnancy, childbirth, reproductive health decisions, or related medical conditions), sexual orientation, gender identity, gender expression, age, status as a protected veteran, status as an individual with a disability, genetic information, political views or activity, or any other legally protected characteristic.

$70 - $110 / hourOpen / Referral verified
Code / Remote

Risk-adjustment / HCC coding leader

Mercor is working with a leading AI research lab to improve the capabilities of next-generation AI systems. We are seeking experienced Risk Adjustment and HCC Coding leaders to evaluate AI tools designed to improve risk score accuracy and coding completeness in Medicare Advantage, Medicaid managed care, and ACA markets. Your expertise in hierarchical condition category (HCC) methodology, RADV audit preparation, and risk adjustment coding will directly shape AI systems that enhance documentation capture and risk score integrity. — Responsibilities • Lead risk adjustment and HCC coding operations across Medicare Advantage, Medicaid, and/or ACA risk adjustment programs. • Evaluate AI-generated HCC coding assignments and risk adjustment recommendations for clinical accuracy and regulatory compliance. • Review medical records to ensure complete and accurate capture of HCC-eligible conditions supported by clinical documentation. • Conduct and oversee retrospective and prospective chart reviews for risk score optimisation. • Manage RADV (Risk Adjustment Data Validation) audit preparation and response processes. • Monitor risk adjustment KPIs including HCC capture rates, risk score accuracy, and chart retrieval rates. • Collaborate with clinical, coding, and compliance teams to improve documentation and coding for risk adjustment purposes. • Ensure compliance with CMS risk adjustment guidelines (RAPS, EDGE submissions) and Official Coding Guidelines. • Annotate AI outputs and provide structured coding feedback to support AI training datasets. — Requirements • 5+ years of experience in risk adjustment coding, HCC coding, or Medicare Advantage coding operations, with at least 2 years in a leadership role. • Deep expertise in CMS-HCC, RxHCC, and/or ACA HHS-HCC risk adjustment methodologies. • Strong knowledge of ICD-10-CM coding guidelines as applied to HCC risk adjustment. • Experience with RADV audit preparation and CMS compliance requirements. • Familiarity with RAPS and EDGE submission processes. • Exceptional written and verbal English communication skills. • High attention to detail with the ability to identify coding inaccuracies and documentation gaps in AI-generated outputs. — Preferred Qualifications • CRC (Certified Risk Coder), CCS, CPC, or RHIA credential. • Experience with risk adjustment analytics platforms and chart retrieval systems. • Background in health plan, Medicare Advantage organisation, or value-based care setting. • Familiarity with AI-assisted HCC coding tools and comfort evaluating AI-generated risk adjustment content. • Experience presenting risk adjustment performance to actuarial or executive teams. — Why Join? • Contribute to the development of frontier AI systems in healthcare. • Collaborate with a world-class AI research organisation. • Gain exposure to cutting-edge AI workflows in risk adjustment and value-based care. • Opportunity to work on high-impact projects shaping the future of healthcare AI.

$110 / hourOpen / Referral verified
Business / Remote

Security Expert

About the Role — Mercor is building realistic, high-fidelity simulated environments to evaluate and train AI models on real-world procurement workflows for a leading spend-management technology company. We're looking for security and vendor-risk professionals to author and validate security-review tasks inside these simulated environments. — Key Responsibilities • Review simulated vendor SOC 2 reports, security questionnaires, and pen-test evidence against a buyer's security standard • Catch scope mismatches and lapsed bridge letters that a surface-level review would miss • Author step-level rubrics and golden responses capturing how an experienced security reviewer would judge a request or renewal • Assess data-handling and sub-processor risk for vendors touching sensitive data — Ideal Qualifications • 8+ years of professional experience in security review, vendor risk management, or third-party risk (TPRM) • Hands-on experience evaluating SOC 2 reports, security questionnaires, and compliance evidence • Strong written communication skills; comfortable producing structured, rubric-style feedback — Nice to Have • Relevant security certification, such as CISSP or CISA • Prior task-writing, rubric-authoring, or AI-training data experience

$90 - $110 / hourOpen / Referral verified
Code / Remote

Document/deck production QA Evaluator

About the role — We are hiring expert Evaluators in Document/deck production QA to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. — This is a remote, hourly engagement. — Requirements (must have) — 1. 5+ years of relevant professional experience in Document/deck production QA. 2. Native or professional fluency in English. 3. Highly proficient in Microsoft Office and Google Workspace, especially Slides (Google Slides / PowerPoint). — Preferred (nice to have) • Advanced degree (Master's or higher) from a reputable institution. — What you'll do • Evaluate AI-generated artifacts against domain-specific quality rubrics. • Identify factual, aesthetic, and presentation errors. • Provide clear, structured written feedback.

$80 - $120 / hourOpen / Referral verified
Finance / Remote

Advisory & Transaction Services Expert (M&A / Valuation)

Role Overview — Mercor is collaborating with a leading AI lab to engage experienced transaction advisory and valuation professionals. You'll translate real M&A, due diligence, and valuation work into structured, high-quality training data that teaches AI to reason the way deal teams do. — Focus Areas — M&A / transaction advisory · financial due diligence (quality of earnings) · business valuation · purchase accounting (ASC 805) support. — Key Responsibilities • Design realistic scenarios from your work — quality-of-earnings (QoE) analyses, working-capital and net-debt analyses, financial due-diligence findings, valuation models (DCF, comparable companies, precedent transactions), purchase price allocation (ASC 805), and deal / LBO models • Review and compare AI-generated advisory outputs for analytical accuracy, defensible assumptions, and sound professional judgment • Provide clear written feedback that improves how AI performs transaction and valuation tasks • Collaborate asynchronously with the research team — Ideal Qualifications • Transaction advisory (Big 4 TAS / boutique) and/or corporate development, investment banking, or valuation background • CPA, and/or ASA / ABV / CFA a plus for valuation-leaning candidates • Hands-on ownership of QoE, valuation, or deal models • Bachelor's degree in Accounting, Finance, or a related field • Strong written communication and attention to detail — Application Process • Submit a resume or a short summary of your transaction advisory / valuation experience • Complete a short form on your practice area, specialties, and certifications • Selected applicants may complete a brief sample task

$80 - $120 / hourOpen / Referral verified
Finance / Remote

Accounting Domain Intake — Shape Upcoming Accounting Engagements

About This Survey — Mercor is opening a series of accounting-focused engagements with leading AI labs over the coming weeks and months — spanning audit reasoning, financial reporting, tax, GAAP/IFRS analysis, and adjacent workstreams. Before we finalize the scope of each, we want to build them around how accountants *actually* work, not how the field is described in textbooks or public sources. — This is a written intake designed to map out the real workflows, tools, decision points, and specialty cuts inside the profession — from the inside. — What You'll Do • Complete a ~30-minute written intake about your accounting practice • Describe your day-to-day workflows, how your time is split, and the workflows core to your role • Explain how experienced accountants actually divide the field from the inside — the cuts insiders use, not the textbook ones • Share the "invisible" parts of your work that outsiders (including project designers) tend to miss • No right or wrong answers — detail and specificity matter far more than polish — Ideal Respondents • Practicing or recently practicing accountants — public accounting, corporate finance, audit, tax, or advisory • CPA (or equivalent) preferred; deep domain experience without the credential also welcome • Willing to write detailed, specific responses about the mechanics of your work — Why This Matters • Your intake shapes the scope of the accounting engagements we launch • Thoughtful responses translate directly into more offer matches and priority routing as we open new engagements over the coming weeks and months • Typical accounting rates on Mercor range from $80–$120/hr and scale with seniority — About Mercor — Mercor is a talent marketplace that connects experts with leading AI labs and research organizations. Backed by Benchmark, General Catalyst, Adam D'Angelo, Larry Summers, and Jack Dorsey.

$80 - $120 / hourOpen / Referral verified
Finance / Remote

FP&A & Treasury Analyst

Role Overview — Mercor is collaborating with a leading AI lab to engage experienced FP&A and treasury professionals. You'll translate real planning, analysis, and cash / treasury work into structured, high-quality training data that teaches AI to reason about how finance teams actually work. — Focus Areas — Financial planning & analysis (FP&A) · treasury & cash management. — Key Responsibilities • Design realistic scenarios from your work — operating budgets and rolling forecasts, flux / variance analysis and commentary, 13-week cash flow forecasts, daily cash positioning & bank administration, FX revaluation, debt / covenant compliance schedules, and formula-driven financial models • Review and compare AI-generated FP&A / treasury outputs for accuracy and sound judgment on assumptions • Provide clear written feedback that improves how AI performs planning and cash-management tasks • Collaborate asynchronously with the research team — Ideal Qualifications • Finance / MBA background with strong financial-modeling skills and hands-on cash / treasury exposure • CTP (Certified Treasury Professional) a plus for treasury-leaning candidates • Corporate FP&A or treasury experience at a mid-size or larger company • Bachelor's degree in Finance, Accounting, or a related field • Strong written communication and attention to detail — Application Process • Submit a resume or a short summary of your FP&A / treasury experience • Complete a short form on your practice area, specialties, and certifications • Selected applicants may complete a brief sample task

$80 - $120 / hourOpen / Referral verified
Legal / Remote

Applied Legal Benchmark Specialist

Role Overview — We are seeking legal experts to author and review high-quality academic assessment content for an AI research initiative. You will write and verify rigorous multiple-choice questions across core law domains, evaluate solution quality, and help establish gold-standard benchmarks used to advance AI capabilities. — You will be assigned one of two task types: • Question Authoring — Create original, challenging multiple-choice questions in your area of legal expertise, rate their difficulty, and submit them for review. • Question Verification — Review pre-written questions for accuracy, clarity, and rigor. Edit where needed, rate difficulty, and document any changes made. — Law Domains Covered — Intellectual Property, Privacy, and Technology, Regulatory and Government Affairs, Securities, Capital Markets, Financial Regulation & Compliance, Private Equity, M&A & Transaction Structuring, Antitrust, Merger Control & Competition, Healthcare, Life Sciences, & Pharmaceuticals, Environmental, Energy, & ESG/Climate. — Key Responsibilities • Author original law questions that test deep conceptual understanding, not surface-level recall • Ensure questions are unambiguous, self-contained, and precisely defined — all necessary information must be in the problem statement • Rate each question's difficulty: Medium (intro undergraduate), Hard (advanced undergraduate), or Expert (post-graduate and above) • Provide 1 correct answer and 9 plausible but subtly incorrect alternatives that challenge expert-level solvers • Write step-by-step Chain-of-Thought solutions with clear, concise reasoning in markdown format • Supply 1–5 academic references per question from reputable sources (peer-reviewed journals, legal repositories) • For verification tasks: flag issues with clarity, completeness, precision, or solvability and justify any edits made — Ideal Qualifications • JD, LLM, SJD, or doctoral candidate in Law or a closely related field • Master's degree considered for candidates with exceptional depth in a specific subdomain • Strong command of legal reasoning, statutory interpretation, and jurisprudential theory • Experience with bar exam writing, legal academia, or judicial clerkships is a strong plus • Excellent written English and ability to express complex legal ideas clearly and concisely — More About the Opportunity • Expected commitment: 10+ hours/week • Asynchronous, fully remote work

$83 - $105 / hourOpen / Referral verified
STEM / Remote UK

Molecular Biology Experts

Join a leading AI lab's cutting-edge GenAI team to be at the core of the AI revolution, where your expertise fuels the development of the most advanced AI models. — 1. Overview — Join a leading AI lab's cutting-edge GenAI team and help build foundational AI models from the ground up. We're seeking talented Molecular Biology subject-matter experts (SMEs) with hands-on experience designing DNA and RNA sequences — primers, plasmids, guide RNAs, mRNA constructs, and repair templates — to bring deep domain expertise and elevate the quality of our AI training data. This is a part-time to full-time commitment of up to 40 hours per week. — This is a W-2 employment position with Cincinnatus LLC (or appropriate international entity), with the opportunity to be placed at a leading AI Lab as part of their extended workforce. — 2. Key Responsibilities • Guide research and engineering teams to close knowledge gaps and improve AI model performance in molecular biology, nucleic-acid sequence design, and construct engineering. • Design challenging, domain-relevant tasks and write accurate, well-documented solutions — spanning primers (PCR/qPCR/cloning), plasmids/expression vectors, gRNA/sgRNA for CRISPR editing, mRNA/RNA constructs, and HDR/repair templates — that serve as ground truth. • Evaluate molecular-biology tasks and AI model outputs against expert-quality solutions and provide clear, written technical feedback on correctness, rigor, and biological reasoning. • Develop guidelines and detailed rubrics/evaluation frameworks to assess sequence-design quality — guide/target selection, homology-arm length, primer melting temperature and specificity, codon optimization, and regulatory elements. • Collaborate with other subject matter experts to ensure consistency and accuracy in training data. — 3. Core Qualifications • Advanced degree (PhD strongly preferred) in molecular biology, genetics, biochemistry, synthetic biology, bioengineering, or a related life-science field. • Hands-on experience designing nucleic-acid constructs prior to experiments — across several of primers, plasmids/vectors, gRNA/sgRNA, mRNA, and HDR/repair templates — together with cloning strategy (Gibson, Golden Gate, Gateway, restriction) and codon optimization. • A strong peer-reviewed publication record, weighted toward first-author work in notable venues (e.g., Nature and the Nature family, Cell, eLife, Nucleic Acids Research, PNAS, EMBO Journal). Please include publication links with your application. • Demonstrable career progression. • Ability to engage reliably for at least 20 hours/week during weekdays. • Past experience in AI training, model evaluation, and data annotation is preferred. • Strong written communication skills and the ability to justify design choices clearly and precisely. — About Cincinnatus LLC: Cincinnatus LLC is an enterprise staffing company that partners with leading technology companies to source and employ highly skilled professionals for contingent and contract-based opportunities. Cincinnatus serves as the employer of record for these engagements, providing W-2 employment, payroll, benefits, and compliance, while placing employees directly within client teams to work on high-impact initiatives. — Equal Employment Opportunity: Cincinnatus is proud to be an Equal Employment Opportunity employer. We do not discriminate based upon race, religion, color, national origin, sex (including pregnancy, childbirth, reproductive health decisions, or related medical conditions), sexual orientation, gender identity, gender expression, age, status as a protected veteran, status as an individual with a disability, genetic information, political views or activity, or any other legally protected characteristic.

$70 - $105 / hourOpen / Referral verified
Medical / Remote

Prior Authorisation Manager

Mercor is working with a leading AI research lab to improve the capabilities of next-generation AI systems. We are seeking experienced Prior Authorisation Managers with clinical and medical expertise to support the evaluation of AI tools designed to streamline prior authorisation workflows. Your clinical knowledge of medical-necessity criteria, payer review processes, and utilisation management will help train AI systems to reduce administrative burden and improve patient access to care. — Responsibilities • Manage end-to-end prior authorisation workflows for medical and clinical services across multiple payer types. • Review clinical documentation to assess medical necessity against InterQual, MCG, or payer-specific criteria. • Evaluate AI-generated prior authorisation recommendations and clinical justification drafts for accuracy and appropriateness. • Coordinate with clinical staff, physicians, and payers to obtain timely authorisation approvals. • Track authorisation status, denials, and appeal outcomes to identify workflow improvement opportunities. • Develop and maintain SOPs for prior authorisation submission, follow-up, and escalation processes. • Monitor KPIs including authorisation approval rates, turnaround times, and denial rates. • Ensure compliance with payer requirements, CMS guidelines, and clinical review criteria. • Annotate AI outputs and provide structured clinical feedback to support AI training datasets. — Requirements • 5+ years of experience in prior authorisation, utilisation management, or clinical review, with at least 2 years in a management role. • Strong clinical background with knowledge of medical necessity criteria (InterQual, MCG, or equivalent). • Deep familiarity with commercial, Medicare Advantage, and Medicaid prior authorisation requirements. • Experience managing authorisation workflows across multiple specialities and service types. • Proficiency with authorisation management systems and EHR platforms (Epic, Cerner, or equivalent). • Exceptional written and verbal English communication skills. • High attention to detail with the ability to critically evaluate clinical documentation and AI-generated outputs. — Preferred Qualifications • Clinical licensure (RN, LPN, or equivalent) or relevant certification (CPUR, CPUM). • Experience with appeals and peer-to-peer review processes. • Familiarity with CMS prior authorisation rules and the No Surprises Act. • Background in physician office, hospital, or health plan prior authorisation operations. • Familiarity with AI tools and comfort evaluating AI-generated clinical content. — Why Join? • Contribute to the development of frontier AI systems in healthcare. • Collaborate with a world-class AI research organisation. • Gain exposure to cutting-edge AI workflows in healthcare revenue cycle management. • Opportunity to work on high-impact projects shaping the future of healthcare AI.

$105 / hourOpen / Referral verified
Medical / Remote

Insurance Verification & Benefit Manager

Mercor is working with a leading AI research lab to improve the capabilities of next-generation AI systems. We are seeking experienced Insurance Verification and Eligibility & Benefits Managers to evaluate AI tools designed to automate front-end revenue cycle operations. Your deep expertise in payer systems, EDI transactions, and eligibility workflows will directly inform AI models that drive accuracy and efficiency in insurance verification processes. — Responsibilities • Oversee insurance verification, eligibility determination, and benefits investigation workflows across commercial, Medicare, Medicaid, and managed care payers. • Verify patient insurance coverage using payer portals, clearinghouses, and EDI 270/271 real-time eligibility transactions. • Identify and resolve coordination of benefits issues, coverage gaps, and eligibility discrepancies prior to service delivery. • Evaluate and annotate AI-generated eligibility verification outputs for accuracy, completeness, and payer compliance. • Develop and document SOPs for eligibility and benefits verification workflows. • Monitor KPIs including verification accuracy rates, front-end denial rates related to eligibility, and turnaround times. • Ensure compliance with payer-specific requirements, CMS guidelines, and HIPAA regulations. • Identify process gaps and recommend workflow improvements to reduce eligibility-related claim denials. • Provide structured feedback and annotations to support AI training datasets. — Requirements • 5+ years of experience in insurance verification, eligibility and benefits management, or front-end revenue cycle operations, with at least 2 years in a management role. • Expert knowledge of EDI 270/271 transactions, payer portal navigation, and real-time eligibility tools. • Strong familiarity with Medicare, Medicaid, and commercial payer eligibility requirements and benefit structures. • Proficiency with Epic, Cerner, Meditech, or equivalent EHR platforms. • Experience with clearinghouse platforms such as Availity, Change Healthcare, or similar. • Exceptional written and verbal English communication skills. • High attention to detail with the ability to identify subtle discrepancies in coverage data. — Preferred Qualifications • CHAM, CHAA, or equivalent front-end revenue cycle certification. • Experience with automated eligibility verification tools and RPA solutions. • Background in denial root cause analysis related to eligibility and coverage errors. • Familiarity with AI tools and comfort evaluating AI-generated healthcare content. • Experience presenting eligibility performance data to senior leadership. — Why Join? • Contribute to the development of frontier AI systems in healthcare. • Collaborate with a world-class AI research organisation. • Gain exposure to cutting-edge AI workflows in healthcare revenue cycle management. • Opportunity to work on high-impact projects shaping the future of healthcare AI.

$105 / hourOpen / Referral verified
Legal / Remote

IP / trademark / copyright law Evaluator

About the role — We are hiring expert Evaluators in IP / trademark / copyright law to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. — This is a remote, hourly engagement. — Requirements (must have) — 1. 5+ years of relevant professional experience in IP / trademark / copyright law. 2. Native or professional fluency in English. 3. Highly proficient in Microsoft Office and Google Workspace, especially Slides (Google Slides / PowerPoint). — Preferred (nice to have) • Advanced degree (Master's or higher) from a reputable institution. — What you'll do • Evaluate AI-generated artifacts against domain-specific quality rubrics. • Identify factual, aesthetic, and presentation errors. • Provide clear, structured written feedback.

$80 - $120 / hourOpen / Referral verified
Legal / Remote

Legal contracts / diligence / redlines Evaluator

About the role — We are hiring expert Evaluators in Legal contracts / diligence / redlines to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. — This is a remote, hourly engagement. — Requirements (must have) — 1. 5+ years of relevant professional experience in Legal contracts / diligence / redlines. 2. Native or professional fluency in English. 3. Highly proficient in Microsoft Office and Google Workspace, especially Slides (Google Slides / PowerPoint). — Preferred (nice to have) • Advanced degree (Master's or higher) from a reputable institution. — What you'll do • Evaluate AI-generated artifacts against domain-specific quality rubrics. • Identify factual, aesthetic, and presentation errors. • Provide clear, structured written feedback.

$80 - $120 / hourOpen / Referral verified
Finance / Remote

Personal finance / consumer planning Evaluator

About the role — We are hiring expert Evaluators in Personal finance / consumer planning to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. — This is a remote, hourly engagement. — Requirements (must have) — 1. 5+ years of relevant professional experience in Personal finance / consumer planning. 2. Native or professional fluency in English. 3. Highly proficient in Microsoft Office and Google Workspace, especially Slides (Google Slides / PowerPoint). — Preferred (nice to have) • Advanced degree (Master's or higher) from a reputable institution. — What you'll do • Evaluate AI-generated artifacts against domain-specific quality rubrics. • Identify factual, aesthetic, and presentation errors. • Provide clear, structured written feedback.

$80 - $120 / hourOpen / Referral verified
Finance / Remote

Compliance / regulatory response with financial-services AI Evaluator

About the role — We are hiring expert Evaluators in Compliance / regulatory response with financial-services AI to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. — This is a remote, hourly engagement. — Requirements (must have) — 1. 5+ years of relevant professional experience in Compliance / regulatory response with financial-services AI. 2. Native or professional fluency in English. 3. Highly proficient in Microsoft Office and Google Workspace, especially Slides (Google Slides / PowerPoint). — Preferred (nice to have) • Advanced degree (Master's or higher) from a reputable institution. — What you'll do • Evaluate AI-generated artifacts against domain-specific quality rubrics. • Identify factual, aesthetic, and presentation errors. • Provide clear, structured written feedback.

$80 - $120 / hourOpen / Referral verified
Business / Remote

Brand / creative direction / marketing collateral Evaluator

About the role — We are hiring expert Evaluators in Brand / creative direction / marketing collateral to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. — This is a remote, hourly engagement. — Requirements (must have) — 1. 5+ years of relevant professional experience in Brand / creative direction / marketing collateral. 2. Native or professional fluency in English. 3. Highly proficient in Microsoft Office and Google Workspace, especially Slides (Google Slides / PowerPoint). — Preferred (nice to have) • Advanced degree (Master's or higher) from a reputable institution. — What you'll do • Evaluate AI-generated artifacts against domain-specific quality rubrics. • Identify factual, aesthetic, and presentation errors. • Provide clear, structured written feedback.

$80 - $120 / hourOpen / Referral verified
Policy & Safety / Remote

Privacy / regulatory compliance Evaluator

About the role — We are hiring expert Evaluators in Privacy / regulatory compliance to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. — This is a remote, hourly engagement. — Requirements (must have) — 1. 5+ years of relevant professional experience in Privacy / regulatory compliance. 2. Native or professional fluency in English. 3. Highly proficient in Microsoft Office and Google Workspace, especially Slides (Google Slides / PowerPoint). — Preferred (nice to have) • Advanced degree (Master's or higher) from a reputable institution. — What you'll do • Evaluate AI-generated artifacts against domain-specific quality rubrics. • Identify factual, aesthetic, and presentation errors. • Provide clear, structured written feedback.

$80 - $120 / hourOpen / Referral verified
Language / Remote

Process improvement / SOPs Evaluator

About the role — We are hiring expert Evaluators in Process improvement / SOPs to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. — This is a remote, hourly engagement. — Requirements (must have) — 1. 5+ years of relevant professional experience in Process improvement / SOPs. 2. Native or professional fluency in English. 3. Highly proficient in Microsoft Office and Google Workspace, especially Slides (Google Slides / PowerPoint). — Preferred (nice to have) • Advanced degree (Master's or higher) from a reputable institution. — What you'll do • Evaluate AI-generated artifacts against domain-specific quality rubrics. • Identify factual, aesthetic, and presentation errors. • Provide clear, structured written feedback.

$80 - $120 / hourOpen / Referral verified
Language / Remote

Training / onboarding / L&D Evaluator

About the role — We are hiring expert Evaluators in Training / onboarding / L&D to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. — This is a remote, hourly engagement. — Requirements (must have) — 1. 5+ years of relevant professional experience in Training / onboarding / L&D. 2. Native or professional fluency in English. 3. Highly proficient in Microsoft Office and Google Workspace, especially Slides (Google Slides / PowerPoint). — Preferred (nice to have) • Advanced degree (Master's or higher) from a reputable institution. — What you'll do • Evaluate AI-generated artifacts against domain-specific quality rubrics. • Identify factual, aesthetic, and presentation errors. • Provide clear, structured written feedback.

$80 - $120 / hourOpen / Referral verified
Language / Remote

Cybersecurity / IT GRC Evaluator

About the role — We are hiring expert Evaluators in Cybersecurity / IT GRC to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. — This is a remote, hourly engagement. — Requirements (must have) — 1. 5+ years of relevant professional experience in Cybersecurity / IT GRC. 2. Native or professional fluency in English. 3. Highly proficient in Microsoft Office and Google Workspace, especially Slides (Google Slides / PowerPoint). — Preferred (nice to have) • Advanced degree (Master's or higher) from a reputable institution. — What you'll do • Evaluate AI-generated artifacts against domain-specific quality rubrics. • Identify factual, aesthetic, and presentation errors. • Provide clear, structured written feedback.

$80 - $120 / hourOpen / Referral verified
Finance / Remote

Finance operations / audit support Evaluator

About the role — We are hiring expert Evaluators in Finance operations / audit support to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. — This is a remote, hourly engagement. — Requirements (must have) — 1. 5+ years of relevant professional experience in Finance operations / audit support. 2. Native or professional fluency in English. 3. Highly proficient in Microsoft Office and Google Workspace, especially Slides (Google Slides / PowerPoint). — Preferred (nice to have) • Advanced degree (Master's or higher) from a reputable institution. — What you'll do • Evaluate AI-generated artifacts against domain-specific quality rubrics. • Identify factual, aesthetic, and presentation errors. • Provide clear, structured written feedback.

$80 - $120 / hourOpen / Referral verified
Language / Remote

Media / journalism / communications Evaluator

About the role — We are hiring expert Evaluators in Media / journalism / communications to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. — This is a remote, hourly engagement. — Requirements (must have) — 1. 5+ years of relevant professional experience in Media / journalism / communications. 2. Native or professional fluency in English. 3. Highly proficient in Microsoft Office and Google Workspace, especially Slides (Google Slides / PowerPoint). — Preferred (nice to have) • Advanced degree (Master's or higher) from a reputable institution. — What you'll do • Evaluate AI-generated artifacts against domain-specific quality rubrics. • Identify factual, aesthetic, and presentation errors. • Provide clear, structured written feedback.

$80 - $120 / hourOpen / Referral verified
Business / Remote

General business strategy / management Evaluator

About the role — We are hiring expert Evaluators in General business strategy / management to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. — This is a remote, hourly engagement. — Requirements (must have) — 1. 5+ years of relevant professional experience in General business strategy / management. 2. Native or professional fluency in English. 3. Highly proficient in Microsoft Office and Google Workspace, especially Slides (Google Slides / PowerPoint). — Preferred (nice to have) • Advanced degree (Master's or higher) from a reputable institution. — What you'll do • Evaluate AI-generated artifacts against domain-specific quality rubrics. • Identify factual, aesthetic, and presentation errors. • Provide clear, structured written feedback.

$80 - $120 / hourOpen / Referral verified
Finance / Remote

FP&A / corporate finance Evaluator

About the role — We are hiring expert Evaluators in FP&A / corporate finance to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. — This is a remote, hourly engagement. — Requirements (must have) — 1. 5+ years of relevant professional experience in FP&A / corporate finance. 2. Native or professional fluency in English. 3. Highly proficient in Microsoft Office and Google Workspace, especially Slides (Google Slides / PowerPoint). — Preferred (nice to have) • Advanced degree (Master's or higher) from a reputable institution. — What you'll do • Evaluate AI-generated artifacts against domain-specific quality rubrics. • Identify factual, aesthetic, and presentation errors. • Provide clear, structured written feedback.

$80 - $120 / hourOpen / Referral verified
Medical / Remote

Clinical / biomedical / pharma Evaluator

About the role — We are hiring expert Evaluators in Clinical / biomedical / pharma to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. — This is a remote, hourly engagement. — Requirements (must have) — 1. 5+ years of relevant professional experience in Clinical / biomedical / pharma. 2. Native or professional fluency in English. 3. Highly proficient in Microsoft Office and Google Workspace, especially Slides (Google Slides / PowerPoint). — Preferred (nice to have) • Advanced degree (Master's or higher) from a reputable institution. — What you'll do • Evaluate AI-generated artifacts against domain-specific quality rubrics. • Identify factual, aesthetic, and presentation errors. • Provide clear, structured written feedback.

$80 - $120 / hourOpen / Referral verified
Finance / Remote

Investment analysis / valuation / credit Evaluator

About the role — We are hiring expert Evaluators in Investment analysis / valuation / credit to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. — This is a remote, hourly engagement. — Requirements (must have) — 1. 5+ years of relevant professional experience in Investment analysis / valuation / credit. 2. Native or professional fluency in English. 3. Highly proficient in Microsoft Office and Google Workspace, especially Slides (Google Slides / PowerPoint). — Preferred (nice to have) • Advanced degree (Master's or higher) from a reputable institution. — What you'll do • Evaluate AI-generated artifacts against domain-specific quality rubrics. • Identify factual, aesthetic, and presentation errors. • Provide clear, structured written feedback.

$80 - $120 / hourOpen / Referral verified
STEM / Remote

Computational Chemistry & Electronic Structure Expert

Computational Chemistry & Electronic Structure Expert — About the Project — We're building a large-scale benchmark to test how well advanced AI systems can solve hard scientific and engineering problems. As a task designer, you'll create challenging computational problems that check whether AI can use real scientific software to do research-level work — running simulations, interpreting results, designing experiments, and uncovering hidden information from data. — This isn't a typical data-labeling job. You'll design original, graduate-level problems based on real scientific workflows, test them against cutting-edge AI models, and fine-tune them until the difficulty is just right. — What You'll Do — You'll create problems that require skilled use of specialized scientific software. Some will ask the AI to compute exact answers from a fully defined setup — testing whether it can correctly carry out complex, multi-step workflows. Others will be harder: the AI must plan a series of queries or experiments to uncover information that isn't directly visible, which means thinking strategically about what to measure, how to read partial results, and how to narrow down the possibilities efficiently. — Each problem goes through a testing loop against state-of-the-art AI models, and you'll refine it until it hits the target difficulty. — Domains & Tools We're Hiring For — We're especially interested in experts with deep, hands-on experience in: — Computational Chemistry & Electronic Structure — working with PySCF for quantum chemistry calculations including Hartree-Fock, DFT, TDDFT, CASSCF, and post-HF methods. Ideal candidates can design problems around excited-state analysis, orbital diagnostics, choosing the right method for tricky electronic structures, and interpreting computational artifacts that come from method limitations. — _Experience with other specialized software in this domain will also be considered._ — What Makes a Strong Candidate — You have graduate-level expertise (MS or PhD preferred) in the domain above, with real hands-on experience using these tools — not just theoretical knowledge. You've written code using these libraries to solve actual research problems, and you understand where they break, what their edge cases are, and what makes a problem genuinely hard rather than just complicated. — Beyond domain expertise, the best candidates think like puzzle designers: building problems where the challenge comes from smart reasoning rather than raw computation, where several approaches seem plausible but only careful analysis reveals the right one, and where surface-level pattern matching won't get you to the answer. — Requirements • Graduate-level training in a relevant STEM field (MS, PhD, or equivalent research experience) • Proven proficiency with at least one of the listed scientific software libraries, shown through research publications, open-source contributions, or professional work • Strong Python skills — you'll be writing problem setups, oracle functions, and solution validators • Ability to work independently and refine problem designs based on feedback • Comfortable working in a Linux/terminal environment with remote compute sandboxes • Available for at least 15–20 hours per week — Nice to Have • Experience across multiple listed domains or tools • Familiarity with benchmark or evaluation design • Background in scientific teaching or exam/problem-set design • Experience with computational reproducibility and containerized environments

$70 - $100 / hourOpen / Referral verified
Multimodal / Remote

Voice Actor: CX Agent Voice Cloning (Southern USA)

Mercor is partnering with a leading AI research lab to build next-generation text-to-speech (TTS) systems capable of producing natural, expressive, and human-like voices. We are seeking USA-based voice actors with native Southern American English accents (e.g. Texas, Georgia, Tennessee, the Carolinas, Louisiana, Alabama, Mississippi) to contribute high-quality voice recordings for training and evaluating cutting-edge speech models. This role is ideal for professionals with experience in voice acting, narration, or broadcast who can deliver consistent, expressive, and clean audio across a variety of scripts. — Your voice recordings will be used exclusively for the client's internal CX AI Agent. They will not be sold, licensed, or reused for any other product, dataset, or purpose. • * * — Key Responsibilities • Record high-quality voice samples across diverse scripts (conversational, narrative, instructional, etc.) • Deliver clear, natural, and expressive speech with strong control over tone, pacing, and pronunciation • Maintain consistency in voice, accent, and delivery across recording sessions • Follow detailed recording guidelines (environment, microphone setup, file formatting) • Perform multiple takes with variation in emotion, emphasis, and style when required — Requirements • Native Southern American English speaker currently based in the United States, with an authentic Southern US accent • Proven experience in voice acting, dubbing, narration, podcasting, or broadcasting • Access to a professional or near-professional recording setup (quality microphone, quiet environment, pop filter, etc.) • Strong command of intonation, diction, and emotional range • Ability to follow scripts precisely while maintaining natural delivery • Reliable availability for 5–10 hours per week throughout the project duration. Please note: This is a short-term project expected to last 1–2 weeks. • \[IMP\]: Your voice may be cloned for the client's CX AI Agent so please only apply if you are okay with voice cloning. Your recordings will not be sold or reused outside of this specific use case. — Preferred Qualifications • Experience recording for TTS, audiobooks, IVR systems, or AI voice datasets • Familiarity with audio editing tools (e.g., Audacity, Adobe Audition, Reaper) • Ability to deliver multiple vocal styles (e.g., conversational, corporate, energetic, calm) • * *

$50 - $100 / hourOpen / Referral verified
Multimodal / Remote

Voice Actor: CX Agent Voice Cloning (Standard French)

Mercor is partnering with a leading AI research lab to build next-generation text-to-speech (TTS) systems capable of producing natural, expressive, and human-like voices. We are seeking native or native-level French speakers with a standard, neutral, accent-free French delivery to contribute high-quality voice recordings for training and evaluating cutting-edge speech models. This role is ideal for professionals with experience in voice acting, narration, or broadcast who can deliver consistent, expressive, and clean audio across a variety of scripts. — Your voice recordings will be used exclusively for the client's internal CX AI Agent. They will not be sold, licensed, or reused for any other product, dataset, or purpose. • * * — Key Responsibilities • Record high-quality voice samples across diverse scripts (conversational, narrative, instructional, etc.) • Deliver clear, natural, and expressive speech with strong control over tone, pacing, and pronunciation • Maintain consistency in voice, accent, and delivery across recording sessions • Follow detailed recording guidelines (environment, microphone setup, file formatting) • Perform multiple takes with variation in emotion, emphasis, and style when required — Requirements • Native or native-level French speaker with a standard, neutral French accent (accent-free) — no strong regional accent and no foreign-language influence • Location does not matter — we welcome qualified speakers from anywhere, as long as the delivery is neutral, standard French • Proven experience in voice acting, dubbing, narration, podcasting, or broadcasting • Access to a professional or near-professional recording setup (quality microphone, quiet environment, pop filter, etc.) • Strong command of intonation, diction, and emotional range • Ability to follow scripts precisely while maintaining natural delivery • Reliable availability for 5-10 hours per week over the project duration. • \[IMP\]: Your voice may be cloned for the client's CX AI Agent so please only apply if you are okay with voice cloning. Your recordings will not be sold or reused outside of this specific use case. — Preferred Qualifications • Experience recording for TTS, audiobooks, IVR systems, or AI voice datasets • Familiarity with audio editing tools (e.g., Audacity, Adobe Audition, Reaper) • Ability to deliver multiple vocal styles (e.g., conversational, corporate, energetic, calm) • * *

$50 - $100 / hourOpen / Referral verified
STEM / Remote

Computational Particle & Nuclear Physics Expert

Particle & Nuclear Physics Expert — About the Project — We're building a large-scale benchmark to test how well advanced AI systems can solve hard scientific and engineering problems. As a task designer, you'll create challenging computational problems that check whether AI can use real scientific software to do research-level work — running simulations, interpreting results, designing experiments, and uncovering hidden information from data. — This isn't a typical data-labeling job. You'll design original, graduate-level problems based on real scientific workflows, test them against cutting-edge AI models, and fine-tune them until the difficulty is just right. — What You'll Do — You'll create problems that require skilled use of specialized scientific software. Some will ask the AI to compute exact answers from a fully defined setup — testing whether it can correctly carry out complex, multi-step workflows. Others will be harder: the AI must plan a series of queries or experiments to uncover information that isn't directly visible, which means thinking strategically about what to measure, how to read partial results, and how to narrow down the possibilities efficiently. — Each problem goes through a testing loop against state-of-the-art AI models, and you'll refine it until it hits the target difficulty. — Domains & Tools We're Hiring For — We're especially interested in experts with deep, hands-on experience in: — Particle & Nuclear Physics — working with scikit-hep and related HEP Python tools for particle physics data analysis, cross-section computations, renormalization group calculations, and perturbative QCD. Experience with Monte Carlo event generation or collider phenomenology is a plus. — _Experience with other specialized software in this domain will also be considered._ — What Makes a Strong Candidate — You have graduate-level expertise (MS or PhD preferred) in the domain above, with real hands-on experience using these tools — not just theoretical knowledge. You've written code using these libraries to solve actual research problems, and you understand where they break, what their edge cases are, and what makes a problem genuinely hard rather than just complicated. — Beyond domain expertise, the best candidates think like puzzle designers: building problems where the challenge comes from smart reasoning rather than raw computation, where several approaches seem plausible but only careful analysis reveals the right one, and where surface-level pattern matching won't get you to the answer. — Requirements • Graduate-level training in a relevant STEM field (MS, PhD, or equivalent research experience) • Proven proficiency with at least one of the listed scientific software libraries, shown through research publications, open-source contributions, or professional work • Strong Python skills — you'll be writing problem setups, oracle functions, and solution validators • Ability to work independently and refine problem designs based on feedback • Comfortable working in a Linux/terminal environment with remote compute sandboxes • Available for at least 15–20 hours per week — Nice to Have • Experience across multiple listed domains or tools • Familiarity with benchmark or evaluation design • Background in scientific teaching or exam/problem-set design • Experience with computational reproducibility and containerized environments

$70 - $100 / hourOpen / Referral verified
STEM / Remote

Computational Astrophysics & Cosmology Expert

Computational Astrophysics & Cosmology Expert — About the Project — We're building a large-scale benchmark to test how well advanced AI systems can solve hard scientific and engineering problems. As a task designer, you'll create challenging computational problems that check whether AI can use real scientific software to do research-level work — running simulations, interpreting results, designing experiments, and uncovering hidden information from data. — This isn't a typical data-labeling job. You'll design original, graduate-level problems based on real scientific workflows, test them against cutting-edge AI models, and fine-tune them until the difficulty is just right. — What You'll Do — You'll create problems that require skilled use of specialized scientific software. Some will ask the AI to compute exact answers from a fully defined setup — testing whether it can correctly carry out complex, multi-step workflows. Others will be harder: the AI must plan a series of queries or experiments to uncover information that isn't directly visible, which means thinking strategically about what to measure, how to read partial results, and how to narrow down the possibilities efficiently. — Each problem goes through a testing loop against state-of-the-art AI models, and you'll refine it until it hits the target difficulty. — Domains & Tools We're Hiring For — We're especially interested in experts with deep, hands-on experience in: — Astrophysics & Cosmology — working with astropy and related tools for cosmological calculations, angular power spectra, galaxy survey analysis, and observational data reduction pipelines. — _Experience with other specialized software in this domain will also be considered._ — What Makes a Strong Candidate — You have graduate-level expertise (MS or PhD preferred) in the domain above, with real hands-on experience using these tools — not just theoretical knowledge. You've written code using these libraries to solve actual research problems, and you understand where they break, what their edge cases are, and what makes a problem genuinely hard rather than just complicated. — Beyond domain expertise, the best candidates think like puzzle designers: building problems where the challenge comes from smart reasoning rather than raw computation, where several approaches seem plausible but only careful analysis reveals the right one, and where surface-level pattern matching won't get you to the answer. — Requirements • Graduate-level training in a relevant STEM field (MS, PhD, or equivalent research experience) • Proven proficiency with at least one of the listed scientific software libraries, shown through research publications, open-source contributions, or professional work • Strong Python skills — you'll be writing problem setups, oracle functions, and solution validators • Ability to work independently and refine problem designs based on feedback • Comfortable working in a Linux/terminal environment with remote compute sandboxes • Available for at least 15–20 hours per week — Nice to Have • Experience across multiple listed domains or tools • Familiarity with benchmark or evaluation design • Background in scientific teaching or exam/problem-set design • Experience with computational reproducibility and containerized environments

$70 - $100 / hourOpen / Referral verified
Multimodal / Remote

Voice Actor: CX Agent Voice Cloning - Northern UK

Mercor is partnering with a leading AI research lab to build next-generation text-to-speech (TTS) systems capable of producing natural, expressive, and human-like voices. We are seeking Northern UK-based female voice actors with authentic Northern English accents (e.g. Yorkshire, Manchester, Newcastle, Liverpool) to contribute high-quality voice recordings for training and evaluating cutting-edge speech models. This role is ideal for professionals with experience in voice acting, narration, or broadcast who can deliver consistent, expressive, and clean audio across a variety of scripts. — Your voice recordings will be used exclusively for the client's internal CX AI Agent. They will not be sold, licensed, or reused for any other product, dataset, or purpose. • * * — Key Responsibilities • Record high-quality voice samples across diverse scripts (conversational, narrative, instructional, etc.) • Deliver clear, natural, and expressive speech with strong control over tone, pacing, and pronunciation • Maintain consistency in voice, accent, and delivery across recording sessions • Follow detailed recording guidelines (environment, microphone setup, file formatting) • Perform multiple takes with variation in emotion, emphasis, and style when required — Requirements • Female voice actor with a natural, authentic Northern English accent, currently based in the northern UK • Proven experience in voice acting, dubbing, narration, podcasting, or broadcasting • Access to a professional or near-professional recording setup (quality microphone, quiet environment, pop filter, etc.) • Strong command of intonation, diction, and emotional range • Ability to follow scripts precisely while maintaining natural delivery • Reliable availability for 5-10 hours per week over the project duration • \[IMP\]: Your voice may be cloned for the client's CX AI Agent so please only apply if you are okay with voice cloning. Your recordings will not be sold or reused outside of this specific use case. — Preferred Qualifications • Experience recording for TTS, audiobooks, IVR systems, or AI voice datasets • Familiarity with audio editing tools (e.g., Audacity, Adobe Audition, Reaper) • Ability to deliver multiple vocal styles (e.g., conversational, corporate, energetic, calm) • * *

$50 - $100 / hourOpen / Referral verified
Code / Remote

Bioinformatics & Computational Single-Cell Genomics Expert

Bioinformatics & Computational Single-Cell Genomics Expert — About the Project — We're building a large-scale benchmark to test how well advanced AI systems can solve hard scientific and engineering problems. As a task designer, you'll create challenging computational problems that check whether AI can use real scientific software to do research-level work — running simulations, interpreting results, designing experiments, and uncovering hidden information from data. — This isn't a typical data-labeling job. You'll design original, graduate-level problems based on real scientific workflows, test them against cutting-edge AI models, and fine-tune them until the difficulty is just right. — What You'll Do — You'll create problems that require skilled use of specialized scientific software. Some will ask the AI to compute exact answers from a fully defined setup — testing whether it can correctly carry out complex, multi-step workflows. Others will be harder: the AI must plan a series of queries or experiments to uncover information that isn't directly visible, which means thinking strategically about what to measure, how to read partial results, and how to narrow down the possibilities efficiently. — Each problem goes through a testing loop against state-of-the-art AI models, and you'll refine it until it hits the target difficulty. — Domains & Tools We're Hiring For — We're especially interested in experts with deep, hands-on experience in: — Bioinformatics & Single-Cell Genomics — working with tools like scanpy, scvelo, squidpy, and gudhi for single-cell RNA-seq analysis, trajectory inference, spatial transcriptomics, and topological data analysis. You should be comfortable designing problems around cell-type annotation, pseudotime ordering, multi-omic integration, spatial variable gene identification, and persistence-based analysis pipelines. This is our highest-throughput domain and where we're scaling first. — _Experience with other specialized software in this domain will also be considered._ — What Makes a Strong Candidate — You have graduate-level expertise (MS or PhD preferred) in the domain above, with real hands-on experience using these tools — not just theoretical knowledge. You've written code using these libraries to solve actual research problems, and you understand where they break, what their edge cases are, and what makes a problem genuinely hard rather than just complicated. — Beyond domain expertise, the best candidates think like puzzle designers: building problems where the challenge comes from smart reasoning rather than raw computation, where several approaches seem plausible but only careful analysis reveals the right one, and where surface-level pattern matching won't get you to the answer. — Requirements • Graduate-level training in a relevant STEM field (MS, PhD, or equivalent research experience) • Proven proficiency with at least one of the listed scientific software libraries, shown through research publications, open-source contributions, or professional work • Strong Python skills — you'll be writing problem setups, oracle functions, and solution validators • Ability to work independently and refine problem designs based on feedback • Comfortable working in a Linux/terminal environment with remote compute sandboxes • Available for at least 15–20 hours per week — Nice to Have • Experience across multiple listed domains or tools • Familiarity with benchmark or evaluation design • Background in scientific teaching or exam/problem-set design • Experience with computational reproducibility and containerized environments

$70 - $100 / hourOpen / Referral verified
Finance / Bay Area, CA

Finance Domain Expert — AI Training & Evaluation

Join a leading AI lab's cutting-edge GenAI team to be at the core of the AI revolution, where your expertise fuels the development of the most advanced AI models. — 1. Overview — We are hiring a senior finance domain expert to work directly with a leading AI lab's research and program management teams, improving how frontier AI models reason about real financial work. — Your finance expertise is the substance of this role. You will review the quality of finance knowledge work tasks, write the instruction specs and golden solutions that define what "correct" looks like, and build the benchmarks that show whether the model is genuinely improving. We are looking for a practising specialist rather than a generalist. — This is a full-time W-2 employment position with Cincinnatus LLC, with the opportunity to be placed at a leading AI lab as part of their extended workforce. You will be provisioned with client-issued accounts and equipment, and will work inside the client's own tools alongside their research teams. — Location: This is a hybrid role based in the Bay Area, California. You must live in the Bay Area and work on-site with the client's team multiple days each week, when required. This is not a remote role. If you do not currently live in the Bay Area, you must be willing to relocate there at your own cost before the engagement starts — relocation assistance is not provided. — 2. Key Responsibilities • Data QA and reviews: Vet the quality of finance knowledge work tasks and model outputs — spotting missing behaviors, thin reasoning, flawed assumptions, and answers that read well but would not survive professional scrutiny. • Instruction specs and golden datasets: Write high-quality instruction specs, produce golden solutions to financial problems, and define new finance tasks that reflect how the work is actually done in practice. • Benchmarks and domain depth: Design challenging finance tasks and evaluation sets, and help build finance-specific skills and tools together with the research team. • Calibration: Work with client researchers and specialists in adjacent fields to keep standards consistent, translating tacit financial judgment into explicit, teachable criteria. — 3. Core Qualifications • Experience: 5+ years of substantive, dedicated professional finance experience at a recognized institution — for example an investment bank, asset manager, private equity or credit fund, Big Four firm, a large corporate finance function, or a financial regulator. Generalist roles that only touch finance peripherally do not count. • Domain depth: Genuine specialization in at least one core finance discipline, for example corporate finance and FP&A, investment banking and M&A, asset or wealth management, private equity or private credit, quantitative finance and risk management, treasury, or accounting and audit. • Seniority: Clear progression to a senior individual-contributor or leadership level — for example Vice President, Director, Principal, Managing Director, Portfolio Manager, Controller, or CFO — with real ownership of analysis and decisions. • Education and credentials: An advanced degree from a strong program (MBA, MS, or PhD in finance, economics, accounting, or a quantitative field) and/or a recognized professional credential such as CFA, CPA, FRM, or an actuarial designation. Strongly preferred. • AI fluency: Hands-on working use of large language models in your professional work, and the judgment to tell a well-reasoned answer from a plausible-sounding wrong one. • Availability: Able to commit reliably to 40 hours per week for an initial engagement of 6 months. • Location: Living in the Bay Area, California, and able to work on-site with the client's team multiple days each week, when required. Candidates not currently based in the Bay Area must be willing to relocate there at their own cost; relocation assistance is not provided. • Excellent written communication, and the ability to give precise, well-structured written feedback. — About Cincinnatus LLC — Cincinnatus LLC is an enterprise staffing company that partners with leading technology companies to source and employ highly skilled professionals for contingent and contract-based opportunities. Cincinnatus serves as the employer of record for these engagements, providing W-2 employment, payroll, benefits, and compliance, while placing employees directly within client teams to work on high-impact initiatives. — Roles hired through Cincinnatus are not project-based or freelance engagements. They are structured, role-based positions that typically involve part-time or full-time commitments, close collaboration with a client's internal teams, and integration into standard enterprise workflows. — Cincinnatus is a legal entity separate from Mercor. While opportunities may be discovered through Mercor's platform, employment, onboarding, payroll, and benefits for these roles are administered by Cincinnatus LLC. — Equal Employment Opportunity — Cincinnatus is proud to be an Equal Employment Opportunity employer. We do not discriminate based upon race, religion, color, national origin, sex (including pregnancy, childbirth, reproductive health decisions, or related medical conditions), sexual orientation, gender identity, gender expression, age, status as a protected veteran, status as an individual with a disability, genetic information, political views or activity, or any other legally protected characteristic. — Cincinnatus is committed to providing reasonable accommodations for qualified individuals with disabilities and disabled veterans throughout the job application process.

$60 - $100 / hourOpen / Referral verified
Legal / Bay Area, CA

Legal Domain Expert — AI Training & Evaluation

Join a leading AI lab's cutting-edge GenAI team to be at the core of the AI revolution, where your expertise fuels the development of the most advanced AI models. — 1. Overview — We are hiring a senior legal domain expert to work directly with a leading AI lab's research and program management teams, improving how frontier AI models reason about real legal work. — Your legal expertise is the substance of this role. You will review the quality of legal knowledge work tasks, write the instruction specs and golden solutions that define what "correct" looks like, and build the benchmarks that show whether the model is genuinely improving. We are looking for a practising specialist rather than a generalist. — This is a full-time W-2 employment position with Cincinnatus LLC, with the opportunity to be placed at a leading AI lab as part of their extended workforce. You will be provisioned with client-issued accounts and equipment, and will work inside the client's own tools alongside their research teams. — Location: This is a hybrid role based in the Bay Area, California. You must live in the Bay Area and work on-site with the client's team multiple days each week, when required. This is not a remote role. If you do not currently live in the Bay Area, you must be willing to relocate there at your own cost before the engagement starts — relocation assistance is not provided. — 2. Key Responsibilities • Data QA and reviews: Vet the quality of legal knowledge work tasks and model outputs — spotting missing behaviors, thin reasoning, and answers that read well but would not survive professional scrutiny. • Instruction specs and golden datasets: Write high-quality instruction specs, produce golden solutions to legal problems, and define new legal tasks that reflect how the work is actually done in practice. • Benchmarks and domain depth: Design challenging legal tasks and evaluation sets, and help build legal-specific skills and tools together with the research team. • Calibration: Work with client researchers and specialists in adjacent fields to keep standards consistent, translating tacit legal judgment into explicit, teachable criteria. — 3. Core Qualifications • Education: Juris Doctor (JD) from an accredited law school; a highly ranked school is strongly preferred. • Experience: 5+ years of substantive post-qualification legal practice at a reputable institution — an established law firm, a corporate legal department, a regulatory body, or a court. Internships and clerkships alone do not count. • Domain depth: Genuine specialization in at least one substantive practice area, for example corporate and transactional, litigation and dispute resolution, regulatory and compliance, intellectual property, employment and labor, or tax. • Seniority: Clear progression to a senior level — partner, of counsel, counsel, senior associate, senior in-house counsel, or general counsel — with real ownership of matters. • Licensure: Active admission to at least one U.S. state bar, in good standing. • AI fluency: Hands-on working use of large language models in your professional work, and the judgment to tell a well-reasoned answer from a plausible-sounding wrong one. • Availability: Able to commit reliably to 40 hours per week for an initial engagement of 6 months. • Location: Living in the Bay Area, California, and able to work on-site with the client's team multiple days each week, when required. Candidates not currently based in the Bay Area must be willing to relocate there at their own cost; relocation assistance is not provided. • Excellent written communication, and the ability to give precise, well-structured written feedback. — About Cincinnatus LLC — Cincinnatus LLC is an enterprise staffing company that partners with leading technology companies to source and employ highly skilled professionals for contingent and contract-based opportunities. Cincinnatus serves as the employer of record for these engagements, providing W-2 employment, payroll, benefits, and compliance, while placing employees directly within client teams to work on high-impact initiatives. — Roles hired through Cincinnatus are not project-based or freelance engagements. They are structured, role-based positions that typically involve part-time or full-time commitments, close collaboration with a client's internal teams, and integration into standard enterprise workflows. — Cincinnatus is a legal entity separate from Mercor. While opportunities may be discovered through Mercor's platform, employment, onboarding, payroll, and benefits for these roles are administered by Cincinnatus LLC. — Equal Employment Opportunity — Cincinnatus is proud to be an Equal Employment Opportunity employer. We do not discriminate based upon race, religion, color, national origin, sex (including pregnancy, childbirth, reproductive health decisions, or related medical conditions), sexual orientation, gender identity, gender expression, age, status as a protected veteran, status as an individual with a disability, genetic information, political views or activity, or any other legally protected characteristic. — Cincinnatus is committed to providing reasonable accommodations for qualified individuals with disabilities and disabled veterans throughout the job application process.

$60 - $100 / hourOpen / Referral verified
Language / Remote

Voice Actor: CX Agent Voice Cloning (UK English)

Mercor is partnering with a leading AI research lab to build next-generation text-to-speech (TTS) systems capable of producing natural, expressive, and human-like voices. We are seeking UK-based native English speakers with a neutral, standard UK English accent to contribute high-quality voice recordings for training and evaluating cutting-edge speech models. This role is ideal for professionals with experience in voice acting, narration, or broadcast who can deliver consistent, expressive, and clean audio across a variety of scripts. — Your voice recordings will be used exclusively for the client's internal CX AI Agent. They will not be sold, licensed, or reused for any other product, dataset, or purpose. • * * — Key Responsibilities • Record high-quality voice samples across diverse scripts (conversational, narrative, instructional, etc.) • Deliver clear, natural, and expressive speech with strong control over tone, pacing, and pronunciation • Maintain consistency in voice, accent, and delivery across recording sessions • Follow detailed recording guidelines (environment, microphone setup, file formatting) • Perform multiple takes with variation in emotion, emphasis, and style when required — Requirements • Native English speaker with a neutral, standard UK English accent, currently based in the United Kingdom • Proven experience in voice acting, dubbing, narration, podcasting, or broadcasting • Access to a professional or near-professional recording setup (quality microphone, quiet environment, pop filter, etc.) • Strong command of intonation, diction, and emotional range • Ability to follow scripts precisely while maintaining natural delivery • Reliable availability for 5-10 hours per week over the project duration • \[IMP\]: Your voice may be cloned for the client's CX AI Agent so please only apply if you are okay with voice cloning. Your recordings will not be sold or reused outside of this specific use case. — Preferred Qualifications • Experience recording for TTS, audiobooks, IVR systems, or AI voice datasets • Familiarity with audio editing tools (e.g., Audacity, Adobe Audition, Reaper) • Ability to deliver multiple vocal styles (e.g., conversational, corporate, energetic, calm) • * *

$50 - $100 / hourOpen / Referral verified
Language / Remote

Voice Actor: CX Agent Voice Cloning (Australian English)

Mercor is partnering with a leading AI research lab to build next-generation text-to-speech (TTS) systems capable of producing natural, expressive, and human-like voices. We are seeking Australia-based native English speakers with authentic Australian accents to contribute high-quality voice recordings for training and evaluating cutting-edge speech models. This role is ideal for professionals with experience in voice acting, narration, or broadcast who can deliver consistent, expressive, and clean audio across a variety of scripts. — Your voice recordings will be used exclusively for the client's internal CX AI Agent. They will not be sold, licensed, or reused for any other product, dataset, or purpose. • * * — Key Responsibilities • Record high-quality voice samples across diverse scripts (conversational, narrative, instructional, etc.) • Deliver clear, natural, and expressive speech with strong control over tone, pacing, and pronunciation • Maintain consistency in voice, accent, and delivery across recording sessions • Follow detailed recording guidelines (environment, microphone setup, file formatting) • Perform multiple takes with variation in emotion, emphasis, and style when required — Requirements • Native English speaker with an authentic Australian accent, currently based in Australia • Proven experience in voice acting, dubbing, narration, podcasting, or broadcasting • Access to a professional or near-professional recording setup (quality microphone, quiet environment, pop filter, etc.) • Strong command of intonation, diction, and emotional range • Ability to follow scripts precisely while maintaining natural delivery • Reliable availability for 5-10 hours per week over the project duration • \[IMP\]: Your voice may be cloned for the client's CX AI Agent so please only apply if you are okay with voice cloning. Your recordings will not be sold or reused outside of this specific use case. — Preferred Qualifications • Experience recording for TTS, audiobooks, IVR systems, or AI voice datasets • Familiarity with audio editing tools (e.g., Audacity, Adobe Audition, Reaper) • Ability to deliver multiple vocal styles (e.g., conversational, corporate, energetic, calm) • * *

$50 - $100 / hourOpen / Referral verified
Multimodal / Remote

Voice Actor: CX Agent Voice Cloning (German)

Mercor is partnering with a leading AI research lab to build next-generation text-to-speech (TTS) systems capable of producing natural, expressive, and human-like voices. We are seeking Germany-based native German speakers to contribute high-quality voice recordings for training and evaluating cutting-edge speech models. This role is ideal for professionals with experience in voice acting, narration, or broadcast who can deliver consistent, expressive, and clean audio across a variety of scripts. — Your voice recordings will be used exclusively for the client's internal CX AI Agent. They will not be sold, licensed, or reused for any other product, dataset, or purpose. • * * — Key Responsibilities • Record high-quality voice samples across diverse scripts (conversational, narrative, instructional, etc.) • Deliver clear, natural, and expressive speech with strong control over tone, pacing, and pronunciation • Maintain consistency in voice, accent, and delivery across recording sessions • Follow detailed recording guidelines (environment, microphone setup, file formatting) • Perform multiple takes with variation in emotion, emphasis, and style when required — Requirements • Native German speaker currently based in Germany • Proven experience in voice acting, dubbing, narration, podcasting, or broadcasting • Access to a professional or near-professional recording setup (quality microphone, quiet environment, pop filter, etc.) • Strong command of intonation, diction, and emotional range • Ability to follow scripts precisely while maintaining natural delivery • Reliable availability for 5-10 hours per week over the project duration • \[IMP\]: Your voice may be cloned for the client's CX AI Agent so please only apply if you are okay with voice cloning. Your recordings will not be sold or reused outside of this specific use case. — Preferred Qualifications • Experience recording for TTS, audiobooks, IVR systems, or AI voice datasets • Familiarity with audio editing tools (e.g., Audacity, Adobe Audition, Reaper) • Ability to deliver multiple vocal styles (e.g., conversational, corporate, energetic, calm) • * *

$50 - $100 / hourOpen / Referral verified
Finance / Remote

Revenue-cycle analytics / decision-support / RCM reporting leader

Mercor is working with a leading AI research lab to improve the capabilities of next-generation AI systems. We are seeking experienced Revenue Cycle Analytics and Decision-Support leaders — including RCM reporting managers and analytics directors — to evaluate AI tools designed to transform revenue cycle intelligence and financial decision-making. Your expertise in RCM data analytics, financial reporting, and decision-support will directly shape AI systems that deliver actionable insights across the full revenue cycle. — Responsibilities • Lead revenue cycle analytics, reporting, and decision-support functions to drive data-driven performance improvement across RCM operations. • Evaluate AI-generated revenue cycle analytics outputs, KPI dashboards, and financial modeling recommendations for accuracy and completeness. • Develop and maintain revenue cycle reporting frameworks covering patient access, coding, billing, denials, A/R, and collections performance. • Build and interpret dashboards, scorecards, and trend analyses to support operational and executive decision-making. • Conduct root cause analyses of revenue cycle performance variances and develop data-driven improvement recommendations. • Collaborate with finance, IT, and operational revenue cycle teams to align analytics infrastructure with strategic priorities. • Manage RCM data governance, including data definitions, data quality standards, and reporting consistency. • Support revenue cycle forecasting, budget modelling, and net revenue realisation analysis. • Annotate AI outputs and provide structured analytical feedback to support AI training datasets. — Requirements • 5+ years of experience in revenue cycle analytics, RCM reporting, or healthcare financial decision-support, with at least 2 years in a leadership role. • Deep knowledge of revenue cycle KPIs and financial metrics across patient access, coding, billing, denials, and collections domains. • Proficiency with healthcare analytics platforms, BI tools (Tableau, Power BI, or equivalent), and SQL-based data analysis. • Experience with EHR-integrated analytics platforms and RCM reporting systems. • Strong financial modelling and data interpretation skills with the ability to translate complex data into actionable recommendations. • Exceptional written and verbal English communication skills. • High attention to detail with the ability to identify data quality issues and analytical errors in AI-generated outputs. — Preferred Qualifications • HFMA CRCR, CHFP, or healthcare analytics certification. • Experience with predictive analytics, revenue cycle forecasting, and net revenue modelling. • Background in health system, large physician enterprise, or RCM outsourcing analytics operations. • Familiarity with AI tools and comfort evaluating AI-generated analytics content. • Experience implementing enterprise-wide revenue cycle analytics infrastructure or BI platform migrations. — Why Join? • Contribute to the development of frontier AI systems in healthcare. • Collaborate with a world-class AI research organisation. • Gain exposure to cutting-edge AI workflows in revenue cycle analytics and decision-support. • Opportunity to work on high-impact projects shaping the future of healthcare AI.

$100 / hourOpen / Referral verified
Language / Remote

Market research / competitive intelligence Evaluator

About the role — We are hiring expert Evaluators in Market research / competitive intelligence to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. — This is a remote, hourly engagement. — Requirements (must have) — 1. 5+ years of relevant professional experience in Market research / competitive intelligence. 2. Native or professional fluency in English. 3. Highly proficient in Microsoft Office and Google Workspace, especially Slides (Google Slides / PowerPoint). — Preferred (nice to have) • Advanced degree (Master's or higher) from a reputable institution. — What you'll do • Evaluate AI-generated artifacts against domain-specific quality rubrics. • Identify factual, aesthetic, and presentation errors. • Provide clear, structured written feedback.

$80 - $120 / hourOpen / Referral verified
Business / Remote

General Sales / GTM Evaluator

About the role — We are hiring expert Evaluators in General Sales / GTM to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. — This is a remote, hourly engagement. — Requirements (must have) — 1. 5+ years of relevant professional experience in General Sales / GTM. 2. Native or professional fluency in English. 3. Highly proficient in Microsoft Office and Google Workspace, especially Slides (Google Slides / PowerPoint). — Preferred (nice to have) • Advanced degree (Master's or higher) from a reputable institution. — What you'll do • Evaluate AI-generated artifacts against domain-specific quality rubrics. • Identify factual, aesthetic, and presentation errors. • Provide clear, structured written feedback.

$80 - $120 / hourOpen / Referral verified
Language / Remote

Humanities / arts / culture Evaluator

About the role — We are hiring expert Evaluators in Humanities / arts / culture to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. — This is a remote, hourly engagement. — Requirements (must have) — 1. 5+ years of relevant professional experience in Humanities / arts / culture. 2. Native or professional fluency in English. 3. Highly proficient in Microsoft Office and Google Workspace, especially Slides (Google Slides / PowerPoint). — Preferred (nice to have) • Advanced degree (Master's or higher) from a reputable institution. — What you'll do • Evaluate AI-generated artifacts against domain-specific quality rubrics. • Identify factual, aesthetic, and presentation errors. • Provide clear, structured written feedback.

$80 - $120 / hourOpen / Referral verified
Language / Remote

Real estate / hospitality / events Evaluator

About the role — We are hiring expert Evaluators in Real estate / hospitality / events to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. — This is a remote, hourly engagement. — Requirements (must have) — 1. 5+ years of relevant professional experience in Real estate / hospitality / events. 2. Native or professional fluency in English. 3. Highly proficient in Microsoft Office and Google Workspace, especially Slides (Google Slides / PowerPoint). — Preferred (nice to have) • Advanced degree (Master's or higher) from a reputable institution. — What you'll do • Evaluate AI-generated artifacts against domain-specific quality rubrics. • Identify factual, aesthetic, and presentation errors. • Provide clear, structured written feedback.

$80 - $120 / hourOpen / Referral verified
Language / Remote

Special education / IEP Evaluator

About the role — We are hiring expert Evaluators in Special education / IEP to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. — This is a remote, hourly engagement. — Requirements (must have) — 1. 5+ years of relevant professional experience in Special education / IEP. 2. Native or professional fluency in English. 3. Highly proficient in Microsoft Office and Google Workspace, especially Slides (Google Slides / PowerPoint). — Preferred (nice to have) • Advanced degree (Master's or higher) from a reputable institution. — What you'll do • Evaluate AI-generated artifacts against domain-specific quality rubrics. • Identify factual, aesthetic, and presentation errors. • Provide clear, structured written feedback.

$80 - $120 / hourOpen / Referral verified
Language / Remote

Public health communications Evaluator

About the role — We are hiring expert Evaluators in Public health communications to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. — This is a remote, hourly engagement. — Requirements (must have) — 1. 5+ years of relevant professional experience in Public health communications. 2. Native or professional fluency in English. 3. Highly proficient in Microsoft Office and Google Workspace, especially Slides (Google Slides / PowerPoint). — Preferred (nice to have) • Advanced degree (Master's or higher) from a reputable institution. — What you'll do • Evaluate AI-generated artifacts against domain-specific quality rubrics. • Identify factual, aesthetic, and presentation errors. • Provide clear, structured written feedback.

$80 - $120 / hourOpen / Referral verified
Language / Remote

Government / public administration Evaluator

About the role — We are hiring expert Evaluators in Government / public administration to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. — This is a remote, hourly engagement. — Requirements (must have) — 1. 5+ years of relevant professional experience in Government / public administration. 2. Native or professional fluency in English. 3. Highly proficient in Microsoft Office and Google Workspace, especially Slides (Google Slides / PowerPoint). — Preferred (nice to have) • Advanced degree (Master's or higher) from a reputable institution. — What you'll do • Evaluate AI-generated artifacts against domain-specific quality rubrics. • Identify factual, aesthetic, and presentation errors. • Provide clear, structured written feedback.

$80 - $120 / hourOpen / Referral verified
Business / Remote

Customer success / support operations Evaluator

About the role — We are hiring expert Evaluators in Customer success / support operations to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. — This is a remote, hourly engagement. — Requirements (must have) — 1. 5+ years of relevant professional experience in Customer success / support operations. 2. Native or professional fluency in English. 3. Highly proficient in Microsoft Office and Google Workspace, especially Slides (Google Slides / PowerPoint). — Preferred (nice to have) • Advanced degree (Master's or higher) from a reputable institution. — What you'll do • Evaluate AI-generated artifacts against domain-specific quality rubrics. • Identify factual, aesthetic, and presentation errors. • Provide clear, structured written feedback.

$80 - $120 / hourOpen / Referral verified
Business / Remote

Data quality / CRM operations Evaluator

About the role — We are hiring expert Evaluators in Data quality / CRM operations to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. — This is a remote, hourly engagement. — Requirements (must have) — 1. 5+ years of relevant professional experience in Data quality / CRM operations. 2. Native or professional fluency in English. 3. Highly proficient in Microsoft Office and Google Workspace, especially Slides (Google Slides / PowerPoint). — Preferred (nice to have) • Advanced degree (Master's or higher) from a reputable institution. — What you'll do • Evaluate AI-generated artifacts against domain-specific quality rubrics. • Identify factual, aesthetic, and presentation errors. • Provide clear, structured written feedback.

$80 - $120 / hourOpen / Referral verified
Finance / Remote

General finance / accounting Evaluator

About the role — We are hiring expert Evaluators in General finance / accounting to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. — This is a remote, hourly engagement. — Requirements (must have) — 1. 5+ years of relevant professional experience in General finance / accounting. 2. Native or professional fluency in English. 3. Highly proficient in Microsoft Office and Google Workspace, especially Slides (Google Slides / PowerPoint). — Preferred (nice to have) • Advanced degree (Master's or higher) from a reputable institution. — What you'll do • Evaluate AI-generated artifacts against domain-specific quality rubrics. • Identify factual, aesthetic, and presentation errors. • Provide clear, structured written feedback.

$80 - $120 / hourOpen / Referral verified
Policy & Safety / Remote

Legal / compliance Evaluator

About the role — We are hiring expert Evaluators in Legal / compliance to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. — This is a remote, hourly engagement. — Requirements (must have) — 1. 5+ years of relevant professional experience in Legal / compliance. 2. Native or professional fluency in English. 3. Highly proficient in Microsoft Office and Google Workspace, especially Slides (Google Slides / PowerPoint). — Preferred (nice to have) • Advanced degree (Master's or higher) from a reputable institution. — What you'll do • Evaluate AI-generated artifacts against domain-specific quality rubrics. • Identify factual, aesthetic, and presentation errors. • Provide clear, structured written feedback.

$80 - $120 / hourOpen / Referral verified
Language / Remote

Product launch / experiment readiness Evaluator

About the role — We are hiring expert Evaluators in Product launch / experiment readiness to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. — This is a remote, hourly engagement. — Requirements (must have) — 1. 5+ years of relevant professional experience in Product launch / experiment readiness. 2. Native or professional fluency in English. 3. Highly proficient in Microsoft Office and Google Workspace, especially Slides (Google Slides / PowerPoint). — Preferred (nice to have) • Advanced degree (Master's or higher) from a reputable institution. — What you'll do • Evaluate AI-generated artifacts against domain-specific quality rubrics. • Identify factual, aesthetic, and presentation errors. • Provide clear, structured written feedback.

$80 - $120 / hourOpen / Referral verified
Business / Remote

Program management / implementation planning Evaluator

About the role — We are hiring expert Evaluators in Program management / implementation planning to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. — This is a remote, hourly engagement. — Requirements (must have) — 1. 5+ years of relevant professional experience in Program management / implementation planning. 2. Native or professional fluency in English. 3. Highly proficient in Microsoft Office and Google Workspace, especially Slides (Google Slides / PowerPoint). — Preferred (nice to have) • Advanced degree (Master's or higher) from a reputable institution. — What you'll do • Evaluate AI-generated artifacts against domain-specific quality rubrics. • Identify factual, aesthetic, and presentation errors. • Provide clear, structured written feedback.

$80 - $120 / hourOpen / Referral verified
Language / Remote

Investor materials / fundraising / pitchbook Evaluator

About the role — We are hiring expert Evaluators in Investor materials / fundraising / pitchbook to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. — This is a remote, hourly engagement. — Requirements (must have) — 1. 5+ years of relevant professional experience in Investor materials / fundraising / pitchbook. 2. Native or professional fluency in English. 3. Highly proficient in Microsoft Office and Google Workspace, especially Slides (Google Slides / PowerPoint). — Preferred (nice to have) • Advanced degree (Master's or higher) from a reputable institution. — What you'll do • Evaluate AI-generated artifacts against domain-specific quality rubrics. • Identify factual, aesthetic, and presentation errors. • Provide clear, structured written feedback.

$80 - $120 / hourOpen / Referral verified
Language / Remote

Education / school Evaluator

About the role — We are hiring expert Evaluators in Education / school to review and assess AI-generated work products (documents, spreadsheets, and slide decks) for accuracy, rigor, and domain quality. You will apply deep subject-matter expertise to grade outputs. — This is a remote, hourly engagement. — Requirements (must have) — 1. 5+ years of relevant professional experience in Education / school. 2. Native or professional fluency in English. 3. Highly proficient in Microsoft Office and Google Workspace, especially Slides (Google Slides / PowerPoint). — Preferred (nice to have) • Advanced degree (Master's or higher) from a reputable institution. — What you'll do • Evaluate AI-generated artifacts against domain-specific quality rubrics. • Identify factual, aesthetic, and presentation errors. • Provide clear, structured written feedback.

$80 - $120 / hourOpen / Referral verified
Finance / Remote

Investment Banking Expert

Role Overview • Mercor is seeking senior investment banking professionals to build evaluation tasks for AI systems operating in Fortune 500 enterprise M&A and capital markets contexts. • The workflows are calibrated to the deal complexity, stakeholder sensitivity, and transaction stakes of Fortune 500 M&A, capital raises, and large-cap advisory mandates. • Contributors design enterprise investment banking scenarios, draft reference outputs, and write rubrics that capture how senior F500-focused bankers think. — Key Responsibilities • Construct enterprise investment banking scenarios spanning $1B+ M&A transactions, multi-stakeholder deal committees, and complex capital markets or regulatory review cycles. • Build tasks across M&A advisory, equity and debt capital markets, leveraged finance, valuation and financial modeling, and pitch/deal execution. • Develop deal execution scenarios involving tools such as Bloomberg, Capital IQ, FactSet, and enterprise financial modeling and deal-management platforms used in bulge-bracket workflows. • Apply enterprise investment banking methodologies (DCF/comparable company/precedent transaction analysis, LBO modeling, fairness opinion standards) and produce reference pitch books, valuation models, and executive/board-level deal narratives. • Author rubrics that distinguish authentic investment banking judgment from generic textbook or interview-prep-level recall. — Ideal Qualifications • 5+ years working in M&A advisory, capital markets, or leveraged finance at a bulge-bracket bank, elite boutique, or major financial institution (Goldman Sachs, Morgan Stanley, JPMorgan, Evercore, Lazard). • Direct ownership of F500-scale deal execution, capital raises, or advisory mandates. • Fluency in investment banking tooling, valuation methodologies, and deal-process mechanics, plus understanding of how F500 board approvals, regulatory review (SEC, antitrust), and legal/due diligence processes actually work. • Prior rubric, financial-modeling curriculum, or pitch book/deal documentation authorship is a plus.

$90 - $100 / hourOpen / Referral verified
Legal / Remote

Regulatory Law Expert

Role Overview • Mercor is seeking senior regulatory law professionals to build evaluation tasks for AI systems operating in regulatory compliance and government affairs contexts. • The workflows are calibrated to the regulatory complexity, enforcement stakes, and scope of major regulated-industry compliance programs. • This role builds worlds on two tracks: a US track (Administrative Procedure Act and agency-specific statutes such as SEC, FDA, and FTC rules) and an International track (EU regulatory frameworks and UK regulatory bodies). Experts qualified in either or both tracks are encouraged to apply. • Contributors design regulatory law scenarios, draft reference outputs, and write rubrics that capture how senior regulatory counsel think. — Key Responsibilities • Construct scenarios spanning regulatory compliance program design, agency investigations and enforcement actions, and rulemaking or comment processes. • Build tasks across financial services regulation, healthcare and FDA regulation, antitrust and competition law, data privacy and cybersecurity regulation, and environmental and energy regulation. • Develop scenarios involving tools such as regulatory tracking platforms, compliance management systems, and agency filing systems used across regulated industries. • Apply regulatory law methodologies (regulatory interpretation, compliance risk assessment, enforcement response strategy) to the standards track a world targets (US: Administrative Procedure Act, agency-specific statutes and regulations; International: EU regulations and directives, UK regulatory frameworks), and produce reference compliance memoranda, regulatory filings, and enforcement response documents. • Author rubrics that distinguish authentic regulatory judgment from generic law school or bar exam-level recall. — Ideal Qualifications • 5+ years working as a regulatory attorney or compliance counsel at a major law firm, regulated company, or government agency (Covington & Burling, Sidley Austin, WilmerHale, or in-house regulatory/compliance counsel at a bank, pharmaceutical company, or technology company). • Direct ownership of regulatory compliance programs, agency interactions, or enforcement matters. • Fluency in regulatory tooling, plus understanding of how administrative process and agency rulemaking actually work. • A recognized professional credential is strongly preferred (JD with bar admission, or an international equivalent); prior rubric, training, or compliance documentation authorship is a plus.

$90 - $100 / hourOpen / Referral verified
Business / Remote

Intellectual Property Expert

Role Overview • Mercor is seeking senior intellectual property professionals to build evaluation tasks for AI systems operating in patent prosecution, licensing, and IP enforcement contexts. • The workflows are calibrated to the technical complexity, commercial stakes, and procedural scope of major patent portfolios, licensing programs, and IP litigation. • This role builds worlds on two tracks: a US track (35 U.S.C., MPEP, USPTO procedure) and an International track (European Patent Convention, PCT, WIPO treaties). Experts qualified in either or both tracks are encouraged to apply. • Contributors design IP scenarios, draft reference outputs, and write rubrics that capture how senior IP counsel think. — Key Responsibilities • Construct IP scenarios spanning patent prosecution and portfolio strategy, trademark and copyright disputes, and complex licensing or technology transfer processes. • Build tasks across patent prosecution, IP litigation and enforcement, trademark and copyright practice, licensing and technology transfer, and freedom-to-operate analysis. • Develop scenarios involving tools such as patent search platforms (PatSnap, Innography, USPTO PatFT), IP docketing and portfolio management systems (Anaqua, CPA Global), and claim-charting tools used across major IP practices. • Apply IP methodologies (patentability analysis, claim construction, freedom-to-operate review) to the standards track a world targets (US: 35 U.S.C., MPEP, Federal Circuit precedent; International: EPC, PCT, WIPO treaties), and produce reference patent applications, office action responses, licensing agreements, and litigation memoranda. • Author rubrics that distinguish authentic IP judgment from generic law school or bar exam-level recall. — Ideal Qualifications • 5+ years working as an IP attorney or patent agent at a major IP firm or corporate IP department (Fish & Richardson, Finnegan, Wilson Sonsini, or in-house IP counsel at a technology or pharmaceutical company). • Direct ownership of patent prosecution portfolios, licensing negotiations, or IP litigation matters. • Fluency in IP tooling and methodologies, plus understanding of how USPTO/EPO procedure and international treaty systems actually work. • A recognized professional credential is strongly preferred (JD with bar admission and USPTO patent bar registration, or an international equivalent such as European Patent Attorney); a technical/STEM background plus prior rubric or training authorship is a plus.

$90 - $100 / hourOpen / Referral verified
Legal / Remote

Public Interest Law (Civil Law / Environmental Law)

Role Overview • Mercor is seeking senior public interest law professionals to build evaluation tasks for AI systems operating in civil rights, environmental, and public advocacy contexts. • The workflows are calibrated to the societal stakes, regulatory scope, and community impact of major public interest litigation and advocacy campaigns. • This role builds worlds on two tracks: a US track (federal civil rights statutes, NEPA, Clean Air Act and Clean Water Act) and an International track (international human rights law, EU environmental directives). Experts qualified in either or both tracks are encouraged to apply. • Contributors design public interest law scenarios, draft reference outputs, and write rubrics that capture how senior public interest attorneys think. — Key Responsibilities • Construct scenarios spanning civil rights litigation, environmental compliance and enforcement, and community advocacy or policy reform processes. • Build tasks across civil rights and constitutional law, environmental law and regulation, housing and consumer protection, immigration and asylum, and government accountability. • Develop scenarios involving tools such as legal research platforms, case management systems used by legal aid organizations and nonprofits, and regulatory filing systems (e.g., EPA dockets). • Apply public interest law methodologies (constitutional analysis, regulatory compliance review, impact litigation strategy) to the standards track a world targets (US: federal civil rights statutes, NEPA, environmental statutes; International: human rights treaties, EU directives), and produce reference legal briefs, regulatory comments, and advocacy memoranda. • Author rubrics that distinguish authentic public interest legal judgment from generic law school or bar exam-level recall. — Ideal Qualifications • 5+ years working as a public interest attorney at a legal aid organization, advocacy nonprofit, or government agency (ACLU, Earthjustice, NRDC, Legal Aid Society, or a state attorney general's office). • Direct ownership of impact litigation, regulatory advocacy, or policy reform initiatives. • Fluency in public interest legal tooling, plus understanding of how regulatory processes and legislative advocacy actually work. • A recognized professional credential is strongly preferred (JD with bar admission, or an international equivalent); prior rubric, training, or policy authorship is a plus.

$90 - $100 / hourOpen / Referral verified
Legal / Remote

Corporate Law Expert

Role Overview • Mercor is seeking senior corporate law professionals to build evaluation tasks for AI systems operating in complex corporate transactions and governance contexts. • The workflows are calibrated to the deal complexity, stakeholder stakes, and regulatory scope of major M&A transactions, capital markets deals, and corporate governance matters. • This role builds worlds on two jurisdictional tracks: a US track (Delaware General Corporation Law, SEC regulations, Model Business Corporation Act) and an International track (UK Companies Act 2006, EU company law directives). Experts qualified in either or both tracks are encouraged to apply. • Contributors design corporate law scenarios, draft reference outputs, and write rubrics that capture how senior corporate lawyers think. — Key Responsibilities • Construct corporate law scenarios spanning large-scale M&A transactions, multi-party deal negotiation and regulatory review, and complex corporate governance or restructuring processes. • Build tasks across M&A and deal structuring, securities and capital markets, corporate governance, commercial contracts, and corporate restructuring. • Develop legal scenarios involving tools such as Westlaw, Lexis, virtual data rooms (Intralinks, Datasite), contract lifecycle management platforms, and document automation systems used on major transactions. • Apply corporate law methodologies (deal structuring, due diligence review, regulatory filing analysis) to the jurisdictional track a world targets (US: Delaware General Corporation Law, SEC rules, Model Business Corporation Act; International: UK Companies Act 2006, EU company law directives), and produce reference legal memoranda, transaction documents, and client/regulatory-facing narratives. • Author rubrics that distinguish authentic corporate law judgment from generic law school or bar exam-level recall. — Ideal Qualifications • 5+ years working as a corporate attorney or general counsel at a major law firm, investment bank, or corporation (Wachtell, Skadden, Cravath, Latham & Watkins, Kirkland & Ellis, Sullivan & Cromwell, or in-house legal at a large public company). • Direct ownership of M&A transactions, governance matters, or securities filings. • Fluency in corporate law tooling and methodologies, plus understanding of how regulatory approvals (SEC review, antitrust/HSR, or an international equivalent) and board/shareholder oversight actually work. • A recognized professional credential is strongly preferred (JD with bar admission, or an international equivalent such as Solicitor or Qualified Lawyer in England & Wales); prior rubric, legal training curriculum, or transaction documentation authorship is a plus.

$90 - $100 / hourOpen / Referral verified
Code / Remote

Litigation Expert

Role Overview • Mercor is seeking senior litigation professionals to build evaluation tasks for AI systems operating in civil litigation and dispute resolution contexts. • The workflows are calibrated to the case complexity, evidentiary stakes, and procedural scope of major commercial litigation and complex disputes. • This role builds worlds on two tracks: a US track (Federal Rules of Civil Procedure, Federal Rules of Evidence, state court analogs) and an International track (English Civil Procedure Rules, international arbitration rules such as ICC and LCIA). Experts qualified in either or both tracks are encouraged to apply. • Contributors design litigation scenarios, draft reference outputs, and write rubrics that capture how senior litigators think. — Key Responsibilities • Construct litigation scenarios spanning pre-trial discovery, motion practice, trial strategy, and settlement or arbitration processes. • Build tasks across commercial litigation, complex and class action litigation, discovery and evidence management, trial advocacy, and appellate practice. • Develop scenarios involving tools such as e-discovery platforms (Relativity, Everlaw), litigation management software, and deposition/trial preparation tools used on major matters. • Apply litigation methodologies (case strategy development, discovery and evidence analysis, motion drafting) to the standards track a world targets (US: FRCP, Federal Rules of Evidence, state court rules; International: CPR, international arbitration rules), and produce reference pleadings, discovery responses, motions, and trial memoranda. • Author rubrics that distinguish authentic litigation judgment from generic law school or bar exam-level recall. — Ideal Qualifications • 5+ years working as a litigator or trial attorney at a major litigation firm or corporate litigation department (Quinn Emanuel, Gibson Dunn, Boies Schiller Flexner, Kirkland & Ellis, or in-house litigation counsel). • Direct ownership of complex commercial litigation matters, with trial or arbitration experience. • Fluency in litigation tooling and methodologies, plus understanding of how court procedure, discovery, and evidentiary rules actually work. • A recognized professional credential is strongly preferred (JD with bar admission, or an international equivalent such as Barrister or Solicitor with advocacy rights); prior rubric or training authorship is a plus.

$90 - $100 / hourOpen / Referral verified
Code / Remote

ML Research PhD Experts (ICML / NeurIPS / ICLR Publications)

Overview — We're looking to rapidly assemble a small group of world-class machine learning researchers for an initial pilot. This is a high-priority engagement with an accelerated timeline, so both exceptional candidate quality and fast turnaround are critical. We're specifically seeking researchers with demonstrated contributions to frontier ML research, particularly those driving algorithmic innovation rather than applied analytics. — Candidate Requirements — Required Qualifications • PhD in Machine Learning, Computer Science, AI, or a closely related field • Published at least one main conference paper at ICML, NeurIPS, or ICLR • Strong preference for candidates with 2+ publications at these venues • Demonstrated experience conducting original ML research — Preferred Research Areas • Reinforcement Learning (RL) • Meta-Learning • Recursive Self-Improvement • AI for Science (e.g. weather forecasting, protein modeling, scientific discovery) — Timeline — Priority: Urgent — Expected Schedule • Initial data/results: Monday / Tuesday 27th July • * * — Success Criteria — Candidates should have a proven record of advancing state-of-the-art machine learning through publications at top-tier conferences and possess deep expertise in frontier ML research. Speed of sourcing is important, but quality should not be compromised.

$60 - $100 / hourOpen / Referral verified
Multimodal / Remote

Voice Actor: CX Agent Voice Cloning (Spanish - Mexico)

Mercor is partnering with a leading AI research lab to build next-generation text-to-speech (TTS) systems capable of producing natural, expressive, and human-like voices. We are seeking Mexico-based voice actors with native Mexican Spanish accents to contribute high-quality voice recordings for training and evaluating cutting-edge speech models. This role is ideal for professionals with experience in voice acting, narration, or broadcast who can deliver consistent, expressive, and clean audio across a variety of scripts. — Your voice recordings will be used exclusively for the client's internal CX AI Agent. They will not be sold, licensed, or reused for any other product, dataset, or purpose. • * * — Key Responsibilities • Record high-quality voice samples across diverse scripts (conversational, narrative, instructional, etc.) • Deliver clear, natural, and expressive speech with strong control over tone, pacing, and pronunciation • Maintain consistency in voice, accent, and delivery across recording sessions • Follow detailed recording guidelines (environment, microphone setup, file formatting) • Perform multiple takes with variation in emotion, emphasis, and style when required — Requirements • Native Mexican Spanish speaker currently based in Mexico • Proven experience in voice acting, dubbing, narration, podcasting, or broadcasting • Access to a professional or near-professional recording setup (quality microphone, quiet environment, pop filter, etc.) • Strong command of intonation, diction, and emotional range • Ability to follow scripts precisely while maintaining natural delivery • Reliable availability for 5-10 hours per week over the project duration • \[IMP\]: Your voice may be cloned for the client's CX AI Agent so please only apply if you are okay with voice cloning. Your recordings will not be sold or reused outside of this specific use case. — Preferred Qualifications • Experience recording for TTS, audiobooks, IVR systems, or AI voice datasets • Familiarity with audio editing tools (e.g., Audacity, Adobe Audition, Reaper) • Ability to deliver multiple vocal styles (e.g., conversational, corporate, energetic, calm) • * *

$50 - $100 / hourOpen / Referral verified
Medical / US Remote

Healthcare Expert

Join a growing network of industry experts supporting AI research and development. — Overview — We're building a select group of experienced healthcare professionals to help evaluate and improve how AI systems understand and reason about healthcare topics. You'll bring your real-world clinical expertise to review content for accuracy, answer domain-specific questions, and share feedback that helps make AI tools more reliable and useful in healthcare contexts. — This is a flexible, part-time engagement — approximately 10 hours per week, fully remote, with the opportunity to continue and grow over time. — What You'll Do • Review and evaluate written content and AI-generated responses in your area of clinical expertise for accuracy and quality. • Answer questions and provide clear, well-reasoned feedback based on your professional experience. • Help identify useful reference material relevant to your field. • Share insight into how healthcare professionals actually work — the tools, workflows, and standards you rely on day to day. • Collaborate with a small group of other subject-matter experts across different fields. — What We're Looking For • Licensed medical doctor (MD). • 2+ years of professional clinical experience; broad or generalist clinical exposure is a plus over a narrow subspecialty. • Comfortable engaging with medical literature and clinical guidelines (e.g. UpToDate, PubMed, published research). • Prior teaching or training experience (e.g. supervising residents, CME instruction, medical education) is a plus. • Comfortable using everyday AI tools (e.g. ChatGPT or similar). • Based in the United States. • Professional working English. • Able to commit reliably to about 10 hours per week. — About Cincinnatus LLC: Cincinnatus LLC is an enterprise staffing company that partners with leading technology companies to source and employ highly skilled professionals for contingent and contract-based opportunities. Cincinnatus serves as the employer of record for these engagements, providing W-2 employment, payroll, benefits, and compliance, while placing employees directly within client teams to work on high-impact initiatives. — Equal Employment Opportunity: Cincinnatus is proud to be an Equal Employment Opportunity employer. We do not discriminate based upon race, religion, color, national origin, sex (including pregnancy, childbirth, reproductive health decisions, or related medical conditions), sexual orientation, gender identity, gender expression, age, status as a protected veteran, status as an individual with a disability, genetic information, political views or activity, or any other legally protected characteristic.

$90 - $100 / hourOpen / Referral verified
Business / Remote

Procurement Expert

About the Role — Mercor is building realistic, high-fidelity simulated environments to evaluate and train AI models on real-world procurement workflows for a leading spend-management technology company. We're looking for senior procurement and vendor-management professionals to author and validate tasks that mirror how procurement teams actually vet new vendors and review contract renewals. — Key Responsibilities • Review simulated new-vendor intake requests, benchmarking price and contract terms against realistic company reference data • Consolidate multi-lens reviews (finance, legal, security, IT) into a well-reasoned final approval recommendation • Author step-level rubrics and golden responses that capture how an experienced procurement lead would judge a request or renewal • Audit simulated company environments for realism and unintended contradictions before tasks open — Ideal Qualifications • 8+ years of professional experience in procurement, spend operations, or vendor management, with hands-on experience running real vendor requests and renewals • Sourced from mature procurement functions at companies with formal vendor-review processes • Strong written communication skills; comfortable producing structured, rubric-style feedback — Nice to Have • Experience with modern spend-management or e-procurement platforms • Prior task-writing, rubric-authoring, or AI-training data experience

$70 - $100 / hourOpen / Referral verified
Multimodal / Remote

Voice Actor: CX Agent Voice Cloning (Spanish - Latin America)

Mercor is partnering with a leading AI research lab to build next-generation text-to-speech (TTS) systems capable of producing natural, expressive, and human-like voices. We are seeking Latin America-based voice actors with native Latin American Spanish fluency and neutral, internationally intelligible delivery to contribute high-quality voice recordings for training and evaluating cutting-edge speech models. This role is ideal for professionals with experience in voice acting, narration, or broadcast who can deliver consistent, expressive, and clean audio across a variety of scripts. — Your voice recordings will be used exclusively for the client's internal CX AI Agent. They will not be sold, licensed, or reused for any other product, dataset, or purpose. • * * — Key Responsibilities • Record high-quality voice samples across diverse scripts (conversational, narrative, instructional, etc.) • Deliver clear, natural, and expressive speech with strong control over tone, pacing, and pronunciation • Maintain consistency in voice, accent, and delivery across recording sessions • Follow detailed recording guidelines (environment, microphone setup, file formatting) • Perform multiple takes with variation in emotion, emphasis, and style when required — Requirements • Native Latin American Spanish speaker currently based in Latin America, with the ability to speak neutral, internationally intelligible Spanish — without a strong regional accent • Proven experience in voice acting, dubbing, narration, podcasting, or broadcasting • Access to a professional or near-professional recording setup (quality microphone, quiet environment, pop filter, etc.) • Strong command of intonation, diction, and emotional range • Ability to follow scripts precisely while maintaining natural delivery • Reliable availability for 5-10 hours per week over the project duration • \[IMP\]: Your voice may be cloned for the client's CX AI Agent so please only apply if you are okay with voice cloning. Your recordings will not be sold or reused outside of this specific use case. — Preferred Qualifications • Experience recording for TTS, audiobooks, IVR systems, or AI voice datasets • Familiarity with audio editing tools (e.g., Audacity, Adobe Audition, Reaper) • Ability to deliver multiple vocal styles (e.g., conversational, corporate, energetic, calm) • * *

$50 - $100 / hourOpen / Referral verified
Multimodal / Remote

Voice Actor: CX Agent Voice Cloning (Spanish - Peninsular)

Mercor is partnering with a leading AI research lab to build next-generation text-to-speech (TTS) systems capable of producing natural, expressive, and human-like voices. We are seeking Spain-based voice actors with native Castilian/Peninsular Spanish accents to contribute high-quality voice recordings for training and evaluating cutting-edge speech models. This role is ideal for professionals with experience in voice acting, narration, or broadcast who can deliver consistent, expressive, and clean audio across a variety of scripts. — Your voice recordings will be used exclusively for the client's internal CX AI Agent. They will not be sold, licensed, or reused for any other product, dataset, or purpose. • * * — Key Responsibilities • Record high-quality voice samples across diverse scripts (conversational, narrative, instructional, etc.) • Deliver clear, natural, and expressive speech with strong control over tone, pacing, and pronunciation • Maintain consistency in voice, accent, and delivery across recording sessions • Follow detailed recording guidelines (environment, microphone setup, file formatting) • Perform multiple takes with variation in emotion, emphasis, and style when required — Requirements • Native Castilian/Peninsular Spanish speaker currently based in Spain • Proven experience in voice acting, dubbing, narration, podcasting, or broadcasting • Access to a professional or near-professional recording setup (quality microphone, quiet environment, pop filter, etc.) • Strong command of intonation, diction, and emotional range • Ability to follow scripts precisely while maintaining natural delivery • Reliable availability for 5-10 hours per week over the project duration • \[IMP\]: Your voice may be cloned for the client's CX AI Agent so please only apply if you are okay with voice cloning. Your recordings will not be sold or reused outside of this specific use case. — Preferred Qualifications • Experience recording for TTS, audiobooks, IVR systems, or AI voice datasets • Familiarity with audio editing tools (e.g., Audacity, Adobe Audition, Reaper) • Ability to deliver multiple vocal styles (e.g., conversational, corporate, energetic, calm) • * *

$50 - $100 / hourOpen / Referral verified
Multimodal / Remote

Voice Actor: CX Agent Voice Cloning (Spanish - Peru)

Mercor is partnering with a leading AI research lab to build next-generation text-to-speech (TTS) systems capable of producing natural, expressive, and human-like voices. We are seeking Peru-based female voice actors with native Peruvian Spanish accents to contribute high-quality voice recordings for training and evaluating cutting-edge speech models. This role is ideal for professionals with experience in voice acting, narration, or broadcast who can deliver consistent, expressive, and clean audio across a variety of scripts. — Your voice recordings will be used exclusively for the client's internal CX AI Agent. They will not be sold, licensed, or reused for any other product, dataset, or purpose. • * * — Key Responsibilities • Record high-quality voice samples across diverse scripts (conversational, narrative, instructional, etc.) • Deliver clear, natural, and expressive speech with strong control over tone, pacing, and pronunciation • Maintain consistency in voice, accent, and delivery across recording sessions • Follow detailed recording guidelines (environment, microphone setup, file formatting) • Perform multiple takes with variation in emotion, emphasis, and style when required — Requirements • Native Peruvian female Spanish speaker currently based in Peru • Proven experience in voice acting, dubbing, narration, podcasting, or broadcasting • Access to a professional or near-professional recording setup (quality microphone, quiet environment, pop filter, etc.) • Strong command of intonation, diction, and emotional range • Ability to follow scripts precisely while maintaining natural delivery • Reliable availability for 5-10 hours per week over the project duration • \[IMP\]: Your voice may be cloned for the client's CX AI Agent so please only apply if you are okay with voice cloning. Your recordings will not be sold or reused outside of this specific use case. — Preferred Qualifications • Experience recording for TTS, audiobooks, IVR systems, or AI voice datasets • Familiarity with audio editing tools (e.g., Audacity, Adobe Audition, Reaper) • Ability to deliver multiple vocal styles (e.g., conversational, corporate, energetic, calm) • * *

$50 - $100 / hourOpen / Referral verified