Role directory

Data engineering jobs

2 active, referral-verified opportunities.

Code / Remote

Applied Computer Science Benchmark Specialist

Role Overview — We are seeking expert computer scientists to author and review high-quality academic assessment content for an AI research initiative. You will write and verify rigorous multiple-choice questions across core computer science domains, evaluate solution quality, and help establish gold-standard benchmarks used to advance AI capabilities. — You will be assigned one of two task types: • Question Authoring — Create original, challenging multiple-choice questions in your area of computer science expertise, rate their difficulty, and submit them for review. • Question Verification — Review pre-written questions for accuracy, clarity, and rigor. Edit where needed, rate difficulty, and document any changes made. — Computer Science Domains Covered — Accelerator / GPU Kernel Engineering, Formal Methods & Automated Reasoning, Computer Architecture & Accelerators, Distributed Systems, DevOps & Site Reliability, Data Engineering & Databases, Cloud & Infrastructure, OS & Systems Kernel, Machine Learning Engineering, Web & API Development, Embedded Systems Engineering, Computer Graphics & Game Development, Mobile Engineering. — Key Responsibilities • Author original computer science questions that test deep conceptual understanding, not surface-level recall • Ensure questions are unambiguous, self-contained, and precisely defined — all necessary information must be in the problem statement • Rate each question's difficulty: Medium (intro undergraduate), Hard (advanced undergraduate), or Expert (post-graduate and above) • Provide 1 correct answer and 9 plausible but subtly incorrect alternatives that challenge expert-level solvers • Write step-by-step Chain-of-Thought solutions with clear, concise intermediate steps in markdown format • Supply 1–5 academic references per question from reputable sources (peer-reviewed journals, university repositories) • For verification tasks: flag issues with clarity, completeness, precision, or solvability and justify any edits made — Ideal Qualifications • PhD or doctoral candidate in Computer Science, Electrical Engineering, or a closely related field • Master's degree considered for candidates with exceptional depth in a specific subdomain • Strong command of graduate-level CS theory, algorithms, systems design, and/or machine learning • Research publications, industry experience at top tech companies, or competitive programming background is a strong plus • Excellent written English and ability to express complex ideas clearly and concisely — More About the Opportunity • Expected commitment: 10+ hours/week • Asynchronous, fully remote work

$66 - $84 / hourOpen / Referral verified
Code / Remote

Data Engineer (Coding Agent Experience)

About the Role — \- Mercor is partnering with a leading AI research lab to support a Frontier Code Agents project. — \- Contributors help evaluate and improve frontier AI coding models through structured technical assessments. — \- The work focuses on realistic data engineering workflows and model evaluation. — \- Spots are limited and filling quickly on a first come, first serve basis. — What You'll Do — \- Use frontier AI coding agents to complete and evaluate complex data engineering tasks. — \- Review model-generated implementations involving ETL pipelines, data warehouses, analytics platforms, and distributed data systems. — \- Identify bugs, edge cases, scalability issues, and failure modes. — \- Compare outputs from multiple frontier models and assess their strengths and weaknesses. — \- Apply professional engineering judgment to realistic data engineering scenarios. — Time Commitment — \- Sprint based project that runs in 12-24 hour stretches based on client requirement. — Compensation — \- $400 per accepted task. — \- Typical tasks take approximately 2–3 hours after ramp-up. — \- Compensation is tied to accepted work. — Who Should Apply — \- 2+ years of professional data engineering experience. — \- Experience building ETL pipelines, data warehouses, analytics platforms, or distributed data systems. — \- Regular use of AI coding agents such as Cursor, Claude Code, Codex, Windsurf, Gemini CLI, or similar tools. — \- Ability to evaluate model-generated data infrastructure and pipeline implementations. — \- Experience operating large-scale data platforms is preferred.

$80 / hourOpen / Referral verified