Avg Pay
$157/hr
Pay Range
$60–$1000/hr
Categories
9
Active Gigs
45
Corporate Strategy Expert
Role Overview • Mercor is seeking senior corporate strategy professionals to build evaluation tasks for AI systems operating in Fortune 500 enterprise strategic planning contexts. • The workflows are calibrated to the organizational complexity, competitive stakes, and long-term capital implications of Fortune 500 and large public company strategic decisions. • Contributors design enterprise strategy scenarios, draft reference outputs, and write rubrics that capture how senior F500 strategy leaders think. Key Responsibilities • Construct enterprise strategy scenarios spanning multi-year corporate planning, multi-stakeholder board-level decision-making, and complex M&A, divestiture, or portfolio reallocation cycles at F500 accounts. • Build tasks across F500 corporate development, competitive strategy and market analysis, business unit portfolio strategy, capital allocation, and organizational strategy/restructuring. • Develop strategic planning scenarios involving tools such as enterprise financial modeling platforms, competitive intelligence systems, scenario-planning frameworks, and board-reporting/presentation platforms in F500 stacks. • Apply enterprise strategy methodologies (Porter's Five Forces, McKinsey 7-S, scenario planning, capital allocation frameworks, growth-share matrices) and produce reference strategic plans, board presentations, and executive-level strategic narratives. • Author rubrics that distinguish authentic enterprise strategic judgment from generic framework or MBA-case-study-level recall. Ideal Qualifications • 5+ years working in corporate strategy, corporate development, or strategic planning at a Fortune 500 company or top-tier strategy consulting firm (McKinsey, BCG, Bain) serving F500 clients. • Direct ownership of F500-scale strategic initiatives, M&A/divestiture processes, or board-level strategic recommendations. • Fluency in enterprise strategy frameworks and analytical tooling, plus understanding of how F500 board governance, capital markets expectations, and cross-functional executive alignment actually work. • Prior rubric, strategy-case curriculum, or board-deck/documentation authorship is a plus.
Computer Science PhD Researchers
Mercor is seeking Computer Science PhD’s (Graduated) for a premier project with one of the world's top AI labs. In this role, you will contribute your subject matter expertise to a cutting-edge project involving state-of-the-art large language models. Specifically, you will help create high-quality data that will inform the future of AI innovation by coming up with difficult problems in your domain. You're a good fit if you: • Received your undergraduate degree in US/UK/Canada/Western Europe • Received your graduate degree at a top US/UK/Canada/Western European university • Have high attention to detail • Have exceptional written and verbal communication skills • Have excellent proficiency in English Here are more details about the role: • The role is ongoing starting in February and continuing with rolling applications • Experts are expected to contribute 4-6 tasks per week, each taking several hours to complete • The work will require rigorous physics expertise and ability to follow complex instructions Screening Process: • You will need to complete a short AI interview - the whole application process should last 20-40 minutes • Apply today and leverage your leadership and technical expertise to advance cutting-edge AI models!
Accounting Expert
Role Overview Mercor is collaborating with a leading AI lab to engage experienced accounting professionals across all areas of practice — including audit, tax, financial reporting, bookkeeping, controllership, forensic, and accounting systems. Contributors help build AI systems that reason about real accounting work by translating everyday accounting workflows, judgments, and decision-making into structured, high-quality training data. Key Responsibilities • Design realistic accounting scenarios and tasks drawn from your day-to-day work (e.g., financial statement preparation, reconciliations, journal entries, audit procedures, tax filings, month-end close, internal controls) • Review and compare AI-generated accounting outputs for accuracy, standards compliance (GAAP/IFRS), and sound professional judgment • Create structured examples that reflect how accountants actually reason through problems • Provide clear written feedback that improves how AI performs accounting tasks • Collaborate asynchronously with the research team Ideal Qualifications • 3+ years of professional experience in accounting, audit, tax, bookkeeping, or finance operations (public accounting, corporate/in-house, advisory, or a firm) • CPA, CA, ACCA, CMA, EA, or an equivalent professional credential required • Bachelor's degree in Accounting, Finance, or a related field • Comfortable with common accounting tools (e.g., QuickBooks, NetSuite, SAP, Oracle, Excel) • Strong written communication and attention to detail More About the Opportunity • Open to all accounting specialties — contribute where your expertise is strongest • Work spans task design, evaluation, and structured feedback on AI accounting outputs • Strong contributors advance into reviewer, lead, and domain-expert roles Application Process • Submit a resume or a short summary of your accounting experience • Complete a short form on your practice area, specialties, and certifications • Selected applicants may complete a brief sample task • Follow-up typically provided within a few days
Legal Expert
Seeking a Legal Expert with extensive PACER and federal court records experience to review and QA AI-agent rollouts. You will assess AI accuracy against actual dockets, identify failure modes, and refine tasks for realism and difficulty. Key Requirements: * Regular, weekly or more, hands-on PACER/CM/ECF use for federal docket research. * 2+ years of professional federal litigation experience (paralegal, docket clerk, librarian, or litigator). * Ability to fluently read federal dockets, understand case numbers, and trace procedural posture. * Proficiency in PACER search mechanics and identifying name-variant traps. * Sound judgment to evaluate answer correctness and confidence to flag plausible but wrong answers. Pay: $70-$120/hr To apply, demonstrate your frequent PACER/CM/ECF use and 2+ years of federal litigation support experience.
Medicare Advantage Members (Devoted Health) – Insight Study
Mercor is conducting a paid research study in collaboration with a leading AI research lab focused on improving healthcare and member experiences. We are seeking current or former Devoted Health Medicare Advantage members/Age 64+ to participate in a short online survey about their experience with their health plan. Participants will share perspectives on plan enrollment, benefits, member support, and day-to-day healthcare experiences. Insights gathered will help inform the development of AI tools designed to improve how health plans serve their members. Responsibilities Participants will be asked to: • Complete a structured ~20-minute online survey • Provide a mix of multiple-choice answers and short voice-recorded responses (approximately 10 voice responses) • Share perspectives on their experience as a Medicare Advantage plan member • Submit all responses within the given timeline Requirements • Based in the United States • Age 64+ / Medicare-eligible • Currently or previously enrolled in a Devoted Health Medicare Advantage plan • Comfortable providing voice-recorded responses • Access to a microphone and a quiet environment • Able to independently complete a ~20-minute online survey Engagement Details • Format: Online survey (voice-recorded responses + multiple-choice questions) • Duration: Approximately 20 minutes • Compensation: One-time payment upon successful verification of the submission • Location: Remote (United States) Why Participate • Contribute to research shaping the next generation of healthcare AI tools • Share your real-world experience as a Medicare Advantage plan member
ML Engineer (Coding Agent Experience)
About the Role \- Mercor is partnering with a leading AI research lab to support a Frontier Code Agents project. \- Contributors help evaluate and improve frontier AI coding models through structured technical assessments. \- The work focuses on realistic machine learning engineering workflows and model evaluation. \- Spots are limited and filling quickly on a first come, first serve basis. What You'll Do \- Use frontier AI coding agents to complete and evaluate complex machine learning and AI engineering tasks. \- Review model-generated implementations involving model training, inference systems, MLOps, and LLM applications. \- Identify bugs, edge cases, performance issues, and failure modes. \- Compare outputs from multiple frontier models and assess their strengths and weaknesses. \- Apply professional engineering judgment to realistic ML engineering scenarios. Time Commitment \- Sprint based project that runs in 12-24 hour stretches based on client requirement. Compensation \- $400 per accepted task. \- Typical tasks take approximately 2–3 hours after ramp-up. \- Compensation is tied to accepted work. Who Should Apply \- 2+ years of professional machine learning engineering experience. \- Experience building production ML systems, model deployment infrastructure, LLM applications, or AI-powered products. \- Regular use of AI coding agents such as Cursor, Claude Code, Codex, Windsurf, Gemini CLI, or similar tools. \- Ability to evaluate model-generated machine learning implementations and technical tradeoffs. \- Experience deploying ML systems to production is preferred.
Electrical Engineering Expert
Role Overview • Mercor is seeking senior electrical engineering professionals to build evaluation tasks for AI systems operating in Fortune 500 enterprise electrical systems and product design contexts. • The workflows are calibrated to the design complexity, safety stakes, and production scale of Fortune 500 and large industrial and technology manufacturers. • Contributors design enterprise electrical engineering scenarios, draft reference outputs, and write rubrics that capture how senior F500 engineering leaders think. Key Responsibilities • Construct enterprise electrical engineering scenarios spanning complex circuit and systems design cycles, multi-stakeholder design reviews, and manufacturing or regulatory certification processes at F500 accounts. • Build tasks across F500 power systems and grid infrastructure, semiconductor and PCB design, embedded systems and firmware, signal processing, and control systems engineering. • Develop engineering scenarios involving tools such as Altium Designer, Cadence, MATLAB/Simulink, SPICE simulation platforms, and enterprise PLM/EDA systems in F500 stacks. • Apply enterprise electrical engineering methodologies (circuit analysis and simulation, EMC/EMI compliance, control theory, reliability and failure mode analysis) and produce reference design specifications, engineering analyses, and executive-level technical narratives. • Author rubrics that distinguish authentic enterprise electrical engineering judgment from generic textbook or coursework-level recall. Ideal Qualifications • 5+ years working as an electrical engineer or engineering lead at a Fortune 500 technology, industrial, or energy company (Intel, Texas Instruments, GE, Siemens, Tesla). • Direct ownership of F500-scale circuit designs, power systems, or engineering certification programs. • Fluency in enterprise electrical engineering tooling and methodologies, plus understanding of how F500 safety standards, regulatory certification (UL, FCC, IEC), and cross-functional design governance actually work. • Prior rubric, engineering-training curriculum, or design documentation authorship is a plus.
Software Developer — O*NET Occupation Study
Mercor is seeking experienced Software Developers to contribute to a study improving the O\*NET framework. Participants will share insights into their daily work through a brief online survey/interview. Key Requirements: * 4+ years of experience as a software developer * Current or recent work experience (within the last 12 months) * Direct knowledge of design, coding, and deployment tasks * U.S.-based Pay: $64/hr To apply or get started, complete the one-time online survey/interview.
Mechanical Engineering Expert
Role Overview • Mercor is seeking senior mechanical engineering professionals to build evaluation tasks for AI systems operating in Fortune 500 enterprise product design and manufacturing contexts. • The workflows are calibrated to the design complexity, safety stakes, and production scale of Fortune 500 and large industrial manufacturers. • This role builds worlds on two standards tracks: a US track (ASME) and an International track (ISO). Experts qualified in either or both tracks are encouraged to apply. • Contributors design enterprise mechanical engineering scenarios, draft reference outputs, and write rubrics that capture how senior F500 engineering leaders think. Key Responsibilities • Construct enterprise mechanical engineering scenarios spanning complex product design cycles, multi-stakeholder design reviews, and manufacturing or regulatory certification processes at F500 accounts. • Build tasks across F500 product design and CAD modeling, thermal/fluid systems analysis, structural and stress analysis, manufacturing process engineering, and reliability/failure analysis. • Develop engineering scenarios involving tools such as SolidWorks, ANSYS, CATIA, MATLAB/Simulink, and enterprise PLM systems (Siemens Teamcenter, PTC Windchill) in F500 stacks. • Apply enterprise mechanical engineering methodologies (FEA/CFD simulation, DFM/DFA principles, Six Sigma/DFSS, GD&T per ASME Y14.5 or ISO GPS) and produce reference design specifications, engineering analyses, and executive-level technical narratives. • Author rubrics that distinguish authentic enterprise mechanical engineering judgment from generic textbook or coursework-level recall. Ideal Qualifications • 5+ years working as a mechanical engineer or engineering lead at a Fortune 500 or major industrial, automotive, aerospace, or manufacturing company (GE, Boeing, Caterpillar, Ford, Honeywell, Siemens, Bosch, Rolls-Royce, Mitsubishi Heavy Industries). • Direct ownership of large-scale product designs, manufacturing processes, or engineering certification programs. • Fluency in enterprise mechanical engineering tooling and methodologies, plus understanding of how safety standards, regulatory certification (ASME, ISO, FAA, FDA), and cross-functional design governance actually work. • A professional engineering credential (US PE or an international equivalent such as CEng or EUR ING) and prior rubric, engineering-training curriculum, or design documentation authorship are a plus.
Lawyer — O*NET Occupation Study
About this study Mercor is building a new, more accurate version of O\*NET — the U.S. government's framework for describing occupations and the tasks they involve. We are gathering input directly from experienced practitioners to improve how the work of lawyers is described. About the role (SOC 23-1011.00 — Lawyers) Represent clients in criminal and civil litigation and other legal proceedings, draw up legal documents, or manage and advise clients on legal transactions. May specialize in a single area or practice broadly across many areas of law. What you'll do • Complete a one-time online survey/interview about your day-to-day work as a lawyer. It takes up to 30 minutes. • There may be an optional follow-up survey, task, or short interview afterward, which would be paid separately. Who we're looking for (eligible applicants have) • 4+ years of experience practicing law. • Current work, or work within the last 12 months, in the role. • Direct knowledge of legal research, drafting, and case work tasks. • U.S.-based. Not eligible (common confounders) • Paralegals and legal assistants (SOC 23-2011.00). • Judicial law clerks (SOC 23-1012.00). • Judges, magistrates, arbitrators, mediators, and conciliators (SOC 23-1023.00, 23-1022.00). Reference for the occupation: https://www.onetonline.org/link/summary/23-1011.00
Data Engineer (Coding Agent Experience)
About the Role \- Mercor is partnering with a leading AI research lab to support a Frontier Code Agents project. \- Contributors help evaluate and improve frontier AI coding models through structured technical assessments. \- The work focuses on realistic data engineering workflows and model evaluation. \- Spots are limited and filling quickly on a first come, first serve basis. What You'll Do \- Use frontier AI coding agents to complete and evaluate complex data engineering tasks. \- Review model-generated implementations involving ETL pipelines, data warehouses, analytics platforms, and distributed data systems. \- Identify bugs, edge cases, scalability issues, and failure modes. \- Compare outputs from multiple frontier models and assess their strengths and weaknesses. \- Apply professional engineering judgment to realistic data engineering scenarios. Time Commitment \- Sprint based project that runs in 12-24 hour stretches based on client requirement. Compensation \- $400 per accepted task. \- Typical tasks take approximately 2–3 hours after ramp-up. \- Compensation is tied to accepted work. Who Should Apply \- 2+ years of professional data engineering experience. \- Experience building ETL pipelines, data warehouses, analytics platforms, or distributed data systems. \- Regular use of AI coding agents such as Cursor, Claude Code, Codex, Windsurf, Gemini CLI, or similar tools. \- Ability to evaluate model-generated data infrastructure and pipeline implementations. \- Experience operating large-scale data platforms is preferred.
DevOps / SRE / Cloud Engineer (Coding Agent Experience)
About the Role \- Mercor is partnering with a leading AI research lab to support a Frontier Code Agents project. \- Contributors help evaluate and improve frontier AI coding models through structured technical assessments. \- The work focuses on realistic infrastructure engineering workflows and model evaluation. \- Spots are limited and filling quickly on a first come, first serve basis. What You'll Do \- Use frontier AI coding agents to complete and evaluate complex infrastructure engineering tasks. \- Review model-generated implementations involving cloud platforms, Kubernetes, CI/CD systems, observability, and infrastructure automation. \- Identify bugs, edge cases, reliability issues, and failure modes. \- Compare outputs from multiple frontier models and assess their strengths and weaknesses. \- Apply professional engineering judgment to realistic infrastructure engineering scenarios. Time Commitment \- Sprint based project that runs in 12-24 hour stretches based on client requirement. Compensation \- $400 per accepted task. \- Typical tasks take approximately 2–3 hours after ramp-up. \- Compensation is tied to accepted work. Who Should Apply \- 2+ years of professional DevOps, SRE, or Cloud Engineering experience. \- Experience with AWS, Azure, GCP, Kubernetes, Terraform, CI/CD pipelines, or observability tooling. \- Regular use of AI coding agents such as Cursor, Claude Code, Codex, Windsurf, Gemini CLI, or similar tools. \- Ability to evaluate model-generated infrastructure and reliability engineering solutions. \- Experience supporting production-scale systems is preferred.
Writing Expert
Role Overview Mercor is collaborating with leading AI labs to engage highly accomplished creative writers—novelists, journalists, short story authors, and essayists—for advanced AI training projects. Contributors will apply their literary and editorial expertise to improve AI systems’ ability to generate nuanced, high-quality narrative content. This work emphasizes strong storytelling, stylistic precision, and editorial judgment informed by published and awarded experience. This is a project-based opportunity with flexible participation. Key Responsibilities • Review and refine AI-generated content across articles, long form essays, novels, short stories, and other pieces of writing • Evaluate outputs for literary quality, structure, tone, and thematic depth • Provide detailed editorial feedback to improve coherence and originality Ideal Qualifications • 10+ years of experience in creative writing, including journalism, essays, novel writing, and short fiction • A substantial body of published work (e.g., books, bylines, or multiple short stories in recognized outlets) • Work featured in prestigious literary publications, journals, anthologies, or professional productions • Recipient of recognized literary awards, fellowships, or honors • Editors (including book, acquisitions, and developmental editors), literary agents, manuscript readers, literary scouts, documentary writers, speechwriters, memoir ghostwriters, critics, and other publishing or editorial professionals with significant experience evaluating high-quality written work may also be considered • Exceptional command of narrative craft, including structure, voice, dialogue, and character development across formats • Strong ability to critique and refine written work with attention to style, coherence, and thematic depth
Backend Engineer (Coding Agent Experience)
About the Role \- Mercor is partnering with a leading AI research lab to support a Frontier Code Agents project. \- Contributors help evaluate and improve frontier AI coding models through structured technical assessments. \- The work focuses on realistic software engineering workflows and model evaluation rather than traditional software development. \- Spots are limited and filling quickly on a first come, first serve basis. What You'll Do \- Use frontier AI coding agents to complete and evaluate complex engineering tasks. \- Review model-generated code for correctness, quality, maintainability, and performance. \- Identify bugs, edge cases, and failure modes in model outputs. \- Compare outputs from multiple frontier models and assess their strengths and weaknesses. \- Apply professional engineering judgment to realistic backend engineering scenarios. Time Commitment \- Sprint based project that runs in 12-24 hour stretches based on client requirement. Compensation \- $400 per accepted task. \- Typical tasks take approximately 2–3 hours after ramp-up. \- Compensation is tied to accepted work. Who Should Apply \- 2+ years of professional backend engineering experience. \- Experience building APIs, distributed systems, microservices, backend platforms, or databases. \- Regular use of AI coding agents such as Cursor, Claude Code, Codex, Windsurf, Gemini CLI, or similar tools. \- Ability to evaluate model-generated code and identify bugs, edge cases, and architectural tradeoffs. \- Experience working on large-scale production systems is preferred.
DNA Nanotechnology Expert
Role Overview Mercor is partnering with a leading AI lab to strengthen expert-level scientific reasoning in frontier models. We are hiring DNA nanotechnology experts to author and review challenging problems in structural DNA nanotechnology and to evaluate AI-generated solutions for correctness and rigor. What You'll Do • Design expert-level problems in DNA nanotechnology: DNA origami, nucleic-acid nanostructures, strand-displacement systems, and sequence/structure design • Review problems authored by peers for clarity, genuine difficulty, and ground-truth correctness • Evaluate and compare AI model outputs, delivering Accept / Revise / Reject verdicts with detailed written rationale Ideal Qualifications • PhD or equivalent research experience in chemistry, bioengineering, biophysics, or a related field with a focus on DNA/nucleic-acid nanotechnology • Hands-on experience with DNA origami, structural nucleic-acid design, or molecular self-assembly; familiarity with tools such as caDNAno, oxDNA, or NUPACK • Strong scientific writing and meticulous attention to detail
Consulting Expert
Consulting Expert — AI Training & Evaluation We're training frontier AI to perform advanced consulting work, and we need experienced consultants to create and grade expert-level deliverables. You'll take on realistic engagements end-to-end — diagnose the problem, work through dense and sometimes conflicting source material, and produce (or evaluate) the kind of rigorous, decision-grade output a client would actually pay for. What you'll do • Produce and review consulting deliverables such as: M&A / investment cases (EBITDA normalization, valuation, acquisition-debt structuring, covenant and sensitivity analysis, "proceed-with-conditions" recommendations); post-merger integration plans (synergy phasing vs. one-time costs, operating pro formas, integration narratives); and strategic / operational recommendations backed by quantitative analysis. • Synthesize messy inputs — data rooms, due-diligence reports, memos, conflicting stakeholder guidance — into a clear, structured, well-reasoned case, correctly weighing more-credible vs. stale or distractor sources. • Deliver client-ready work products: crisp memos, models, and recommendations with explicit assumptions, risks, and conditions. Who we're looking for • 4+ years at a top consulting or advisory firm — MBB, Big-4 strategy & transaction advisory (TAS/TS), corporate strategy, PMI & operations consulting, or boutique advisory — or equivalent in-house corporate-development experience. • Elite structured problem-solving and the ability to drive to a defensible recommendation under ambiguity. • Strong financial and business acumen plus advanced modeling (LBO / 3-statement, pro forma) and executive-quality written deliverables. • Comfort owning an engagement independently, not just supporting one. Details $95/hr · remote · flexible hours · contract.
LLM Research Scientist (Pre-training & Post-Training)
We're looking for experienced machine learning researchers with hands-on experience training and improving language models end-to-end. You'll work on well-scoped empirical open-ended LLM research problems. • Responsibilities • Train transformer-based language models from scratch and fine-tune open-weight models. • Get the most out of limited data and compute. • Construct training corpora from raw web-scale sources. • Build post-training pipelines. • Diagnose and resolve training issues. • Requirements We are looking for candidates with strong expertise in one or more of the following areas: Foundation Model Pre-training Experience with: • Training transformer-based language models from scratch, end-to-end. • Data- and compute-constrained regimes: allocating a fixed budget across model size, tokens, and epochs. • Diagnosing optimisation failures, convergence issues, and training instabilities. Pre-training Data Experience with: • Corpus construction from raw web crawls and other large unfiltered sources. • Data filtering, deduplication, quality classification, and mixture/ordering optimisation. • Measuring data interventions rigorously. LLM Post-Training Hands-on experience with one or more of: • Supervised fine-tuning, including building your own datasets via synthetic generation, noisy or weak supervision, and rejection sampling. • Preference optimisation (DPO, RLHF, RLAIF) and reward modelling / human-preference prediction. • Alignment fine-tuning: shaping refusal behaviour, truthfulness, and unbiased reasoning while preserving general capability. • Fine-tuning for narrow, verifiable domains (math, code, games, structured prediction) where outputs can be checked programmatically. Additional Areas of Interest Experience in any of the following is a plus: • Scaling laws and training-efficiency research. • Curriculum learning and data ordering. • LLM evaluation: benchmark construction, contamination control, statistically sound comparisons. • Reinforcement learning for language models. • Model alignment and AI safety. General Qualifications • 3+ years of machine learning research experience (PhD research counts toward this requirement). • Strong experience with PyTorch, JAX, TensorFlow, or similar ML frameworks. • Degree from a top-100 university, experience at a FAANG or comparable AI company, or an equivalent research track record through publications or impactful open-source contributions. • Why Join • Work on cutting-edge foundation model research. • Collaborate with leading AI researchers on challenging, high-impact projects. • Flexible, project-based work with competitive compensation.
Data Science Expert
1\. Role Overview Mercor is partnering with a leading AI research organization to engage experienced data scientists for a project focused on evaluating how well AI systems perform real-world data science work. Rather than producing deliverables yourself, you will define what excellent work looks like: designing task-specific grading criteria and scoring completed work samples with rigorous, well-reasoned written justifications. 2\. Key Responsibilities • Design precise, task-specific grading criteria for real-world data science deliverables (analyses, models, dashboards, experiment readouts, and written recommendations) • Score AI-generated and human work samples against those criteria, with detailed written justifications for every score • Apply consistent, evidence-based judgment so that scores are reproducible and defensible • Incorporate structured feedback from senior reviewers and iterate quickly on your work 3\. Ideal Qualifications • 5+ years of professional data science experience in industry • Background in business operations, product, or growth data science at top-tier technology companies • Deep fluency in experiment design and A/B testing, metric definition, SQL/Python analysis, and communicating findings to executive stakeholders • Exceptionally strong written communication • Detail-oriented, consistent, and comfortable having your judgment reviewed and calibrated against peers • Prior experience with AI training, evaluation, or human-data projects is a strong plus 4\. Application Process • Submit your resume or relevant technical background to get started • Qualified applicants may be asked to complete a brief technical assessment or submit additional information
Biology Expert (PhD) — AI Safety
About the role We're hiring PhD‑level biologists to help make advanced AI models safer. You'll apply your scientific expertise to evaluate and strengthen how these models handle specialized life‑science topics. No prior AI/ML experience is required — we'll train you on the workflow. What you'll do • Write expert‑level prompts across specialized life‑science topics. • Evaluate and annotate model responses for scientific accuracy, helpfulness, and appropriate handling of sensitive content. • Apply structured guidelines to classify prompts and conversations. What we're looking for • A PhD in biology or a closely related life‑science field — ongoing or completed (e.g., molecular biology, microbiology, virology, genetics, biochemistry, bioinformatics, synthetic biology, immunology). • Deep familiarity with modern laboratory and computational techniques in your subfield. • Strong scientific reasoning and writing in English. • Sound judgment around biosecurity and the responsible handling of dual‑use information. Nice to have • Research or coursework touching on biosafety/biosecurity, pathogen biology, or gain‑of‑function considerations. • Experience reviewing, grading, or red‑teaming technical content. Project Timeline • Start date: Immediate • Commitment: Part‑time (15–25 hours/week, with flexibility up to 40 hours/week)
Chemistry Expert (PhD) — AI Safety
About the role We're hiring PhD‑level chemists to help make advanced AI models safer. You'll apply your scientific expertise to evaluate and strengthen how these models handle specialized chemistry topics. No prior AI/ML experience is required — we'll train you on the workflow. What you'll do • Write expert‑level prompts across specialized chemistry topics. • Evaluate and annotate model responses for scientific accuracy, helpfulness, and appropriate handling of sensitive content. • Apply structured guidelines to classify prompts and conversations. What we're looking for • A PhD in chemistry or a closely related field — ongoing or completed (e.g., organic, inorganic, analytical, physical, medicinal, or computational chemistry; chemical engineering; materials science). • Deep familiarity with modern laboratory and computational techniques in your subfield. • Strong scientific reasoning and writing in English. • Sound judgment around chemical safety and the responsible handling of dual‑use information. Nice to have • Research or coursework touching on chemical safety/security, synthesis, or hazardous‑materials handling. • Experience reviewing, grading, or red‑teaming technical content. Project Timeline • Start date: Immediate • Commitment: Part‑time (15–25 hours/week, with flexibility up to 40 hours/week)
Management Consulting Expert
1\. Role Overview Mercor is partnering with a leading AI research organization to engage experienced management consultants for a project focused on evaluating how well AI systems perform real-world consulting work. Rather than producing deliverables yourself, you will define what excellent work looks like: designing task-specific grading criteria and scoring completed work samples with rigorous, well-reasoned written justifications. 2\. Key Responsibilities • Design precise, task-specific grading criteria for real-world consulting deliverables (market analyses, strategy recommendations, financial models, client-ready slide decks, and implementation roadmaps) • Score AI-generated and human work samples against those criteria, with detailed written justifications for every score • Apply consistent, evidence-based judgment so that scores are reproducible and defensible • Incorporate structured feedback from senior reviewers and iterate quickly on your work 3\. Ideal Qualifications • 5+ years of professional management consulting experience • Background at a top-tier strategy or management consulting firm (e.g., McKinsey, Bain, BCG, or equivalent) • Deep fluency in structured problem solving, market sizing, financial modeling and business-case building, and slide-based executive communication • Exceptionally strong written communication • Detail-oriented, consistent, and comfortable having your judgment reviewed and calibrated against peers • Prior experience with AI training, evaluation, or human-data projects is a strong plus 4\. Application Process • Submit your resume or relevant professional background to get started • Qualified applicants may be asked to complete a brief assessment or submit additional information
Cybersecurity SWE — AI Safety
About the role We're hiring cybersecurity‑focused software engineers to help make advanced AI models safer. You'll apply your security and engineering expertise to evaluate and strengthen how these models handle specialized cybersecurity topics. No prior AI/ML experience is required — we'll train you on the workflow. What you'll do • Write expert‑level prompts across specialized cybersecurity topics. • Evaluate and annotate model responses for technical accuracy, helpfulness, and appropriate handling of sensitive content. • Apply structured guidelines to classify prompts and conversations. What we're looking for • One of the following: • A BS or MS in Computer Science or a closely related field, or • 5+ years of professional software engineering experience at a reputable tech company or startup. • Strong understanding of cybersecurity concepts and modern software systems. • Strong technical reasoning and writing in English. • Sound judgment around security and the responsible handling of dual‑use information. Nice to have • Background in offensive security, penetration testing, vulnerability research, or related areas. • Experience reviewing, grading, or red‑teaming technical content. Project Timeline • Start date: Immediate • Commitment: Part‑time (15–25 hours/week, with flexibility up to 40 hours/week)
Physician - Research Survey (ONET Occupation Study)
About this study Mercor is building a new, more accurate version of ONET — the U.S. government's framework for describing occupations and the tasks they involve. We are gathering input directly from experienced practitioners to improve how the work of physicians is described. What you'll do • Complete a one-time online survey/interview about your day-to-day work as a physician. It takes up to 30 minutes. • There may be an optional follow-up survey, task, or short interview afterward, which would be paid separately. Pay • Pay is set in proportion to the expected hourly wage for physicians (~$127.85/hour). The initial task is \\~30 minutes\\, so it pays roughly half an hour at that rate, with the potential for additional paid tasks later. Who we're looking for • 2+ years of experience working specifically as a physician (SOC 29-1229 — Physicians, All Other). • U.S.-based. • Currently employed as a physician, or employed in the role recently — we are not looking for people who left the profession many years ago. • You must have actually held a physician role. Medical students or people with only general clinical exposure are not a fit unless they have genuinely worked as a physician. Reference for the occupation: https://www.onetonline.org/link/summary/29-1229.00
CUDA Engineering Expert
1\. Role Overview Mercor is seeking GPU kernel optimization experts to contribute to a project with a leading AI lab. This opportunity is designed for freelancers with strong C++ skills, practical GPU programming experience, and the ability to improve kernel performance using profiler-guided analysis. You’ll help evaluate, optimize, and reason about GPU kernels across modern hardware environments. This is a contract-based opportunity for specialists who enjoy squeezing performance out of modern GPU architectures. 2\. Key Responsibilities • Analyze and optimize GPU kernels for performance, efficiency, and hardware utilization • Use profiler metrics such as L2 cache hit rate, L2 throughput, occupancy, and related signals to guide kernel improvements • Review GPU kernel implementations and identify bottlenecks without requiring extensive background in the underlying algorithms • Write, modify, and reason about C++17, Python, and GPU programming code • Apply CUDA, HIP, shader programming, or related kernel programming expertise to improve performance outcomes • Document optimization decisions clearly, including when specific profiler metrics are or are not useful 3\. Ideal Qualifications • Available to work at least 20 hrs/wk • Fluent in core C++ features through C++17 • Working knowledge of Python and Git • Fluent in at least one GPU programming model, such as CUDA, HIP, Slang, HLSL, GLSL, or related kernel programming • At least 1 year of professional or graduate-level research experience working with GPUs • Strong understanding of GPU profiler performance metrics and how to use them to optimize kernels • Ability to optimize GPU kernels without needing deep prior context on every algorithm • Experience with CUDA, HIP, CUDA C++ Core Libraries, inline PTX assembly, or tensor core-level optimization is a plus • Experience optimizing kernels for NVIDIA Blackwell hardware is a plus • Familiarity with NSight Compute is a plus • Prior experience with GPU hardware organizations such as NVIDIA, AMD, or Qualcomm is a plus • Open-source contributions related to GPU kernel optimization are a plus 4\. Application Process • Submit your resume or relevant technical background to get started • Qualified applicants may be asked to complete a brief technical assessment or submit additional information
Healthcare Expert
About the Role Mercor is partnering with a leading AI research lab to train frontier models on high-quality clinical and biomedical reasoning. We're hiring Healthcare Experts to evaluate and compare model outputs on challenging medical problems, helping shape how the next generation of AI reasons about diagnosis, treatment, and patient care. Your clinical and medical expertise will ensure each evaluation is accurate, evidence-based, and aligned with real-world medical reasoning. What You'll Do • Compare two AI-generated responses to the same prompt and decide which one is better • Work from the provided input files and read each model's reasoning / thought process to inform your judgment • Clearly explain why one output outperforms the other, citing clinical accuracy, safety, and quality of medical reasoning • Provide written feedback the research team uses to improve model behavior • Participate in onboarding and specialty calibration sessions You're a Good Fit If You • Have an MD, DO, PharmD, PhD, or advanced clinical/biomedical degree — or equivalent professional experience • Have 2+ years of professional experience in clinical practice, medical research, or a biomedical specialty • Are comfortable applying current clinical guidelines and evidence-based medicine • Demonstrate excellent written communication with high attention to detail Role Highlights • Flexible workload: 10–20 hours per week • Role starts immediately; applications reviewed on a rolling basis
Accounting Expert
About the Role Mercor is partnering with a leading AI research lab to train frontier models on rigorous accounting and financial-reporting reasoning. We're hiring Accounting Experts to evaluate and compare model outputs on challenging accounting problems, helping shape how the next generation of AI reasons about financial statements, auditing, and tax. Your accounting expertise will ensure each evaluation is accurate, precise, and aligned with real-world professional standards. What You'll Do • Compare two AI-generated responses to the same prompt and decide which one is better • Work from the provided input files and read each model's reasoning / thought process to inform your judgment • Clearly explain why one output outperforms the other, citing correctness, standards compliance (GAAP/IFRS), and quality of reasoning • Provide written feedback the research team uses to improve model behavior • Participate in onboarding and specialty calibration sessions You're a Good Fit If You • Have 2+ years of professional experience in accounting, audit, or tax (public accounting, corporate finance, or advisory) • Hold a CPA (or equivalent) and/or a degree in Accounting, Finance, or a related field • Are fluent in GAAP and/or IFRS and financial-statement analysis • Demonstrate excellent written communication with high attention to detail Role Highlights • Flexible workload: 10–20 hours per week • Role starts immediately; applications reviewed on a rolling basis
Physics PhD Researchers
Mercor is seeking Physics PhD’s (Graduated) for a premier project with one of the world's top AI labs. In this role, you will contribute your subject matter expertise to a cutting-edge project involving state-of-the-art large language models. Specifically, you will help create high-quality data that will inform the future of AI innovation by coming up with difficult problems in your domain. You're a good fit if you: • Received your undergraduate degree in US/UK/Canada/Western Europe • Received your graduate degree at a top US/UK/Canada/Western European university • Have high attention to detail • Have exceptional written and verbal communication skills • Have excellent proficiency in English Here are more details about the role: • The role is ongoing starting in February and continuing with rolling applications • Experts are expected to contribute 4-6 tasks per week, each taking several hours to complete • The work will require rigorous physics expertise and ability to follow complex instructions Screening Process: • You will need to complete a short AI interview and form - the whole application process should last 20-40 minutes • Apply today and leverage your leadership and technical expertise to advance cutting-edge AI models!
Architecture Expert
Benchmark dataset project evaluating AI models on visual document understanding and instruction-following in the Architecture (building) domain. Experts author complex, grounded tasks with a clear ground-truth output and objective rubric. ~15–20 hrs/week, remote, US/Canada.
Legal Expert
About the Role Mercor is partnering with a leading AI lab to train frontier models on high-quality legal reasoning data. We're hiring Legal Experts to evaluate model outputs for realistic workplace scenarios and help shape how the next generation of AI reasons about employment disputes, workplace investigations, and labor matters. Your experience in legal research, legal writing, case analysis, and familiarity with the U.S. legal system will ensure that each your evaluations are accurate, precise, and aligned with real-world legal reasoning. What You'll Do • Evaluate two AI-generated responses for the same prompt for correctness, accuracy, and client-readiness. • Provide written feedback the research team uses to improve model behavior • Participate in onboarding office hours and specialty calibration sessions You're a Good Fit If You • Have 2+ years of professional experience in employment / labor law at a law firm, in-house legal department, government agency, or labor union • Hold a J.D. (U.S.) or LL.B./LL.M. from an accredited law school • Are licensed to practice law in the U.S. (active or inactive) • Have strong experience in legal research and legal writing, including case law analysis and drafting legal memoranda • Demonstrate excellent written communication skills with high attention to detail Role Highlights • Flexible workload: 6-15 hours per week • Role starts immediately, applications reviewed on a rolling basis Compensation • $100–$150/hr
Energy/Utilities Expert
Benchmark dataset project evaluating AI models on visual document understanding and instruction-following in the Energy/Utilities domain. Experts author complex, grounded tasks with a clear ground-truth output and objective rubric. ~15–20 hrs/week, remote, US/Canada.
Telecom/Network Expert
Benchmark dataset project evaluating AI models on visual document understanding and instruction-following in the Telecom/Network domain. Experts author complex, grounded tasks with a clear ground-truth output and objective rubric. ~15–20 hrs/week, remote, US/Canada.
Government Backoffice — Visual Document Understanding
Benchmark dataset project evaluating AI models on visual document understanding and instruction-following in the Government Backoffice domain. Experts author complex, grounded tasks with a clear ground-truth output and objective rubric. ~15–20 hrs/week, remote, US/Canada.
Real Estate Appraisal — Visual Document Understanding
Benchmark dataset project evaluating AI models on visual document understanding and instruction-following in the Real Estate Appraisal domain. Experts author complex, grounded tasks with a clear ground-truth output and objective rubric. ~15–20 hrs/week, remote, US/Canada.
Electrical & Electronics Expert
Benchmark dataset project evaluating AI models on visual document understanding and instruction-following in the Electrical & Electronics domain. Experts author complex, grounded tasks with a clear ground-truth output and objective rubric. ~15–20 hrs/week, remote, US/Canada.
Surveying & GIS — Visual Document Understanding
Benchmark dataset project evaluating AI models on visual document understanding and instruction-following in the Surveying & GIS domain. Experts author complex, grounded tasks with a clear ground-truth output and objective rubric. ~15–20 hrs/week, remote, US/Canada.
Behavioral Health Provider – Insurance & Practice Insights Study (U.S.A)
Mercor is seeking licensed Behavioral Health Providers to participate in a short-term research study with a leading AI lab. This project aims to better understand how insurance dynamics—particularly payer mix and administrative workflows—impact small behavioral health practices. Participants will share insights on their day-to-day experience with insurance-related processes and how these influence clinical and operational decisions. • Key Responsibilities • Participate in a structured survey. • Provide insights on insurance workflows, including billing, claims, and payer interactions • Offer feedback on administrative challenges within small behavioral health practices • Requirements • Licensed Behavioral Health Provider (Psychiatrist, Psychologist, LCSW, LPC, or LMFT) • Currently practicing in a clinical setting (e.g., hospital, clinic, or behavioral health center) • Some involvement in insurance-related processes (e.g., billing, claims, contracting, or reimbursement) • Experience working in small practice environments (lean teams, close involvement in operations) • Based in the United States • Work Arrangement • One-time engagement (30–45 minutes) • To be completed in 7 days after if shortlisted • One-time task payment of $100 USD • Why Join • Contribute to research shaping AI tools for healthcare operations • Share real-world insights from frontline behavioral health practice • Fast, structured engagement with competitive compensation
Finance & Banking Expert
We are building a benchmark dataset to evaluate AI models on professional document understanding and instruction following within the Finance & Banking domain. Tasks consist of complex, multi-step requests grounded in real-world workspace files (financial statements, reports, spreadsheets), web search, and code execution — each paired with a clearly defined ground truth output and an objective evaluation rubric. You will be responsible for authoring tasks that test an AI's ability to reason over financial documents, follow precise instructions, and produce accurate, structured outputs. We expect a minimum commitment of 15–20 hours per week. Ideal candidates have 3+ years of hands-on experience in one or more of the following sub-domains: • Investment & financial analysis • Financial management • Personal financial advisory • Banking & capital markets • Corporate finance & accounting
Lawyers
1\. Role Overview Mercor is partnering with a leading AI organization to engage practicing attorneys for a project focused on improving the legal reasoning of AI systems. This flexible, contract-based engagement involves evaluating and refining model-generated legal content to make it more accurate, clear, and useful. 2\. Key Responsibilities • Review and refine AI-generated legal outputs for accuracy and clarity • Evaluate legal reasoning across a range of practice areas • Verify claims against primary authority (statutes, regulations, cases, official guidance) • Collaborate with researchers to assess and improve model performance 3\. Ideal Qualifications • JD from an ABA-accredited US law school • Active US bar license, in good standing • 3+ years in practice with client- or matter-facing advisory experience • Demonstrated ability to locate and cite primary authority • Generalist / multi-practice exposure a plus • Clear, plain-language writer comfortable researching outside a primary specialty 4\. More About the Opportunity • Remote and asynchronous — flexible scheduling • Project-based work at the intersection of law and AI 5\. Compensation & Contract Terms • $100–200/hour for U.S.-based professionals • Paid weekly via Stripe Connect • Independent contractor engagement 6\. Application Process • Submit your resume to begin • Complete a short form outlining your practice areas and experience • Mercor will follow up with qualified applicants 7\. About Mercor • Mercor is a talent marketplace that connects top experts with leading AI labs and research organizations • Our investors include Benchmark, General Catalyst, Adam D'Angelo, Larry Summers, and Jack Dorsey • Thousands of professionals across law, engineering, research, and creative domains use Mercor to collaborate on frontier AI projects
Psychiatry Expert
About the Role Mercor is partnering with a leading AI lab to train frontier models on high-quality healthcare reasoning data. We're hiring Psychiatry experts to design clinical scenarios, evaluate model outputs against evidence-based standards, and help shape how the next generation of AI reasons about mental health care. We welcome general adult Psychiatry attendings, Child & Adolescent, Addiction, Forensic, Consultation-Liaison, and Geriatric Psychiatry subspecialists, and final-year Psychiatry residents. What You'll Do • Design clinically realistic prompts and scenarios drawn from your Psychiatry practice (diagnostic formulation, medication management, risk assessment, psychotherapy fidelity, capacity evaluation) • Write "golden" reference responses at attending-level quality • Grade AI-generated responses against structured rubrics • Provide written feedback the research team uses to improve model behavior • Participate in onboarding office hours and specialty calibration sessions Who You Are • Attending physicians: Must be board certified with current, active, unrestricted medical license • Resident physicians: Must be in final year of residency (recent graduates must be board-eligible) • Fellows: Must be board-certified/board-eligible in primary specialty and have current active, unrestricted medical license Rates $150–$350/hr payable, set per seniority tier (early-career → senior attending). Your specific rate is confirmed in your offer. Engagement • Remote, 100% asynchronous • 20 hrs/wk default (can be raised after onboarding based on demand) • Paid weekly via Mercor • Start: rolling, after onboarding sign-off • Duration: ongoing, reviewed monthly Before Your First Task You'll receive an onboarding doc and a calendar invite for Onboarding Office Hours after you accept.
Investment Banking Expert
Mercor is recruiting U.S./UK/Canada/Europe/Australia-based Investment Banking Experts for a research project with a leading foundational model AI lab. You are a good fit if you: • Have at least 2 years of experience working at top firms in investment banking and experience in at least one of the following • Financial Modeling • Pitch Decks • Investment/Analysis Summaries and Memos • Company/Industry Analysis Here are more details about the role: • You must be able to commit at least 10 hours per week for this role • This is a minimum four week engagement, with potential for significant extension or rotation to similar, future projects • Successful contributions increase the odds that you are selected on future projects with Mercor • This role will pay between $100-$130/hour with potential for increases for top performers
Social Work Expert
Seeking experienced social workers to help build realistic digital environments and design challenging tasks for AI agents. This role involves recreating your daily digital workspace and authoring tasks based on real-world cases and programs. Key Requirements: * LCSW/LMSW or equivalent licensure strongly preferred. * 3+ years of full-time experience in a large hospital system, county DHS/DCF, or major nonprofit/behavioral-health organization. * Experience in clinical social work, child/adult protective services, or macro social work. * Familiarity with platforms like Salesforce Nonprofit, Apricot, DocuSign, SharePoint, or Box. Pay: $850-$1000 per completed task, with potential for bonuses and hourly compensation for top performers. To apply, follow the instructions provided by Mercor.
Physics PhD
Mercor is seeking Physics PhD’s (Graduated) specializing in one of the following sub-domains: • Statistical Mechanics, Condensed Matter & AMO Physics • Advanced Quantum, Electrodynamics & Classical Mechanics • String theory, QFT, Particle physics & Nuclear physics • General Relativity, Astrophysics & Cosmology In this role, you will contribute your subject matter expertise to a cutting-edge project involving state-of-the-art large language models. Specifically, you will help create high-quality data that will inform the future of AI innovation by coming up with difficult problems in your domain. You're a good fit if you: • Hold a PhD in Physics • Received your graduate degree in US/UK/Canada/Western Europe • Have high attention to detail • Have exceptional written and verbal communication skills • Have excellent proficiency in English Here are more details about the role: • The role is ongoing starting in February and continuing with rolling applications • Experts are expected to contribute 4-6 tasks per week, each taking several hours to complete • The work will require rigorous physics expertise and ability to follow complex instructions Screening Process: • You will need to complete a short AI interview and form - the whole application process should last 20-40 minutes
Hematology / Oncology Expert
Seeking a Hematology/Oncology expert for remote consultation. This role involves providing expert insights and guidance in the field. * Expertise in Hematology and Oncology * Strong analytical and communication skills Pay: $130-$180/hr Apply through Mercor.
Law Experts
Seeking experienced legal professionals for remote freelance opportunities. This role involves providing expert legal analysis and consultation. * Expertise in relevant legal fields * Strong analytical and communication skills Pay: $100-$130/hr To apply, please visit the Mercor platform.
Management & Strategy Consultants (MBB/Big 5)
1. Role Overview Project Panacea is a Mercor research initiative focused on training AI agents to handle complex, real-world business and consulting work. As a contributor, you'll help design and refine the scenarios used to evaluate and improve how AI systems approach multi-document, multi-stakeholder business problems. This is a long-term role with flexible hours and a consistent workload. 2. Key Responsibilities • Conduct market research and competitive analysis • Develop business cases, operational frameworks, and go-to-market strategies • Synthesize data into insights using presentations, memos, and models • Collaborate with stakeholders to define key metrics and decision points • Support strategic planning, scenario modeling, and opportunity evaluation 3. Ideal Qualifications • 1.5+ years at a top management consultancy (e.g., McKinsey, BCG, Bain, or Big 5) • Strong analytical and communication skills, both written and verbal • Comfortable working independently in ambiguous or rapidly evolving contexts • Proficiency in PowerPoint, Excel, and basic data analysis tools • Experience advising on tech, AI, or SaaS industries preferred but not required 4. More About the Opportunity • Expected commitment: minimum 10 hours/week 6. Application Process • Submit your resume to get started • You'll complete a short form to detail your consulting background • We'll follow up within a few days with potential next steps