Physician (MD/DO) - Medical AI Evaluation
$130 per hourMindrift
Please submit your CV in English and indicate your level of English proficiency.
Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems. Participation is project-based.
What this opportunity involves
We’re looking for US-based, actively practicing physicians (MD/DO) to evaluate clinical EHR vignettes for a medical AI evaluation program. General and internal medicine are our main focus. While each project involves unique tasks, contributors may:
- Evaluate clinical EHR vignettes paired with a question, a proposed answer, and a distractor “trap” answer across diagnosis and treatment tasks, spanning cardiovascular, nervous, hematologic, respiratory, digestive, urinary, reproductive, musculoskeletal, and integumentary systems;
- Score the clinical reasoning quality of benchmark items: vignette accuracy and completeness, whether the vignette gives the answer away, answer correctness and gradeability, and trap quality;
- Check whether the reasoning chain reaches the answer from vignette facts alone, correct reasoning traces, and write short rationales;
- Work within three independent blind reads, followed by physician adjudication. Real clinical complexity only. You’re improving the AI tools you’ll eventually use yourself.
If you’re a practicing physician ready to take on this challenging and engaging project, join us!
What we look for
This opportunity is a good fit for US-based physicians open to part-time, non-permanent projects. Ideally, contributors will have:
- Medical degree (MD or DO) and an active, unrestricted US medical license (verified with the issuing medical board);
- Board certification or completed residency training, with recent direct patient care (attending-level preferred);
- Broad, multi-system diagnostic and treatment experience (internal, family, emergency, or hospital medicine);
- Strong clinical reasoning: differential diagnosis, next-step management, and application of evidence-based guidelines;
- Prior experience in medical AI evaluation, clinical content or exam-question review, or medical annotation/QA (a plus);
- Strong written English (C1+).
This opportunity is not fit to medical students, unlicensed medical graduates, non-physician clinicians (NPs, PAs, nurses, pharmacists), or non-clinical healthcare titles.
Project time expectations
For this project, tasks are estimated to require around 10-20 hours per week during active phases, based on project requirements. This is an estimate, not a guaranteed workload, and applies only while the project is active. Tasks must be submitted by the deadline and meet the listed acceptance criteria to be accepted.
Compensation
Paid per accepted task. Your rate depends on the qualification tier you reach and how efficiently you complete tasks — up to the equivalent of $130/hr . Because payment is per approved task, a faster pace raises your effective hourly rate.
$130 per hour
...Mindrift connects specialists with project-based AI opportunities for leading tech companies, focused on testing, evaluating, and improving AI systems. Participation is... ...’re looking for US-based, actively practicing physicians (MD/DO) to evaluate clinical EHR vignettes for a...MedicalHourly payPermanent employmentPart timeFreelance10 hours per week- ...collaboration with other disciplines - all foundational ingredients in successful digital experiences and client partnerships. As a QA / AI Evaluation Engineer, you will join a highly motivated and experienced team in a forward-leaning role that proves the platform actually...SuggestedFull timeImmediate start
$80 - $120 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Position: Compliance / regulatory response with financial-services AI Evaluator Type: Contract Compensation: $80–$120/hour...SuggestedFull timeContract workSummer workWork at officeRemote work$80 - $120 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Position: Compliance / regulatory response with financial-services AI Evaluator Type: Contract Compensation: $80–$120/hour...SuggestedFull timeContract workSummer workWork at officeRemote work$100 - $150 per hour
...creative and technical talent with leading AI research labs. Headquartered in San... ...pipelines , and A/B test write-ups . Evaluate AI-generated or human-created work against... ...detailed written justifications for evaluations and scores to maintain transparency and accountability...SuggestedFull timeContract workSummer workRemote work$100k - $150k per year
...efficient and productive. We are looking for an experienced technical product manager to design and deliver high-value AI Agent solutions for Vault Medical. In this role, you will be the expert in leveraging AI to automate the content lifecycle, from speeding up the Medical...MedicalRemote jobFull timeInternshipWork at officeLocal areaWork from home$85 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Responsibilities Use frontier AI coding agents to complete and evaluate complex infrastructure engineering tasks. Review model-...Full timeContract workSummer workRemote work$80 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Responsibilities Use frontier AI coding agents to complete and evaluate complex data engineering tasks. Review model-generated...Full timeContract workSummer workRemote work$60 - $80 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include... ...D'Angelo , Larry Summers , and Jack Dorsey . Position: Medical Secretaries and Administrative Assistants Type: Contract...MedicalFull timeContract workSummer workImmediate startRemote workFlexible hours$94 - $119 per hour
...creative and technical talent with leading AI research labs. Headquartered in San... ...Qualifications Must-Have MD , DO , PhD , or doctoral candidate in Medicine... ...subdomain. Strong command of graduate-level medical knowledge, clinical reasoning, and biomedical...MedicalFull timeContract workSummer workRemote work$85 - $93 per hour
...technical talent with leading AI research labs. Headquartered in... ...Dorsey . Position: Basic Medical Sciences Expert Type:... ...and review technical tasks to evaluate AI model performance on basic... ...in a basic medical science, MD/DO with graduate-level research...MedicalHourly payFull timeContract workFor contractorsSummer workRemote workFlexible hours$50 - $65 per hour
...creative and technical talent with leading AI research labs. Headquartered in San... ..., and Jack Dorsey . Position: Non-Physician Admin (CL Funnel) Type: Contract... ...author healthcare-operations tasks. Evaluate tasks for AI model training to ensure accuracy...MedicalFull timeContract workSummer workRemote workFlexible hours$100 - $150 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...hour Location: Remote Role Responsibilities Evaluate AI-generated artifacts for usability in a software engineering...Contract workSummer workWork at officeRemote workusd16 per hour
...We are looking for AI Linguistic Evaluators to assess how effectively an AI application performs in Telugu for a short collaboration. Job Type... ...naturalness, and cultural appropriateness. - Providing structured evaluations and feedback following the project instructions and quality...Hourly payDaily paidFull timeFreelanceImmediate startRemote workFlexible hours$20 - $160 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Position: Generalist Annotator — Health AI Conversation Quality Evaluation Type: Contract Compensation: $20–$160/hour...Remote jobContract workSummer work$60 - $80 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Jack Dorsey . Position: Business Domain Expert — AI Model Evaluation Pilot (Admin, Marketing, HR, Accounting) Type: Contract...Hourly payFull timeContract workFor contractorsSummer workRemote work$80 - $160 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Utilize strong analytical skills to contribute to high-impact evaluation projects. Collaborate with teams to improve document standards...Full timeContract workSummer workRemote work$70 per hour
...About the job Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include Benchmark , General Catalyst , Peter Thiel , Adam D'Angelo , Larry Summers , and Jack Dorsey . Position...Remote jobContract workSummer workImmediate start$60 - $80 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors include... ...D'Angelo , Larry Summers , and Jack Dorsey . Position: Medical Secretaries and Administrative Assistants Type: Contract...MedicalFull timeContract workSummer workImmediate startRemote work$15 - $20 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Angelo , Larry Summers , and Jack Dorsey . Position: Video Evaluation Generalist Type: Contract Compensation: $15–$...Full timeContract workSummer workRemote work$80 - $120 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Summers , and Jack Dorsey . Position: Legal / compliance Evaluator Type: Contract Compensation: $80–$120/hour...Full timeContract workSummer workWork at officeRemote work$80 - $150 per hour
...creative and technical talent with leading AI research labs. Headquartered in San... ...Position: MS Excel / Google Sheets Evaluator Type: Contract Compensation... ...structured feedback to enhance document evaluation processes. Utilize MS Excel and Google...Full timeContract workSummer workRemote work$90 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Preferred Prior RLHF , preference-labeling, model-evaluation, or structured code-review work. Pixel-perfect design-to-code...Full timeContract workSummer workLocal areaRemote work$100 - $150 per hour
...creative and technical talent with leading AI research labs. Headquartered in San... ...pipelines , and technical reports. Evaluate AI-generated or human-created work against... ...Provide detailed written justifications for evaluations and scores to maintain transparency....Remote jobContract workSummer work$80 - $120 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Dorsey . Position: Customer success / support operations Evaluator Type: Contract Compensation: $80–$120/hour...Full timeContract workSummer workWork at officeRemote work$80 - $120 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Summers , and Jack Dorsey . Position: General Sales / GTM Evaluator Type: Contract Compensation: $80–$120/hour...Full timeContract workSummer workWork at officeRemote workFlexible hours$100 - $150 per hour
...creative and technical talent with leading AI research labs. Headquartered in San... ...objections, supporting proofs of concept, and evaluating product fit and technical risk. ~... ...~ Prior experience with AI training , evaluation, or human-data projects. Application...Full timeContract workSummer workRemote work$80 - $120 per hour
...creative and technical talent with leading AI research labs. Headquartered in San... ...Position: Procurement / vendor management Evaluator Type: Contract Compensation... ...deadlines while ensuring high-quality evaluations. Qualifications Must-Have...Full timeContract workSummer workWork at officeRemote work$70 - $100 per hour
...Mercor connects elite creative and technical talent with leading AI research labs. Headquartered in San Francisco, our investors... ...Commitment: 20–40 hours/week Role Responsibilities Evaluate AI-generated office artifacts in Excel , Word , and PowerPoint...Full timeContract workSummer workWork at officeImmediate startRemote work$25 - $45 per hour
...creative and technical talent with leading AI research labs. Headquartered in San... ...Remote Role Responsibilities Evaluate AI-generated written conversations using... ...enhance performance. Draft and refine evaluation rubrics for generalist writing tasks....Full timeContract workSummer workRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Physician (MD/DO) - Medical AI Evaluation. Be the first to apply!

