Clinical Medicine Domain Expert
Weekday AI · California City, California, United States · Hybrid
Pay: USD 70 – 110 a hour
Posted Sep 2, 2026
Sign up free: we match you to jobs like this, tailor your application and fill the form. 2 free applications every day.
This role is for one of our clients
Compensation: $70 - $110 per hour
We are seeking an experienced Clinical Medicine Domain Expert to contribute to a cutting-edge GenAI initiative focused on improving how advanced AI models understand, reason about, and respond to real-world clinical scenarios.
Your clinical expertise will be central to this role. You will evaluate medical knowledge tasks and AI-generated outputs, develop detailed instructions and reference solutions, and create rigorous benchmarks that establish what high-quality clinical reasoning looks like. We are looking for an experienced practicing clinician with deep expertise in a defined medical specialty.
This is a full-time position requiring 40 hours per week for an initial six-month engagement . The role involves working closely with research and program teams and using standard enterprise tools and workflows.
Location: Hybrid role based in the Bay Area, California . Candidates must be based in the Bay Area and able to work on-site with the team multiple days per week when required. This is not a fully remote position. Candidates currently outside the Bay Area must be willing to relocate at their own expense before the engagement begins. Relocation assistance is not provided.
Requirements
Key Responsibilities
Clinical Data Quality & Evaluation
Review and assess clinical knowledge tasks and AI-generated medical outputs for accuracy, safety, completeness, and clinical relevance.
Identify missing reasoning steps, unsupported conclusions, unsafe recommendations, deviations from established guidelines, and clinically inappropriate responses.
Evaluate whether AI-generated answers meet the standards expected in real-world clinical practice.
Instruction & Reference Dataset Development
Develop clear, detailed instruction specifications that define expected clinical reasoning and outcomes.
Create high-quality reference solutions for complex clinical problems and scenarios.
Develop new clinical tasks that accurately represent how healthcare professionals evaluate and solve real-world medical problems.
Benchmark & Evaluation Development
Design challenging clinical scenarios and evaluation datasets that test medical reasoning, decision-making, and domain knowledge.
Contribute to the development of medicine-specific benchmarks, tools, and evaluation methodologies.
Help establish measurable standards for assessing the quality and reliability of AI-generated clinical responses.
Expert Calibration & Collaboration
Collaborate with researchers and specialists from adjacent medical and technical disciplines to maintain consistent evaluation standards.
Translate practical clinical judgment and experience into explicit, structured, and teachable evaluation criteria.
Provide precise written feedback that helps improve the quality and reliability of AI systems.
Core Qualifications
Education: MD or DO from an accredited medical school, along with completion of residency…