THE OPPORTUNITY
Put your skills to work.
In this hourly, remote contractor role, you will work as a Healthcare Subject Matter Expert (SME) to create challenging, expert level healthcare questions, evaluate AI generated clinical content, and design high quality rubrics defining what an accurate, complete, and professionally appropriate answer should contain.…
- Employment
- Contract
- Interview language
- English
- Applied this month
- 64
In this hourly, remote contractor role, you will work as a Healthcare Subject Matter Expert (SME) to create challenging, expert-level healthcare questions, evaluate AI-generated clinical content, and design high-quality rubrics defining what an accurate, complete, and professionally appropriate answer should contain.
You will use your clinical or professional expertise to identify areas where frontier AI models may provide incomplete, inaccurate, poorly reasoned, overconfident, or clinically inadequate responses. You will write specialized questions within your field, develop detailed scoring criteria, critique rubrics created by other healthcare professionals, and provide precise written feedback.
A major focus of this role is rubric design and expert evaluation. You will define essential clinical concepts, reasoning steps, relevant caveats, acceptable alternative approaches, significant omissions, and critical errors that distinguish superficial answers from expert-level responses.
This role is with SME Careers, a fast-growing AI Data Services company and subsidiary of SuperAnnotate, delivering training data for many of the world’s largest AI companies and foundation-model labs. Your healthcare expertise directly helps improve the world’s premier AI models by making their responses more accurate, clinically appropriate, rigorous, and reliable.
Responsibilities
- Create expert healthcare questions: Develop specialized questions capable of exposing limitations in advanced AI systems.
- Design detailed rubrics: Define the components required for an accurate, complete, clinically appropriate answer.
- Identify critical criteria: Distinguish essential clinical reasoning from supplementary information.
- Evaluate AI responses: Assess model outputs for factual accuracy, reasoning, safety, completeness, and professional quality.
- Identify subtle failures: Detect missing considerations, unsupported assumptions, inappropriate certainty, and clinically important omissions.
- Critique peer rubrics: Identify ambiguity, redundancy, incorrect expectations, missing criteria, or poor scoring design.
- Develop reference answers: Produce expert examples, explanations, critiques, and gold-standard content.
- Support AI training: Generate questions, rubrics, evaluations, critiques, and reference-answer pairs for RL and SFT workflows.
- Maintain clinical quality: Ensure content reflects appropriate terminology, professional reasoning, and evidence-aware judgment.
Requirements
- Professional degree, license, credential, or recognized qualification in medicine, dentistry, pharmacy, nursing, or another healthcare discipline.
- Strong professional proficiency in English, minimum C1.
- 3+ years of professional or clinical experience; senior and specialist-level professionals are strongly preferred.
- Demonstrated expertise in a clearly defined specialty or subspecialty.
- Ability to develop challenging clinical or professional questions that require genuine expertise and nuanced judgment.
- Strong ability to write detailed evaluation rubrics defining required facts, reasoning steps, caveats, acceptable alternatives, and critical clinical errors.
- Ability to distinguish between minor omissions and errors that materially affect clinical quality or safety.
- Comfortable reviewing and critiquing rubrics created by other healthcare professionals.
- Ability to evaluate AI outputs for accuracy, completeness, reasoning quality, safety, calibration, and clinical relevance.
- Experience with medical education, examination writing, competency assessment, clinical guidelines, teaching, peer review, or quality assurance is strongly preferred.
- Prior experience with AI evaluation, RLHF, SFT, data annotation, model benchmarking, or rubric-based scoring is preferred.
- Reliable, self-directed, and capable of producing consistent written work.
Open to experts in
United States
Role details are copied from the SME Careers listing and can change. Check the platform before you apply.
Relevant skills
HealthcaremedicineDentistry
A few things to check
- Location and experience requirements.
- Schedule, compensation, and assessment process.
- Current availability on SME Careers.