Jobs Companies Weekday AI Applied Health & Medicine Benchmark Specialist

About this Applied Health & Medicine Benchmark Specialist role at Weekday AI

Weekday AI · Remote · United States

This role is for one of our clients

Compensation: $94 - $119 per hour

We are seeking highly qualified medical and health science professionals to contribute to an AI research initiative focused on developing rigorous academic assessment content across health and medicine.

As an Applied Health & Medicine Benchmark Specialist, you will author and review challenging multiple-choice questions, evaluate the quality and validity of existing assessment content, and help establish high-quality benchmarks for measuring AI performance on advanced medical and health science concepts.

You may be assigned one of two primary task types:

  • Question Authoring: Develop original, challenging multiple-choice questions within your area of expertise, assess their difficulty, provide a detailed solution, and submit them for review.
  • Question Verification: Review existing questions for medical accuracy, clarity, completeness, and rigor; make necessary edits; assess difficulty; and document the rationale for changes.

Requirements

Health & Medicine Domains

Work may span one or more of the following areas:

  • Clinical Medicine & Surgery
  • Medical Imaging & Diagnostics
  • Pharmacovigilance
  • Healthcare Management & Economics
  • Rehabilitation & Allied Health
  • Biomedical and Health Sciences

Key Responsibilities

  • Develop original assessment questions that test deep conceptual understanding, clinical reasoning, and applied knowledge, rather than simple factual recall.
  • Ensure every question is self-contained, unambiguous, technically accurate, and sufficiently defined to support a clear solution.
  • Assign an appropriate difficulty level:
    • Medium: Introductory undergraduate
    • Hard: Advanced undergraduate
    • Expert: Postgraduate level and above
  • Create one correct answer alongside nine plausible but subtly incorrect alternatives designed to meaningfully challenge advanced AI systems.
  • Develop clear, structured solutions that explain the reasoning required to reach the correct answer.
  • Support questions with 1–5 reputable academic references, such as peer-reviewed research, clinical guidelines, authoritative textbooks, or recognized medical sources.
  • For verification assignments, identify issues related to:
    • Medical or scientific accuracy
    • Clarity
    • Completeness
    • Precision
    • Solvability
    • Answer validity
  • Make appropriate corrections and clearly document the reasoning behind significant edits.
  • Maintain a high standard of academic and professional rigor across all assessment content.

Ideal Qualifications

  • MD, DO, PhD, or doctoral candidate in Medicine, Biomedical Sciences, Public Health, or a closely related discipline.
  • A Master's degree may be considered for candidates with exceptional expertise in a specialized health or medical domain.
  • Strong command of advanced medical knowledge, biomedical science, clinical reasoning, and research methodology.
  • Board certification, relevant clinical experience, research publications, or advanced specialization in a health-related field is highly desirable.
  • Ability to critically evaluate medical evidence and distinguish nuanced, technically correct answers from plausible but incorrect alternatives.
  • Excellent written English with the ability to communicate complex medical and scientific concepts clearly and precisely.
  • Strong attention to detail and a rigorous approach to fact-checking and quality assurance.
  • Comfortable working independently on challenging, open-ended academic assessment tasks.

Engagement Details

  • Role: Applied Health & Medicine Benchmark Specialist
  • Work Arrangement: Fully Remote
  • Engagement Type: Independent Contractor
  • Expected Commitment: 10+ hours per week
  • Schedule: Flexible and asynchronous

Contract and Payment Terms

  • Work can be completed remotely on a flexible schedule.
  • Projects may be extended, shortened, or concluded early depending on project requirements and performance.
  • The engagement will not require access to confidential or proprietary information belonging to any employer, client, or institution.
  • Payments are made weekly based on services rendered through available payment platforms.
  • H-1B and STEM OPT candidates are not currently supported.

Equal Employment Opportunity

We consider all qualified applicants without regard to legally protected characteristics and provide reasonable accommodations upon request.

Ready to apply to Weekday AI?
Apply to Weekday AI

About Weekday AI

At Weekday (backed by YC; also Product Hunt #1 product of the day), we are building the next frontier in hiring. We have built the largest database of white collar talent in India and have built outreach tools on top of it to generate highest response rates.

See all jobs at Weekday AI →

Similar jobs

Sign up for suggestions tailored to the jobs you open and the searches you save.

More jobs at Weekday AI

See all jobs at Weekday AI →

Apply now
🤖

Whoa — hold up

JobsRadar was built for real people having a rough time in their job search — not for automated requests. You're clicking way too fast and you're now temporarily blocked.

Come back later. If you're genuinely job hunting, we've got your back — just act like a human.

Catch your next role the second it’s posted.

Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

Create free account

Free forever · takes 30 seconds · already have one?

Get an edge on your job hunt.

Join our Telegram channel for the stuff that helps you land the role — salary benchmarks, the weekly market pulse, and new-feature drops. No spam, just signal.

Join the channel — it's free