Applied Health & Medicine Benchmark Specialist at Weekday AI — NeverHard
Applied Health & Medicine Benchmark Specialist at Weekday AI in US. Skills: AI, Health Sciences, Medical Assessment. Apply on NeverHard.
Company
Weekday AI
Location
US
Type
part_time
Required skills:
AI
Health Sciences
Medical Assessment
This role is for one of our clients
Compensation: $94 - $119 per hour
We are seeking highly qualified
medical and health science professionals
to contribute to an AI research initiative focused on developing rigorous academic assessment content across health and medicine.
As an
Applied Health & Medicine Benchmark Specialist
, you will author and review challenging multiple-choice questions, evaluate the quality and validity of existing assessment content, and help establish high-quality benchmarks for measuring AI performance on advanced medical and health science concepts.
You may be assigned one of two primary task types:
Question Authoring:
Develop original, challenging multiple-choice questions within your area of expertise, assess their difficulty, provide a detailed solution, and submit them for review.
Question Verification:
Review existing questions for medical accuracy, clarity, completeness, and rigor; make necessary edits; assess difficulty; and document the rationale for changes.
Requirements
Health & Medicine Domains
Work may span one or more of the following areas:
Clinical Medicine & Surgery
Medical Imaging & Diagnostics
Pharmacovigilance
Healthcare Management & Economics
Rehabilitation & Allied Health
Biomedical and Health Sciences
Key Responsibilities
Develop original assessment questions that test
deep conceptual understanding, clinical reasoning, and applied knowledge
, rather than simple factual recall.
Ensure every question is self-contained, unambiguous, technically accurate, and sufficiently defined to support a clear solution.
Assign an appropriate difficulty level:
Medium:
Introductory undergraduate
Hard:
Advanced undergraduate
Expert:
Postgraduate level and above
Create one correct answer alongside
nine plausible but subtly incorrect alternatives
designed to meaningfully challenge advanced AI systems. Develop clear, structured solutions that explain the reasoning required to reach the correct answer. Support questions with
1–5 reputable academic references
, such as peer-reviewed research, clinical guidelines, authoritative textbooks, or recognized medical sources. For verification assignments, identify issues related to:
Medical or scientific accuracy
Clarity
Completeness
Precision
Solvability
Answer validity
Make appropriate corrections and clearly document the reasoning behind significant edits. Maintain a high standard of academic and professional rigor across all assessment content. Ideal Qualifications
MD, DO, PhD, or doctoral candidate
in Medicine, Biomedical Sciences, Public Health, or a closely related discipline.
A
Master's degree
may be considered for candidates with exceptional expertise in a specialized health or medical domain.
Strong command of advanced medical knowledge, biomedical science, clinical reasoning, and research methodology.
Board certification, relevant clinical experience, research publications, or advanced specialization in a health-related field is highly desirable.
Ability to critically evaluate medical evidence and distinguish nuanced, technically correct answers from plausible but incorrect alternatives.
Excellent written English with the ability to communicate complex medical and scientific concepts clearly and precisely.
Strong attention to detail and a rigorous approach to fact-checking and quality assurance.
Comfortable working independently on challenging, open-ended academic assessment tasks.
Engagement Details
Role:
Applied Health & Medicine Benchmark Specialist
Work Arrangement:
Fully Remote
Engagement Type:
Independent Contractor
Expected Commitment:
10+ hours per week
Schedule:
Flexible and asynchronous
Contract and Payment Terms
Work can be completed remotely on a flexible schedule.
Projects may be extended, shortened, or concluded early depending on project requirements and performance.
The engagement will not require access to confidential or proprietary information belonging to any employer, client, or institution.
Payments are made weekly based on services rendered through available payment platforms.
H-1B and STEM OPT candidates are not currently supported.
Equal Employment Opportunity
We consider all qualified applicants without regard to legally protected characteristics and provide reasonable accommodations upon request.