Medical Expert - AI Model Evaluation
$30k $80k • No equity
Get paid to help build better AI.
Alpheva AI, the Silicon Valley company that is building intelligence behind frontier AI, is looking for experienced medical professionals to participate in paid, project-based work that helps train and evaluate the next generation of AI systems.
You work remotely and asynchronously on projects that fit around your existing schedule.
Your medical expertise becomes the signal that helps AI systems learn what accurate, evidence-based, and professionally sound medical reasoning actually looks like.
IMPORTANT: To be considered, Apply directly through: alpheva{dot}ai/experts/
What you'll do
Depending on the project, you may:
Analyze: Review clinical scenarios, patient cases, medical information, diagnostic questions, and treatment options.
Evaluate: Review AI-generated medical answers, clinical reasoning, diagnoses, recommendations, and explanations.
Validate: Determine whether AI-generated medical outputs are accurate, clinically sound, appropriately supported, and consistent with accepted medical knowledge.
Review: Assess clinical reasoning, differential diagnoses, treatment approaches, medical documentation, and patient scenarios.
Reason: Solve realistic medical problems and explain how an experienced medical professional would approach them.
Compare solutions: Assess multiple AI-generated medical solutions and identify meaningful differences in accuracy, reasoning, safety, completeness, and professional judgment.
Create evaluations: Develop challenging medical tasks and criteria for measuring AI systems.
Provide expert judgment: Identify errors, unsupported conclusions, missing considerations, unsafe recommendations, and weaknesses in AI-generated medical work.
Projects may involve clinical medicine, diagnostics, treatment planning, pharmacology, medical research, patient care, medical documentation, public health, and other areas depending on your expertise.
Who we're looking for
We're looking for experienced medical professionals, not necessarily AI researchers.
You should have:
3+ years of professional experience in medicine, healthcare, clinical practice, medical research, or a related field
Strong medical and analytical reasoning skills
Experience working with real-world medical information
Ability to evaluate medical analysis and conclusions critically
Strong attention to detail
Ability to identify errors, inconsistencies, unsupported claims, and missing clinical considerations
Strong professional judgment
Ability to clearly explain why a medical conclusion is correct or incorrect
Experience with any of the following is valuable:
Clinical medicine
Internal medicine
Emergency medicine
Surgery
Pediatrics
Psychiatry
Radiology
Pathology
Pharmacology
Medical research
Public health
Nursing or allied health professions
MD, DO, MBBS, RN, NP, PA, PharmD, or equivalent professional qualification
You do not need prior experience working in AI.
What matters most is that you know how to do medical work well.
You might be asked to:
Review an AI-generated medical answer and identify the issues.
Evaluate whether an AI-generated diagnosis is supported by the available clinical information.
Assess whether a treatment recommendation is medically appropriate and identify potential concerns.
Compare two AI-generated clinical analyses and determine which is stronger.
Review a clinical scenario and evaluate the quality of the AI's reasoning.
Identify incorrect medical claims, missing considerations, unsupported conclusions, or potentially unsafe recommendations.
Analyze a realistic medical problem and explain your reasoning.
Evaluate whether an AI correctly interprets medical information, symptoms, laboratory results, or clinical findings.
Create challenging medical problems that can be used to evaluate AI systems.
The goal is not simply to determine whether medical analysis looks reasonable.
The goal is to capture the judgment an experienced medical professional uses to determine whether medical reasoning actually makes sense.
Compensation
Paid project-based work.
Rates vary based on your experience, specialization, and project requirements. Your applicable rate and project scope will be shown before you commit to a project.
Some projects are short assignments. Others may involve recurring work over a longer period.
There is no fixed schedule and no minimum commitment.
Work arrangement
Remote
Asynchronous
Project-based
Flexible hours
Work around your existing job or commitments
Paid for qualifying project work
But producing a plausible-looking medical answer is not the same as demonstrating strong medical judgment.
AI needs to learn how experienced medical professionals:
Analyze clinical information
Interpret symptoms and medical findings
Evaluate diagnostic possibilities
Assess treatment options
Identify risks and contraindications
Evaluate clinical reasoning
Recognize important edge cases
Distinguish strong evidence from weak or unsupported claims
Identify potentially unsafe recommendations
Explain medical conclusions clearly
Determine when medical analysis is actually reliable
That's where you come in.
Your medical expertise helps shape how AI learns to reason about medicine.
How it works
Apply
Tell us about your medical background, experience, qualifications, and areas of expertise.
Get qualified
Complete a short assessment designed around the type of medical work you'll perform.
Get matched
When your experience matches an active project, we'll share the scope, requirements, and compensation.
Do the work
Complete projects remotely and asynchronously.
Get paid
Receive the agreed rate for your completed work.
Help teach AI how great medical professionals actually work.
Apply directly through: alpheva{dot}ai/experts/
Medical Expert - AI Model Evaluation in San Francisco at Alpheva AI
This position is listed as full time and able to be worked remotely. It was posted 2 days ago.