Biochemistry and Life Sciences Researcher for AI Model Evaluation
$65–$105/hr
About this role
Help improve how frontier AI models reason about real-world life sciences research. In this senior, hands-on role, you will apply deep scientific judgment to evaluate research tasks and model outputs, define rigorous standards for correct answers, and build benchmarks that measure meaningful model improvement.
Key Responsibilities- Review life sciences knowledge-work tasks and model outputs, identifying missing behaviors, weak reasoning, mechanistic oversimplification, unsupported causal claims, and responses that would not withstand peer review.
- Write detailed instruction specifications, create gold-standard solutions to life sciences problems, and define tasks that reflect real research practice.
- Design challenging domain-specific evaluation sets and contribute to life sciences skills and tools with the research team.
- Partner with researchers and adjacent-domain specialists to calibrate consistent quality standards and translate expert scientific judgment into explicit, teachable criteria.
- PhD in biochemistry, molecular or cell biology, genetics, immunology, neuroscience, or a closely related life science field.
- At least 4 years of substantive research experience at a recognized research university, academic medical center, national laboratory, or industrial research organization. Graduate coursework alone does not qualify.
- Deep specialization in at least one area, such as molecular and cell biology, biochemistry and structural biology, immunology and immuno-oncology, genetics and genomics, neuroscience, or systems biology.
- Demonstrated senior-level progression, such as Senior Scientist, Staff Scientist, Research Scientist, Instructor, Principal Investigator, or faculty appointment, with ownership of research direction.
- Peer-reviewed publication record, preferably including first-author publications in strong journals.
- Practical experience using large language models in professional work and the ability to distinguish sound reasoning from plausible but incorrect responses.
- Excellent written communication skills and the ability to provide precise, well-structured feedback.
- Full-time W-2 employment, with placement on an AI lab''s extended workforce and close collaboration with its internal research teams.
- Initial commitment of 6 months at 40 hours per week.
- Hybrid role based in the Bay Area, California, with on-site work alongside the client team multiple days per week when required.
- This is not a remote role. Candidates who do not currently live in the Bay Area must relocate at their own expense before the engagement begins. Relocation assistance is not available.
- Client-issued accounts and equipment will be provided, and work will be performed in the client''s tools and workflows.
$65 to $105 per hour.
EligibilityEmployment, payroll, benefits, compliance, and onboarding are administered through the employer of record. Qualified applicants are considered without regard to legally protected characteristics. Reasonable accommodations are available for qualified individuals with disabilities and disabled veterans throughout the application process.