About this role
Join a vetted pool of experienced data scientists who will be contacted for future, project-by-project evaluations of AI systems performing real-world data science tasks. Members of this talent network design evaluation criteria and judge technical deliverables to ensure assessments are rigorous, reproducible, and well-documented.
Key Responsibilities- Design task-specific, precise grading criteria for data science outputs, including exploratory data analyses, statistical models, machine learning pipelines, experimentation and A/B test write-ups, feature engineering, and technical reports or notebooks
- Evaluate AI-generated and human-created work against established criteria
- Provide detailed written justifications for evaluations and scores
- Apply consistent, evidence-based judgment so assessments are reproducible and defensible
- Incorporate structured feedback from senior reviewers and iterate on submitted evaluations
- Recognize that specific responsibilities will vary by project
- 1+ years of professional data science experience
- Experience at a leading technology, research, or quantitative firm, for example top FAANG companies, AI labs, or top-tier quantitative funds, or equivalent experience
- Strong command of Python and SQL, statistical modeling, machine learning, experimentation and causal inference
- Proven ability to translate messy, real-world data into rigorous analyses
- Exceptional written communication skills, able to convey technical findings clearly
- Detail-oriented, consistent approach to evaluating complex work
- Comfort receiving feedback and calibrating judgment against established standards
Remote, hourly engagements for project-based assignments as opportunities arise. There is no immediate project start date guaranteed.
CompensationPay range $100 to $150 hourly.
EligibilitySubmitting an application adds you to the talent network. Qualified applicants may be contacted when relevant projects become available, placement depends on project availability and need.