About this role
Join a remote talent network supporting potential future projects that assess how effectively AI systems handle real-world data science work.
Role OverviewThere is no immediate project opening. Qualified candidates may be contacted as relevant opportunities become available, and project responsibilities will vary by assignment.
Key Responsibilities- Design precise, task-specific grading criteria for data science deliverables, including exploratory data analysis, statistical modeling, machine learning pipelines, experimentation and A/B test write-ups, feature engineering, and technical reports or notebooks.
- Evaluate AI-generated or human-created work against established criteria.
- Provide detailed written rationales for evaluations and scores.
- Apply consistent, evidence-based judgment to produce reproducible, defensible assessments.
- Incorporate structured feedback from senior reviewers and revise submitted work accordingly.
- At least 1 year of professional data science experience.
- Experience at a leading technology, research, or quantitative firm, such as a top FAANG company, AI lab, top-tier quantitative fund, or equivalent organization.
- Strong command of Python, SQL, statistical modeling, machine learning, experimentation and causal inference, and turning messy real-world data into rigorous analyses.
- Exceptional written communication skills and the ability to explain technical findings clearly.
- A detail-oriented, consistent approach to evaluating complex work.
- Comfort receiving feedback and calibrating judgment to established standards.
- Remote, hourly engagement.
- $100 to $150 per hour.