Skip to content
SaidGig
Sign up.

Give me your email, I promise I won't do anything weird with it.

Member of Technical Staff, Research Engineering

$220,000–$500,000/yr

RemoteFull-timetechnology
Apply Now

About this role

Role Overview

Build reinforcement learning systems that turn experimental ideas into scalable, high-performance production capabilities. You will develop novel RL environments, training pipelines, and evaluation systems that advance modern AI model performance.

Key Responsibilities
  • Design self-contained RL environments for complex real-world tasks, including reward functions, verifiers, and evaluation logic.
  • Build and scale episode pipelines and multi-component training processes that support reproducible experiments.
  • Create automated synthetic-data generation systems that accelerate training cycles while maintaining quality.
  • Develop AI-driven evaluation and quality-assurance systems for automated grading, validation, and feedback loops.
  • Fine-tune and optimize open-source RL models using internally generated data and custom training strategies.
  • Establish benchmarking frameworks to measure model capability, robustness, and data quality across tasks.
  • Contribute to releasing and analyzing evaluations on internal and external benchmark platforms.
Qualifications
  • Deep reinforcement learning experience, including environment design and training dynamics.
  • Experience building and scaling RL systems, pipelines, or experimentation frameworks.
  • Proficiency with automation, ML-oriented data design, and synthetic-data pipelines.
  • Experience with RL environments, RL workflows, automated evaluation, model validation, and quality-assurance processes.
  • Experience fine-tuning and evaluating open-source machine learning models.
  • Clear technical communication and writing skills.
  • Comfort working in a fast-paced, research-driven, collaborative environment.

Preferred experience includes publishing benchmarks, evaluations, or research artifacts; familiarity with evaluation ecosystems; and scalable infrastructure for large-scale RL experimentation.

Work Terms
  • Full-time, remote position.
Compensation
  • Listed compensation range: $220, 000 to $500, 000 per year.
  • National base-salary range: $140, 000 to $180, 000 USD.
  • Equity compensation is available to all employees, and performance-based bonuses may be offered depending on role and company policy.
  • Benefits include up to 100% reimbursement of health-insurance premiums, paid time off, and a 401(k) plan with company match.
Eligibility

Qualified applicants are considered without regard to race, color, religion, sex, including pregnancy, sexual orientation, or gender identity, national origin, age, disability, genetic information, veteran status, or other legally protected characteristics.

Related Jobs

More like this