About this role
Role Overview
Apply senior platform engineering expertise to create real-world technical environments that help evaluate and improve next-generation AI systems. You will design production-grade cloud infrastructure scenarios that assess an AI model''s ability to deploy, troubleshoot, secure, scale, and recover distributed systems. AI experience is not required, hands-on production expertise is the priority.
Key Responsibilities
- Create realistic cloud infrastructure tasks covering distributed systems, networking, security, scalability, and reliability.
- Develop scenarios involving IAM, queues, durable storage, observability, rolling deployments, and disaster recovery.
- Build reproducible, containerized environments with reference solutions and intentionally defective variants.
- Define measurable requirements for infrastructure configuration, deployed topology, and runtime behavior.
- Develop deterministic integration, load, security, failure-injection, deployment, and recovery tests.
- Debug environments, document technical decisions, and review work created by other experts.
Qualifications
- Senior-level experience in technical architecture, cloud infrastructure, platform engineering, DevOps, systems engineering, or SRE, including ownership of a production platform.
- Strong knowledge of distributed systems, scalable APIs, queues, autoscaling, durable storage, and partial-failure scenarios.
- Practical experience with IAM, private networking, least-privilege access, and service-to-service security.
- Experience with observability, measurable SLOs, rolling deployments, rollback strategies, and disaster recovery.
- Ability to write infrastructure automation or testing tools and debug containerized environments using a relevant programming language.
- Experience with Terraform or OpenTofu, AWS, Azure, GCP, Kubernetes, multi-cloud infrastructure, internal developer platforms, edge infrastructure, shared platform services, chaos engineering, fault injection, local cloud emulators, or resilience testing is preferred.
- Experience creating technical evaluations, automated grading systems, or AI environments is helpful but not required.
Work Terms
- Remote independent contractor engagement.
- Minimum weekly task-submission requirements apply.
- Selected candidates are expected to begin their first tasks within 24 to 48 hours after completing onboarding.
Compensation
Compensation is listed at $60 to $130 per hour. Payment is output-based and made per task that meets project specifications. Time to complete work varies based on experience and workflow.
Application Process
- Submit an application and complete the screening questions.
- Complete an approximately 30-minute AI interview reviewed by recruiters.
- Proceed to hiring-manager review.