Finance Subject-Matter Expert for AI Model Evaluation
$65–$90/hr
About this role
Bring deep finance domain expertise to a GenAI training team by creating realistic finance tasks, producing professional solutions, and evaluating Large Language Model outputs against structured rubrics. This role strengthens model accuracy and judgment in real-world financial scenarios while working closely with research and engineering teams.
Key Responsibilities- Advise research and engineering teams to close gaps in financial reasoning, analysis, and decision-making.
- Design challenging, domain-relevant finance tasks and produce accurate, well-reasoned solutions based on real financial practice.
- Evaluate AI model outputs against structured rubrics, providing clear written feedback on correctness, judgment, and reasoning quality.
- Develop and refine finance-specific evaluation guidelines and scoring rubrics.
- Collaborate with other subject matter experts to ensure consistency and accuracy across training data and evaluations.
- Minimum 8 years of dedicated professional finance experience, for example in investment banking, asset management, corporate finance, or financial advisory, at a recognized, top-tier organization such as Goldman Sachs, JPMorgan, Morgan Stanley, BlackRock, Fidelity, Deloitte, PwC, EY, KPMG, or equivalent.
- Prior hands-on experience evaluating LLM or AI model outputs against rubrics or structured scoring criteria is mandatory, please describe this experience in your application.
- Demonstrable career progression, for example Analyst to Associate to VP or Director.
- Strong verbal and written communication skills, solid problem-solving ability, and effective interpersonal skills.
- Ability to engage reliably for at least 35 hours per week during weekdays.
- Employment type: Hourly, W-2 employment through Cincinnatus LLC, which acts as the employer of record and provides payroll, benefits, and compliance support.
- Placement: Employees are placed directly within a leading AI laboratory as part of the client team to work on GenAI model development and evaluation.
- Schedule commitment: Minimum 35 hours per week on weekdays, reliable weekday availability required.
- Pay rate: 65 to 90 hourly.
- Location: United States based role, candidates must be eligible to work in the United States under W-2 employment with Cincinnatus LLC.
- This role is offered as contingent/contract employment through the employer of record arrangement with Cincinnatus LLC.
When you apply, include a description of your prior experience evaluating LLM or AI model outputs against rubrics or structured scoring criteria. Applications will be reviewed for finance domain expertise, demonstrated evaluation experience, and ability to meet the weekly weekday hours commitment.
Equal Employment OpportunityCincinnatus LLC is an equal opportunity employer. Employment decisions are made without regard to race, religion, color, national origin, sex, sexual orientation, gender identity, age, veteran status, disability, genetic information, political views, or any other legally protected characteristic.