This contract role involves evaluating AI-generated content within the Finance domain to improve the accuracy and rigor of model outputs. The position is remote, requires 15–20 hours per week, and offers compensation of $50 per hour.
Responsibilities
Evaluate the accuracy and depth of AI-generated Finance content to enhance reasoning and rigor.
Create complex, well-defined tasks with clear ground-truth outputs and objective evaluation rubrics.
Collaborate with subject matter experts to ensure dataset consistency, relevance, and comprehensive coverage.
Work independently and asynchronously to meet deadlines and contribute to AI model performance improvements.