About the Company
We are seeking experienced Data Science professionals to join a cutting-edge AI evaluation project as an AI Coder / AI Response Evaluator. This is a unique opportunity to contribute to the development of next-generation Artificial Intelligence systems by applying your professional expertise to assess, validate, and improve AI-generated responses within your area of specialization. This assignment is entirely remote, with flexible and self-directed working hours. There is no patient care, no shift work, and no requirement to be available at specific times during the day. Successful candidates will work independently while contributing to the training and improvement of advanced AI models.
About the Role
We are seeking experienced Data Science professionals to join a cutting-edge AI evaluation project as an AI Coder / AI Response Evaluator.
Responsibilities
- Review, assess, and evaluate AI-generated responses within data science and related technical domains.
- Analyze the accuracy, quality, completeness, and relevance of AI-produced outputs.
- Compare multiple AI-generated solutions and identify the strongest response based on technical merit.
- Evaluate Python, SQL, analytics, machine learning, statistical, and data engineering-related content.
- Provide clear written feedback and justification for evaluation decisions.
- Identify factual inaccuracies, logical errors, coding issues, and opportunities for improvement.
- Apply professional expertise to ensure responses align with real-world industry standards and best practices.
- Work independently while maintaining quality and productivity expectations.
- Participate in onboarding and qualification activities, including an initial paid assessment.
Qualifications
- Extensive Data Science experience in analytics, machine learning, statistical modeling, data engineering, business intelligence, artificial intelligence, or related applied disciplines.
- Undergraduate study does not count toward the minimum experience requirement.
- Working proficiency in Python and/or SQL, with the ability to read, understand, and evaluate technical code.
- Availability to commit 30-40 hours per week throughout the full 12-week assignment.
- Strong written communication skills and attention to detail.
- Ability to work independently in a fully remote environment.
- General familiarity with AI, Large Language Models (LLMs), or generative AI tools at a user level.
Preferred Skills
- Advanced degree (Master's or PhD) in Data Science, Computer Science, Statistics, Mathematics, Engineering, Economics, Physics, or another quantitative discipline.
- Experience reviewing technical work, conducting quality assessments, or evaluating analytical outputs.
- Professional experience across multiple data domains, including machine learning, predictive analytics, experimentation, statistical inference, or data engineering.
- Exposure to generative AI technologies and AI-assisted coding tools.
Pay range and compensation package
Successful candidates will be onboarded to the project and compensated for all approved hours worked throughout the assignment.
Equal Opportunity Statement
To ensure full transparency, this assignment does not involve patient care or clinical responsibilities, does not require shift work, on-call coverage, or weekend commitments, does not involve sales, business development, or customer support activities, and is not a traditional software engineering role focused on product development. It focuses exclusively on evaluating and improving AI-generated content and technical outputs.