You will advance AI reasoning capabilities by evaluating complex technical model outputs.
Emma highlights
This highlight was written by Emma’s AI. Ask Emma to edit it.
Freelance
Recurrent
Compensation
USD50-150/hour
The 'Compensation' is required.
location_on
Remote (for United States residents)
Remote (for Canada residents)
Remote (for United Kingdom residents)
Remote (for Australia residents)
The 'Location' is required.
Shared by
18 days ago
Responsibilities
Key ResponsibilitiesEvaluate AI-generated technical content for accuracy and factual correctnessCreate and answer advanced STEM and AI-related questionsReview, compare, and rank AI model responsesIdentify reasoning errors and improve AI model performanceContribute to cutting-edge AI research and evaluation projectsRequired QualificationsPhD (completed or near completion) in: Machine Learning / Artificial Intelligence, Computer Science, Engineering, Statistics, Mathematics, Physics, or a closely related quantitative STEM fieldStrong analytical thinking and problem-solving skillsExcellent written English communicationAbility to evaluate complex technical reasoningPreferred ExperienceAI/ML research or industry experienceData annotation or AI model evaluationResearch publications or peer review experienceExperience with Large Language Models (LLMs)Why Apply?Earn up to $150/hour100% Remote & Flexible Work ScheduleContribute to cutting-edge AI and LLM researchCollaborate with leading AI labs and researchersWork from anywhere within eligible countries