You will advance AI reasoning capabilities by evaluating complex technical model outputs.
Emma highlights
This highlight was written by Emma’s AI. Ask Emma to edit it.
Freelance
Recurrent
Compensation
USD50-150/hour
The 'Compensation' is required.
location_on
Remote (for United States residents)
Remote (for Canada residents)
Remote (for United Kingdom residents)
Remote (for Australia residents)
The 'Location' is required.
Shared by
about 2 months ago
Responsibilities
Key ResponsibilitiesEvaluate AI-generated technical content for accuracy and factual correctnessCreate and answer advanced STEM and AI-related questionsReview, compare, and rank AI model responsesIdentify reasoning errors and improve AI model performanceContribute to cutting-edge AI research and evaluation projectsRequired QualificationsPhD (completed or near completion) in: Machine Learning / Artificial Intelligence, Computer Science, Engineering, Statistics, Mathematics, Physics, or a closely related quantitative STEM fieldStrong analytical thinking and problem-solving skillsExcellent written English communicationAbility to evaluate complex technical reasoningPreferred ExperienceAI/ML research or industry experienceData annotation or AI model evaluationResearch publications or peer review experienceExperience with Large Language Models (LLMs)Why Apply?Earn up to $150/hour100% Remote & Flexible Work ScheduleContribute to cutting-edge AI and LLM researchCollaborate with leading AI labs and researchersWork from anywhere within eligible countries