Role OverviewAs a Software Engineering Evaluator, you'll create high-quality datasets, review AI-generated code, and assess software engineering workflows across multiple programming languages. Your work will directly contribute to training state-of-the-art AI coding models.Key ResponsibilitiesEvaluate and improve AI-generated code for quality, scalability, and performanceDevelop coding examples and benchmark datasets for LLM trainingReview and correct code in: Python, JavaScript (React.js), C/C++, Java, Rust, GoDesign automated code verification systemsAnalyze software engineering workflows including: Architecture Design, API Development, Prototyping, Production Deployment, Testing, Monitoring, MaintenanceBuild AI agents capable of validating software quality and identifying coding errorsCollaborate with AI researchers and engineering teams to improve coding model performanceRequired QualificationsSeveral years of professional software engineering experienceMinimum 2+ years of continuous full-time experience at a top-tier technology company such as: Google, Microsoft, Amazon, Meta, Apple, Stripe, Netflix, Shopify, Dropbox, Datadog, PayPal, IBM ResearchStrong Full-Stack development expertiseDeep understanding of: Software Architecture, Code Review, System Design, API Development, Debugging, Production SystemsExcellent written and verbal communication skillsWhy Join Turing?Competitive contractor compensation100% RemoteWork with leading AI research labsFlexible schedule (10–40 hours/week)Turing is hiring Senior Software Engineers to evaluate, improve, and benchmark next-generation Large Language Models (LLMs). Work alongside leading AI researchers while helping build the future of AI-powered software development.