Senior Software Engineer (Python) for LLM Evaluation & Repository Validation at turing | Torre

Senior Software Engineer (Python) for LLM Evaluation & Repository Validation

Emma highlights
This highlight was written by Emma’s AI. Ask Emma to edit it.
Freelance
Recurrent
Provide your expected compensation while applying
location_on
Remote (for Kenya residents)
Remote (for India residents)
Remote (for Pakistan residents)
Remote (for Nigeria residents)
Shared by
Daniel Correa Laverde
1 day ago

Responsibilities


Key ResponsibilitiesAnalyze and triage GitHub issues across open-source repositories.Set up repositories, Docker environments, and development pipelines.Evaluate unit-test coverage and quality.Modify and run real codebases to assess LLM bug-fixing performance.Identify challenging repositories and tasks for AI evaluation.Collaborate with researchers and potentially lead junior engineers.Requirements3+ years of professional experience.Strong Python skills.Experience with Git, Docker, testing, and software pipelines.Ability to understand and navigate complex codebases.Open-source or previous LLM evaluation experience is a plus.Selection: Two technical interview rounds, approximately 75 minutes total.Turing is hiring experienced Python software engineers to help build and evaluate datasets that teach AI models to solve realistic software engineering problems.Contract: 3 months. Remote: Kenya, India, Pakistan, Nigeria, Egypt, Ghana, Bangladesh, Turkey & Mexico. 20–40 hrs/week.