Senior Software Engineer – LLM Evaluation & Repository Validation
Turing Jobs Feed · Worldwide ·
- Work mode
- Remote
- Seniority
- Senior
- Employment
- Contract
- Category
- AI engineering
Turing Jobs Feed · Worldwide ·
Streem Energy · Paris, IDF, fr
firstam · USA, California, Santa Ana
Doctolib · Paris, Paris, France
Thomson Reuters · United States of America, Eagan, Minnesota
Build and evaluate datasets of verifiable software-engineering tasks from public GitHub repos to train LLMs on realistic coding problems, including environment setup, issue triage, and test-coverage analysis.
Job Title: Senior Software Engineer – LLM Evaluation & Repository Validation
About the projects: we are building LLM evaluation and training datasets to train LLM to work on realistic software engineering problems. One of our approaches, in this project, is to build verifiable SWE tasks based on public repository histories in a synthetic approach with human-in-the-loop; while expanding the dataset coverage to different types of tasks in terms of programming language, difficulty level, and etc.
About the Role: We are looking for experienced software engineers (tech lead level) who are familiar with high-quality public GitHub repositories and can contribute to this project. You should have experience working with well-maintained, widely-used repos with 500+ stars. This role involves hands-on software engineering work, including development environment automation, issue triaging, and evaluating test coverage and quality
Why Join Us?
Turing is one of the world’s fastest-growing AI companies accelerating the advancement and deployment of powerful AI systems. You’ll be at the forefront of evaluating how LLMs interact with real code, influencing the future of AI-assisted software development. This is a unique opportunity to blend practical software engineering with AI research.
What does day-to-day look like:
Required Skills:
Nice to Have:
Perks of Freelancing With Turing:
Offer details:
Thomson Reuters · Toronto, Ontario, Canada