We are looking for Senior Software Engineer – LLM Evaluation candidates for a project delivered through Turing.
What you'll do
- Create cutting-edge datasets for training, benchmarking, and advancing large language models.
- Curate code examples, provide precise solutions, and make corrections in Python, JavaScript (including ReactJS), C/C++, Java, Rust, and Go.
- Evaluate and refine AI-generated code for efficiency, scalability, and reliability.
- Work with cross-functional teams to enhance enterprise-level AI-driven coding solutions.
- Curate code examples, build solutions, and correct code in Python, JavaScript (including ReactJS), C/C++, Java, Rust, and Go.
- Evaluate and refine AI-generated code to ensure that it is efficient, scalable, and reliable.
- Collaborate with cross-functional teams to enhance AI-driven coding solutions against industry performance benchmarks.
- Build agents that can verify the quality of the code and identify error patterns.
- Design verification mechanisms that can automatically verify a solution to a software engineering task.
What you need
- Several years of software engineering experience, including 2+ years of continuous full-time experience at a top-tier product company.
- Strong expertise in building full-stack applications and deploying scalable, production-grade software using modern languages and tools.
- Deep understanding of software architecture, design, development, debugging, and code quality/review assessment.
- Excellent oral and written communication skills for clear, structured evaluation rationales.
Expertise
Who you work with
Project and contracting process: Turing. Applications continue on the provider's website.

