About Turing
Turing is one of the world's leading AGI infrastructure companies, working with frontier AI labs to accelerate model development through high-quality training data, evaluations, and engineering talent. We are staffing a frontier AI data initiative building the infrastructure and training data used to develop and evaluate AI agents. You will be assigned to one of two tracks: connectors or tasks.
About the Role
We are looking for experienced Python Engineers to work with leading AI research labs on evaluating and improving next-generation AI coding agents. You will review agentic coding trajectories, assess technical correctness, debug and validate outputs, and use coding agents such as Claude Code, Codex, or Cursor as part of your workflow.
What You'll Do
- Review agentic coding trajectories, including prompts, tool calls, intermediate actions, code changes, execution results, and final outputs.
- Evaluate whether AI agents correctly understand and execute software engineering tasks.
- Analyze agent behavior to identify technical errors, inefficient approaches, and opportunities for improvement.
- Read, debug, test, and validate Python code produced or modified by AI agents.
- Work with LLM-powered agents, workflows, and development environments to evaluate model capabilities.
- Use coding agents such as Claude Code, Codex, Cursor, or similar tools to accelerate analysis, debugging, and validation.
Requirements
- 5+ years of professional software engineering experience, with strong hands-on expertise in Python.
- At least 6 months to 1 year of practical AI/LLM engineering experience building agents, agent loops, LLM-powered applications, data pipelines, or similar systems.
- Hands-on experience using AI coding agents such as Claude Code, Codex, Cursor, or equivalent tools as part of regular software development workflows.
- Ability to understand and evaluate agentic workflows, including tool usage, intermediate execution steps, code modifications, and resulting outputs.
Perks of Working With Turing
- Work with leading AI research labs and contribute to the development of cutting-edge AI systems.
- Fully remote opportunity.
- Collaborate with a global network of talented engineers and AI professionals.
- Flexible engagement with exposure to frontier AI development and evaluation.
Offer Details
- Engagement: Full-time contractual opportunity
- Location: Remote
- Availability: 40 hours per week with required overlap with the global team
- Contract Duration: 2 months (with possible extensions based on project requirements)
Evaluation process
- Around 25 mins of AI interview