Turing

Senior Research Engineer

Turing$250K — $350K *
Information Technology
Less than 5 years of experience
Job Overview by Ladders

Qualifications

  • 5-7 years of hands-on experience in post-training tasks or building coding agents.
  • Proven engineering management experience with a focus on QA processes.
  • Strong fluency in Python, plus proficiency in C++, Java, Go, Rust, or JavaScript.
  • Demonstrated operational leadership managing complex data pipelines and multi-stakeholder projects.
  • Exceptional cross-functional communication skills with technical and non-technical stakeholders.
  • Educational background in Computer Science, Machine Learning, or a related field preferred.

Responsibilities

  • Lead creation of datasets and RL environments specifically for coding agents in AI labs.
  • Ensure all deliverables meet high standards of realism, correctness, and diversity.
  • Design and implement automated quality assurance processes for data generation.
  • Build and mentor cross-functional teams utilizing Turing's extensive developer network.
  • Act as a primary point of contact for clients, aligning on project goals and gathering feedback.
  • Conduct model evaluations and publish findings to enhance data offerings and engage industry.
  • Oversee development of tools that streamline data verification and generation.

Benefits

  • Collaborate directly with leading AI labs on cutting-edge research.
  • Influence the development of Artificial General Intelligence with tangible impact.
  • Contribute to a productivity revolution in software engineering affecting global GDP.
  • Be part of a talented team with high autonomy and opportunities for rapid learning.
Full Job Description
The Role

Turing builds large-scale datasets and reinforcement learning (RL) environments that power post-training for the world's leading AI labs and enterprises. We create RL environments to evaluate and improve our customers' models on complex, long-range, multi-step workflows across high-GDP-value domains such as Finance, Sales, Retail, Developer Tools, Collaboration, Customer Experience.

The environments vary depending on the model capability being evaluated / improved, a few examples of environment types are listed here:
  1. Environments for Software Engineering / coding agents
  2. UI-Environments for Computer-Use/Browser-Use agents
  3. MCP-based Environments for general function-calling agents across various enterprise and consumer applications

The Senior Research Engineer will own end-to-end the creation of datasets, RL environments, and evals for frontier AI labs in the domain of coding agents and software engineering. This is a hands-on technical leadership role where you influence revenue directly - you will be mapped to one or more AI labs and interface directly with researchers / engineers at those labs to understand their needs and build data offerings to address those needs. To achieve this, you will build and manage teams of software engineers, researchers, QAs, and contractors/data-annotators from Turing's talent pool of 4M+ developers.

You'll be responsible for delivering projects at frontier quality and scale-owning data quality, throughput, and timely delivery. You'll define and manage data pipelines, validation workflows, and review processes to ensure datasets meet the highest standards for realism, correctness, and diversity. You'll also develop automations, synthetic data generation systems, and internal tools to scale production efficiently.

In short, you'll run your project like a startup within Turing, owning both the technical architecture and the operational execution required to produce best-in-class datasets/environments/evals to make the world's best coding agents and models even better at real-world coding tasks across the software development lifecycle.
What You'll Do

1. End-to-End Ownership: Data Quality, Process Design, and Team Building
  • Lead the creation of datasets, rl environments, and evals focused on Coding Agents / Software Engineering for one or more AI lab customers.
  • Ensure that everything you ship to clients meets frontier standards for realism, correctness, diversity, and difficulty.
  • Set up quality rubrics, automated validation scripts, and human review processes for every stage of data generation.
  • Build and lead cross-functional teams of software engineers, researchers, QAs, and data creators drawn from Turing's 4M+ developer network.
  • Interview, onboard, train, and mentor team members to ensure consistent output quality and technical excellence.

2. Collaborate with Researchers at Frontier Labs
  • Act as the primary technical point of contact for your customer projects, interfacing directly with researchers and engineers at frontier AI labs to understand their coding agent roadmap and model data needs, to gather feedback, and to co-define success criteria for your projects.
  • Provide regular progress updates, surface insights from model evaluations, and incorporate client feedback to improve future iterations.

3. Drive Research, Sales Enablement, and Industry Thought Leadership
  • Fine-tune models in-house on Turing-generated datasets or Turing-rl-environment generated trajectories to determine model improvement as a proof of data quality
  • Proactively build benchmarks and run evals on frontier models and coding agents to identify strengths and weaknesses on SWE tasks, and leverage these insights to inform product roadmap
  • Equip customer-facing teams with the Evaluation reports, sample datasets, and trainings to enable them to communicate your data offerings to customers most effectively
  • Publish research papers and technical posts on Turing's data products, innovations in our synthetic data generation / automation pipelines, evaluations of frontier agents and models, and Turing's model fine-tuning results on our datasets.

4. Build Tools and Infrastructure
  • Oversee development of internal tools that accelerate data generation and verification (e.g., automated data scraping pipelines, unit test generators, repo sandboxing).
  • Design dashboards and APIs for customers to run model evals, view performance reports, and integrate Turing data directly into their post-training pipelines.
What We're Looking For
  • Post-training experience on SWE tasks or experience building coding agents: We expect that you have a deep understanding of data ingredients and design principles that lead to measurable coding model improvements, either from fine-tuning models to improve SWE capabilities or building your own coding agents to improve upon SWE capabilities of the base model.
  • Engineering Management experience: have led teams of engineers in the past, including interviewing/hiring them and setting up QA processes.Hands-on technical capability: Fluency in Python and proficiency in one or more major languages (C++, Java, Go, Rust, or JS).
  • Operational leadership: Proven ability to manage complex data pipelines, multi-stakeholder delivery, and concurrent high-stakes projects.
  • Cross-functional communicator: ability to communicate clearly with researchers at frontier AI labs, subject matter experts for various domains, and diverse teams.
  • Background in Computer Science, Machine Learning, or related technical field preferred.
Why Turing
  • Work directly with the world's leading AI labs and enterprises at the cutting edge of post-training and RL environment design.
  • Real impact (path to AGI): your datasets and environments will directly influence the trajectory toward Artificial General Intelligence and, ultimately, Superintelligence. Coding is the core reasoning substrate of intelligence-advancing models' ability to understand, design, and write code is effectively advancing their capacity for logic, planning, and abstract thought.
  • Real Impact (GDP): automating software engineering unlocks one of the largest productivity frontiers in history. The software engineering market represents trillions in global GDP, and every percentage gain in automation translates to profound efficiency and innovation benefits across all industries.
  • Talent-dense team, where you'll find high autonomy, rapid iteration, and an exceptional learning curve.

This role is required to be in office five days a week, based in any of Turing's offices in San Francisco, Palo Alto, or Seattle.

Compensation: $250,000 to $350,000 OTE + Equity

About Turing

Turing is a technology company that provides a platform for companies to hire remote software developers. The company's platform uses artificial intelligence to match companies with developers who have the skills and experience they need. Turing was founded in 2018 and is headquartered in San Francisco, California. The company has raised $32 million in funding to date.
Learn more about Turing
Size
200 employees
Industry
Founded
2018

Similar Jobs

More Jobs at Turing

More Information Technology Jobs

Find similar Senior Research Engineer jobs: