Research Engineer, Frontier Data

turing

Brasil

Teletrabalho

BRL 180 000 - 240 000

Tempo integral

há 8 horas
Torna-te num dos primeiros candidatos
Gerador de candidaturas

Transforma esta função numa entrevista — um currículo e uma carta de apresentação criados à volta do que este empregador procura.

Ultrapassa os filtros ATS

Vantagens oferecidas por esta oferta de emprego

Remote work
Flexible schedule

Resumo da oferta

Turing is seeking a Research Engineer to help deliver frontier-quality datasets, RL environments, and evaluations for a frontier AI lab client. You will work directly with the client's researchers and engineers, turning requests into concrete data requirements and owning the technical execution with high standards for correctness and realism.

This remote role can be performed anywhere in Brazil, with flexibility to support other accounts.

Qualificações

  • Experience building or improving deep learning systems with data quality as a focus.
  • Strong programming skills in Python; proficiency in SQL and data pipelines.
  • Ability to design, implement, and evaluate data and environment quality rubrics.

Responsabilidades

  • Own data and environment quality from an AI researcher perspective.
  • Design and build datasets and RL environments for capabilities.
  • Build verification and denoising systems and synthetic data pipelines.
  • Collaborate with client researchers and engineers to translate goals into specs.
  • Provide analysis showing data impact on model performance.

Conhecimentos

Python
SQL
Deep learning
RL environments
Data quality
Verifiers
Trajectory analysis
Communication

Formação académica

BS in CS/ML

Ferramentas

Git
PyTorch
TensorFlow

Descrição da oferta de emprego

Turing’s mission is to accelerate superintelligence to drive real economic progress. Headquartered in San Francisco, Turing works with frontier AI labs to generate high-quality datasets, reinforcement learning environments, and frontier research benchmarks that improve model capabilities in software engineering, enterprise knowledge work, and advanced STEM reasoning. In software engineering, Turing is the largest and longest-running data provider in the category. Turing also works with Fortune 500 enterprises across financial services, life sciences, healthcare, retail, automotive, and CPG to build and deploy end-to-end agentic AI systems inside mission-critical workflows. By operating on both sides, Turing closes the loop between frontier research and enterprise deployment, turning real-world deployment signals into better data, evaluations, and more capable models. Learn more at www.turing.com .

*This is a remote role and can be performed anywhere in Brazil/Colombia*

The Role

We are looking for a Research Engineer to help deliver frontier-quality datasets, RL environments, and evaluations that improve state-of-the-art models for a frontier AI lab client.You will work directly with the client's researchers and engineers, turning inbound requests and post-training goals into concrete technical proposals and data/environment specifications, and then owning the technical execution: building the quality and verification systems that ensure what we deliver meets extremely high standards for correctness, realism, diversity, difficulty, and measurable model lift.

Most of this work is bespoke and custom, built for one client's evolving needs.You'll typically stay attached to a single frontier lab client so you can build real context and a working relationship with their team, with flexibility to support other accounts when needed. Because bespoke work is judged on both turnaround time and quality, both matter equally here.

This role is designed for candidates with experience building and improving deep learning systems, especially where strong results depend on data quality, data curation, denoising, synthetic data generation, and rigorous evaluation. You'll operate in one or more of the following capability areas:

  • Coding and software engineering agents (repositories, unit tests, debugging, tool use, code reviews, long-horizon workflows)
  • RL environments and verifier-based training (tasks, rewards/verifiers, trajectories, evaluation harnesses)
  • Multimodal data and reasoning (text + images + documents + tables/charts; optional audio/video)
  • STEM reasoning (math, physics, chemistry, bio, engineering – solution verification and error analysis)
  • Modern embodied AI / VLM-driven agents (vision-language(-action) models, embodied task suites, tool/sensor/action abstractions, long-horizon interaction data)
What You'll Do
1) Own data and environment quality from an AI researcher perspective
  • Respond to inbound requests from the client and translate ambiguous, evolving research goals into a technical proposal and clear data requirements: target skills, failure modes, difficulty calibration, coverage, and success metrics.
  • Provide the technical feasibility assessment, assumptions, acceptance criteria, and evaluation requirements needed to scope the contract or statement of work.
  • Define what “good” looks like by creating detailed rubrics, counterexamples, and boundary cases (what to include vs. exclude).
  • Perform deep, detail-oriented audits of produced data: spot subtle errors, reward hacking opportunities, leakage, ambiguity, inconsistent assumptions, and distribution shifts.
  • Drive iterative improvements using evidence: group failures (for example, by semantic similarity) to find the underlying gap, then turn that into concrete instructions for the production team — better few-shot examples, explicit counterexamples, clearer guidance on what's causing rejections.
2) Design and build datasets and RL environments for your capability area(s)

Contribute to or lead the design of:

  • Task suites (single-step and long-horizon workflows)
  • Ground-truth signals (verifiers, unit tests, structured checks, reward functions, automatic validators)

Depending on your mapped capability area(s), you may focus on:

  • Coding / SWE agents: data reflecting real development work (codebase navigation, bug localization, patching, tests, code reviews, CI-like constraints, refactors, security fixes).
  • Multimodality: tasks that test true multimodal reasoning (chart reading, document QA, UI understanding, diagram-based STEM reasoning, OCR-aware tasks).
  • STEM: tasks with verifiable solutions (symbolic checks, reference solvers, numerical validation, step consistency, unit sanity).
  • Modern embodied AI / VLM-driven agents: interaction data and environments for vision-language(-action) models.
3) Build robust validation, denoising, and synthetic data systems
  • Build project-specific verification systems that combine deterministic checks, model-based evaluators calibrated against gold sets, and targeted human review.
  • Implement automated validation and filtering to achieve frontier-grade signal-to-noise: deduplication, decontamination, leakage checks, consistency checks, difficulty and diversity controls.
  • Develop synthetic data generation and augmentation pipelines where appropriate: programmatic task generators, controlled perturbations, scenario templating, simulator-/tool-driven rollouts.
  • Create documentation and data cards: dataset intent, known limitations, recommended use, and evaluation linkage.
4) Use evaluations and training runs to prove impact
  • Design and run evals that reflect the client's intended usage.
  • Produce analysis that connects data to outcomes: pre/post comparisons, error breakdowns, ablations that identify which data attributes drive lift.
  • When needed, run in-house fine-tuning or RL-style experiments (or partner with research) to demonstrate that the data/environment improves model behavior in measurable ways.
5) Collaborate effectively with large production teams without being ops-heavy
  • Give technical direction to the distributed expert network producing the data, by providing clear specs, examples, edge cases, and fast feedback loops based on audits and quantitative signals.
  • You own technical translation, acceptance criteria, evaluator and verifier design, quality diagnosis, and the technical release recommendation. The Strategic Project Lead owns production execution, staffing, throughput, project cost, and recovery.
  • You are expected to be highly engaged in reviewing and improving outputs from large annotation/creation efforts, but not primarily responsible for hiring, staffing, or people operations — that sits with the SPL.
  • Communicate directly with the client on technical findings, quality risks, and recovery options. Align changes to scope, schedule, or external commitments with the account and delivery owners.
Who We're Looking For
  • 4–5 years of experience building or improving deep learning systems where data quality mattered materially (training, post-training, evals, or agentic systems).
  • Strong intuition for the “data ingredients” that drive model improvements: what to collect, what to filter, what to synthesize, and how to measure.
  • Ability to communicate clearly with researchers and engineers: turning research objectives into concrete specs, and turning messy outputs into actionable insights.
  • Comfortable being the direct technical point of contact for a client, including explaining setbacks or trade-offs when things don't go as planned.
  • Demonstrated ability to be extremely detail-oriented in diagnosing subtle data quality issues and failure modes.
  • Solid programming ability with a bias for shipping: Python proficiency required; comfort with SQL/structured data workflows strongly preferred; for coding-focused work, proficiency in a major language (e.g., C++, Java, Go, Rust, JS/TS) is a plus.
  • Comfort designing quality systems: rubrics, validation scripts, gold sets, sampling strategies, statistical checks, slice-based evaluation, human-in-the-loop review loops grounded in measurable criteria.
  • Direct experience with a frontier AI lab is not required, and is uncommon in the LATAM market — what matters most is hands‑on engineering depth and the ability to operate with ambiguity.
  • RL or post‑training experience (any of: RLHF/RLAIF, verifier training, reward modeling, RL fine‑tuning, environment design).
  • Experience with agentic evaluation (tool use, multi‑step workflows, long‑horizon tasks, trajectory analysis).
  • STEM depth (math/physics/engineering) with an eye for verifiability and rigorous correctness.
  • Systems thinking: ability to “simulate” an application's API/data schema and design tasks that realistically reflect real‑world constraints and workflows.
Why Turing
  • Work directly with a frontier AI lab, at the cutting edge of post‑training and RL environment design, embedded in a real research relationship rather than a rotating set of anonymous tasks.
  • Real impact (path to AGI): your datasets and environments will directly influence the trajectory toward Artificial General Intelligence and, ultimately, Superintelligence.
  • High autonomy and fast iteration, with a Strategic Project Lead handling scaling and resourcing so you can stay focused on the technical approach and quality bar.
  • Talent‑dense team, where you'll find rapid iteration and an exceptional learning curve.
Values
  • We are client first: We put our clients at the center of everything we do, because their success is the ultimate measure of our value.
  • We work at Start-Up Speed: We move fast, stay agile and favor action because momentum is the foundation of perfection
  • We are AI forward: We help our clients build the future of Al and implement it in our own roles and workflow to amplify productivity.
Advantages of joining Turing
  • Work at the frontier of AI, helping the world’s leading AI labs improve their most advanced models by building expert datasets, RL environments, and first-of-a-kind benchmarks.
  • Contribute to leading‑edge AI research and showcase your work at top conferences such as ICLR, ICML, and NeurIPS.
  • Bring frontier AI innovation to the enterprise, applying lessons learned from leading AI labs to solve real‑world business challenges.
  • Collaborate with and learn from exceptional colleagues with deep AI experience from Google, Meta, Amazon, and other leading technology companies.
  • Move at the pace of AI innovation, with the speed, ownership, and impact of a startup.

Turing is proud to be an equal opportunity employer. We do not discriminate on the basis of race, religion, color, national origin, gender, gender identity, sexual orientation, age, marital status, disability, protected veteran status, or any other legally protected characteristics. At Turing we are dedicated to building a diverse, inclusive and authentic workplaceand celebrate authenticity, so if you’re excited about this role but your past experience doesn’t align perfectly with every qualification in the job description, we encourage you to apply anyways. You may be just the right candidate for this or other roles.

Obtém a tua avaliação gratuita e confidencial do currículo.

ou arrasta e larga o ficheiro aqui.

Similar jobs

Ofertas semelhantes que vale a pena comparar

Research Engineer
Research Engineer

Turing • Brasil

Presencial
BRL 167 400 - 279 000
Competitive compensation
Supportive work environment
Opportunity to work with top talent
Senior Software Engineer
Senior Software Engineer

Turing • São Paulo

Presencial
BRL 180 000 - 360 000
Senior Engineering Manager
Senior Engineering Manager

Foundation Capital • São Paulo

Teletrabalho
BRL 627 000 - 1 044 000
Remote Data Scientist (ML) - 52621
Remote Data Scientist (ML) - 52621

Turing • Brasil

Presencial
BRL 20 000 - 30 000
Work fully remotely
Collaborate with world-class AI labs
High autonomy and ownership
Research Scientist / Research Engineer
Research Scientist / Research Engineer

adaption • São Paulo

Presencial
BRL 917 899 - 1 529 832
Flexible work Bay Area
Travel stipend
Lunch stipend
+1
Senior Applied AI Researcher (Brazil)
Senior Applied AI Researcher (Brazil)

Articul8 • Brasil

Presencial
BRL 120 000 - 150 000
Mentorship opportunities
Innovative research environment
Community Manager
Community Manager

Turing • São Paulo

Presencial
BRL 80 000 - 100 000
Competitive compensation
Collaborative work culture
Amazing colleague network
Senior Research Scientist
Senior Research Scientist

adaption • São Paulo

Presencial
BRL 250 000 - 420 000
Flexible work arrangements
Annual global travel stipend
Lunch stipend
+1
Backend Engineer, AI Integrations
Backend Engineer, AI Integrations

Lever, Inc. • Brasil

Teletrabalho
BRL 311 000 - 519 000
Fully remote (Brazil)
Flexible hours
Ownership & autonomy
+2
Remote Business Analyst - 57510
Remote Business Analyst - 57510

Turing • Brasil

Presencial
BRL 143 000 - 287 000
Competitive compensation based on experience
Flexible working hours
Remote work environment
+3