Research, Simulations

talp

Fatih

On-site

TRY 2,790,000 - 5,580,000

Full time

3 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Talp is a research-driven company building simulation infrastructure for enterprise decisions. You will shape the scientific direction, design experiments, and turn research into working systems spanning modeling, evaluation, and deployment.

This is a hands-on role demanding rigorous thinking and close collaboration with engineering. You should have five or more years of AI/ML research experience, strong Python skills, and hands-on experience with language models and modern training techniques.

Qualifications

  • Five+ years of AI/ML research experience.
  • Ability to translate research ideas into production-ready systems.
  • Strong Python programming and ML framework skills (PyTorch/JAX).

Responsibilities

  • Shape the research direction and tie experiments to product improvements.
  • Develop and improve models across data prep, training, and fine-tuning.
  • Define benchmarks and evaluate behavioral accuracy and usefulness.
  • Build repeatable evaluation systems with traceable experiments.
  • Collaborate with engineering on data pipelines, training workflows, and model serving.
  • Carry improvements into production and publish findings when appropriate.
  • Set rigorous standards for research quality and communication.

Skills

Applied AI research
Python
PyTorch
JAX
Statistical judgment
ML tooling
Autonomy
Intellectual honesty
Deep model experience

Education

PhD welcome (not required)

Tools

ML frameworks

Job description

Compensation: 250.000-500.000 TL net per month · equity · private health insurance

About Talp

Talp is simulation infrastructure for enterprise decisions. We build a company's customer base from its own behavioral data so decisions can be tested before they are made, and we already work with large enterprises across three regions.

We are the wind tunnel for decisions. What we install stays and sharpens every time it runs, which is why we call it infrastructure rather than a tool. Testing decisions this way will soon be as ordinary as version control.

We have raised $2.2 million to date, most recently at a $25 million valuation, from Formus Capital, Sunshine Lake Ventures, Aito Capital, the a16z Scout Fund, and a group of GPs and founders.

The Role

Everyone in this field renders the simulation. Nobody renders the proof. Confidence scores, calibration methods, benchmarks measured against human evidence. None of it has been published by anyone in this category, and the company that publishes it first will define how everybody else gets measured.

That is the center of this job.

You will participate in shaping our scientific direction and turn research into working systems. Your scope spans modeling, evaluation, and the infrastructure that supports experimentation and deployment.

This is a hands-on role. You will design experiments, write code, inspect data, and work closely with engineering. You should be comfortable asking difficult scientific questions and taking full responsibility for how they are answered.

Responsibilities
  • Shape the research direction: Identify the most valuable questions, prioritize experiments, and connect scientific progress directly to product improvements.
  • Develop and improve models: Work across data preparation, model adaptation, training, and fine-tuning. Reproduce relevant research, challenge its assumptions, and test approaches against strong alternatives.
  • Define how quality is measured: Design benchmarks and experiments that assess behavioral accuracy, consistency, and usefulness. Compare simulations against human evidence while accounting for uncertainty, population variance, and data limitations.
  • Build reliable evaluation systems: Develop repeatable workflows for comparing models, investigating failures, and detecting regressions. Keep datasets, experiments, and results strictly traceable, protecting evaluation data from leaking into development.
  • Combine human and automated judgment: Create clear scoring criteria, review processes, and automated evaluators, then verify that those evaluators consistently agree with qualified human reviewers.
  • Make research practical to run: Partner with engineering on data pipelines, experiment tracking, training workflows, and model serving to maximize iteration speed, reliability, and compute efficiency.
  • Carry improvements into production: Follow promising results through implementation and deployment, verifying that their advantages hold outside the original experiment.
  • Set the scientific standard: Establish clear standards for rigor, and communicate findings through technical writing, publications, and external collaborations.
Who You Are
  • Applied Research Record: A strong track record in applied AI or machine learning research, with evidence of taking ideas from experiments into working production systems.
  • Deep Model Experience: Hands-on experience with language models, modern training/fine-tuning techniques, and rigorous model evaluation.
  • Core Engineering Skills: Strong Python skills and deep familiarity with frameworks such as PyTorch or JAX.
  • Sound Statistical Judgment: You reason intuitively about sampling, uncertainty, and bias, and can tell immediately whether an apparent improvement is statistically meaningful.
  • Systems & Tooling Ownership: Experience building or owning research tooling, evaluation pipelines, or ML infrastructure that other engineers rely on.
  • High Agency: The ability to drive technical work autonomously and make clear decisions even when the empirical evidence is incomplete.
  • Intellectual Honesty: Clear communication and the willingness to discard an approach the moment the data challenges it.

Typically five or more years of relevant AI or ML research experience. Relevant academic research counts toward this; we also welcome candidates with exceptional demonstrated impact on a shorter timeline.

Valuable Additional Experience
  • Behavioral science, computational social science, causal inference, or experimental design
  • Synthetic data generation, human feedback loops, annotation systems, or complex behavioral datasets
  • Multi-agent systems and simulations involving interacting populations
  • Distributed training, GPU optimization, or efficient model serving

You do not need equal depth in every area. We look for deep research instincts combined with strong engineering ability.

A PhD is welcome but not required.

Our Process

We prioritize thoughtful conversations and clear examples of past work. Our hiring process is designed to help both sides align on mutual fit, working style, and expectations.

Reapplication Policy

To ensure a fair and thorough evaluation for all applicants, Talp observes a 90-day waiting period before reconsidering candidates for the same role.

Talp is an equal opportunity workplace. We welcome applicants of every background and identity. If you need support or an accommodation at any point in the process, let us know and we will arrange it.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior AI Researcher - Simulation & Evaluation
Senior AI Researcher - Simulation & Evaluation

talp • Fatih

On-site
TRY 2,790,000 - 5,580,000
Artificial Intelligence Engineer
Artificial Intelligence Engineer

Seamflow • Fatih

On-site
TRY 7,843,000 - 11,765,000
Relocation support
Visa sponsorship
Lunch & dinner covered
+2
AI Engineer - Core
AI Engineer - Core

Hilbert's AI • Turkey

On-site
TRY 900,000 - 1,200,000
Software Engineer (Full-Stack / Product)
Software Engineer (Full-Stack / Product)

Hilbert • Turkey

On-site
TRY 600,000 - 1,200,000
Equity package
Equal Opportunity Employer
AI Engineer - Core
AI Engineer - Core

Hilbert • Turkey

On-site
TRY 600,000 - 900,000
Equity package
Forward-Deployed AI Engineer
Forward-Deployed AI Engineer

Zaigo • Fatih

On-site
TRY 600,000 - 1,000,000
Senior Generative AI Operations (GenAI Ops) Engineer
Senior Generative AI Operations (GenAI Ops) Engineer

EPAM Systems • Turkey

On-site
TRY 2,917,000 - 4,375,000
Private health insurance
English courses
Continuous upskilling
Elite Research Scientist - Frontier AI Evaluation
Elite Research Scientist - Frontier AI Evaluation

Perle • Turkey

Remote
TRY 2,663,000 - 4,439,000
Flexible engagement
Direct exposure to frontier AI research challenges
Collaboration with an elite peer group
Forward Deployed AI Engineer, Türkiye - BCG X
Forward Deployed AI Engineer, Türkiye - BCG X

BCG X • Fatih

Hybrid
TRY 1,000,000 - 1,900,000
Junior AI Creative - Marketing Artist
Junior AI Creative - Marketing Artist

Moyra • Fatih

Hybrid
TRY 180,000 - 240,000
Unlimited private health insurance for
Multisport card
Massage chairs and gaming zones
+3