DATA Engineer

Insud Pharma

Madrid

On-site

EUR 65,000 - 90,000

Full time

43 hours ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Insud Pharma's AI Labs seeks a highly skilled Data Engineer / Machine Learning Engineer to join the Applied AI Team. The ideal candidate combines strong software engineering with hands-on data pipelines and ML systems, and works at the intersection of data, models, and production.

You will design, build, deploy, and operate end-to-end data and ML solutions across business units, ensuring reliable production moves, scalable pipelines, and robust ML infrastructure.

Qualifications

  • Proficient in Spanish and English, written and verbal communication.
  • Strong proficiency in Python, including clean code practices, packaging, and modular design.
  • Solid understanding of software engineering principles (OOP, SOLID, testing, version control).
  • Hands-on experience building data pipelines (ETL / ELT) using Python-based frameworks or custom solutions.
  • Experience working with machine learning workflows, including model training, evaluation, and deployment.
  • Familiarity with REST APIs and service-based architectures (FastAPI, Flask, or similar).
  • Strong experience with Git and collaborative development workflows.
  • Experience with containerization (Docker) and cloud environments (AWS or Azure).
  • Experience with MLOps practices (model versioning, monitoring, drift detection, retraining strategies).
  • Familiarity with orchestration tools (Airflow, Prefect, Dagster).
  • Experience with data storage systems (SQL / NoSQL databases, data lakes, object storage).
  • Experienced deploying or operating ML systems in regulated or high-reliability environments.
  • Familiarity with ML frameworks and scientific libraries (NumPy, Pandas, Scikit-learn, PyTorch, TensorFlow).
  • Interest in applied AI topics such as NLP, LLM-based systems, or scientific computing.

Responsibilities

  • Design, build, and maintain scalable data pipelines for data ingestion, transformation, and serving, supporting analytics and ML use cases.
  • Develop and productionize machine learning pipelines, covering training, validation, deployment, and monitoring.
  • Collaborate closely with Data Scientists to translate notebooks and prototypes into robust, production-ready ML systems.
  • Build and maintain feature pipelines and data abstractions that enable reproducible and reliable model behavior.
  • Ensure data quality, versioning, and traceability across datasets and models.
  • Optimize pipelines and ML workloads for performance, scalability, and cost efficiency.
  • Work with DevOps and Platform teams to deploy solutions using containerization and CI/CD best practices.
  • Contribute to defining data engineering and MLOps standards across AI Labs; participate in code reviews, documentation, and mentoring to foster excellence.

Skills

Spanish proficiency
English proficiency
Python programming
Software engineering principles
ETL/ELT pipelines
ML workflows
REST APIs
Git teamwork
CI/CD practices
DevOps practices

Tools

Docker
AWS/Azure
Airflow/Prefect/Dagster
FastAPI/Flask
NumPy
Pandas
Scikit-learn
PyTorch
TensorFlow

Job description

WeareseekingahighlyskilledData Engineer / Machine Learning Engineerto join our Applied AI Team. The ideal candidate combines strong software engineering foundations with hands-on experience in data pipelines and machine learning systems, and enjoys working at the intersection between data, models, and production systems.

As a Data Engineer / MLE at AI Labs, you will work closely with data scientists, software engineers, and product owners todesign, build, deploy, and operate end-to-end data and machine learning solutionsacross multiple business units — including Regulatory, Clinical Trials, R&D, Pharmacovigilance, and Drug Manufacturing.

This role is critical to ensuring that AI models move reliably from experimentation to production, supported by scalable data pipelines, robust ML infrastructure, and strong engineering standards.

Specific Responsibilities

Design,build, andmaintainscalabledata pipelinesfordataingestion,transformation, andserving,supportingbothanalyticsand machinelearninguse cases.

Developandproductionizemachinelearningpipelines,coveringtraining,validation,deployment, andmonitoring.

CollaboratecloselywithDataScientiststotranslatenotebooks andprototypesintorobust,production-readyMLsystems.

Buildandmaintainfeaturepipelines and dataabstractionsthatenablereproducible andreliablemodelbehavior.

Ensuredataquality,versioning, andtraceabilityacrossdatasetsandmodels.

Optimizepipelines and MLworkloadsforperformance,scalability, andcostefficiency.

WorkwithDevOps andPlatformteamstodeploysolutionsusingcontainerizationand CI/CDbestpractices.

ContributetodefiningdataengineeringandMLOpsstandardsacrossAI Labs.Participateincodereviews,documentation, andmentoringtofostera cultureofengineeringexcellence.

Requirements and personal skills

ProficientinSpanish and English,writtenand verbalcommunication.

StrongproficiencyinPython,includingcleancodepractices,packaging, and modulardesign.

Solidunderstandingofsoftwareengineeringprinciples(OOP, SOLID,testing,versioncontrol).

Hands-onexperiencebuildingdata pipelines(ETL / ELT)usingPython-basedframeworksorcustomsolutions.

Experienceworkingwithmachinelearningworkflows,includingmodeltraining,evaluation, anddeployment.

FamiliaritywithRESTAPIsandservice-basedarchitectures(FastAPI, Flask,orsimilar).

StrongexperiencewithGitandcollaborativedevelopmentworkflows.

Experiencewithcontainerization(Docker)andcloudenvironments(AWSorAzure).

ExperiencewithMLOpspractices(modelversioning,monitoring,driftdetection,retrainingstrategies).

Familiaritywithorchestrationtools(e.g.,Airflow,Prefect,Dagster).

Experiencewithdatastoragesystems(SQL / NoSQLdatabases, datalakes,objectstorage).

ExperiencedeployingoroperatingMLsystemsinregulatedorhigh-reliabilityenvironments.

FamiliaritywithMLframeworksandscientificlibraries(NumPy, Pandas,Scikit-learn,PyTorch,TensorFlow).

InterestinappliedAItopicssuchas NLP, LLM-basedsystems,orscientificcomputing.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Data Engineer
Data Engineer

Pyramid Consulting, Inc • Spain

On-site
EUR 85,000 - 120,000
Data Engineer for ML Pipelines & Production
Data Engineer for ML Pipelines & Production

Insud Pharma • Madrid

On-site
EUR 65,000 - 90,000
Senior Data Engineer
Senior Data Engineer

Encardio Rite Group • Madrid

On-site
EUR 120,000 - 180,000
Senior Machine Learning Engineer
Senior Machine Learning Engineer

EPAM Systems • Spain

On-site
EUR 70,000 - 110,000
Hybrid work model
Senior Applied Data Scientist - Data Architecture & Feature Engineering
Senior Applied Data Scientist - Data Architecture & Feature Engineering

Keysight Technologies • Barcelona

On-site
EUR 60,000 - 80,000
AI Machine Learning Engineer
AI Machine Learning Engineer

Visium SA • Barcelona

On-site
Confidential
Competitive compensation package
Yearly education budget
Yearly sport budget
+3
Artificial Intelligence Engineer
Artificial Intelligence Engineer

Migx • Barcelona

On-site
EUR 90,000 - 130,000
Hybrid work model
25 holiday days per year
Career development opportunities
+1
Data Scientist (GenAI)
Data Scientist (GenAI)

Plain Concepts Group • Spain

Remote
EUR 60,000 - 90,000
Flexible 35-hour week
100% Remote work (optional)
Free medical and dental insurance
+7
AI Machine Learning Engineer
AI Machine Learning Engineer

Visium • Barcelona

On-site
EUR 60,000 - 80,000
Data Engineer MLE
Data Engineer MLE

Iwantic • Madrid

Hybrid
EUR 55,000 - 75,000
Ticket restaurante
Seguro de vida y accidentes
Servicio médico
+1