Data Engineer (Mid‑Level ), Global

Vantage Data Centers

Greater London

Hybrid

GBP 50,000 - 70,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Flexible work policy
Health and wellness benefits

Job summary

Vantage Data Centers seeks a Mid-Level Data Engineer for their London office. The role involves building and maintaining data pipelines on Azure, developing solutions that support analytics and AI use cases.

The ideal candidate will have a background in data engineering, proficiency in Python and SQL, and experience with Azure services. The position includes a flexible work arrangement, requiring 3 days on-site and 2 from home, supporting a work-life balance.

Qualifications

  • 3-5 years of experience in data engineering or analytics engineering.
  • Strong understanding of ETL/ELT pipelines and data transformation patterns.
  • Experience with source control and CI/CD workflows.

Responsibilities

  • Design, build, and maintain reliable, scalable data pipelines.
  • Develop batch and incremental data pipelines using Azure Data Factory.
  • Take ownership of assigned data pipelines, monitoring and troubleshooting.

Skills

Python
PySpark
SQL
Azure Data Factory
Data modeling
Data governance

Education

Bachelor’s degree in Engineering, Computer Science, Data Analytics, or a related field

Tools

Azure Synapse
Azure Data Lake Storage Gen2
GitHub
Jira

Job description

About Vantage Data Centers

Vantage Data Centers powers, cools, protects and connects the technology of the world’s well‑known hyperscalers, cloud providers and large enterprises. Developing and operating across North America, EMEA and Asia Pacific, Vantage has evolved data center design in innovative ways to deliver dramatic gains in reliability, efficiency and sustainability in flexible environments that can scale as quickly as the market demands.

Position Overview

This position will be based at our office in London in alignment with our flexible work policy. (3 days on site, 2 days from home).

Vantage Data Centers is seeking a Mid‑Level Data Engineer to help build, operate, and scale our enterprise data platform. This role is designed for an engineer who can operate independently, execute reliably in a fast‑paced environment, and take ownership of data pipelines and datasets with minimal ramp‑up.

As part of the Data Engineering & Business Intelligence team, you will be responsible for delivering production‑ready data solutions that support analytics, reporting, and emerging AI‑enabled use cases. You will work closely with senior data engineers and business partners, but this role assumes a self‑starter mindset with the ability to move from requirements to implementation without constant oversight.

Success in this position requires comfort with ambiguity, strong execution discipline, and accountability for results.

Essential Job Functions
  • Design, build, and maintain reliable, scalable data pipelines using Python and PySpark on the Microsoft Azure data platform.
  • Develop and operate batch and incremental data pipelines leveraging Azure Data Factory for orchestration and Azure Data Lake Storage Gen2 as the primary data store.
  • Independently implement SQL- and Spark‑based transformations to produce curated datasets that support enterprise reporting, analytics, and downstream consumption.
  • Take ownership of assigned data pipelines and datasets, including monitoring, troubleshooting, and performance optimization in production environments.
  • Work with Azure Synapse (dedicated or serverless where applicable) to support analytical workloads and data consumption patterns.
  • Collaborate with business analysts and cross‑functional stakeholders to translate data requirements into practical, working data solutions.
  • Prepare and structure data to support advanced analytics and AI‑enabled use cases by ensuring data quality, consistency, and documentation.
  • Apply established data governance, security, and engineering standards to ensure compliant, maintainable, and scalable solutions.
  • Participate in code reviews, technical discussions, and platform improvement initiatives as an active contributor.
  • Proactively identify data quality issues, pipeline risks, and improvement opportunities, and communicate them clearly in a fast‑paced environment.
Duties
  • Develop and maintain PySpark notebooks and jobs to ingest, transform, and curate data within the enterprise data platform.
  • Build and modify Azure Data Factory pipelines for batch and incremental data ingestion.
  • Implement Spark‑based transformations that write curated datasets to Azure Data Lake Storage Gen2 using established folder structures and naming conventions.
  • Create and maintain SQL views and tables in Azure Synapse to support analytics and reporting use cases.
  • Respond to pipeline failures, data validation issues, and operational alerts.
  • Perform basic performance tuning of Spark jobs (e.g., partitioning, filtering, incremental logic) within established architectural patterns and standards.
  • Validate data outputs with business partners and address data defects or discrepancies.
  • Commit code using Git, follow branching standards, and participate in pull request reviews.
  • Update documentation for pipelines, datasets, and operational runbooks as changes are made.
  • Execute assigned backlog items within sprint timelines and raise risks or blockers early.
  • Additional duties as assigned by management.
Job Requirements
Education & Experience
  • Bachelor’s degree in Engineering, Computer Science, Data Analytics, or a related field, or equivalent experience.
  • Minimum of 3–5 years of experience in data engineering or analytics engineering.
  • Proficiency in Python for building and maintaining data pipelines, automation, and data processing workflows, including use of PySpark.
  • Proficiency in SQL for querying, transformation, and analytical data processing.
  • Solid understanding of ETL/ELT pipelines, data transformation patterns, and data integration concepts.
  • Experience analyzing enterprise data sources to identify data relationships, transformations, and business rules.
  • Experience building solutions on the Microsoft Azure platform with exposure to services such as Azure Data Factory, Azure Synapse, Azure Data Lake Storage Gen2, and related analytics services.
  • Experience working with source control and CI/CD workflows using tools such as GitHub or Azure DevOps.
  • Working knowledge of data modeling fundamentals, including fact and dimension tables.
  • Strong communication and interpersonal skills with the ability to collaborate across teams in a fast‑paced environment.
  • Experience working in Agile development environments.
  • Experience using collaboration and project tracking tools such as Jira or similar tools.
  • Travel required is expected to be up to 10% but may increase over time as the business evolves.
Desired Qualifications
  • Experience working with distributed data processing frameworks, including Apache Spark.
  • Exposure to advanced analytics or AI‑adjacent data use cases, including preparing data for machine learning or intelligent applications.
  • Familiarity with additional Azure services such as Azure Functions or Logic Apps in support of data workflows.
  • Experience supporting data platform enhancement, refactoring, or modernization initiatives.
  • Familiarity with data quality, reliability, and operational best practices in production environments.
  • Experience working in a scaling or fast‑paced organization where priorities evolve quickly.

Vantage Data Centers is an Equal Opportunity Employer

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Data Engineer (Mid‐Level ), Global
Data Engineer (Mid‐Level ), Global

Vantage Data Centers • Greater London

Hybrid
GBP 55,000 - 75,000
Health benefits
Flexible work policy
Training and development opportunities
Hybrid Data Engineer: Azure & PySpark
Hybrid Data Engineer: Azure & PySpark

Vantage Data Centers • Greater London

Hybrid
GBP 55,000 - 75,000
Health benefits
Flexible work policy
Training and development opportunities
Azure Data Engineer
Azure Data Engineer

Accion Labs • Greater London

On-site
GBP 70,000 - 110,000
Data Engineer
Data Engineer

RedRock Resourcing • Birmingham

On-site
GBP 40,000 - 60,000
London Data Engineer: Azure Pipelines & PySpark
London Data Engineer: Azure Pipelines & PySpark

Vantage Data Centers • Greater London

Hybrid
GBP 50,000 - 70,000
Flexible work policy
Health and wellness benefits
Data Engineer - Tech Lead (Databricks, Pyspark)
Data Engineer - Tech Lead (Databricks, Pyspark)

EPAM Systems • Greater London

Hybrid
GBP 90,000 - 120,000
ESPP
Private medical insurance
Life assurance
+4
Senior Data Engineer (Azure)
Senior Data Engineer (Azure)

Cognitive Group | Part of the Focus Cloud Group • Greater London, Newcastle upon Tyne

Hybrid
GBP 70,000 - 100,000
30 days leave
Pension contribution
Income protection
+5
Data Engineer
Data Engineer

Halian | Managed Services, Recruitment Agency & Contract Staffing • Greater London

Hybrid
GBP 60,000 - 80,000
Collaborative and engineering-led environment
Influence tooling and long-term strategy
Senior Data Engineer
Senior Data Engineer

TrueNorth® • Manchester

On-site
GBP 60,000 - 90,000
Senior Data Engineer
Senior Data Engineer

Nasstar • Greater London

Remote
GBP 60,000 - 80,000
Competitive salary and compensation package
Access to AWS & Databricks certifications
Remote work with team socials