Data Engineer, Specialist

Vanguard

Hyderabad

On-site

INR 1,200,000 - 1,800,000

Full time

34 hours ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Vanguard India in Hyderabad is seeking an experienced data engineer to design, build and operate scalable data pipelines on AWS (Glue, Athena, S3, Lambda, Step Functions). You will optimize data processing and ensure reliable, reproducible data delivery for analytics and reporting.

You will refactor PySpark transformations, implement reconciliation checks, and collaborate on a lakehouse migration to Delta/Databricks, while maintaining production readiness and robust observability.

Qualifications

  • 5+ years in data engineering with Python and SQL.
  • Hands-on PySpark / Spark for large-scale data transformation.
  • Production experience with AWS data services— Glue, Athena, S3, Lambda, Step Functions.
  • Solid data modelling skills and a reconciliation-first approach to data quality.
  • Ability to read a business or technical specification and turn it into correct, maintainable pipeline logic.
  • Comfortable with Git-based workflows, CI/CD, and infrastructure-as-code.

Responsibilities

  • Build and operate production data pipelines on AWS — Glue, PySpark, Step Functions, Lambda, and S3.
  • Own operational support for production data pipelines, including monitoring and incident response.
  • Develop and refactor transformation logic in PySpark and optimize large-scale Spark jobs.
  • Implement data reconciliation and sign-off with parity checks between systems.
  • Translate complex business logic into well-documented code for production use.
  • Drive production readiness with pre-checks, access setup, and runbooks for support.
  • Contribute to cloud lakehouse migration and repeatable ETL improvements.
  • Keep pipelines healthy by monitoring jobs, triaging failures, and maintaining month/quarter-end processing.

Skills

Python
SQL
PySpark
Spark
AWS
Git CI/CD

Tools

Glue
Athena
S3
Lambda
Step Functions

Job description

Provides advanced data solutions by using software to process, store, and serve data to others. Tests data quality and optimizes data availability. Ensures that data pipelines are scalable, repeatable, and secure. Utilizes a deep dive analytical skillset on a variety of internal and external data.

About Vanguard

Founded in 1975, Vanguard is one of the world's leading investment management companies. The firm offers investments, advice, and retirement services to tens of millions of individual investors around the globe—directly, through workplace plans, and through financial intermediaries.

Vanguard India

Vanguard’s office in India is a significant milestone in our global expansion. We are committed to establishing an enduring technology center in Hyderabad, Telangana and are excited to be adding talent who will focus on Artificial Intelligence (AI), mobile, and cloud-based technologies that drive our business outcomes and deliver a world-class experience for our clients.

Responsibilities
  • Build and operateproduction data pipelineson AWS — Glue (PySpark), Step Functions, Lambda, Athena, SNS, and S3 — from ingestion through transformation to reconciled, reportable output.
  • Own operational support for production data pipelines, including monitoring, incident management, PagerDuty response, root cause analysis, and driving timely resolution of production issues while continuously improving platform reliability and observability.
  • Develop and refactortransformation logic in PySpark, including migrating legacy Pandas/SQL logic and performance-tuning large-scale Spark jobs.
  • Implementdata reconciliation and sign-off— design recon checks (including parity between old and new implementations), investigate discrepancies to root cause, and produce the evidence the business needs to trust a run.
  • Translatecomplex business and financial logic(asset classification, valuation and pricing rules, methodology calculations) into correct, testable, well-documented code.
  • Take features throughproduction readiness— pre-checks, access/role setup, environment promotion, and clear runbooks for ongoing support.
  • Contribute to ourcloud lakehouse migration— modelling refined tables, building repeatable ETL, and lifting reporting workloads onto a modern Delta/Databricks platform.
  • Keep running pipelines healthy — monitor scheduled jobs, triage failures (partitioning, schema/data-type issues, backfills), and keep periodic (month-end / quarter-end) processing on time.
What We're Looking For
Must have
  • 5+ years in data engineering with strongPythonandSQL.
  • Hands-onPySpark / Sparkfor large-scale data transformation, including debugging and performance tuning.
  • Production experience withAWS data services— Glue, Athena, S3, Lambda, Step Functions (or close equivalents on another major cloud).
  • Soliddata modellingskills and a rigorous, reconciliation-first approach todata quality.
  • Ability to read a business or technical specification and turn it into correct, maintainable pipeline logic.
  • Comfortable withGit-based workflows, CI/CD, and infrastructure-as-code.
Nice to have
  • Databricks/ lakehouse (Delta Lake) experience.
  • Background infinancial services— custody, fund accounting, unit pricing, or regulatory/financial reporting.
  • Experience withaccess governance and IAM(e.g. role provisioning) and writing operational runbooks.
  • Track recordmigrating legacy reporting or ETL onto modern cloud data platforms.
Location

This role is based in Hyderabad, Telangana at Vanguard India. Only qualified external applicants will be considered.

Our mission

Vanguard adheres to a simple purpose: To take a stand for all investors, to treat them fairly, and to give them the best chance for investment success.

Our commitment to you

Vanguard takes the same long-term view of your success—at work and in life—with Benefits and Rewards packages that reflect what you care about, throughout all the phases and stages of your life. Our Total Rewards programs provide you and your loved ones with wellness support for key areas in your life:

Financial wellness

We're committed to enabling your financial success and provide competitive offers and programs.

Physical wellness

We're committed to providing benefits that support your physical health and wellness.

Personal wellness

We're committed to providing resources that help support the full scope of your life.

How We Work

Vanguard has implemented a hybrid working model for most of our employees (crew members), designed to capture the benefits of enhanced flexibility while enabling in-person learning, collaboration, and connection. We believe our mission-driven and highly collaborative culture is a critical enabler to support long-term client outcomes and enrich the employee experience.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Data Engineer, Specialist
Data Engineer, Specialist

Vanguard India • Hyderabad

Hybrid
INR 1,400,000 - 2,100,000
Data Engineer, Specialist (WT Service Experience)
Data Engineer, Specialist (WT Service Experience)

Vanguard • Hyderabad

Hybrid
INR 1,100,000 - 1,600,000
Hybrid work model
Data Engineer, Specialist (PITech- Core - Middle Office Team 1)
Data Engineer, Specialist (PITech- Core - Middle Office Team 1)

Vanguard • Hyderabad

On-site
INR 1,800,000 - 2,800,000
Data Engineer, Specialist (WT Service Experience)
Data Engineer, Specialist (WT Service Experience)

Vanguard India • Hyderabad

On-site
INR 2,400,000 - 4,200,000
Data Engineer, Specialist (PITech- Core - Middle Office Team 1)
Data Engineer, Specialist (PITech- Core - Middle Office Team 1)

The Vanguard Group • Hyderabad

Hybrid
INR 1,200,000 - 1,800,000
Hybrid work model
Total Rewards program
Data Engineer, Specialist (WT Service Experience)
Data Engineer, Specialist (WT Service Experience)

The Vanguard Group • Hyderabad

On-site
INR 1,800,000 - 2,400,000
Manager, Data Engineering
Manager, Data Engineering

The Vanguard Group • Hyderabad

Hybrid
INR 4,200,000 - 6,600,000
Fraud Strategy, Specialist
Fraud Strategy, Specialist

Vanguard • Hyderabad

Hybrid
INR 2,500,000 - 4,000,000
Application Engineer III
Application Engineer III

Vanguard • Hyderabad

On-site
INR 2,500,000 - 4,000,000
Application Engineer III
Application Engineer III

Vanguard India • Hyderabad

Hybrid
INR 1,800,000 - 3,200,000