Data Engineer

Veeva

Boston (MA)

On-site

USD 75,000 - 130,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Medical insurance
Dental insurance
Vision insurance
Flexible PTO
Retirement plan
Charitable giving program

Job summary

Veeva is seeking a Data Engineer to own end-to-end data platform development in a fast-paced environment. You will design, build, and deploy scalable data processing systems using Python, Spark, and AWS, supporting life sciences customers.

You will implement ETL/ELT workflows, lead the Medallion Architecture, and collaborate with product managers and data scientists to deliver impactful features while maintaining data governance and quality.

Qualifications

  • 4+ years of professional data engineering experience.
  • Expert-level proficiency in Python and Apache Spark with JVM tuning.
  • Experience with AWS (EMR) or Databricks for data platforms.
  • Experience with Airflow, AWS Step Functions, or Prefect.
  • Data cleansing, curation, transformation, and governance experience.
  • Ability to build reusable libraries and internal tools.
  • CI/CD tooling exposure (Codefresh, Jenkins).
  • Excellent English communication and cross-functional collaboration.

Responsibilities

  • Architect and build resilient, distributed data processing systems using Python and Spark on AWS.
  • Design and implement end-to-end ETL/ELT workflows ingesting data from diverse sources.
  • Lead the Medallion Architecture (Bronze, Silver, Gold) for scalable data maturity.
  • Build reusable libraries and frameworks for data quality, metadata tracking, and monitoring.
  • Create CI/CD processes to automate deployment and testing.
  • Enforce data governance, security, and regulatory compliance.
  • Monitor system health, resolve bottlenecks, and optimize resource use.
  • Partner with PMs and Data Scientists to translate requirements into solutions.
  • Own the full feature lifecycle from whiteboarding to production deployment.

Skills

Python
Apache Spark
AWS
Airflow
CI/CD
JVM tuning
Distributed systems
Data modeling
Data governance
ETL/ELT

Tools

Codefresh
Jenkins
Databricks
EMR
Kinesis
EKS
Kafka

Job description

The Role

Veeva OpenData supports the industry by providing real-time reference data across the complete healthcare ecosystem, to support commercial sales execution, compliance, and business analytics. We drive value to our customers through constant innovation, using cloud-based solutions and state-of-the-art technologies to deliver product excellence and customer success. As a Data Engineer, you will own the end-to-end development lifecycle, collaborating with a high-performing engineering team to design, build, and deploy high-impact features. Operating within a fast-paced Agile environment, you will have a direct hand in engineering the data foundation for Veeva’s life sciences customers.

  • Architect and build resilient, distributed data processing systems using Python and Spark on AWS
  • Design and implement end-to-end ETL/ELT workflows that ingest and unify data from diverse sources — ranging from modern table formats like Iceberg and Delta to legacy business files such as Excel and CSV — ensuring a scalable and consistent single source of truth for the organization
  • Lead the implementation of the Medallion Architecture, managing data maturity through Bronze, Silver, and Gold layers. You will define how data is structured, classified, and stored to maximize business value while ensuring scalability and high availability.
  • Build reusable libraries and frameworks for data quality validation, metadata tracking, and pipeline monitoring
  • Build CI/CD process, to automate deployment and testing to maintain a high bar for engineering excellence
  • Enforce data governance standards, including security, privacy, and regulatory compliance
  • Proactively monitor system health, implement automated observability, and resolve complex bottlenecks in distributed systems to ensure peak resource efficiency and cost-effectiveness
  • Partner directly with Product Managers and Data Scientists to translate business requirements into innovative solutions
  • Own the full feature lifecycle—from initial whiteboarding to production deployment and long‑term maintenance
  • 4+ years of professional data engineering experience with a demonstrated ability to architect and deploy production‑grade data platforms from scratch
  • Expert‑level proficiency in Python and Apache Spark, with specific experience in JVM tuning, memory management, and optimizing execution plans for large‑scale distributed workloads
  • Deep expertise in modern data architecture, software design patterns, and various data modeling techniques designed for scalability and performance
  • Proven track record of building on AWS (primary) or GCP, including hands‑on experience with managed services like EMR or Databricks
  • Extensive experience designing and managing complex data lifecycles using orchestration tools such as Airflow, AWS Step Functions, or Prefect
  • Deep understanding of data cleansing, curation, and transformation strategies, coupled with experience implementing data governance, security, and lifecycle management policies
  • Strong background in building reusable libraries, frameworks, and internal tools that standardize data ingestion and automate ETL/ELT workflows
  • Exceptional debugging skills for distributed systems and resolving performance bottlenecks at scale
  • Proficiency with CI/CD tools and processes (e.g. Codefresh, Jenkins)
  • Excellent verbal and written communication skills in English, with the ability to translate complex technical architectures into actionable insights for stakeholders and cross‑functional teams
  • Must be located in EST or CST
  • Applicants must have the unrestricted right to work in the United States. Veeva will not provide sponsorship at this time
  • Relevant certifications (e.g., AWS, Spark, or similar)
  • Familiarity with streaming and distributed technologies such as Spark Streaming, EKS, Kinesis, or Apache Kafka
  • Experience implementing or managing modern cloud data warehouses or lakehouse architectures
  • Prior experience working in the Life Sciences industry
  • Medical, dental, vision, and basic life insurance
  • Flexible PTO and company paid holidays
  • Retirement programs
  • 1% charitable giving program

Base pay: $75,000 - $130,000. The salary range listed here has been provided to comply with local regulations and represents a potential base salary range for this role. Please note that actual salaries may vary within the range above or below, depending on experience and location. We look at compensation for each individual and base our offer on your unique qualifications, experience, and expected contributions. This position may also be eligible for other types of compensation in addition to base salary, such as variable bonus and/or stock bonus.

Veeva is an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to race, color, sex, sexual orientation, gender identity or expression, religion, national origin or ancestry, age, disability, marital status, pregnancy, protected veteran status, protected genetic information, political affiliation, or any other characteristics protected by local laws, regulations, or ordinances. If you need assistance or accommodation due to a disability or special need when applying for a role or in our recruitment process, please contact us at talent_accommodations@veeva.com.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Data Scientist - Business Consulting
Data Scientist - Business Consulting

Veeva • Boston (MA)

On-site
USD 95,000 - 140,000
Medical, dental, vision, and life insurance
Flexible PTO
Retirement programs
+2
Senior Director - Engineering and Data Operations - OpenData
Senior Director - Engineering and Data Operations - OpenData

Veeva Systems • Chicago (IL)

On-site
USD 165,000 - 220,000
Medical, dental, vision, and basic life insurance
Flexible PTO and company paid holidays
Retirement programs
+1
Senior Data Scientist
Senior Data Scientist

Veeva • Boston (MA)

On-site
USD 135,000 - 185,000
Medical, dental, vision, and life insurance
Flexible PTO and company paid holidays
Retirement programs
+1
Senior Data Analyst
Senior Data Analyst

Veeva Systems • Germany (OH)

Hybrid
USD 63,000 - 87,000
1% charitable giving program
Life insurance
Pension contribution
+1
Senior Data Scientist
Senior Data Scientist

Veeva • New York (NY)

On-site
USD 135,000 - 185,000
Medical, dental, vision, and life insurance
Flexible PTO
Retirement programs
+1
Senior Director - Data Cloud Solution Consulting
Senior Director - Data Cloud Solution Consulting

Veeva • Philadelphia

Hybrid
USD 150,000 - 300,000
Medical, dental, vision, and life insurance
Flexible PTO and company paid holidays
Retirement programs
+1
Senior Health Data Operations Analyst - Compass
Senior Health Data Operations Analyst - Compass

Veeva Systems, Inc. • San Francisco (CA)

On-site
USD 80,000 - 125,000
Medical, dental, vision & life
Flexible PTO & holidays
Retirement programs
+1
Senior Health Data Operations Analyst - Compass
Senior Health Data Operations Analyst - Compass

Veeva Systems • United States

Hybrid
USD 80,000 - 125,000
Medical insurance
Flexible PTO
Retirement programs
+1
Senior Software Engineer - Python
Senior Software Engineer - Python

Veeva • Raleigh (NC)

Hybrid
USD 110,000 - 270,000
Medical, dental, and vision insurance
Flexible PTO
Retirement programs
+1
Senior Software Engineer - Python
Senior Software Engineer - Python

Veeva • Boston (MA)

Hybrid
USD 110,000 - 270,000
Medical, dental, and vision insurance
Flexible PTO
Retirement programs
+1