Principal Data Engineer

Amgen

Hyderabad

On-site

INR 3,500,000 - 7,000,000

Full time

4 hours ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Amgen in Hyderabad is seeking a Principal Data Engineer to own the design and development of complex data pipelines, integration frameworks, and metadata-driven architectures enabling analytics and AI-driven insights. You will architect scalable Databricks-Spark-Delta Lake implementations and mentor engineers while partnering with cross-functional teams to align with enterprise data strategy.

The ideal candidate has deep experience in big data processing, data governance, and modern engineering

Qualifications

  • Must have hands-on experience with Databricks, PySpark, SparkSQL, Apache Spark, AWS, Python and SQL in large-scale data environments.
  • Strong knowledge of workflow orchestration, big data performance tuning, and Spark optimizations.
  • Experience with Data Fabric, Data Mesh, or equivalent enterprise data architectures.
  • Ability to quickly learn and apply new technologies; excellent problem-solving abilities.
  • Excellent communication and teamwork skills; experience in SAFe/Agile delivery and DevOps practices.

Responsibilities

  • Architect and maintain robust, scalable data pipelines using Databricks, Spark, and Delta Lake for batch and real-time processing.
  • Lead evaluation and integration of emerging data tools to improve productivity, scalability, and data delivery.
  • Drive Spark tuning, resource optimization, job scheduling, and query improvements for performance.
  • Build metadata-driven data engineering frameworks enabling reuse and consistency across pipelines.
  • Collaborate with platform, architecture, analytics, and governance teams to align with enterprise data strategy.
  • Define and uphold SLOs, data quality KPIs, and observability for production pipelines.

Skills

Databricks
PySpark
SparkSQL
Apache Spark
AWS
Python
SQL
SAFe
DevOps practices

Education

12 to 17 years of experience in Computer Science/IT
AWS Certified Data Engineer preferred
Databricks Certificate preferred
Scaled Agile SAFe certification preferred

Job description

Job Description

ABOUT THE ROLE

Role Description:

Let’s do this. Let’s change the world. We are looking for highly motivated expert Principal Data Engineer who can own the design & development of complex data pipelines, solutions and frameworks. The ideal candidate will be responsible to design, develop, and optimize data pipelines, data integration frameworks, and metadata-driven architectures that enable seamless data access and analytics. This role prefers deep expertise in big data processing, distributed computing, data modeling, and governance frameworks to support self-service analytics, AI-driven insights, and enterprise-wide data management.

Roles & Responsibilities:
  • Architect and maintain robust, scalable data pipelines using Databricks, Spark, and Delta Lake, enabling efficient batch and real-time processing.
  • Lead efforts to evaluate, adopt, and integrate emerging technologies and tools that enhance productivity, scalability, and data delivery capabilities.
  • Drive performance optimization efforts, including Spark tuning, resource utilization, job scheduling, and query improvements.
  • Identify and implement innovative solutions that streamline data ingestion, transformation, lineage tracking, and platform observability.
  • Build frameworks for metadata-driven data engineering, enabling reusability and consistency across pipelines.
  • Foster a culture of technical excellence, experimentation, and continuous improvement within the data engineering team.
  • Collaborate with platform, architecture, analytics, and governance teams to align platform enhancements with enterprise data strategy.
  • Define and uphold SLOs, monitoring standards, and data quality KPIs for production pipelines and infrastructure.
  • Partner with cross-functional teams to translate business needs into scalable, governed data products.
  • Mentor engineers across the team, promoting knowledge sharing and adoption of modern engineering patterns and tools.
  • Collaborate with cross-functional teams, including data architects, business analysts, and DevOps teams, to align data engineering strategies with enterprise goals.
  • Stay up to date with emerging data technologies and best practices, ensuring continuous improvement of Enterprise Data Fabric architectures.
Must-Have Skills:
  • Hands-on experience in data engineering technologies such as Databricks, PySpark, SparkSQL Apache Spark, AWS, Python, SQL, and Scaled Agile methodologies.
  • Proficiency in workflow orchestration, performance tuning on big data processing.
  • Strong understanding of AWS services
  • Experience with Data Fabric, Data Mesh, or similar enterprise-wide data architectures.
  • Ability to quickly learn, adapt and apply new technologies
  • Strong problem-solving and analytical skills
  • Excellent communication and teamwork skills
  • Experience with Scaled Agile Framework (SAFe), Agile delivery practices, and DevOps practices.
Good-to-Have Skills:
  • Good to have deep expertise in Biotech & Pharma industries
  • Experience in writing APIs to make the data available to the consumers
  • Experienced with SQL/NOSQL database, vector database for large language models
  • Experienced with data modeling and performance tuning for both OLAP and OLTP databases
  • Experienced with software engineering best-practices, including but not limited to version control (Git, Subversion, etc.), CI/CD (Jenkins, Maven etc.), automated unit testing, and Dev Ops
Education and Professional Certifications
  • 12 to 17 years of experience in Computer Science, IT or related field
  • AWS Certified Data Engineer preferred
  • Databricks Certificate preferred
  • Scaled Agile SAFe certification preferred
Soft Skills:
  • Excellent analytical and troubleshooting skills.
  • Strong verbal and written communication skills
  • Ability to work effectively with global, virtual teams
  • High degree of initiative and self-motivation.
  • Ability to manage multiple priorities successfully.
  • Team-oriented, with a focus on achieving team goals.
  • Ability to learn quickly, be organized and detail oriented.
  • Strong presentation and public speaking skills.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Lead Data Engineer / Engineering Manager
Lead Data Engineer / Engineering Manager

Gurgaon Hire • Mhalunge

Hybrid
INR 4,000,000 - 6,500,000
Senior Data Engineer
Senior Data Engineer

Amgen • Hyderabad

On-site
INR 2,500,000 - 4,500,000
Sr Data Engineer
Sr Data Engineer

Amgen • Hyderabad

On-site
INR 1,500,000 - 2,500,000
Principal Data Engineer- Hyderabad (Hybrid)
Principal Data Engineer- Hyderabad (Hybrid)

Syneos Health, Inc. • Hyderabad

On-site
INR 1,500,000 - 2,500,000
Data Engineer
Data Engineer

EXL • Gurugram District

On-site
INR 2,800,000 - 4,000,000
Data Architect
Data Architect

Infosys • Bengaluru Urban

On-site
INR 1,200,000 - 2,000,000
Cutting-edge cloud and data technologies
Collaborative culture focusing on innovation
Cross-functional project visibility
Senior Data Engineer
Senior Data Engineer

People, Jobs, and News • Maharashtra

On-site
INR 1,200,000 - 2,000,000
Data Engineering Manager
Data Engineering Manager

Good co India • India

On-site
INR 2,400,000 - 5,400,000
Senior Data Engineer
Senior Data Engineer

KSB • Pune District

On-site
INR 800,000 - 1,200,000
Principal Data Engineer
Principal Data Engineer

Nielseniq India • Pune District

On-site
INR 4,000,000 - 7,000,000
Flexible working environment
Volunteer time off
LinkedIn Learning
+1