Senior - Data Engineer

Darwinbox Digital Solutions Pvt. Ltd.

Hyderabad

On-site

INR 1,200,000 - 1,800,000

Full time

5 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Parallon is seeking a Data Engineer to support the HIM Datamart used by Parallon HIM Operations. You will design, build, and maintain ETL/ELT pipelines pulling data from Teradata, SQL Server, and GCP Datamarts into the HIM Datamart, ensuring reliable, timely data delivery.

The role emphasizes production support, data quality, automation, and documentation in a healthcare environment, using Airflow, Python, SQL, SSIS, and Google Cloud services.

Qualifications

  • Bachelor’s degree in Computer Science, Information Systems, Data Engineering, Healthcare Informatics, or related field.
  • Five or more years of data engineering, ETL/ELT, or data warehousing experience.

Responsibilities

  • Design, develop, maintain, and support ETL/ELT data pipelines for HIM GCP Datamart.
  • Develop and support Airflow DAGs and data workflows.
  • Troubleshoot production data issues and improve data quality and reliability.

Skills

Airflow
Python
SQL
GCP Dataproc
BigQuery
Teradata
SSIS
Data Warehousing

Education

Bachelor's in CS/IT/related

Tools

GitHub
Azure DevOps
Cloud Composer
Dataproc

Job description

As applicable based on functional alignment

Direct Reports:
PositionSummary:

Data Engineer to support the HIM Datamart used by ParallonHIM Operations. This critical role is responsible for creating, maintaining,monitoring, and supporting data pipelines that pull data from Teradata,multiple SQL Server environments, various GCP Datamarts, and other enterprisedata sources into the HIM GCP Datamart.

The successful candidate will design and support ETL andELT processes using Airflow, DAGs, Google Cloud Dataproc, Python, PowerShell,SSIS, SQL, and other data integration technologies. The ideal candidate mustbe comfortable supporting production data pipelines, troubleshooting complexdata issues, improving data quality, and ensuring timely and accurate dataavailability for HIM operations.

This position is best suited for a data engineeringprofessional who can balance development, production support, automation,data quality, cloud data engineering, documentation, and operationalreliability in a critical healthcare operations environment.

Responsibilities:
  • Design,develop, maintain, and support ETL/ELT data pipelines for the HIM GCPDatamart. Extract, transform, and load data from Teradata, SQL Server,GCP Datamarts, and other enterprise data sources into HIM dataplatforms.
  • Developand support data workflows using Airflow, DAGs, Dataproc, Python,PowerShell, SSIS, SQL, and related data integration tools.
  • Create,schedule, monitor, and troubleshoot Airflow DAGs for recurring dataingestion, transformation, validation, reconciliation, and deliveryprocesses.
  • Developand optimize SQL queries, stored procedures, views, functions, andtransformation logic across SQL Server, Teradata, PostgreSQL, BigQuery,and other relational platforms.
  • Supporton-premise to GCP data integration using Google Cloud Dataproc andrelated GCP services for large-scale data processing and transformation.
  • Monitordata pipelines for failures, delays, performance issues, data qualityproblems, schema changes, and source system impacts.
  • Troubleshootproduction data issues, including missing data, late-arriving files,failed jobs, transformation errors, data mismatches, and performancebottlenecks.
  • Performroot cause analysis and implement long-term fixes to improve pipelinereliability, reduce recurring failures, and improve operationalstability.
  • Implementdata validation, reconciliation, and quality checks to ensure HIM datais complete, accurate, consistent, and timely.
  • Collaboratewith Parallon HIM Operations, source system owners, and technical teamsto understand data needs, dependencies, refresh schedules, reportingimpacts, and business priorities.
  • Createand maintain data pipeline documentation, data flow diagrams,source-to-target mappings, job schedules, support procedures, andoperational runbooks.
  • Participatein release management, change control, defect resolution, enhancementdelivery, incident response, and production support activities.
  • Improveautomation, monitoring, alerting, logging, and operational visibilityacross HIM data pipelines.
  • Ensuredata engineering solutions follow enterprise standards for security,privacy, compliance, data governance, change control, and productionreadiness.
  • Identifyopportunities to modernize legacy ETL processes and improve scalability,maintainability, reliability, and cost efficiency.
Education& Experience:
  • Bachelor’sdegree in Computer Science, Information Systems, Data Engineering,Engineering, Healthcare Informatics, or a related field is required.
  • Fiveor more years of experience in data engineering, ETL/ELT development,data warehousing, data integration, or supporting enterprise dataplatforms, data marts, data lakes, or operational data stores.
  • Stronghands-on experience designing, developing, maintaining, andtroubleshooting production data pipelines using SQL, Python, PowerShell,SSIS, Apache Airflow, DAGs, and other ETL/ELT technologies.
  • StrongSQL experience across SQL Server, Teradata, PostgreSQL, BigQuery, orsimilar platforms, including complex queries, joins, aggregations,stored procedures, views, functions, performance tuning, andsource-to-target validation.
  • Experienceextracting, integrating, and migrating data from Teradata, SQL Server,file-based sources, GCP Datamarts, and other enterprise data sourcesinto GCP-based data platforms.
  • Experiencewith Google Cloud Platform data services, including Dataproc, CloudComposer, Cloud Storage, BigQuery, Cloud SQL, Cloud Logging, and CloudMonitoring.
  • Experiencewith batch processing, job scheduling, dependency management, errorhandling, logging, alerting, data validation, reconciliation, dataquality checks, and production support for critical business operations.
  • Experiencewith file-based integrations, including CSV, fixed-width files, Excel,JSON, XML, Parquet, Avro, and delimited files.
  • Experiencewith CI/CD and source control practices using GitHub, Azure DevOps, Git,YAML pipelines, or similar tools.
  • Stronganalytical, troubleshooting, problem-solving, written communication, andverbal communication skills, with the ability to work independently andcollaboratively with technical and business teams.
MustHave Skills
  • Experiencewith Spark, PySpark, Hadoop, or distributed data processing frameworks.
  • Experiencewith Teradata SQL, BTEQ, FastExport, TPT, or other Teradata dataextraction patterns.
  • Experiencedesigning scalable data ingestion frameworks and reusable pipelinecomponents.
  • Experiencewith incremental loads, full refreshes, CDC patterns, SCD handling,partitioning, and data archival strategies.
  • Experiencewith secure file transfers, file watchers, landing zones, stagingtables, and data lake patterns.
  • Experiencewith data modeling, dimensional modeling, star schemas, snowflakeschemas, and reporting-friendly data structures.
  • Experiencewith BI/reporting platforms and downstream analytics dependencies.
  • Experiencewith monitoring, alerting, operational dashboards, and data pipelineobservability.
  • Experiencewith healthcare data, HIM operations, revenue cycle, coding, clinicaldocumentation, billing, or Parallon operational data is stronglypreferred.
  • Experiencewith data privacy, HIPAA, PHI handling, role-based access, encryption,audit logging, and healthcare compliance requirements.
  • Experienceworking in Agile/Scrum or Lean delivery environments.
  • Orchestration& Processing: Apache Airflow, DAGs, Cloud Composer, Google CloudDataproc, PySpark, Batch Processing, Job Scheduling
  • ETL/ELT& Integration: ETL, ELT, SSIS, File-Based Integrations,Source-to-Target Mapping, Data Validation, Data Reconciliation, DataQuality Checks
  • Scripting& Automation: Python, PowerShell, YAML
  • File/DataFormats: JSON, XML, CSV, Parquet
  • DevOps& Support: GitHub, Azure DevOps, CI/CD Pipelines, Cloud Logging,Cloud Monitoring, Monitoring and Alerting, Production Support
Niceto Have Skills
  • Fiveor more years of relevant data engineering, ETL development, or datawarehousing experience.
  • Hands-onexperience with SQL, Python, Airflow, DAGs, SQL Server, and ETLdevelopment is required.
  • Experiencewith Teradata and SSIS is strongly preferred.
  • Experiencewith Google Cloud Platform, Dataproc, Cloud Composer, Cloud Storage,BigQuery, or related GCP data services is preferred.
  • Experiencesupporting production data pipelines and critical business operations isrequired.
  • Experiencewith data validation, reconciliation, monitoring, alerting, andproduction troubleshooting is required.
  • HealthcareIT, HIM, revenue cycle, or Parallon operational data experience ispreferred.
Licenses,Certifications & Training:
  • N/A
Knowledge,Skills, Abilities, Behaviors:
  • Strongownership mindset with the ability to support critical operational dataprocesses.
  • Abilityto understand business workflows and translate them into reliable dataengineering solutions.
  • Abilityto communicate data issues, risks, dependencies, and resolution plansclearly to technical and non-technical stakeholders.
  • Strongattention to detail, especially when working with operational,financial, healthcare, or compliance-sensitive data.
  • Abilityto troubleshoot complex data issues across source systems, ETLprocesses, cloud platforms, databases, and downstream consumers.
  • Abilityto prioritize production support issues based on business impact andoperational urgency.
  • Abilityto create clear documentation, data flow diagrams, source-to-targetmappings, and support runbooks.
  • Abilityto work effectively with business teams, source system owners, cloudteams, database teams, reporting teams, and application support teams.
  • Self-motivatedlearner who stays current with data engineering, cloud data platforms,orchestration tools, and automation practices.
  • Abilityto work with minimal supervision while managing multiple priorities.
  • Collaborativeteam player with strong analytical and problem-solving skills.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Urgently Required | Sr. Data Warehouse Engineer | WFO Only
Urgently Required | Sr. Data Warehouse Engineer | WFO Only

Rxadvance Pbm • Dadri

On-site
INR 3,000,000 - 5,400,000
Group Data Engineer I
Group Data Engineer I

The Peninsular and Oriental Steam Navigation Company • Bengaluru

On-site
INR 2,500,000 - 4,000,000
Assistant Manager
Assistant Manager

Jobtailor • Gurugram District

On-site
INR 1,200,000 - 1,800,000
Data Engineer
Data Engineer

Vriba Solutions • Bengaluru

On-site
INR 1,500,000 - 2,800,000
Data Engineer
Data Engineer

fluid.live • Chennai District

On-site
INR 1,200,000 - 1,800,000
Data Engineer
Data Engineer

Hudson Data • India

On-site
INR 1,800,000 - 3,400,000
Data Engineer
Data Engineer

Altraize • Bengaluru

On-site
INR 2,800,000 - 4,200,000
Data Engineer
Data Engineer

Valuenode Private Limited • India

Remote
INR 1,200,000 - 1,800,000
Senior GCP Data Engineer
Senior GCP Data Engineer

Tata Consultancy Services • Bengaluru

On-site
INR 4,000,000 - 7,000,000
Senior /Staff Data & Cloud Platform Engineer
Senior /Staff Data & Cloud Platform Engineer

Vibehackers • Hyderabad

On-site
INR 3,600,000 - 6,000,000