Technical Specialist

HCL Technologies Limited

Ciudad de México

Presencial

MXN 540.000 - 720.000

Jornada completa

Hace 4 días
Sé de los primeros/as/es en solicitar esta vacante

Recibe más respuestas de empleadores

Envía un currículum específico para el puesto de trabajo en cuestión de minutos.

Descripción de la vacante

HCLTech seeks a data engineering expert to design, build, and optimize large-scale data pipelines on Hadoop and cloud platforms like Databricks and Snowflake. You will work with Python, Spark, Hive, and Airflow to deliver robust data products and scalable ETL/ELT workflows.

You will collaborate with Product Managers to translate user needs into technical modules and contribute to DevOps practices to ensure reliable deployments.

Formación

  • Production-grade pipelines on Hadoop and cloud data platforms.
  • Strong Python, Spark, Hadoop, SQL skills.
  • Experience with databricks and snowflake is preferred.
  • Experience with orchestration tools like Airflow or NiFi.
  • DevOps/CI-CD practices: Git, automated testing, release pipelines.

Responsabilidades

  • Work with cutting-edge big data platforms at large scale to enable data processing and model development.
  • Build and maintain ETL/ELT pipelines for ingestion, transformation, and aggregation of large-scale datasets on Hadoop and enterprise data platforms.
  • Develop high-performance data processing jobs using PySpark/Spark on Cloudera and Databricks.
  • Optimize pipeline performance and cost through partitioning and tuning.
  • Contribute to CI/CD for data workflows and promote engineering best practices.
  • Partner with Product Managers to understand users and use cases and scope new modules.

Conocimientos

Data engineering
Python
Spark
Hadoop
SQL
DevOps/CI-CD

Herramientas

Hive
Impala
Airflow
NiFi
Talend

Descripción del empleo

Work with cutting-edge big data platforms (e.g., Databricks, Apache Spark) at large scale, pushing the boundaries of data processing and model enablement. Build and maintain robust ETL/ELT pipelines for ingestion, transformation, and aggregation of large-scale datasets on Hadoop and enterprise data platforms. Develop high-performance data processing jobs using PySpark/Spark, Python on data platforms such as cloudera and databricks. Optimize pipeline performance and cost through partitioning, file formats, compute tuning, and efficient query patterns. Contribute to CI/CD for data workflows (testing, code reviews, deployment automation), promoting engineering best practices and maintainable codebases. Partner with Product Managers to develop a deep understanding of users and use cases and apply that knowledge to scoping and building new modules and features Ideal Candidate Qualifications: Strong hands-on experience in data engineering building production-grade pipelines on big data platforms (Hadoop ecosystem and cloud data platforms - databricks). High proficiency in using Python, Spark, Hadoop platforms & tools (Hive, Impala, Airflow, NiFi), SQL to build Big Data products. Hands-on experience with cloud data platforms such as databricks, snowflake (databricks preferred). Experience with orchestration/integration tools such as Apache Airflow, Apache NiFi, or Talend. Working knowledge of DevOps/CI-CD practices: version control (Git), automated testing, release pipelines, and observability. Strong problem-solving skills with the ability to debug complex data issues and communicate clearly with technical and non-technical stakeholders. Experience developing Java based applications is an added advantage.

Key Responsibilities
  • Work with cutting-edge big data platforms (e.g., Databricks, Apache Spark) at large scale, pushing the boundaries of data processing and model enablement.
  • Build and maintain robust ETL/ELT pipelines for ingestion, transformation, and aggregation of large-scale datasets on Hadoop and enterprise data platforms.
  • Develop high-performance data processing jobs using PySpark/Spark, Python on data platforms such as cloudera and databricks.
  • Optimize pipeline performance and cost through partitioning, file formats, compute tuning, and efficient query patterns
  • Contribute to CI/CD for data workflows (testing, code reviews, deployment automation), promoting engineering best practices and maintainable codebases.
  • Partner with Product Managers to develop a deep understanding of users and use cases and apply that knowledge to scoping and building new modules and features
  • Ideal Candidate Qualifications:
  • Strong hands-on experience in data engineering building production-grade pipelines on big data platforms (Hadoop ecosystem and cloud data platforms - databricks).
  • High proficiency in using Python, Spark, Hadoop platforms & tools (Hive, Impala, Airflow, NiFi), SQL to build Big Data products.
  • Hands-on experience with cloud data platforms such as databricks, snowflake (databricks preferred)
  • Experience with orchestration/integration tools such as Apache Airflow, Apache NiFi, or Talend.
  • Working knowledge of DevOps/CI-CD practices: version control (Git), automated testing, release pipelines, and observability.
  • Strong problem-solving skills with the ability to debug complex data issues and communicate clearly with technical and non-technical stakeholders.
  • Experience developing Java based applications is an added advantage.
Skill Requirements
  • Work with cutting-edge big data platforms (e.g., Databricks, Apache Spark) at large scale, pushing the boundaries of data processing and model enablement.
  • Build and maintain robust ETL/ELT pipelines for ingestion, transformation, and aggregation of large-scale datasets on Hadoop and enterprise data platforms.
  • Develop high-performance data processing jobs using PySpark/Spark, Python on data platforms such as cloudera and databricks.
  • Optimize pipeline performance and cost through partitioning, file formats, compute tuning, and efficient query patterns
  • Contribute to CI/CD for data workflows (testing, code reviews, deployment automation), promoting engineering best practices and maintainable codebases.
  • Partner with Product Managers to develop a deep understanding of users and use cases and apply that knowledge to scoping and building new modules and features
  • Ideal Candidate Qualifications:
  • Strong hands-on experience in data engineering building production-grade pipelines on big data platforms (Hadoop ecosystem and cloud data platforms - databricks).
  • High proficiency in using Python, Spark, Hadoop platforms & tools (Hive, Impala, Airflow, NiFi), SQL to build Big Data products.
  • Hands-on experience with cloud data platforms such as databricks, snowflake (databricks preferred)
  • Experience with orchestration/integration tools such as Apache Airflow, Apache NiFi, or Talend.
  • Working knowledge of DevOps/CI-CD practices: version control (Git), automated testing, release pipelines, and observability.
  • Strong problem-solving skills with the ability to debug complex data issues and communicate clearly with technical and non-technical stakeholders.
  • Experience developing Java based applications is an added advantage.
Other Requirements
  • Work with cutting-edge big data platforms (e.g., Databricks, Apache Spark) at large scale, pushing the boundaries of data processing and model enablement.
  • Build and maintain robust ETL/ELT pipelines for ingestion, transformation, and aggregation of large-scale datasets on Hadoop and enterprise data platforms.
  • Develop high-performance data processing jobs using PySpark/Spark, Python on data platforms such as cloudera and databricks.
  • Optimize pipeline performance and cost through partitioning, file formats, compute tuning, and efficient query patterns
  • Contribute to CI/CD for data workflows (testing, code reviews, deployment automation), promoting engineering best practices and maintainable codebases.
  • Partner with Product Managers to develop a deep understanding of users and use cases and apply that knowledge to scoping and building new modules and features
  • Ideal Candidate Qualifications:
  • Strong hands-on experience in data engineering building production-grade pipelines on big data platforms (Hadoop ecosystem and cloud data platforms - databricks).
  • High proficiency in using Python, Spark, Hadoop platforms & tools (Hive, Impala, Airflow, NiFi), SQL to build Big Data products.
  • Hands-on experience with cloud data platforms such as databricks, snowflake (databricks preferred)
  • Experience with orchestration/integration tools such as Apache Airflow, Apache NiFi, or Talend.
  • Working knowledge of DevOps/CI-CD practices: version control (Git), automated testing, release pipelines, and observability.
  • Strong problem-solving skills with the ability to debug complex data issues and communicate clearly with technical and non-technical stakeholders.
  • Experience developing Java based applications is an added advantage.

At HCLTech, you'll supercharge your potential. You'll find your career. And you'll find your spark. All at a place that knows that helping its customers stay on top starts by putting its people first.

HCLTech is a global technology company, home to more than 223,000 people across 60 countries, delivering industry-leading capabilities centered around digital, engineering, cloud and AI, powered by a broad portfolio of technology services and products. We work with clients across all major verticals, providing industry solutions for Financial Services, Manufacturing, Life Sciences and Healthcare, Technology and Services, Telecom and Media, Retail and CPG, and Public Services. Consolidated revenues as of 12 months ending June 2026totaled $14.8billion.

Consigue la evaluación confidencial y gratuita de tu currículum.
o arrastra y suelta tu archivo aquí
Similar jobs

Puestos de trabajo similares que vale la pena comparar

Cloud Data Engineer
Cloud Data Engineer

HCLTech • Estado de México

Presencial
MXN 1.269.000 - 1.632.000
Big Data Engineer: Spark, PySpark & ETL Mastery
Big Data Engineer: Spark, PySpark & ETL Mastery

HCL Technologies Limited • Ciudad de México

Presencial
MXN 540.000 - 720.000
Senior Data Architect - Databricks Platform New MEXICO
Senior Data Architect - Databricks Platform New MEXICO

Caylent • México

A distancia
USD 140.000 - 210.000
Medical Insurance for you and eligible
Generous holidays and flexible PTO
Paid for exams and certifications
+5
Data Engineer
Data Engineer

Pyramid Consulting, Inc • Región Centro

Presencial
MXN 900.000 - 1.500.000
Senior Databricks Data Engineer ID86297
Senior Databricks Data Engineer ID86297

AgileEngine, LLC. • Región Centro

A distancia
MXN 600.000 - 1.200.000
Growth without limits
Competitive compensation
Flexibility
+3
Senior Full Stack Developer
Senior Full Stack Developer

HCL Technologies Limited • Región Centro

Presencial
MXN 420.000 - 700.000
Data Engineer
Data Engineer

Capgemini Engineering • Aguascalientes

Presencial
MXN 420.000 - 660.000
Data Engineer Sr
Data Engineer Sr

Turtle Trax S.A. • Región Centro

Híbrido
MXN 800.000 - 1.100.000
Remote Data Engineer: Spark, Databricks & ETL Pipelines
Remote Data Engineer: Spark, Databricks & ETL Pipelines

Capgemini Engineering • Aguascalientes

Presencial
MXN 420.000 - 660.000
Database Engineer
Database Engineer

HCLTech • México

A distancia
MXN 500.000 - 750.000
Life insurance
Major Medical Expenses Insurance
Minor Medical Expense Insurance
+2