Data Engineer-Data Platforms-Google - 1

IBM

San Francisco (CA)

A distancia

USD 120.000 - 180.000

Jornada completa

hace 31 horas
Sé de los primeros/as/es en solicitar esta vacante
Generador de candidaturas

No envíes un currículum genérico — crea un currículum y una carta de presentación adaptados a este puesto concreto.

Supera los filtros ATS

Descripción de la vacante

IBM Consulting seeks an experienced Data Engineer to design, build, and maintain data pipelines on Google's Cloud ecosystem. You will leverage DataProc, DataFlow, Pub/Sub, BigQuery and BigTable to deliver scalable data solutions.

You will work with Apache Beam, Airflow, dbt, Spark/Python, and related tech to ensure robust batch and real-time processing, hosted from anywhere in the United States. This role emphasizes data accessibility, security, and performance.

Formación

  • Deep expertise designing, building and maintaining data engineering on Google Cloud.
  • Proficient with open-source tools like Apache Beam, Airflow, dbt and Spark.
  • Experience with batch and real-time data pipelines for Data Warehouse/Data Lake.
  • Ability to schedule and manage data platforms with Cloud Scheduler and Airflow.
  • Design and optimize data layers for efficient migration and processing.

Responsabilidades

  • Design batch and real-time data pipelines for Data Warehouse and Datalake on Google Cloud.
  • Develop and maintain data engineering solutions using Google Cloud Storage, BigQuery, DataProc, DataFlow, and Spark/Python.
  • Manage data platforms with Cloud Scheduler and Cloud Composer (Airflow).
  • Optimize data layers for efficient data migration and processing.
  • Ensure pipelines scale to meet business needs.

Conocimientos

Google Cloud data platforms
Open-source technologies
Batch and real-time pipelines
Data platform management
Data layer optimization
Apache Beam
dbt and data modeling

Herramientas

DataProc
DataFlow
Pub/Sub
BigQuery
BigTable
Cloud Spanner
CloudSQL
AlloyDB
Apache Airflow
dbt
Spark

Descripción del empleo

Introduction

A career in IBM Consulting is built on long-term client relationships and close collaboration worldwide. You'll work with leading companies across industries, helping them shape their hybrid cloud and AI journeys. With support from our strategic partners, robust IBM technology, and Red Hat, you'll have the tools to drive meaningful change and accelerate client impact. At IBM Consulting, curiosity fuels success. You'll be encouraged to challenge the norm, explore new ideas, and create innovative solutions that deliver real results. Our culture of growth and empathy focuses on your long-term career development while valuing your unique skills and experiences.

Introduction

A career in IBM Consulting is built on long-term client relationships and close collaboration worldwide. You'll work with leading companies across industries, helping them shape their hybrid cloud and AI journeys. With support from our strategic partners, robust IBM technology, and Red Hat, you'll have the tools to drive meaningful change and accelerate client impact. At IBM Consulting, curiosity fuels success. You'll be encouraged to challenge the norm, explore new ideas, and create innovative solutions that deliver real results. Our culture of growth and empathy focuses on your long-term career development while valuing your unique skills and experiences.

Your Role And Responsibilities

As a seasoned Data Engineer specializing in Google's data platforms, you will design, build, and maintain data engineering solutions on Google's Cloud ecosystem. You will utilize your expertise in Google's services and open-source technologies to deliver scalable and efficient data pipelines.

Your Primary Responsibilities Will Include
  • Design Data Pipelines: Design and develop batch and real-time data pipelines for Data Warehouse and Datalake using Google Cloud services such as DataProc, DataFlow, PubSub, BigQuery, and Big Table.
  • Develop Data Engineering Solutions: Build and maintain data engineering solutions using Google Cloud Storage, BigTable, BigQuery DataProc with Spark and Hadoop, Google DataFlow with Apache Beam or Python, and other open-source technologies like Apache Airflow, dbt, Spark/Python, or Spark/Scala.
  • Manage Data Platforms: Schedule and manage the data platform using Google Cloud Scheduler and Cloud Composer (Airflow), ensuring seamless data pipeline operations.
  • Optimize Data Layer: Design and optimize the data layer for efficient data migration and data processing using Google Cloud services.
  • Ensure scalability and efficiency of data pipelines and data engineering solutions to meet business needs.
This role can be performed from anywhere in the United States of America
Required Technical And Professional Expertise
  • Deep Expertise in Google Data Platforms: Proven experience designing, building, and maintaining data engineering solutions on Google's Cloud ecosystem, including Google DataProc, DataFlow, PubSub, BigQuery, Big Table, Cloud Spanner, CloudSQL, and AlloyDB.
  • Proficiency in Open-Source Technologies: Experience with Apache Beam, Apache Airflow, dbt, Spark/Python, or Spark/Scala, and ability to integrate these technologies with Google Cloud services.
  • Batch and Real-Time Data Pipelines: Experience developing and managing batch and real-time data pipelines for Data Warehouse and Datalake using Google Cloud services.
  • Data Platform Management: Experience scheduling and managing data platforms using Google Cloud Scheduler and Cloud Composer (Airflow).
  • Data Layer Optimization: Experience designing and optimizing data layers for efficient data migration and data processing using Google Cloud services.
Preferred Technical And Professional Experience
  • Advanced Apache Beam Knowledge: Experience with Apache Beam, including integrating it with Google Cloud services such as DataFlow, is highly valued. Ability to optimize Beam pipelines for scalability and efficiency is a plus.
  • dbt and Data Modeling: Familiarity with dbt and data modeling concepts, including data warehousing and data lake architecture, is beneficial for designing and optimizing data layers.
  • Spark and Scala Expertise: Proficiency in Spark and Scala, including integrating them with Google Cloud services such as DataProc, is desirable for building and maintaining data engineering solutions.
Consigue la evaluación confidencial y gratuita de tu currículum.

o arrastra y suelta tu archivo aquí

Similar jobs

Puestos de trabajo similares que vale la pena comparar

Data Engineer-Data Platforms-Google - 1
Data Engineer-Data Platforms-Google - 1

IBM • New York (NY)

A distancia
USD 120.000 - 160.000
Data Engineer-Data Platforms-Google - 2
Data Engineer-Data Platforms-Google - 2

IBM • San Francisco (CA)

A distancia
USD 130.000 - 170.000
Data Engineer-Data Platforms-Google - 4
Data Engineer-Data Platforms-Google - 4

IBM • New York (NY)

A distancia
USD 120.000 - 180.000
Data Engineer-Data Platforms-Google - 5
Data Engineer-Data Platforms-Google - 5

IBM • New York (NY)

A distancia
USD 140.000 - 180.000
Senior Data Engineer - Google Cloud & Real-Time Pipelines
Senior Data Engineer - Google Cloud & Real-Time Pipelines

IBM • San Francisco (CA)

A distancia
USD 130.000 - 170.000
Remote Data Engineer, Google Cloud & Real-Time Pipelines
Remote Data Engineer, Google Cloud & Real-Time Pipelines

IBM • Boston (MA)

Híbrido
USD 215.000 - 253.000
Data Engineer – Google Cloud Real-Time Pipelines (Remote)
Data Engineer – Google Cloud Real-Time Pipelines (Remote)

IBM • Boston (MA)

Híbrido
USD 215.000 - 253.000
Health benefits
401(k)
Paid time off
+2
Cloud Data Engineer - Google Cloud Pipelines
Cloud Data Engineer - Google Cloud Pipelines

IBM • Boston (MA)

A distancia
USD 131.000 - 154.000
Google Cloud Data Engineer — Remote (US)
Google Cloud Data Engineer — Remote (US)

IBM • San Francisco (CA)

A distancia
USD 120.000 - 180.000
Remote Google Data Engineer - BigQuery & Dataflow Pro
Remote Google Data Engineer - BigQuery & Dataflow Pro

IBM • New York (NY)

A distancia
USD 120.000 - 160.000