Data Engineer II - Databricks, Azure & Spark

Cogent IBS, Inc

United States

Remote

USD 120,000 - 150,000

Full time

5 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Cogent IBS, Inc. seeks a Data Scientist (Big Data Engineer) II to design and optimize scalable data pipelines with Apache Spark on Databricks, and to integrate with Azure services.

You will own ETL/ELT workflows, data models, governance, security, and CI/CD deployments, collaborating with data scientists and analysts in Agile, multicultural teams. Proficiency in Python and SQL, plus Azure Data Lake/Delta Lake, is required.

Qualifications

  • 4+ years ETL/ELT workflows for structured and unstructured data.
  • 4+ years automating deployments using CI/CD tools.
  • 4+ years collaborating with data scientists, analysts, stakeholders, and cross-functional teams.
  • 4+ years designing and maintaining data models, schemas, and database structures.
  • 4+ years working with data storage solutions, including Azure Data Lake Storage and data warehouses.
  • 4+ years implementing data validation and data quality checks.
  • 4+ years contributing to data governance, metadata management, data lineage, and data cataloging.
  • 4+ years implementing data security measures, including encryption, access controls, and auditing.
  • 4+ years of proficiency in Python and R programming languages.
  • 4+ years of strong SQL querying and data manipulation experience.
  • 4+ years of experience with the Microsoft Azure cloud platform.
  • 4+ years of experience with DevOps, CI/CD pipelines, and version control systems.
  • 4+ years working in Agile and multicultural environments.
  • 4+ years of strong troubleshooting and debugging capabilities.
  • 3+ years designing and developing scalable data pipelines using Apache Spark on Databricks.
  • 3+ years optimizing Spark jobs for performance and cost efficiency.
  • 3+ years integrating Databricks with Azure Data Factory.
  • 3+ years ensuring data quality, governance, and security using Unity Catalog or Delta Lake.
  • 3+ years of strong understanding of Apache Spark architecture, RDDs, DataFrames, and Spark SQL.
  • 3+ years of hands-on experience with Databricks notebooks, clusters, jobs, and Delta Lake.

Responsibilities

  • Design, develop, and maintain scalable data pipelines using Apache Spark on Databricks.
  • Implement ETL/ELT workflows for structured and unstructured data.
  • Develop and optimize Spark jobs for performance and cost efficiency.
  • Build and maintain data models, schemas, and database structures supporting analytical and operational use cases.
  • Integrate Databricks solutions with Azure Data Factory and other Azure cloud services.
  • Work with Azure Data Lake Storage and data warehouse solutions.
  • Implement data validation and quality checks to ensure data accuracy, consistency, and reliability.
  • Contribute to data governance initiatives, including metadata management, data lineage, and data cataloging.
  • Implement data security measures, including encryption, access controls, and auditing.
  • Support compliance with applicable regulations, security requirements, and industry best practices.
  • Automate deployments using CI/CD pipelines, DevOps practices, and version control systems.
  • Work with Databricks notebooks, clusters, jobs, and Delta Lake.
  • Utilize Unity Catalog and/or Delta Lake to support data quality, governance, and security.
  • Troubleshoot and debug data pipelines, Spark applications, and related technical issues.
  • Collaborate with data scientists, data analysts, stakeholders, and cross-functional teams.
  • Work effectively within Agile and multicultural environments.

Skills

Python
SQL
Spark
Databricks
Azure
CI/CD
Data Modeling
Delta Lake
Data Governance
Unix/Linux

Tools

Azure Data Factory
Azure Data Lake Storage
Delta Lake
Unity Catalog

Job description

Cogent IBS, Inc. seeks a Data Scientist (Big Data Engineer) II to design and optimize scalable data pipelines with Apache Spark on Databricks, and to integrate with Azure services.

You will own ETL/ELT workflows, data models, governance, security, and CI/CD deployments, collaborating with data scientists and analysts in Agile, multicultural teams. Proficiency in Python and SQL, plus Azure Data Lake/Delta Lake, is required.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Data Scientist II - Big Data Engineer
Data Scientist II - Big Data Engineer

Cogent IBS, Inc • United States

Remote
USD 120,000 - 150,000
Senior Data Engineer - Databricks/PySpark on Azure
Senior Data Engineer - Databricks/PySpark on Azure

EPAM Systems Inc • United States

Remote
USD 130,000 - 180,000
Data Engineer: Azure Spark & Streaming Pipelines
Data Engineer: Azure Spark & Streaming Pipelines

Software Technology Inc • Brentsville (KY)

On-site
USD 80,000 - 120,000
Data Engineer II — AWS & Databricks Pipelines
Data Engineer II — AWS & Databricks Pipelines

Travelers • Atlanta (GA)

On-site
USD 127,000 - 209,000
Health Insurance
401(k) match
Paid Time Off
+1
Databricks Data Engineer | PySpark, Delta Lake, Azure
Databricks Data Engineer | PySpark, Delta Lake, Azure

BuzzClan LLC • Beaverton (OR)

On-site
USD 120,000 - 170,000
Databricks Data Engineer
Databricks Data Engineer

Adastra Corp • United States

Remote
USD 120,000 - 160,000
Azure Data Architect with Databricks
Azure Data Architect with Databricks

Arkhya Tech. Inc. • New Jersey

On-site
USD 140,000 - 190,000
Senior Data Engineer: Databricks, Spark & Azure
Senior Data Engineer: Databricks, Spark & Azure

CI&T • United States

Remote
USD 100,000 - 140,000
Health insurance
Dental insurance
Meal allowance
+8
Azure Databricks Lead: Scalable Data Pipelines & DevOps
Azure Databricks Lead: Scalable Data Pipelines & DevOps

Capgemini • Jersey City (NJ)

On-site
USD 103,330 - 128,656
Paid time off
Medical, dental, and vision coverage
Retirement savings plans
Senior Data Engineer: Databricks, Spark & Azure Lakehouse
Senior Data Engineer: Databricks, Spark & Azure Lakehouse

Enfint • United States

Remote
USD 150,000 - 210,000
Annual holiday
Private healthcare insurance
Dental support
+4