Data Operations Engineer

Inclusion Cloud

Washington (District of Columbia)

On-site

USD 80,000 - 100,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

A data solutions provider in Washington, DC is seeking a skilled Data Operations Engineer to manage and monitor ETL data pipelines. In this role, you will improve data integrity and reliability while collaborating with teams across the organization. Candidates should hold a B.S. in computer science or possess 5+ years of relevant experience, with strong skills in SQL, Python, and data visualization tools. This is an excellent opportunity for detail-oriented individuals passionate about data optimization.

Qualifications

  • Strong analytical and critical thinking skills.
  • Proficient in shell scripting, Python, SQL.
  • Experience with data visualization tools like Tableau.

Responsibilities

  • Monitor and maintain data ETL pipelines.
  • Analyze data workflows for reliability improvements.
  • Facilitate onboarding of new data products.

Skills

Analytical thinking
Problem-solving skills
Shell scripting
Python
Spark/PySpark
SQL
HQL
HDFS
Data visualization
CI/CD principles

Education

B.S. in computer science or information systems
5+ years related work experience

Tools

GitHub
Tableau
Big Data Studio
AWS
GCP
Azure

Job description

Washington, District of Columbia, United States

  • B.S. in computer science or information systems fields required, or 5+ years related work experience.
  • Strong analytical, critical thinking skills used to solve complex problems
  • Strong technical background with a mix of development and automation skills
  • Outstanding attention to detail and consistently meets deadlines
  • Exceptional communication and interpersonal skills
  • Ability to work alongside a highly collaborative team, but also a self-starter, able to work independently with little guidance
  • Experience in troubleshooting, performance tuning, and optimization
  • Proficient in shell scripting, Python, Scala or other programming languages
  • Knowledge of Spark/PySpark
  • Excellent SQL knowledge, ability to read/write SQL queries
  • Skilled in Hive (HQL) and HDFS
  • Experience working with both unstructured and structured data sets, including flat files, JSON, XML, ORC, Parquet and AVRO
  • Comfortable working with big data environments and dealing with large diverse data sets
  • Familiarity with source code management/versioning tools such as Github
  • Understanding of CI/CD principles and best practices in data processing
  • Experience building data visualization dashboards to capture data quality metrics using tools like Tableau, Big Data Studio
  • Understanding of public cloud technologies such as AWS, GCP and Azure is a plus
MAJOR JOB RESPONSIBILITIES
  • Contribute to the maintenance, documentation, and monitoring of supported data pipelines
  • Continuously analyze supported data workflows for opportunities to improve reliability and timeliness against established SLAs
  • Conceive, develop, and apply improvements to workflows and monitoring to minimize the occurrence and impact of defects
  • Communicate with stakeholders when data is in error or is delayed, with clear plans and timelines for recovery, and future prevention
  • Develop modifications to workflows using git and github
  • Assist with production support tickets and inquiries from consumers of supported data pipelines
  • Facilitate the onboarding of new products and pipelines into our suite of supported production processes
ABOUT YOU

You are passionate about improving the integrity, accuracy and reliability of data across the organization. You are a highly motivated individual with excellent analytical, critical thinking and problem solving skills. You bring substantial value to the team with your prior experience in building and supporting production data and reporting pipelines. Strong verbal, written and interpersonal skills provide the flexibility to work collaboratively with a team or independently with minimal supervision. Youre a detail-oriented, self-starter with the ability to multitask and thrive in a dynamic environment. Furthermore, your familiarity with techniques for automating, cleansing and standardizing data at rest, and in motion, make you a great fit for this role.

WHAT YOUD BE DOING

As a Data Operations Engineer you will be responsible for monitoring and managing maintenance of multiple data ETL pipelines, which power hundreds to thousands of business-critical applications and reports, used by many teams throughout the organization, as well as external customers, every day. Due to the scale, variety, and complexity of the processes we support, standardized or automated practices and tools are needed for monitoring and maintenance of the pipelines. It is your duty to ensure that monitoring alerts Data Operations to any issues with data pipelines, with appropriate timeliness and sensitivity. Any alerts must be dealt with in a timely and appropriate manner. This can include a variety of things, including job regeneration, communication to stakeholders, definition and development of process enhancements via code, or updates to process documentation. In addition, you will assist with production support inquiries from stakeholders about the quality or timeliness of data in reporting. Finally, you will work with other teams to facilitate the onboarding of new data pipelines and products into our suite of supported production processes. Strong analytic skills and problem solving skills are needed throughout to model, monitor, and troubleshoot the production processes.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Data Engineer
Senior Data Engineer

Novatalent • United States

On-site
USD 110,000 - 160,000
Data Operations Analyst
Data Operations Analyst

Aplaro Ltd • New York (NY)

On-site
USD 75,000 - 110,000
DATA ENGINEER
DATA ENGINEER

Black Financial Consulting Group • Snowflake (AZ)

On-site
USD 120,000 - 160,000
Senior Data Engineer
Senior Data Engineer

Inizio Partners Corp • San Francisco (CA)

On-site
USD 120,000 - 150,000
Data Engineer
Data Engineer

Jobtailor • Town of Florida (NY)

On-site
USD 120,000 - 180,000
Data Engineer
Data Engineer

Aptdata Solutions Inc. • Farmington Hills (MI)

On-site
USD 90,000 - 115,000
Data Engineer
Data Engineer

Compunnel, Inc. • Columbus (OH)

On-site
USD 85,000 - 115,000
Data Engineer
Data Engineer

Arsenault • San Diego (CA)

On-site
USD 80,000 - 120,000
Senior Data Engineer – 8+ Years Experience
Senior Data Engineer – 8+ Years Experience

Hudson Manpower • New Jersey

On-site
USD 140,000 - 190,000
Senior Data Engineer – 8+ Years Experience
Senior Data Engineer – 8+ Years Experience

Hudson Manpower • New York (NY)

On-site
USD 140,000 - 190,000