A leading IT services firm in Baltimore is seeking an experienced Data Engineer specializing in the design and optimization of data pipelines. You will collaborate with Data Scientists and Analysts to ensure the delivery of quality data solutions. The ideal candidate should have at least 5 years of experience in data engineering and proficiency in cloud-based technologies. This full-time position requires strong problem-solving and communication skills.
Qualifications
5+ years of experience in data engineering, with a focus on building and maintaining data pipelines.
1-2 years experience with building NiFI data flows or similar for Kafka and Hadoop-based NoSQL databases.
Experience with API-led design.
Strong problem-solving skills and attention to detail.
Responsibilities
Design, develop, and optimize end-to-end data pipelines for efficient data ETL processes.
Collaborate with cross-functional teams to understand data requirements.
Implement best practices for data integration.
Skills
Data pipeline development
ETL processes
Sql
Cloud technologies
Data validation and cleansing
Education
Bachelor’s degree in Computer Science, Engineering, or a related field
Tools
Apache NiFi
AWS S3
Redshift
Google BigQuery
Python
Scala
ELK stack
Job description
Design, develop, and optimize end-to-end data pipelines for efficient data extraction, transformation, and loading (ETL) processes.
Collaborate with cross-functional teams to understand data requirements and translate them into scalable pipeline solutions.
Implement best practices for data integration, ensuring high performance, reliability, and scalability.
Data Transformation and Quality
Transform raw data into usable formats, ensuring data quality, consistency, and accuracy.
Handle data validation, cleansing, and error handling to maintain data integrity.
Monitor and proactively maintain data pipelines to ensure high service availability.
Performance Optimization
Continuously improve pipeline performance by identifying bottlenecks and implementing optimizations.
Work with cloud-based technologies (e.g., AWS, GCP, Azure) to enhance scalability and efficiency.
Collaboration and Leadership
Partner with Data Scientists, Analysts, and other stakeholders to understand their data needs.
Lead discussions on system enhancements, process improvements, and data governance.
Mentor junior engineers and contribute to the growth of the data engineering team.
Requirements
Bachelor’s degree in Computer Science, Engineering, or a related field.
5+ years of experience in data engineering, with a focus on building and maintaining data pipelines.
1-2 years experience with building NiFI data flows or similar for Kafka and Hadoop-based NoSQL databases.
Proficiency in ETL tools, SQL, and scripting languages (Python, Scala, etc.).
Experience with Data Catalog and Accumulo indexes for information retrieval and discovery.
Experience with API-led design.
Experience with ELK stack a plus.
Experience with cloud-based data platforms (e.g., AWS S3, Redshift, Google BigQuery).
Strong problem-solving skills and attention to detail.
Excellent communication and collaboration abilities.
Security+ Certification
Active US Government Clearance at Secret level or higher
Effective written and verbal communications skills for collaboration with both customers and fellow team members.
Ability to sit for extended periods of time.
Ability to regularly lift at least 25 pounds.
Ability to commute to the designated onsite work location as required.