Bigdata Engineer

Disys - Oak Brook

Tampa (FL)

On-site

USD 90,000 - 120,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

A technology consulting firm based in Florida is seeking a qualified candidate for a critical role involving the design of data pipelines and the deployment of machine learning models. Ideal applicants should have strong skills with Apache Spark and SCALA, along with experience in AWS services. This role requires collaboration across engineering and data science teams to build scalable data solutions. Suitable candidates will have a graduate degree in Computer Science or a related field and a passion for innovative problem-solving.

Qualifications

  • 2+ years of experience in data ingesting, cleansing, and processing.
  • Hands-on experience with model deployment and application deployment.
  • Experience in building stable, scalable, and high-speed live streams.

Responsibilities

  • Design interfaces to the data warehouses and machine learning applications.
  • Interface with various teams to ensure data pipelines fit within the production framework.
  • Deploy the machine learning model and serve outputs as RESTful API calls.

Skills

Apache Spark (Spark SQL)
SCALA
Java
AWS Big Data components
Docker
Kubernetes
Linux scripting
ETL processes
Apache KAFKA
Hadoop
AGILE development

Education

Graduation (MS or Undergraduate) in Computer Science/Engineering/relevant field

Tools

AWS
Docker
Kubernetes
Apache Spark
MapReduce
Hive
HBase

Job description

  • Design interfaces to the data warehouses/data storages and machine learning/Big Data applications using open source tools such as Scala, Java, Python, Perl and shell scripting.
  • Design and create data pipelines to maintain stable dataflow to the machine learning models – both in batch mode and near real-time mode.
  • Interface with Engineering/Operations/System Admin/Data Scientist teams to ensure data pipelines and processes fit within the production framework.
  • Ensure that tools and environments adhere to strict security protocols.
  • Deploy the machine learning model and serve its outputs as RESTful API calls.
  • Understand the business needs in close collaborations with subject matter experts (SMEs) and Data Scientists to do efficient feature engineering for machine learning models.
  • Maintain the code and libraries in code repository.
  • Work with system administration team to proactively resolve issues/install tools and libraries on the AWS platform.
  • Research and come up with architecture and solutions most appropriate for problems at hand.
  • Maintain and improve tools to assist Analytics in ETL, retrospective testing, efficiency, repeatability, and R&D.
  • Lead by example regarding software best practices, including code style and architecture, documentation, source control, and testing.
  • Support the Chief Data Scientist/Data Scientists/Big Data Engineers in creating new and novel approaches to solve challenging problems using Machine Learning, Big Data and Cloud.
  • Handle ADHOC requirements to create reports for the end users.
Required Skills
  • Strong skills with Apache Spark (Spark SQL) and SCALA with at least 2+ years of experience.
  • Understanding of AWS Big Data components and tools.
  • Strong Java skills with experience in web services and web development is required.
  • Hands on experience with model deployment.
  • Hands on experience in application deployment on Docker and/or Kubernetes or other similar technology.
  • Linux scripting is a plus.
  • Fundamental understanding of AWS cloud components.
  • 2+ years of experience in data ingesting, cleansing/processing, storing and querying large datasets
  • 2+ years of experience in engineering large-scale data solutions with Java/Tomcat/ SQL/Linux
  • Experience working in a data intensive role including the extraction of data (db/web/api/etc.), transformation and loading (ETL)
  • Exposure with structured and/or unstructured data contents
  • Experience with data cleansing/preparation on Hadoop/Apache Spark Ecosystem – MapReduce/Hive/HBase/Spark SQL
  • Experience with distributed streaming tools like Apache KAFKA.
  • Experience with multiple file formats (Parquet, Avro, OCR)
  • Knowledge in AGILE development cycle.
  • Efficient coding skills to enhance the performance/cost savings of the job running on AWS platform.
  • Experience in building stable, scalable, and high-speed live streams of data and serving web platforms
  • Enthusiastic self-starter with ability to work in a team environment.
  • Graduate (MS) or Undergraduate degree in Computer Science/ Engineering/relevant field
Nice to have:
  • Strong Software development experience
  • Ability to write custom Map/Reduce programs to clean/prepare complex data
  • Familiarity with Streaming data processing - Experience with distributed real time computation system like Apache STORM/Apache Spark Streaming.

All your information will be kept confidential according to EEO guidelines.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Big Data Engineer - RQ308
Big Data Engineer - RQ308

Experis • McLean (VA)

On-site
USD 110,000 - 150,000
Data Engineer
Data Engineer

Infinite Computer Solutions • Town of Texas (WI)

On-site
USD 120,000 - 170,000
Big Data Platform Engineer
Big Data Platform Engineer

Compunnel, Inc. • Rockville (MD)

On-site
USD 140,000 - 190,000
Big Data Engineer
Big Data Engineer

TechDigital Group • Jersey City (NJ)

On-site
USD 90,000 - 150,000
Senior Big Data Engineer (Databricks + AWS)
Senior Big Data Engineer (Databricks + AWS)

SoftServe • Town of Poland (NY)

On-site
USD 100,000 - 150,000
Big Data Developer
Big Data Developer

Unisys • Rockville (MD)

Hybrid
USD 120,000 - 180,000
Senior Data Engineer
Senior Data Engineer

Compunnel, Inc. • Charlotte (NC)

On-site
USD 120,000 - 150,000
AWS Data Engineer
AWS Data Engineer

Inizio Partners Corp • Newark (NJ)

On-site
Big Data Lead
Big Data Lead

Veriipro • United States

On-site
USD 180,000 - 240,000
Lead BigData Engineer (Databricks + AWS)
Lead BigData Engineer (Databricks + AWS)

SoftServe • Town of Poland (NY)

On-site
USD 150,000 - 190,000