Data Engineer

Artha Nexgen

Northern (KY)

Hybrid

USD 100,000 - 150,000

Full time

5 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Equity compensation
Health insurance reimbursement
401(k) with match
Remote-first workforce

Job summary

Artha Nexgen is seeking a Data Engineer to build and scale data infrastructure powering AI-driven products and research. You will develop distributed data pipelines, manage large-scale datasets across cloud environments, and design reliable data systems for data processing, experimentation, and model development at scale.

Responsibilities include designing and maintaining scalable pipelines, optimizing Spark workflows in AWS, building storage across SQL and NoSQL, writing Python and SQL,

Qualifications

  • Minimum 3 years of experience in data engineering.
  • Experience building scalable data pipelines and data infrastructure for AI/ML workloads.
  • Strong proficiency in Python and SQL, with distributed processing expertise.

Responsibilities

  • Design, build, and maintain scalable data pipelines to ingest, process, and transform large-scale datasets from multiple sources.
  • Develop and optimize distributed data processing workflows using Spark and cloud-native technologies.
  • Build and maintain data storage solutions across SQL and NoSQL systems, ensuring scalability and performance.
  • Design and implement data architectures on AWS to support high-volume data ingestion and distribution.
  • Write efficient Python and SQL code to extract, transform, validate, and analyze large datasets.
  • Ensure data quality, integrity, monitoring, and reliability across pipelines and storage.
  • Collaborate with AI researchers, data scientists, and engineering teams to support data-intensive applications.
  • Implement automation, orchestration, and monitoring workflows to support scalable data operations.

Skills

Python
SQL
Apache Spark
Distributed data processing

Tools

AWS data services
SQL/NoSQL databases

Job description

This job requires a minimum of 3 years of experience.

About the Job

Job Title: Data Engineer

Job Type: Full-time

Location: Remote

The Role

We are looking for a Data Engineer to build and scale the data infrastructure that powers AI-driven products and research initiatives. In this role, you will develop distributed data pipelines, manage large-scale datasets across cloud environments, and design reliable data systems that support data processing, experimentation, and model development at scale.

Key Responsibilities

  • Design, build, and maintain scalable data pipelines to ingest, process, and transform large-scale datasets from multiple sources.
  • Develop and optimize distributed data processing workflows using Spark and cloud-native technologies.
  • Build and maintain data storage solutions across SQL and NoSQL systems, ensuring scalability, performance, and reliability.
  • Design and implement data architectures on AWS to support high-volume data ingestion, processing, and distribution.
  • Write efficient Python and SQL code to extract, transform, validate, and analyze large datasets.
  • Ensure data quality, integrity, monitoring, and operational reliability across data pipelines and storage layers.
  • Collaborate with AI researchers, data scientists, and engineering teams to support data-intensive applications and experimentation.
  • Implement automation, orchestration, and monitoring workflows to support scalable and efficient data operations.

Required Skills and Qualifications

  • Strong proficiency in Python, SQL, and distributed data processing frameworks such as Apache Spark.
  • Hands-on experience with AWS data services and cloud-native data architectures.
  • Experience working with both SQL and NoSQL databases.
  • Experience managing and processing large-scale datasets in distributed environments.
  • Strong understanding of data partitioning, performance optimization, and scalable data architectures

Nice to Have

  • Exposure to AI/ML workflows or research environments.
  • Experience with data visualization tools such as Matplotlib, Seaborn, or Plotly.
  • Familiarity with LLM-related data workflows (datasets for training, evaluation, or prompt experimentation).

Compensation & Benefits Notice

The national pay range for this full-time position is base salary of $100,000 –$150,000 USD. All employees are eligible for equity compensation, and employees may also receive performance-based bonuses, dependent on role and subject to company policies. micro1 provides a comprehensive benefits package, including up to 100% reimbursement for health-insurance premiums, paid time off, a 401(K) plan with a company match, and additional benefits designed to support a high-performing, remote-first workforce.

micro1 is an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex (including pregnancy, sexual orientation, or gender identity), national origin, age, disability, genetic information, veteran status, or any other characteristic protected by applicable local laws, regulations and ordinances. If you need assistance and/or a reasonable accommodation during the application process, reach out to support@micro1.ai.

Our hiring process utilizes artificial intelligence tools to assist in candidate screening and assessment. Our AI tools are designed to complement, not replace, human decision-making.

Disclaimer

The information contained in this job posting, including but not limited to role responsibilities, qualifications, compensation, and benefits, is provided for informational purposes only and does not constitute a binding offer of employment. micro1 reserves the right to amend, modify, or withdraw any portion of this posting at its sole discretion and without prior notice. All employment decisions are made in accordance with applicable laws and regulations.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Data Engineer, Remote - Full Time Remote (International)
Data Engineer, Remote - Full Time Remote (International)

S27a • Northern (KY)

Hybrid
USD 120,000 - 180,000
Equity
Bonuses
Health insurance
+2
Data Engineer
Data Engineer

Appsierra Group • United States

On-site
USD 140,000 - 180,000
Equity
Performance bonuses
Health insurance reimbursement
+3
Data Engineer
Data Engineer

re-zoo-me • Nevada (IA)

Hybrid
USD 70,000 - 95,000
Data Engineer
Data Engineer

Real Chemistry • United States

On-site
USD 90,000 - 120,000
Comprehensive medical, dental, and vision plans
Paid time off
Mental wellness support
+1
Remote | Data Engineer — $140,000–$180,000/year
Remote | Data Engineer — $140,000–$180,000/year

24-Mag Llc • Northern (KY), New York (NY)

Hybrid
USD 140,000 - 180,000
Data Engineer
Data Engineer

Southern Arkansas University • Warner Robins (GA)

Remote
Flexible hours
Weekly bonus of $500–$1000 USD
Work from anywhere
Data Engineer
Data Engineer

SpaceCoast AV Consultants • Town of Florida (NY)

Remote
USD 110,000 - 140,000
Health, dental, and vision insurance
401(k) with company matching
Generous paid time off and parental leave
+1
Sr. Data Engineer - AI
Sr. Data Engineer - AI

Dairy Challenge • Kansas City (KS), Northern (KY)

Hybrid
USD 99,000 - 139,000
Data Architect (Remote)
Data Architect (Remote)

Innowhyte Inc • Irvine (CA)

Remote
USD 120,000 - 160,000
Flexible work from home
Direct interaction with CXO team
Career growth opportunities
Data Engineer
Data Engineer

OpenAI • Los Angeles (CA)

On-site
USD 120,000 - 160,000