Principal Software Engineer

PeerIslands US Inc.

Southlake (TX)

Hybrid

USD 140,000 - 170,000

Full time

8 days ago
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

PeerIslands US Inc. in Southlake, TX seeks a Senior Software Engineer to build automation tooling with LLMs and agent frameworks, supporting legacy code analysis, documentation, and modernization across pipelines.

You will work with PySpark and Databricks, design cloud-native migration services, and architect data models for MongoDB. The role emphasizes TDD, data validation, and scalable ETL frameworks in hybrid work arrangements.

Qualifications

  • Bachelor's degree in Computer Science or related field required.
  • 5 years post-baccalaureate experience in software/data engineering.
  • Experience with Spark (PySpark), Databricks or Dataproc, SQL/NoSQL databases, and cloud services.

Responsibilities

  • Develop automation tooling using LLMs and agent frameworks for legacy code analysis, documentation, and migration scaffolding.
  • Analyze and reverse-engineer legacy databases and mainframe workloads to identify dependencies and data flows.
  • Design and build cloud-native migration services to convert legacy procedures into modern microservices.
  • Develop scalable data platforms and pipelines using Apache Spark (PySpark) and Dataproc in a cloud environment.
  • Architect data models and denormalized MongoDB schemas to boost query performance.

Skills

Python
LLMs
Agentic AI
Data engineering
TDD
Cloud computing
ETL design

Education

Bachelor’s degree in Computer Science or related field

Tools

PySpark
Databricks
Google Cloud Dataproc
Apache Airflow
MongoDB
SQL
NoSQL
LangChain
LangGraph

Job description

JOB TITLE: Senior Software Engineer

LOCATION: Southlake, TX (hybrid role, may work from home)

DUTIES

Develop automation tooling using LLMs (Large Language Models) and agent frameworks to assist with legacy code analysis, documentation generation, and migration/pipeline scaffolding; integrate outputs with engineering review, validation, and testing; Analyze and reverse-engineer legacy database and mainframe workloads to identify dependencies, data flows, and functional logic required for modernization and migration; Develop parsers and analysis utilities to extract metadata, lineage, and relationships from legacy codebases to support migration planning, documentation, and implementation; Design and build cloud-native migration services to convert legacy procedures and batch logic into modern microservices and standardized data processing jobs; Design and implement scalable data platforms and pipelines to support enterprise eligibility and operational data processing using Apache Spark (PySpark) and Databricks/Google Cloud Dataproc; Architect data models and storage patterns; redesign relational schemas into denormalized nested document models suitable for MongoDB to improve downstream query efficiency and application performance; Build configuration-driven Extract, Transform, Load (ETL) frameworks that generate Spark jobs from declarative specifications; Develop data validation, auditing, and threshold-based control frameworks to detect data discrepancies and enforce quality gates across pipeline stages; Orchestrate and monitor data workflows using Apache Airflow (Google Cloud Composer), including dependency management, retries, alerts, and operational controls; Implement Change Data Capture and reverse ETL mechanisms to synchronize changes from MongoDB to data lake storage in near real time for downstream analytics and reporting.

REQUIREMENTS

Bachelor’s or foreign equivalent degree in Computer Science, Computer or Electronic Engineering, or a related field, and 5 years of progressive, post-baccalaureate experience in the job offered or as a Software Engineer/Developer, Data Engineer, Programmer Analyst, or in a related/similar position. Experience therein to include 5 years in data engineering using Apache Spark (PySpark) and Databricks or Google Cloud Dataproc, databases and data modeling using SQL, NoSQL or MongoDB, schema design and optimization, cloud services such as Microsoft Azure, GCP or AWS for data/compute, and Python software development with Agentic AI, data processing, backend services, and TDD; and 2 years with Large Language Models (LLMs), LangChain and LangGraph agent frameworks. Hybrid role, ability to work from home.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Software Engineer
Senior Software Engineer

Eliassen Group • Dallas (TX)

Hybrid
USD 110,000 - 117,000
Medical, Dental, and Vision benefits
401k with company matching
Life insurance
Senior Data Engineer 4 4P/259
Senior Data Engineer 4 4P/259

4P Consulting Inc. • Atlanta (GA)

On-site
USD 120,000 - 150,000
2026-2201 Data Scientist
2026-2201 Data Scientist

Mountain Cat LLC • McLean (VA)

On-site
USD 120,000 - 180,000
Data Platform Engineer - 90/HR- REMOTE
Data Platform Engineer - 90/HR- REMOTE

ContractStaffingRecruiters.com • Branford (CT)

Remote
USD 130,000 - 160,000
Senior Software Engineer
Senior Software Engineer

Eliassen Group • Boston (MA)

Hybrid
Confidential
Medical coverage
Dental coverage
Vision coverage
+2
Lead Software Engineer - Big Data
Lead Software Engineer - Big Data

JPMorgan Chase & Co. • Plano (TX)

On-site
USD 140,000 - 190,000
Senior Manager Data Engineer
Senior Manager Data Engineer

DataJobs • Chicago (IL)

On-site
USD 209,000 - 239,000
Performance-based incentive
Comprehensive benefits
Data Engineer
Data Engineer

Compunnel, Inc. • Dallas (TX)

On-site
USD 90,000 - 120,000
Senior Data Engineer
Senior Data Engineer

Compunnel, Inc. • Charlotte (NC)

On-site
USD 120,000 - 150,000
Data Engineer (Snowflake / Databricks / BigQuery)
Data Engineer (Snowflake / Databricks / BigQuery)

Zoho • United States

On-site
USD 120,000 - 170,000