Senior Software Engineer - IP&R Reliability Engineering (Remote)

The Home Depot

Atlanta (GA)

On-site

USD 130,000 - 170,000

Full time

2 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

The Home Depot is seeking a Senior Software Engineer in Reliability Engineering to design, build, and operate automated monitoring, workload orchestration, and proactive reliability systems for critical retail and supply chain operations.

You will migrate legacy systems to scalable script-based platforms, establish centralized health dashboards, and streamline alerting pipelines using modern observability tools.

Qualifications

  • Minimum 4+ years in professional software engineering, SRE, DevOps, or automation.
  • Strong scripting/programming: Python, Bash, Java, Go, or Groovy.
  • Experience with API integrations and automating health dashboards.
  • Hands-on with enterprise workload automation and job schedulers such as Rundeck, CA7, Maestro, AutoSys, or Control-M.

Responsibilities

  • Develops, tests, deploys, and maintains software with a focus on reliability and observability.
  • Mentors junior engineers and contractors on automation scripting and scheduling practices.
  • Leads modernization efforts to decommission legacy monitoring platforms in favor of script-driven automation.
  • Collaborates with application teams to implement centralized health dashboards and alerting pipelines.

Skills

Python
Bash/Shell
Java
Go
Groovy

Education

Bachelor's degree in Computer Science/Engineering
No further education required

Tools

Rundeck
CA7
Maestro / HCL Workload Automation
AutoSys
Control-M
PagerDuty
Grafana
Splunk
Kubernetes
Docker

Job description

The Senior Software Engineer in Reliability Engineering (RE) is responsible for designing, building, and maintaining automated monitoring solutions, enterprise workload orchestrations, and proactive reliability systems that safeguard critical retail and supply chain operations. In this role, you will lead initiatives to eliminate operational toil through code, migrate legacy systems to scalable script-based platforms, and build centralized health dashboards/establish robust onboarding standards with application development teams. You will streamline alerting pipelines using modern observability and incident management tools (e.g., PagerDuty, Grafana, custom APIs), manage and optimize enterprise batch execution flows across scheduling platforms such as CA7, Maestro (HCL/TWS), and Rundeck, and drive modernization and decommissioning of redundant legacy monitoring platforms in favor of maintainable, script-driven automation. A Senior Software Engineer is also responsible for mentoring junior engineers and contractors on automation scripting, workflow scheduling, and troubleshooting practices.

Req192134

Position Purpose

The Senior Software Engineer in Reliability Engineering (RE) is responsible for designing, building, and maintaining automated monitoring solutions, enterprise workload orchestrations, and proactive reliability systems that safeguard critical retail and supply chain operations. In this role, you will lead initiatives to eliminate operational toil through code, migrate legacy systems to scalable script-based platforms, and build centralized health dashboards/establish robust onboarding standards with application development teams. You will streamline alerting pipelines using modern observability and incident management tools (e.g., PagerDuty, Grafana, custom APIs), manage and optimize enterprise batch execution flows across scheduling platforms such as CA7, Maestro (HCL/TWS), and Rundeck, and drive modernization and decommissioning of redundant legacy monitoring platforms in favor of maintainable, script-driven automation. A Senior Software Engineer is also responsible for mentoring junior engineers and contractors on automation scripting, workflow scheduling, and troubleshooting practices.

Key Responsibilities
  • 50% Delivery and Execution - Develops, tests, deploys, and maintains software, with a clear understanding of the value the software is to provide; Takes on new opportunities and tough challenges with a sense of urgency, high energy and enthusiasm; Consistently achieves results, even under tough circumstances; Develops test suites (functional, destructive, etc) to enable success, rapid deployment of code to production; Takes a broad view when approaching issues; using a global lens
  • 20% Learns and Grows - Learns through successful and failed experiment when tackling new problems; Actively seeks ways to grow and be challenged using both formal and informal development channels
  • 20% Plans and Aligns - Collaborates with other team members in agile processes; Creates new and better ways for the organization to be successful; Works the Product Team to ensure user stories are valuable, developer ready, easy to understand and testable; Delivers multi-mode communications that convey a clear understanding of the unique needs of different audiences; Adapts approach and demeanor in real time to match the shifting demands of different situations; Relates openly and comfortably with diverse groups of people
  • 10% Supports and Enables - Helps grow junior engineers by providing guidance on modern software development frameworks, and leading technical discussions
Direct Manager/Direct Reports
  • This position typically reports to Software Engineer Manager or Sr. Manager
  • This position has 0 Direct Reports
Travel Requirements
  • No travel required.
Physical Requirements
  • Most of the time is spent sitting in a comfortable position and there is frequent opportunity to move about. On rare occasions there may be a need to move or lift light articles.
Working Conditions
  • Located in a comfortable indoor area. Any unpleasant conditions would be infrequent and not objectionable.
Minimum Qualifications
  • Must be eighteen years of age or older.
  • Must be legally permitted to work in the United States.
Preferred Qualifications
  • 4+ years of professional software engineering, SRE, DevOps, or system automation experience in an enterprise environment.
  • Strong proficiency in scripting and programming languages (e.g., Python, Bash/Shell, Java, Go, or Groovy).
  • Experience developing API integrations and automating operational health dashboards and CLI utilities.
  • Hands-on experience with enterprise workload automation and job schedulers (e.g., Rundeck, CA7, Maestro / HCL Workload Automation, AutoSys, or Control-M).
  • Proven experience migrating, re-architecting, or modernizing batch job flows and dependency chains.
  • In-depth experience configuring alerting, escalation policies, and incident response integrations (e.g., PagerDuty, Slack/Teams webhooks).
  • Experience with telemetry, log aggregation, and dashboarding tools (e.g., Splunk, Prometheus, Grafana, Datadog).
  • Familiarity with retail supply chain operations, warehouse management, or Distribution Center (DC) operational systems.
  • Experience leveraging AI coding assistants, agentic workflows, or AIOps tooling to improve operational diagnostics and team productivity.
  • Experience with Google Cloud Platform (GCP) or modern containerization platforms (Docker, Kubernetes).
Minimum Education
  • The knowledge, skills and abilities typically acquired through the completion of a bachelor's degree program or equivalent degree in a field of study related to the job.
Preferred Education
  • No additional education
Minimum Years Of Work Experience
  • 3
Preferred Years Of Work Experience
  • No additional years of experience
Minimum Leadership Experience
  • None
Preferred Leadership Experience
  • None
Certifications
  • None
Competencies
  • Global Perspective
  • Manages Ambiguity
  • Nimble Learning
  • Self-Development
  • Collaborates
  • Cultivates Innovation
  • Situational Adaptability
  • Communicates EffectivelyDrives Results
  • Interpersonal Savvy

Apply End Date: 10/23/2026

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Software Engineer II - Reliability Engineering Tooling (Remote)
Software Engineer II - Reliability Engineering Tooling (Remote)

The Home Depot • Atlanta (GA)

On-site
USD 100,000 - 140,000
Software Engineer Manager - Supply Chain RE (Remote)
Software Engineer Manager - Supply Chain RE (Remote)

The Home Depot • Atlanta (GA)

Hybrid
USD 180,000 - 240,000
Software Reliability Manager, Store Systems
Software Reliability Manager, Store Systems

Visa Hunt • United States

On-site
USD 140,000 - 240,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Spectraforce Technologies • Austin (TX)

Hybrid
USD 130,000 - 170,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

United States Digital Space LLC • Charlotte (TX)

On-site
USD 153,000 - 192,000
Discretionary incentive eligible
Benefits package
Software Engineering Manager - Reliability Engineering, Store Systems
Software Engineering Manager - Reliability Engineering, Store Systems

Visa Hunt • United States

On-site
USD 140,000 - 240,000
Software Reliability Engineer
Software Reliability Engineer

RE Partners Consulting • New York (NY)

On-site
USD 145,000 - 175,000
Associate Engineer, Site Reliability
Associate Engineer, Site Reliability

R&D • United States

On-site
USD 90,000 - 140,000
Principal Site Reliability Engineer
Principal Site Reliability Engineer

Jobtailor • Arizona

On-site
USD 180,000 - 240,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Hobbsnews • Jersey City (NJ)

On-site
USD 153,000 - 192,000
Benefits eligible
Discretionary incentive plan