Remote Application SRE: Reliability, Automation & Cloud

ELLKAY

Elmwood Park (NJ)

Hybrid

USD 120,000 - 180,000

Full time

4 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Remote work options
Hybrid work environment
Medical, Dental, and Vision benefits
401k w/ matching
Paid time off
Career growth opportunities
Flexible working hours
Hybrid/remote options for certainRoles

Job summary

ELLKAY is seeking an Application Site Reliability Engineer (SRE) with strong DevOps experience to improve the reliability, scalability, and performance of our applications. The role drives reliability standards, observability, and automation across teams to enable faster, safer releases.

The SRE will lead complex incident responses, partner with development teams, and implement best practices to ensure resilient software delivery in a hybrid work environment.

Qualifications

  • Minimum 7 years of SRE experience with strong DevOps background.
  • Solid understanding of Windows and Linux systems and networking.
  • Hands-on cloud experience (AWS, Azure, GCP) and container orchestration.
  • Proficiency with CI/CD tools and Infrastructure as Code.
  • Strong scripting and observability skills; incident management experience.

Responsibilities

  • Own application reliability, availability, performance, and scalability in production and non-production environments.
  • Design, build, and maintain CI/CD pipelines for deployments.
  • Automate infrastructure provisioning with Infrastructure as Code.
  • Monitor health with metrics, logs, and traces; define SLIs, SLOs, and error budgets.
  • Lead incident response and RCA, ensuring corrective actions are completed.
  • Improve resilience via capacity planning, tuning, and fault tolerance.
  • Partner with development teams to meet reliability and scalability objectives.
  • Reduce manual toil through automation and self-healing solutions.
  • Act as incident commander for Sev1/Sev2 situations when needed.

Skills

SRE/DevOps experience
Windows & Linux
Cloud platforms (AWS, Azure, GCP)
Docker & Kubernetes
CI/CD (Jenkins, GitHub Actions)
Infrastructure as Code (Terraform, in/
Scripting (Python, Bash)
Observability (Prometheus, Grafana)
Incident management & RCA

Tools

Docker
Kubernetes
Jenkins
GitHub Actions
Terraform
CloudFormation
ARM
Prometheus
Grafana
ELK
Datadog

Job description

ELLKAY is seeking an Application Site Reliability Engineer (SRE) with strong DevOps experience to improve the reliability, scalability, and performance of our applications. The role drives reliability standards, observability, and automation across teams to enable faster, safer releases.

The SRE will lead complex incident responses, partner with development teams, and implement best practices to ensure resilient software delivery in a hybrid work environment.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Application SRE: DevOps for Reliable, Scalable Apps
Application SRE: DevOps for Reliable, Scalable Apps

Ellkay, Llc • Elmwood Park (NJ)

Hybrid
USD 90,000 - 110,000
Medical, Dental, and Vision benefits
401k with matching
Generous paid time off (FTO)
+2
Application SRE (DevOps)
Application SRE (DevOps)

ELLKAY • Elmwood Park (NJ)

Hybrid
USD 120,000 - 180,000
Remote work options
Hybrid work environment
Medical, Dental, and Vision benefits
+5
Application SRE
Application SRE

Ellkay, Llc • Elmwood Park (NJ)

Hybrid
USD 90,000 - 110,000
Medical, Dental, and Vision benefits
401k with matching
Generous paid time off (FTO)
+2
Hybrid SRE: Application Reliability & Automation
Hybrid SRE: Application Reliability & Automation

Diversified Services Network, Inc. • Chicago (IL)

Hybrid
USD 95,000 - 100,000
401(k)
Dental insurance
Vision Insurance
+7
Remote SRE — Scale, Resilience & Observability
Remote SRE — Scale, Resilience & Observability

Bright Vision Technologies • United States

On-site
USD 100,000 - 150,000
Competitive base salary
Health benefits
Long-term stability
Lead Principal SRE: Cloud Reliability & Automation
Lead Principal SRE: Cloud Reliability & Automation

Ll Oefentherapie • Vienna (VA)

On-site
USD 120,000 - 180,000
Senior Platform Engineer - Hybrid Cloud & SRE Lead
Senior Platform Engineer - Hybrid Cloud & SRE Lead

Ellkay,-LLC • United States

Hybrid
USD 160,000 - 180,000
Remote work options
401k with matching
Hybrid work model
+1
Site Reliability Engineer -Jersey City, NJ & Dallas, TX
Site Reliability Engineer -Jersey City, NJ & Dallas, TX

StradIT • Jersey City (NJ)

Hybrid
USD 120,000 - 160,000
Senior Platform Engineer - Hybrid Cloud & SRE Leader
Senior Platform Engineer - Hybrid Cloud & SRE Leader

Ellkay, Llc • United States

On-site
USD 140,000 - 190,000
Remote Senior Network Reliability Engineer (SRE)
Remote Senior Network Reliability Engineer (SRE)

Gainbridge • Zionsville (IN), Northern (KY)

On-site
USD 135,000 - 190,000
Health Insurance
Dental Insurance
Vision Insurance
+4