Senior SRE: Infrastructure Observability & Reliability

NewGen Technologies

Owings Mills, Northern (MD, KY)

Hybrid

USD 150,000 - 190,000

Full time

11 hours ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

NewGen Technologies seeks a Principal Site Reliability Engineer, Infrastructure Observability to lead a team focused on observability, sustainability, scalability, measurability, and recoverability of cloud and on-prem solutions. You will implement automation, drive SRE best practices, and collaborate with partners to ensure 24x7 reliability.

Strong ops background and cloud expertise are required. The role emphasizes designing resilient architectures, incident management, and leading

Qualifications

  • Bachelor's degree or the equivalent combination of education and relevant experience.
  • 10+ years of experience designing and operating cloud infrastructure with senior-level impact.
  • 5+ years building and supporting solutions in Amazon AWS.
  • 5+ years of experience building and running a DevOps and/or SRE function.
  • Experience with chaos engineering at scale.
  • Experience with defining, tracking, and reporting SLOs/SLIs and system availability.
  • Experience with incident prevention and remediation through automation.
  • Proficiency with multiple programming languages (Python, Java, Go, Node.js, .Net Core).
  • Proficiency with SQL Server, PostgreSQL, MySQL.
  • Experience with cloud management tools such as Ansible, Terraform, Vault, and Vagrant.

Responsibilities

  • Design technology solutions to prevent or minimize service disruptions.
  • Lead SRE initiatives across a distributed technology environment.
  • Drive automation to proactively prevent incidents and support recovery.
  • Promote blameless post-mortems and reliability practices.
  • Analyze incidents for trends and improve availability across services.
  • Coordinate with partners and sponsor groups to implement target state architecture.
  • Develop dashboards and observability standards across ecosystems.

Skills

SRE
Cloud infrastructure
Automation
Incident response
Observability
Dashboarding
On-call

Education

Bachelor's degree or equivalent

Tools

New Relic
SolarWinds DPA
Elastic Stack
Prometheus
Grafana
Splunk
Ansible
Terraform
Vault
Vagrant

Job description

NewGen Technologies seeks a Principal Site Reliability Engineer, Infrastructure Observability to lead a team focused on observability, sustainability, scalability, measurability, and recoverability of cloud and on-prem solutions. You will implement automation, drive SRE best practices, and collaborate with partners to ensure 24x7 reliability.

Strong ops background and cloud expertise are required. The role emphasizes designing resilient architectures, incident management, and leading

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Principal Site Reliability Engineer, Infrastructure Observability Information Technology Owings Mills, MD
Principal Site Reliability Engineer, Infrastructure Observability Information Technology Owings Mills, MD

NewGen Technologies • Owings Mills (MD), Northern (KY)

Hybrid
USD 150,000 - 190,000
Senior SRE: Cloud-Native Reliability & Observability
Senior SRE: Cloud-Native Reliability & Observability

UnitedHealth Group • Schaumburg (IL)

Remote
USD 92,000 - 164,000
Telecommute within US
Senior SRE Lead: Cloud Reliability & Observability
Senior SRE Lead: Cloud Reliability & Observability

Prosum • Scottsdale (AZ)

On-site
USD 150,000 - 190,000
Senior SRE: Scale Reliability, Observability & Resilience
Senior SRE: Scale Reliability, Observability & Resilience

Early Warning Services LLC • Scottsdale (AZ)

Hybrid
USD 106,000 - 130,000
Healthcare Coverage
401(k) Retirement Plan
Paid Time Off
+2
Senior SRE: Observability, Automation & Scalable Systems
Senior SRE: Observability, Automation & Scalable Systems

Replit • Northern (KY)

Hybrid
USD 140,000 - 190,000
Competitive Salary & Equity
401(k) 4% match (US)
Health, Dental, Vision & Life
+7
Senior SRE - Hybrid, Observability & Reliability
Senior SRE - Hybrid, Observability & Reliability

Early Warning Services LLC • Chicago (IL)

Hybrid
USD 106,000 - 130,000
Healthcare Coverage
401(k) Plan with match
PTO and Holidays
+1
Senior SRE: Cloud-Native Reliability & Observability
Senior SRE: Cloud-Native Reliability & Observability

Optum • Schaumburg (IL)

Remote
USD 92,000 - 164,000
Senior SRE - Observability & Cloud Reliability
Senior SRE - Observability & Cloud Reliability

Cisco Systems, Inc. • San Francisco (CA)

On-site
USD 168,000 - 245,000
Medical benefits
401(k) matching
Parental leave
+1
Senior Production SRE: Cloud & On-Prem Reliability
Senior Production SRE: Cloud & On-Prem Reliability

Weights & Biases • New York (NY)

On-site
USD 140,000 - 180,000
Medical Insurance
Dental Insurance
Vision Insurance
+15
Senior SRE: Automate Reliability & Observability
Senior SRE: Automate Reliability & Observability

Bank of America • Charlotte (TX)

On-site
USD 153,000 - 192,000
Discretionary incentive eligible
Benefits package