ELH Site Reliability Engineer Lowell, SVL, Austin

IBM

Lowell (MA)

On-site

USD 85,000 - 110,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

A technology leader is seeking a Site Reliability Engineer to ensure optimal performance of production systems. Responsibilities include 24x7 monitoring, collaborating on troubleshooting, and deploying services using CI/CD tools. Ideal candidates should have 1-3 years of experience with monitoring, Linux, and automation in a production environment. This full-time role offers an opportunity to work in a dynamic engineering environment in Lowell, MA.

Qualifications

  • 1-3 years of experience in monitoring/observability and troubleshooting.
  • Experience with Linux operating systems for system support.
  • Knowledge or experience of Kubernetes/OpenShift environments is preferred.

Responsibilities

  • Monitor health of production systems and services globally.
  • Collaborate with teams to troubleshoot and resolve production issues.
  • Deploy services using Continuous Delivery tools.

Skills

Monitoring/Observability
Automation proficiency
Linux
Kubernetes/OpenShift
Scripting (Ansible, Python)

Education

Bachelor's Degree

Tools

CI/CD tools (Jenkins, IBM Continuous Delivery)
Grafana/Prometheus

Job description

A career in IBM Software means you’ll be part of a team that transforms our customer’s challenges into solutions.

Seeking new possibilities and always staying curious, we are a team dedicated to creating the world’s leading AI-powered, cloud-native software solutions for our customers. Our renowned legacy creates endless global opportunities for our IBMers, so the door is always open for those who want to grow their career. IBM’s product and technology landscape includes Research, Software, and Infrastructure. Entering this domain positions you at the heart of IBM, where growth and innovation thrive.

Your Role And Responsibilities

As a Site Reliability Engineer, you will work in an agile, collaborative environment to build, deploy, configure, and maintain systems for the IBM client business. In this role, you will lead the problem resolution process for our clients, from analysis and troubleshooting, to deploying the latest software updates & fixes.

Your Primary Responsibilities Include
  • 24x7 Observability: Be part of a worldwide team that monitors the health of production systems and services around the clock, ensuring continuous reliability and optimal customer experience.
  • Cross-Functional Troubleshooting: Collaborate with engineering teams to provide initial assessments and possible workarounds for production issues. Troubleshoot and resolve production issues effectively.
  • Deployment and Configuration: Leverage Continuous Delivery (CI/CD) tools to deploy services and configuration changes at enterprise scale.
  • Security and Compliance Implementation: Implementing security measures that meet or exceed industry standards for regulations such as GDPR, SOC2, ISO 27001, PCI, HIPAA, and FBA.
  • Maintenance and Support: Tasks related to applying security patches and upgrades, and collaborating with Product support for issue resolution.
Preferred Education

Bachelor's Degree

Required Technical And Professional Expertise
  • System Monitoring and Troubleshooting: 1-3 years of experience in monitoring/observability, issue response, and troubleshooting for optimal system performance.
  • Automation Proficiency: 1-3 years of experience in automation for production environment changes, streamlining processes for efficiency, and reducing toil.
  • Linux: 1 to 3 years of experience working with Linux operating systems.
  • Operation and Support Experience: 1-3 years of experience handling day-to-day operations, alert management, incident support, migration tasks, and break‑fix support.
Preferred Technical And Professional Experience
  • Kubernetes/OpenShift: knowledge or experience of Kubernetes/OpenShift environments.
  • Automation/Scripting: knowledge or experience of Ansible, Python, Terraform, and CI/CD tools such as Jenkins, IBM Continuous Delivery, ArgoCD.
  • Monitoring/Observability: knowledge or experience crafting alerts and dashboards using tools such as Instana, New Relic, Grafana/Prometheus.
  • DBA: Interest or experience configuring and maintaining SQL, NoSQL, and data streaming technologies (e.g. PostgreSQL, CouchDB, Redis, Kafka, Spark, etc.).
Seniority level

Mid‑Senior level

Employment type

Full‑time

Job function

Engineering and Information Technology

Industries

IT Services and IT Consulting

Referrals increase your chances of interviewing at IBM by 2x.

Get notified about new Site Reliability Engineer jobs in Lowell, MA.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer ELH - OneIT - Rochester, MN - 2027
Site Reliability Engineer ELH - OneIT - Rochester, MN - 2027

IBM • Town of Montana (WI)

On-site
USD 65,000 - 90,000
Healthcare benefits
401(k) plan
Paid time off
Site Reliability Engineer Intern 2027
Site Reliability Engineer Intern 2027

IBM • Durham (NC)

Hybrid
USD 76,000 - 166,000
Site Reliability Engineer ELH – OneIT – Rochester, MN – 2027
Site Reliability Engineer ELH – OneIT – Rochester, MN – 2027

IBM • United States

On-site
USD 65,000 - 90,000
Site Reliability Engineer ELH - OneIT - Rochester, MN - 2027
Site Reliability Engineer ELH - OneIT - Rochester, MN - 2027

IBM • United States

Hybrid
USD 101,000 - 151,000
Entry Level Site Reliability Engineer - Tucson-AZ
Entry Level Site Reliability Engineer - Tucson-AZ

IBM • Tucson (AZ)

On-site
USD 110,000 - 160,000
Site & Network Reliability Engineer & Administrator
Site & Network Reliability Engineer & Administrator

IBM • Tucson (AZ)

On-site
USD 70,000 - 90,000
Site Reliability Engineer Intern 2027
Site Reliability Engineer Intern 2027

IBM • Lowell (MA)

On-site
USD 24,000 - 36,000
Site Reliability Engineer Intern 2027
Site Reliability Engineer Intern 2027

IBM • Austin (TX)

On-site
USD 25,000 - 36,000
Site Reliability Engineer Intern 2027
Site Reliability Engineer Intern 2027

IBM • San Jose (CA)

On-site
USD 30,000 - 47,000
Entry Level Site Reliability Engineering Professional - Austin, TX - 2027
Entry Level Site Reliability Engineering Professional - Austin, TX - 2027

IBM • Austin (TX)

Hybrid
USD 101,000 - 151,000
Healthcare benefits
401(k) plan
Paid time off