Site Reliability Engineer- Application Development(Kubernetes/Linux)

Socket.dev

Town of Texas (WI)

Hybrid

USD 80,000 - 90,000

Full time

4 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

MOTIVE is seeking an experienced Managed Services SRE to deploy, operate, and maintain customer applications on Linux bare metal servers and OpenShift/Kubernetes clusters in a live production environment. The role emphasizes deployment, release management, reliability, and automation.

You will participate in on-call rotations, night-time deployments, and incident response, carrying pager on rotation while aiming to meet SLA targets and continuously improve observability and tooling.

Qualifications

  • Minimum 5 years Linux system administration experience.
  • 5+ years Kubernetes (K8s) experience and exposure to Red Hat OpenShift.
  • Experience with application servers such as JBoss or WebLogic.
  • Experience with monitoring: Zabbix, Prometheus, Grafana.
  • Experience with logging pipelines: Elasticsearch, Logstash, Kibana (ELK).
  • Exposure to web servers (Apache, Nginx) and Ansible.
  • Strong troubleshooting in a live production environment.
  • Willingness to carry pager for on-call rotations and night deployments.

Responsibilities

  • Deploy, manage, and maintain applications on Linux bare metal servers and OpenShift/Kubernetes clusters.
  • Execute CI/CD pipelines and ensure reliable releases across hybrid environments.
  • Build and maintain observability using Prometheus, Grafana, Zabbix.
  • Maintain centralized logging with Grafana Loki, OpenSearch/Elasticsearch, Fluentd/Fluent Bit.
  • Develop automation scripts in Bash, Python, and JavaScript.
  • Participate in incident response and troubleshooting in live production environments.
  • Support night deployments and pager-based on-call rotations; weekends/holidays as needed.
  • Collaborate with internal teams to improve deployment reliability and efficiency.

Skills

Linux system administration
Kubernetes
OpenShift
Web servers
Ansible
Bash
Python
JavaScript
Incident response
Troubleshooting

Tools

OpenShift/Kubernetes clusters
Prometheus
Grafana
ELK (Elasticsearch, Logstash, Kibana)
Apache
Nginx

Job description

The Managed Services SRE is responsible for deploying, operating, and maintaining customer applications across Linux bare metal servers and Red Hat OpenShift (OCP) containerized platforms. This role focuses on application deployment, release management, reliability, and operational support in a live production environment.

The SRE will participate in on-call rotations, night-time deployments, and support, ensuring systems meet SLA requirements while continuously improving reliability and automation practices.

Key Responsibilities
  • Deploy, manage, and maintain applications on Linux bare metal servers and OpenShift/Kubernetes clusters
  • Execute CI/CD pipelines and ensure reliable, repeatable releases across hybrid environments
  • Build and maintain observability for deployed applications using Prometheus, Grafana, Zabbix
  • Implement and maintain centralized logging solutions using Grafana Loki, OpenSearch/Elasticsearch, Fluentd/Fluent Bit
  • Develop automation scripts to streamline deployments and reduce operational toil (Bash, Python, JavaScript)
  • Participate in incident response and troubleshoot application or platform issues in a live production environment
  • Support night-time deployments and carry pager on rotation; respond to emergencies, including weekends and holidays
  • Collaborate with internal teams to continuously improve deployment reliability and efficiency
  • Learn new technologies, take direction, and develop skills as needed
Required Qualifications
  • Minimum 5 years Linux system administration experience
  • Minimum 5 years Kubernetes (K8s) experience
  • Exposure to RedHat OpenShift
  • Experience with application servers such as JBoss or WebLogic
  • Experience with monitoring tools: Zabbix, Prometheus, Grafana
  • Experience with logging pipelines: Elasticsearch, Logstash, Kibana (ELK)
  • Exposure to web servers – Apache, Nginx
  • Experience with Ansible
  • Basic networking skills
  • Basic SQL skills
  • Strong troubleshooting skills and ability to operate in a live production environment
  • Willingness to:
    • Carry pager for on-call rotations (typically a week at a time)
    • Support night-time deployments
    • Work off-hours, including weekends and holidays in emergencies
    • Learn on the fly and develop new skills
Skills & Competencies
  • Strong problem-solving and troubleshooting skills
  • Excellent communication and teamwork abilities
  • Self-driven, proactive, and willing to take ownership
  • Ability to operate effectively in a fast-paced, SLA-driven environment
Working Conditions
  • 24/7 On-call responsibilities on a rotating schedule including weekends and holidays
  • Night-time and off-hours deployment support

Location - Remote - US

Compensation USD 80-90K annual

MOTIVE gives equal opportunity in employment regardless of gender, gender identity, sexual orientation, marital status, race, nationality, religion, age, disability, political beliefs, or any other factor. MOTIVE will not pay fees to any third-party agency or company that does not have a signed agreement, do not submit resumes/CV's directly.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer- Application Development(Kubernetes/Linux)
Site Reliability Engineer- Application Development(Kubernetes/Linux)

Advantage 360 • Town of Texas (WI)

Hybrid
USD 80,000 - 90,000
Remote SRE: Kubernetes/OpenShift Production Reliability
Remote SRE: Kubernetes/OpenShift Production Reliability

Socket.dev • Town of Texas (WI)

On-site
USD 80,000 - 90,000
Site Reliability Engineer Lead (SRE) – Internal Kubernetes Container Platform (IKCP)
Site Reliability Engineer Lead (SRE) – Internal Kubernetes Container Platform (IKCP)

Bank of America • Chandler (AZ)

On-site
USD 140,000 - 170,000
Site Reliability Engineer Lead (SRE) – Internal Kubernetes Container Platform (IKCP)
Site Reliability Engineer Lead (SRE) – Internal Kubernetes Container Platform (IKCP)

Hobbsnews • Chandler (AZ), Northern (KY)

Hybrid
USD 120,000 - 180,000
Site Reliability Engineer Lead (SRE) - Internal Kubernetes Container Platform (IKCP)
Site Reliability Engineer Lead (SRE) - Internal Kubernetes Container Platform (IKCP)

Koitecc Solutions • Chandler (AZ), Northern (KY)

On-site
USD 140,000 - 200,000
Site Reliability Engineer
Site Reliability Engineer

Veritas Search Group • Tustin (CA)

On-site
USD 140,000 - 190,000
Site Reliability Engineer Lead (SRE) – Internal Kubernetes Container Platform (IKCP)
Site Reliability Engineer Lead (SRE) – Internal Kubernetes Container Platform (IKCP)

Bank of America • Jersey City (NJ)

On-site
USD 180,000 - 240,000
Site Reliability Engineer Lead (SRE) – Internal Kubernetes Container Platform (IKCP)
Site Reliability Engineer Lead (SRE) – Internal Kubernetes Container Platform (IKCP)

Bank of America • Charlotte (NC)

On-site
USD 125,000 - 168,000
Discretionary incentive eligible
Annual discretionary plan
Site Reliability Engineer Lead (SRE) – Internal Kubernetes Container Platform (IKCP)
Site Reliability Engineer Lead (SRE) – Internal Kubernetes Container Platform (IKCP)

Bank of America • Plano (TX)

On-site
USD 140,000 - 190,000
Site Reliability Engineer Lead (SRE) - Internal Kubernetes Container Platform (IKCP)
Site Reliability Engineer Lead (SRE) - Internal Kubernetes Container Platform (IKCP)

Bank of America • Jersey City (NJ)

On-site
USD 125,000 - 168,000