Linux SME / SRE Engineer

Weekday 1

Hyderabad

On-site

INR 2,000,000 - 3,000,000

Full time

4 days ago
Be an early applicant
Application generator

Get a reply from this employer โ€” a resume and cover letter tailored to exactly what theyโ€™re hiring for.

Get past ATS filters

Job summary

Weekday 1 seeks an experienced Linux SME / SRE Engineer to support client-critical production environments. The role emphasizes core Linux administration, RHEL expertise, and managing PCS/Pacemaker clusters to ensure high availability and reliability across infrastructure.

The candidate should demonstrate strong troubleshooting skills, incident management mindset, and ability to work with clients and technical stakeholders. Rotational production support is expected.

Qualifications

  • Experience in Linux administration and production infrastructure.
  • Hands-on with PCS/Pacemaker high-availability clusters.
  • Familiarity with RHEL and cloud platforms (AWS).
  • Ability to troubleshoot complex production incidents.
  • Knowledge of SRE practices and ITIL concepts is desirable.

Responsibilities

  • Administer and support Linux-based production environments across critical infrastructure.
  • Manage PCS/Pacemaker high-availability clusters to ensure uptime.
  • Monitor, configure, and troubleshoot production systems and services.
  • Perform root-cause analysis and implement durable fixes for recurring issues.
  • Collaborate with clients and internal teams during incidents and changes.
  • Contribute to incident, problem, and change management processes.

Skills

Linux administration
SRE concepts
Incident management
Cluster troubleshooting
Production monitoring

Tools

VMware administration
AWS
Oracle Database

Job description

๐—ง๐—ต๐—ถ๐˜€ ๐—ฟ๐—ผ๐—น๐—ฒ ๐—ถ๐˜€ ๐—ณ๐—ผ๐—ฟ ๐—ผ๐—ป๐—ฒ ๐—ผ๐—ณ ๐˜๐—ต๐—ฒ ๐—ช๐—ฒ๐—ฒ๐—ธ๐—ฑ๐—ฎ๐˜†'๐˜€ ๐—ฐ๐—น๐—ถ๐—ฒ๐—ป๐˜๐˜€


๐—ฆ๐—ฎ๐—น๐—ฎ๐—ฟ๐˜† ๐—ฟ๐—ฎ๐—ป๐—ด๐—ฒ: ๐—ฅ๐˜€ ๐Ÿฎ๐Ÿฌ๐Ÿฌ๐Ÿฌ๐Ÿฌ๐Ÿฌ๐Ÿฌ - ๐—ฅ๐˜€ ๐Ÿฏ๐Ÿฌ๐Ÿฌ๐Ÿฌ๐Ÿฌ๐Ÿฌ๐Ÿฌ (๐—ถ๐—ฒ ๐—œ๐—ก๐—ฅ ๐Ÿฎ๐Ÿฌ-๐Ÿฏ๐Ÿฌ ๐—Ÿ๐—ฃ๐—”)


Experience: 5+ yrs


Location: Bengaluru, Karnataka, India, Hyderabad, Telangana, India


Job Type: Full-time


We are looking for an experiencedLinux SME / SRE Engineerwith strong expertise inCore Linux Administration, RHEL, and PCS/Pacemaker clusteringto support business-critical production environments.


The role focuses on maintaining highly available Linux infrastructure, resolving complex production issues, ensuring system reliability, and supporting clustered environments. The ideal candidate will have strong hands-on troubleshooting capabilities, a solid understanding of high-availability architectures, and the ability to work effectively with clients and technical stakeholders.


Requirements

Key Responsibilities


  • Administer and supportLinux-based production environmentsacross critical infrastructure.

  • Perform day-to-dayCore Linux administration, configuration, monitoring, maintenance, and troubleshooting.

  • Manage, monitor, configure, and troubleshootPCS/Pacemaker high-availability clusters.

  • Ensure availability, reliability, stability, and performance of Linux infrastructure and clustered services.

  • Troubleshoot complex and critical production incidents and drive issues through to resolution.

  • Perform root-cause analysis and implement sustainable solutions for recurring infrastructure problems.

  • Monitor system and cluster health and proactively identify potential availability or performance issues.

  • Support failover, recovery, maintenance, and operational activities across high-availability environments.

  • Collaborate with clients, infrastructure teams, application teams, and other technical stakeholders on incidents and enhancements.

  • Participate in incident management, problem management, change management, and production maintenance activities.

  • Follow SRE practices for monitoring, reliability improvement, incident response, and operational efficiency.

  • Maintain technical documentation, operational procedures, troubleshooting guides, and support records.

  • Participate in rotational shifts to provide continuous production support.

  • Identify opportunities to automate repetitive infrastructure tasks and improve operational efficiency.

  • Support infrastructure changes, upgrades, patching, and maintenance activities in accordance with established processes.

  • Contribute to service reliability, availability, and continuous improvement initiatives.


What Makes You a Great Fit


  • 5โ€“9 years of overall experiencein Linux administration, infrastructure engineering, SRE, or production support, with a maximum of 10 years preferred.

  • Minimum4 years of hands-on experience with PCS/Pacemaker cluster administration.

  • Strong expertise inCore Linux Administrationand production infrastructure support.

  • Strong hands-on experience withRHEL (Red Hat Enterprise Linux).

  • Solid understanding ofHigh Availability, clustering, failover, resource management, and cluster troubleshooting.

  • Proven experience supportingcritical production environmentswith strict availability and reliability requirements.

  • Strong troubleshooting, debugging, root-cause analysis, and incident-resolution capabilities.

  • Experience working with production monitoring, incident management, and infrastructure maintenance processes.

  • Strong understanding ofSRE and ITIL practicesis desirable.

  • Excellent communication and client-facing skills with the ability to explain technical issues clearly to stakeholders.

  • Strong stakeholder-management and collaboration skills.

  • Ability to work effectively under pressure during critical production incidents.

  • Willingness to work inrotational shifts, including scheduled production-support coverage.

  • Experience withVMware administrationis an advantage.

  • Exposure toAWS or other cloud platformsis desirable.

  • Knowledge ofOracle Databaseand its infrastructure dependencies is an advantage.

  • Strong ownership mindset with a focus on system reliability, operational excellence, and continuous improvement.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Linux SME / SRE Engineer | Intineri Infosol | Bangalore / Hyderabad
Linux SME / SRE Engineer | Intineri Infosol | Bangalore / Hyderabad

Tech Junction Ltd โ€ข Bengaluru

Hybrid
INR 1,800,000 - 3,000,000
SRE Lead
SRE Lead

3across โ€ข Bengaluru

On-site
INR 1,500,000 - 2,300,000
Linux Engineer
Linux Engineer

Cloud4C Services โ€ข Hyderabad

On-site
INR 600,000 - 1,200,000
Senior Site Reliability Engineer (SRE) / DevOps Engineer
Senior Site Reliability Engineer (SRE) / DevOps Engineer

Umanist Staffing LLC โ€ข Pune District

On-site
INR 2,250,000 - 2,750,000
Linux System Administrator
Linux System Administrator

SourcingXPress โ€ข Hyderabad

On-site
INR 1,800,000 - 2,300,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

BuildxPartners โ€ข Bengaluru Urban

Hybrid
INR 2,400,000 - 4,200,000
VS01700 - SRE & Production Reliability Engineer
VS01700 - SRE & Production Reliability Engineer

E4 Software Services Pvt Ltd. โ€ข Bengaluru

On-site
INR 1,800,000 - 2,400,000
VS01700 - SRE & Production Reliability Engineer
VS01700 - SRE & Production Reliability Engineer

E4 Software Services Pvt Ltd. โ€ข India

On-site
INR 2,000,000 - 4,000,000
Senior Linux Infrastructure Engineer
Senior Linux Infrastructure Engineer

Spectrum Talent Management โ€ข Mumbai

Hybrid
INR 600,000 - 1,000,000
Site Reliability Engineer - Cloud Infrastructure
Site Reliability Engineer - Cloud Infrastructure

NetConnectGlobal โ€ข Bengaluru

On-site
INR 1,800,000 - 2,600,000