Sr Lead-Systems & Applications

Mphasis

Dallas (TX)

On-site

USD 120,000 - 170,000

Full time

10 days ago
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Mphasis is seeking a Systems Reliability Engineer to join as Technical Team Lead in Texas, USA. The role concentrates on SRE practices to ensure reliability, availability, and performance of our systems and applications.

You will lead a team and design scalable infrastructure using Kubernetes and Docker, while mentoring junior engineers. Proficiency in Golang, Grafana, and Bash or PowerShell is essential, with cloud platform experience highly valued.

Qualifications

  • Minimum 6 years of experience in Site Reliability Engineering (SRE).
  • Billing experience in Telecom domain.
  • Bachelor's degree in Computer Science, Information Technology, or a related field; cloud/SRE certs are a plus.

Responsibilities

  • Lead the SRE team to improve reliability and performance.
  • Design and maintain scalable infrastructure with Kubernetes and Docker.
  • Develop automation scripts to streamline operations.
  • Monitor systems with Grafana and respond to incidents.
  • Mentor junior team members and conduct post-incident reviews.
  • Collaborate with development teams to ensure reliable, scalable applications.

Skills

SRE principles
Kubernetes
Docker
Golang
Grafana
Bash
PowerShell
Communication
Collaboration
Incident response
Problem solving

Education

Bachelor's degree in CS/IT

Tools

Kubernetes
Docker
Grafana
CI/CD tooling

Job description

Role description

Job Title: Systems Reliability EngineerLocation: Texas, USAJob Summary:

We are seeking a highly skilled and motivated Systems Reliability Engineer to join our dynamic team as a Technical Team Lead. The ideal candidate will possess extensive experience in Site Reliability Engineering (SRE) and will be responsible for ensuring the reliability, availability, and performance of our systems and applications. This role requires a strong technical background, particularly in Kubernetes, Docker, and Golang, along with proficiency in monitoring and automation tools such as Grafana, Bash, or PowerShell.

Responsibilities:

  • Lead the SRE team in implementing best practices for system reliability and performance.
  • Design, build, and maintain scalable and resilient infrastructure using Kubernetes and Docker.
  • Develop and maintain automation scripts and tools to streamline operations and improve efficiency.
  • Monitor system performance and reliability using Grafana and other monitoring tools, responding to incidents and outages as they occur.
  • Collaborate with development teams to ensure that applications are designed with reliability and scalability in mind.
  • Conduct post incident reviews and implement improvements to prevent future occurrences.
  • Provide technical guidance and mentorship to junior team members.
  • Stay current with industry trends and emerging technologies to continuously improve our systems.

Mandatory Skills:

  • Strong experience in Site Reliability Engineering (SRE) principles and practices.
  • Proficiency in container orchestration using Kubernetes.
  • Experience with Docker for containerization.
  • Strong programming skills in Golang.
  • Familiarity with monitoring and visualization tools, particularly Grafana.
  • Proficient in scripting languages such as Bash or PowerShell.
  • Excellent problem-solving skills and the ability to work under pressure.
  • Strong communication and collaboration skills.

Preferred Skills:

  • Experience with cloud platforms (AWS, Azure, GCP).
  • Knowledge of CI CD pipelines and DevOps practices.
  • Familiarity with configuration management tools (e.g., Ansible, Puppet).
  • Understanding of networking concepts and protocols.
  • Experience with database management and optimization.

Years of Experience:

  • Must have at least 6 years of Experience
  • Must have billing experience in Telecom Domain

Qualifications:

Bachelor's degree in Computer Science, Information Technology, or a related field. Relevant certifications in cloud technologies or SRE practices are a plus.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

TalentDome Staffing • United States

On-site
USD 140,000 - 210,000
Senior SRE Lead: Kubernetes, Go & Cloud Reliability
Senior SRE Lead: Kubernetes, Go & Cloud Reliability

Mphasis • Dallas (TX)

On-site
USD 120,000 - 170,000
Lead SRE Engineer
Lead SRE Engineer

SIMARN Solutions • Charlotte (NC)

On-site
USD 120,000 - 160,000
Site Reliability Engineer
Site Reliability Engineer

JobCubby • Barrington (RI), Northern (KY)

On-site
USD 110,000 - 170,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Kovoro • Denver (CO), Northern (KY)

Hybrid
USD 150,000 - 190,000
Senior SRE Engineer
Senior SRE Engineer

Compunnel, Inc. • New Jersey

On-site
USD 140,000 - 190,000
Lead Associate – Service Reliability Engineering (SRE) – $140-170K Plus 20% Bonus
Lead Associate – Service Reliability Engineering (SRE) – $140-170K Plus 20% Bonus

ACCsurance, LLC • Town of Texas (WI)

On-site
USD 140,000 - 170,000
Site Reliability Engineer
Site Reliability Engineer

Stelvio Inc. • Town of Texas (WI)

On-site
USD 125,000 - 145,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

L'Oréal • United States

On-site
USD 150,000 - 230,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Axiom Pursuits • San Francisco (CA)

On-site
USD 150,000 - 180,000