Senior Site Reliability Engineer

Poshmark

Chennai

On-site

INR 1,200,000 - 1,800,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Poshmark is seeking a Site Reliability Engineer in Chennai, India, to ensure the health and performance of complex, web-scale systems. The role requires 4+ years of experience in systems engineering, UNIX-based large-scale operations, and 24/7 support of production environments. Responsibilities include deploying product features, developing monitoring tools, and participating in on-call rotations. Ideal candidates will have expertise in cloud infrastructure like AWS, CI tools such as Jenkins, and configuration management with Ansible. This position plays a crucial role in maintaining operational excellence in a fast-paced environment.

Qualifications

  • 4+ years of experience in Systems Engineering/Site Reliability Operations role.
  • 4+ years in a UNIX‑based large‑scale web operations role.
  • 4+ years of experience in 24/7 support for large scale production environments.

Responsibilities

  • Serve as a primary point responsible for overall health and performance of services.
  • Assist in deploying new product features.
  • Develop tools for rapid deployment and monitoring in UNIX environment.

Skills

Systems Engineering
Site Reliability Operations
UNIX
Cloud Infrastructure (AWS, GCP, Azure)
Continuous Integration (Jenkins)
Configuration Management (Ansible)
Monitoring Tools (Nagios, New Relic, Graphite)
Scripting/Coding
DevOps Tools (Terraform, Docker, Kubernetes)

Job description

We’re looking for an experienced Site Reliability Engineer to fill the mission‑critical role of ensuring that our complex, web‑scale systems are healthy, monitored, automated, and designed to scale. You will use your background as an operations generalist to work closely with our development teams from the early stages of design all the way through identifying and resolving production issues. The ideal candidate will be passionate about an operations role that involves deep knowledge of both the application and the product, and will also believe that automation is a key component to operating large‑scale systems.

  • Familiarize with poshmark tech stack and functional requirements.
  • Get comfortable with automation tools/frameworks used within cloudops organization and deployment processes associated with.
  • Gain in depth knowledge related to related product functionality and infrastructure required for it.
  • Start contributing by working on small to medium scale projects.
  • Understand and follow on call rotation as a secondary to get familiarized with the on call process.
  • Execute projects related to comms functionality, independently, with little guidance from lead.
  • Create meaningful alerts and dashboards for various sub‑system involved in targeted infrastructure.
  • Identify gaps in infrastructure and suggest improvements or work on it.
  • Get involved in on‑call rotation.
Responsibilities
  • Serve as a primary point responsible for the overall health, performance, and capacity of one or more of our Internet‑facing services.
  • Gain deep knowledge of our complex applications.
  • Assist in the roll‑out and deployment of new product features and installations to facilitate our rapid iteration and constant growth.
  • Develop tools to improve our ability to rapidly deploy and effectively monitor custom applications in a large‑scale UNIX environment.
  • Work closely with development teams to ensure that platforms are designed with “operability” in mind.
  • Function well in a fast‑paced, rapidly‑changing environment.
  • Participate in a 24x7 on‑call rotation
Desired Skills
  • 4+ years of experience in Systems Engineering/Site Reliability Operations role is required, ideally in a startup or fast‑growing company.
  • 4+ years in a UNIX‑based large‑scale web operations role.
  • 4+ years of experience in doing 24/7 support for large scale production environments.
  • Battle‑proven, real‑life experience in running a large scale production operation.
  • Experience working on cloud‑based infrastructure e.g. AWS, GCP, Azure.
  • Hands‑on experience with continuous integration tools such as Jenkins, configuration management with Ansible, systems monitoring and alerting with tools such as Nagios, New Relic, Graphite.
  • Experience scripting/coding.
  • Ability to use a wide variety of open source technologies and tools.
Technologies we use:
  • Terraform, Packer, Jenkins, Datadog, Kubernetes, Docker, Ansible and other DevOps tools.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Systems Engineer - Devops Engineer
Senior Systems Engineer - Devops Engineer

Thyrocare • Bengaluru

On-site
INR 1,200,000 - 2,200,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Saama • Chennai District

On-site
INR 1,200,000 - 1,800,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Headout • Bengaluru

On-site
INR 1,200,000 - 1,800,000
Associate Senior Site Reliability Engineer
Associate Senior Site Reliability Engineer

Global Payments Inc. • Pune District

On-site
INR 1,500,000 - 2,100,000
Staff Site Reliability Engineer
Staff Site Reliability Engineer

PowerToFly • Gurgaon

On-site
INR 1,800,000 - 3,000,000
Staff Site Reliability Engineer
Staff Site Reliability Engineer

Stryker Group • Gurugram District

On-site
INR 1,200,000 - 1,800,000
Lead Site Reliability Engineer
Lead Site Reliability Engineer

Sierra Ventures • Bengaluru

On-site
INR 3,500,000 - 5,500,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

AcquireX • Pune District

On-site
INR 1,200,000 - 1,800,000
Health insurance
Flexible working hours
Training opportunities
Senior Site Reliability Engineer
Senior Site Reliability Engineer

iLink Digital • Chennai

On-site
INR 1,200,000 - 1,800,000
Senior Site Reliability Lead
Senior Site Reliability Lead

Generac • Pune District

On-site
INR 3,000,000 - 6,500,000