Senior Site Reliability Engineer

Akamai Technologies

Bengaluru

On-site

INR 2,400,000 - 3,400,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Akamai Technologies is seeking a Platform & Reliability Engineer to enhance the performance, availability, and scalability of our large distributed content delivery systems. You will collaborate with Product and Engineering to define measurable SLIs and objectives and provide technical guidance to ensure reliability and performance across platforms.

The role requires strong scripting (Python, Bash, JavaScript) for automation, deep UNIX/Linux expertise, and experience with monitoring tools

Qualifications

  • 5+ years of relevant experience in platform reliability or SRE.
  • Proficient in Python, Bash, and JavaScript for automation.
  • Experience with Prometheus, Grafana, Datadog, and Oracle SQL.
  • Strong UNIX/Linux system administration and troubleshooting skills.
  • Ability to collaborate across Product and Engineering teams.

Responsibilities

  • Improve performance, availability, and scalability of large distributed content delivery systems.
  • Define and monitor measurable SLIs/ SLOs with cross-functional teams.
  • Provide design feedback ensuring reliability and performance.
  • Monitor platform health, analyze data, and implement corrective actions.
  • Develop automation to reduce repetitive tasks and improve efficiency.
  • Participate in design reviews and guide scalable, robust architectures.
  • Keep up-to-date with cloud computing, DevOps, and SRE best practices.

Skills

Python
Bash
JavaScript
UNIX/Linux
Automation

Education

Bachelor's degree in Computer Science, Engineering, or related field

Tools

Prometheus
Grafana
ADBMS
Datadog
Oracle SQL

Job description

The Platform & Reliability Engineering team is responsible for defining, measuring, and optimizing the key performance indicators of delivery customers. Your expertise in software engineering and systems administration will be instrumental in building robust and resilient infrastructure.

Responsibilities
  • Working on Internet technologies to improve the performance, availability, and scalability of large distributed content delivery systems.
  • Collaborating with cross‑functional teams, including Product and engineering, to define and establish measurable Service Level Indicators and Objectives.
  • Providing technical expertise and feedback to ensure system designs and implementations align with reliability and performance requirements effectively.
  • Monitoring platform availability and performance, analyzing data to debug issues, and implementing corrective actions to prevent future occurrences.
  • Developing and implementing automation solutions aimed at enhancing operational efficiency while minimizing repetitive tasks.
  • Participating in design reviews and providing technical guidance to ensure designs meet requirements for scalability, performance, and robustness.
  • Staying updated on recent developments in cloud computing, DevOps, and SRE best practices, without missing emerging trends.
Qualifications
  • Have 5+ years of relevant experience and a Bachelor's degree in Computer Science, Engineering, or related field.
  • Demonstrate expertise in scripting languages like Python, Bash, and JavaScript to enable automation and create efficient tools.
  • Utilize monitoring and alerting tools such as Prometheus, Grafana, ADBMS, and Datadog effectively for creating and managing dashboards.
  • Work proficiently in UNIX/Linux environments, showcasing exceptional problem‑solving abilities and comprehensive system expertise.
  • Utilize Oracle SQL to perform data integrity checks, identify root causes of anomalies, and generate detailed reports.
  • Improve outcomes, embrace ongoing learning, and deliver automation‑centered operational excellence across various initiatives with proactive self‑direction.
  • Demonstrate customer focus, accountability, and exceptional communication and teamwork abilities across various cross‑functional groups.
Benefits

We support your health, well‑being, finances, and life beyond work. Our program includes FlexBase, allowing employees to work in ways that suit them best.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer
Senior Site Reliability Engineer

Akamai Technologies GmbH • Bengaluru

Hybrid
INR 3,500,000 - 5,200,000
FlexBase program
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Five9 • Bengaluru

On-site
INR 4,000,000 - 7,000,000
Resilience and Reliability Engineer
Resilience and Reliability Engineer

EY • Pune District, Gurugram District, Bengaluru

Hybrid
INR 1,800,000 - 2,800,000
Staff Site Reliability Engineer
Staff Site Reliability Engineer

PowerToFly • Gurgaon

On-site
INR 1,800,000 - 3,000,000
Staff Site Reliability Engineer
Staff Site Reliability Engineer

Stryker Group • Gurugram District

On-site
INR 1,200,000 - 1,800,000
Site Reliability Engineer(SRE)
Site Reliability Engineer(SRE)

MetaForgeIT • Hyderabad

On-site
INR 1,000,000 - 1,500,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Saama • Chennai District

On-site
INR 1,200,000 - 1,800,000
Lead Site Reliability Engineer
Lead Site Reliability Engineer

Sierra Ventures • Bengaluru

On-site
INR 3,500,000 - 5,500,000
Site Reliability Engineer, Cloud Platform
Site Reliability Engineer, Cloud Platform

Qualys • Maharashtra

On-site
INR 1,200,000 - 1,800,000
Site Reliability Engineer
Site Reliability Engineer

Snapmint • Gurugram District

On-site
INR 800,000 - 1,200,000