Site Reliability Engineer II

Akamai Technologies

Bengaluru

On-site

INR 1,200,000 - 2,000,000

Full time

27 hours ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Akamai Technologies in Bengaluru, India, is hiring a Site Reliability Engineer II to reinforce our platform and reliability goals for the distributed content delivery network.

You will define KPIs, improve monitoring, automation, and collaborate with cross-functional teams; requires 2+ years experience, and strong Unix/Linux, Python, and monitoring tools.

Qualifications

  • 2+ years of relevant experience and a Bachelor's degree in CS/Engineering or related field.
  • Proficient in Python, Bash, or JavaScript.
  • Experience with Unix/Linux environments.
  • Familiar with monitoring/alerting tools (Prometheus, Grafana, Datadog).

Responsibilities

  • Improve the performance, availability, and scalability of large distributed content delivery systems.
  • Define and establish measurable SLIs and SLOs with cross-functional teams.
  • Enhance monitoring, alerting, and incident response processes.
  • Develop automation to reduce repetitive tasks and improve efficiency.
  • Participate in architecture and design reviews for scalability and resilience.
  • Stay updated on cloud, DevOps, and SRE best practices.

Skills

Python scripting
Unix/Linux
SRE fundamentals
Data analysis

Education

Bachelor's degree in Computer Science, Engineering, or related field

Tools

Prometheus
Grafana
Datadog
Oracle SQL

Job description

Job Description
Do you like collaborating across teams to solve complex problems?
Do you enjoy solving large scale distributed content delivery challenges?
Join our critical Platform and Reliability Engineering Team!

The Platform & Reliability Engineering team is responsible for defining, measuring, & optimizing the key performance indicators of delivery customers. Your expertise in software engineering and systems administration will be instrumental in building robust and resilient infrastructure.

Job Description
Do you like collaborating across teams to solve complex problems?
Do you enjoy solving large scale distributed content delivery challenges?
Join our critical Platform and Reliability Engineering Team!

The Platform & Reliability Engineering team is responsible for defining, measuring, & optimizing the key performance indicators of delivery customers. Your expertise in software engineering and systems administration will be instrumental in building robust and resilient infrastructure.

Partner with the best

As a Site Reliability Engineer II, this role involves shaping product futures by ensuring system reliability, scalability, and performance. Collaboration with product teams starts early in development. Responsibilities include defining key performance indicators, improving monitoring and alerting processes, enhancing operational responses, and resolving complex performance challenges.

As a Site Reliability Engineer II, you will be responsible for:
  • Working on Internet technologies to improve the performance, availability, and scalability of large distributed content delivery systems.
  • Collaborating with cross-functional teams to define and establish measurable Service Level Indicators (SLIs) and Service Level Objectives (SLOs).
  • Providing technical expertise and feedback to ensure system designs and implementations align with reliability and performance requirements effectively.
  • Monitoring platform availability and performance, analyzing data to resolve complex issues, and implementing solutions to prevent future occurrences.
  • Creating and implementing automation solutions to enhance operational efficiency while minimizing repetitive tasks.
  • Ensuring participation in architecture and design reviews to verify systems achieve exceptional scalability, performance, and resilience standards.
  • Staying updated on the newest developments in cloud computing, DevOps, and SRE practices to ensure expertise and effectiveness.
Do What You Love

To be successful in this role you will:

  • Have 2+ years of relevant experience and a Bachelor's degree in Computer Science, Engineering, or related field.
  • Validate data integrity, analyze anomalies, and generate reports using Oracle SQL for comprehensive root cause identification.
  • Demonstrate expertise in scripting languages like Python, Bash, or JavaScript without compromising quality or functionality.
  • Work with monitoring and alerting tools such as Prometheus, Grafana, ADBMS, and Datadog for metrics, alerts, dashboards, and troubleshooting.
  • Demonstrate expertise in Unix/Linux operating environments.
  • Demonstrate commitment to continuous learning and achieving operational excellence through automation and efficiency enhancements.
  • Adopt a customer-focused approach with a deep sense of responsibility and accountability, showcasing excellent interpersonal, written, and verbal communication abilities.
About Us

At Akamai, we make life better for billions of people, trillions of times a day. Whether you're streaming live events, scrolling social media, watching your favorite series, or managing your savings, we're the engine behind the scenes. We provide the world's most distributed platform from Cloud to Edge to help the giants of the digital world work faster and stay more secure, making the internet a better experience for everyone.

Our Focus Is Simple
Cloud and Edge:

Running apps closer to users for instant performance.

Security

Neutralizing threats before they ever reach your data.

Content Delivery

Scaling the world's biggest moments without a glitch.

AI

Enabling our customers to build, secure, and scale AI apps on the world's most distributed cloud platform.

Benefits at Akamai:

We support your health, well-being, finances, and life beyond work. See our benefits.

FlexBase adapts to your job's needs

Akamai's FlexBase program is yet another way we show our commitment to providing employees with an exceptional workplace experience. It's not about telling employees where to work; it's about supporting employees to do their best work.

We trust our incredible employees to work in ways that suit them best: at home, in an office, or a combination of both.

Connect with us on social and see what life at Akamai is like!

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer II
Site Reliability Engineer II

Akamai Career Site • India

On-site
INR 1,200,000 - 2,400,000
FlexBase program
Benefits package
Cloud Support Engineer II
Cloud Support Engineer II

Akamai • Bengaluru

Hybrid
INR 1,200,000 - 1,800,000
Health insurance
FlexBase program
Hybrid work options
Cloud Support Engineer II
Cloud Support Engineer II

Akamai Career Site • India

Hybrid
INR 1,000,000 - 2,000,000
FlexBase program
Cloud Support Engineer II
Cloud Support Engineer II

Akamai Technologies • Bengaluru

On-site
INR 900,000 - 1,500,000
Senior Software Engineer
Senior Software Engineer

Akamai Career Site • India

On-site
INR 1,800,000 - 2,400,000
FlexBase program
Benefits
Technical Support Engineer II
Technical Support Engineer II

Akamai Career Site • India

On-site
INR 600,000 - 1,200,000
FlexBase program
Technical Support Engineer II
Technical Support Engineer II

Akamai Technologies GmbH • India

On-site
INR 1,200,000 - 2,400,000
Technical Support Engineer II
Technical Support Engineer II

Akamai Technologies • Bengaluru

Hybrid
INR 1,200,000 - 1,700,000
FlexBase program
Flexible work options
Software Engineer
Software Engineer

Akamai Technologies GmbH • Bengaluru

Hybrid
INR 700,000 - 1,100,000
Security Consultant II
Security Consultant II

Akamai Career Site • India

Hybrid
INR 1,500,000 - 2,500,000
FlexBase program
Benefits at Akamai