Sr S/W Engg, SRE & Platform Automation

Aziro

Bengaluru

On-site

INR 3,000,000 - 5,000,000

Full time

13 days ago
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Infoblox in Bangalore, India is seeking a Senior Software Engineer for Site Reliability Platform Automation. You will design and build production-quality components across reliability measurement, toil automation, and resilience testing, collaborating with DevOps, CloudOps, and product engineering teams.

The role emphasizes AI-assisted tooling, incident response, and platform enablement to reduce toil and improve reliability across global cloud services.

Qualifications

  • 5+ years of professional software engineering experience with reliability systems
  • Strong production coding ability in Go or Python
  • Cloud and Kubernetes experience (AWS or GCP)
  • Infrastructure-as-code with Terraform or equivalent
  • Experience owning software end-to-end including reliability and cost
  • Experience with observability and incident response

Responsibilities

  • Design and build production-quality components across reliability measurement and toil automation
  • Define SLIs, SLOs, error budgets, and observability improvements
  • Automate operational workflows including upgrades and DR tests
  • Develop AI-assisted tooling for incident summarization and remediation recommendations
  • Collaborate with DevOps, CloudOps, and product teams
  • Participate in on-call rotations and post-incident reviews

Skills

SRE
Go
Python
Cloud experience
Kubernetes
Observability

Education

Bachelor's degree in CS/CE/IT
Master's degree preferred

Tools

Terraform
AWS
GCP
OpenTelemetry
Prometheus
Grafana

Job description

Job Summary

Senior Software Engineer, Site Reliability Platform Automation We have an opportunity for a Senior Software Engineer, Site Reliability Platform Automation to join our SaaS Platform Engineering team in Bangalore, India, reporting to the Sr. Manager, Site Reliability Platform Engineering. In this pivotal role, you will build software that improves reliability, reduces operational toil, and enables self-service across Infoblox s global cloud networking and SaaS platforms. You will work across reliability measurement, toil automation, resilience test engineering, and platform enablement, partnering with DevOps, CloudOps, product engineering, architecture, security, and product management teams. Infoblox Engineering runs on a shared platform that includes EKS clusters, RDS databases, CI/CD pipelines, and developer tooling used by product engineering teams. You will help treat operations as a software problem by understanding manual workflows, instrumenting them, and replacing repetitive effort with reliable automation. You will also use AI-assisted engineering and operational tools responsibly to improve productivity, analytics, content generation, automation, and incident decision support, with appropriate human review and security controls.

Responsibilities
  • Be a Contributor - What You ll Do Design and build production-quality components and services across reliability measurement, toil automation, resilience testing, or self-service enablement
  • Define and implement SLIs, SLOs, error budgets, burn-rate alerts, observability improvements, and reliability evidence
  • Automate recurring operational work, including CVE remediation, EKS and RDS upgrades, third-party provider testing, and infrastructure workflows
  • Build chaos experiments, disaster recovery tests, regional failover exercises, and incident-response tooling
  • Create golden paths, Terraform modules, and policy-as-code guardrails that enable product engineering teams to provision and operate services independently
  • Instrument workflows before optimizing them, and measure outcomes such as toil reduction, adoption, reliability, cost, and developer experience
  • Develop AI-assisted tools for log and incident summarization, alert correlation, anomaly analysis, documentation, code assistance, and remediation recommendations
  • Validate AI-generated code, content, analytics, and operational recommendations through testing, peer review, security checks, and human approval
  • Partner with DevOps, CloudOps, and product engineering teams to understand real workflows and deliver capabilities they adopt
  • Participate in on-call rotations, incident response, after-action reviews, and follow-through engineering
Be Prepared - What You Bring
  • 5+ years of professional software engineering experience, including meaningful exposure to infrastructure, platform, DevOps, SRE, or reliability systems
  • Strong production coding ability in Go, Python, or a comparable language
  • Practical cloud and Kubernetes experience in AWS or GCP, including the ability to operate and debug Kubernetes beyond deploying manifests
  • Infrastructure-as-code experience with Terraform or an equivalent technology on production systems
  • Experience owning software or services end to end, including reliability, operational cost, security, and user impact
  • Demonstrated success replacing recurring manual work with software and measuring the resulting improvement
  • Experience with observability, incident response, SLOs, SLIs, error budgets, resilience testing, or disaster recovery
  • Experience building internal platforms, developer tooling, or self-service capabilities for other engineers
  • Practical experience using AI-assisted engineering or operational tools with sound judgment regarding privacy, security, accuracy, auditability, and human oversight
  • Bachelor s degree in Computer Science, Computer Engineering, Information Technology, or a related technical field; Master s degree preferred
Nice to have
  • Experience with Prometheus, Grafana, Loki, OpenTelemetry, ELK, Datadog, or comparable observability platforms
  • Experience with Jenkins, GitHub Actions, Argo, GitOps, or CI/CD platform engineering
  • Experience with chaos engineering or fault-injection tooling
  • Experience with RDS, PostgreSQL, database operations, or multi-region systems
  • Experience with Kyverno, OPA, Gatekeeper, or other policy-as-code technologies
  • Experience with Vault, Harbor, or secrets and artifact management
  • Experience operating multi-tenant or multi-region platforms or working in regulated environments
  • Experience applying LLM-based or agentic tooling to operational workflows

Disclaimer: This job posting has been aggregated from external source. Role details, content, and availability are subject to change. Applicants are advised to confirm the latest information directly on the company website before applying.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Sr Engg Manager, SRE & Platform Automation
Sr Engg Manager, SRE & Platform Automation

Aziro • Bengaluru

On-site
INR 4,500,000 - 7,000,000
Staff S/W Engg , SRE & Platform Automation Engg
Staff S/W Engg , SRE & Platform Automation Engg

Aziro • Bengaluru

On-site
INR 4,000,000 - 7,000,000
Sr Software Engineer PLG
Sr Software Engineer PLG

Aziro • Pune District

Hybrid
INR 1,400,000 - 2,100,000
Senior Software Engineer (Golang AND K8s AND Networking)
Senior Software Engineer (Golang AND K8s AND Networking)

Infoblox • Bengaluru

On-site
INR 3,500,000 - 5,500,000
Comprehensive health coverage
Flexible work options
Volunteer hours
Platform Engineer Lead
Platform Engineer Lead

Randstad • Bengaluru

On-site
INR 2,500,000 - 4,000,000
Software Engineer (Golang + Agentic AI)
Software Engineer (Golang + Agentic AI)

Infoblox • Bengaluru

On-site
INR 1,400,000 - 2,000,000
Comprehensive health coverage
Generous PTO
Flexible work options
+2
Associate - SRE - Platform Engineering
Associate - SRE - Platform Engineering

Jefferies Financial Group Inc. • Pune District

On-site
INR 1,600,000 - 2,200,000
Site Reliability Engineer
Site Reliability Engineer

SourcingXPress • Maharashtra

On-site
INR 700,000 - 1,800,000
Product Manager
Product Manager

Infoblox • Bengaluru

Hybrid
INR 1,400,000 - 2,200,000
Comprehensive health coverage
Generous PTO
Flexible work options
+2
Senior Site Reliability Engineer
Senior Site Reliability Engineer

BuildxPartners • Bengaluru Urban

Hybrid
INR 2,400,000 - 4,200,000