Senior Site Reliability Engineer

OutSystems

Bengaluru

Hybrid

INR 2,400,000 - 3,800,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

OutSystems in Bengaluru is seeking a Site Reliability Engineer to join a hybrid/onsite setup. You will lead reliability initiatives, define SLOs/SLAs, and design scalable, secure infrastructure while collaborating with product teams to ensure high availability and performance across the platform.

The role emphasizes observability, automation, and on-call responsibilities to support 24/7 production systems in a growing enterprise AI environment.

Qualifications

  • BS/MS in Computer Science or Equivalent
  • 6+ years of Site Reliability Engineering experience managing infrastructure at scale
  • End-to-end project delivery experience
  • Experience with Hadoop and Kubernetes infrastructure
  • Advanced knowledge of Linux, Networking and Containers
  • Proficiency in Python or GoLang (or similar)
  • Strong troubleshooting and debugging skills
  • Fluency in English with excellent communication
  • Prompt engineering exposure and AI native tooling familiarity

Responsibilities

  • Lead and onboard services to reliability tenets
  • Establish and maintain SLOs/SLAs
  • Design scalable, reliable cloud-native infrastructure
  • Collaborate with development teams for resilient systems
  • Implement monitoring, alerting, logging and tracing solutions
  • Lead incident response and post-mortems
  • Automate operational tasks with focus on rapid detection/recovery
  • Programming in Python with Gen AI tooling for automation
  • Foster continuous improvement and knowledge sharing
  • Communicate reliability updates to stakeholders
  • Participate in on-call 24/7 production support

Skills

Python
GoLang
Linux
Networking
Containers
English fluency
Communication
Problem solving
Prompt engineering awareness

Education

BS/MS in Computer Science or Equivalent

Tools

Kubernetes (CKA/CKAD)
AWS
Terraform
CI/CD tooling

Job description

There are NO limits to your career: come shape the future and be part of a truly unique global culture at OutSystems! Hybrid / Onsite in Bangalore

Site Reliability Engineering (SRE) is a discipline that incorporates aspects of software engineering and applies them to infrastructure and operations problems. The main goals of SRE are to create scalable and highly reliable systems. Our SREs ensure our production systems' reliability, performance, and scalability while enabling rapid development and deployment of new features and services. SREs at OutSystems work closely with development teams, acting as an extension of the team, in adopting the reliability tenets with the shared goal of meeting Service Level Objectives (SLOs) and thus delivering a smooth and frictionless Customer Experience.

Site Reliability Engineer Role
  • Lead and onboard services and teams to the reliability tenets;
  • Establish and maintain Service Level Objectives (SLOs) and Service Level Agreements (SLAs);
  • Design and implement scalable, reliable, and secure infrastructure, while ensuring cloud-native best practices;
  • Collaborate with software development teams to ensure systems are resilient (observable, fault-tolerant, recoverable, scalable) and performant;
  • Implement monitoring, alerting, logging, and tracing solutions to detect and respond to incidents;
  • Lead incident response efforts, ensuring quick resolution and minimal downtime, and conduct RCA/post-mortems;
  • Automate every operational task, with a special focus on fast incident detection & recovery;
  • Programming in Python supported by Gen AI tooling to accelerate development of mission critical automation and tools.
  • Foster a culture of continuous improvement and knowledge sharing;
  • Communicate effectively with stakeholders, providing updates on system reliability and performance;
  • Participate in on-call rotation to provide 24/7 support for production systems.
Site Reliability Engineering Performance Indicators

The main KPIs that aid in understanding the impact and success of the SRE function at OutSystems are:

  • SLA and Service Level Objectives (SLO) compliance;
  • SLO Coverage and Detection Ratio;
  • MTTA - Mean time to acknowledge;
  • MTTR - Mean time to resolve.
Qualifications and Skills
  • BS/MS in Computer Science or Equivalent
  • 6+ years of experience in Site Reliability Engineering, managing infrastructure and services at scale
  • History of end-to-end project delivery
  • Experience managing Hadoop and Kubernetes infrastructure and related services, or equivalent experience
  • Advanced knowledge of Linux, Networking, and Containers
  • Proficiency in at least one high-level programming language (Python, GoLang etc.).
  • Strong troubleshooting and debugging skills.
  • Fluency in English and excellent communication skills.
  • An understanding or hands-on experience with Prompt engineering in software development; Familiarity with AI Native IDEs or AI Assistants such as Cursor, GitHub CoPilot, and Claude.
Soft Skills
  • Communication - able to communicate effectively (in English) both orally and written showing empathy for the other person;
  • Collaboration - Proactive collaboration and presentation skills to effectively communicate ideas and represent the deliverables and needs of the SRE team with leadership.
  • Humbleness - accepts mistakes and acts accordingly, with a humble attitude, apologizing for them and mitigating them ASAP to avoid higher impact.
  • Accountability - takes ownership of problems and makes sure to see them through. Even if he does not have all the necessary knowledge to move on alone, can involve the right people to reach closure.
  • Negotiation Skills - has tough and politically complex conversations with colleagues and customers, defusing disagreements and leading towards a mutual agreement and understanding of all parties involved.
  • Process Oriented - is organized and able to properly follow defined processes, whilst being able to properly challenge inefficient processes and suggest improvements.
  • Problem-solving - Has a top-down approach to problems, breaking them into smaller pieces and solving them by starting with a wider scope and narrowing it down as the analysis progresses. Has critical thinking, so can analyze information objectively and make a reasoned judgment.
Technical Skills
  • Experience in any of the following is valued, but not fully required: Ability to establish, monitor, and improve Service Level Objectives (SLOs), Indicators (SLIs), and Agreements (SLAs) in line with business needs.
  • Containerization technologies and orchestration platforms, mainly Kubernetes and EKS (CKA, CKAD, CKS certifications are valued); Experience with automation and Infrastructure as Code (IaC) tools, such as AWS CloudFormation, Terraform, Puppet, Chef, Spacelift, etc;
  • Experience with Python, Go, Bash/Shell scripting, or other automation tools/languages;
  • Familiarity with AWS services like EC2, RDS, ELB, CloudFront, Lambda, etc;
  • Proficiency in monitoring and troubleshooting complex distributed systems;
  • Experience with Grafana, ELK stack, Prometheus, or others;
  • Strong understanding of designing resilient and fault-tolerant systems;
  • Expertise in debugging complex distributed systems.

More about OutSystems

OutSystems is a leading AI Development Platform built for the enterprise. Global organizations trust OutSystems to rapidly build mission-critical apps and agents, modernize legacy processes with agentic systems, and govern their entire AI portfolio across complex regulatory environments, all on one unified platform. As the future becomes agentic, our customers need us now more than ever. While AI has opened the door to extraordinary possibilities, most large organizations find themselves stuck on one side of the "enterprise gap" because AI by itself doesn't solve their complex use cases and business challenges. OutSystems bridges the "enterprise gap" by combining the speed of generative AI with a deterministic, enterprise-grade framework. We provide the tools for teams of any size to deliver high-quality, reliable AI solutions that drive real business impact. We are looking for passionate, talented, and motivated people to join us as we empower organizations to build, deploy, and scale the next generation of enterprise software. While we are leading the charge into the agentic era, our mission is broader: we are the platform enterprise leaders trust to evolve their entire business, accelerating innovation through secure, governed human‑AI collaboration. OutSystems is a global company, with more than 900k developer community members, 1,700 employees, more than 600 partners, and thousands of active customers in over 75 countries and across 21 industries. Founded in 2001, OutSystems now has offices in the United States, United Kingdom, the Netherlands, Portugal, Germany, the UAE, Japan, Hong Kong, Malaysia, Australia, India, and Singapore, and includes a thriving, worldwide community of remote employees. Our customers are some of the world's most recognizable

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Team Lead - Security Analyst
Team Lead - Security Analyst

Visa Hunt • India

On-site
INR 4,000,000 - 7,000,000
Field Marketing Manager - India
Field Marketing Manager - India

OutSystems • Bengaluru

On-site
INR 1,500,000 - 2,100,000
Health & Wellness
Financial Security (EPF & Gratuity)
Work-Life Balance
+2
Field Marketing Manager - India
Field Marketing Manager - India

OutSystems, Inc. • Bengaluru

On-site
INR 1,500,000 - 2,100,000
Field Marketing Manager - India
Field Marketing Manager - India

Outsystems India Private Limited • Bengaluru

Hybrid
INR 2,500,000 - 4,500,000
Competitive compensation
Health coverage for you and dependents
EPF and Gratuity
+4
AI Quality Data Tester (Manual & Automation)
AI Quality Data Tester (Manual & Automation)

OutSystems, Inc. • Bengaluru

On-site
INR 1,500,000 - 2,100,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Employ Connect • Bengaluru

Hybrid
INR 4,000,000 - 7,000,000
OutSystems Software Developer
OutSystems Software Developer

Ceres India • Bengaluru

On-site
INR 1,200,000 - 2,000,000
Hands-on experience with OutSystems
Collaboration with a mission-driven team
Career growth opportunities
Senior Software Engineer- Site Reliability Platform
Senior Software Engineer- Site Reliability Platform

UiPath • Bengaluru

Hybrid
INR 2,600,000 - 4,200,000
Site Reliability Engineer - Mumbai - India
Site Reliability Engineer - Mumbai - India

plantemoran • Mumbai City

On-site
INR 1,800,000 - 3,000,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Employ • Bengaluru

On-site
INR 4,000,000 - 6,000,000
Remote-first
Flexible scheduling
Paid time off
+1