Lead DevOps Engineer

SG-EDTS

Jakarta Pusat

On-site

IDR 446,400,000 - 892,800,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

SG-EDTS is seeking an experienced DevOps Engineer/SRE to design and manage cloud infrastructure on AWS and GCP, and to build robust Kubernetes environments for production. You will develop CI/CD pipelines, implement IaC with Terraform and Ansible, monitor and secure cloud resources, and collaborate across engineering teams to drive reliability and performance.

You will also lead incident responses, implement HA/DR strategies, optimize costs, and mentor junior engineers while contributing to

Qualifications

  • Bachelor’s degree in a related field.
  • Minimum 7 years in DevOps/SRE/Cloud engineering roles.
  • Minimum 2–3 years in Senior/Lead roles.
  • Proven experience with production environments and HA infra.
  • Cloud platforms AWS and GCP, Kubernetes (EKS/GKE/self-managed), Docker, Terraform, Ansible, GitLab CI/GitHub Actions/Jenkins, Linux admin, networking basics.
  • Monitoring with Prometheus, Grafana, ELK/OpenSearch; scripting with Bash/Python; GitOps with Argo CD.
  • Cloud security IAM, secrets management, DevSecOps; cost optimization; incident RCA.
  • Experience across DevOps lifecycle, code reviews, architecture reviews, mentoring.
  • Certifications are a plus (AWS/GCP/Professional DevOps).

Responsibilities

  • Design, implement, and manage cloud infrastructure on AWS and GCP.
  • Build and optimize production Kubernetes platforms for scalability and HA.
  • Develop, maintain, and optimize CI/CD pipelines.
  • Implement IaC using Terraform and Ansible to automate provisioning.
  • Monitor, troubleshoot, and tune cloud resources, apps, and platforms.
  • Manage cloud security, IAM, secrets, and DevSecOps practices.
  • Collaborate with Development, QA, Security, and Infra teams.
  • Design HA/DR, backup and business continuity strategies.
  • Optimize cloud costs and resource utilization.
  • Lead incident response and RCA for critical events.
  • Create and maintain technical docs and platform standards.
  • Provide mentoring, reviews, and guidance to engineers.

Skills

AWS
GCP
Kubernetes
Docker
Terraform
Ansible
CI/CD
Git
Linux admin
Networking
Prometheus
Grafana
ELK/OpenSearch
Python Bash
GitOps Argo CD
Istio
Cilium
Chaos Engineering
Crossplane
IaC automation
Security/DPS

Education

Bachelor's degree in Computer Science/Information Systems/IT/Software Engineering

Tools

Argo CD
GitLab CI
GitHub Actions
Jenkins
Terraform
Ansible
Docker
Kubernetes

Job description

What You Will Do:
  • Design, implement, and manage cloud infrastructure on Amazon Web Services (AWS) and Google Cloud Platform (GCP) to support business and operational requirements
  • Build, manage, and optimize Kubernetes platforms for production environments, ensuring scalability, high availability, and reliability
  • Develop, maintain, and optimize CI/CD pipelines to enable fast, secure, and reliable application deployments
  • Implement Infrastructure as Code (IaC) using Ansible to automate infrastructure provisioning, configuration, and management
  • Perform monitoring, troubleshooting, capacity planning, and performance tuning for cloud infrastructure, applications, and platforms
  • Manage cloud security, including Identity and Access Management (IAM), secrets management, and the implementation of DevSecOps best practices
  • Collaborate closely with Development, QA, Security, and Infrastructure teams to support the software development and delivery lifecycle
  • Design and ensure the implementation of High Availability (HA), Disaster Recovery (DR), backup strategies, and business continuity solutions in accordance with organizational standards
  • Optimize cloud resource utilization and implement cost optimization strategies to improve operational efficiency
  • Lead incident response, problem management, and conduct Root Cause Analysis (RCA) for critical incidents to drive continuous improvement
  • Create and maintain technical documentation, platform standards, and DevOps best practices to ensure consistency and knowledge sharing
  • Provide technical guidance, mentoring, and code reviews to team members, fostering engineering excellence and continuous skill development
Person We Are Looking For:
  • Bachelor's degree in Computer Science, Information Systems, Information Technology, Software Engineering, or a related field
  • Minimum 7 years of experience as a DevOps Engineer, Site Reliability Engineer (SRE), or Cloud Engineer
  • Minimum 2–3 years of experience in a Senior or Lead role
  • Proven experience managing production environments and high-availability infrastructures
  • Mandatory Skills Cloud Platforms: Amazon Web Services (AWS) and Google Cloud Platform (GCP). Kubernetes (EKS, GKE, and self-managed Kubernetes). Docker and container platforms. Infrastructure as Code (IaC) using Terraform and Ansible.CI/CD tools such as GitLab CI, GitHub Actions, and Jenkins. Linux administration (Ubuntu, RHEL, Rocky Linux). Networking fundamentals, including TCP/IP, DNS, VPN, Load Balancers, Reverse Proxies, and TLS/SSL.
  • Monitoring and observability tools such as Prometheus, Grafana, ELK Stack, and OpenSearch.
  • Scripting using Bash and Python.Git and GitOps practices, including Argo CD.3. Technical Competencies
  • Design and manage scalable, secure, and highly available cloud infrastructure.
  • Build, operate, and optimize production-grade Kubernetes environments.
  • Develop, maintain, and optimize CI/CD pipelines and implement and manage Infrastructure as Code (IaC).Troubleshoot issues across cloud platforms, Kubernetes, Linux, networking, and application deployments
  • Implement DevSecOps practices, including Identity and Access Management (IAM), Secrets Management, and container security
  • Perform monitoring, capacity planning, performance tuning, backup management, disaster recovery, and cloud cost optimization
  • Lead the implementation and continuous improvement of DevOps practices and cloud platforms
  • Conduct code reviews and architecture reviews to ensure engineering quality and best practices
  • Mentor and provide technical guidance to engineering teams and lead incident response, problem management, and Root Cause Analysis (RCA) for critical incidents
  • Collaborate effective with Development, QA, Security, and Infrastructure teams to support the software delivery lifecycle
  • Knowledge of Service Mesh technologies such as Istio and Cilium
  • Experience with Chaos Engineering practices and Crossplane
  • Familiarity with AI-powered developer tools such as Claude Code, GitHub Copilot, Codex CLI, and Gemini CLI.6.
  • Cloud certifications are considered a strong advantage (AWS Certified Solutions, Professional AWS Certified DevOps Engineer – Professional Google Professional Cloud Architect Google Professional Cloud DevOps Engineer)
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Lead DevOps Engineer
Lead DevOps Engineer

Enterprise Digital Technology Services Edts • Jakarta Utara

On-site
IDR 390,600,000 - 725,400,000
On-site work at Wisma 46, Jakarta Pus
Competitive compensation
Collaborative environment
DevOps Engineer
DevOps Engineer

Merkle Innovation • Tangerang

On-site
IDR 350,000,000 - 550,000,000
Platform Engineer
Platform Engineer

ATI Business Group • Jakarta Pusat

On-site
IDR 180,000,000 - 240,000,000
Senior DevOps Lead: Cloud, Kubernetes & CI/CD Mastery
Senior DevOps Lead: Cloud, Kubernetes & CI/CD Mastery

SG-EDTS • Jakarta Pusat

On-site
IDR 446,400,000 - 892,800,000
Cloud & DevOps Engineer
Cloud & DevOps Engineer

PT Leap Digital Indonesia • Jakarta Utara

On-site
IDR 334,800,000 - 502,200,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

StraitsX • Indonesia

On-site
IDR 300,000,000 - 540,000,000
Cloud Data Infrastructure Engineer (SDE 3)
Cloud Data Infrastructure Engineer (SDE 3)

Kredivo Group • Daerah Khusus Ibukota Jakarta

On-site
IDR 673,740,000 - 1,010,612,000
Cloud Data Infrastructure Engineer (SDE 3)
Cloud Data Infrastructure Engineer (SDE 3)

Kredivo Group • Jakarta Timur

On-site
IDR 200,000,000 - 300,000,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

BookCabin • Jakarta Pusat

On-site
IDR 250,000,000 - 420,000,000
DevSecOps Engineer
DevSecOps Engineer

Lembaga Pengkajian Pangan, Obat-Obatan dan Kosmetika MUI • Bogor

On-site
IDR 279,000,000 - 502,200,000