Senior SRE, Infrastructure & Platform

F5 NETWORKS SINGAPORE PTE LTD

Singapore

On-site

SGD 150,000 - 210,000

Full time

11 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

F5 Networks Singapore PTE LTD is seeking a Senior Site Reliability Engineer to lead automation, build tooling, and internal platforms for a global, multi-datacenter infrastructure spanning PoPs worldwide.

You will work in a PCI-DSS compliant environment and participate in a 24x7 on-call rotation, building automation, tooling, and CI/CD pipelines to improve operational efficiency.

Qualifications

  • 5+ years in SRE, DevOps, or Infra Eng with strong coding focus.
  • Proficiency in Python; ability to build CLIs, APIs, and automation frameworks.
  • Expert Ansible skills: custom roles, modules/plugins, complex templating at scale.
  • Solid Linux (RHEL/CentOS) knowledge for debugging and robust automation.
  • Experience building GitLab-based CI/CD pipelines for infrastructure automation.
  • Hands-on with self-hosted Kubernetes: cluster operations and controllers.
  • AWS and Azure experience with an IaC mindset for automation.

Responsibilities

  • Design and develop internal tools, CLIs, and APIs to enable self-service and automation.
  • Integrate infrastructure systems (NetBox, Vault, Proxmox, monitoring, CI/CD) into automated workflows.
  • Develop and maintain API clients and libraries for infra services.
  • Develop Ansible roles, playbooks, and modules across a large DC footprint.
  • Automate bare-metal lifecycle: from iLO bootstrap to OS install and VM provisioning.
  • Design and maintain GitLab CI pipelines for multi-stage deployment and testing.
  • Build Kubernetes automation for cluster lifecycle, upgrades, and workload deployment.
  • Develop IaC automation for AWS and Azure resources across hybrid environments.
  • Instrument tooling with logging and metrics; automate remediation and toil reduction.
  • Automate PCI-DSS compliance workflows including audit trails and drift detection.

Skills

Go
Python
GitLab CI

Tools

NetBox
HashiCorp Vault
Proxmox
iLO
Kubernetes

Job description

We are looking for a Senior Site Reliability Engineerthat leads withkindness, andpossessesa strong software development background to join our Infrastructure Engineering team. Your primary focus will be building automation, tooling, and internal platforms that enable our team tooperatea global, multi-datacenter infrastructure spanninga growing number ofPoints of Presenceacross the globe.

You will work within a PCI-DSS compliant environment andparticipatein a 24x7 on-call rotation.

What You'll Do
Internal Tooling & Application Development
  • Design and develop internal tools, CLIs, and APIs (primarily in Go and Python) that enable infrastructure self-service, automate complex workflows, and improve operational efficiency
  • Build integrations between infrastructure systems -- connecting CMDB/IPAM (NetBox), secrets management(HashiCorpVault),hypervisor APIs (Proxmox), monitoring platforms, and CI/CD pipelines into cohesive automated workflows
  • DevelopandmaintainAPIclients and libraries for interacting with infrastructure services (ProxmoxAPI, VaultAPI,NetBoxAPI,iLORedfish,container registries)
  • Write well-tested, documented, and maintainable code with proper versioning, release processes, and code review practices
Infrastructure as Code & Ansible Development
  • Architect, develop, and refactor Ansible roles and playbooks across a large-scale inventory spanning 30+ datacenters, 80+ group variable files, and 40+ roles
  • Design reusable, composable Ansible role patterns that scale cleanly as the DC footprint grows -- new DCs should be deployable with minimal variable additions
  • Improve idempotency, error handling, and test coverage across the existing Ansible codebase
  • Develop custom Ansible modules, plugins, and lookup plugins where upstream modulesmay be insufficient (e.g., custom Vaultintegration,ProxmoxAPIinteractions,iLOautomation)
  • Automate bare-metal server lifecycle end-to-end:fromiLObootstrapthrough OS installation, hypervisor configuration, VM provisioning, and service deployment
CI/CD Pipeline Engineering
  • Design, write, and maintain GitLab CI pipelines for infrastructure automation, including multi-stage deployment workflows with linting, validation, canary testing, and regional rollout
  • Build pipeline patterns for safe infrastructure changes: staged rollouts, automated rollback, drift detection, and change validation
  • Create reusable pipeline templates and shared CI componentsthatstandardisehowinfrastructure changes are tested and deployed
  • Implement automated testing for Ansible roles and infrastructurechanges(molecule,ansible-lint, integration testing in ephemeroe environments)
Kubernetes & Container Platform Automation
  • Develop automation for self-hosted Kubernetes cluster lifecycle management: provisioning, upgrades, scaling, and disaster recovery
  • Buildandmaintaincontainerimage build pipelines, registry management, and image promotion workflows
  • Create Kubernetes operators or controllers (in Go) where custom automation of cluster-level concerns is needed
  • Automate workload deployment patterns, including Helm chart developmentandGitOpsworkflows
Cloud Infrastructure Automation
  • DevelopIaCandautomation for AWS and Azure resources, integrating cloud infrastructure with on-premises systems
  • Build automation that spans hybrid environments -- coordinating deployments acrossbare-metal,virtualized,and cloud targets from a unified pipeline
Observability & Reliability Engineering
  • Instrument internal tools and automation with proper logging, metrics, and tracing
  • Build automated remediation workflows that respond to monitoring alerts and reduce mean time to recovery
  • Develop reporting and dashboards that provide visibility into infrastructure state, automation success rates, and toil metrics
  • Identifyand automate away recurring operational toil; track and quantify toil reduction over time
Security & Compliance Automation
  • Automate PCI-DSS compliance workflows including CIS benchmark hardening, audit evidence collection, and configuration drift detection
  • Build automated secret rotation pipelinesusingHashiCorpVault
  • Develop security scanning integration into CI/CD pipelines (container image scanning, infrastructure configuration validation)
What We're Looking For
  • 5+ years of experience in an SRE, DevOps, or Infrastructure Engineering role with a strong emphasis on writing code and building automation
  • Proficiencyin Python, with experience building CLI tools, APIs(Flask/FastAPIorequivalent), and automation frameworks
  • Expert-level Ansible skills: custom role development, module/plugin authorship, complex Jinja2 templating, inventory management at scale, and CI/CD integration
  • Solid Linux systems knowledge (RHEL/CentOS) -- you need to understand thesystemsyou'reautomatingat a depth that lets you debug failures and design robust automation
  • Experience buildingandmaintainingCI/CDpipelines (GitLab CI preferred) for infrastructure automation, not just application builds
  • Production experience with self-hosted Kubernetes: cluster operations, controller/operator development, and workload automation
  • Practical AWS and Azure experience withanIaCmindset-- provisioning and managing cloud resources through automation, not console clicks
  • Experience with API-driven infrastructure management (RESTful APIs, Redfish/iLO, hypervisor APIs)
  • FamiliaritywithHashiCorpVaultor equivalent secrets management platforms, including programmatic integration
  • Understanding of PCI-DSS requirements as they apply to automated infrastructure management -- audit trails, change control, hardening automation
  • Strong software engineering fundamentals: version control workflows, code review, testing practices, documentation, and release management

The Job Description is intended to be a general representation of the responsibilities and requirements of the job. However, the description may not be all-inclusive, and responsibilities and requirements are subject to change.

Please note that F5 only contacts candidates through F5 email address (ending with @f5.com) or auto email notification from Workday (ending with f5.com or @myworkday.com).

Equal Employment Opportunity

It is the policy of F5 to provide equal employment opportunities to all employees and employment applicants without regard to unlawful considerations of race, religion, color, national origin, sex, sexual orientation, gender identity or expression, age, sensory, physical, or mental disability, marital status, veteran or military status, genetic information, or any other classification protected by applicable local, state, or federal laws. This policy applies to all aspects of employment, including, but not limited to, hiring, job assignment, compensation, promotion, benefits, training, discipline, and termination. F5 offers a variety of reasonable accommodations for candidates. Requesting an accommodation is completely voluntary. F5 will assess the need for accommodations in the application process separately from those that may be needed to perform the job. Request by contacting accommodations@f5.com.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior SRE, Infrastructure & Platform
Senior SRE, Infrastructure & Platform

F5 Networks, Inc.  • Singapore

Hybrid
SGD 120,000 - 180,000
Site Reliability Engineer III
Site Reliability Engineer III

F5 NETWORKS SINGAPORE PTE LTD • Singapore

Hybrid
SGD 120,000 - 180,000
Hybrid working mode
Career growth opportunities
Recognitions & Rewards
+4
Site Reliability Engineer III
Site Reliability Engineer III

F5 • Singapore

Hybrid
SGD 90,000 - 130,000
Hybrid working mode
Sr Sales Operation Specialist
Sr Sales Operation Specialist

F5 Networks, Inc.  • Singapore

Hybrid
SGD 70,000 - 110,000
Site Reliability Engineer II
Site Reliability Engineer II

F5 NETWORKS SINGAPORE PTE LTD • Singapore

On-site
SGD 150,000 - 190,000
Principal Solutions Engineer (Singapore)
Principal Solutions Engineer (Singapore)

F5 Networks, Inc.  • Singapore

Hybrid
SGD 180,000 - 260,000
Hybrid working mode
Career growth and development
Competitive pay and benefits
+1
Sales Operations Specialist III
Sales Operations Specialist III

F5 Networks, Inc.  • Singapore

On-site
SGD 42,000 - 56,000
Sr. SRE
Sr. SRE

United States Digital Space LLC • Singapore

On-site
SGD 120,000 - 180,000
On-site in Singapore (3 days/wk)
Solutions Engineer - Modern Apps (Singapore)
Solutions Engineer - Modern Apps (Singapore)

F5 Networks, Inc.  • Singapore

Hybrid
SGD 120,000 - 180,000
Tuition assistance for professional开发
Sales Operations Specialist III
Sales Operations Specialist III

F5 • Singapore

On-site
SGD 42,000 - 54,000