Site Reliability Engineer (Managed Patching & Platform Automation)

Swift Software

Kuala Lumpur

On-site

MYR 180,000 - 300,000

Full time

4 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Swift is seeking an experienced Site Reliability/DevOps Engineer to join the automation and platform engineering team in Kuala Lumpur. You will help support the Managed Patching Service, design and build automation workflows, and drive reliability across enterprise-scale infrastructure.

You will implement Python-based solutions, integrate with ServiceNow and CI/CD tools, and contribute to a self-service operating model. A strong Linux background and SRE principles are essential for success.

Qualifications

  • 3+ years in Site Reliability/DevOps/Platform/Infrastructure engineering or related disciplines.
  • Hands-on experience with infrastructure automation and Linux administration.
  • Bachelor’s degree in CS/Engineering/IT or equivalent practical experience.

Responsibilities

  • Support the Managed Patching Service (MPS).
  • Contribute to enhancement of MPS and platform-driven self-service model.
  • Own assigned technical deliverables and service improvements.
  • Improve scalability, reliability, performance, and maintainability of the service.
  • Develop automation components and scripts; build enterprise automation.
  • Integrate automation with ServiceNow, inventory systems, CI/CD platforms, and related services.
  • Apply engineering best practices including testing, version control, peer review, and release management.
  • Develop automation using Python and related technologies.

Skills

Site Reliability Engineering
DevOps
Platform Engineering
Infrastructure Automation
Linux Administration
System Reliability

Education

Bachelor's Degree in Computer Science/Engineering/IT

Tools

Ansible Automation Platform
Python
ServiceNow
CMDBs
CI/CD tools
Power BI
Jenkins

Job description

ABOUT US

We’re the world’s leading provider of secure financial messaging services, headquartered in Belgium. We are the way the world moves value - across borders, through cities and overseas. No other organisation can address the scale, precision, pace and trust that this demands, and we’re proud to support the global economy. We’re unique too. We were established to find a better way for the global financial community to move value - a reliable, safe and secure approach that the community can trust, completely. We’re always striving to be better and are constantly evolving in an ever-changing landscape, without undermining that trust. Five decades on, our vibrant community reflects the complexity and diversity of the financial ecosystem. We innovate diligently, test exhaustively, then implement fast. In a connected and exciting era, our mission has never been more relevant. Swift now has a presence in 200+ countries and legal territories to serve a community of more than 12,000 banks and financial institutions.

Experience and Qualifications
  • 3 plus years of experience in Site Reliability Engineering, DevOps, Platform Engineering, Infrastructure Engineering, or related disciplines
  • Proven experience building and operating enterprise-scale automation solutions
  • Strong hands-on experience in infrastructure automation, Linux administration, and system reliability
  • Experience working within large-scale enterprise environments
  • Bachelor's Degree in Computer Science, Engineering, Information Technology, or equivalent practical experience
Key Responsibilities
  • Support the Managed Patching Service (MPS)
  • Contribute to the enhancement and continuous improvement of the Managed Patching Service (MPS)
  • Take ownership of assigned technical deliverables and service improvements
  • Support the scalability, reliability, performance, and maintainability of the service
  • Help implement engineering solutions that meet enterprise and regulatory requirements
  • Participate in the evolution of MPS toward a platform-driven and self-service operating model
  • Design and Build Enterprise Automation Solutions
  • Design, develop, and maintain automation workflows using Ansible Automation Platform
  • Develop reusable automation components, scripts, and operational tooling
  • Integrate automation solutions with ServiceNow, inventory systems, CI/CD platforms, and related enterprise services
  • Apply engineering best practices including testing, version control, peer review, and release management
  • Develop automation and integrations using Python and related technologies where required
  • Platform Integration and Service Engineering
  • Support onboarding, subscription, scheduling, and maintenance window capabilities within MPS
  • Improve service delivery through automation and standardization
  • Contribute to platform enhancements that improve user experience and operational efficiency
  • Assist in building scalable integration patterns across enterprise platforms
  • Support initiatives that reduce manual effort and improve service adoption
  • Reliability, Observability and Reporting
  • Implement and maintain operational monitoring, logging, and reporting capabilities
  • Contribute to the definition and measurement of service reliability objectives and operational metrics
  • Improve visibility of patching outcomes, compliance status, and service health
  • Support the development of dashboards and reporting solutions for operational and regulatory requirements
  • Identify opportunities to improve service reliability and reduce operational complexity
  • Incident, Problem and Operational Management
  • Investigate and resolve complex technical issues affecting service availability or performance
  • Participate in incident response, troubleshooting, and service recovery activities
  • Contribute to root cause analysis and corrective actions following incidents
  • Develop and maintain operational documentation, runbooks, and troubleshooting guides
  • Support continuous improvement initiatives to improve service stability and resilience
  • Compliance and Governance
  • Support compliance with enterprise security, risk, and regulatory requirements
  • Ensure automation workflows maintain appropriate traceability and auditability
  • Contribute to the implementation of governance controls and operational standards
  • Support evidence collection and reporting requirements for audits and compliance reviews
  • Assist in maintaining service documentation and operational records
  • Platform Operations and Automation Engineering
  • Contribute to infrastructure automation and platform engineering initiatives beyond Managed Patching Service
  • Apply automation and SRE practices to improve operational efficiency and reliability across related platform services
  • Support service transition, operational readiness, and continuous improvement activities
  • Collaborate with engineering teams to identify automation opportunities and operational improvements
  • Participate in shared engineering responsibilities aligned with evolving business priorities
  • Collaboration and Technical Contribution
  • Collaborate closely with Infrastructure, Security, Architecture, Service Management, and Engineering teams
  • Work with engineers across teams to identify and resolve technical challenges
  • Participate in design discussions, solution reviews, and technical workshops
  • Share knowledge, best practices, and lessons learned with team members
  • Provide guidance and mentoring to less experienced engineers when required
  • Support service adoption by collaborating with stakeholders and platform consumers
Required Skills
  • Strong expertise in Ansible Automation Platform
  • Strong Linux administration and troubleshooting experience (RHEL preferred)
  • Experience developing automation solutions using scripting languages such as Python
  • Experience integrating enterprise platforms such as ServiceNow, CMDBs, monitoring solutions, and CI/CD tools
  • Strong understanding of Site Reliability Engineering principles and operational practices
  • Experience managing and supporting large-scale infrastructure environments
  • Good understanding of automation governance, change management, and operational controls
  • Strong analytical and problem-solving skills
  • Strong communication and collaboration skills
Preferred Skills
  • Experience with CloudBees, Jenkins, GitHub Actions, or similar CI/CD platforms
  • Familiarity with infrastructure as code practices
  • Experience with observability, monitoring, and enterprise reporting solutions
  • Experience with Power BI or similar reporting tools
  • Experience working in regulated or financial services environments
  • Exposure to platform engineering or self-service operational models
What Success Looks Like (6–12 Months)

Managed Patching Service operates reliably and efficiently within its defined scope Automation capabilities are enhanced with reduced manual intervention Service onboarding and operational processes become increasingly standardized Operational and compliance reporting is available and trusted by stakeholders Service reliability and operational performance improve through continuous enhancement Strong collaboration is established across engineering and support teams Contributions made to broader infrastructure automation and platform engineering initiatives Knowledge is actively shared within the team, helping raise overall engineering capability

What we offer

We give you the freedom to be yourself. We are creating an environment of unique individuals - like you - with different perspectives on the financial industry and the world. A diverse and inclusive environment in which everyone’s voice counts and where you can reach your full potential. We are committed to an inclusive and accessible recruitment process. If you require a reasonable accommodation related to accessibility during your application or interview, please contact accessibility-Sysgroup@swift.com or indicate this in your application. Please note that this mailbox is not monitored for general recruitment enquiries and should only be used for accessibility or accommodation-related requests (for example related to vision, hearing or neurodiversity). All requests are confidential and will not affect your candidacy. Swift doesn’t stand still. We are constantly evolving and tirelessly innovating. Working at the intersection of finance and technology is a very exciting place to be right now. Swift is transforming cross-border payments, making them faster and more transparent than ever before. We are the way the world moves value - every instant of every day, in almost every country. We are proud that what we do has a critical impact on the global financial community and touches almost every aspect of the financial world. So, what you do at Swift has real impact too - an impact that matters every day. Which is why you matter to us. Joining Swift gives you unparalleled exposure to knowledge, expertise and technologies. If you have what it takes, you’ll be able to take on different career paths and have the opportunity to work in teams, departments and disciplines in countries around the world. Swift is unique. There is no other organisation like ours in the world driving the long-term future of the financial ecosystem. You’ll be surrounded by bright, customer-focused and intellectually curious people in a collaborative, friendly, open and inclusive environment. At Swift we are trusted every instant. Everything we do has an impact that matters. And as a member of our team, you are trusted to make your impact every day.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Lead DevOps Engineer - DevOps Enablement
Lead DevOps Engineer - DevOps Enablement

Swift Software • Kuala Lumpur

On-site
MYR 180,000 - 340,000
Associate IT Operations Specialist
Associate IT Operations Specialist

Swift Software • Kuala Lumpur

On-site
MYR 67,000 - 100,000
Site Reliability Engineer (Managed Patching & Platform Automation)
Site Reliability Engineer (Managed Patching & Platform Automation)

SWIFT SUPPORT SERVICES MALAYSIA SDN. BHD. • Kuala Lumpur

On-site
MYR 120,000 - 180,000
Medical
Education support
Dental
+5
Site Reliability Engineering (SRE) Intern
Site Reliability Engineering (SRE) Intern

Swift Software • Kuala Lumpur

On-site
MYR 13,000 - 27,000
Lead Site Reliability Engineer
Lead Site Reliability Engineer

Swift Software • Kuala Lumpur

On-site
MYR 120,000 - 240,000
Senior Java Developer
Senior Java Developer

Swift Software • Kuala Lumpur

On-site
MYR 120,000 - 180,000
Senior Customer Support Engineer
Senior Customer Support Engineer

Swift Software • Kuala Lumpur

On-site
MYR 120,000 - 180,000
Associate tester - Intern
Associate tester - Intern

Swift Software • Kuala Lumpur

On-site
MYR 17,000 - 28,000
Accessibility accommodation
Diverse and inclusive environment
Commitment to inclusive recruitment
Lead DevOps Engineer - DevOps Enablement
Lead DevOps Engineer - DevOps Enablement

Swift • Kuala Lumpur

On-site
MYR 180,000 - 300,000
Chapter Lead
Chapter Lead

Swift Software • Kuala Lumpur

On-site
MYR 180,000 - 260,000