Location: Jakarta, Indonesia
Thales is a global technology leader trusted by governments, institutions, and enterprises to tackle their most demanding challenges. From quantum applications and artificial intelligence to cybersecurity and 6G innovation, our solutions empower critical decisions rooted in human intelligence. Operating at the forefront of aerospace and space, cybersecurity and digital identity, we’re driven by a mission to build a future we can all trust.
Key Responsibilities
Run Operations & Incident Management
- Ensure the day-to-day operational health of cloud infrastructure (Private Cloud based on VMware and GCP-based Public Cloud)
- Troubleshoot infrastructure-related issues (OS, container, network, middleware, storage)
- Act as Level 2 support for incidents and service requests
- Monitor system health, analyze logs, and perform root cause analysis using GCP Cloud monitoring and Prometheus/Grafana
- Set up and fine‑tune alerts on usual Prometheus metrics
Maintenance & Changes
- Perform routine maintenance (patching, upgrades, backups)
- Participate in change planning, testing, and execution within the Change Management process
- Assist with capacity planning, cost and performance optimization
Cloud Engineering & Standardization
- Contribute to the standardization of cloud infrastructure components
- Maintain technical documentation and operational procedures
- Enforce security policies: access control, hardening, vulnerability remediation
Collaboration & Continuous Improvement
- Work closely with internal DevOps and SRE teams to improve reliability and automation
- Participate in post‑incident reviews and problem management
- Support the transition of new applications into operations
- Identify and propose improvements to reduce operational toil, promoting Infrastructure as Code
Technical Environment
- Cloud platforms: Private Cloud (VMware), Public Cloud S3ns (GCP – Compute Engine, GKE, Cloud Logging, IAM)
- Containers: Kubernetes, Helm, Docker, GitOps (FluxCD)
- OS: Linux (RHEL), Windows OS
- Tools: Monitoring – Prometheus, Grafana, Google Cloud Monitoring
- ITSM – ServiceNow
- CI/CD – Jenkins, GitLab CI, Ansible
- Security – IAM, Bastion, Patch Management, OS Hardening
- Scripting – Bash, PowerShell (basic); read Bash or Python scripts
Required Skills & Experience
- 3+ years in cloud infrastructure or system administration roles
- Hands‑on experience with VMware Private Cloud and/or Public Cloud (GCP preferred)
- Good understanding of networking, backup, OS, and storage fundamentals
- Familiar with monitoring, logging, and alerting in production environments
- Experience working in ITIL‑driven environments (Incident, Change, Problem)
- Scripting capabilities for automation (Shell or PowerShell)
- Experience with Helm Chart and Ansible Roles
- Experience with containerization and container orchestration
- Fluent in English
- Experience in an international environment
Soft Skills
- Strong analytical and troubleshooting skills
- Service‑oriented mindset
- Ability to work in a distributed and international team
- Clear written documentation habits
- Self‑motivated and proactive in resolving operational issues
Nice to Have
- Experience with SQL databases (PostgreSQL, MariaDB, SQL Server, Oracle)
- Understanding of how Terraform works
- Understanding of modern Authentication and authorization mechanisms (OIDC/SAML)
- Understanding of TLS encryption protocol
At Thales, we’re committed to fostering a workplace where respect, trust, collaboration, and passion drive everything we do. Here, you’ll feel empowered to bring your best self, thrive in a supportive culture, and love the work you do. Join us, and be part of a team reimagining technology to create solutions that truly make a difference – for a safer, greener, and more inclusive world.
Referrals increase your chances of interviewing at Thales by 2x