Get more replies from employers
Send a job-specific resume in minutes.
Unifocus is seeking a DevOps Engineer to support high-availability, 24/7 production systems and drive reliability, performance, and security for our hospitality-focused software platform. You will partner with development teams, implement scalable cloud infrastructure, and lead incident RCA and post-incident improvements.
The role requires strong Linux, cloud, Docker/Kubernetes, and IaC skills, with a hands-on automation mindset.
Unifocus is an integrated workforce management software platform offering intelligent automation for daily work orders management, housekeeping activities, facility maintenance, survey solutions, scheduling & labour management, and time & attendance built for the hospitality market and other dynamic scheduling environment.
We support hotels, restaurants, casinos, and more with our innovative web-based and mobile software suite. Some of the chains we work with include Hilton, Rosewood, Shangri La, Accor, IHG, Hoxton, Corinthia, Oetker Collection etc. We are a small but growing team, and you'll have opportunities to express yourself and make meaningful contributions to our products and the company.
The DevOps Engineer supports high-availability 24/7 production systems of moderate to high complexity and risk. The role performs ongoing application support for live production systems by diagnosing and resolving highly complex issues, identifying, recommending and implementations options for improving performance, maintainability, and operability; update existing practices and procedures, as defined by supervisor.
Own and support high-availability, 24×7 production systems, ensuring reliability, performance, scalability, and security in line with defined SLA/SLO commitments.
Act as the primary Technical Operations point of contact for development and product teams, providing guidance on system design, operability, and production readiness.
Provide L3/L4 production support, troubleshooting complex application, infrastructure, and platform issues beyond documented procedures.
Participate in a 12×7 on‑call rotation, responding to production alerts, mitigating incidents, and restoring services with minimal business impact.
Lead incident troubleshooting, coordination, root cause analysis (RCA), and post‑incident reviews, driving preventive and corrective actions.
Design, implement, and operate cloud infrastructure (AWS), including compute, networking, storage, IAM, and security components.
Build, maintain, and enhance Infrastructure as Code (IaC) and deployment automation using industry‑standard tools.
Operate and support Docker and Kubernetes‑based platforms, including deployments, scaling, upgrades, and runtime troubleshooting.
Design, implement, and maintain monitoring and observability solutions using metrics, logs, and alerts to proactively detect and resolve issues.
Automate operational tasks and workflows using scripting and tooling to reduce manual effort and improve operational efficiency.
Perform system installation, administration, patching, configuration, and upgrades for production infrastructure.
Our Culture Statement: Thriving Together, Achieving Greatness.
To support our culture mission, we have four core culture values of Unite, Inspire, Empower, and Excel. Each value representing a set of key traits that define how we live and breathe our culture every day. We UNITE globally, combining our diverse talents, perspectives, and expertise. With professionalism and a touch of fun, we inspire and empower each other to excel. Together, we deliver exceptional value, challenge norms, and leave a lasting impact within the hospitality industry.
Health insurance
Paid time off
A hybrid environment that promotes a healthy work‑life balance