A complete application in a minute — tailored resume and cover letter, ready to send.
Firmus Technologies, based in Singapore, seeks an Engineering Manager to lead the Platform Engineering and Observability team. You are accountable for both people growth and the delivery of the platform that underpins the Firmus AI Cloud.
You’ll drive a product-like approach to teams, raising the bar on reliability, security, and developer experience while balancing cost and speed in a fast-growing environment.
Firmus Technologies is a global leader pioneering the development and operation of efficient AIinfrastructure from model to grid. Founded in Australia in 2019, our mission is to create the mostefficient AI infrastructure by combining cutting-edge technology with a steadfast commitment tosustainability.
At Firmus, we are unique in our approach. We design, build, and operate a new class of digitalinfrastructure – the AI Factory. Through our model-to-grid technology approach, we have pushed theboundaries of multi-generational liquid cooling systems, energy management, AI softwareorchestration, and construction. This co-designed approach from model to grid allows us to makeevery watt count and deliver low-cost AI tokens globally.
Firmus AI Cloud
Our large-scale GPU cloud platform, Firmus AI Cloud, is purpose-built to deliver energy-efficient AIcompute at scale. It empowers developers, enterprises, educational institutions, and governmentusers to train and deploy AI models with unmatched efficiency and cost savings. With an ever-growing suite of services and applications, we are committed to delivering a cloud experience that ismarket-leading, proprietary, and built to scale.
Why you’ll love working here
At Firmus, you’ll work at the intersection of sustainability and artificial intelligence in a fast-pacedenvironment powered by next-generation technology. You’ll be helping to transform an entireindustry — and you’ll feel it every day.
Our team is made up of true innovators and leaders in their fields, and as an emerging company, youwon’t be lost in a crowd. You’ll work closely with the founders, build a strong network, and see theimpact of your work first-hand as we democratise AI tools for everyone — more sustainably andmore affordably.
We believe great things happen when people from diverse backgrounds come together to do theirbest work and be their authentic selves. We are proud to be an equal opportunity employer.
ROLE
Firmus Technologies is seeking an Engineering Manager to lead the Platform Engineering and Observability team. You are accountable for both the people and the delivery: the engineers you grow and the platform software they ship. Your team builds and operates the platform beneath the Firmus AI Cloud, from bare-metal GPU compute and high-performance networking to the internal platform services, self-service tooling, and the observability platform that our engineering teams and customers depend on. You run these as products for the engineering teams that build on them, investing in self-service, reliability, and developer experience. You are the escalation point above first-line operations and own the cross-cutting decisions your team shares. This is a build-and-grow leadership role: you build the team that scales our AI platform toward gigawatt-scale AI factories across the regions and grow your scope as you earn it.
KEY RESPONSIBILITIES
Be the Single-Threaded Owner (STO) for multiple squads: their engineers and tech leads report to you, and you work alongside product, architecture, and delivery peers. Hire strong engineers, develop them through coaching and clear feedback, and manage performance directly. Own the org design, headcount planning, and engineering culture for your group, building a team that raises its own bar rather than depending on you.
Own end-to-end delivery for your team. Turn the product roadmap into sequenced, predictable delivery, shaping prioritisation with product and working with team leads on sprint planning and release gates. Own the cross-squads' dependencies that determine whether the platform ships on time, keep delivery health visible through metrics, and balance velocity against reliability and technical debt.
Stay technically credible and set the engineering quality bar across your squads. You are accountable for making sure the right cross-cutting decisions get made and driven to closure, such as build, buy, or open source, and the product SLAs your teams commit to. Keep design and code review rigorous, champion AI-assisted development so your teams ship faster without lowering the design, review, or security bar, and unblock the hard problems that stall delivery.
Own the reliability and security of the services your squads run in production. Act as the escalation point above the operations centre, accountable for L3 incident resolution, SLA-breach response, and post-mortems that convert into runbook and prevention work. Set the standards for SLOs, on-call, and change management, backed by the observability platform your team owns. Govern how AI-generated code and agentic workloads reach production, extend Firmus' existing SOC 2 Type 2 and ISO 27001 controls as the platform scales into new data centres, and own the cost efficiency of what your team operate, balancing reliability, performance, and speed against infrastructure spend.
Represent your team to engineering leadership and the CTO, clear about progress, risk, and the decisions you need. Own the engineering side of the customer relationship: lead customer technical briefings, and represent engineering directly when escalations turn on delivery, reliability, or architecture. Align with product, who own the roadmap and customer outcomes, with the architects on technical direction, and with delivery and operations so the platform ships and runs as one system, not a set of parts.
SKILLS AND EXPERIENCE
Education
Platform and Technical Depth
Operations, Security and Cost
Stakeholder and Business