Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.
Firmus Technologies seeks a Senior Database Reliability Engineer to own the operational datastore for Firmus AI Cloud and internal services. You will work on bare-metal, self-hosted, and cloud infrastructure, delivering reliable, scalable database platforms with low cognitive load for product teams.
You will champion HA, backup, and performance, coordinating with software engineers on migrations, data modeling, and telemetry. Join a company pushing sustainable AI infrastructure across APAC.
Firmus Technologies is a global leader pioneering the development and operation of efficient AI infrastructure across Asia Pacific.
Founded in Australia in 2019, our mission is to create the most efficient AI infrastructure by combining cutting-edge technology with a steadfast commitment to sustainability.
At Firmus, we are unique in our approach. We design, build, and operate a new class of digital infrastructure – the AI Factory. Through our model-to-grid technology approach, we have pushed the boundaries of multi-generational liquid cooling systems, energy management, AI software orchestration, and construction. For our customers, this approach allows us to make every watt count and deliver low-cost AI tokens globally.
Our large-scale GPU cloud platform, Firmus AI Cloud, is purpose-built to deliver energy-efficient AI compute at scale to customers.
It empowers developers, enterprises, educational institutions, and government users to train and deploy AI models with unmatched efficiency and cost savings. With an ever-growing suite of services and applications, we are committed to delivering a cloud experience that is market-leading, proprietary, and built to scale.
Firmus Technologies is seeking a Senior Database Reliability Engineer to join our Engineering and Technology team. You will own and evolve the operational datastore platform that Firmus AI Cloud and our internal platform services depend on. The work sits on bare-metal and self-hosted infrastructure as well as cloud, with proven recovery and measurable reliability as first-class product qualities. You build a paved road so product teams can provision and change databases safely, with low cognitive load, instead of waiting on a ticket queue.
Set and improve how we run databases for customer-facing and internal platform workloads. Choose the right pattern for the job: relational, document, NoSQL, or cache. Define connection pooling and database proxies. Define how we scale with vertical growth, read replicas, sharding or partitioning, and multi-site setups. Set HA and failover standards for shared multi-tenant systems. Keep tenant isolation clear and limit blast radius when something fails.
Own provisioning, config, upgrades, backup, restore, and decommissioning with infrastructure as code and solid automation. Set RPO and RTO. Prove restore and failover with regular drills. Give product teams safe self-service for routine database work so they do not wait on you for every change.
Own production health across on-premises and cloud database: SLOs, capacity, query performance, and database monitoring. Fix hard production issues such as lock contention, replication lag, memory or storage pressure, slow queries, and storage or network bottlenecks. Join on-call, lead post-mortems, and turn fixes into automation or better runbooks. Track storage and compute cost when you plan capacity.
Own access control, encryption in transit and at rest, audit logging, and change control for the databases. Extend Firmus SOC 2 Type 2 and ISO 2701 controls as we add sites and services. Keep backup, restore, and recovery evidence ready for audit. Keep privileged access tight.
Work with software and platform engineers on schema design, safe migrations, data modelling, and performance reviews before release. Work with data engineering and observability where your databases connect to their systems, such as CDC (Change Data Capture), read replicas, and database telemetry. Run design reviews that raise the bar. Join customer escalations when the issue is database reliability, performance, or recovery.
Full-time
At Firmus, we are committed to building a diverse and inclusive workplace. We encourage applications from candidates of all backgrounds who are passionate about creating a more sustainable future through innovative engineering solutions.