Senior Database Reliability Engineer (DBRE) We are looking for a Senior Database Reliability Engineer (PostgreSQL) to join an established software company. In this role, you'll contribute to the reliability, performance, and scalability of critical backend infrastructure and database systems.
Requirements
- 5+ years of professional experience in database administration
- 5+ years of hands-on PostgreSQL experience, including designing and operating production clusters
- Strong knowledge of PostgreSQL operations: replication, query tuning, locking, indexes, vacuum and bloat management, major upgrades, backups, and point-in-time recovery
- 2+ years of experience administering MongoDB in production
- Experience running highly available databases, including failover, recovery, and disaster-recovery planning
- Linux administration experience, including networking, storage, and performance troubleshooting
- Automation experience with Ansible
- Experience with Terraform/OpenTofu or a comparable tool such as Puppet or Chef
- Practical experience with AI-assisted engineering tools and readiness to use them extensively in daily work
- English at upper-intermediate level or higher, spoken and written
Nice to Have
- Hands-on ClickHouse operations experience
- Experience with cloud platforms such as AWS or Google Cloud
- Experience with Redis, including Sentinel and common cache failure modes
- Familiarity with Percona Backup for MongoDB
- Experience building self-service database provisioning or DBaaS-style workflows
- Experience with database observability: Grafana, SLOs, and alert tuning
Responsibilities
- Operate and improve production PostgreSQL clusters, covering high availability, replication, failover, and major version upgrades
- Tune PostgreSQL performance across query plans, indexes, locks, vacuum and bloat control, and capacity planning
- Manage backups, point-in-time recovery, and regular restore testing across the database estate
- Support MongoDB and Redis in production, troubleshooting incidents and reviewing access and data-safety changes
- Learn the existing ClickHouse environment and take shared ownership of its day-to-day operations
- Automate routine DBA work - provisioning, access management, backups, and health checks - using Ansible, Terraform/OpenTofu, and CI/CD pipelines
- Build self-service capabilities that let engineering teams request databases and credentials without manual DBA involvement
- Maintain and improve observability through dashboards, metrics, SLOs, and alert rules