Database Reliability Engineer

NEXT Ventures

Dubai

On-site

AED 360,000 - 480,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

NEXT Ventures in Dubai is seeking a Senior Database Reliability Engineer to own database performance, archiving pipelines, caching reliability, and BC/DR for the data tier that underpins FundedNext services. You will drive scaling and optimization initiatives as part of the Platform Engineering squad to keep databases fast, durable, and cost-efficient as trading data grows.

You will work with MySQL/PostgreSQL at scale, build data archiving pipelines to cold storage, and implement automated

Qualifications

  • 4–6 years of professional engineering experience, with at least 2 years in database-focused infrastructure work at moderate-to-high scale.
  • Strong expertise in MySQL and PostgreSQL—EXPLAIN plan analysis, index design, partitioning, sharding, and replication setup—on 100M+ row tables.
  • Experience building data archiving pipelines to cold storage (S3/Parquet/Athena/Glacier) while keeping records queryable.
  • Proficient with AWS managed database services—RDS and Aurora—including provisioning, read replicas, parameter tuning, and monitoring; Aurora Global Database is a plus.
  • Hands-on with Redis/ElastiCache—cache invalidation patterns, TTL hygiene, hit-rate monitoring, and incident diagnosis.
  • Automation of archival, maintenance, and diagnostic work using Python and/or Bash.
  • Familiar with monitoring and observability tooling (Prometheus, Grafana, Datadog) for slow-query alerting and dashboards.
  • Knowledge of multi-AZ and cross-region DR architectures, backup strategies, and failover triggers (RTO/RPO < 5 minutes).
  • Ability to write Terraform or CloudFormation to provision data resources in a version-controlled, reproducible way.
  • Ability to write and debug stored procedures and dynamic partitioning logic for automation.
  • Evidence-driven approach to optimization, with before/after data validation.
  • Strong communication skills to collaborate with product squads and explain data constraints.

Responsibilities

  • Execute database optimization and scaling initiatives—partitioning, sharding, replication configurations, and slow-query fixes.
  • Own performance diagnostics on high-volume tables—EXPLAIN plans, index tuning, cache analysis, and replication lag.
  • Build and maintain data archiving pipelines to move historical data to cold storage while preserving query access.
  • Implement and maintain data lifecycle automation—TTL policies, archival jobs, partition rotation, and retention enforcement.
  • Conduct regular performance profiling across the data tier and implement measurable fixes.
  • Support load testing and capacity planning for the data layer to prevent production issues during growth.
  • Implement and maintain data-tier BC/DR infrastructure with multi-AZ deployments and cross-region replication.

Skills

MySQL
PostgreSQL
Archiving pipelines
AWS RDS/Aurora
Redis/ElastiCache
Terraform/CloudFormation
Monitoring/observability
Python/Bash
Cross-team collaboration
Data-tier reliability

Tools

Prometheus
Grafana
Datadog
Terraform
CloudFormation
Athena/Parquet/Glacier
S3

Job description

Who We Are

NEXT Ventures is a global fintech group powering FundedNext — one of the world's fastest-growing proprietary trading platforms — and FNmarkets, a regulated CFD brokerage. Across offices in Bangladesh, Malaysia, Sri Lanka, Cyprus, and Dubai, we build and operate the technology that lets traders access global markets at scale. Our Platform Engineering team owns the infrastructure, reliability, and observability backbone that every product squad depends on.

Your Role in Our Mission

As our Database Reliability Engineer, you are the specialist who ensures our data layer never becomes the bottleneck. You own database performance, archiving pipelines, caching reliability, and BC/DR for the data tier that underpins all FundedNext services. Working within the Platform Engineering squad, you execute the scaling and optimization initiatives designed by the squad lead — and you are the reason our databases stay fast, durable, and cost‑efficient as trading data and traffic grow.

Data Performance & Scalability
  • Execute database optimization and scaling initiatives — implement partitioning strategies, sharding schemes, and replication configurations; independently diagnose and fix slow queries across the platform.
  • Own performance diagnostics on high-volume tables — EXPLAIN plan analysis, index tuning, buffer cache hit‑ratio investigation, lock contention, and replication lag on 100M+ row tables.
  • Build and maintain data archiving pipelines — move historical records (trade logs, transaction data, audit trails) from production tables to cold storage (S3/Parquet/Athena/Glacier) while preserving query access for compliance and reporting.
  • Implement and maintain data lifecycle automation — TTL policies, scheduled archival jobs, partition rotation, and retention enforcement.
  • Conduct regular performance profiling across the data tier — identify bottlenecks in MySQL/PostgreSQL and the caching layer, then implement measurable fixes.
  • Support load testing and capacity planning for the data layer — analyze growth trends, identify breaking points, and implement fixes before traffic growth causes production issues.
Data Reliability & Infrastructure
  • Implement and maintain data-tier BC/DR infrastructure — multi-AZ deployments, read replicas, Aurora Global Database/cross-region replication, backup schedules, and failover triggers targeting RTO/RPO under 5 minutes; participate in regular DR drills.
  • Implement and maintain caching strategies (Redis, ElastiCache) — cache invalidation logic, read-through/write-behind patterns, TTL hygiene, and cache hit-rate monitoring.
  • Define and track reliability metrics (SLIs/SLOs) for the data layer — query latency, availability, replication health — and drive improvements against them.
  • Implement monitoring, alerting, and observability for the data tier — dashboards, slow-query alerting, and replication/cache health monitoring targeting fast MTTD on data incidents.
  • Maintain and improve database-level automation — stored procedures, dynamic partitioning scripts, and scheduled maintenance jobs.
Cross-Team Support
  • Collaborate with application and product squads to design schemas, indexes, and access patterns that scale.
  • Remediate data-related security findings from the Cyber Security Squad — implement fixes and verify through retesting.
  • Write and maintain infrastructure-as-code for data resources — Terraform or equivalent for reproducible, version-controlled provisioning.
  • Document runbooks, archiving procedures, and operational guidance so any engineer can respond to data-tier incidents with clear guidance.
What You Bring
  • 4–6 years of professional engineering experience, with at least 2 years in database-focused infrastructure work at moderate-to-high scale.
  • Strong expertise in MySQL and PostgreSQL—EXPLAIN plan analysis, index design, partitioning, sharding, and replication setup—with direct experience on 100M+ row tables.
  • Able to build data archiving pipelines to cold storage (S3/Parquet/Athena/Glacier) while keeping records queryable for compliance and reporting.
  • Proficient with AWS managed database services—RDS and Aurora—including provisioning, read replicas, parameter group tuning, and monitoring; Aurora Global Database experience is a plus.
  • Hands‑on with Redis and ElastiCache—cache invalidation patterns, TTL hygiene, hit‑rate monitoring, and incident diagnosis.
  • Automates archival, maintenance, and diagnostic work using Python and/or Bash.
  • Familiar with monitoring and observability tooling (Prometheus, Grafana, Datadog, or similar) for slow-query alerting, dashboards, and data-tier incident analysis.
  • Understands multi-AZ and cross-region DR architectures, backup strategies, and failover triggers targeting RTO/RPO under 5 minutes.
  • Comfortable writing Terraform or CloudFormation to provision data resources in a version-controlled, reproducible way.
  • Can write and debug stored procedures and dynamic partitioning logic for database-level automation.
  • Evidence-driven—profiles and measures before guessing; validates every optimization against before/after data.
  • Performance-oriented—treats data-tier latency as a first-class reliability problem, not a tuning afterthought.
  • Good communicator—works with product squads to shape schemas and access patterns, and explains data constraints clearly.
X-Factor: AI-Native Engineering
  • You actively use modern AI agentic workflows daily—not limited to Copilot autocomplete.
  • You are proficient with Claude Code, Cursor, Windsurf, or equivalent tools.
  • You are comfortable with project-level AI configuration (CLAUDE.md, rules files), agentic task delegation, and AI-driven code review.
  • You think in terms of 5–10x productivity through AI-augmented development—and you can demonstrate it.
Why Join NEXT
  • Work on data challenges at real scale—100M+ row tables, 1TB+ databases, high-frequency trading transaction volumes.
  • A team that treats database performance and data-tier reliability as first-class engineering problems, not DBAs on the side.
  • Flat structure—your work directly shapes the platform's data infrastructure, not filtered through layers of process.
  • Offices across Bangladesh, Malaysia, Sri Lanka, Cyprus, and Dubai—a genuinely global engineering team.
  • Competitive compensation benchmarked to your market, with room to grow as the team scales.

Experience level: Senior

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer
Senior Site Reliability Engineer

NEXT Ventures • Dubai

On-site
AED 240,000 - 360,000
Senior Data Reliability Engineer — Scale High-Volume DBs
Senior Data Reliability Engineer — Scale High-Volume DBs

NEXT Ventures • Dubai

On-site
AED 360,000 - 480,000
DevSecOps Engineer
DevSecOps Engineer

NEXT Ventures • Dubai

On-site
AED 310,000 - 460,000
Database Reliability Engineer
Database Reliability Engineer

Fuse Energy • United Arab Emirates

On-site
AED 120,000 - 150,000
Competitive salary and equity sign-on bonus
Biannual bonus scheme
Fully expensed tech
+2
Senior Data Engineer
Senior Data Engineer

Deriv.com • Dubai

On-site
AED 320,000 - 520,000
Fullstack Lead (Laravel & Node.js)
Fullstack Lead (Laravel & Node.js)

NEXT Ventures • Dubai

On-site
AED 160,000 - 230,000
Staff Database Engineer
Staff Database Engineer

Remotedxb • Dubai

On-site
AED 350,000 - 550,000
Sr. Backend Engineer
Sr. Backend Engineer

ClearGrid • Dubai

On-site
AED 69,000 - 97,000
Trading and Risk Advisor — CFD
Trading and Risk Advisor — CFD

NEXT Ventures • Dubai

On-site
AED 367,000 - 551,000
Senior Data Engineer
Senior Data Engineer

Newbridge • Abu Dhabi

On-site
AED 60,000 - 90,000
Continuous learning opportunities
Collaborative team environment
High autonomy and ownership