Stateful Services Engineer

United States Digital Space LLC

Paris

Hybride

EUR 110 000 - 140 000

Plein temps

Il y a 5 jours
Soyez parmi les premiers à postuler
Générateur de candidature

Transformez ce poste en entretien — un CV et une lettre de motivation conçus selon ce que cet employeur recherche.

Passez les filtres ATS

Avantages offerts par ce poste

Lunch vouchers
Navigo transport subsidy
Hybrid onsite in Paris

Résumé du poste

United States Digital Space LLC is seeking a technically skilled Stateful Services Engineer to build and operate the infrastructure behind our AI‑powered customer platform. You will own the stateful services and collaborate with SRE, Infra and engineering teams to improve reliability and scale for growth.

You will manage PostgreSQL clusters, explore ClickHouse, Redis, RabbitMQ and Kafka/Redpanda, and drive safe maintenance, migrations and disaster recovery across on‑premise environments.

Qualifications

  • Strong hands‑on experience as a DBA/SRE with PostgreSQL in production.
  • Experience upgrading, migrations, replication, backups and performance troubleshooting.
  • High availability, failover, replication, DR for stateful systems.
  • Experience in on‑premise or bare‑metal environments.
  • Ability to design for scalability, reliability and performance.
  • Collaborates with software teams to build reliable database solutions.

Responsabilités

  • Operate and maintain stateful services including PostgreSQL, ClickHouse, Redis, RabbitMQ and Kafka/Redpanda.
  • Design and scale database infrastructure with 50+ clusters and complex workloads.
  • Implement safe maintenance, upgrades and migrations with minimal downtime.
  • Improve HA, replication, failover and backup/restore processes.
  • Contribute to monitoring, automation and on‑call incident management.
  • Collaborate with SRE, network and infra teams to ensure reliability.

Connaissances

PostgreSQL
ClickHouse
Redis
RabbitMQ
Kafka/Redpanda

Outils

Monitoring
On‑premises

Description du poste

Build the infrastructure behind the next generation of AI-powered customer communications

the company is building an AI-first communication platform where AI and human agents, business workflows and communication channels work together seamlessly.

As we enter our next phase of growth, we are looking for a Stateful Services Engineer to help build and operate the infrastructure foundation behind the platform.

This is a unique opportunity to work on an infrastructure environment combining carrier-grade voice systems, on-premise data centers, distributed systems and AI workloads, while helping us build the reliability and operational maturity required to support 3x growth.

You will report to the newly forming Stateful Services Lead and work closely with Engineering, SRE and Infrastructure teams to improve the reliability, scalability and evolution of our stateful systems.

Your mission

Your goal is to build, operate and improve the company’s stateful services, with a strong focus on databases and data stores.

You will help ensure that our databases and stateful systems remain highly available, reliable and performant, while evolving our infrastructure so that maintenance, upgrades and scaling can be performed safely and with minimal to no downtime.

Working closely with the Stateful Services Lead, you will take ownership of specific technical projects and contribute to the evolution of our stateful infrastructure.

You will:

  • Operate and maintain our stateful services, including PostgreSQL, ClickHouse, Redis, RabbitMQ and Kafka/Redpanda.
  • Contribute to the design, evolution and scaling of our database infrastructure, with 50+ PostgreSQL clusters and increasingly complex data workloads.
  • Work on projects aimed at making database maintenance, upgrades and migrations safer and less disruptive to our services.
  • Improve high availability, replication, failover, backup/restore and disaster recovery across our stateful systems.
  • Contribute to rebuilding our PostgreSQL clusters to support automatic failover between data centers and reduce the impact of maintenance operations.
  • Help improving our queueing infrastructure, including projects such as quorum-based RabbitMQ.
  • Help building and maintaining stable Redis infrastructure and improve the reliability of our stateful services.
  • Investigate and solve complex performance, scalability and reliability issues across databases and distributed systems.
  • Work closely with development teams to understand their needs, support their use of databases and help them build reliable solutions.
  • Collaborate with SRE, networking and infrastructure teams to ensure that our stateful services operate reliably within our broader infrastructure.
  • Contribute to monitoring, automation and operational practices that make our infrastructure easier and safer to operate.
  • Participate in our on-call and incident management rotation, helping diagnose and resolve production issues when needed.
  • Stay hands‑on and close to the technology: you will be expected to understand how our systems work, how they fail and how to improve them.
Who we are looking for

We’re looking for a technically strong and hands‑on Stateful Services Engineer who enjoys solving complex infrastructure problems and taking ownership of production systems where reliability really matters.

You have experience operating databases or stateful services in production and understand that running these systems is about much more than simply maintaining them. You know how to troubleshoot them, how they behave under load, how they fail, and how to make them more reliable as they scale.

You’ll work on specific technical projects defined with the Stateful Services Lead, while having the autonomy to investigate problems, propose solutions and contribute to technical decisions.

We care more about your technical depth, curiosity and ability to take ownership than the exact number of years of experience you have.

Must Have
  • Strong hands‑on experience as a DBA, Database Engineer, SRE or similar, with solid experience in PostgreSQL.
  • Experience operating and maintaining production databases, including upgrades, migrations, replication, backups, recovery, monitoring and performance troubleshooting.
  • Good understanding of high availability and reliability for stateful systems: failover, redundancy, replication, disaster recovery and minimising downtime during maintenance.
  • Good understanding of distributed systems, including replication, consistency, fault tolerance and the challenges of operating stateful services at scale.
  • Experience working with on‑premise infrastructure or a willingness and ability to quickly adapt to an on‑premise/bare‑metal environment.
  • Good system design skills and the ability to reason about scalability, reliability and performance.
  • Experience collaborating closely with software development teams, understanding their requirements and helping them build reliable solutions around databases and stateful services.
  • Strong troubleshooting and problem‑solving skills, with the ability to investigate complex production issues.
  • Strong hands‑on mindset: you enjoy getting into the technical details and solving problems yourself.
  • Comfortable participating in on‑call and incident management.
  • Fluent English, both written and spoken. French or other languages are a plus.
Nice to HaveExperience with some of the following:
  • ClickHouse or another complex database/data platform.
  • Kafka or Redpanda and event‑streaming infrastructure.
  • RabbitMQ, particularly quorum‑based configurations.
  • Redis and distributed caching.
  • Other database technologies such as Elasticsearch/OpenSearch, MySQL, Cassandra or MongoDB.
  • Advanced Linux systems administration, including processes, storage, networking and troubleshooting.
  • Bare‑metal infrastructure and data‑center environments.
  • Observability and monitoring of stateful systems.
  • Capacity planning, performance analysis and load testing.
  • Infrastructure automation and configuration management.
  • Incident management, root‑cause analysis and post‑mortem practices.
About You
  • You take ownership. When you see a reliability or scalability problem, you investigate it and work towards a solution.
  • You are genuinely technical. You like understanding how things work under the hood and are comfortable diving deep when something breaks.
  • You think in systems. You understand that databases don't operate in isolation and enjoy working across infrastructure, networking, SRE and development teams.
  • You care about reliability. You want systems that can be maintained, upgraded and scaled without putting the business at risk.
  • You are pragmatic. You can make sensible technical trade‑offs in complex environments.
  • You are curious and self‑driven. You learn by doing, investigate unfamiliar technologies and are comfortable figuring things out independently.
  • You enjoy solving complex problems. You like working on systems where there isn't always an obvious answer.
  • You communicate clearly. You can explain technical topics and work effectively with the teams relying on your services.
  • You are comfortable taking responsibility. You don't need to be a manager to have an impact or take the lead on a technical problem.
Why join us
  • Own critical infrastructure: your work will directly impact the reliability and scalability of the company’s platform.
  • Solve complex technical challenges: work with large‑scale databases and stateful services, including PostgreSQL, ClickHouse, Kafka, Redis and RabbitMQ.
  • Work on real infrastructure challenges: help improve PostgreSQL clusters, cross‐data‑center failover, RabbitMQ and Redis reliability.
  • Stay hands‑on: spend your time solving challenging technical problems rather than managing people.
  • Learn and grow: work across databases, distributed systems, messaging and infrastructure in a diverse technical environment.
  • Work across teams: collaborate closely with Engineering, SRE and Infrastructure to build reliable systems at scale.
Recruitment Process
  • Introductory call with Jade — 30 min

Get to know each other and discuss your background.

  • Interview with Sergey, CTO — 1 hour

Discuss your technical background, skills and knowledge.

  • On‑site System Design interview with Sergey, CTO — 1 hour

Work through a technical case study and explore your system design approach.

  • Team & cultural fit — 1 hour

Meet members of the team, onsite or via video.

Logistics

Paris, rue de la Paix, 75002 Hybrid model: 3 days onsite / 2 days remote Equipment: choose your preferred setup (Mac/Linux + monitors) Lunch vouchers Navigo: 50% of the monthly pass reimbursed

One More ThingWe are not necessarily looking for someone who has already worked in this exact role or with every technology in our stack.

The strongest candidates are often engineers who have built and operated complex infrastructure, taken ownership of production systems and learned new technologies along the way.

If you enjoy solving difficult technical problems, want to work close to the infrastructure and are excited by the challenge of making stateful systems more reliable and scalable — we would love to talk.

At the company, diversity and inclusion are in our DNA. All qualified applicants will receive equal consideration for employment without regard to color, language, religion, sex, sexual orientation, gender identity, national or social origin, opinion disability, age.

We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses and identifying potential inconsistencies or verification signals in application materials based on available information. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us.

Obtenez votre examen gratuit et confidentiel de votre CV.

ou faites glisser et déposez votre fichier ici.

Similar jobs

Postes similaires à comparer

Stateful Services Engineer
Stateful Services Engineer

Diabolocom • Paris

Hybride
EUR 90 000 - 130 000
Lunch vouchers
Navigo: 50% reimbursement
Hybrid model: 3 days onsite / 2 days w
+1
Head of Infrastructure
Head of Infrastructure

Diabolocom • Paris

Sur place
EUR 120 000 - 180 000
Lunch vouchers
Relocation support
Senior SRE Engineer - Security
Senior SRE Engineer - Security

United States Digital Space LLC • Paris

Sur place
EUR 84 000 - 106 000
Competitive salary & equity
5-week vacation
Paid sick leave
+5
Senior Data Engineer - Real time analytics
Senior Data Engineer - Real time analytics

United States Digital Space LLC • France

Sur place
EUR 65 000 - 90 000
Health coverage
Lunch provided
Commuting support
+2
Stateful Services Team Lead
Stateful Services Team Lead

Diabolocom • Paris

Sur place
EUR 90 000 - 130 000
Staff Engineer (Core & MLOps)
Staff Engineer (Core & MLOps)

Lever, Inc. • France

À distance
EUR 120 000 - 170 000
Fully remote
Flexible hours
Global collaboration
Senior Software Engineer - Distributed Systems (Applied AI)
Senior Software Engineer - Distributed Systems (Applied AI)

United States Digital Space LLC • Bordeaux

Hybride
EUR 100 000 - 140 000
Stock equity
ESPP
Professional development
+4
Staff Platform Engineer
Staff Platform Engineer

United States Digital Space LLC • Cologne

Sur place
EUR 90 000 - 130 000
Hybrid work environment
Virtual Shares
30 days annual leave
+1
Founding Analytics Engineer
Founding Analytics Engineer

United States Digital Space LLC • Paris

Hybride
EUR 85 000 - 94 000
Stock options
Healthcare plan
Office in Le Peletier, Paris
+3
Staff Rails Engineer - Engineering
Staff Rails Engineer - Engineering

Altertable • Paris

Sur place
EUR 70 000 - 90 000
30+ days of paid vacation
5,000€ budget for equipment
50% commute cost coverage
+4