Ceph Cluster Development Engineer (C++ Focus)

Fortinet

Santa Clara (CA)

On-site

USD 179,000 - 219,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Medical, dental, vision insurance
401(k) plan
Vacation and sick time
Comprehensive leave program

Job summary

A cybersecurity firm based in Santa Clara is looking for a Ceph Cluster Development & Operations Engineer. You will design and maintain enterprise-scale Ceph storage clusters, work with system architects, and optimize performance across data centers. The ideal candidate has strong C++ skills and experience with Kubernetes. This position offers a competitive salary of $179,000-$219,000 and various benefits including insurance, paid holidays, and vacation time.

Qualifications

  • Strong proficiency in C++ (C++11 or later) for large-scale distributed systems.
  • Deep understanding of Ceph core components and architecture.
  • Experience with Python or Go for automation.
  • Hands-on experience with Kubernetes and related DevOps tools.

Responsibilities

  • Design, build, and operate large-scale Ceph clusters.
  • Integrate Ceph with Kubernetes and CI/CD pipelines.
  • Develop automation and tooling for cluster lifecycle management.
  • Collaborate with the upstream Ceph community for feature development.

Skills

C++ programming
Ceph architecture
Linux systems programming
Kubernetes
Performance profiling

Tools

Docker
Ansible

Job description

We are seeking a highly skilled Ceph Cluster Development & Operations Engineer with strong expertise in C++ systems programming to design, extend, and maintain enterprise-scale Ceph distributed storage clusters. The role involves deep development in Ceph core subsystems (RADOS, OSD, RGW, MDS), performance optimization, and operational excellence across multi-site, multi-zone architectures.

You will work closely with system architects, SREs, and cloud infrastructure teams to ensure the reliability, scalability, and security of mission-critical storage systems deployed across multiple data centers and Kubernetes environments.

Key Responsibilities
  • Design, build, and operate large-scale Ceph clusters including RADOS, RGW, RBD
  • Contribute to or extend Ceph core components written in C++ (e.g., OSD, RGW, librados, BlueStore, MGR modules).
  • Profile and optimize performance across network, disk I/O, and replication layers (PG placement, CRUSH rules, BlueStore tuning).
  • Develop automation and tooling for cluster lifecycle management (deployment, upgrades, scaling, failover, and recovery).
  • Integrate Ceph with Kubernetes (via Rook-Ceph, CSI drivers) and CI/CD pipelines for continuous delivery.
  • Implement and validate multi-site replication and disaster recovery architectures for high availability.
  • Develop and maintain secure storage solutions using dm-crypt, KMS integration, and CephX authentication.
  • Build observability pipelines using Prometheus, Grafana, and custom exporters for metrics and health analytics.
  • Write and maintain SOPs, automation scripts, and system documentation to support production-grade operations.
  • Collaborate with upstream Ceph community or maintain in-house forks for feature development and bug fixes.
Qualifications
Required Skills
  • Strong proficiency in C++ (C++11 or later), with experience in large-scale distributed systems or kernel‑adjacent development.
  • Deep understanding of Ceph architecture and its core components: MON, OSD, MGR, RGW, MDS, and CRUSH maps.
  • Proficient in Linux systems programming, debugging (gdb, perf, valgrind), and performance profiling.
  • Experience with Python or Go for tooling and automation.
  • Strong foundation in data replication, erasure coding, and consistency models in distributed storage.
  • Hands‑on experience with Kubernetes, Rook‑Ceph, Helm, Ansible, and related DevOps tools.
  • Familiarity with TCP/IP, HTTP/S3 APIs, block storage (RBD/iSCSI), and object storage semantics.
  • Ability to conduct root‑cause analysis and lead performance investigations under production environments.
Preferred Skills
  • Contributions to the Ceph open‑source project or prior experience modifying Ceph source code.
  • Experience with multi‑site replication, object versioning, compliance retention, or legal hold features.
  • Background in distributed storage systems, file systems, or cloud storage platforms.
  • Familiarity with containerized environments, network virtualization, and cloud‑native observability stacks.
  • Excellent technical documentation and communication skills in English.

The US base salary range for this full‑time position is $179,000-$219,000. Fortinet offers employees a variety of benefits, including medical, dental, vision, life and disability insurance, 401(k), 11 paid holidays, vacation time, and sick time, as well as a comprehensive leave program.

Wage ranges are based on various factors, including the labour market, job type, and job level. Exact salary offers will be determined by factors such as the candidate's subject knowledge, skill level, qualifications, experience, and geographic location.

All roles are eligible to participate in the Fortinet equity program. Bonus eligibility is reviewed at the time of hire and annually at the Company’s discretion.

Why Join Us:

We encourage candidates from all backgrounds and identities to apply. We offer a supportive work environment and a competitive Total Rewards package to support you with your overall health and financial well‑being.

Embark on a challenging, enjoyable, and rewarding career journey with Fortinet. Join us in bringing solutions that make a meaningful and lasting impact to our 660,000+ customers around the globe.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Ceph Cluster Development Engineer (C++ Focus)
Ceph Cluster Development Engineer (C++ Focus)

Zoomcar • Santa Clara (CA)

On-site
USD 179,000 - 219,000
Medical, dental, vision insurance
401(k)
Paid holidays and vacation time
Ceph Cluster Engineer - C++ & Distributed Storage Expert
Ceph Cluster Engineer - C++ & Distributed Storage Expert

Fortinet • Santa Clara (CA)

On-site
USD 179,000 - 219,000
Medical, dental, vision insurance
401(k) plan
Vacation and sick time
+1
Site Reliability Engineer
Site Reliability Engineer

Fortinet • Sunnyvale (CA)

On-site
USD 170,000 - 200,000
Principal Software Development Engineer
Principal Software Development Engineer

Fortinet • Sunnyvale (CA)

On-site
USD 175,000 - 245,000
Medical Insurance
Dental Insurance
Vision Insurance
+5
Staff Software Development Engineer
Staff Software Development Engineer

Zoomcar • Sunnyvale (CA)

On-site
USD 150,000 - 215,000
Medical, dental, and vision insurance
401(k)
Paid holidays and vacation time
Software Development Engineer
Software Development Engineer

Zoomcar • Santa Clara (CA)

On-site
USD 123,000 - 151,000
Medical, dental, and vision insurance
401(k)
Paid holidays
+4
Software Development Engineer
Software Development Engineer

Zoomcar • Sunnyvale (CA)

On-site
USD 100,000 - 120,000
Medical insurance
Dental insurance
Vision insurance
+6
Cloud-Native DevOps Engineer - AWS, Kubernetes, CI/CD
Cloud-Native DevOps Engineer - AWS, Kubernetes, CI/CD

Fortinet • Santa Clara (CA)

On-site
DevOps Engineer
DevOps Engineer

Zoomcar • Santa Clara (CA)

On-site
USD 130,000 - 180,000
Medical, dental, and vision insurance
401(k) plan
Paid holidays and vacation time
Principal Software Development Engineer
Principal Software Development Engineer

Zoomcar • Sunnyvale (CA)

On-site
USD 200,000 - 261,000
Medical, dental, vision insurance
401(k) plan
Paid holidays and vacation
+1