Cluster & Systems Capacity Engineer

Backblaze

United States

On-site

USD 123,000 - 175,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Healthcare for family, including dental and vision
Competitive compensation and 401K
RSU grants for full-time employees
Flexible vacation policy
Maternity & paternity leave

Job summary

Backblaze is searching for a Capacity Planner to ensure that our cloud storage infrastructure scales efficiently. In this high-impact Cloud Operations role, you'll develop forecasts, monitor system performance, and collaborate with cross-functional teams to optimize capacity and costs.

The ideal candidate has extensive experience in site reliability engineering and cloud storage systems. We offer a competitive salary and comprehensive benefits including healthcare, flexible vacation, and a supportive work culture focused on work-life balance.

Qualifications

  • 3-6+ years of experience in a Cloud Operations role.
  • Familiarity with Cloud Storage infrastructure and distributed systems.
  • Proficiency in data analysis tools and capacity modeling.

Responsibilities

  • Develop forecasts for capacity demand and hardware deployment.
  • Monitor cluster and system-level performance and utilization.
  • Align capacity plans with capital budgets and financial outcomes.

Skills

Site Reliability Engineering
Capacity Planning
Data Analysis
Communication skills

Education

Bachelor’s degree in Computer Science or related field

Tools

Snowflake
SQL
Grafana
Prometheus

Job description

About Backblaze

Backblaze is the object storage leader in the open cloud movement, fueling customer success with cloud storage built purposefully to unlock budgets, unburden administrators, and unleash innovators. Together with our partners, we’re helping customers break free from the restrictive, overpriced legacy solutions that hold them back, and blaze forward with the full power of the open cloud in their hands.

Founded in 2007, Backblaze scaled the business with less than $3 million in outside funding until 2021, when we did a traditional IPO on the Nasdaq stock exchange. Today, Backblaze generates over $100m in revenue and is the leading specialized storage cloud – managing over three billion gigabytes of data storage for 500K+ customers in 175+ countries, including businesses, developers, IT professionals, and individuals.

About the Role

This role ensures that Backblaze’s storage clusters, compute systems, and network infrastructure scale reliably, cost‑efficiently, and ahead of demand. You will build and maintain predictive models, ensure consistent supply and demand alignment, and partner cross‑functionally to inform strategic investment and deployment decisions. This is a high‑impact role within Cloud Operations, directly contributing to service availability, durability, performance, margin optimization, and long‑term platform scalability.

Key Responsibilities
Capacity Planning & Forecasting
  • Develop and maintain short, medium, and long‑term capacity demand and hardware deployment forecasts across storage, compute, and network domains within the platform.
  • Build predictive models that translate business demand signals into infrastructure requirements using historical utilization, growth trends, product sales plans, hardware lifecycle roadmaps, and other key business inputs.
  • Partner with Infrastructure, Production, and Network Engineering teams to align capacity plans with system design and scaling initiatives.
  • Develop and automate forecasting pipelines, simulation calculators and tools, and capacity dashboards to improve data quality, reduce manual analysis, and provide stakeholders clear visibility into platform usage and cluster health metrics.
Cluster Performance & Resource Optimization
  • Monitor and analyze cluster and system‑level utilization and performance across CPU, memory, IOPS, and network resources.
  • Adjust deployment plans and recommended configurations in real‑time to maintain adequate headroom and system stability in support of delivering a world‑class customer experience.
  • Partner with service and platform owners to develop headroom and live buffer policies, optimize hardware BoMs, leverage virtualized orchestration, and reduce product cost.
Cross‑functional Organizational Alignment
  • Work in lockstep with Operations and Finance peers to align capacity plans and hardware requirements with capital budgets, cost targets, and financial outcomes.
  • Support strategic optimization initiatives across infrastructure investments, engineering development, and operations processes, contributing to long‑term infrastructure strategy and capital planning.
  • Lead efforts to evaluate, procure, and provision requests for new or additional hardware, working with Systems and Network Engineering, SRE, NOC, and Data Center Operations teams to identify and deliver optimal solutions.
  • Maintain alignment with Product and Sales to support customer onboarding, growth, and demand variability.
  • Communicate complex capacity and infrastructure insights clearly to technical and non‑technical stakeholders.
Required Qualifications
  • Bachelor’s degree in Computer Science, Engineering, Mathematics, Data Science, Information Systems, Statistics, or a related technical field (or equivalent experience).
  • 3-6+ years of experience in Site Reliability Engineering, Infrastructure Capacity Planning, Systems/Infrastructure Engineering, Production Engineering, Data Center Operations or similar Cloud Operations role.
  • Familiarity and experience working with Cloud Storage infrastructure, particularly highly‑available, large‑scale distributed systems supporting large amounts of data with high throughput and complex performance requirements.
  • Background in capacity modeling, performance analysis, scenario modeling, and/or infrastructure cost optimization, with an ability to quantify ideas within financial frameworks and forecasts.
  • Proficiency in database and data analysis tools (preferably Snowflake, Metabase, Grafana, Python, SQL, Prometheus, Victoria Metrics, and Excel/Google Sheets).
  • Demonstrated deep, creative, and logical thinking complemented by a strong data analysis skillset.
  • Excellent communication and documentation skills, with the ability to share knowledge and explain concepts accurately and concisely.
  • Desire to work on a highly‑autonomous team that cares deeply about quality, cost, and the customer experience.
Backblaze Perks
  • Healthcare for family, including dental and vision
  • Competitive compensation and 401K
  • RSU grants for full‑time employees
  • ESPP program
  • Flexible vacation policy
  • Maternity & paternity leave
  • MacBook Pro to use for work, plus a generous stipend to personalize your workstation
  • Childcare bonus (human children only)
  • Fertility treatment and support
  • Learning & development program
  • Commuter benefits
  • Culture that supports a healthy work‑life balance

The expected salary range for this role is $123,000 – $175,000.

At Backblaze, we value being fair and good to our customers, partners, and employees. That’s why diversity, equity, and inclusion are at the core of our values. We are committed to fostering a workforce where all employees feel a sense of belonging regardless of race, ethnicity, nationality, gender, sexual orientation, age, religion, socio‑economic status, ability, veteran status, and education. We believe that our dedication to cultivating a diverse workspace not only allows us to better serve our customers in over 175 countries but further reinforces our commitment to doing the right thing. We are proud to be an Equal Opportunity Employer.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Cluster & Systems Capacity Engineer
Cluster & Systems Capacity Engineer

Backblaze External Website • United States

On-site
USD 123,000 - 175,000
Healthcare for family, including dental and vision
Competitive compensation and 401K
RSU grants for full-time employees
+7
Data Center Operations Manager
Data Center Operations Manager

Webhosting • Chandler (AZ)

On-site
USD 110,000 - 140,000
Data Center Operations Manager
Data Center Operations Manager

Socket.dev • Chandler (AZ)

On-site
USD 110,000 - 140,000
Data Center Operations Manager
Data Center Operations Manager

Backblaze • Chandler (AZ)

On-site
USD 110,000 - 140,000
Data Center Operations Manager
Data Center Operations Manager

Backblaze External Website • Chandler (AZ)

On-site
USD 110,000 - 140,000
Strategic Operations Engineer III
Strategic Operations Engineer III

Backblaze • United States

On-site
USD 123,000 - 175,000
Principal Product Manager
Principal Product Manager

Backblaze • United States

On-site
USD 180,000 - 260,000
RSU grants for full‑time employees.
Annual company bonus plan.
Healthcare for family, includingDental
+10
Senior Software Engineer
Senior Software Engineer

Webhosting • San Mateo (CA)

Hybrid
USD 180,000 - 240,000
ESPP program
Sr. Sales Operations Manager - GTM
Sr. Sales Operations Manager - GTM

Backblaze • United States

On-site
USD 160,000 - 185,000
Healthcare for family, including dental and vision
401K
RSU grants for full-time employees
+3
Cloud Capacity Architect & Optimization Engineer
Cloud Capacity Architect & Optimization Engineer

Backblaze External Website • United States

On-site
USD 123,000 - 175,000
Healthcare for family, including dental and vision
Competitive compensation and 401K
RSU grants for full-time employees
+7