Principal Software Engineer, Enterprise Scalability

United States Digital Space LLC

Boston (MA)

On-site

USD 244,000 - 366,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Klaviyos in the United States seeks a senior individual contributor to own scale, reporting to the VP of Engineering. You will drive performance, reliability, multi-region readiness, and large-tenant readiness, coordinating across teams to productionize improvements and influence through technical depth.

You will lead scalability reviews, implement AI-driven enhancements, and align technical work with business outcomes, ensuring clean cross‑org collaboration and measurable impact on enterprise

Qualifications

  • Define enterprise scalability fitness functions and align teams to SLOs and budgets.
  • Design sharding, partitioning, caching/back-pressure, multi-region readiness, and high-volume migrations.
  • Build benchmarks, profiling harnesses, reproducible testbeds; pair with teams to land fixes.
  • Lead scalability reviews and readiness gates; drive incident deep dives tied to systemic fixes.
  • Communicate clearly to execs and engineers, tying technical work to business impact and customer outcomes.
  • Integrate AI into scale and resiliency work—from proactive anomaly detection to synthetic load and guided runbooks—so performance improvements stick and incidents don’t repeat.

Responsibilities

  • Scale multi-tenant SaaS while removing bottlenecks and proving impact with data.
  • Drive platform-wide architectural changes and optimize systems for performance and reliability.
  • Lead scalability reviews and readiness gates to accelerate delivery and perform deep-dive incident analysis.
  • Collaborate with cross-functional teams to translate tradeoffs into executive-level recommendations.
  • Implement AI-driven enhancements in monitoring, load testing, and runbooks to improve resilience.

Skills

Performance engineering
Capacity planning
Sharding / partitioning
Caching / back-pressure
Multi-region readiness
High-volume migrations
AI tools & automation
Cross-org influence
Executive communication
AI fluency

Job description

*At the company, we value the unique backgrounds, experiences and perspectives each the company (we call ourselves Klaviyos) brings to our workplace each and every day. We believe everyone deserves a fair shot at success and appreciate the experiences each person brings beyond the traditional job requirements. If you’re a close but not exact match with the description, we hope you’ll still consider applying. Want to learn more about life at the company? Visit the company.com/careersto see how we empower creators to own their own destiny.*

Be the company’s senior IC for scale, you will report into a VP of Engineering and lead performance, reliability, multi‑region, and large‑tenant readiness. You’ll drive platform-wide architectural change, hunt bottlenecks and optimize systems, and partner across teams to productionize improvements. Given that this is an IC role with no direct reports; you will lead via technical depth, hands‑on impact, and crisp cross‑org alignment.

What You’ll Do
  • Define enterprise scalability fitness functions (latency/throughput/error rates) and a scorecard; align teams to SLOs and budgets.
  • Design/implement sharding and partitioning strategies, caching/back‑pressure, multi‑region readiness, and high‑volume migration paths.
  • Build lightweight enablement: benchmarks, profiling harnesses, reproducible testbeds; pair with teams to land fixes.
  • Lead scalability reviews and readiness gates that accelerate—not block—delivery; drive incident deep dives tied to systemic fixes.
  • Communicate clearly to execs and engineers, tying technical work to business impact and customer outcomes.
  • Integrate AI into scale and resiliency work—from proactive anomaly detection to synthetic load and guided runbooks—so performance improvements stick and incidents don’t repeat.
Who You Are
  • Experience: 12+ years scaling multi‑tenant SaaS with a reputation for removing major bottlenecks and proving impact with data.
  • Technical expertise: Performance engineering, capacity planning, sharding/partitioning, caching/back‑pressure, multi‑region readiness, and high‑volume migrations; you turn hotspots into robust patterns.
  • AI tools & automation: You apply AI to scale work—profiling assistance, workload modeling, synthetic traffic generation, anomaly detection, and runbook copilots—always with explicit guardrails and observability.
  • Cross‑org influence: You align teams through fitness functions, scorecards, and readiness gates that accelerate—not block—delivery; you communicate tradeoffs crisply to execs and engineers.
  • AI fluency: Curious, adaptable, and proactive in exploring AI that responsibly improves scale outcomes.
Nice to Haves
  • Scale scorecard: Company‑wide fitness functions (latency/throughput/error rates) are adopted and reviewed regularly.
  • High‑impact wins: 2–3 bottlenecks removed with documented, reproducible testbeds; pXX latencies and error rates improve on top enterprise workloads; repeat P0s trend down.
  • AI‑assisted scale engineering: AI‑driven anomaly detection reduces alert noise while improving signal; generative load testing and copilot runbooks are used in release/readiness checks for the top critical services; time‑to‑isolate regressions drops 20–30%.
Success in 6–12 Months
  • Company‑wide scale scorecard in place; 2–3 high‑impact bottlenecks removed; top enterprise workloads show improved pXX latencies and error rates; fewer repeat P0s.

We use Covey as part of our hiring and / or promotional process. For jobs or candidates in NYC, certain features may qualify it as an AEDT. As part of the evaluation process we provide Covey with job requirements and candidate submitted applications. We began using Covey Scout for Inbound on April 3, 2025.

Please see the independent bias audit report covering our use of Covey here

*Massachusetts Applicants:*It is unlawful in Massachusetts to require or administer a lie detector test as a condition of employment or continued employment. An employer who violates this law shall be subject to criminal penalties and civil liability.

Our salary range reflects the cost of labor across various U.S. geographic markets. The range displayed below reflects the minimum and maximum target salaries for the position across all our US locations. The base salary offered for this position is determined by several factors, including the applicant’s job‑related skills, relevant experience, education or training, and work location.

In addition to base salary, our total compensation package may include participation in the company’s annual cash bonus plan, variable compensation (OTE) for sales and customer success roles, equity, sign‑on payments, and a comprehensive range of health, welfare, and wellbeing benefits based on eligibility.

Your recruiter can provide more details about the specific salary/OTE range for your preferred location during the hiring process.

Base Pay Range For US Locations:

$244,000—$366,000 USD

*This role may require up to 10% travel for purposes such as new hire onboarding, client or partner work if applicable, team meetings, and industry events. Travel is coordinated in advance.

Get to Know the company

We’re the company (pronounced clay-vee-oh). We empower creators to own their destiny by making first‑party data accessible and actionable like never before. We see limitless potential for the technology we’re developing to nurture personalized experiences in ecommerce and beyond. To reach our goals, we need our own crew of remarkable creators—ambitious and collaborative teammates who stay focused on our north star: delighting our customers. If you’re ready to do the best work of your career, where you’ll be welcomed as your whole self from day one and supported with generous benefits, we hope you’ll join us.

*AI fluency at the company includes responsible use of AI (including privacy, security, bias awareness, and human‑in‑the‑loop). We provide accommodations as needed.

*By participating in the company’s interview process, you acknowledge that you have read, understood, and will adhere to ourGuidelines for using AI in the the company interview Process. For more information about how we process your personal data, see ourJob Applicant Privacy Notice.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Principal Engineer, Core Infrastructure
Principal Engineer, Core Infrastructure

United States Digital Space LLC • Boston (MA)

On-site
USD 244,000 - 366,000
Principal Engineer, Data Infrastructure
Principal Engineer, Data Infrastructure

United States Digital Space LLC • Boston (MA)

On-site
USD 244,000 - 366,000
Senior Lead Program Manager, Technical Programs
Senior Lead Program Manager, Technical Programs

United States Digital Space LLC • Boston (MA)

On-site
USD 192,000 - 288,000
Senior Software Engineer - Profiles, Lists and Segments
Senior Software Engineer - Profiles, Lists and Segments

United States Digital Space LLC • Boston (MA)

On-site
USD 148,000 - 222,000
Senior Software Engineer, Platform Engineering
Senior Software Engineer, Platform Engineering

United States Digital Space LLC • Boston (MA)

On-site
USD 148,000 - 222,000
Sr. AI Systems & Solutions Architect, Customer Success
Sr. AI Systems & Solutions Architect, Customer Success

United States Digital Space LLC • Denver (CO)

On-site
USD 112,000 - 168,000
Sr. AI Systems & Solutions Architect, Customer Success
Sr. AI Systems & Solutions Architect, Customer Success

United States Digital Space LLC • Boston (MA)

On-site
USD 112,000 - 168,000
Sr. Lead Engineer, Applied AI (Customer Agent)
Sr. Lead Engineer, Applied AI (Customer Agent)

United States Digital Space LLC • Boston (MA)

On-site
USD 216,000 - 324,000
Sr. AI Systems & Solutions Architect, Customer Success
Sr. AI Systems & Solutions Architect, Customer Success

United States Digital Space LLC • San Francisco (CA)

On-site
USD 112,000 - 168,000
Principal Software Engineer, Enterprise Scalability Boston, MA
Principal Software Engineer, Enterprise Scalability Boston, MA

Klaviyo Inc. • Boston (MA)

On-site
USD 248,000 - 372,000
Generous benefits
Collaborative work environment
Opportunities for professional growth