Senior SRE: Core Platform & Embedded Reliability

CrowdStrike

Sunnyvale (CA)

Hybrid

USD 140,000 - 215,000

Full time

12 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Equity awards
Wellness programs
Vacation policy
Parental leaves
Professional development
Employee networks
Office culture
Great Place to Work

Job summary

CrowdStrike is seeking a Principal Site Reliability Engineer to lead reliability, scalability, and architectural decisions across CrowdStrike's Falcon Platform. You will collaborate with product engineering teams to deliver robust backend services, libraries, and tooling, shaping reliability outcomes at scale.

You will mentor engineers, drive SLO/SLI practices, and influence architectural standards with hands-on ownership of critical infrastructure across multiple cloud environments.

Qualifications

  • 10+ years building and operating distributed systems at scale.
  • 5+ years developing microservices for a SaaS product in a modern backend language.
  • Expert-level proficiency in at least one programming language (Go) or willingness to reach expert level.
  • Deep understanding of distributed systems, consensus algorithms, replication, consistency models, and scalability patterns.
  • Proven experience scaling backend systems: sharding, partitioning, capacity planning, and performance optimization.
  • Deep understanding of multi-threading, concurrency, and parallel processing.
  • Track record of impactful architectural decisions at organizational scope.
  • Strong systems thinking and ability to influence across organizational boundaries.
  • Engineering best practices: testing paradigms, code review, and resilient architecture.
  • Thrive in a fast-paced, test-driven, collaborative environment; team player.
  • Desire to ship code and see it in production.
  • Degree in Computer Science, or commensurate experience.
  • Proven experience utilizing AI technologies to enhance decision-making, streamline workflows, and drive outcomes.

Responsibilities

  • Partner with engineering leadership to define and drive multi-year reliability roadmaps.
  • Design and implement architectural improvements to services, libraries, and platforms.
  • Develop and maintain services meeting reliability and scalability demands.
  • Extend and build libraries for cross-cutting concerns across CrowdStrike's cloud platform.
  • Lead initiatives around reliability, scalability, and cost efficiency.
  • Establish observability practices and drive automation through monitoring and CD.
  • Define and implement SLOs and error budgets for real decision-making.
  • Lead performance and cost optimization: profiling and capacity planning.
  • Conduct resilience engineering: chaos experiments and failure modeling.
  • Automate infrastructure-as-code to improve reliability and reduce toil.
  • Provide technical leadership during incidents and retrospectives.
  • Extract patterns into shared libraries and tools; collaborate with platform teams.
  • Continuously re-evaluate architectures to improve performance and stability.
  • Drive strategic technical decisions across the organization.
  • Mentor engineers and raise technical IQ across the team.
  • Contribute to open source and evangelize best practices.
  • Collaborate across teams to deliver scalable backend infrastructure.

Skills

Distributed systems
Microservices
Go programming
System architecture
Concurrency
SLO/SLI
Cloud fundamentals
AI technologies
Leadership
CS degree

Education

Bachelor's degree in Computer Science

Tools

Kubernetes
AWS
Cassandra
Kafka
Elasticsearch/OpenSearch
GCP
OCI

Job description

CrowdStrike is seeking a Principal Site Reliability Engineer to lead reliability, scalability, and architectural decisions across CrowdStrike's Falcon Platform. You will collaborate with product engineering teams to deliver robust backend services, libraries, and tooling, shaping reliability outcomes at scale.

You will mentor engineers, drive SLO/SLI practices, and influence architectural standards with hands-on ownership of critical infrastructure across multiple cloud environments.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior SRE: Core Platform & Embedded Reliability
Senior SRE: Core Platform & Embedded Reliability

Socket.dev • New York (NY)

Hybrid
USD 140,000 - 215,000
Market-leading compensation
Equity awards
Professional development opportunities
Senior SRE: Core Platform & Embedded Reliability
Senior SRE: Core Platform & Embedded Reliability

CrowdStrike • Austin (TX)

Hybrid
USD 140,000 - 215,000
Market-leading compensation
Equity awards
Paid time off
+1
Senior SRE: Core Platform & Embedded Reliability
Senior SRE: Core Platform & Embedded Reliability

CrowdStrike • Redmond (WA)

Hybrid
USD 140,000 - 215,000
Market-leading compensation
Equity awards
Comprehensive wellness programs
+2
Senior SRE: Core Platform & Embedded Reliability
Senior SRE: Core Platform & Embedded Reliability

CrowdStrike • New York (NY)

Hybrid
USD 140,000 - 215,000
Market-leading compensation
Equity awards
Wellness programs
+3
Senior SRE - Core Platform & Embedded Reliability (Hybrid)
Senior SRE - Core Platform & Embedded Reliability (Hybrid)

CrowdStrike Holdings, Inc. • Sunnyvale (TX)

Hybrid
USD 140,000 - 215,000
Competitive compensation
Equity awards
Health insurance
+1
Senior SRE: Core Platform & Embedded Reliability
Senior SRE: Core Platform & Embedded Reliability

CrowdStrike Holdings, Inc. • New York (NY)

On-site
USD 140,000 - 215,000
Market-leading compensation
Equity opportunities
Hybrid work model
+1
Sr. Site Reliability Engineer - Core Platform & Embedded Reliability (Hybrid)
Sr. Site Reliability Engineer - Core Platform & Embedded Reliability (Hybrid)

Socket.dev • New York (NY)

Hybrid
USD 140,000 - 215,000
Market-leading compensation
Equity awards
Professional development opportunities
SRE Engineering Manager: Scale Reliability & Cloud Ops
SRE Engineering Manager: Scale Reliability & Cloud Ops

CrowdStrike • Sunnyvale (CA)

Hybrid
USD 140,000 - 215,000
Market leader in compensation
Comprehensive wellness programs
Competitive vacation
+1
Resident Platform Services Lead - Falcon Security
Resident Platform Services Lead - Falcon Security

CrowdStrike • California (MO)

On-site
USD 140,000 - 195,000
Market compensation
Wellness programs
Paid time off
+3
Sr Engineer, SRE TechOps CICD (Remote)
Sr Engineer, SRE TechOps CICD (Remote)

CrowdStrike Holdings, Inc. • Northern (KY)

Hybrid
USD 140,000 - 215,000
Competitive compensation
Wellness programs
Generous vacation and holidays
+1