Senior Software Engineer, Core Reliability

United States Digital Space LLC

United States

Hybrid

USD 150,000 - 210,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

the company, a remote-first but not remote-only organization, seeks a Senior Software Engineer, Core Reliability in the Infra Reliability team within Platform. You will help scale the company by improving reliability, security, and deployment safety across the production environment.

You will lead high-impact reliability projects that make the entire service environment more resilient, working with multiple teams to reduce incidents and support thousands of services and daily deployments.

Qualifications

  • 5+ years of software engineering designing, building, and maintaining production services.
  • Experience with Ruby, Go, Terraform, and cloud platforms (AWS, GCP or Azure).
  • Experience designing reliable, high-throughput, low-latency distributed systems at scale.
  • Proficient with observability and monitoring tools (Kibana, Datadog).
  • Ability to communicate architecture decisions to cross-functional teams.
  • Willingness to participate in on-call rotations and respond outside business hours.
  • Uses generative AI responsibly with human oversight.

Responsibilities

  • Own the design and delivery of reliability projects and features across the service environment.
  • Partner with critical T0/T1 services to improve scalability and reduce toil.
  • Build and enhance systems that securely manage service configurations and secrets.
  • Improve canary-based release systems and deployment capabilities for thousands of services.
  • Drive reliability best practices and strengthen reliability culture across engineering teams.

Skills

Ruby
Go
Distributed systems design
Observability & monitoring
On-call readiness
AI in workflows
Communication with stakeholders

Tools

AWS
GCP
Azure
Kibana
Datadog
Terraform

Job description

Ready to do the most impactful work of your career? At the company, we are uncompromising on our mission to increase economic freedom. The bar is high, the environment is intense, and we like it that way. This isn't a place for complacency, it’s a place to be pushed past your perceived limits. If you're ready to build the future of finance alongside people who refuse to settle for "good enough," you belong here. the company is a remote-first, but not remote-only company. Expect to get together quarterly for intense in-person working sessions called “surges.” Learn more about working at the company.

As a Senior Software Engineer, Core Reliability on the Infra Reliability team within Platform, you'll help the company scale 50x by improving reliability, security, and deployment safety across our production environment. This team owns the systems that secure service configurations and secrets, reduce customer-facing incidents, and strengthen deployment infrastructure supporting thousands of services and hundreds of daily releases. You'll lead high-impact reliability projects that make our entire service environment more resilient and safer for customers.

What you'll do:

  • Own the design and delivery of reliability projects and features that improve resiliency across the company's service environment in partnership with other engineering teams.
  • Partner with critical T0/T1 services to understand architecture, improve scalability, and reduce operational toil.
  • Build and enhance systems that securely manage service configurations and secrets at scale.
  • Improve canary-based release systems and expand deployment capabilities to support thousands of services and hundreds of daily deployments with fewer incidents.
  • Drive reliability best practices and strengthen reliability culture across engineering teams at the company.

Required Skills and Experience:

  • 5+ years of software engineering experience designing, building, and maintaining production services in service-oriented architectures, including experience with Ruby, Go, Terraform, and cloud platforms (AWS, GCP or Azure).
  • Demonstrated ability to design and operate reliable, high-throughput, low-latency distributed systems at scale, with a track record of writing well-tested, production-quality code.
  • Proven experience with observability and monitoring tools (e.g., Kibana, Datadog) to debug complex production issues, tune system performance, and reduce incident frequency.
  • Experience writing and verbally communicating architecture decisions to cross-functional engineering stakeholders.
  • Ability to participate in on-call rotations and respond to issues outside normal business hours.
  • Utilizes generative AI responsibly, maintaining human oversight to deliver business-ready outputs and drive measurable improvements in workflow efficiency, cost, and quality.

Pay Transparency Notice: The target annual base salary for this position can range as detailed below. Total compensation may also include equity and bonus eligibility and benefits (including medical, dental, and vision).

Annual base salary range (excluding equity and bonus):

$191,100—$191,100 CAD

  • Application Limit: Candidates may submit a maximum of 3 applications within a 6-month period.
  • Equal Opportunity Employer: the company is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, protected veteran status, or genetic information. Applicants with criminal histories will be considered consistent with applicable federal, state, and local laws.
  • US Applicants: View Employee Rights, Know Your Rights, and E-Verify Notice of Participation.
  • Accommodations: If you are an individual with a disability who needs a reasonable accommodation, email us your request and contact info at accommodations[at]the company.com. Need screen reading technology? Click here to download a free compatible screen reader and view the tutorial.
  • Data Privacy & Arbitration: By submitting your application, you agree to our Candidate Privacy Notice. US applicants: By submitting your application, you agree Arbitration of Disputes.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Software Engineer, Core Reliability
Senior Software Engineer, Core Reliability

Omaze • United States

Hybrid
USD 122,735 - 150,010
Manager, Software Engineering (Reliability Platform)
Manager, Software Engineering (Reliability Platform)

Affirm • Sioux Falls (SD)

On-site
USD 204,000 - 264,000
Health care coverage
Flexible Spending Wallets
Competitive vacation and holiday schedules
+1
Manager, Software Engineering (Reliability Platform)
Manager, Software Engineering (Reliability Platform)

Affirm • Los Angeles (CA)

On-site
USD 230,000 - 290,000
Health care coverage
Flexible Spending Wallets
Competitive vacation and holiday schedules
+1
Manager, Software Engineering (Reliability Platform)
Manager, Software Engineering (Reliability Platform)

Affirm • Houston (TX)

On-site
USD 204,000 - 264,000
Health care coverage
Flexible Spending Wallets
Competitive vacation and holiday schedules
+1
Senior Software Engineer, Data Layer
Senior Software Engineer, Data Layer

United States Digital Space LLC • United States

Hybrid
USD 186,000 - 219,000
Senior Software Engineer, Data Engineering Platform
Senior Software Engineer, Data Engineering Platform

United States Digital Space LLC • United States

Hybrid
USD 186,000 - 219,000
Senior Site Reliability Engineer II
Senior Site Reliability Engineer II

Juniper Square • United States

On-site
USD 165,000 - 195,000
Health, dental, and vision care
Life insurance
Mental wellness coverage
+3
Senior Software Engineer (Platform - Access & Authorization)
Senior Software Engineer (Platform - Access & Authorization)

Talanto • Northern (KY)

Hybrid
USD 186,000 - 219,000
Quarterly in-person surges
Remote-first company
Staff Software Engineer (Platform - Access & Authorization)
Staff Software Engineer (Platform - Access & Authorization)

United States Digital Space LLC • United States

Hybrid
USD 218,000 - 257,000
Senior Engineering Manager, Core Automation (Platform)
Senior Engineering Manager, Core Automation (Platform)

United States Digital Space LLC • United States

Hybrid
USD 254,000 - 299,000