Infrastructure Software Engineer

Worky

Campbell (CA)

On-site

USD 180,000 - 230,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Fully remote workplace
In-office options in the Bay Area
Flexible PTO
Comprehensive benefits

Job summary

Worky is seeking an Infrastructure Software Engineer to sit at the intersection of software development and site reliability. You will build backend services and platform infrastructure for a cloud-native stack, while also guiding reliability, observability, and operational health with the engineering team, ensuring systems are scalable, observable, and maintainable from day one.

You will own and evolve CI/CD pipelines, deployment tooling, and release processes to accelerate shipping, drive

Qualifications

  • 3+ years of software engineering with an infrastructure focus.
  • Proficient with Python 3 for backend services and APIs.
  • Hands-on experience with Kubernetes and Google Cloud Platform.
  • Experience with CI/CD pipelines.
  • Familiar with Prometheus and Grafana monitoring.
  • Collaborative approach to reliability.

Responsibilities

  • Design and build backend services, APIs, and platform infrastructure.
  • Contribute to product features end-to-end from backend to deployment.
  • Lead the observability stack and establish production health practices.
  • Own and evolve CI/CD pipelines, deployment tooling, and release processes.
  • Drive incident response and postmortems collaboratively.

Skills

Python
Observability
Backend services
Team collaboration
Production-grade systems

Tools

Kubernetes
GCP
CI/CD tooling

Job description

Role Overview

We're looking for an Infrastructure Software Engineer to sit at the intersection of software development and site reliability. Our software stack is cloud-native, built on active collaboration with popular open source technologies from developer tooling through production. You'll spend roughly half your time writing production code for product platform and core systems infrastructure and the other half driving reliability, observability, and operational health of these systems with the rest of the engineering team.

On the software side, you'll contribute to the full backend of our product - building and shipping features while also helping to define and build the platform and infrastructure layers those features run on. That means writing production-quality services and APIs and thinking carefully about what it means for a system to be maintainable, observable, and operable at scale. You'll bring an infrastructure-aware perspective to product development, helping the team build things that are designed to run well in production from the start. On the reliability side, you'll take the lead on how we build, deploy, monitor, and alert on our systems - but this isn't a sole-ownership role. You'll work alongside the broader engineering team, building the culture and practices around reliability as much as the tooling itself. That means driving postmortems and remediations collaboratively, establishing the frameworks that help everyone participate meaningfully in on-call and incident response, and closing the loop by improving internal tooling and systems.

Our platform sits at the heart of how utilities and energy providers manage an increasingly complex grid. The systems you build and operate need to be fast, correct, and available - grid operators make real-time decisions based on the data and interfaces we provide, and reliability has consequences that extend well beyond our codebase. We take that responsibility seriously, and we want someone who does too.

What You'll Do
  • Design and build backend services, APIs, and platform infrastructure that power our platform
  • Contribute to product features end to end - from backend logic and data pipelines to the deployment and operational scaffolding that ships them reliably
  • Lead the development of our observability stack - metrics, logging, and alerting - and establish the practices that keep the whole team engaged in production health
  • Own and evolve our CI/CD pipelines, deployment tooling, and release processes to make shipping software faster and safer
  • Drive incident response and postmortems collaboratively, and follow through by building the automation and tooling that closes the loop on recurring issues
What you'll bring
  • Minimum 3 years of experience in a software engineering role with a strong infrastructure or platform focus
  • Experience working with Python 3 - you're comfortable owning backend services and APIs end to end, not just writing scripts
  • Hands-on experience with Kubernetes and GCP (or comparable cloud), ideally in production environments
  • Experience working with CI/CD pipelines
  • Familiarity with observability tooling such as Prometheus and Grafana, and an instinct for what good monitoring looks like
  • A collaborative approach to reliability - you know how to bring a team along, not just fix things yourself
Nice to have
  • Experience with energy, utilities, or other regulated/critical infrastructure domains
  • Familiarity with the Bazel build tool
  • Familiarity with Go, Node.js and Python packaging
  • Experience with SQL, databases or data warehouses
What We Offer
  • Competitive base salary
  • Comprehensive benefits, including FSA and 401k for full time employees
  • Fully remote workplace with options for in office work in the Bay Area
  • Flexible PTO, which we encourage you to use!
  • A real impact on climate change - we're building the world we want to live in and we want you to join us!

The expected base salary for this role is $180,000 - $230,000 annually, depending on experience, skills, and qualifications.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Software Engineer, Site Reliability
Software Engineer, Site Reliability

fal • San Francisco (CA)

On-site
USD 180,000 - 250,000
Health, dental, and vision insurance
Relocation assistance
Learning and growth opportunities
+1
Senior Platform Engineer — Data Reliability
Senior Platform Engineer — Data Reliability

Duracell Power Center, a Duracell Authorized Licensee • San Jose (CA)

On-site
USD 170,000 - 200,000
Medical, dental, and vision
401(k)
PTO
+1
Senior Software Engineers | Platform | Infrastructure | Cloud | AI
Senior Software Engineers | Platform | Infrastructure | Cloud | AI

District-Partners • Arlington (VA)

Hybrid
USD 180,000 - 230,000
Software Engineer - Infrastructure
Software Engineer - Infrastructure

Emergentlabsinc • San Francisco (CA)

On-site
USD 110,000 - 150,000
401(k)
Health, dental, and vision insurance
Unlimited Paid Time Off
+1
Software Engineer 2 or 3 - Infrastructure
Software Engineer 2 or 3 - Infrastructure

NV Energy • Portland (OR)

On-site
USD 80,000 - 120,000
Founding Senior Software Engineer (Infrastructure) [32656]
Founding Senior Software Engineer (Infrastructure) [32656]

Stealth Startup • San Carlos (CA)

Hybrid
USD 215,000 - 290,000
Unlimited meals and snacks
Gym fee reimbursement
Internet and phone bills covered
Director of Infrastructure Engineering
Director of Infrastructure Engineering

Appsierra Group • United States

On-site
USD 350,000 - 500,000
Equity compensation eligibility
Performance-based bonuses
Health insurance reimbursement up to 1
+3
Senior/Staff Backend Software Engineer
Senior/Staff Backend Software Engineer

GridCARE • Redwood City (CA)

Hybrid
USD 184,000 - 284,000
Equity
Health coverage
Lunch provided
+2
Infrastructure Engineer
Infrastructure Engineer

Eden Prescott • San Francisco (CA)

On-site
Senior/Staff Backend Software Engineer
Senior/Staff Backend Software Engineer

Doist • Redwood City (CA)

Hybrid
USD 184,000 - 284,000
Competitive salary
Performance bonus
Equity
+5