Lead Site Reliability Engineer

Lumen Technologies

United States

On-site

USD 120,000 - 180,000

Full time

7 hours ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Health benefits
Life insurance
Voluntary lifestyle benefits

Job summary

Lumen Technologies is seeking a Lead SRE to own the reliability of the NaaS platform from the ground up. You will collaborate with operations and development teams to drive technical direction, improve observability, and automate deployment and incident response.

You’ll lead on‑call management, champion SRE principles, and leverage AI-assisted workflows to accelerate platform administration, deployment, and investigations across cloud and network environments.

Qualifications

  • Bachelor’s degree or equivalent in engineering, computer science, or related field.
  • 8+ years in software development, systems engineering, and/or networking.
  • Hands‑on experience with at least one major cloud platform (AWS, Azure, or GCP).
  • Strong automation and infrastructure‑as‑code skills: Terraform, Ansible, and Python.
  • Experience running containerized workloads on Kubernetes.
  • Good observability/monitoring tooling knowledge (Datadog, CloudWatch, Grafana, Prometheus).
  • Incident management experience and blameless postmortems mindset.
  • Comfort using AI‑assisted/agentic tools in daily engineering practice.

Responsibilities

  • Reliability & Observability: build and maintain observability stack and proactive alerting.
  • Incident Management: lead on‑call rotation and postmortems.
  • Automation & Infrastructure: automate CI/CD and cloud provisioning with IaC; build tools to reduce toil.
  • Collaboration & Leadership: mentor junior engineers and document architectures.

Skills

Networking fundamentals
Observability
Incident management
Automation & IaC
AI-assisted tooling
Cloud platforms
Kubernetes
Python

Education

Bachelor’s degree in engineering, computer science, or related field

Tools

Terraform
Ansible
Python
Datadog
CloudWatch
Grafana
Prometheus
Kubernetes

Job description

Lumen is the trusted network for the AI‑powered world, connecting people, data, and applications through our expansive fiber network and connected ecosystem. We enable secure, high‑performance connectivity across cloud, edge, and AI workloads for enterprises, governments, and communities.

At Lumen, you’ll work on infrastructure customers rely on today and build for what’s next, where performance, security, and resilience matter.

This is a high accountability environment where bold ideas drive real innovation for our customers, partners, and industry. The work is challenging, expectations are clear, and trust is built into how we operate. If you’re ready to take ownership, deliver meaningful impact, and help shape the future of AI‑ready connectivity, join us today.

The Role

Lumen's Network as a Service (NaaS) platform delivers on-demand networking at scale. As Lead SRE, you'll own the reliability of that platform — partnering with operations teams and development counterparts to drive technical direction and resolve systemic issues across a broad range of network topologies and applications.

You’ll be accountable for platform observability, incident management, and automation, and you’ll coordinate across architecture, engineering, and systems development organizations to measurably improve reliability. You'll also use AI and agentic tooling to build utilities that accelerate deployment automation, platform administration, and incident investigation.

Success in this role draws on networking fundamentals, cloud platforms, software development and troubleshooting methodology, and a bias toward automating what you'd otherwise do twice. We're looking for a change maker — someone who sees where the platform should go next and drives meaningful impact for the customers who rely on it

Location

This role is designated as a fully remote position within the United States.

The Main Responsibilities
  • Reliability & Observability
    • Serve as subject matter expert for network automation platform applications, services, and hosting environments
    • Build and maintain the observability stack: instrument services, collect and curate metrics, and create dashboards and visualizations that make system health obvious at a glance
    • Define and tune proactive alerting so issues surface before customers feel them
    • Champion core SRE principles — SLIs, SLOs, and error budgets — and advocate for resilient, fault tolerant architecture
  • Incident Management
    • Participate in an on‑call rotation and lead incident response for service outages and unplanned downtime
    • Drive blameless postmortems and root cause analysis; own follow‑up actions through to completion
    • Prevent recurrence through process improvements, tooling, and knowledge sharing across teams
  • Automation & Infrastructure
    • Automate deployment pipelines (CI/CD) and cloud infrastructure provisioning, scaling, and configuration using infrastructure as code
    • Develop tools and utilities that reduce toil and empower operations and development teams to manage services independently
    • Apply AI‑assisted and agentic workflows to development, support, and investigation work
  • Collaboration & Leadership
    • Collaborate with cross‑functional development teams to support, enhance, and scale NaaS applications
    • Provide guidance and mentorship to junior engineers
    • Maintain clear documentation for processes and architecture
Required Qualifications
  • Bachelor’s degree or equivalent in engineering, computer science, or related field.
  • 8+ years in software development, systems engineering, and/or networking
  • 5+ years of related experience required.
  • Hands‑on experience with at least one major cloud platform (AWS, Azure, or GCP), including compute, networking, and identity services
  • Strong automation and infrastructure‑as‑code skills: Terraform, Ansible, and Python
  • Experience running containerized workloads on Kubernetes
  • Working knowledge of modern observability and monitoring tooling (e.g., Datadog, CloudWatch, Grafana, Prometheus), including building dashboards and defining alerts
  • Demonstrated experience with incident management and blameless postmortems
  • Comfort using AI‑assisted development and agentic tools as part of daily engineering practice
  • Understanding of network technologies including Internet, Ethernet, IPVPN, Edge Compute, and Optical transport
  • Strong listening and communication skills; able to operate with autonomy while knowing when to elevate
What We Look For in a Candidate
  • Bachelor’s degree or equivalent in engineering, computer science, or related field.
  • 8+ years in software development, systems engineering, and/or networking
  • 5+ years of related experience required.
  • Hands‑on experience with at least one major cloud platform (AWS, Azure, or GCP), including compute, networking, and identity services
  • Strong automation and infrastructure‑as‑code skills: Terraform, Ansible, and Python
  • Experience running containerized workloads on Kubernetes
  • Working knowledge of modern observability and monitoring tooling (e.g., Datadog, CloudWatch, Grafana, Prometheus), including building dashboards and defining alerts
  • Demonstrated experience with incident management and blameless postmortems
  • Comfort using AI‑assisted development and agentic tools as part of daily engineering practice
  • Understanding of network technologies including Internet, Ethernet, IPVPN, Edge Compute, and Optical transport
  • Strong listening and communication skills; able to operate with autonomy while knowing when to elevate
Preferred Qualifications
  • Multi‑cloud experience across AWS, Azure, and GCP
  • Asynchronous programming concepts and distributed systems design
  • Zero‑downtime deployment strategies
  • High availability and multi‑region architectures
  • Source control and CI/CD practices at scale
  • Experience applying agentic workflows to operational support and investigation
Compensation

This information reflects the anticipated base salary range for this position based on current national data. Minimums and maximums may vary based on location. Individual pay is based on skills, experience and other relevant factors.

Location Based Pay Ranges

$105,786 - $141,047 in these states: AL AR AZ FL GA IA ID IN KS KY LA ME MO MS MT ND NE NM OH OK PA SC SD TN UT VT WI WV WY

$111,074 - $148,099 in these states: CO HI MI MN NC NH NV OR RI

$116,364 - $155,152 in these states: AK CA CT DC DE IL MA MD NJ NY TX VA WA

Lumen offers a comprehensive package featuring a broad range of Health, Life, Voluntary Lifestyle benefits and other perks that enhance your physical, mental, emotional and financial wellbeing. We're able to answer any additional questions you may have about our bonus structure (short-term incentives, long-term incentives and/or sales compensation) as you move through the selection process. Learn more about Lumen's:Benefits

Bonus Structure

Requisition #: 343429

Life at Lumen

Life at Lumen is human and connected, even in a fast moving, AI‑focused organization. We set clear expectations and trust people to meet them. With real support and shared accountability, teams collaborate better, move faster, and deliver meaningful outcomes.

Our Lumen 8 behaviors guide how we interact, make decisions, and work together, shaping a culture built to perform and win.

To learn more about Life at Lumen and how we live the Lumen 8, please visit: https://jobs.lumen.com/global/en/life-at-lumen

Background Screening

If you are selected for a position, there will be a background screen, which may include checks for criminal records and/or motor vehicle reports and/or drug screening, depending on the position requirements. For more information on these checks, please refer to the Post Offer section of our FAQ page. Job‑related concerns identified during the background screening may disqualify you from the new position or your current role. Background results will be evaluated on a case‑by‑case basis.

Pursuant to the San Francisco Fair Chance Ordinance, we will consider for employment qualified applicants with arrest and conviction records.

Equal Employment Opportunities

We are committed to providing equal employment opportunities to all persons regardless of race, color, ancestry, citizenship, national origin, religion, veteran status, disability, genetic characteristic or information, age, gender, sexual orientation, gender identity, gender expression, marital status, family status, pregnancy, or other legally protected status (collectively, “protected statuses”). We do not tolerate unlawful discrimination in any employment decisions, including recruiting, hiring, compensation, promotion, benefits, discipline, termination, job assignments or training.

Privacy Notice

Lumen is committed to protecting the privacy and security of personal information collected during the recruitment and hiring process. Our Applicant Privacy Notice explains how we collect, use, disclose, and protect applicant information, as well as how individuals may request access to or deletion of their personal data.

To review Lumen’s Global Employment Applicant and Talent Community Privacy Notice, please visit: https://jobs.lumen.com/global/en/privacy-notice

Disclaimer

The job responsibilities described above indicate the general nature and level of work performed by employees within this classification. It is not intended to include a comprehensive inventory of all duties and responsibilities for this job. Job duties and responsibilities are subject to change based on evolving business needs and conditions.

In any materials you submit, you may redact or remove age‑identifying information such as age, date of birth, or dates of school attendance or graduation. You will not be penalized for redacting or removing this information.

Please be advised that Lumen does not require any form of payment from job applicants during the recruitment process. All legitimate job openings will be posted on our official website or communicated through official company email addresses. If you encounter any job offers that request payment in exchange for employment at Lumen, they are not for employment with us, but may relate to another company with a similar name.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Sr. Manager Software Development
Sr. Manager Software Development

Worky • United States

On-site
USD 140,000 - 190,000
SENIOR LEAD COMMUNICATIONS MANAGER
SENIOR LEAD COMMUNICATIONS MANAGER

Lumen • Dover (DE)

On-site
USD 116,000 - 155,000
Health benefits
SENIOR LEAD COMMUNICATIONS MANAGER
SENIOR LEAD COMMUNICATIONS MANAGER

Lumen • Lincoln (NE)

On-site
USD 106,000 - 141,000
Director, Applied AI Solutions (Remote, US)
Director, Applied AI Solutions (Remote, US)

Visa Hunt • United States

Remote
USD 152,000 - 278,000
Health benefits
Life benefits
Voluntary lifestyle benefits
FIELD TECHNICIAN II PUB SEC
FIELD TECHNICIAN II PUB SEC

Lumen • New York (NY)

On-site
USD 55,000 - 73,000
Health benefits
Life insurance
Voluntary benefits
+1
Lead IT Systems Engineer – Public Sector
Lead IT Systems Engineer – Public Sector

Lumen • Jackson (MS)

On-site
USD 106,000 - 141,000
Health benefits
Life insurance
Voluntary lifestyle benefits
Lead IT Systems Engineer – Public Sector
Lead IT Systems Engineer – Public Sector

Lumen • Denver (CO)

On-site
USD 111,000 - 148,000
Health benefits
Life insurance
Bonus program
Sr Lead Solution Architect - SAP
Sr Lead Solution Architect - SAP

Lumen • Salem (OR)

On-site
USD 139,000 - 164,000
Health benefits
Life insurance
Bonus structure
+1
Manager Network Planning & Strategic Communications (Remote, US)
Manager Network Planning & Strategic Communications (Remote, US)

Visa Hunt • United States

Hybrid
USD 120,000 - 190,000
Principal Architect - Identity, Access, and Management (Remote, US)
Principal Architect - Identity, Access, and Management (Remote, US)

Visa Hunt • United States

Remote
USD 152,000 - 223,000
Health, Life, and other comprehensive/
AI-focused benefits and incentives