Senior Systems Engineer

NextGenEnergyJobs

Atlanta (GA)

On-site

USD 120,000 - 180,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

NextGenEnergyJobs is seeking a senior Infrastructure Engineer to own operational health across domains, lead major incidents, and drive DR and patching programs. You will collaborate with Platform, Cybersecurity, Networking, and ITSM teams to harden platforms and improve observability.

The role requires hands-on management of Azure-based infrastructure, IaC, and scripting, with a focus on scalable, auditable operations. Excellent communication and mentorship are essential.

Qualifications

  • Bachelor's degree in Computer Science or related field
  • 5+ years in systems engineering or infrastructure operations
  • Experience leading major incidents and disaster recovery exercises
  • Strong fundamentals across server platforms, virtualization, cloud, networking and identity
  • Experience with Microsoft Azure and enterprise virtualization
  • Infrastructure-as-code with Terraform/Bicep/ARM

Responsibilities

  • Own the operational health of one or two infrastructure domains (e.g., server platforms, virtualization, cloud infrastructure).
  • Lead major incident response, coordinate responders, and own post-incident reviews and corrective actions.
  • Own patching programs and disaster recovery runbooks, coordinate tests and evidence.
  • Drive capacity planning, refresh cycles, and decommissioning of legacy systems.
  • Operate and improve monitoring and observability, build dashboards, and reduce noise.
  • Lead cross-team initiatives across Platform Engineering, Cybersecurity, Networking, and ITSM.
  • Define and document runbooks and standards; mentor engineers and ensure quality output.
  • Serve as senior escalation in on-call rotations and contribute to audits.

Skills

System administration
Incident management
Cloud infrastructure
Cross-team collaboration
Technical leadership

Education

Bachelor's degree in Computer Science or related field

Tools

Terraform
PowerShell
Python
Bash
Azure ARM/Bicep

Job description

Work with a Top 20 CPA and advisory firm that Accounts for Anything.

Key Responsibilities
  • Domain ownership : Own the operational health of one or two infrastructure domains (e.g., server platforms, virtualization, cloud infrastructure, identity, backup and recovery, specialty business systems). Keep them measurably healthy and improving.
  • Major incident leadership : Lead major incident response: drive technical resolution, coordinate responders, communicate to stakeholders, and own the post-incident review and corrective actions.
  • Patching and DR programs : Own the patching program for assigned domains — cadence, exception handling, reporting, continuous improvement. Own disaster recovery execution: maintain and exercise DR runbooks, coordinate tests, and ensure recovery objectives are met and evidenced.
  • Server lifecycle at scale : Drive capacity planning, refresh cycles, configuration baselines, and decommissioning of legacy systems.
  • Monitoring and observability : Operate and improve monitoring and observability across managed systems — tune alerting, eliminate noise, build dashboards, and contribute to anomaly-detection and AIOps initiatives.
  • Cross-team initiatives : Lead initiatives that span Platform Engineering, Cybersecurity, Networking, and application teams — controlled rollouts, hardening efforts, platform migrations. Land them without breaking production.
  • Standards and patterns : Define and document operational patterns, runbooks, and standards the team executes against and the auditors review.
  • Mentorship : Pair with Systems Engineers, run technical reviews, give substantive feedback, and grow the next tier. Quality of output from less senior engineers is part of your scope.
  • Operational partnership : Be the senior partner Platform Engineering, Cybersecurity, Networking, and IT Service Management call when they need operational input. Solve problems with them, not at them.
  • Specialty systems : Provide operational ownership of specialty business systems and legacy platforms supporting business-critical applications.
  • Security and Audit : Apply and validate security baselines, lead remediation of high-severity findings, and keep your domains' evidence map current.
  • Automation : Push toward repeatable, codified operations (IaC, automated evidence collection, scripted runbooks) instead of one-off manual work.
  • On-call : Participate in and serve as senior escalation for the on-call rotation, including after-hours support for high-severity incidents, change windows, and disaster recovery events.
Requirements
  • 5+ years in systems engineering, systems administration, or infrastructure operations, including time in a senior individual contributor or technical lead capacity.
  • Strong fundamentals across multiple infrastructure domains (server platforms, virtualization, cloud infrastructure, networking, identity, backup and recovery).
  • Experience operating production workloads in at least one major cloud platform.
  • Demonstrated experience leading major incidents and disaster recovery exercises.
  • Ability to produce clear architecture, operational, and decision documentation that holds up under audit and peer review.
  • Excellent written and verbal communication; able to explain trade-offs across technical and business audiences in plain language.
  • Comfortable mentoring less senior engineers and owning quality-of-output for one or more domains.
  • Comfortable serving as senior escalation in an on-call rotation.
  • Deep hands-on experience with Microsoft Azure (compute, networking, storage, identity, RBAC, monitoring).
  • Advanced administration experience with enterprise virtualization, hybrid Active Directory, and endpoint management platforms.
  • Infrastructure-as-code experience (Terraform, Bicep, ARM) and exposure to policy-as-code.
  • Advanced scripting / automation (PowerShell, Python, Bash) and experience automating operational work at scale.
  • Experience with observability or AIOps platforms.
  • Operations or administration experience with specialty business systems (e.g., IBM iSeries / AS400).
  • Experience operating within regulated environments (SOC 2, ISO 27001, HIPAA, PCI, NIST 800-171, CMMC, or similar).
  • Industry certifications (Microsoft AZ-104 / AZ-305 / AZ-500, MS-102, VMware VCP, Red Hat RHCSA/RHCE, ITIL Foundation, or equivalents).
  • Experience supporting a professional services, accounting, or financial services firm.
  • Bachelor's degree in Computer Science, Information Systems, or related field — or equivalent applicable years of experience.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Systems Engineer
Senior Systems Engineer

Jobtailor • Atlanta (GA)

On-site
USD 120,000 - 180,000
Systems Engineer
Systems Engineer

CRG • Greensboro (NC)

On-site
USD 90,000 - 120,000
Systems Operations and Engineering Manager
Systems Operations and Engineering Manager

LOOP • Greenville (SC), Spartanburg (SC), Anderson (SC)

On-site
USD 120,000 - 160,000
Senior System Engineer
Senior System Engineer

Hollingsworth & Vose • East Walpole (MA)

On-site
USD 120,000 - 180,000
Senior Infrastructure Engineer
Senior Infrastructure Engineer

Grayson Search Partners • Nashville (TN)

On-site
USD 120,000 - 180,000
Senior System Engineer
Senior System Engineer

Talent Acquisition LLC • Los Angeles (CA)

On-site
USD 90,000 - 120,000
Senior Lead System Engineer
Senior Lead System Engineer

Infinite Computer Solutions • Campus (IL)

On-site
USD 120,000 - 180,000
Systems Engineer
Systems Engineer

Compunnel, Inc. • Westlake (OH)

Hybrid
USD 100,000 - 120,000
Sr Systems Engineer
Sr Systems Engineer

Insight Global • New York (NY)

On-site
USD 120,000 - 150,000
Vice President, Lead Engineer – Infrastructure Operations
Vice President, Lead Engineer – Infrastructure Operations

Jobtailor • Illinois

On-site
USD 130,000 - 170,000