Site Reliability Engineer, GovCloud 24x7

Engg

Denver (CO)

On-site

USD 120,000 - 180,000

Full time

9 days ago
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Salesforce GovCloud is seeking a Site Reliability Engineer to maintain reliability and performance of mission-critical cloud services. You will lead incident response, perform RCAs, and collaborate with cross-functional teams to implement robust prevention and automation. A U.S.

citizen with the ability to obtain clearance is required, along with a related technical degree and strong Linux/Unix experience. The role emphasizes 24/7 coverage, security compliance, and mentoring teammates while

Qualifications

  • U.S. citizen with ability to obtain Moderate Public Trust clearance
  • Related technical degree required
  • Experience with enterprise-scale internet service operations or systems
  • Strong knowledge of networking protocols and TCP/IP
  • Solid Linux/Unix administration and scripting skills

Responsibilities

  • Maintain system reliability and high performance for customer-facing services
  • Lead incident response during Sev0/Sev1 events and participate in post-incident reviews
  • Contribute to RCAs and work with Global Solutions to implement preventative solutions
  • Ensure SRE activities align with security standards and compliance
  • Mentor teammates and explore emerging technologies
  • Support 24/7 operations and on-call rotations
  • Automate detection and resolution of recurring issues
  • Collaborate to improve engineering workflows and reduce toil

Skills

TCP/IP
Unix/Linux
Security monitoring
Incident management
Python/Go
AWS
Kubernetes
Jenkins
Spinnaker
Chef/Puppet
SRE practices

Education

Bachelor's degree in Computer Science or related field

Tools

AWS
Kubernetes
Jenkins
Spinnaker
Chef/Puppet

Job description

Job Category Software Engineering

To get the best candidate experience, please consider applying for a maximum of 3 roles within 12 months to ensure you are not duplicating efforts. About Futureforce University Recruiting Our Futureforce University Recruiting program is dedicated to attracting, retaining and cultivating talent. Our interns and new graduates work on real projects that affect how our business runs, giving them the opportunity to make a tangible impact on the future of our company. With offices all over the world, our recruits have the chance to collaborate and connect with fellow employees on a global scale. We offer job shadowing, mentorship programs, talent development courses, and much more.

Job Details About Salesforce

Salesforce is the #1 AI CRM, where humans with agents drive customer success together. Here, ambition meets action. Tech meets trust. And innovation isn’t a buzzword — it’s a way of life. The world of work as we know it is changing and we’re looking for Trailblazers who are passionate about bettering business and the world through AI, driving innovation, and keeping Salesforce’s core values at the heart of it all. Ready to level-up your career at the company leading workforce transformation in the agentic era? You’re in the right place! Agentforce is the future of AI, and you are the future of Salesforce. Applications will be accepted until 12/14/2026.

The Experience

Join our team and contribute to the operational excellence of the Salesforce GovCloud! Are you passionate about ensuring the reliability and performance of mission-critical cloud services? Salesforce is seeking a talented Site Reliability Engineer to join our dynamic team, supporting our GovCloud environment. As a key member of our Site Reliability organization, you’ll play a vital role in maintaining 99.99% uptime for customer-facing services, proactively addressing issues, and ensuring the security of our data. We foster a collaborative and innovative culture, where you’ll work alongside skilled engineers to solve complex problems and drive continuous improvement.

Please Note: This position requires a successful background investigation and the ability to obtain and maintain a specific level of U.S. government background clearance. Details will be provided during the interview process.

Shift Requirements

This role involves shift work, including night shifts, as part of a 24/7 support team. We provide a rotating schedule and ensure adequate compensation for shift differentials.

GovCloud Incident Response (GIR)

Maintains critical infrastructure stability through incident management, automated alerting, collaborative smart-hands support, thoughtful retrospectives, and sustainable long-term remediation strategy.

What You’ll Actually Be Doing
  • Maintain system reliability and high performance across customer-facing services by supporting foundational infrastructure health.
  • Lead incident response efforts during major operational events (e.g., Sev0/Sev1) and collaborate in post-incident reviews to drive problem resolution.
  • Populate and participate in Root Cause Analyses (RCAs) and partner with Global Solutions teams to implement preventative solutions.
  • Ensure Site Reliability engineering activities align with organizational compliance, security standards, and operational guidelines.
  • Demonstrate passion for collaborative problem-solving and addressing complex technical challenges across cross-functional teams.
  • Mentor and collaborate with teammates to explore emerging technologies, foster continuous learning, and support professional growth.
  • Adapt to dynamic, fast-paced operational needs while prioritizing competing tasks effectively.
  • Work to automate detection and resolution of recurring issues in the production environment.
  • Collaborate on improving existing workflows to streamline engineering practices and minimize operations and engineering toil.
You’re Our Person If...
  • You’re a U.S. citizen (U.S. born or naturalized) who does not hold dual citizenship, and you agree to complete a Minimum Background Investigation (MBI) for a Moderate Public Trust position with the U.S. federal government or other clearances as deemed appropriate for the role.
  • You have a related technical degree.
  • You have demonstrated experience with enterprise-scale internet service operations or systems engineering.
  • You have expertise in TCP/IP related technologies (networking protocols, network programming, etc.).
  • You have command-line expertise and administration knowledge of Unix variants (Linux/Solaris/BSD) systems (e.g., RedHat Enterprise Linux and Solaris).
  • You have solid comprehension of security monitoring systems and infrastructure administration.
  • You have effective written and verbal communication skills with a focus on collaborative teamwork.
  • You’re familiar with Incident Management principles and IT service operation frameworks.
  • You can support a 24/7 high-availability operational environment, including shift and on-call rotations as needed.
  • You have experience provisioning, operating, and running AWS/C2S based infrastructure and systems.
  • You’re proficient in scripting or programming languages such as Python, Go, or similar technologies.
  • You have an open mindset toward leveraging AI technologies to enhance engineering productivity and continuous learning across technical domains.
Even Better If...
  • You have prior Chef/Puppet or automated deployment experience.
  • You have prior Jenkins/Bamboo/Spinnaker pipeline execution experience.
  • You have experience supporting and maintaining monitoring and alert systems.
  • You have experience supporting and maintaining Java applications.
  • You have hands‑on experience configuring and running AWS, using the CLI/SDKs.
  • You hold certifications in Linux+, RedHat, and AWS.
  • You have experience supporting and leading Kubernetes based applications and services.
  • You’ve taken part in blameless retrospectives, learning from incidents, and conducting post-incident investigations, including incident
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Systems Engineering Associate - GovCloud [Salesforce National Security]
Systems Engineering Associate - GovCloud [Salesforce National Security]

Salesforce • Virginia (MN)

On-site
USD 140,000 - 190,000
Site Reliability Engineer, GovCloud 24x7
Site Reliability Engineer, GovCloud 24x7

salesforce.com, inc. • McLean (VA)

On-site
USD 114,000 - 125,000
Health insurance
401(k)
Employee stock purchase program
+1
Site Reliability Engineer, GovCloud 24x7
Site Reliability Engineer, GovCloud 24x7

salesforce.com, inc. • Burlington (MA)

On-site
USD 114,000 - 125,000
Time off programs
Medical insurance
Dental insurance
+6
Site Reliability Engineer, GovCloud 24x7
Site Reliability Engineer, GovCloud 24x7

salesforce.com, inc. • Denver (CO)

On-site
USD 114,000 - 125,000
Medical
Dental
Vision
+3
Site Reliability Engineer, GovCloud 24x7
Site Reliability Engineer, GovCloud 24x7

Salesforce • McLean (VA)

On-site
USD 114,000 - 125,000
Site Reliability Engineer, GovCloud 24x7
Site Reliability Engineer, GovCloud 24x7

Salesforce • Denver (CO)

On-site
USD 114,000 - 125,000
Site Reliability Engineer, GovCloud 24x7
Site Reliability Engineer, GovCloud 24x7

Salesforce • Burlington (MA)

On-site
USD 114,000 - 125,000
GovCloud Distributed Systems Software Engineer
GovCloud Distributed Systems Software Engineer

Engg • Bellevue (WA)

On-site
USD 190,000 - 230,000
Systems Engineering Associate - GovCloud [Salesforce National Security]
Systems Engineering Associate - GovCloud [Salesforce National Security]

109 Computable Insights LLC • Herndon (VA)

On-site
USD 111,000 - 122,000
Competitive salary range: $111,000 – $122,000 annually
Time-off programs
Medical, dental, vision benefits
+1
Site Reliability Engineer, GovCloud 24x7
Site Reliability Engineer, GovCloud 24x7

Salesforce, Inc. • Burlington (MA), Northern (KY)

Hybrid
USD 114,000 - 125,000