Site Reliability Engineer, GovCloud 24x7

salesforce.com, inc.

Burlington (MA)

On-site

USD 114,000 - 125,000

Full time

8 days ago
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Time off programs
Medical insurance
Dental insurance
Vision care
Mental health support
Paid parental leave
Life and disability insurance
401(k)
Employee stock purchasing program

Job summary

Salesforce GovCloud is seeking a Site Reliability Engineer to join a 24/7 high-availability operations team. You’ll maintain reliability for mission-critical cloud services, lead Sev0/Sev1 responses, and help drive improvements across incident management and automation.

You will work with AWS, Kubernetes, and CI/CD tooling, supporting secure, scalable infrastructure while collaborating with cross-functional teams and stakeholders.

Qualifications

  • Related technical degree required.
  • Experience with enterprise-scale internet service operations or systems engineering.
  • Strong knowledge of TCP/IP networking and related technologies.
  • Experience with incident management processes and post-incident reviews.
  • Ability to script and automate with Python/Go or similar.
  • Familiarity with AWS, Kubernetes, and CI/CD pipelines.
  • Understanding of security best practices in cloud environments.

Responsibilities

  • Maintain system reliability and high performance for customer-facing services.
  • Lead incident response for major operational events and participate in post-incident reviews.
  • Contribute to RCAs and collaborate with Global Solutions teams on preventative measures.
  • Ensure Site Reliability activities align with security standards and operational guidelines.
  • Mentor teammates and explore emerging technologies for continuous learning.
  • Support 24/7 high-availability operations with on-call rotations.
  • Automate detection and resolution of recurring production issues.
  • Improve workflows to reduce operational toil and optimize tooling.

Skills

TCP/IP networking
Unix/Linux command line
Incident management
SRE/DevOps mindset
Automation scripting
Cloud technologies
Security monitoring
Collaboration & communication
Agile methodologies

Education

Related technical degree

Tools

AWS
Kubernetes
Jenkins
Spinnaker
Puppet
Chef
CI/CD tooling
Linux

Job description

About Futureforce University Recruiting

Our Futureforce University Recruiting program is dedicated to attracting, retaining and cultivating talent. Our interns and new graduates work on real projects that affect how our business runs, giving them the opportunity to make a tangible impact on the future of our company. With offices all over the world, our recruits have the chance to collaborate and connect with fellow employees on a global scale. We offer job shadowing, mentorship programs, talent development courses, and much more.

Job Category

Software Engineering

Job Details
About Salesforce

Salesforce is the #1 AI CRM, where humans with agents drive customer success together. Here, ambition meets action. Tech meets trust. And innovation isn't a buzzword - it's a way of life. The world of work as we know it is changing and we're looking for Trailblazers who are passionate about bettering business and the world through AI, driving innovation, and keeping Salesforce's core values at the heart of it all.

Ready to level-up your career at the company leading workforce transformation in the agentic era? You're in the right place! Agentforce is the future of AI, and you are the future of Salesforce.

The Experience

Join our team and contribute to the operational excellence of the Salesforce GovCloud! Are you passionate about ensuring the reliability and performance of mission-critical cloud services? Salesforce is seeking a talented Site Reliability Engineer to join our dynamic team, supporting our GovCloud environment. As a key member of our Site Reliability organization, you'll play a vital role in maintaining 99.99% uptime for customer-facing services, proactively addressing issues, and ensuring the security of our data. We foster a collaborative and innovative culture, where you'll work alongside skilled engineers to solve complex problems and drive continuous improvement.

Please Note: This position requires a successful background investigation and the ability to obtain and maintain a specific level of U.S. government background clearance. Details will be provided during the interview process.
Shift Requirements: This role involves shift work, including night shifts, as part of a 24/7 support team. We provide a rotating schedule and ensure adequate compensation for shift differentials.

GovCloud Incident Response (GIR) maintains critical infrastructure stability through incident management, automated alerting, collaborative smart-hands support, thoughtful retrospectives, and sustainable long-term remediation strategy.

What You'll Actually Be Doing
  • Maintain system reliability and high performance across customer-facing services by supporting foundational infrastructure health.
  • Lead incident response efforts during major operational events (e.g., Sev0/Sev1) and collaborate in post-incident reviews to drive problem resolution.
  • Populate and participate in Root Cause Analyses (RCAs) and partner with Global Solutions teams to implement preventative solutions.
  • Ensure Site Reliability engineering activities align with organizational compliance, security standards, and operational guidelines.
  • Demonstrate passion for collaborative problem-solving and addressing complex technical challenges across cross-functional teams.
  • Mentor and collaborate with teammates to explore emerging technologies, foster continuous learning, and support professional growth.
  • Adapt to dynamic, fast-paced operational needs while prioritizing competing tasks effectively.
  • Work to automate detection and resolution of recurring issues in the production environment.
  • Collaborate on improving existing workflows to streamline engineering practices and minimize operations and engineering toil.
You're Our Person If...
  • You're a U.S. citizen (U.S. born or naturalized) who does not hold dual citizenship, and you agree to complete a Minimum Background Investigation (MBI) for a Moderate Public Trust position with the U.S. federal government or other clearances as deemed appropriate for the role.
  • You have a related technical degree.
  • You have demonstrated experience with enterprise-scale internet service operations or systems engineering.
  • You have expertise in TCP/IP related technologies (networking protocols, network programming, etc.).
  • You have command-line expertise and administration knowledge of Unix variants (Linux/Solaris/BSD) systems (e.g., Red Hat Enterprise Linux and Solaris).
  • You have solid comprehension of security monitoring systems and infrastructure administration.
  • You have effective written and verbal communication skills with a focus on collaborative teamwork.
  • You're familiar with Incident Management principles and IT service operation frameworks.
  • You can support a 24/7 high-availability operational environment, including shift and on-call rotations as needed.
  • You have experience provisioning, operating, and running AWS/C2S based infrastructure and systems.
  • You're proficient in scripting or programming languages such as Python, Go, or similar technologies.
  • You have an open mindset toward leveraging AI technologies to enhance engineering productivity and continuous learning across technical domains.
Even Better If...
  • You have prior Chef/Puppet or automated deployment experience.
  • You have prior Jenkins/Bamboo/Spinnaker pipeline execution experience.
  • You have experience supporting and maintaining monitoring and alert systems.
  • You have experience supporting and maintaining Java applications.
  • You have hands‑on experience configuring and running AWS, using the CLI/SDKs.
  • You hold certifications in Linux+, RedHat, and AWS.
  • You have experience supporting and leading Kubernetes based applications and services.
  • You've taken part in blameless retrospectives, learning from incidents, and conducting post-incident investigations, including incident analysis as well as performance evaluations of responders.
  • You have working knowledge of and interest in resilience engineering, including concepts such as Safety II - looking at how things go right instead of how things go wrong, being proactive instead of reactive, and investigating complex sociotechnical systems.
  • You're familiar with Agile process and DevOps.
  • You have experience using AI tools (e.g., Claude Code, GitHub Copilot, Codex, Cursor, etc.) in development workflows.
  • You have advanced prompt engineering skills and the ability to write precise, structured prompts and cultivate the system context that makes AI outputs reliable, secure, and production-ready.

Must be a U.S. Citizen operating on U.S. Soil with ability to meet customer and government screening standards applicable to this role, including a Criminal Justice Information Services screening with fingerprint scan. Due to the citizenship requirement for this role, which supports U.S. federal, state, and/or local government customers, citizenship will be verified through two of the following REAL ID Act documents: U.S. Passport, Passport Card, REAL Driver's License, Global Entry Card, U.S. Government CAC/PIV.

Unleash Your Potential

When you join Salesforce, you'll be limitless in all areas of your life. Our benefits and resources support you to find balance and be your best, and our AI agents accelerate your impact so you can do your best. Together, we'll bring the power of Agentforce to organizations of all sizes and deliver amazing experiences that customers love.

Accommodations

If you need a reasonable accommodation during the application or the recruiting process, please submit a request via this Accommodations Request Form.

Please note that Salesforce uses artificial intelligence (AI) tools to help our recruiters assess and evaluate candidates' resumes and qualifications throughout the recruiting process. Humans will always make any candidate selection and hiring decisions. Please see our Candidate Privacy Statement for more information about how we use your personal data and your rights, including with regard to use of AI tools and opt out options.

Posting Statement

Salesforce is an equal opportunity employer and maintains a policy of non-discrimination with all employees and applicants for employment. What does that mean exactly? It means that at Salesforce, we believe in equality for all. And we believe we can lead the path to equality in part by creating a workplace that's inclusive, and free from discrimination.

Any employee or potential employee will be assessed on the basis of merit, competence and qualifications - without regard to race, religion, color, national origin, sex, sexual orientation, gender expression or identity, transgender status, age, disability, veteran or marital status, political viewpoint, or other classifications protected by law. This policy applies to current and prospective employees, no matter where they are in their Salesforce employment journey. It also applies to recruiting, hiring, job assignment, compensation, promotion, benefits, training, assessment of job performance, discipline, termination, and everything in between. Recruiting, hiring, and promotion decisions at Salesforce are fair and based on merit. The same goes for compensation, benefits, promotions, transfers, reduction in workforce, recall, training, and education.

In the United States, compensation offered will be determined by factors such as location, job level, job-related knowledge, skills, and experience.

Certain roles may be eligible for incentive compensation, equity, and benefits.

  • time off programs
  • medical
  • dental
  • vision
  • mental health support
  • paid parental leave
  • life and disability insurance
  • 401(k)
  • employee stock purchasing program

More details about company benefits can be found at the following link: https://www.salesforcebenefits.com.

At Salesforce, we believe in equitable compensation practices that reflect the dynamic nature of labor markets across various regions.

The typical base salary range for this position is $114,400 - $125,400 annually.

The range represents base salary only, and does not include company bonus, incentive for sales roles, equity or benefits, as applicable.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Site Reliability Engineer, GovCloud 24x7
Site Reliability Engineer, GovCloud 24x7

salesforce.com, inc. • Denver (CO)

On-site
USD 114,000 - 125,000
Medical
Dental
Vision
+3
Site Reliability Engineer, GovCloud 24x7
Site Reliability Engineer, GovCloud 24x7

salesforce.com, inc. • McLean (VA)

On-site
USD 114,000 - 125,000
Health insurance
401(k)
Employee stock purchase program
+1
Site Reliability Engineer, GovCloud 24x7
Site Reliability Engineer, GovCloud 24x7

Salesforce • McLean (VA)

On-site
USD 114,000 - 125,000
Site Reliability Engineer, GovCloud 24x7
Site Reliability Engineer, GovCloud 24x7

Salesforce • Denver (CO)

On-site
USD 114,000 - 125,000
Site Reliability Engineer, GovCloud 24x7
Site Reliability Engineer, GovCloud 24x7

Salesforce • Burlington (MA)

On-site
USD 114,000 - 125,000
Site Reliability Engineer, GovCloud 24x7
Site Reliability Engineer, GovCloud 24x7

Salesforce, Inc. • Burlington (MA), Northern (KY)

Hybrid
USD 114,000 - 125,000
Software Engineer- GovCloud
Software Engineer- GovCloud

salesforce.com, inc. • Bellevue (WA)

On-site
USD 117,000 - 177,000
Software Engineer- GovCloud
Software Engineer- GovCloud

salesforce.com, inc. • Herndon (VA)

On-site
USD 117,000 - 177,000
Software Engineer- GovCloud
Software Engineer- GovCloud

salesforce.com, inc. • Denver (CO)

On-site
USD 117,000 - 177,000
Software Engineer- GovCloud
Software Engineer- GovCloud

salesforce.com, inc. • San Francisco (CA)

On-site
USD 117,000 - 194,000