Senior Site Reliability Engineer

Oracle

Nashville (TN)

On-site

USD 81,100 - 187,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Oracle in Nashville is seeking a Senior Infrastructure/SRE professional to design and operate reliable Windows and Linux environments. You will partner with software teams to build scalable infrastructure, automate tasks, and ensure security and compliance while supporting incident response and change control.

Responsibilities include deploying and validating apps, patch management, and providing runbooks and documentation.

Qualifications

  • Hands-on experience administering Windows Server and Linux systems.
  • Ability to access deployed hosts and perform post-deployment configuration and validation.
  • Experience installing, configuring, and validating applications in Windows Server and Linux environments.
  • Ability to troubleshoot operating-system-level, service-level, and application-level issues.
  • Working knowledge of system services, permissions, configuration files, logs, and resource utilization.

Responsibilities

  • Design and architect reliable, secure, and maintainable infrastructure and services. Take proactive steps to ensure solutions meet defined reliability and functionality requirements.
  • Identify operational risks, dependencies, performance issues, and potential failure points before they affect service.
  • Translate business and application requirements into practical technical solutions.
  • Install, configure, deploy, and validate applications across Windows Server and Linux environments.
  • Perform structured builds and deployments using runbooks, scripts, readiness checks, and post-build validation.
  • Troubleshoot operating system, service, application, installation, patching, permissions, certificate, and connectivity issues.
  • Monitor system performance and implement improvements to availability, reliability, and operational efficiency.
  • Develop and maintain scripts and automation that reduce manual effort and improve consistency.
  • Plan and execute operating system, middleware, and application patching using established change-control and rollback procedures.
  • Lead or support incident investigations, root cause analyses, and corrective actions.
  • Support vulnerability remediation, system hardening, STIG compliance, and other security-driven changes.
  • Maintain accurate runbooks, deployment procedures, troubleshooting guides, and operational records.
  • Provide technical guidance and mentorship to junior engineers.
  • Communicate status, risks, blockers, and escalation details clearly to stakeholders.

Skills

Windows & Linux Admin
System Troubleshooting
Automation Scripting
Cloud/OCI Familiarity
Network Troubleshooting
Documentation & Communication
Incident Response
Security & STIGs

Tools

PowerShell
Bash
Python
Ansible
Chef

Job description

Job Description

Takes proactive steps to design and architect infrastructure and service to ensure reliability and functionality. Forecasts demands and responds to capacity needs. Collaborates with software development teams to develop reliable and scalable infrastructures. Performs data collection to maintain and optimize operations and reliability. Leverages knowledge to perform incident response and/or maintenance tasks. Provides health and performance reporting. Identifies opportunities for automation. Communicates about services and identifies and explains the potential impact of changes. Provides support for technology and documents incidents. Experiments with new tools and assesses potential impact and develops knowledge of site reliability trends.

Responsibilities
Key Responsibilities
  • Design and architect reliable, secure, and maintainable infrastructure and services. Take proactive steps to ensure solutions meet defined reliability and functionality requirements.
  • Identify operational risks, dependencies, performance issues, and potential failure points before they affect service.
  • Translate business and application requirements into practical technical solutions.
  • Install, configure, deploy, and validate applications across Windows Server and Linux environments.
  • Perform structured builds and deployments using runbooks, scripts, readiness checks, and post-build validation.
  • Troubleshoot operating system, service, application, installation, patching, permissions, certificate, and connectivity issues.
  • Monitor system performance and implement improvements to availability, reliability, and operational efficiency.
  • Develop and maintain scripts and automation that reduce manual effort and improve consistency.
  • Plan and execute operating system, middleware, and application patching using established change-control and rollback procedures.
  • Lead or support incident investigations, root cause analyses, and corrective actions.
  • Support vulnerability remediation, system hardening, STIG compliance, and other security-driven changes.
  • Maintain accurate runbooks, deployment procedures, troubleshooting guides, and operational records.
  • Provide technical guidance and mentorship to junior engineers.
  • Communicate status, risks, blockers, and escalation details clearly to stakeholders.
Core Skills And Qualifications
Windows and Linux System Administration
  • Hands-on experience administering Windows Server and/or Linux systems.
  • Ability to access deployed hosts and perform post-deployment configuration and validation.
  • Experience installing, configuring, and validating applications in Windows Server and Linux environments.
  • Ability to troubleshoot operating-system-level, service-level, and application-level issues.
  • Working knowledge of system services, permissions, configuration files, logs, and resource utilization.
Manual Build And Deployment Experience
  • Experience performing structured build and deployment tasks using runbooks, deployment guides, and technical procedures.
  • Ability to execute scripts, validate outputs, and correct common build or configuration issues.
  • Familiarity with build handoffs, environment‑readiness checks, deployment validation, and post‑build verification.
  • Ability to follow detailed implementation steps while identifying and documenting exceptions or deviations.
Troubleshooting and Operational Support
  • Ability to investigate failed services, installation errors, patching failures, application startup problems, permissions issues, and connectivity failures.
  • Experience reviewing logs, event viewers, service status, configuration files, ports, certificates, and permissions.
  • Strong analytical and problem‑solving skills, with the ability to isolate root causes and recommend practical solutions.
  • Ability to elevate issues clearly by documenting symptoms, troubleshooting steps, findings, impact, and recommended actions.
  • Experience supporting production or other business‑critical environments.
Scripting and Automation
  • Hands‑on experience with one or more of the following:
  • PowerShell
  • Bash
  • Python
  • Ansible
  • Chef

Candidates should be able to run, modify, validate, and troubleshoot existing scripts and understand basic automation concepts.

Patching and Software Maintenance
  • Experience applying operating system, middleware, and application patches.
  • Ability to follow patching procedures, validate successful completion, and troubleshoot failures.
  • Understanding of maintenance windows, change control, rollback planning, and post‑change validation.
Cloud and OCI Familiarity
  • Familiarity with Oracle Cloud Infrastructure or another major cloud platform.
  • Understanding of cloud compute, storage, networking, identity, and environment‑provisioning concepts.
  • Experience supporting applications in cloud‑hosted or hybrid environments.
Network Troubleshooting
  • Working knowledge of DNS, firewalls, routing, load balancers, ports, and certificates.
  • Ability to identify basic connectivity issues between hosts, applications, and services.
  • Familiarity with standard network diagnostic tools.
Cybersecurity and Compliance
  • Familiarity with STIGs, vulnerability remediation, system hardening, and compliance‑driven configuration.
  • Ability to support cybersecurity remediation activities without disrupting application functionality.
  • Experience in federal, government‑hosted, or regulated environments is highly valued.
Documentation and Communication
  • Ability to follow detailed technical instructions, runbooks, and change procedures.
  • Strong attention to detail when documenting completed work, issues, deviations, and validation results.
  • Experience working in ticketing, incident‑management, or change‑management systems.
  • Clear written and verbal communication skills for status updates, handoffs, and escalations.
  • Ability to collaborate effectively with engineering, operations, cybersecurity, networking, and client‑facing teams.
Preferred Qualifications
  • Experience supporting federal clients or government‑hosted environments.
  • Knowledge of cybersecurity workflows, STIG implementation, or federal compliance requirements.
  • Experience with Citrix technologies.
  • Legacy infrastructure or application‑support experience.
  • Millennium or Cerner application knowledge.
  • Experience with infrastructure‑as‑code or configuration‑management tools.
  • Production support, incident response, or SRE operational experience.
  • Knowledge of monitoring, alerting, centralized logging, observability, and reliability practices.
Qualifications

Disclaimer:

Certain U.S. based or U.S. customer or client‑facing roles may be required to comply with applicable requirements, such as immunization/occupational health mandates, and/or drug testing requirements.

Range and benefit information provided in this posting are specific to the stated locations only

US: Hiring Range in USD from: $81,100 - $187,000 per year. May be eligible for bonus and equity.

Oracle maintains broad salary ranges for its roles in order to account for variations in knowledge, skills, experience, market conditions and locations, as well as reflect Oracle's differing products, industries and lines of business.

Candidates are typically placed into the range based on the preceding factors as well as internal peer equity.

Oracle US offers a comprehensive benefits package which includes the following:

  • Medical, dental, and vision insurance, including expert medical opinion
  • Short term disability and long term disability
  • Life insurance and AD&D
  • Supplemental life insurance (Employee/Spouse/Child)
  • Health care and dependent care Flexible Spending Accounts
  • Pre‑tax commuter and parking benefits
  • 401(k) Savings and Investment Plan with company match
  • Paid time off: Flexible Vacation is provided to all eligible employees assigned to a salaried (non‑overtime eligible) position. Accrued Vacation is provided to all other employees eligible for vacation benefits. For employees working at least 35 hours per week, the vacation accrual rate is 13 days annually for the first three years of employment and 18 days annually for subsequent years of employment. Vacation accrual is prorated for employees working between 20 and 34 hours per week. Employees working fewer than 20 hours per week are not eligible for vacation.
  • 11 paid holidays
  • Paid sick leave: 72 hours of paid sick leave upon date of hire. Refreshes each calendar year. Unused balance will carry over each year up to a maximum cap of 112 hours.
  • Paid parental leave
  • Adoption assistance
  • Employee Stock Purchase Plan
  • Financial planning and group legal
  • Voluntary benefits including auto, homeowner and pet insurance

The role will generally accept applications for at least three calendar days from the posting date or as long as the job remains posted.

Career Level - IC3

About Us

Only Oracle brings together the data, infrastructure, applications, and expertise to power everything from industry innovations to life‑saving care. And with AI embedded across our products and services, we help customers turn that promise into a better future for all. Discover your potential at a company leading the way in AI and cloud solutions that impact billions of lives.

True innovation starts when everyone is empowered to contribute. That’s why we’re committed to growing a workforce that promotes opportunities for all with competitive benefits that support our people with flexible medical, life insurance, and retirement options. We also encourage employees to give back to their communities through our volunteer programs.

We’re committed to including people with disabilities at all stages of the employment process. If you require accessibility assistance or accommodation for a disability at any point, let us know by emailing accommodation‑request_mb@oracle.com or by calling 1-888-404-2494 in the United States.

Oracle is an Equal Employment Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, sexual orientation, gender identity, disability and protected veterans’ status, or any other characteristic protected by law. Oracle will consider for employment qualified applicants with arrest and conviction records pursuant to applicable law.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer 2
Site Reliability Engineer 2

Oracle • Nashville (TN)

On-site
USD 69,800 - 148,300
Health insurance
401(k) with company match
Paid time off
Principal Site Reliability Engineer
Principal Site Reliability Engineer

Oracle • Nashville (TN)

On-site
USD 84,900 - 209,500
Medical Insurance
Dental Insurance
Vision Insurance
+4
Lead Principal Site Reliability Engineer
Lead Principal Site Reliability Engineer

Oracle • Nashville (TN)

On-site
USD 96,000 - 265,000
Health insurance
Paid time off
401(k) match
Senior Systems Engineer
Senior Systems Engineer

Oracle • Nashville (TN)

On-site
USD 68,000 - 148,000
Medical, dental, and vision insurance
Commuter benefits
401(k) match
+3
Senior Systems Engineer
Senior Systems Engineer

Oracle • Anthony, NM (NM)

On-site
USD 68,000 - 148,000
Health insurance
401(k) plan
Paid time off
Senior Systems Engineer
Senior Systems Engineer

Oracle • Port Washington (WI)

On-site
USD 68,000 - 148,000
Medical, dental, and vision insurance
401(k) with company match
Paid time off
Senior Core Infrastructure Engineer
Senior Core Infrastructure Engineer

Oracle • United States

On-site
USD 79,000 - 210,000
Lead Principal Core Infrastructure Engineer
Lead Principal Core Infrastructure Engineer

Oracle • Santa Clara (CA)

On-site
USD 146,000 - 307,000
Health insurance
401(k) plan with company match
Paid time off
+1
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Oracle • Reston (VA)

On-site
USD 81,200 - 187,000
Medical, dental, and vision insurance
401(k) Savings and Investment Plan
Paid time off
+2
Principal Site Reliability Engineer
Principal Site Reliability Engineer

Oracle • Austin (TX)

On-site
USD 84,000 - 210,000