Senior Site Reliability Engineer

Ll Oefentherapie

Reston (VA)

On-site

USD 50,000 - 70,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Ll Oefentherapie is seeking a candidate for a role focused on Capacity Ingestion and Management, Incident and Service Lifecycle Management, and Automation. Responsibilities include assisting in infrastructure discussions, data collection for incident management, supporting automation initiatives, and ensuring operational reliability. The ideal candidate will have a basic understanding of infrastructure management and a willingness to learn and adapt. The position is based in Reston, Virginia, and offers opportunities for continuous learning and improvement.

Qualifications

  • Basic understanding of infrastructure and reliability.
  • Ability to follow detailed instructions for testing and troubleshooting.
  • Willingness to learn new tools and technologies.

Responsibilities

  • Participate in discussions about infrastructure design.
  • Assist in data collection and incident response.
  • Support the development of automation solutions.
  • Communicate performance attributes within the team.
  • Provide operational support for technology issues.

Job description

Capacity Ingestion and Management
  • Participates and listens in on discussions for the design and architecture of infrastructure and/or service according to terms for reliability and functionality.
  • Assists team members responding to infrastructure demands and capacity increases to support current and future workloads.
  • Supports collaborations with the software development team to contribute to the development of reliable and scalable infrastructures based on detailed deployment requirements.
  • Participates in identifying opportunities for prototyping and provides support for prototyping initiatives (e.g., testing new applications or infrastructures, assisting in onboarding).
Incident and Service Lifecycle Management
  • Assists in data collection, triage, and redirection to maintain and optimize operations and infrastructure reliability.
  • Monitors services and maintains up-to-date knowledge of their performance.
  • Supports incident response and/or maintenance tasks (e.g., software installs, version upgrades, and security updates, backup and recovery) under supervision.
  • Assists in providing health and performance reporting and takes appropriate actions based on trends in data.
  • May perform provisioning according to established procedures to support infrastructure, applications, and services.
  • May perform decommissioning (e.g., shutting down servers, removing data from databases) according to established procedures to remove objects that are no longer needed.
Automation
  • Assists in identifying opportunities for automation and assessing potential benefits.
  • Supports the development of automation or scripts to provide solutions, gather metrics, monitor, analyze, mitigate, or remediate issues/defects within infrastructures.
  • Follows detailed instructions to conduct testing to ensure automation performs tasks correctly and produces expected results, with supervision.
Technical Communication and Guidance
  • Communicates basic information about the scale, capacity, security, and performance attributes of services and technology within immediate team.
  • Assists in identifying and communicating basic infrastructure, feature, and tool changes within immediate team.
Troubleshooting and Resolution
  • Provides operational support for technology, escalating routine, low-impact incidents and other issues arising within Oracle services.
  • Participates in on-call shifts to address issues.
  • Assists with resolving technical issues, performing investigations, and debugging products in order to reach SLOs (service level objectives), with supervision.
  • Follows detailed instructions to document incidents and perform root cause analyses according to standard reporting methods.
  • Participates in post-mortem procedures to prevent incident reoccurrence.
Innovation and Improvement
  • Assists in experimenting with new tools and technologies to improve infrastructure performance and reliability and helps ensure adherence to security standards, with supervision.
  • Supports the execution of improvements for performance bottlenecks and deployments.
  • Gains basic knowledge of site reliability trends and shares relevant information with immediate team members.
  • Performs analyses as assigned and assists in providing clear data on production to support business development decisions (e.g., design changes).
Core Responsibilities
Planning & Execution
  • Completes assigned tasks and monitors timelines to ensure timely completion of work in accordance with project requirements, with supervision. Follows direction to prioritize work and adjust to shifts in resources or timelines.
Collaboration & Partnership
  • Collaborates with team members to better understand expectations and contribute to shared objectives. Builds basic understanding of business, stakeholder, and/or customer needs with guidance.
Problem Solving
  • Follows standard procedures to identify and elevate issues to senior team members. Collects and reviews basic data and/or information to troubleshoot common errors.
Continuous Learning
  • Builds knowledge and learns new skills and/or tools aligned with industry trends and best practices as directed. Incorporates feedback and participates in training to improve skills.
Continuous Improvement
  • Begins to identify ways to increase the efficiency and effectiveness of processes, protocols, and workflows with guidance.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Manager, Site Reliability Engineering
Manager, Site Reliability Engineering

Oracle • Austin (TX)

On-site
USD 140,000 - 190,000
Flexible benefits
Life insurance
Retirement options
Manager, Site Reliability Engineering
Manager, Site Reliability Engineering

Oracle • Reston (VA)

On-site
USD 140,000 - 180,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Oracle • Austin (TX)

On-site
USD 81,000 - 187,000
Medical, dental, and vision insurance
401(k) savings and investment plan
Paid parental leave
+1
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Oracle • Pleasanton (CA)

On-site
USD 81,000 - 187,000
Medical, dental, and vision insurance
401(k) savings plan with company match
Paid parental leave
+1
Senior Data Center Facilities Engineer
Senior Data Center Facilities Engineer

Oracle • Chicago (IL)

On-site
USD 90,000 - 120,000
Manager, Data Center Facilities Engineering (liquid cooling)
Manager, Data Center Facilities Engineering (liquid cooling)

Oracle • San Antonio (TX)

On-site
USD 102,000 - 210,000
Medical, dental, and vision insurance
401(k) Savings Plan with company match
Flexible vacation and 11 paid holidays
Senior Data Center Facilities Engineer
Senior Data Center Facilities Engineer

Oracle • Harwood Township (ND)

On-site
USD 90,000 - 130,000
Medical, dental, and vision insurance
Short/long term disability
401(k) with company match
+1
Lead Principal Site Reliability Engineer
Lead Principal Site Reliability Engineer

Oracle • Nashville (TN)

On-site
USD 96,000 - 265,000
Health insurance
Paid time off
401(k) match
Senior Data Center Facilities Engineer I
Senior Data Center Facilities Engineer I

Oracle • United States

On-site
USD 102,000 - 210,000
Medical, dental, and vision insurance
401(k) savings plan with company match
Flexible vacation policy
+1
Senior Site Reliability Engineering Lead
Senior Site Reliability Engineering Lead

Oracle • Frankfort (KY)

On-site
USD 122,000 - 264,000
Medical, dental, and vision insurance
401(k) Savings and Investment Plan
Paid time off