Site Reliability Engineer

Insight Global

Aiea (HI)

On-site

USD 120,000 - 150,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Insight Global is seeking a Site Reliability Engineer to sit onsite in Oahu, Hawaii. The role focuses on monitoring system health, responding to team requests, and maintaining dashboards within a restricted SCIF environment. Strong Linux skills, Bash scripting, and familiarity with SOPs and SLA adherence are essential.

The candidate should be able to work with minimal supervision, communicate effectively, and contribute ideas to improve SOPs while supporting a dynamic work schedule.

Qualifications

  • Be available to execute requests from team members, such as specific metrics or detailed information.
  • Perform basic Linux troubleshooting, including navigating the Linux environment and extracting system information.
  • Utilize Bash scripting and common Linux commands (cd, mv, df, du, nc, scp, ping, vim, nano, cat, grep).
  • Communicate effectively with team members and escalates issues when necessary.
  • Work independently with minimal supervision.
  • Willingness to bring new ideas to improve SOPs and processes.
  • Basic Linux troubleshooting experience.
  • Proficiency in Bash scripting and common Linux commands.
  • Familiarity with working within a SCIF environment and following SOPs.
  • Strong understanding of metrics logging and SLA adherence.
  • Experience with system and application-level health dashboards.
  • This role requires working within a SCIF environment, adhering to strict security protocols.
  • The candidate must be comfortable with a dynamic work schedule and be available for requests from team members. - Ideally, experience with Docker/Kubernetes or similar container orchestration tools (kubectl, docker, podman).

Responsibilities

  • Execute requests from team members for metrics or detailed information.
  • Monitor dashboards and respond to thresholds per SOPs.
  • Maintain Linux environments and perform basic troubleshooting.
  • Collaborate with team, escalate issues when needed.

Skills

Bash scripting
Linux troubleshooting
Linux commands
SCIF environment
SLA adherence
Monitoring dashboards
Independent work
Communication
Dynamic work schedule
Docker/Kubernetes (kubectl)

Tools

Docker
Kubernetes
kubectl
podman

Job description

Job Description

An aerospace and defense client is seeking a dedicated and proactive Site Reliability Engineer to sit onsite in Oahu. This role involves managing and monitoring metrics on a regular schedule, ensuring system health, and responding to various requests from team members. The ideal candidate will have a strong understanding of working within a SCIF environment, be familiar with SOPs, have a strong understanding of SLA adherence and possess basic Linux troubleshooting skills. This person will be responsible for monitoring various dashboards within Ironman DFC and required to take action based on SOP-defined thresholds using tools like Link Monitoring, Management Tool (LMMT) and client-focused dashboarding.

We are a company committed to creating diverse and inclusive environments where people can bring their full, authentic selves to work every day. We are an equal opportunity/affirmative action employer that believes everyone matters. Qualified candidates will receive consideration for employment regardless of their race, color, ethnicity, religion, sex (including pregnancy), sexual orientation, gender identity and expression, marital status, national origin, ancestry, genetic factors, age, disability, protected veteran status, military or uniformed service member status, or any other status or characteristic protected by applicable laws, regulations, and ordinances. If you need assistance and/or a reasonable accommodation due to a disability during the application or recruiting process, please send a request to HR@insightglobal.com.To learn more about how we collect, keep, and process your private information, please review Insight Global's Workforce Privacy Policy: https://insightglobal.com/workforce-privacy-policy/.

Skills and Requirements
  • Be available to execute requests from team members, such as specific metrics or detailed information.

  • Perform basic Linux troubleshooting, including navigating the Linux environment and extracting system information.

  • Utilize Bash scripting and common Linux commands (cd, mv, df, du, nc, scp, ping, vim, nano, cat, grep).

  • Communicate effectively with team members and escalates issues when necessary.

  • Work independently with minimal supervision.

  • Willingness to bring new ideas to improve SOPs and processes.

  • Basic Linux troubleshooting experience.

  • Proficiency in Bash scripting and common Linux commands.

  • Familiarity with working within a SCIF environment and following SOPs.

  • Strong understanding of metrics logging and SLA adherence.

  • Experience with system and application-level health dashboards.

  • This role requires working within a SCIF environment, adhering to strict security protocols.

  • The candidate must be comfortable with a dynamic work schedule and be available for requests from team members. - Ideally, experience with Docker/Kubernetes or similar container orchestration tools (kubectl, docker, podman).

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Onsite SRE for Mission-Critical Systems (SCIF)
Onsite SRE for Mission-Critical Systems (SCIF)

Insight Global • Aiea (HI)

On-site
USD 120,000 - 150,000
SITE RELIABILITY ENGINEER — JUNIOR / JOURNEYMAN
SITE RELIABILITY ENGINEER — JUNIOR / JOURNEYMAN

OSAAVA Services LLC • Honolulu (HI)

On-site
USD 90,000 - 115,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

ClearanceJobs • Springfield (VA)

On-site
USD 130,000 - 190,000
Site Reliability Engineer (TS/SCI)
Site Reliability Engineer (TS/SCI)

Beyond SOF • Reston (VA)

On-site
USD 120,000 - 160,000
Senior DevSecOps & SRE — TS/SCI, On-Site
Senior DevSecOps & SRE — TS/SCI, On-Site

ClearanceJobs • Springfield (VA)

On-site
USD 130,000 - 190,000
Systems Engineer (Engineer Systems 2) - 29456
Systems Engineer (Engineer Systems 2) - 29456

Mission Technologies, a division of HII • Honolulu (HI)

On-site
USD 100,000 - 140,000
Site Reliability Engineer
Site Reliability Engineer

Fortress Information Security, LLC • Patuxent Highland (MD)

Hybrid
USD 160,000 - 180,000
Medical, dental, and vision plans
401(k) match
Flexible Paid Time Off
+1
Site Reliability Engineer
Site Reliability Engineer

SOC LLC • Herndon (VA)

On-site
USD 110,000 - 160,000
Site Reliability Engineer
Site Reliability Engineer

Leidos • Waipahu (HI)

On-site
USD 108,000 - 195,000
DevOps Site Reliability Engineer (TS/SCI)
DevOps Site Reliability Engineer (TS/SCI)

Beyond SOF • Reston (VA)

Hybrid
USD 130,000 - 190,000