Site Reliability Engineer 3

Wyetech, LLC

Laurel (MD)

On-site

USD 122,000 - 165,000

Full time

2 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

PTO up to 200 hours yearly
Medical, Vision, Dental plans
Employee Referral Bonus up to $10,000
Mobility among Wyetech contracts
Company events & gear

Job summary

Wyetech, LLC is seeking an experienced Site Reliability Engineer Level 3 (SRE3) to work in a cloud environment supporting data‑intensive analytics on a managed infrastructure. The role involves ensuring high availability, scalability, and security of mission systems, with on‑call Tier 1–Tier 3 support in a fast‑paced team.

The ideal candidate has 14+ years in software engineering, 10+ years in systems engineering, and strong Linux troubleshooting.

Qualifications

  • Fourteen years of software development/engineering experience.
  • Ten years of Systems Engineering/Architecture experience.
  • Ten years with distributed, massively parallel technologies (HBase, Hadoop, Accumulo, Bigtable, Cassandra, Scality).
  • At least ten years writing automation scripts (Perl, Python, Ruby).
  • Ten years managing and monitoring large cloud systems (>1,000 nodes).
  • Strong Linux troubleshooting and incident management experience.

Responsibilities

  • Maintain day-to-day operational stability of mission systems and cloud infrastructure.
  • Provide customer support and troubleshoot complex issues.
  • Monitor system health, availability, performance, and scalability.
  • Provide on-call Tier 1–Tier 3 support.
  • Develop automation scripts in Python, Perl, Ruby, and Bash.
  • Collaborate with customers, developers, and operations teams on resolutions.
  • Perform incident response, root-cause analysis, and postmortems.

Skills

Java
Kubernetes
Hadoop
Accumulo
Docker
Linux
Cloud platforms
SRE fundamentals
On-call support
Python scripting

Education

Bachelor's degree in Computer Science or related field
10+ years Systems Engineering/Architecture

Tools

JIRA
Salt/Ansible
OpenStack
AWS
HDFS

Job description

Site Reliability Engineer (SRE) Skill Level 3

Wyetech is seeking an experienced Site Reliability Engineer Level 3 (SRE3) to work within a cloud environment supporting data-intensive analytics on a managed infrastructure. The position will work with technologies including Java, Kubernetes, Hadoop, Accumulo, Docker, Linux, and cloud platforms to support highly available, scalable, and secure mission systems.

This position sits on the Operations Team and is responsible for maintaining day-to-day operational stability, providing customer support, troubleshooting complex technical issues, and ensuring the reliability and performance of mission‑critical systems. This is an on‑call position providing Tier 1 through Tier 3 support and requires a strong background troubleshooting operational issues in Linux environments.

Required Skills/Experience:
  • Fourteen (14) years of experience in software development/engineering, including requirements analysis, software development, installation, integration, evaluation, enhancement, maintenance, testing, and problem diagnosis/resolution.
  • Ten (10) years of experience in Systems Engineering/Architecture.
  • Ten (10) years of experience working with technologies supporting highly distributed, massively parallel computation, such as HBase, Hadoop, Accumulo, Bigtable, Cassandra, Scality, or similar technologies.
  • At least ten (10) years of experience writing software scripts using languages such as Perl, Python, or Ruby for software automation.
  • At least four (4) years of experience managing and monitoring large cloud systems consisting of more than 1,000 nodes.
  • Strong experience troubleshooting operational and technical issues within a Linux environment.
  • Experience providing technical direction for the development, engineering, interfacing, integration, and testing of complete hardware/software systems.
  • Experience monitoring the technical health and operational stability of systems.
  • Experience improving organizational and operational processes.
  • Experience performing postmortem/failure analysis and incident management.
  • Ability to work effectively in a fast‑paced team environment while remaining self‑motivated, proactive, and detail-oriented.
  • Ability to provide Tier 1, Tier 2, and Tier 3 operational support in an on‑call environment.
Key Responsibilities:
  • Support software development and engineering activities, including requirements analysis, development, installation, integration, evaluation, enhancement, maintenance, testing, and troubleshooting.
  • Support a managed cloud infrastructure used to execute data-intensive analytics.
  • Maintain day-to-day operational stability of mission systems and cloud infrastructure.
  • Provide customer support and technical troubleshooting expertise.
  • Monitor system health, availability, performance, scalability, and reliability.
  • Diagnose and resolve complex software and infrastructure problems within cloud and Linux environments.
  • Provide on‑call Tier 1 through Tier 3 support.
  • Develop and maintain automation scripts using languages such as Python, Perl, Ruby, and Bash.
  • Support highly distributed and massively parallel computing environments.
  • Support the development, engineering, integration, interfacing, and testing of complex hardware/software systems.
  • Perform incident response, incident management, root‑cause analysis, and postmortem analysis.
  • Identify opportunities to improve system reliability, operational processes, and overall system performance.
  • Collaborate with customers, developers, Systems Engineers, and Operations personnel to resolve technical issues.
Required Certification:

Candidate must possess at least one (1) of the following certifications:

  • AWS Certified Developer – Associate
  • AWS Certified Solutions Architect – Associate
  • AWS Certified Solutions Architect – Professional
  • AWS Certified SysOps Administrator – Associate
  • Certified Kubernetes Administrator (CKA)
  • Elastic Certified Engineer
  • Elastic Certified Observability Engineer
Additional/Desired Technical Experience:
  • Docker and Kubernetes
  • Hadoop and Accumulo
  • Prometheus and Grafana
  • Hadoop Distributed File System (HDFS)
  • JIRA
  • Salt/Ansible
  • Virtualization technologies
  • OpenStack and AWS
  • Python and Bash scripting
  • Ten (10) years of experience working within a cleared environment.
  • Ten (10) years of demonstrated experience developing software for UNIX or Linux operating systems.
  • Knowledge and experience developing distributed storage, routing, and querying algorithms.
  • Ten (10) years of experience developing software systems using object‑oriented programming languages such as Java and Python.
  • Experience developing solutions that integrate and extend COTS products.
  • Demonstrated knowledge of analytical needs and requirements, query syntax, data flows, and traffic manipulation.
  • Ten (10) years of experience developing system performance, availability, scalability, manageability, and security requirements for mid‑to‑large‑scale programs.
  • Experience designing, developing, testing, evaluating, and integrating information systems into a service‑oriented environment.
  • Experience optimizing storage, retrieval, backup, and retention strategies across globally distributed, high‑throughput text and multimedia storage environments within clustered or cloud infrastructures.
  • Experience operating within a multi‑threaded environment.
  • Experience debugging and troubleshooting complex software in cloud environments.
  • Familiarity with Configuration Management and monitoring tools.
  • Familiarity with Agile software methodologies and practices.
  • Significant experience provisioning and sustaining network infrastructures.
  • Experience developing, operating, and managing networks within secure PKI, IPSEC, or VPN‑enabled environments.
Education & Experience:
  • Fourteen (14) years of applicable experience is required.
  • A bachelor’s degree in Computer Science or a related technical field is highly desired and will be considered equivalent to two (2) years of experience.
  • A master’s degree in a technical field will be considered equivalent to four (4) years of experience.
  • Degrees in Mathematics, Information Systems, Engineering, or a similar discipline are considered technical fields.
Clearance & Compliance Requirements:
  • Active TS/SCI security clearance with a current Fullscope polygraph is required.
  • DoD 8570 IAT Level I or higher is required.

Due to federal contract requirements, United States Citizenship and position appropriate security clearance is required.

At Wyetech, you’ll be at the center of an award‑winning corporate culture, breaking technological barriers and solving real‑world problems for our federal government customers. We are committed to hiring the best of the best, and in return, we offer a world‑class, truly unique employee experience that is rare within our industry.

Wyetech believes in generously supporting employees as they prepare for retirement. The company automatically contributes 20% of each employee's gross compensation to a Simplified Employee Pension (SEP) IRA, with no requirement for employee matching. All contributions are fully vested from day one, ensuring immediate ownership of retirement funds.

Additional benefits include:
  • Wyetech provides a generous PTO plan of up to 200 hours annually, aligned with applicable state leave regulations. Employees have the flexibility to adjust their PTO allocation at the start of each calendar year, ensuring it meets their evolving needs.
  • Full‑time employees have the option to participate in a variety of voluntary benefit plans including:
    • Choice of Medical Plan Options, some with Health Savings Account (HSA)
    • Vision and Dental
    • Life and AD&D Benefits
    • Short and Long‑Term Disability
    • Hospital Indemnity, Accident, and Critical Illness Insurances
    • Optional Identity Theft and Legal Protection Services
    • Employee Referral Bonus Eligibility up to $10,000
    • Mobility Among Wyetech‑supported Contracts
    • Various contract and work locations throughout Maryland, Virginia, Colorado, Texas, Utah, Alaska, Hawaii and OCONUS
    • Various team‑building events throughout the year such as: monthly lunches, summer company picnic, and an annual holiday party.
    • Employees receive two complimentary branded clothing orders annually.
  • Salary Range: $88.50-$119.84 per hour

Hourly pay rates listed for this position serve as a general guideline and are not a guarantee of compensation. Compensation will vary dependent upon factors including but not limited to: Government contract rates; education; relevant prior work experience, knowledge, skills, and competencies; certifications, and geographic location. Hourly pay rates reflect the pre‑benefit gross wage amounts.

Wyetech, LLC is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, or status as a protected veteran.

Affirmative Action Statement:

Wyetech, LLC is committed to the principles of affirmative action in all hiring and employment for minorities, women, individuals with disabilities, and protected veterans.

Accommodations:

Wyetech, LLC is committed to providing an inclusive and accessible hiring process. If you need any accommodations during the application or interview process, please contact Brittney Wood. at 844-WYETECH x727 or staffing@wyetech.com. We are happy to provide reasonable accommodations to ensure equal access to all candidates.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Software Engineer2-Machine Learning (Potential Telework)
Software Engineer2-Machine Learning (Potential Telework)

Wyetech, LLC • Laurel (MD)

Hybrid
USD 88,000 - 165,000
SEP IRA
PTO up to 200 hours annually
Medical plan options
+8
SWE2-Machine Learning (Potential Telework)
SWE2-Machine Learning (Potential Telework)

Wyetech, LLC • Laurel (MD)

Remote
USD 88,000 - 165,000
SEP IRA
200 hours PTO
Medical plan options (HSA)
+6
Software Engineer 2
Software Engineer 2

Wyetech, LLC • Annapolis (MD)

On-site
USD 183,500,000 - 344,623,000
SEP IRA
PTO 200 hours/year
Medical plan
+3
Software Engineer 3
Software Engineer 3

Wyetech • Hanover (MD)

On-site
USD 246,998,000 - 390,470,000
SEP IRA 20% contribution
Generous PTO up to 200 hours per year
Medical Plans with options
+9
System Administrator 2
System Administrator 2

Wyetech • Maryland

On-site
USD 70,000 - 130,000
Employee Referral Bonus up to $10,000
Various team-building events
Complementary branded clothing
Application Engineer 3 (24x7)
Application Engineer 3 (24x7)

Wyetech • Fort Meade (MD)

On-site
USD 88,166 - 119,851
SEP IRA contributions
PTO up to 200 hours per year
Medical, dental, vision plans
System Administrator 3
System Administrator 3

Wyetech • Bluffdale (UT)

On-site
USD 67,502 - 162,556
200 hours PTO
SEP IRA with 20% employer contribution
Multiple locations across states
Digital Network Exploitation Analyst 3
Digital Network Exploitation Analyst 3

Wyetech, LLC • Maryland

On-site
USD 80,000 - 127,000
SEP IRA 20%
Medical with HSA
Vision & Dental
+3
Software Engineer 1
Software Engineer 1

Wyetech, LLC • Annapolis (MD), Northern (KY)

On-site
USD 72,000 - 147,000
SEP IRA contribution 20% of gross pay
PTO up to 200 hours annually
Medical Vision Dental Life/AD&D
+2
Systems Engineer 4
Systems Engineer 4

Wyetech, LLC • Linthicum (MD)

On-site
USD 95,000 - 174,000
PTO up to 200 hours annually
SEP IRA with 20% company contribution
Medical plan options
+2