Site Reliability Engineer

BRG

United States

Remote

USD 130.000 - 160.000

Vollzeit

14 Tage+
Bewerbungsgenerator

Mach aus dieser Rolle ein Bewerbungsgespräch — ein Lebenslauf und ein Anschreiben, die darauf ausgerichtet sind, was dieser Arbeitgeber sucht.

Schaffe es an den ATS-Filtern vorbei

Zusammenfassung

BRG's Second Sight Solutions, a health-tech subsidiary, seeks a Site Reliability Engineer to design, build, and maintain scalable cloud systems in Azure. You will work with software developers and operations teams to automate processes, minimize downtime, and improve release validation.

The ideal candidate has a Bachelor’s degree in CS, 5+ years in SRE, strong coding in Golang/Ruby/Python, and Kubernetes expertise. Datadog, OpsGenie, PagerDuty experience are expected. Salary $130,000-$160,000.

Qualifikationen

  • Bachelor’s degree in computer science or a related field.
  • 5+ years of experience as a site reliability engineer or similar role.
  • Strong programming skills in Golang, Ruby, Python, or similar.
  • Expertise in Kubernetes and cloud-native infrastructure.

Aufgaben

  • Design, build, and maintain scalable and reliable cloud systems.
  • Improve system reliability with automation and self-healing capabilities.
  • Lead incident management and response to outages.
  • Develop service-level indicators and publish incident policies.

Kenntnisse

Golang
Ruby
Python
Kubernetes
Problem solving
Communication

Ausbildung

Bachelor’s degree in computer science or similar field

Tools

Datadog
OpsGenie
PagerDuty
Infrastructure as Code

Jobbeschreibung

We do Consulting Differently

Second Sight Solutions, a subsidiary of Berkeley Research Group (BRG), is a health technology company, and our innovative technology reimagines how drug discount data is exchanged, establishing new connections and improving transparency for drug manufacturers and their customers. Our customers and partners trust us to deliver reliable, first-to-market solutions and safeguard the data we receive. We trust our employees, and our culture gives them the freedom to create, collaborate, and grow. Our leaders are industry experts, creative, unafraid to challenge the status quo, and the pioneers of market-changing solutions.

We are seeking a Site Reliability Engineer to design, build, and maintain highly available systems and infrastructure. The SRE will work closely with software developers and operations teams to improve system reliability, automate processes, and minimize downtime.

Responsibilities
  • Design, implement, and maintain scalable and reliable systems in cloud environments such as Azure Cloud Services.

  • Experience with CI/CD Platforms (GitHub Actions, GitLab CI)

  • Provide operational support for full-stack software applications.

  • Increase system resilience with expert-level coding, bulletproof release, and change management skills.

  • Develop service-level indicators and objectives to automate release validation.

  • Improve automation and increase the system’s self-healing capability.

  • Collect operating system data and report performance metrics to stakeholders.

  • Ensure security best practices are followed in cloud infrastructure and application deployments.

  • Manage cloud and database system maintenance, debugging production issues as they arise.

  • Improve reliability, quality, and time-to-market of our suite of software solutions.

  • Partner with security and product teams to define and publish policies, processes, and playbooks to facilitate rapid and effective handling of alerts and incidents.

  • Lead incident management processes; respond to outages and service disruptions promptly.

Qualifications:
  • Bachelor’s degree in computer science or similar field.

  • Five years’ experience as a site reliability engineer or similar role.

  • Strong programming skills (Golang, Ruby, Python, or similar)

  • Proven ability to diagnose and monitor performance and reliability issues across the stack.

  • Expertise in Kubernetes.

  • Relevant industry certifications, such as through the Site Reliability Engineering (SRE) Foundation.

  • Proven experience working with cloud-native infrastructure (Azure Cloud Services, AWS, or GCP).

  • Experience working with observability and incident management tools (Datadog, OpsGenie, PagerDuty).

  • Experience scripting operating system tasks with Infrastructure as Code.

  • Impeccable communication skills.

  • Ability to problem-solve in a fast-paced, high-stakes environment.

Candidate must be able to submit verification of his/her legal right to work in the United States, without company sponsorship.

Salary: $130,000 - $160,000

About BRG

BRG combines world-leading academic credentials with world-tested business expertise and purpose-built emerging technologies. Our culture centers on agility and connectivity which sets us apart and gets you ahead.

At BRG, our professionals include specialist consultants, industry experts, renowned academics, and leading-edge data scientists. Together, they bring a diversity of real-world experience, data, and human and artificial intelligence, to economics, disputes, and investigations; corporate finance; and performance improvement services that address the most complex challenges facing organizations across the globe.

Our unique structure nurtures the interdisciplinary relationships that give us the edge, laying the groundwork for more informed insights and more original, incisive thinking. When paired with our global reach and resources, our diverse perspectives and technical capabilities make us uniquely capable to address our clients’ challenges. We get results because we know how to apply our thinking to your world.

At BRG, we don’t just show you what’s possible. We’re built to help you make it happen.

BRG is proud to be an Equal Opportunity Employer. Our hiring practices provide equal opportunity for employment without regard to race, religion, color, sex, gender, national origin, age, United States military veteran status, ancestry, sexual orientation, marital status, family structure, medical condition including genetic characteristics or information, veteran status, or mental or physical disability so long as the essential functions of the job can be performed with or without reasonable accommodation, or any other protected category under federal, state, or local law.

Hol dir deinen kostenlosen, vertraulichen Lebenslauf-Check.

oder ziehe deine Datei hierhin.

Similar jobs

Ähnliche Jobs, die dir auch gefallen könnten

Site Reliability Engineer
Site Reliability Engineer

LE038 Second Sight Solutions, LLC • Washington

Hybrid
USD 130.000 - 160.000
Software Engineering Manager
Software Engineering Manager

BRG • Washington

Vor Ort
USD 200.000 - 300.000
Software Engineering Manager
Software Engineering Manager

BRG • New York (NY)

Vor Ort
USD 200.000 - 300.000
Software Engineering Manager
Software Engineering Manager

LE001 Berkeley Research Group, LLC • Washington

Vor Ort
USD 200.000 - 300.000
Senior DevOps Cloud Engineer
Senior DevOps Cloud Engineer

BRG • USA

Remote
USD 149.000 - 168.000
Bonus
Superannuation
Benefits
Senior Site Reliability Engineer II
Senior Site Reliability Engineer II

LexisNexis Risk Solutions • San Jose (CA), Northern (KY)

Hybrid
USD 105.000 - 175.000
401(k) with match
Wellbeing programs
Life Insurance
+1
Software Engineer
Software Engineer

BRG • Washington

Remote
USD 120.000 - 160.000
Software Engineer
Software Engineer

BRG • Chicago (IL)

Hybrid
USD 120.000 - 160.000
Site Reliability Engineer - Cloud, Kubernetes & Automation
Site Reliability Engineer - Cloud, Kubernetes & Automation

BRG • USA

Remote
USD 130.000 - 160.000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

HITEC • Jersey City (NJ)

Vor Ort
USD 153.000 - 192.000
Benefits eligible
Discretionary incentive plan