Sr. Lead Site Reliability Engineer – Technical & People Leadership
Sr. Lead Site Reliability Engineer – Technical & People Leadership
Direct message the job poster from Shell Recharge Solutions
Talent Acquisition / Reporting Specialist - Curious About Data Analytics | Python(2-star Badge/Hackerrank) | Power BI | SQL(Gold 5-star…
Shell Recharge Solutions is looking for a Sr. Lead Site Reliability Engineer + People/ Team management to join our team. We would like to find a highly engaged engineer who is obsessed with monitoring, observability, code quality and self-healing infrastructures with Team management You should be able to identify, troubleshoot, and resolve issues quickly and develop strategies to ensure our environment never experiences the same problem twice. Responsibilities include capacity planning, performance tuning, and automation/tools development. Expect to spend 50% time focusing software engineering activities and should be eager to approach all code changes using a test-driven development model. As a Site Reliability Engineer, you will have great influence on the way we design and deploy our services and infrastructure across the enterprise. This position will be part of a great team that is developing exciting products and solutions and playing a key part in driving forward the electrification of transportation.
What you’ll do:
- Apply a “everything-as-code” philosophy across configuration, infrastructure, orchestration methodologies to ensure our production systems are fault tolerant and resilient.
- Have a mentoring mindset, wil-to participate in live training, and create robust documentation. Work with product development teams to define and implement enhanced monitoring and logging solutions to improve observability and enable Service-Level Objectives to ensure a world class customer experience.
- Participate in ground-up infrastructure design and planning for all future products and services.
- Be part of an on-call rotation and act as incident commander to assist finding a resolution during incidents
- Host blameless postmortems to share learnings, discover gaps, embrace transparency, and improve reliability across our services.
- Design and implement improvements to existing systems after reviewing past incidents and employ your systems knowledge to triage problems and tune resource usage.
What We’re Looking For:
Must have Team Management Expereince that includes :
- Managing and mentoring team members to support their professional growth and development
- Conducting regular performance appraisals and providing constructive feedback
- Setting and tracking KPIs to ensure team alignment with organizational goals
- Driving team engagement, collaboration, and accountability
- Supporting hiring, onboarding, and career progression initiatives
- Fostering a culture of reliability, ownership, and continuous improvement
- Bachelor’s Degree in Computer Science/ Engineering or equivalent work experience required.
- 10+ years’ experience working as full stack engineer focused on product, feature and/or systems development or equivalent experience.
- 5+ years of experience with AWS.
- Must be comfortable reading and writing in any of the following: Java, Go, C#/C++, Python
- Experience with IAC/CM tools (Terraform, Cloud Formation, Ansible, Chef, Puppet, Salt)
- Well versed with one or more cloud service provider offerings AWS (preferred), Azure, GCP
- Understanding of containers and related technologies such as Docker, Podman, Swarm, Kubernetes in a production environment
- You understand networks, protocols, servers, storage systems, and the Linux operating system. Familiar with common application and system-level health monitoring system (NewRelic, Datadog, etc.).
What We Offer:
- A work environment that allows you to work with and learn from some of the best and brightest in this emerging industry
- The ability to make a difference in a world that needs our technology to help reduce carbon emissions and enable a more sustainable energy future through the use of electric vehicle charging software, services and infrastructure
- The freedom to learn, suggest, and implement innovative new ideas applied to our systems, processes, programs and technologies
- Daily ownership of your role in a challenging, high-growth environment.
- A casual work environment and culture that support work life ‘fit’, enabling you to fit life into your work and work into your life, i.e. flexible scheduling, virtualization options, and a generous holiday package
- Competitive pay and benefits programs designed to enable you to thrive inside and outside of work
- Participation in Shell Recharge Solutions’ performance and rewards bonus program
- Best in class medical benefits for employees
Seniority level
Seniority level
Mid-Senior level
Employment type
Job function
Job function
General BusinessIndustries
Software Development
Referrals increase your chances of interviewing at Shell Recharge Solutions by 2x
Sign in to set job alerts for “Site Reliability Engineer” roles.
Site Reliability Manager, Platforms and Devices, SRE
Site Reliability Engineer III (SRE) - Cloud Applications
Site Reliability Engineer (2 to 4 years)
Sr. Site Reliability Engineer- Spera (ISMP)
We’re unlocking community knowledge in a new way. Experts add insights directly into each article, started with the help of AI.