Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.
Xebia It Architects is seeking an experienced SRE to own incidents end-to-end and drive reliability engineering across platforms. You will work with on-call rotations, perform RCAs, and convert toil into automations while tuning SLOs and error budgets.
The role emphasizes hands-on Linux and cloud fundamentals, scripting in Python or Bash, and collaboration with the automation factory to improve detection, runbooks, and service reliability.
The L2/SRE team investigates and resolves incidents end-to-end as a resolver, not a router, and applies reliability-engineering practices.
1-6 years in SRE or operations engineering; strong Linux and cloud fundamentals; scripting ability in at least one of Python or Bash (or equivalent); familiarity with observability and incident tooling; solid troubleshooting and problem-management discipline.