Turn this role into an interview — a resume and cover letter built around what this employer wants.
CloudRaft is seeking passionate Site Reliability Engineers to design, build, operate, and scale mission‑critical infrastructure for our partners. You will own reliability, performance, security, and automation, bridging software and operations to enable fast‑growing organizations to scale confidently.
As a member of the team, you will manage Kubernetes clusters, implement CI/CD pipelines, and maintain observability stacks while participating in 24x7 on‑call coverage.
Note: If you have already applied to CloudRaft in the last 90 days, we already have your CV/resume on file. Multiple applications from the same candidate will not be considered.
CloudRaft is a premier cloud-native consulting and engineering company that helps ambitious startups and digital‑first organizations build, scale, and operate mission-critical platforms. We partner with innovators at the forefront of artificial intelligence, developer productivity, observability, digital commerce, and enterprise software—enabling them to accelerate growth with resilient, scalable, and production‑ready cloud infrastructure.
Our experience spans organizations developing AI safety and governance platforms, AI Cloud, AI agent ecosystems, developer tooling, observability solutions, digital health products, customer engagement platforms, and technology‑driven franchise networks. By combining deep expertise in Platform Engineering, Kubernetes, DevOps, Observability, and Cloud Native technologies, CloudRaft helps high‑growth companies move faster, operate more reliably, and focus on building category‑defining products.
We are looking for passionate Site Reliability Engineers (SREs) to join our growing team. In this role, you will take end‑to‑end ownership of designing, building, operating, and scaling mission‑critical infrastructure for our partners. You will be responsible for ensuring reliability, performance, security, and operational excellence while driving automation, improving system efficiency, and implementing innovative solutions. Working at the intersection of software engineering and operations, you will help create resilient platforms that enable fast‑growing organizations to scale with confidence.