Get more replies from employers
Send a job-specific resume in minutes.
RapidAI is seeking a senior Site Reliability Engineer (SRE) to own the availability and performance of our production EKS clusters and observability stack. You will define SLOs/SLIs, drive post-mortems, and implement infrastructure-as-code with Terraform, Helm, and GitOps.
You will collaborate with engineering to bake reliability in early, perform capacity planning and chaos testing, and optimize costs and autoscaling across AWS workloads. Strong scripting in Go/Python/Bash is expected.
RapidAI is the trusted leader in deep clinical AI, helping hospitals deliver faster, more informed care through intelligent imaging and integrated workflows. The Rapid Enterprise™ Platform supports disease states across the care spectrum, but it’s our clinical depth that drives the most meaningful impact — improving decision-making, patient outcomes, and health-system performance. Used by more than 2,500 hospitals in over 100 countries and backed by 700+ clinical studies, including research that helped expand national stroke-treatment guidelines, RapidAI is the most clinically validated AI platform in healthcare.
RapidAI is committed to creating an inclusive and diverse workplace. We provide equal employment opportunities to all employees and applicants and prohibit discrimination and harassment of any type in regard to race, color, religion, age, sex, national origin, disability status, genetics, protected veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by federal, state or local laws.