Get more replies from employers
Send a job-specific resume in minutes.
AlphaSense is seeking a Staff Site Reliability Engineer to shape reliability, scalability, and performance. This hands-on role involves architecting core reliability platforms and guiding incident response across our global engineering teams.
You will mentor engineers, drive AI-driven reliability, and champion blameless postmortems to lift the technical bar. Strong cloud, Kubernetes, and monitoring skills are essential to sustain 99.99% uptime.
Our Site Reliability Engineering team is growing, and we are looking for a highly experienced Staff Site Reliability Engineer to help shape the future of reliability, scalability, and performance at AlphaSense. This is a hands-on, high-impact role where you will architect core reliability platforms, lead by example in incident response, and drive cultural adoption of SRE best practices across our global engineering organization. Our mission is to engineer our platform to the reliability standards of mission-critical systems, targeting 99.99% uptime, while continuously enhancing our systems and processes. This role is key to that mission and goes beyond traditional system maintenance; it’s about pioneering the platforms, practices, and culture that enable engineering to scale effectively. You will act as a force multiplier, mentoring fellow engineers, influencing architectural decisions, and setting the technical bar for reliability across the company.
AlphaSense is an equal-opportunity employer. We are committed to a work environment that supports, inspires, and respects all individuals. All employees share in the responsibility for fulfilling AlphaSense's commitment to equal employment opportunity. AlphaSense does not discriminate against any employee or applicant on the basis of race, color, sex (including pregnancy), national origin, age, religion, marital status, sexual orientation, gender identity, gender expression, military or veteran status, disability, or any other non-merit factor. This policy applies to every aspect of employment at AlphaSense, including recruitment, hiring, training, advancement, and termination.
8+ years of experience in Site Reliability Engineering, DevOps, or a similar role, 3+ years in a Senior+ SRE position, Production SaaS systems experience at scale, Programming/scripting in Python, Go, or similar, Cloud platforms (AWS, GCP, or Azure) and Kubernetes experience, Networking fundamentals (TCP/IP, DNS, HTTP/S), Monitoring/alerting (Prometheus, Grafana, Datadog, ELK), Advanced observability (OTEL, continuous profiling), Incident management experience and postmortems, Strong troubleshooting and communication skills