Staff Site Reliability Engineer

AlphaSense

United States

Remote

USD 150,000 - 225,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

AlphaSense is seeking a Staff Site Reliability Engineer to shape reliability, scalability, and performance across a global SaaS platform. You will architect core reliability platforms, mentor engineers, and drive SRE practices with incident response and blameless postmortems.

You will lead AI-driven reliability initiatives, improve observability, and steer design reviews to meet 99.99% uptime targets. Equity may be offered; compensation range is disclosed on postings.

Qualifications

  • 8+ years of Site Reliability Engineering or similar role.
  • 3+ years operating in a Senior+ SRE position.
  • Experience running production SaaS systems at scale.
  • Proficiency in Python or Go (or similar).
  • Hands-on with cloud platforms (AWS, GCP, or Azure) and Kubernetes.
  • Deep understanding of networking fundamentals (TCP/IP, DNS, HTTP/S, load balancing).
  • Experience with monitoring and alerting (Prometheus, Grafana, Datadog, ELK).
  • Familiarity with advanced observability (OTEL, continuous profiling).
  • Proven incident management experience including leading high-severity incidents and postmortems.
  • Strong troubleshooting & communication skills.

Responsibilities

  • Architect Reliability Paved Paths: Build frameworks and self-service tooling for reliability in a You Build It, You Run It culture.
  • Lead AI-Driven Reliability: Drive our AIOps strategy — automating diagnostics and remediation.
  • Champion Reliability Culture: Embed SRE practices across engineering via design reviews and production readiness.
  • Incident Leadership: Act as Incident Commander during critical events, leading blameless postmortems.
  • Advance Observability: Deliver end-to-end monitoring, tracing, and profiling to optimize performance.
  • Mentor & Multiply: Elevate engineers across SRE and product teams through mentorship.

Skills

SRE experience
SaaS at scale
Python/Go
Cloud platforms
Kubernetes
Networking basics
Monitoring tools
Observability
Incident leadership
Communication

Tools

Prometheus
Grafana
Datadog
ELK
OTEL
Kubernetes

Job description

About AlphaSense:

The world’s most sophisticated companies rely on AlphaSense to remove uncertainty from decision-making. With market intelligence and search built on proven AI, AlphaSense delivers insights that matter from content you can trust. Our universe of public and private content includes equity research, company filings, event transcripts, expert calls, news, trade journals, and clients’ own research content. The acquisition of Tegus by AlphaSense in 2024 advances our shared mission to empower professionals to make smarter decisions through AI-driven market intelligence. Together, AlphaSense and Tegus will accelerate growth, innovation, and content expansion, with complementary product and content capabilities that enable users to unearth even more comprehensive insights from thousands of content sets. Our platform is trusted by over 6,000 enterprise customers, including a majority of the S&P 500. Founded in 2011, AlphaSense is headquartered in New York City with more than 2,000 employees across the globe and offices in the U.S., U.K., Finland, India, Singapore, Canada, and Ireland. Come join us!

About the Role:

Our Site Reliability Engineering team is growing, and we are looking for a highly experienced Staff Site Reliability Engineer to help shape the future of reliability, scalability, and performance at AlphaSense. This is a hands‑on, high‑impact role where you will architect core reliability platforms, lead by example in incident response, and drive cultural adoption of SRE best practices across our global engineering organization. Our mission is to engineer our platform to the reliability standards of mission‑critical systems, targeting 99.99% uptime, while continuously enhancing our systems and processes. This role is key to that mission and goes beyond traditional system maintenance; it’s about pioneering the platforms, practices, and culture that enable engineering to scale effectively. You will act as a force multiplier, mentoring fellow engineers, influencing architectural decisions, and setting the technical bar for reliability across the company.

Who You Are:
  • 8+ years of experience in Site Reliability Engineering, DevOps, or a similar role, with at least 3+ of those years operating in a Senior+ SRE position
  • Strong background in running production SaaS systems at scale
  • Proficiency in at least one programming/scripting language (Python, Go, or similar)
  • Hands‑on expertise with cloud platforms (AWS, GCP, or Azure) and Kubernetes
  • Deep understanding of networking fundamentals (TCP/IP, DNS, HTTP/S, load balancing)
  • Experience with monitoring & alerting (Prometheus, Grafana, Datadog, ELK)
  • Familiarity with advanced observability (OTEL, continuous profiling)
  • Proven incident management experience, including leading high‑severity incidents and postmortems
  • Strong troubleshooting skills across the full stack
  • Excellent communication and collaboration skills
What You’ll Do:
  • Architect Reliability Paved Paths: Build frameworks and self‑service tooling that let teams own the reliability of their services in a “You Build It, You Run It” culture
  • Lead AI‑Driven Reliability: Drive our AIOps strategy — automating diagnostics, remediation, and proactive failure prevention
  • Champion Reliability Culture: Embed SRE practices across engineering via design reviews, production readiness, and operational standards
  • Incident Leadership: Act as Incident Commander during critical events, modeling operational excellence, and ensuring blameless postmortems lead to lasting improvements
  • Advance Observability: Deliver end‑to‑end monitoring, tracing, and profiling (Prometheus, Grafana, OTEL, Continuous Profiling) to optimize performance proactively
  • Mentor & Multiply: Elevate engineers across SRE and product teams through mentorship, technical guidance, and knowledge sharing

For base compensation, we set standard ranges for all roles based on function and level benchmarked against similar stage growth companies and internal comparables. In order to be compliant with local legislation, as well as to provide greater transparency to candidates, we share salary ranges on all job postings regardless of desired hiring location. Final offer amounts are determined by multiple factors including candidate experience/expertise and may vary from the amounts listed below.

You may also be offered equity, and a generous benefits program.

Compensation Range $150,000 - $225,000 USD

AlphaSense is an equal‑opportunity employer. We are committed to a work environment that supports, inspires, and respects all individuals. All employees share in the responsibility for fulfilling AlphaSense’s commitment to equal employment opportunity. AlphaSense does not discriminate against any employee or applicant on the basis of race, color, sex (including pregnancy), national origin, age, religion, marital status, sexual orientation, gender identity, gender expression, military or veteran status, disability, or any other non‑merit factor. This policy applies to every aspect of employment at AlphaSense, including recruitment, hiring, training, advancement, and termination. In addition, it is the policy of AlphaSense to provide reasonable accommodation to qualified employees who have protected disabilities to the extent required by applicable laws, regulations, and ordinances where a particular employee works.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Staff AI Platform Engineer San Francisco, California, United States
Staff AI Platform Engineer San Francisco, California, United States

AlphaSense, Inc. • San Francisco (CA)

On-site
USD 178,000 - 267,000
Equity options
Generous benefits program
Staff Platform Engineer, Core Cloud Platform New Remote - United States
Staff Platform Engineer, Core Cloud Platform New Remote - United States

AlphaSense, Inc. • Northern (KY)

Remote
USD 183,000 - 305,000
Principal Software Engineer – Platform Engineering
Principal Software Engineer – Platform Engineering

AlphaSense • United States

On-site
USD 246,000 - 339,000
Staff Software Engineer – Full-Stack (Frontend-Heavy)
Staff Software Engineer – Full-Stack (Frontend-Heavy)

Engg • United States

Remote
USD 180,000 - 240,000
Cloud Support Engineer
Cloud Support Engineer

Socket.dev • United States

On-site
USD 100,000 - 122,000
Performance-based bonus
Equity participation
Generous benefits program
Principal Software Engineer – Platform Engineering New Remote - United States
Principal Software Engineer – Platform Engineering New Remote - United States

AlphaSense, Inc. • Northern (KY)

Remote
USD 246,000 - 339,000
Performance bonus
Equity
Generous benefits
Cloud Support Engineer New Remote - United States
Cloud Support Engineer New Remote - United States

AlphaSense, Inc. • Northern (KY)

Remote
USD 108,000 - 114,000
Senior Engineering Manager, AI
Senior Engineering Manager, AI

AlphaSense • United States

On-site
USD 361,200 - 496,800
Equity
Benefits
Staff AI Platform Engineer New York, New York, United States
Staff AI Platform Engineer New York, New York, United States

AlphaSense, Inc. • New York (NY)

On-site
USD 178,000 - 267,000
Equity options
Generous benefits program
Staff Software Engineer
Staff Software Engineer

AlphaSense • United States

Remote
USD 202,000 - 278,000
Generous benefits program