Staff Site Reliability Engineer

JobCubby

Northern (KY)

Hybrid

USD 180,000 - 240,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Equity
Health benefits

Job summary

Attentive’s Platform Infrastructure team seeks a Staff Engineer to drive reliability and scalability across our event-driven platform serving 100M+ customers daily. You will design systems, mentor engineers, and influence roadmaps for high-impact initiatives.

You will collaborate with AI/ML, Data, Platform, and Product teams to deliver robust services, improve observability, and optimize costs while upholding security and performance standards.

Qualifications

  • 7+ years of experience in Production Engineering, Backend Engineering, SRE, DevOps or similar role.
  • Strong technical background enabling long-term architectural planning.
  • Proficient problem-solver with coding skills in at least one language (Golang, Python, Java, Typescript).
  • Track record delivering medium to large-scale reliability improvements.
  • Deep understanding of production reliability concepts: SLIs, SLOs, incident management.

Responsibilities

  • Design and deliver high-impact solutions improving reliability, observability, traceability, and incident management.
  • Lead cross-team initiatives with technical leadership and guidance.
  • Collaborate with AI/ML, Data, Platform, and Product teams to build best-in-class services.
  • Define and enforce production standards, tools, and processes.
  • Champion reliability goals and implement SLIs/SLOs across the org.
  • Mentor team members and help develop future engineering leaders.
  • Drive continuous improvement and challenging the status quo.

Skills

Backend Engineering
SRE
DevOps
Golang
Python
Java
Typescript
System Design

Job description

Attentive® is the AI marketing platform for 1:1 personalization redefining the way brands and people connect. We're the only marketing platform that combines powerful technology with human expertise to build authentic customer relationships. By unifying SMS, RCS, email, and push notifications, our AI-powered personalization engine delivers bespoke experiences that drive performance, revenue, and loyalty through real-time behavioral insights.

Recognized as the #1 provider in SMS Marketing by G2, Attentive partners with more than 8,000 customers across 70+ industries. Leading global brands like Crate and Barrel, Urban Outfitters, and Carter's work with us to enable billions of interactions that power tens of billions in revenue for our customers.

With a distributed global workforce and employee hubs in New York City, San Francisco, London, and Sydney, Attentive's team has been consistently recognized for its performance and culture. We're proud to be included in Deloitte’s Fast 500 (four years running!), LinkedIn’s Top Startups, Forbes’ Cloud 100 (five years running!), Inc.’s Best Workplaces, and the Human Rights Campaign Foundation's Corporate Equality Index!

About the Role

Our Platform Infrastructure team is the backbone of everything we do at Attentive, providing a resilient and cost-effective platform that seamlessly handles billions of events from over 100 million customers daily. We own everything from compute, persistence, and networking to observability and deployments. Joining our team offers a high-growth career opportunity to collaborate with some of the world’s most talented engineers in a high-performance, high-impact culture.

As part of the Infrastructure and Platform organization, the Production Engineering Team is focused on delivering a fast and reliable platform that empowers Attentive engineers to deliver solutions quickly and safely. We build scalable systems that automate routine tasks so we can focus on other impactful efforts. Reliability, scalability, and security are our areas of expertise. We focus on release, observability, and cost optimization. Our mission is to create robust platforms and tools that allow stakeholders to concentrate on delivering exceptional products.

As a Staff Engineer, you will take a strategic role in designing and implementing solutions that enhance the reliability and scalability of our systems, while mentoring others and influencing technical roadmaps across the organization.

What You’ll Accomplish
  • Design and Deliver High-Impact Solutions: Design and implement systems that enhance reliability, observability, traceability, and incident management, ensuring the platform scales effectively
  • Lead Strategic Initiatives: Take ownership of cross-team collaborations and drive impactful projects by providing technical leadership and guidance
  • Partner Across Teams: Collaborate with engineers from AI/ML, Data, Platform, and Product teams to develop best-in-class services
  • Partner with engineers from AI/ML, Data, Platform, Product, and other groups to deliver best-in-class services
  • Establish Standards and Best Practices: Define and enforce production standards, processes, and tools to ensure operational excellence
  • Champion Reliability Goals: Advocate for and implement SLIs, SLOs, and other reliability-focused metrics across the engineering organization
  • Mentorship and Knowledge Sharing: Guide and mentor team members, fostering technical growth and helping to develop the next generation of engineering leaders
  • Innovate and Inspire: Drive continuous improvement by bringing creative ideas and challenging the status quo
Your Expertise
  • 7+ years of experience in Production Engineering, Backend Engineering, SRE, DevOps or similar role
  • Strategic visionary: Your strong technical background enables you to look beyond solving the immediate problem, planning for the future.
  • Proficient Problem-Solver: Strong coding ability in at least one language (e.g., Golang, Python, Java, Typescript) with the capability to solve complex issues through code
  • Track Record of Success: Demonstrated experience delivering medium to large-scale projects that drive meaningful improvements in platform reliability and scalability
  • Reliability Expertise: Deep understanding of production reliability concepts, including SLIs, SLOs, and incident management
  • Strong Communicator: Excellent verbal and written communication skills with the ability to influence and collaborate across technical and non-technical teams
  • Fast-Paced Experience: Familiarity with working in dynamic, reliability-focused production environments (preferred)

You'll get competitive perks and benefits, from health & wellness to equity, to help you bring your best self to work.

For US based applicants:

  • The US base salary range for this full-time position is $180,000 - $240,000 annually + equity + benefits
  • Our salary ranges are determined by role, level and location

#LI-HB1

By applying for this position, your data will be processed as per Attentive's Privacy Policy.

Attentive Company Values
  • Default to Action - Move swiftly and with purpose
  • Be One Unstoppable Team - Rally as each other's champions
  • Champion the Customer - Our success is defined by our customers' success
  • Act Like an Owner - Take responsibility for Attentive's success

Learn more about AWAKE, Attentive's collective of employee resource groups.

At Attentive, we know that our Company's strength lies in the diversity of our employees.

Attentive is an Equal Opportunity Employer and we welcome applicants from all backgrounds.

Our policy is to provide equal employment opportunities for all employees, applicants and covered individuals regardless of protected characteristics.

We prioritize and maintain a fair, inclusive and equitable workplace free from discrimination, harassment, and retaliation.

Attentive is also committed to providing reasonable accommodations for candidates with disabilities. If you need any assistance or reasonable accommodations, please let your recruiter know.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Staff Site Reliability Engineer
Staff Site Reliability Engineer

Attentive • Wilmington (DE)

On-site
USD 180,000 - 240,000
Health & wellness
Equity
Engineering Manager, Segmentation
Engineering Manager, Segmentation

Attentive • Wilmington (DE)

On-site
USD 210,000 - 240,000
Health benefits
Equity
Wellness programs
Staff Software Engineer, Shopper Data Platform
Staff Software Engineer, Shopper Data Platform

Attentive • San Francisco (CA)

On-site
USD 180,000 - 230,000
Equity
Health benefits
Wellness programs
Industrial Cleaning Technician
Industrial Cleaning Technician

Cushman & Wakefield • Iowa (LA)

On-site
USD 178,000 - 206,000
Health & wellness benefits
Equity
Staff Software Engineer, Shopper Data Platform
Staff Software Engineer, Shopper Data Platform

Attentive • United States

On-site
USD 180,000 - 230,000
Health benefits
Equity
Support Engineer
Support Engineer

Attentive • Wilmington (DE)

On-site
USD 70,000 - 80,000
Equity
Health & wellness benefits
Staff Software Engineer, Machine Learning
Staff Software Engineer, Machine Learning

Attentive • Wilmington (DE)

On-site
USD 320,000 - 360,000
Equity
Health benefits
Senior ML Engineer — Real-Time Personalization at Scale
Senior ML Engineer — Real-Time Personalization at Scale

Attentive • Wilmington (DE)

On-site
USD 320,000 - 360,000
Engineering Manager, Segmentation
Engineering Manager, Segmentation

Socket.dev • United States

On-site
USD 210,000 - 240,000
Equity
Health benefits
Wellness programs
Engineering Manager, Segmentation
Engineering Manager, Segmentation

Attentive • United States

On-site
USD 210,000 - 240,000