Sr. Manager, SRE & Performance

Omnissa

Mountain View (CA)

Hybrid

USD 223,000 - 310,000

Full time

9 days ago
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Employee ownership
Health insurance
401k with matching

Job summary

Omnissa seeks a Senior Manager of Site Reliability Engineering (SRE) & Performance Engineering to lead reliability, scalability, and operational excellence for its distributed SaaS platform. This hands-on technical leadership role collaborates with development, security, SaaS operations, and product teams to ensure services are designed and operated for scale, resilience, and efficient performance.

The ideal candidate has 12+ years in software engineering, 5+ years in leadership, and a strong

Qualifications

  • 12+ years of software engineering experience with large-scale systems.
  • 5+ years in engineering leadership/management roles (SRE/Platform).
  • Experience building and operating distributed systems and backend services.
  • Strong Linux, containers, and cloud infrastructure knowledge.
  • Experience with observability tooling and performance tuning.

Responsibilities

  • Lead and grow SRE and performance teams, shaping culture of ownership.
  • Define and drive SRE/performance strategy, goals, and readiness criteria.
  • Partner with engineering to design for reliability and scalability from inception.
  • Establish SLI/SLOs, error budgets, and production readiness standards.
  • Lead performance engineering: benchmarking, profiling, capacity planning, regression testing.
  • Drive observability strategy with metrics, logs, traces, dashboards, alerts.
  • Lead incident responses and postmortems, driving systemic improvements.
  • Collaborate on system design reviews for scalability and resilience.
  • Integrate capacity forecasting into CI/CD and deployment processes.
  • Improve platform efficiency, automation, and cost-aware engineering decisions.
  • Define KPIs for reliability and performance and track them.
  • Coordinate with Product/Engineering/Sec/Cloud/SaaS Ops for secure, resilient services.

Skills

SRE principles
SLIs/SLOs
Observability
Incident management
Performance engineering
Capacity planning
Cloud experience
CI/CD
Infrastructure as Code

Education

Bachelor’s degree

Tools

Docker
Nomad
Consul
Vault

Job description

Job Description:

We areOmnissa!
Omnissa is the first AI-driven digital work platform, built to support flexible, secure, work-from anywhere experiences. We integrate industry-leading solutions—including Unified Endpoint Management,Virtual Appsand Desktops, Digital Employee Experience, and Security & Compliance—into a seamless, autonomous workspace that adapts to how people work. Our platform boosts employee engagement whileoptimizingIT operations, security, and cost.
Guided by our Core Values— Act in Alignment, Build Trust, Foster Inclusiveness, Drive Efficiency, and Maximize Customer Value —we’regrowing rapidly and committed to delivering meaningful impact. Ifyou’repassionate about shaping the future of work,we’dlove to hear from you.

AtOmnissa, we are committed tomaintaininga fair, consistent, and secure hiring process for all candidates. As part of this approach, we use standard interview and verification practices designed to ensure alignment and protect both candidates and the organization. These practices are applied thoughtfully and with respectforcandidate privacy.

What is the opportunity?

We are looking for aSenior Manager of Site Reliability Engineering (SRE) & Performance Engineeringto lead teams responsible for the reliability, scalability, performance, and operational excellenceof our distributed SaaS platform. This is a hands-on technical leadership role for someone who can operate at the intersection ofsoftware engineering, distributed systems, cloud infrastructure, SRE, and performance engineering. You will lead engineers while partnering closely with development, architecture, security, SaaS operations, and product teams to ensure our services are designed and operated forscale, resilience, efficiency, and predictable performance. You will help establish engineering standards aroundSLOs, observability, performance testing, capacity planning, incident management, production readiness, and continuous reliability improvement. Here’s a breakdown:

  • Lead and grow SRE and Performance Engineering teams, providing technical direction, coaching, career development, and establishing a strong culture of ownership and engineering excellence.
  • Define and drive the organization’sSRE and performance engineering strategy, including reliability goals, performance objectives, scalability standards, and operational readiness requirements.
  • Partner with engineering teams to ensure systems are designed forhigh availability, scalability, fault tolerance, performance, and operabilityfrom the beginning rather than addressing these concerns after deployment.
  • Establish and drive adoption ofSLIs, SLOs, error budgets, service health indicators, and production readiness criteriafor critical services.
  • Leadperformance engineering initiativesacross distributed systems, including workload modeling, benchmarking, profiling, scalability testing, capacity planning, and bottleneck analysis across CPU, memory, I/O, storage, database, and network layers.
  • Driveobservability strategyacross metrics, logs, traces, dashboards, and alerting, while improving signal quality and reducing noisy or non-actionable alerts.
  • Provide technical leadership duringcomplex production incidents, helping teams diagnose distributed system failures, latency spikes, resource contention, database issues, network problems, and cascading failures.
  • Drive effectiveincident management, RCA/postmortem practices, corrective actions, and systemic reliability improvements, ensuring lessons from incidents translate into engineering changes.
  • Partner with architects and senior engineers onsystem design and architecture reviews, challenging designs around scalability, resilience, failure modes, performance, data architecture, and operational complexity.
  • Establish scalable approaches forperformance regression detection and reliability validation within CI/CD pipelines, enabling issues to be identified earlier in the development lifecycle.
  • Drivecapacity management and forecasting, using production telemetry, workload characteristics, and performance models to anticipate infrastructure requirements and scalability limits.
  • Improve platform efficiency throughperformance optimization, infrastructure right-sizing, and cost-aware engineering, balancing reliability, performance, and cloud cost.
  • Championautomation and engineering-driven operations, reducing manual operational work and toil through software, tooling, Infrastructure as Code, and automated remediation.
  • Establish measurableSRE and performance KPIs, such as availability/SLO attainment, MTTR, change failure rate, alert quality, performance regression rates, capacity headroom, operational toil, and recurring incident reduction.
  • Collaborate across Product, Engineering, Security, Cloud/SaaS Operations, and Architecture organizations to deliversecure, resilient, scalable, and production-ready services.
What will you bring to the company?
Leadership & Engineering Management
  • 12+ years of software engineering experience, with significant experience building and operating large-scale backend or distributed systems.
  • 5+ years of engineering leadership/management experience, preferably leading SRE, Performance Engineering, Platform Engineering, Infrastructure, or backend engineering teams.
  • Proven ability tobuild, mentor, and develop high-performing engineering teams, including senior and staff-level engineers.
  • Strong ability to balancepeople leadership, technical strategy, operational priorities, and business objectives.
  • Experience influencing engineering practices across teamswithout relying solely on organizational authority.
  • Demonstrated ability to work effectively with senior engineers, architects, engineering managers, product leaders, security teams, and executive stakeholders.
Technical Depth
  • Strong understanding ofdistributed systems and microservices architectures, including scalability, availability, consistency, fault tolerance, failure modes, and architectural trade-offs.
  • Strong software engineering background, preferably with experience inC#/.NET, Go or Java, and the ability to participate meaningfully in architecture, design, and code-level technical discussions.
  • Deep understanding ofLinux systems, including processes, memory, CPU scheduling, networking, file systems, containers, and system-level performance diagnostics.
  • Strong experience withDockerand container orchestration using HashiStack technologies such as Nomad, Consul, and Vault.
  • Experience operating high-scale data platforms using technologies such asKafka, PostgreSQL, OpenSearch, Redis/Valkey, or equivalent technologies.
  • Strong cloud experience, preferably withAWS, including services such as EC2, EKS, MSK, Aurora/RDS, OpenSearch, networking, storage, and cloud observability services.
  • Strong understanding ofCI/CD, Infrastructure as Code, deployment automation, release strategies, and modern DevOps practices.
SRE & Performance Engineering
  • Deep understanding ofSRE principles, including SLIs/SLOs, error budgets, availability engineering, incident management, production readiness, operational toil, and reliability automation.
  • Proven experience designing and implementingobservability strategiesusing metrics, logging, tracing, dashboards, and actionable alerting.
  • Strong understanding ofperformance engineering methodologies, including workload modeling, benchmarking, profiling, stress/load testing, scalability analysis, and performance regression detection.
  • Experience diagnosing performance issues acrossapplications, databases, operating systems, containers, infrastructure, and network layers.
  • Experience withcapacity planning and forecasting, including translating workload growth into infrastructure and service capacity requirements.
  • Strong understanding ofresilience engineering, including graceful degradation, retries, timeouts, circuit breakers, backpressure, rate limiting, disaster recovery, and failure testing.
  • Demonstrated ability to turn production incidents and performance findings intosystemic engineering improvements rather than tactical fixes.

Location: Mountain View, CA
Location Type: hybrid
Education: Bachelor’s Degree preferred, or equivalent combination of education and relevant professional experience.

Compensation: The typical base salary for this role is between USD $223,000– $310,000 per year and it may be eligible for participation in a corporate bonus program. Actual compensation offer may vary from posted hiring range based upon geographic location, work experience, education, skill level, or other relevant factors. In addition to competitive compensation,Omnissaoffers a variety of benefits such as employee ownership, health insurance, 401k with matching contributions, disability insurance,paid-timeoff, growth opportunities, and more.

Omnissa is an Equal Employment Opportunity company and Prohibits Discrimination and Harassment of Any Kind:
Omnissa is committed to the principle of equal employment opportunity and to providing a work environment free of discrimination and harassment. All employment decisions atOmnissaare based on business needs, job requirements and individual qualifications, without regard to race, color, religion, ancestry, ethnicity, national, social or ethnic origin, sex (including pregnancy), age, physical, mental or sensory disability, HIV status, sexual orientation, gender identity and/or expression, marital, civil union or domestic partnership status, past, present, or prospective service in the uniformed services, family medical history or genetic information, family or parental status, veteran status, or any other status protected by applicable laws or regulations in the locations where we operate.Omnissa will not tolerate discrimination or harassment based on any of these characteristics.Omnissa will welcome applicants of all ages.Omnissa will provide reasonable accommodations to applicants and employees who have protected disabilities consistent with applicable federal,stateand local law.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Sr. Manager, SRE & Performance
Sr. Manager, SRE & Performance

Omnissa, LLC in • Mountain View (CA)

Hybrid
USD 223,000 - 310,000
Employee ownership
Health insurance
401k with matching contributions
+3
Software Engineer, Backend
Software Engineer, Backend

Omnissa, LLC in • Atlanta (GA)

Hybrid
USD 121,000 - 201,000
Employee ownership
Health insurance
401k with matching
+3
DevOps Lead, Cloud Automation & Reliability
DevOps Lead, Cloud Automation & Reliability

Omnissa • Mountain View (CA)

Hybrid
USD 134,000 - 280,000
Employee ownership
Health insurance
401k with matching contributions
+3
Staff Machine Learning Engineer
Staff Machine Learning Engineer

Latitude • Atlanta (GA)

Hybrid
USD 163,000 - 343,000
Employee ownership
Health insurance
401k with matching
+3
DevOps Lead, Cloud Automation & Reliability
DevOps Lead, Cloud Automation & Reliability

Omnissa • Atlanta (GA)

On-site
USD 134,000 - 280,000
Employee ownership
Health insurance
401k with matching
+3
Sr. Engineering Manager, Frontline Engineering
Sr. Engineering Manager, Frontline Engineering

Omnissa, LLC in • Atlanta (GA)

Hybrid
USD 177,000 - 325,000
Employee ownership
Health insurance
401k with matching
+3
Staff DevOps Engineer
Staff DevOps Engineer

Omnissa, LLC in • Mountain View (CA)

Hybrid
USD 206,000 - 343,000
Employee ownership
Health insurance
401k with matching contributions
+3
Sr. DevOps Engineer
Sr. DevOps Engineer

Omnissa • Atlanta (GA)

Hybrid
USD 141,000 - 280,000
Employee ownership
Health insurance
401k with matching contributions
+3
Sr. DevOps Engineer
Sr. DevOps Engineer

Omnissa, LLC • Mountain View (CA)

On-site
USD 176,362 - 293,937
Employee ownership
Health insurance
401k with matching contributions
+3
Staff II - Application Security Engineer
Staff II - Application Security Engineer

Omnissa, LLC in • Atlanta (GA)

Hybrid
USD 220,000 - 270,000