Sr. Staff Platform Engineer

Ellkay, Llc

United States

On-site

USD 140,000 - 190,000

Full time

27 hours ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

ELLKAY is seeking a Senior Staff Platform Engineer to own design, build, and operations for our hybrid cloud infrastructure across AWS, Azure, and on-prem environments. You will codify everything with Terraform, drive CI/CD for Kubernetes and Docker, and mentor teams on reliability and security.

You will partner with Platform Engineering and SRE to set standards for deployment, observability, and cost governance while improving incident response and uptime across global environments.

Qualifications

  • 12+ years in platform engineering, infrastructure engineering, DevOps, or SRE roles, with demonstrated staff-level scope and impact.

Responsibilities

  • Design, build, and operate infrastructure across AWS, Azure, and on-prem environments.
  • Own IaC tooling and standardize CI/CD pipelines for containerized services.

Skills

AWS
Azure
Hybrid cloud
Terraform
CI/CD pipelines
Kubernetes
Docker
Ansible
Chef
Puppet
Prometheus
Grafana
Datadog
OpenTelemetry
ELK/Splunk
Incident management
Cloud cost governance
Security/compliance
Python
Go
Bash
Communication

Tools

Terraform
Kubernetes
Docker
Ansible
Chef
Puppet

Job description

If you are unable to complete this application due to a disability, contact this employer to ask for an accommodation or an alternative application process.

Full Time IN

3 days ago Requisition ID: 1365

ELLKAY started out providing connectivity solutions to laboratories and within a few years, grew to also provide data management solutions to ambulatory organizations. ELLKAY is now a trusted data management partner in five healthcare segments. ELLKAY’s solutions continue to serve laboratories and ambulatory practices and have expanded to empower hospitals and health systems, healthcare IT vendors, ambulatory practices, health plans, and other healthcare organizations with cutting-edge technologies and solutions that drive their growth and interoperability strategies.

Today, ELLKAY remains true to our core values, building strong partner relationships and offering unparalleled service and support while providing innovative, scalable solutions to the challenges our customers face in today’s data-rich world.

ELLKAY's experience, customer-focused approach, and reputation for innovation, speed, and accuracy differentiate ELLKAY as a premier partner for your interoperability needs and data management strategy.

Job Description

We're looking for a Senior Staff Platform Engineer to own the design, build, and operational health of our infrastructure across AWS, Azure, and on-premises environments. This is a hands-on, high-ownership role for someone who thinks in systems, codifies everything, and is equally comfortable writing Terraform modules, debugging production incidents throughout the clock, and advising leadership on cost and security posture.

You will be the key design and implementation engineer handling platform engineering initiatives and our Site Reliability Engineering (SRE) teams, setting the standards for how services are deployed, observed, secured, and operated at scale across hybrid cloud and on-prem infrastructure

Infrastructure as Code
  • Design, build, and maintain reusable Terraform modules to provision and manage infrastructure across AWS, Azure, and on-prem environments
  • Establish IaC standards, module versioning strategy, state management practices, and review processes across engineering teams
  • Drive migration of manually managed infrastructure to fully codified, version-controlled definitions
  • Architect and implement CI/CD pipelines for containerized and Kubernetes-based services, from build through progressive production rollout
  • Define deployment strategies (blue/green, canary, rolling) and the automation that supports them
  • Partner with development teams to streamline the path from commit to production while maintaining safety and auditability
Configuration Management
  • Own configuration management tooling and practices across hybrid environments (e.g., Ansible, Chef, Puppet, or equivalent) to ensure consistency, repeatability, and drift detection
  • Standardize secrets management, environment configuration, and golden image/baseline practices across cloud and on-prem fleets
Observability
  • Design and implement observability patterns (metrics, logging, tracing) that provide actionable signal across distributed, hybrid-cloud services
  • Define SLIs/SLOs in partnership with SRE and product teams, and build the dashboards and alerting that make them actionable
  • Reduce mean-time-to-detect (MTTD) and mean-time-to-resolve (MTTR) through better instrumentation, not just more of it
Production Reliability
  • Act as a senior escalation point for complex production incidents, driving root cause analysis and durable remediation
  • Lead or contribute to postmortems and translate findings into infrastructure, process, or tooling improvements
  • Proactively identify and remediate reliability risk before it becomes an incident
Cost & Security Governance
  • Design and implement cost governance practices — tagging standards, budget alerting, rightsizing, and reserved capacity strategy across AWS, Azure, and onprem infrastructure
  • Partner with Security to define and enforce infrastructure security guardrails (IAM least-privilege, network segmentation, secrets handling, compliance controls)
  • Build automated policy enforcement (e.g., policy-as-code) so governance scales with infrastructure rather than depending on manual review
  • Serve as the primary point of contact between Platform Engineering and SRE teams, aligning on standards, priorities, and shared tooling
  • Mentor senior and mid-level engineers on infrastructure design, operational excellence, and IaC best practices
  • Influence infrastructure architecture and technical roadmap at the organizational level
Qualifications

12+ years in platform engineering, infrastructure engineering, DevOps, or SRE roles, with demonstrated staff-level scope and impact

  • Deep, production-grade experience with both AWS and Azure, plus experience managing on-premises infrastructure in a hybrid model
  • Expert-level Terraform experience — module design, state management, workspace/environment strategy at scale
  • Strong experience building CI/CD pipelines for Kubernetes and Docker-based workloads
  • Hands-on experience with configuration management tooling (Ansible, Chef, Puppet, or similar)
  • Proven track record implementing observability stacks (e.g., Prometheus, Grafana, Datadog, OpenTelemetry, ELK/Splunk) and defining meaningful SLIs/SLOs
  • Demonstrated experience leading production incident response and driving reliability improvements
  • Experience designing cloud cost governance and security/compliance frameworks in a multi-cloud or hybrid environment
  • Strong scripting/programming ability (Python, Go, or Bash) for automation and tooling
  • Excellent cross-functional communication — able to work directly with SRE, security, and engineering leadership
Additional Information

ELLKAY is committed to fostering a collaborative and high-performance work environment that supports innovation, teamwork, and professional growth. Most roles are designed to operate from our office locations to encourage effective collaboration and engagement across teams.
Any alternative work arrangements may be considered at the company’s discretion based on role requirements and business needs.
For more information about our company, please visitwww.ELLKAY.com .
ELLKAY is a Smoke-Free Workplace.

ELLKAY, LLC provides equal employment opportunities to all employees and applicants for employment and prohibits discrimination and harassment of any type without regard to race, color, religion, age, sex, national origin, disability status, genetics, protected veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by federal, state or local laws.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Principal Platform Engineer
Principal Platform Engineer

Ellkay,-LLC • United States

Hybrid
USD 160,000 - 180,000
Remote work options
401k with matching
Hybrid work model
+1
Staff CI/CD Engineer
Staff CI/CD Engineer

Ellkay, Llc • United States

On-site
USD 130,000 - 190,000
Application SRE
Application SRE

Ellkay, Llc • Elmwood Park (NJ)

Hybrid
USD 90,000 - 110,000
Medical, Dental, and Vision benefits
401k with matching
Generous paid time off (FTO)
+2
Staff Engineer - Data Platform & APIs
Staff Engineer - Data Platform & APIs

Socket.dev • United States

Hybrid
USD 150,000 - 210,000
Medical, Dental, and Vision benefits
Remote work options
401k w/ matching
+2
Staff Engineer - Data Platform & APIs
Staff Engineer - Data Platform & APIs

Ellkay, Llc • Northern (KY)

Hybrid
USD 180,000 - 200,000
Medical, Dental, Vision benefits
Employer-paid Life and LTD
401k with matching
+2
Staff Data Engineer- Cloud Data Platform
Staff Data Engineer- Cloud Data Platform

Ellkay, Llc • United States

On-site
USD 120,000 - 180,000
Smoke-Free Workplace
Sr. Cloud Data Engineer
Sr. Cloud Data Engineer

Ellkay,-LLC • Elmwood Park (NJ)

Hybrid
USD 125,000 - 150,000
Remote work options
Health benefits
Life and LTD
+5
Data Platform Principal
Data Platform Principal

Ellkay, Llc • Northern (KY)

Hybrid
USD 200,000 - 220,000
Medical benefits
Dental benefits
Vision benefits
+5
Sr. Cloud Data Engineer
Sr. Cloud Data Engineer

Ellkay, Llc • Elmwood Park (NJ), Northern (KY)

Hybrid
USD 125,000 - 150,000
Medical, Dental, and Vision benefits
Employer-paid Life and LTD
401k with matching
+4
Director of Data Platform & Data Lake
Director of Data Platform & Data Lake

Ellkay, Llc • United States

On-site
USD 180,000 - 240,000