Senior Software Engineer (SRE)

eMed LLC.

Miami (FL)

On-site

USD 140,000 - 190,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Retirement Plan (401k with Company Map
Life Insurance (Basic, Voluntary & AD
Paid Time Off
Short Term & Long Term Disability
Training & Development
Catered Breakfast and Lunch 5 days a 5

Job summary

eMed LLC is seeking a Senior Software Engineer in SRE to ensure our platform is highly available, secure, and performant. You will lead reliability efforts, drive automation, and collaborate with product and infrastructure teams to design resilient services.

You will own monitoring, incident handling, and capacity planning, mentoring engineers in observability and operational engineering, while optimizing Kubernetes/AWS workloads for cost and security.

Qualifications

  • Experience operating Kubernetes and cloud-native infrastructure in production (AWS EKS preferred).
  • Proficiency with Terraform and IaC practices.
  • Strong coding skills for building tools and automation; able to troubleshoot complex infra.

Responsibilities

  • Design, implement, and manage monitoring, alerting, and observability across services.
  • Lead incident response, post-incident reviews, and prevention enhancements.
  • Improve scalability, fault-tolerance, and performance via architecture input and automation.
  • Develop infrastructure automation with Terraform and CI pipelines with GitHub Actions.
  • Collaborate with engineers to boost service resilience, capacity planning, and runbooks.
  • Reduce operational toil through tooling and process improvements.
  • Manage production Kubernetes and AWS environments focusing on reliability, security, and cost.
  • Contribute to security hardening, including network controls and secrets management.

Skills

Kubernetes
AWS
Observability
Automation
Incident management
SRE mindset
Scripting

Tools

GitHub Actions
Terraform
CloudWatch
ELB
VPC
EKS

Job description

As a Senior Software Engineer in SRE at eMed, you will play a key role in ensuring our platform is highly available, secure, and performant. You’ll lead reliability engineering efforts across production systems, drive operational excellence, and collaborate closely with application and infrastructure teams to design resilient services. This role suits an engineer with a software mindset and deep operational experience, who thrives on improving systems through automation and proactive engineering.

2.2 WHAT YOU WILL WORK ON
  • Design and implement robust monitoring, alerting, and observability systems across all services and infrastructure
  • Lead reliability reviews, incident response, and post-incident analysis—focusing on prevention, learning, and long-term improvements
  • Improve service scalability, fault tolerance, and performance through architectural input and systems optimisation
  • Build and maintain automation for infrastructure management using Terraform, and delivery pipelines using GitHub Actions
  • Partner with software engineers to improve the operational readiness and resilience of services, including capacity planning and runbooks
  • Lead initiatives to reduce operational toil through tooling, automation, and process improvement
  • Manage and optimise our production Kubernetes and AWS environments with a focus on reliability, security, and cost-effectiveness
  • Contribute to security hardening efforts, including network controls, secrets management, and compliance readiness
  • Participate in and lead in-person stand-ups, incident reviews, and cross-team planning sessions
  • Share knowledge and mentor engineers on best practices in observability, incident response, and operational engineering
2.3 WHAT WE’RE LOOKING FOR:
Technical Skills (Essential)
  • Strong experience operating Kubernetes and cloud-native infrastructure (preferably EKS on AWS) in production environments
  • Proficiency in AWS services, including networking, compute, IAM, and logging/monitoring tools (e.g. CloudWatch, ELB, VPC)
  • Skilled in Terraform and Infrastructure as Code practices
  • Deep understanding of observability tooling (metrics, logs, tracing) and incident management workflows
  • Strong coding skills for building tools, scripts, and automation
  • Ability to troubleshoot complex infrastructure issues and lead delivery of reliable cloud solutions
Preferred
  • Experience implementing SLAs, SLOs, and error budgets to guide operational priorities
  • Background in healthcare or other regulated industries with security and compliance requirements
  • Previous involvement in platform security reviews

2.4 PERKS AT WORK:
  • Retirement Plan (401k with Company Match)
  • Life Insurance (Basic, Voluntary & AD&D)
  • Paid Time Off
  • Short Term & Long Term Disability
  • Training & Development
  • Catered Breakfast and Lunch 5 days a Week
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Software Engineer (SRE)
Senior Software Engineer (SRE)

eMed • Miami (FL)

On-site
USD 100,000 - 130,000
Health Care Plan (Medical, Dental & Vision)
Retirement Plan (401k with Company Match)
Life Insurance (Basic, Voluntary & AD&D)
+5
Senior SRE Engineer: Build Resilient, Automated Cloud
Senior SRE Engineer: Build Resilient, Automated Cloud

eMed LLC. • Miami (FL)

On-site
USD 140,000 - 190,000
Retirement Plan (401k with Company Map
Life Insurance (Basic, Voluntary & AD
Paid Time Off
+3
Senior DevOps Engineer
Senior DevOps Engineer

Modernizing Medicine, Inc. • United States

Hybrid
USD 110,000 - 150,000
Comprehensive medical, dental, and vision benefits
401(k) matching plan
Generous paid time off
+2
Sr Software Engineer - Reliability Engineering
Sr Software Engineer - Reliability Engineering

Cox Enterprises • Village of North Hills (NY)

On-site
USD 150,000 - 185,000
SRE Leader
SRE Leader

Kontakt Micro-Location Sp. Z.o.o. • New York (NY)

Hybrid
USD 180,000 - 260,000
Equity in a high-growth company
Health, dental, and vision coverage
401k
+3
Senior Site Reliability Engineer
Senior Site Reliability Engineer

The ReWork Group • New York (NY)

On-site
USD 120,000 - 160,000
Senior Site Reliability Engineer (SRE)
Senior Site Reliability Engineer (SRE)

Practice by Numbers • United States

On-site
USD 120,000 - 160,000
High ownership and autonomy
Strong engineering culture
Impactful work on healthcare infrastructure
Sr. DevOps Engineer
Sr. DevOps Engineer

Socket.dev • Irvine (CA)

On-site
USD 140,000 - 210,000
100% Company-Paid Medical, Dental, Vis
401(k) with Company Match
Flexible Spending Account
+5
Senior DevOps Engineer/Site Reliability Engineer-East Coast
Senior DevOps Engineer/Site Reliability Engineer-East Coast

Stellar Cyber • North Carolina

On-site
USD 165,000 - 215,000
Pre‑IPO Stock Options
Medical, Dental & Vision care
401(k)
+2
Sr. DevOps Engineer
Sr. DevOps Engineer

EKN Engineering • Irvine (CA)

On-site
USD 140,000 - 190,000
Medical Insurance
Dental Insurance
Vision Insurance
+6