Senior Site Reliability Engineer, Infrastructure Security

Promptly Health

Brasil

Presencial

BRL 180 000 - 320 000

Tempo integral

14 dias+
Gerador de candidaturas

Uma candidatura feita para esta oferta — um currículo e uma carta de apresentação personalizados que vão ao encontro do anúncio.

Ultrapassa os filtros ATS

Vantagens oferecidas por esta oferta de emprego

Competitive salary
Annual performance bonus
Equity compensation via ESOP
Annual training allowance
Private health insurance
Home-office equipment allowance

Resumo da oferta

Promptly is building healthcare data infrastructure with real‑world evidence. We seek a Senior Site Reliability Engineer, Infrastructure Security to strengthen reliability and security across Kubernetes, cloud, and on‑prem environments.

You’ll own security‑driven reliability practices, implement IaC and GitOps, and mentor engineers in secure design and incident response—powering healthcare data platforms that scale safely.

Qualificações

  • 5+ years in SRE/platform/DevSecOps or similar production infra roles.
  • Strong Kubernetes production experience and security expertise.
  • Hands-on infrastructure security across IAM, networking, secrets and logs.
  • Cloud security across AWS/Azure/GCP or private infra.
  • IaC and platform automation with Terraform/Helm.
  • Go/Python/Bash scripting for internal tooling.
  • Solid networking, SLOs/SLIs and incident response experience.
  • Ability to design secure CI/CD pipelines and release workflows.
  • Good security judgement and able to balance risk vs. speed.
  • High agency in ambiguous environments with strong communication.

Responsabilidades

  • Design, build, and operate secure Kubernetes platforms across environments.
  • Improve infrastructure security posture including identity, network boundaries, secrets.
  • Own reliability practices for critical systems and incident response.
  • Build secure platform defaults via IaC, GitOps, CI/CD, policy-as-code.
  • Partner with engineering to embed security from the design phase.
  • Improve vulnerability management for infra, containers, and cloud services.
  • Strengthen production access patterns with least privilege and audit trails.
  • Improve observability and alerting for reliability and security signals.
  • Prepare infrastructure for security reviews and certifications.
  • Lead or support incidents with clear communication and systemic fixes.
  • Mentor engineers on secure infra and operational readiness.

Conhecimentos

SRE experience
Kubernetes
Kubernetes security
Infrastructure security
Cloud security
IaC
Go/Python/Bash
Networking basics
Reliability practices
Secure CI/CD
Security judgement
Adaptability
Incident communication

Ferramentas

Terraform
Helm

Descrição da oferta de emprego

What’s Promptly in a nutshell

Promptly is building the first patient-centered global evidence network, offering real world data sharing and monetization capabilities. Together with a selected network of Partners, we generate new knowledge from harmonized datasets, augmented with the collection of longitudinal patient-reported data and patient-generated digital biomarkers within a secure and privacy-preserving environment.

We answer the question – is this patient treatment the best it could possibly be?

About The Company

We exist to empower every patient and every health organization on the planet with evidence on the outcomes of care. Most healthcare professionals are motivated to make a difference in patients’ lives, and we share that calling by addressing the biggest problem in healthcare: the lack of real-world evidence on the outcomes of care. It is unethical to deny patients better care due to lack of access to data. By making the right evidence available, we promote better healthcare at lower costs, advancing our Hippocratic Oath.

What’s Our Purpose

Our purpose is to empower every patient and every health organization with evidence on the outcomes of care, driving better decision making and lower costs.

About The Role

Promptly is building healthcare data infrastructure for real‑world evidence, patient‑reported outcomes, and AI‑enabled healthcare products. We work with sensitive, fragmented, and regulated data across hospitals, partners, patients, and life sciences organizations.

We are looking for a Senior Site Reliability Engineer, Infrastructure Security to strengthen the reliability and security of the infrastructure behind that platform. This is an SRE role with a strong security focus. You will work on the systems that run Promptly: Kubernetes, cloud and hybrid infrastructure, networking, CI/CD, observability, secrets, identity, access controls, deployment workflows, and incident response. Your job is to make our platform harder to break, easier to operate, and safer for teams to build on.

We are not looking for someone who only writes policies or reviews tickets from the outside. We need an engineer who can stay close to production, understand how systems fail, improve security through better infrastructure, and help engineering teams adopt secure defaults without slowing useful work to a crawl.

The right title for this role could also be Platform Security Engineer or Infrastructure Security Engineer. We are using SRE in the title because the work sits inside the team that owns production infrastructure, reliability, and operational quality.

What You’ll Do
  • Design, build, and operate secure, reliable Kubernetes platforms across cloud, hybrid, and on‑premise environments.
  • Improve the security posture of Promptly’s infrastructure, including identity, access control, network boundaries, secrets management, workload isolation, auditability and secure deployment patterns.
  • Own reliability practices for critical systems, including SLOs, SLIs, incident response, post‑incident analysis, and long‑term remediation.
  • Build secure platform defaults through Infrastructure as Code, GitOps, CI/CD, policy‑as‑code, hardened base images, reusable modules and clear ownership boundaries.
  • Partner with engineering teams to make security part of system design, deployment, observability and operations from the start.
  • Improve vulnerability management for infrastructure, containers, dependencies, cloud services and Kubernetes workloads. Help the team prioritize what matters and fix it properly.
  • Strengthen production access patterns, including least privilege, break‑glass access, audit trails, service accounts, workload identity and human access to sensitive systems.
  • Improve observability and alerting for reliability and security signals across services, infrastructure, data pipelines and customer‑facing products.
  • Help prepare Promptly for security reviews, customer assessments, regulated environments and future certification work by building systems that produce useful evidence naturally.
  • Lead or support incidents with calm judgment, clear communication and focus on systemic fixes.
  • Mentor engineers on secure infrastructure, operational readiness, Kubernetes, incident practices and pragmatic risk reduction.
How We Build With AI

At Promptly, AI‑assisted development is part of how we work. We expect engineers to use coding agents, copilots, code search, AI‑assisted debugging and automated review to move faster through real engineering work. We expect engineers to use AI as part of a modern development workflow, with the same standards of craft, quality, and accountability we apply to any production system. Strong engineers know where AI accelerates the work, where it introduces risk and where human judgment needs to lead.

You own what you ship. You should be able to explain the design, inspect the diff, validate the behavior, check the security and privacy implications, debug failures and operate the system in production. If you cannot understand it, test it, secure it and maintain it you should not merge it.

Ownership, agency and judgment

Senior Site Reliability Engineers at Promptly are expected to own production outcomes, not only infrastructure tasks. For this role, production outcomes include reliability and security. You should be able to identify risks before they become incidents, understand the trade‑offs behind platform decisions and create practical paths to reduce risk without turning security into theatre or process for its own sake.

Healthcare systems involve sensitive data, regulated environments and real customer consequences. We need engineers who can move quickly when the system needs it, slow down when the risk deserves it, communicate clearly and follow through until the platform is stronger than it was before.

What This Role Is Not

This is not a pure compliance, governance or audit role. Compliance matters, but the main lever is engineering.

This is not a penetration‑testing‑only role. Offensive security experience is useful, but we need someone who can turn findings into better infrastructure, controls, automation and operating practices.

This is not a gatekeeper role. We want someone who raises the security bar by building good defaults, making the safe path the easy path and helping teams understand real risk.

What We’re Looking For
Required
  • 5+ years of experience in site reliability engineering, platform engineering, DevSecOps, security engineering or similar production infrastructure roles.
  • Strong production experience with Kubernetes, including cluster design, workload management, networking, ingress, RBAC, secrets, observability and failure recovery.
  • Advanced Kubernetes security experience, including Pod Security Standards, NetworkPolicies, admission controllers, policy engines for enforcing security and compliance rules, workload identity, runtime security or multi‑cluster management.
  • Hands‑on experience improving infrastructure security in production environments. You have worked with identity and access management, network controls, vulnerability management, secrets, TLS, audit logs and secure deployment workflows.
  • Experience with cloud security across AWS, Azure, GCP, VMware or private infrastructure, including IAM, KMS, private networking, logging and account or subscription structure.
  • Strong Infrastructure as Code experience, especially with tools such as Terraform, Helm or similar platform automation frameworks.
  • Strong programming or scripting skills in Go, Python, Bash or similar languages. You should be able to build internal tooling, automation and operational workflows.
  • Strong understanding of networking fundamentals, including DNS, service discovery, virtual networks, load balancing, ingress, TLS, routing, firewalls and private connectivity.
  • Experience with reliability practices such as SLOs, SLIs, incident response, postmortems, remediation plans and operational readiness reviews.
  • Experience designing secure CI/CD, deployment automation and release workflows for production systems, including secrets handling, image and dependency scanning and policy checks in pipelines.
  • Good security judgement. You can separate real risks from noise, explain trade‑offs clearly and avoid both careless shortcuts and security theatre.
  • High agency in ambiguous environments. You can create clarity, make trade‑offs and move work forward without needing every detail defined upfront.
  • Strong communication skills during incidents, design discussions, risk reviews and cross‑team platform work.
Preferred
  • Experience with software supply‑chain security, including container scanning, dependency scanning, SBOMs, artifact signing, provenance, hardened images or secure build pipelines.
  • Experience in healthcare, health technology, life sciences, fintech, enterprise SaaS or another regulated environment.
  • Experience supporting SOC 2, ISO 27001, GDPR, customer security reviews, vendor assessments or similar trust and assurance work.
  • Hands‑on experience with observability tools such as Prometheus, Grafana, Loki, OpenTelemetry, Jaeger, Elasticsearch or similar systems.
  • Experience with GitOps workflows using tools such as Argo CD, Flux or similar systems.
  • Experience with data‑intensive platforms or distributed data systems such as Kafka, Trino, Apache Iceberg, Spark, Airflow or similar technologies.
  • Experience building SRE or security tooling, deployment platforms, observability frameworks, incident automation, asset inventory or risk dashboards.
  • Relevant certifications, especially CKS, as well as CKA or other security certifications.
How You’ll Work

You will work inside the infrastructure and platform area, close to the teams that ship product and operate customer‑facing systems.

Some weeks you may harden Kubernetes workloads, improve IAM boundaries or fix a vulnerable deployment pattern. Other weeks you may support an incident, improve observability, review a production architecture, prepare evidence for a customer security review, or build automation that removes a repeated operational risk.

You will work with engineering managers, product engineers, data teams, customer‑facing teams, leadership and external security or compliance partners. Your job is to help Promptly build systems that are secure enough for sensitive healthcare work and practical enough for engineers to use every day.

You’ll be a strong fit if
  • You care about production systems and the people who depend on them.
  • You think security belongs in infrastructure design, deployment workflows, and operational practice, not only in a checklist at the end.
  • You can debug across application, infrastructure, network, identity and data layers.
  • You know when to automate, when to simplify and when to remove a fragile system entirely.
  • You can move quickly during incidents without creating panic or noise.
  • You care about observability, runbooks, tests, rollback paths, audit trails and operational readiness.
  • You are comfortable with regulated environments, sensitive data and high‑trust systems.
  • You can make security feel like engineering help rather than bureaucracy.
Why Promptly

Promptly is working on one of the hardest and most important problems in healthcare: making outcomes data usable so patients, clinicians, healthcare organisations and researchers can make better decisions. Security and reliability matter here because our systems support sensitive healthcare workflows, partner integrations, data operations and evidence generation. The infrastructure you build will shape how safely and effectively Promptly can scale.

Position
  • Remote‑first
  • Full time
What We Offer
Ownership & Growth
  • Define and own a strategic global initiative from the ground up
  • Shape the future of Promptly’s partner ecosystem and data network expansion
Financial Benefits
  • Competitive salary
  • Annual performance bonus
  • Equity compensation via our ESOP (open to all team members)
  • Annual training allowance
  • Private health insurance
  • Home‑office equipment allowance
Non‑Financial Benefits
  • Equal opportunity and inclusive environment
  • Flexible work schedule and vacation policy
  • Corporate events and international team gatherings
What Is The Recruiting Process Like
  • Initial Interview – Meet your future manager and/or team members to get to know each other and understand the scope of the role.
  • Technical & Strategic Assessment – Work on a short case related to partner management or program design.
  • Final Interview (if needed) – Deep‑dive discussion to align on expectations and vision.
  • Offer Stage – We’ll share the offer and welcome you to the team!
Our Culture & Values

Empathy – Ownership – Responsibility – Teamwork – Excellence. If you embrace these values, you will thrive at Promptly. We are a purpose‑driven company where everyone acts as an owner, committed to improving healthcare through data, technology and collaboration.

Data Privacy

Your personal data will be processed for the purposes of managing Promptly’s recruitment related activities which includes setting up and conducting interviews and tests for applicants, assess and review such candidates and similar activities needed in the recruitment and hiring process. Your personal data will be shared with Proef, a provider engaged by Promptly to help us manage our recruitment and hiring process. For more information about how Promptly processes personal data and information about your rights etc., please see Promptly’s/ Proef’s Privacy Policy (link to:Policy for management of application (proef.com)).

Obtém a tua avaliação gratuita e confidencial do currículo.
ou arrasta e larga o ficheiro aqui.
Similar jobs

Ofertas semelhantes que vale a pena comparar

Senior Site Reliability Engineer
Senior Site Reliability Engineer

Zoomcar • Brasil

Presencial
BRL 422 000 - 485 000
Stock options
Health benefits
Unlimited PTO
+2
Staff Site Reliability Engineer - Work from home
Staff Site Reliability Engineer - Work from home

Nearsure • Rio de Janeiro

Teletrabalho
BRL 382 000 - 547 000
Competitive USD salary
100% remote work
Paid time off
+5
Principal Software Engineer - Card Payments
Principal Software Engineer - Card Payments

PPRO • São Paulo

Híbrido
BRL 180 000 - 300 000
Hybrid working
Gym allowance
Mental health platform
+4
Experienced HITRUST Assessment Manager
Experienced HITRUST Assessment Manager

Insight Assurance • Brasil

Teletrabalho
BRL 421 000 - 633 000
Flexible Paid Time Off
Quarterly Performance Bonuses
Opportunities for professional growth
+1
Global Project Manager Patient Care
Global Project Manager Patient Care

Docplanner • Brasil

Híbrido
BRL 15 000 - 25 000
Healthcare insurance
Gym memberships
Flexible working hours
+2
Site Reliability Engineer ID55632
Site Reliability Engineer ID55632

AgileEngine • São Paulo

Híbrido
Professional growth opportunities
Competitive USD-based compensation
Exciting projects with top companies
+1
Data Engineer, Product
Data Engineer, Product

FuturHealth • Brasil

Híbrido
BRL 140 000 - 210 000
Unlimited PTO
Paid holidays
Remote-friendly policy
+1
Senior Full Stack Engineer Id86905
Senior Full Stack Engineer Id86905

Agileengine • Sorocaba

Presencial
BRL 180 000 - 280 000
Growth opportunities
Competitive compensation
100% remote option
+3
Site Reliability Engineer ID45689
Site Reliability Engineer ID45689

AgileEngine • Riograndina

Híbrido
BRL 385 000 - 551 000
Professional growth
Competitive compensation
Flextime
+1
Software Engineer, Athelas Home
Software Engineer, Athelas Home

Athelas • São Paulo

Presencial
BRL 180 000 - 320 000