Join the Team Modernizing Medicine
At ModMed , we're not just building software-we're reimagining the healthcare experience. Founded in 2010 by a practicing physician and a successful tech entrepreneur, we took a radically different approach: we hired doctors and taught them how to code. This \"for doctors, by doctors\" philosophy has allowed us to create an AI-enabled, specialty-specific cloud platform that places patients at the center of care.
A Culture of Excellence
When you join ModMed, you're joining an award-winning team recognized for innovation and employee satisfaction. From our global headquarters in Boca Raton Florida, and extensive employee base in Hyderabad India, we are a team of 4,500+ passionate problem-solvers on a mission to increase medical practice success and improve patient outcomes:
Consistently ranked as a Top Place to Work
- 2025 Globee Business Awards: Gold Globee for \"Technology Team of the Year\"
- 2025 Black Book Awards: Ranked #1 EHR in 11 Specialties
- Florida Venture Forum: Venture-Backed Company of the Year
We are growing fast, thinking big, and we are just getting started.
Ready to modernize medicine with us?
Job Description Summary:
Our Mission & Vision
- Our Mission: To place doctors and patients at the center of care through an intelligent, specialty-specific cloud platform.
- Our Vision: A world where the software ModMed builds increases medical-practice success and improves patient outcomes.
- Our Core Values: Create customer delight | Save time | Innovate boldly, then make things happen | Align passion with purpose | Think big | Have fun | Do good.
The Role: Technical Visionary for Reliability and Developer Velocity
As a Principal Site Reliability Engineer at ModMed, you are a primary architect of our technical future. You don't just solve problems; you anticipate the needs of a global healthcare platform years in advance. You will think big to define the standards for reliability, scalability, developer velocity, and security that allow our doctors to provide world-class care without interruption. This is a high-impact technical leadership role where you will innovate boldly, then make things happen, acting as a vital bridge between engineering, product, and business goals. By aligning passion with purpose, you will build a resilient infrastructure that serves as the backbone for modern medicine and consistently creates customer delight.
Primary Responsibilities
Core Infrastructure Strategy & Architecture
- Architectural Strategy & Technical Governance: Define and execute the long-term architectural vision for our global AWS cloud ecosystem. Design high-performance, fault‑tolerant, and cost‑optimized multi‑region topologies capable of scaling elastically to meet massive transaction volumes while enforcing rigorous cloud hygiene (Save time).
- Systemic Observability Ecosystems: Institutionalize \"Observability as a Culture\" across the enterprise. Architect unified, global telemetry standards using Datadog and OpenTelemetry, transforming raw logs, metrics, and distributed traces into actionable, predictive insights that elevate system reliability across all product engineering groups.
- Developer Platform Evolution & CI/CD Engineering: Revolutionize the corporate CI/CD philosophy and internal developer platform (IDP) strategy. Engineer systemic, highly automated pipeline improvements using tools like GitHub Actions, ArgoCD, ACK, CUE, and others to eliminate developer friction, optimize resource utilization, and accelerate the release velocity of hundreds of engineers.
- Enterprise Kubernetes Stewardship: Serve as the ultimate authority on container orchestration, microservices deployment, and service mesh architectures. Drive high-level initiatives to maximize Kubernetes cluster efficiency, auto‑scaling dynamics, secure multi‑tenancy, and immutable deployment patterns across the entire organization.
- Proactive Infrastructure Security & Compliance: Partner closely with SecOps to architect a zero‑trust infrastructure fortress. Embed automated security guardrails and compliance controls into core Infrastructure‑as‑Code pipelines, ensuring seamless, continuous adherence to strict healthcare regulatory standards, including HIPAA and SOC 2 (Do good).
- Enterprise Triage & Resilience Engineering: Act as the premier technical escalation authority for critical, cross‑system infrastructure failures. Spearhead complex, multi‑layered root‑cause analysis, lead deep‑dive post‑mortems, and champion engineering patterns that turn production incidents into systemic resilience upgrades.
Technical Leadership & Culture Multiplication
- Engineering Force Multiplier: Act as a strategic force multiplier across the engineering ecosystem. Direct high‑impact cross‑functional initiatives, define and publish enterprise‑wide best practices, influence organizational technology roadmaps, and mentor senior and staff‑level engineers to foster an inclusive, collaborative culture of technical excel