Principal, Cloud Engineer

Stage 2 Capital

Hyderabad

On-site

INR 4,000,000 - 7,000,000

Full time

3 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Meals & Snacks
Health Insurance
Wellness Allowance
Paid Leaves
Parental Leave
Company Holidays

Job summary

ModMed is seeking a Principal Site Reliability Engineer to lead reliability, developer velocity, and security for a global healthcare cloud platform. You will define architectural standards, champion observability, and drive CI/CD improvements across Kubernetes and IaC tooling, partnering with product and business teams.

You bring 13–15+ years of SRE/DevOps experience, deep AWS expertise, and strong coding skills in Python/Go.

Qualifications

  • Bachelor's degree in CS/IT/Engineering or related discipline.
  • 13–15+ years in SRE, Cloud Architecture, or DevOps for enterprise scale.
  • Deep AWS expertise (EC2, Lambda, RDS, S3, IAM) with cost optimization.
  • Production-grade Kubernetes, Datadog telemetry, and IaC platforms.
  • Strong Python/Go/Bash engineering skills and executive communication.

Responsibilities

  • Define long-term architectural vision for global AWS cloud ecosystem.
  • Institute observability culture with Datadog and OpenTelemetry.
  • Advance CI/CD and internal developer platform strategies.
  • Lead enterprise Kubernetes initiatives and secure multi-tenancy.
  • Embed security guardrails and compliance in IaC pipelines.
  • Own incident triage and post-mortems to boost resilience.

Skills

AWS
Kubernetes
Datadog
OpenTelemetry
CI/CD
Python
Go
Bash
Terraform
Terragrunt
GitHub Actions
ArgoCD
ACK
CUE

Education

Bachelor's in Computer Science / IT / Engineering

Tools

Datadog
Kubernetes
Terraform
Terragrunt
OpenTelemetry

Job description

Join the Team Modernizing Medicine

At ModMed, we’re not just building software—we’re reimagining the healthcare experience. Founded in 2010 by a practicing physician and a successful tech entrepreneur, we took a radically different approach: we hired doctors and taught them how to code. This "for doctors, by doctors" philosophy has allowed us to create an AI-enabled, specialty-specific cloud platform that places patients at the center of care.

A Culture of Excellence

When You Join ModMed, You’re Joining An Award-winning Team Recognized For Innovation And Employee Satisfaction. From Our Global Headquarters In Boca Raton Florida, And Extensive Employee Base In Hyderabad India, We Are a Team Of 4,500+ Passionate Problem-solvers On a Mission To Increase Medical Practice Success And Improve Patient Outcomes

  • Consistently ranked as a Top Place to Work
  • 2025 Globee Business Awards: Gold Globee for “Technology Team of the Year”
  • 2025 Black Book Awards: Ranked #1 EHR in 11 Specialties
  • Florida Venture Forum: Venture-Backed Company of the Year
Our Mission & Vision
  • Our Mission: To place doctors and patients at the center of care through an intelligent, specialty-specific cloud platform.
  • Our Vision: A world where the software ModMed builds increases medical-practice success and improves patient outcomes.
  • Our Core Values: Create customer delight | Save time | Innovate boldly, then make things happen | Align passion with purpose | Think big | Have fun | Do good.
The Role:

Technical Visionary for Reliability and Developer Velocity As a Principal Site Reliability Engineer at ModMed, you are a primary architect of our technical future. You don't just solve problems; you anticipate the needs of a global healthcare platform years in advance. You will think big to define the standards for reliability, scalability, developer velocity, and security that allow our doctors to provide world‑class care without interruption. This is a high‑impact technical leadership role where you will innovate boldly, then make things happen, acting as a vital bridge between engineering, product, and business goals. By aligning passion with purpose, you will build a resilient infrastructure that serves as the backbone for modern medicine and consistently creates customer delight.

Primary Responsibilities
  • Architectural Strategy & Technical Governance: Define and execute the long-term architectural vision for our global AWS cloud ecosystem. Design high-performance, fault-tolerant, and cost-optimized multi-region topologies capable of scaling elastically to meet massive transaction volumes while enforcing rigorous cloud hygiene (Save time).
  • Systemic Observability Ecosystems: Institutionalize "Observability as a Culture" across the enterprise. Architect unified, global telemetry standards using Datadog and OpenTelemetry, transforming raw logs, metrics, and distributed traces into actionable, predictive insights that elevate system reliability across all product engineering groups.
  • Developer Platform Evolution & CI/CD Engineering: Revolutionize the corporate CI/CD philosophy and internal developer platform (IDP) strategy. Engineer systemic, highly automated pipeline improvements using tools like GitHub Actions, ArgoCD, ACK, CUE, and others to eliminate developer friction, optimize resource utilization, and accelerate the release velocity of hundreds of engineers.
  • Enterprise Kubernetes Stewardship: Serve as the ultimate authority on container orchestration, microservices deployment, and service mesh architectures. Drive high-level initiatives to maximize Kubernetes cluster efficiency, auto-scaling dynamics, secure multi-tenancy, and immutable deployment patterns across the entire organization.
  • Proactive Infrastructure Security & Compliance: Partner closely with SecOps to architect a zero‑trust infrastructure fortress. Embed automated security guardrails and compliance controls into core Infrastructure-as-Code pipelines, ensuring seamless, continuous adherence to strict healthcare regulatory standards, including HIPAA and SOC 2 (Do good).
  • Enterprise Triage & Resilience Engineering: Act as the premier technical escalation authority for critical, cross-system infrastructure failures. Spearhead complex, multi-layered root-cause analysis, lead deep-dive post-mortems, and champion engineering patterns that turn production incidents into systemic resilience upgrades.
Technical Leadership & Culture Multiplication
  • Engineering Force Multiplier: Act as a strategic force multiplier across the engineering ecosystem. Direct high-impact cross-functional initiatives, define and publish enterprise-wide best practices, influence organizational technology roadmaps, and mentor senior and staff-level engineers to foster an inclusive, collaborative culture of technical excellence (Have fun. Do good.).
Role Requirements & Qualifications Minimum Qualifications
  • Education: Bachelor’s degree in Computer Science, Information Technology, Engineering, or a related discipline with an IT focus.
  • Experience: 8+ years of dedicated experience in cloud engineering, Cloud Engineering, or DevOps roles.
  • Engineering Tenure: 13–15+ years of distinguished experience in Site Reliability Engineering (SRE), Cloud Architecture, or DevOps Platform Engineering, with a definitive track record of scaling high-concurrency, enterprise cloud environments.
  • AWS Architectural Authority: Deep, authoritative mastery of the AWS ecosystem (including EC2, Lambda, RDS, S3, IAM, and advanced networking features) with proven experience executing massive scalability strategies and sophisticated cloud spend optimization structures.
  • Advanced Platform Stack: Professional-grade expertise architecting and managing production-scale Kubernetes environments, comprehensive Datadog telemetry architectures, and modular Infrastructure-as-Code platforms utilizing Terraform or Terragrunt at scale.
  • Strategic Engineering & Tooling: Expert-level software engineering capability in Python, Go, or Bash to construct intelligent automation engines, internal CLI tools, and infrastructure platforms that systematically eliminate operational toil.
  • Executive-Level Influence: Exceptional communication and advisory skills, with a proven ability to translate complex, low-level technical trade-offs into clear strategic visions for executive leadership while remaining capable of conducting a rigorous, line-by-line code and architecture review.
  • Mission Dedication: A strong desire to apply elite technical expertise toward modernizing healthcare infrastructure, protecting patient data integrity, and solving complex societal-scale challenges.
Preferred Qualifications
  • Education: Bachelor's in Engineering or MCA.
  • Industry Credentials: Active AWS Certified Solutions Architect – Professional, AWS Certified DevOps Engineer – Professional, or relevant advanced Specialty certifications (Security, Advanced Networking).
  • Distributed Systems Expertise: Direct experience managing high-throughput, event-driven streaming architectures (e.g., Apache Kafka, AWS Kinesis) or high-scale data warehouse solutions (e.g., Amazon Redshift, Snowflake).
  • Community & Thought Leadership: A history of contributing to foundational open-source cloud-native projects, authoring technical white papers, or speaking at premier industry tech conferences (e.g., AWS re:Invent, KubeCon). ModMed Core Behavioral Competencies To thrive in this role, you should naturally embody our foundational competencies:
  • Agility: Embraces change as a growth opportunity; learns from successes and failures and adapts to new challenges.
  • Results Driven: Consistently achieves quality results, overcomes obstacles, and inspires others to do the same.
  • Company Culture: Actively contributes to our culture to keep this company awesome, acts with a values-first mentality, embraces differences, and creates a sense of belonging for all.
  • Customer Focus: Builds strong, positive relationships (internally and externally) and delivers customer-centric solutions.
ModMed Benefits Highlight
  • Meals & Snacks: Enjoy complimentary office lunches & dinners on select days and healthy snacks delivered to your desk,
  • Insurance Coverage: Comprehensive health, accidental, and life insurance plans, including coverage for family members, all at no cost to employees,
  • Allowances: Annual wellness allowance to support your well-being and productivity,
  • Earned, casual, and sick leaves to maintain a healthy work-life balance,
  • Bereavement leave for difficult times and extended medical leave options,
  • Paid parental leaves, including maternity, paternity, adoption, surrogacy, and abortion leave,
  • Celebration leave to make your special day even more memorable, and company-paid holidays to recharge and unwind.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Principal, Cloud Engineer
Principal, Cloud Engineer

Enboarder • Hyderabad

On-site
INR 4,000,000 - 7,000,000
Meals & Snacks
Comprehensive health insurance
Annual wellness allowance
+1
Principal Cloud Engineer
Principal Cloud Engineer

ModMed • Hyderabad

On-site
INR 400,000 - 680,000
Meals & Snacks
Insurance Coverage
Staff Site Reliability Engineer
Staff Site Reliability Engineer

ModMed India • Hyderabad

On-site
INR 4,000,000 - 7,000,000
Senior Site Reliability Engineer 2
Senior Site Reliability Engineer 2

ModMed • Hyderabad

On-site
INR 4,000,000 - 6,000,000
Meals & Snacks
Insurance Coverage
Wellness Allowance
+1
Principal, Cloud Engineer
Principal, Cloud Engineer

ModMed • Hyderabad

On-site
INR 3,000,000 - 6,000,000
Principal, Cloud Engineer
Principal, Cloud Engineer

ModMed Technologies India Private Limited • Hyderabad

On-site
INR 5,000,000 - 8,000,000
Principal, Cloud Engineer
Principal, Cloud Engineer

ModMed India • Hyderabad

On-site
INR 4,500,000 - 6,500,000
Senior Software Engineer 1-2
Senior Software Engineer 1-2

Enboarder • Hyderabad

On-site
INR 1,800,000 - 2,800,000
Meals & Snacks
Insurance Coverage
Annual wellness allowance
+2
Principal, Cloud Engineer
Principal, Cloud Engineer

modmed • India

On-site
INR 3,500,000 - 5,500,000
Technical Product Owner
Technical Product Owner

ModMed • Hyderabad

On-site
INR 2,800,000 - 4,200,000
Meals & Snacks
Insurance Coverage
Wellness allowance
+1