Trintech is an award‑winning AI‑driven Fintech SaaS organization transforming the way finance teams operate. We are seeking a Principal Cloud Ops Engineer. This senior hands‑on technical leader will oversee and architect a new Microsoft Azure environment, ensuring operational readiness across Kubernetes/AKS, containerized application platforms, CI/CD automation, infrastructure as code, and production operations. The role collaborates with cloud engineering, DevOps, application, security, and infrastructure teams to define scalable Azure platform patterns and implement reliable cloud services.
WHAT YOU WILL DO
- Provide technical oversight and architecture for a new Azure cloud environment supporting Kubernetes and web application deployments.
- Design, implement, and maintain Azure platform capabilities, including AKS, Azure Container Registry, Key Vault, service principals, app registrations, and related cloud services.
- Serve as a principal hands‑on engineer for the implementation, configuration, automation, and operational support of Azure cloud services.
- Architect and support Kubernetes/AKS platforms, including cluster deployment, workload configuration, maintenance, scaling, resiliency, and operational best practices.
- Develop and maintain infrastructure‑as‑code using Terraform to provide repeatable, secure, and consistent cloud environments.
- Design, enhance, and support CI/CD practices using Azure DevOps, pipelines, deployment automation, and release management patterns.
- Provide advanced containerization expertise across Kubernetes deployments, maintenance, image management, runtime configuration, and platform lifecycle activities.
- Partner with application, DevOps, security, and infrastructure teams to define cloud standards, deployment patterns, operational runbooks, and platform documentation.
- Troubleshoot complex cloud, Kubernetes, networking, identity, pipeline, infrastructure, and application deployment issues across Azure environments.
- Evaluate and recommend Azure PaaS offerings where appropriate to improve scalability, reliability, maintainability, and operational efficiency.
- Support monitoring, metrics gathering, alerting, observability, and operational reporting for cloud platforms and hosted workloads.
- Lead technical design discussions, mentor engineers, and provide guidance on cloud operations, platform engineering, automation, and troubleshooting practices.
- Ensure production changes are documented, tested, and aligned with change control, security, compliance, and operational readiness requirements.
- Anticipate technical risks, remove implementation blockers, and, if necessary, escalates issues to protect delivery timelines and platform stability.
WHO YOU ARE
- Bachelor's Degree in Computer Science, Information Systems, Engineering, or equivalent experience.
- 8–12 years of progressive experience in cloud operations, infrastructure engineering, DevOps, systems engineering, platform engineering, or related technical roles.
- Extensive hands‑on experience with Microsoft Azure cloud services and public cloud operating models.
- Strong AKS experience, including Kubernetes cluster deployment, configuration, operations, maintenance, scaling, and troubleshooting.
- Advanced containerization experience with Kubernetes, including workload deployments, platform maintenance, image management, and operational support.
- Strong experience with Terraform and infrastructure‑as‑code practices for cloud environment provisioning and lifecycle management.
- Experience with CI/CD concepts and implementation, including Azure DevOps, pipelines, automated deployments, and release orchestration.
- Hands‑on knowledge of Azure service principals, app registrations, managed identities, Key Vault, Azure Container Registry, and related platform services.
- Knowledge of Azure PaaS offerings and ability to evaluate appropriate use cases for cloud‑native services.
- Strong IT fundamentals across networking, DNS, Active Directory, identity, compute, storage, firewalls, load balancing, and related infrastructure services.
- Advanced troubleshooting skills with the ability to diagnose complex issues across cloud, Kubernetes, networking, identity, pipelines, and application layers.
- Ability to create and maintain technical documentation, standards, runbooks, diagrams, and implementation guidance.
- Excellent written, verbal, and interpersonal communication skills, with the ability to communicate architecture, status, risks, and recommendations to technical and non‑technical stakeholders.
- Ability to work independently, drive complex technical initiatives, mentor others, and influence engineering direction across teams.
NICE TO HAVE
- Experience with monitoring, metrics gathering, observability, alerting, and operational health reporting for cloud platforms.
- Team lead experience or demonstrated ability to mentor engineers, coordinate technical work, and guide implementation efforts across teams.
- Experience supporting SaaS, hosted, or customer‑facing production applications in cloud environments.
- Familiarity with cloud security, governance, policy management, cost management, and reliability engineering practices.
All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin or disability.