Solution Architect – (AI Infrastructure & Hybrid Cloud)

Gruve

Pune District

On-site

INR 2,800,000 - 5,200,000

Full time

4 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Gruve is seeking a Solution Architect – AI Infrastructure & Hybrid Cloud to lead the design and deployment of large-scale OpenShift and Kubernetes environments. You will bridge traditional infrastructure with modern DevOps, driving automation and secure, scalable cloud-native architectures.

The role involves defining high- and low-level designs for multi-tenant clusters across on-prem and hybrid clouds, collaborating with global clients and cross-functional teams to deliver AI-driven solutions

Qualifications

  • Bachelor’s or Master’s degree in computer science, engineering, or a related field.
  • 10+ years of progressive experience in systems architecture, infrastructure engineering, or network DevOps.
  • Expert/Architect-level proficiency in OpenShift (RHOS) and Kubernetes in large-scale production.
  • Experience designing hybrid-cloud OpenShift solutions across AWS, Azure, or Google Cloud.
  • Proven Automation & IaC skills with Ansible, Terraform, and Git-based workflows.
  • Deep understanding of Linux (RHEL) internals and container runtimes.
  • Strong networking knowledge: BGP, VXLAN, EVPN for container workloads.
  • Architecture-as-Code mindset with automated documentation and compliance artifacts.
  • Excellent leadership to manage senior customer relationships and cross-functional teams.

Responsibilities

  • Architect and design enterprise OpenShift solutions; develop HLD and LLD for multi-tenant clusters.
  • Define technology stacks and blueprints for AI solutions across public clouds and on-prem hardware.
  • Oversee end-to-end rollout of services like AI SOC, OpenShift AI, and AI-based cybersecurity.
  • Drive Network DevOps strategy with Ansible, Terraform, and Python for zero-touch provisioning.
  • Lead engagements with global clients; run design workshops and technical reviews.
  • Collaborate with Pre-sales, AI app developers, and Firewall Architects on secure AI deployments.
  • Integrate OpenShift with Cisco ACI/VXLAN and virtualization platforms; optimize hybrid infra.
  • Design CI/CD governance for infrastructure and app delivery with security baked in.
  • Develop observability and reliability stacks using Prometheus, Grafana, and ELK for 99.99% availability.
  • Mentor L2/L3 engineers and establish best practices for DevOps and Network teams.

Skills

OpenShift
Kubernetes
Ansible
Terraform
Python
Networking
Leadership
DevOps
Zero-trust
Security

Education

Bachelor's or Master's in CS/Engineering

Tools

Git
GitHub Actions
Jenkins
CI/CD tooling
OpenShift CLI

Job description

About Gruve

Gruve is an innovative software services startup dedicated to transforming enterprises to AI powerhouses. We specialize in cybersecurity, customer experience, cloud infrastructure, and advanced technologies such as Large Language Models (LLMs). Our mission is to assist our customers in their business strategies utilizing their data to make more intelligent decisions. As a well-funded early-stage startup, Gruve offers a dynamic environment with strong customer and partner networks.


Position Summary:

We are seeking a Solution Architect – (AI Infrastructure & Hybrid Cloud) to lead the strategic design, architecture, and deployment of large-scale, enterprise-grade Red Hat OpenShift and Kubernetes environments. As a technical authority at the L4 level, you will be responsible for defining the blueprint of our cloud-native infrastructure, ensuring it is secure, scalable, and highly automated.
The ideal candidate acts as the bridge between traditional infrastructure and modern DevOps, serving as the lead design authority for global clients. You will collaborate with Network and Firewall Architects to build a unified fabric where containerized workloads, legacy data centers, and hybrid cloud environments coexist seamlessly through advanced automation and Infrastructure-as-Code (IaC).


Key Responsibilities:


  • Architect and Design Enterprise OpenShift Solutions: Lead the high-level design (HLD) and low-level design (LLD) for multi-tenant Red Hat OpenShift and Kubernetes clusters across on-prem and hybrid cloud environments.

  • Define the technology stack, standards, and blueprints for deploying AI solutions across global, multi-region public clouds (AWS/Azure/GCP) and diverse on-premise hardware.

  • Oversee the successful end-to-end rollout of critical services including AI SOC, OpenShift AI, and AI-based Cybersecurity Log Optimization.

  • Drive Network DevOps Strategy: Define and standardize the automation roadmap using Ansible, Terraform, and Python to achieve "Zero-Touch" infrastructure provisioning and configuration.

  • Lead Customer & Stakeholder Engagement: Act as the primary technical consultant for global clients, leading design workshops, architecture validation, and executive-level technical reviews.

  • Integrate Advanced AI apps, Networking & Security: Collaborate with Pre-sales, AI application developers & Engineers, Firewall Architects to design secure AI agents & use cases, container networking (CNI) models, implementing Zero-Trust security, service mesh (Istio), and micro-segmentation within OpenShift environment.

  • Optimize Hybrid Infrastructure: Oversee the seamless integration of OpenShift with physical networking (Cisco ACI, VXLAN) and virtualized platforms (RHEL-V, VMware ESXi).

  • GPU & Hardware Orchestration: Design and manage hardware acceleration using the NVIDIA GPU Operator and Node Feature Discovery (NFD). Implement Multi-Instance GPU (MIG) and time-slicing to optimize resource utilization across multi-tenant clusters.

  • Establish CI/CD Governance: Architect robust CI/CD pipelines (Jenkins, GitLab CI, GitHub Actions) for infrastructure and application delivery, ensuring compliance and security are baked into the workflow.

  • Lead Observability & Reliability: Design comprehensive monitoring and logging architectures using Prometheus, Grafana, and ELK stack to ensure 99.99% availability of cluster services.

  • Mentorship & Technical Leadership: Guide and mentor L2/L3 engineers, providing expert-level escalation support and establishing best practices for the DevOps and Network teams.

  • Innovation & R&D: Evaluate and introduce emerging technologies such as Advanced Cluster Management (ACM), Advanced Cluster Security (ACS), and Cloud-Native Networking (OVN-Kubernetes).


Basic Qualifications:


  • Bachelor’s or master’s degree in computer science, Engineering, or a related field

  • 10+ years of progressive experience in systems architecture, infrastructure engineering, or network DevOps.

  • Expert/Architect-level proficiency in OpenShift (RHOS), and Kubernetes architecture in large-scale production environments.

  • Experience architecting hybrid-cloud OpenShift solutions involving AWS (ROSA), Azure (ARO), or Google Cloud.

  • Proven track record in Automation & IaC: Mastery of Ansible, Terraform, and Git-based workflows to manage complex infrastructures.

  • Deep understanding of Linux (RHEL) Internals: Mastery of kernel networking, storage drivers (CSI), and container runtimes (CRI-O).

  • Strong Network Background: In-depth knowledge of BGP, VXLAN, and EVPN as they apply to connecting containerized workloads to physical data center fabrics.

  • Experience in "Architecture as Code": Ability to develop and maintain compliance artifacts, design validation reports, and automated documentation.

  • Excellent Leadership Skills: Demonstrated ability to manage high-stakes customer relationships and lead cross-functional technical teams


Preferred Qualifications:


  • Certifications: Red Hat Certified Architect (RHCA), Red Hat Certified Specialist in OpenShift (Administration/Development), or CKA/CKS (Certified Kubernetes Administrator/Security), AWS/GCP Certified DevOps Engineer, AWS Certified SysOps Administrator, HashiCorp Certified: Terraform Associate

  • Red Hat Certified OpenShift Administrator, AWS Certified AI Practitioner, Certified Information Systems Security Professional (CISSP), Certified Cloud Security Professional (CCSP)

  • Security Focus: Exposure to DevSecOps tools (e.g., Quay, StackRox) and zero-trust framework implementation

  • Legacy Integration: Familiarity with Cisco ACI, Arista CloudVision, or Juniper Apstra for end-to-end automation integration

  • Red Hat OpenShift AI (ROAI) expertise will be preferred

  • Familiarity with LLM deployment requirements and vector database infrastructure

  • Background in Cybersecurity infrastructure (SIEM, SOAR, SOC & VAPT platforms)

  • Experience with MLOps infrastructure (Kubeflow, MLflow) and high-speed telemetry pipelines


Why Gruve

At Gruve, we foster a culture of innovation, collaboration, and continuous learning.


We are committed to building a diverse and inclusive workplace where everyone can thrive and contribute their best work.


If you’re passionate about technology and eager to make an impact, we’d love to hear from you.


Gruve is an equal opportunity employer.


We welcome applicants from all backgrounds and thank all who apply; however, only those selected for an interview will be contacted.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Security Consultant (Red Hat)
Senior Security Consultant (Red Hat)

Gruve • Maharashtra

Hybrid
INR 4,000,000 - 6,400,000
Senior Network Consultant
Senior Network Consultant

gruve • Pune District

On-site
INR 1,500,000 - 2,100,000
Network Consultant - L2
Network Consultant - L2

Gruve • Pune District

On-site
INR 1,200,000 - 1,800,000
Senior Pre-Sales Consultant - CyberSecurity
Senior Pre-Sales Consultant - CyberSecurity

Gruve • Mumbai

On-site
INR 1,500,000 - 2,500,000
Dynamic work environment
Inclusive workplace culture
Continuous learning opportunities
Software Development Engineer-II (Java, Spring Boot)
Software Development Engineer-II (Java, Spring Boot)

Gruve • Pune District

On-site
INR 1,000,000 - 1,500,000
Innovative and dynamic work environment
Diverse and inclusive workplace
Continuous learning opportunities
Technical Support Engineer
Technical Support Engineer

Gruve • India

On-site
INR 400,000 - 800,000
Network Consultant - L2
Network Consultant - L2

Gruve • Pune District

On-site
INR 800,000 - 1,200,000
Lead Solutions Architect – AI Infrastructure & Private Cloud
Lead Solutions Architect – AI Infrastructure & Private Cloud

At Dawn Technologies • Bengaluru

On-site
INR 2,000,000 - 3,500,000
Competitive, performance-based compensation
Comprehensive healthcare benefits (medical)
Generous paid time off and flexible leave policies
+2
Threat Hunter Analyst
Threat Hunter Analyst

Gruve • Pune District

On-site
INR 1,000,000 - 1,500,000
Dynamic work environment
Commitment to diversity and inclusion
Director Cloud & Infrastructure Architect (Multi-Cloud | Datacenter | SRE)
Director Cloud & Infrastructure Architect (Multi-Cloud | Datacenter | SRE)

Mancer Consulting Services • Bengaluru

On-site
INR 3,500,000 - 6,500,000