AI Operations Engineering Technical Leader

123 Cisco Systems (India) Private Limited

Bengaluru

On-site

INR 3,500,000 - 6,500,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Cisco Systems (India) Private Limited in Bengaluru seeks a senior platform engineer to deploy and manage a Life Cycle Management system with Splunk across engineering labs, designing dashboards, and integrating data feeds. You will leverage AI technologies to generate actionable insights and ensure high availability of monitoring resources.

You will collaborate with cross‑functional teams to align observability with business goals, document processes, and drive best practices to improve

Qualifications

  • Bachelor's degree in Computer Science, Electronics & Communications Engineering, Information Technology, or a related technical discipline.
  • Minimum 8–10 years of professional experience in enterprise software engineering, Site Reliability Engineering (SRE), Platform Engineering, DevOps, or related technical leadership roles.
  • Strong experience designing, developing, and supporting large-scale enterprise applications using Java, Spring Framework, Spring Boot, REST APIs, and modern application architectures.
  • Experience with Kubernetes, Docker, Linux/Unix environments, CI/CD pipelines, and cloud-native deployment models.
  • Hands‑on experience with AI technologies including LangGraph, MCP, RAG architectures, Vector Databases, LangChain, Prompt Engineering, or similar AI frameworks.
  • Strong experience with Oracle databases, SQL optimization, PL/SQL, ETL frameworks, and enterprise data integration.
  • Experience designing dashboards, operational metrics, monitoring solutions, and platform observability.
  • Strong knowledge of DevOps tools including Git, Jenkins, Maven, Jira, and CI/CD automation.
  • Demonstrated ability to troubleshoot production issues, perform root cause analysis, and improve system reliability.
  • Excellent verbal and written communication skills with the ability to collaborate effectively across global engineering organizations.
  • Ability to lead technical initiatives in a fast‑paced, globally distributed engineering environment.

Responsibilities

  • You will be primarily responsible for the deployment, configuration, and continuous management of Life Cycle Management system that involves Splunk across engineering labs to monitor non-production R&D resources.
  • You will design and optimize dashboards, integrate diverse data feeds, and use AI to analyze alerting data for actionable insights.
  • Day-to-day activities include working with cross-functional teams to align observability solutions with business objectives, documenting monitoring processes, and participating in global discussions and planning.
  • They will also contribute to the implementation of best practices, drive performance improvements, and support the overall reliability and efficiency of lab infrastructure.
  • Success in this role will be measured by uptime and reliability of monitoring systems and dashboards, efficiency of life cycle management, and the quality of actionable insights.

Skills

Java
Spring Framework
Spring Boot
REST APIs
Kubernetes
Docker
Linux/Unix
CI/CD
Cloud-native
AI technologies
Oracle databases
SQL & ETL
Git/Jenkins

Education

Bachelor's degree in Computer Science or related field
Master's degree (preferred)

Tools

Splunk
LangGraph
LangChain
Vector Databases

Job description

Meet the Team Global Lab Solutions (GLS) is a centralized, globally focused team responsible for standardizing best practices, driving innovation, and ensuring operational excellence across Engineering Labs worldwide. Our team is structured around core service offerings such as compliance (ISO-9001/14001, calibration, ESD), system and lab administration, infrastructure monitoring, and tool support. Specialized areas like Engineering Inventory Management, Lab Space Planning, and Lab Information Security are managed by sub-functional experts within the team. The primary objectives of GLS include improving lab efficiency, ensuring compliance and safety, enhancing observability through technologies like Splunk, and supporting scalability and reliability of lab operations via AI-driven insights and data-driven decision-making. Our key stakeholders are engineering teams, lab operations personnel, and leadership across global R&D and innovation hubs. We work closely with internal teams such as IT, Facilities, InfoSec, Procurement, Asset Management, and Engineering leadership to ensure alignment of lab services with business and technical goals.

Your Impact

You will be primarily responsible for the deployment, configuration, and continuous management of Life Cycle Management system that involves Splunk across engineering labs to monitor non-production R&D resources. You will design and optimize dashboards, integrate diverse data feeds, and use AI to analyze alerting data for actionable insights.

Day-to-day activities include working with cross-functional teams to align observability solutions with business objectives, documenting monitoring processes, and participating in global discussions and planning. They will also contribute to the implementation of best practices, drive performance improvements, and support the overall reliability and efficiency of lab infrastructure.

Success in this role will be measured by:

  • Uptime and reliability of monitoring systems and dashboards
  • Efficient life cycle management measured in terms of quarterly $ savings
  • Number and quality of actionable insights generated through Splunk and AI
  • Timeliness and effectiveness of integrating new data feeds
  • Feedback from internal stakeholders on system usability and value
  • Documentation completeness and adherence to best practices
  • Proactive identification and remediation of vulnerabilities and performance bottlenecks
  • Cross-functional engagement and alignment with organizational goals
  • Meeting or exceeding KPIs related to system observability, user adoption, and operational impact will define performance excellence in this role
Minimum Qualifications
  • Bachelor’s degree in computer science, Electronics & Communications Engineering, Information Technology, or a related technical discipline.
  • Minimum 8–10 years of professional experience in enterprise software engineering, Site Reliability Engineering (SRE), Platform Engineering, DevOps, or related technical leadership roles.
  • Strong experience designing, developing, and supporting large-scale enterprise applications using Java, Spring Framework, Spring Boot, REST APIs, and modern application architectures.
  • Experience with Kubernetes, Docker, Linux/Unix environments, CI/CD pipelines, and cloud-native deployment models.
  • Hands‑on experience with AI technologies including LangGraph, MCP, RAG architectures, Vector Databases, LangChain, Prompt Engineering, or similar AI frameworks.
  • Strong experience with Oracle databases, SQL optimization, PL/SQL, ETL frameworks, and enterprise data integration.
  • Experience designing dashboards, operational metrics, monitoring solutions, and platform observability.
  • Strong knowledge of DevOps tools including Git, Jenkins, Maven, Jira, and CI/CD automation.
  • Demonstrated ability to troubleshoot production issues, perform root cause analysis, and improve system reliability.
  • Excellent verbal and written communication skills with the ability to collaborate effectively across global engineering organizations.
  • Ability to lead technical initiatives in a fast‑paced, globally distributed engineering environment.
Preferred Qualifications
  • Master’s degree in Computer Science, Software Engineering, Artificial Intelligence, or a related technical field.
  • Experience building AI‑enabled enterprise platforms using LangGraph, MCP, Vector Databases, RAG architectures, and LLM orchestration frameworks.
  • Experience with enterprise observability platforms such as Splunk and operational analytics.
  • Experience architecting reusable ETL frameworks and large‑scale enterprise data platforms.
  • Strong understanding of distributed systems, microservices, API integrations, and enterprise architecture.
  • Experience mentoring engineers and leading technical architecture discussions.
  • Familiarity with engineering laboratory environments, asset management systems, inventory platforms, or enterprise operational systems.
  • Experience supporting globally distributed production environments with high availability and operational excellence requirements.
  • Cisco AI training, cloud‑native certifications, Java certifications, or related professional certifications are highly desirable.
Why Cisco?

At Cisco, we’re revolutionizing how data and infrastructure connect and protect organizations in the AI era – and beyond. We’ve been innovating fearlessly for 40 years to create solutions that power how humans and technology work together across the physical and digital worlds. These solutions provide customers with unparalleled security, visibility, and insights across the entire digital footprint. Fueled by the depth and breadth of our technology, we experiment and create meaningful solutions. Add to that our worldwide network of doers and experts, and you’ll see that the opportunities to grow and build are limitless. We work as a team, collaborating with empathy to make really big things happen on a global scale. Because our solutions are everywhere, our impact is everywhere. We are Cisco, and our power starts with you. Why Cisco? At Cisco, we’re revolutionizing how data and infrastructure connect and protect organizations in the AI era – and beyond. We’ve been innovating fearlessly for 40 years to create solutions that power how humans and technology work together across the physical and digital worlds. These solutions provide customers with unparalleled security, visibility, and insights across the entire digital footprint. Fueled by the depth and breadth of our technology, we experiment and create meaningful solutions. Add to that our worldwide network of doers and experts, and you’ll see that the opportunities to grow and build are limitless. We work as a team, collaborating with empathy to make really big things happen on a global scale. Because our solutions are everywhere, our impact is everywhere. We are Cisco, and our power starts with you. Cisconians power the future. We make impact as a team, innovating fast and fearlessly to create meaningful solutions on a large scale. The depth and breadth of our technology doesn't just benefit our customers – it also means limitless opportunities for us to experiment and learn. We understand the power each of our unique backgrounds bring when we work together. Because of that, we have a global network of thinkers, doers, experts, and curious creators who help one another do their life’s best work.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Operations Engineering Technical Leader
AI Operations Engineering Technical Leader

Cisco • Bengaluru

On-site
INR 3,500,000 - 5,500,000
Data Engineering Application Developer - Python, API, Snowflake/Redshift, AWS services, Tableau/PowerBI, React/Angular. HTML, CSS, GitHub - 7 to 10 years - Bangalore
Data Engineering Application Developer - Python, API, Snowflake/Redshift, AWS services, Tableau/PowerBI, React/Angular. HTML, CSS, GitHub - 7 to 10 years - Bangalore

123 Cisco Systems (India) Private Limited • Bengaluru

On-site
INR 2,500,000 - 3,500,000
Software Engineer- Agentic AI Developer- 6+yrs exp
Software Engineer- Agentic AI Developer- 6+yrs exp

123 Cisco Systems (India) Private Limited • Bengaluru

On-site
INR 3,500,000 - 7,000,000
Leader, Software Engineering - Go/Python - Container/Kubernetes - SaaS/Public cloud - AI - 12 to 16 years - Bangalore
Leader, Software Engineering - Go/Python - Container/Kubernetes - SaaS/Public cloud - AI - 12 to 16 years - Bangalore

123 Cisco Systems (India) Private Limited • Bengaluru

On-site
INR 4,000,000 - 6,000,000
Lead Site Reliability Engineer, Data Platform - AI/ML Infrastructure
Lead Site Reliability Engineer, Data Platform - AI/ML Infrastructure

Cisco Systems, Inc. • Bengaluru

On-site
INR 3,500,000 - 7,000,000
Software Engineer- Backend 8+ years (Golang, Python)
Software Engineer- Backend 8+ years (Golang, Python)

123 Cisco Systems (India) Private Limited • Bengaluru

On-site
INR 4,000,000 - 6,500,000
Software Engineer- Backend (Golang, Python)
Software Engineer- Backend (Golang, Python)

123 Cisco Systems (India) Private Limited • Bengaluru

On-site
INR 1,500,000 - 2,100,000
Software Engineer- Agentic AI Developer- 6+yrs exp
Software Engineer- Agentic AI Developer- 6+yrs exp

Cisco • Bengaluru

On-site
INR 4,500,000 - 7,500,000
Software Engineer (AI)
Software Engineer (AI)

Cisco Systems, Inc. • Bengaluru

On-site
INR 1,200,000 - 1,800,000
Software Engineering Technical Leader | Python/GO | Agentic AI | APIs | Generative AI | Kubernetes/Docker | MCP/A2A - 8 to 12 - Bangalore
Software Engineering Technical Leader | Python/GO | Agentic AI | APIs | Generative AI | Kubernetes/Docker | MCP/A2A - 8 to 12 - Bangalore

123 Cisco Systems (India) Private Limited • Bengaluru

On-site
INR 6,000,000 - 8,500,000