Cloud, Data Science & AI Architect

Capgemini

Bengaluru

On-site

INR 2,000,000 - 3,200,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Capgemini Bengaluru is seeking an Enterprise Cloud and Data Architect to lead end-to-end architecture across cloud, data, analytics, ML and Generative AI platforms. You will design cloud-native, microservices-based architectures on AWS/Azure/GCP; define scalable data architectures; lead the deployment of ML solutions and MLOps.

Collaboration with data scientists, product teams and business stakeholders will deliver scalable, secure and compliant solutions with a strong focus on governance and

Qualifications

  • Extensive experience designing scalable, secure enterprise cloud and data platforms.
  • Proven ability to architect data-lake, data-warehouse and lakehouse solutions.
  • Hands-on with ML lifecycle management and MLOps toolchains.
  • Strong knowledge of CI/CD, IaC and security governance.
  • Experience with AWS, Azure or Google Cloud and modern data technologies.

Responsibilities

  • Define and own end-to-end architecture for cloud, data, analytics and AI platforms.
  • Lead development of scalable cloud and data platforms for BI and AI initiatives.
  • Design cloud-native, microservices architectures across AWS/Azure/GCP.
  • Define scalable data architectures for batch, streaming and real-time processing.
  • Architect data-lake, lakehouse and related platforms.
  • Lead ML model development, deployment, monitoring and lifecycle management.
  • Define and implement MLOps architectures with MLflow, Azure ML or SageMaker.
  • Design Generative AI solutions using LLMs, retrieval-augmented generation, vector databases.
  • Lead observability, reliability, cost-optimisation and security governance.
  • Collaborate with stakeholders to define roadmaps and phased implementations.

Skills

Cloud architecture
Data architecture
MLOps
Generative AI
AI platforms
Security & governance
API design
Kubernetes
Terraform
CI/CD
Python
SQL
Java/Node.js

Tools

Databricks
Snowflake
BigQuery
Synapse
Power BI
Looker
Tableau
Kafka
Kinesis
MLflow
Azure ML
SageMaker
LangChain
Pinecone
Weaviate
OpenAI API

Job description

Roles Responsibilities
  • Define and own the end-to-end architecture for enterprise cloud, data, analytics, machine learning and Generative AI platforms.
  • Architect and lead the development of scalable cloud and data platforms supporting digital transformation, business intelligence, advanced analytics and AI initiatives.
  • Design cloud-native, distributed and microservices-based solution architectures on AWS, Microsoft Azure or Google Cloud Platform.
  • Define scalable data architectures for batch, streaming, event-driven and real-time processing workloads.
  • Design enterprise data platforms covering data ingestion, transformation, storage, metadata management, governance, analytics and consumption.
  • Architect data-lake, data-warehouse and lakehouse solutions using platforms such as Databricks, Snowflake, Microsoft Fabric, BigQuery, Synapse or equivalent technologies.
  • Design cloud-native data products, APIs and reusable services that enable business intelligence, advanced analytics and AI applications.
  • Lead the architecture and deployment of machine-learning solutions, including model development, feature engineering, deployment, monitoring, retraining and lifecycle management.
  • Define and implement MLOps architectures using platforms and tools such as MLflow, Azure Machine Learning, Amazon SageMaker or equivalent technologies.
  • Lead the design of Generative AI solutions using large language models, retrieval-augmented generation, vector databases, prompt engineering and agent-based frameworks.
  • Define AI orchestration patterns for intelligent assistants, copilots, autonomous agents and domain-specific AI applications.
  • Design and optimise enterprise-grade data pipelines to ensure reliable, scalable and high-quality data processing.
  • Architect streaming and real-time analytics solutions using Apache Kafka, Amazon Kinesis, Apache Pulsar, Apache Flink, Spark Streaming or equivalent technologies.
  • Establish data-modelling, indexing, partitioning, caching and performance-optimisation standards for relational, NoSQL and analytical data stores.
  • Design integration frameworks using APIs, event-driven architectures, messaging platforms and enterprise data ecosystems.
  • Work closely with data scientists, data engineers, cloud engineers, software architects, product teams and business stakeholders to deliver end-to-end solutions.
  • Establish architecture standards and best practices for cloud engineering, data engineering, DataOps, MLOps, DevSecOps, security, governance and operational excellence.
  • Define modern CI/CD, automated-testing and Infrastructure-as-Code practices using Terraform, CloudFormation, Bicep or equivalent technologies.
  • Ensure that cloud, data and AI solutions comply with enterprise requirements for security, privacy, regulatory compliance, data sovereignty and responsible AI.
  • Define observability, monitoring, reliability, high-availability, disaster-recovery and cost-optimisation strategies.
  • Evaluate emerging cloud, analytics, data science and AI technologies and recommend appropriate enterprise adoption strategies.
  • Conduct architecture assessments, technology evaluations, proofs of concept and solution trade-off analyses.
  • Collaborate with business and technology stakeholders to define technical roadmaps, target-state architectures and phased implementation strategies.
Job Description - Grade Specific
  • Extensive experience designing scalable, secure and highly available enterprise solutions on AWS, Microsoft Azure or Google Cloud Platform.
  • Strong understanding of cloud-native, distributed, event-driven and microservices-based architectures.
  • Deep expertise in designing and implementing enterprise data platforms covering ingestion, processing, storage, governance, analytics and data consumption.
  • Strong experience with data-lake, data-warehouse and lakehouse architecture patterns.
  • Hands-on experience with data-engineering platforms such as Apache Spark, Databricks, Snowflake, Google BigQuery, Azure Synapse Analytics, Microsoft Fabric or equivalent technologies.
  • Strong experience with relational, NoSQL and analytical databases, including data modelling, indexing, partitioning and performance optimisation.
  • Experience designing batch, near-real-time and real-time data-processing solutions.
  • Hands-on experience with streaming platforms such as Apache Kafka, Amazon Kinesis, Apache Pulsar, Apache Flink or Spark Streaming.
  • Strong understanding of machine-learning and AI lifecycle management, including data preparation, model development, validation, deployment, monitoring, retraining and governance.
  • Experience designing and implementing enterprise MLOps platforms and practices.
  • Hands-on experience with tools such as MLflow, Azure Machine Learning, Amazon SageMaker or equivalent platforms.
  • Strong experience building Generative AI applications using large language models and retrieval-augmented generation architectures.
  • Experience with prompt engineering, model orchestration, grounding, evaluation, guardrails and responsible-AI practices.
  • Experience designing agent-based and multi-agent AI solutions using frameworks such as LangChain, LangGraph, Semantic Kernel or equivalent technologies.
  • Experience with vector databases and semantic-search platforms such as Pinecone, Weaviate, Azure AI Search, OpenSearch, pgvector or equivalent technologies.
  • Proficiency in Python and SQL, together with working knowledge of at least one additional language such as Java, Golang or Node.js.
  • Experience developing and deploying cloud-native APIs, microservices and data services.
  • Strong understanding of API management, service integration and event-driven integration patterns.
  • Experience with Kubernetes, Docker, serverless computing and container-based deployment architectures.
  • Familiarity with modern CI/CD, DataOps, MLOps, Infrastructure as Code and DevSecOps practices.
  • Hands-on experience with Terraform, CloudFormation, Bicep or equivalent automation technologies.
  • Strong knowledge of enterprise data governance, metadata management, lineage, data quality, master-data management and access controls.
  • Experience with cloud and data security, including encryption, identity and access management, key management, network security and secure data sharing.
  • Understanding of regulatory, privacy and compliance requirements applicable to enterprise data and AI platforms.
  • Familiarity with business-intelligence and visualisation platforms such as Power BI, Tableau or Looker.
  • Experience in the energy, utilities, manufacturing, rail, industrial or other asset-intensive industries would be advantageous.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Solution Architect
Solution Architect

Michelin • Pune District

On-site
INR 2,500,000 - 4,500,000
Solution Architect- Data And AI
Solution Architect- Data And AI

Michelin Americas Research • Pune District

On-site
INR 4,000,000 - 9,000,000
Data Architect
Data Architect

Luxoft • Gurugram District

On-site
INR 2,500,000 - 4,200,000
Data Architect
Data Architect

Luxoft • Hyderabad

On-site
INR 1,800,000 - 2,800,000
Solution Architect- Data and AI
Solution Architect- Data and AI

Michelin España Portugal SA • Pune District

On-site
INR 5,000,000 - 7,000,000
Solution Architect- Data and AI
Solution Architect- Data and AI

Michelin • Pune District

On-site
INR 3,000,000 - 6,000,000
Data Architect
Data Architect

Luxoft • Dadri

On-site
INR 1,800,000 - 3,000,000
AI Data Architect
AI Data Architect

Tata Consultancy Services • Bengaluru

On-site
INR 2,500,000 - 4,000,000
Data Architecture Associate Manager
Data Architecture Associate Manager

Accenture in India • Bengaluru

On-site
INR 4,000,000 - 6,400,000
Solution Architect- Data and AI
Solution Architect- Data and AI

MICHELIN France • Pune District

On-site
INR 4,200,000 - 7,000,000