AI Infrastructure Architecture - Manager

Accenture PLC

Brussel Hoofdstad

Hybrid

EUR 120,000 - 180,000

Full time

9 days ago
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Accenture PLC seeks a Lead/Principal Infrastructure Architect to own end-to-end compute infrastructure for large-scale AI/ML systems. You will translate business goals and SLAs into scalable, cost-efficient architectures across compute, networking, storage, and orchestration.

You will lead architecture reviews, set technical standards, and mentor engineers while collaborating with clients to align infrastructure with business outcomes and regulatory requirements.

Qualifications

  • Significant experience in coding, building, monitoring, troubleshooting AI/ML applications and deploying on premises or in public clouds.
  • Strong understanding of AI and machine learning concepts.
  • Deep knowledge of computing infrastructure and AI infrastructure.
  • Proficiency in programming languages such as Python, Java, or C++.
  • Experience with data pipelines and workflow tools (Airflow, Kubeflow).
  • Strong problem-solving skills and ability to work in a fast-paced environment.
  • Excellent communication and collaboration skills.
  • Proven experience in AI/ML infrastructure engineering on a hyperscaler platform for large-scale solutions.
  • Experience in leading AI projects and teams.
  • Strong project management skills and ability to manage multiple projects.
  • Demonstrated ability to evaluate and select AI technologies and frameworks.

Responsibilities

  • Own the end-to-end architecture and design of optimized compute infrastructure for large-scale AI/ML systems, from concept through delivery.
  • Develop and evaluate architecture alternatives, weighing trade-offs across compute, networking, storage, orchestration, and model serving.
  • Lead architecture assessments and reviews, identifying gaps and optimization opportunities, and recommending remediation.
  • Drive architectural decisions with documented rationale, aligned to SLAs and standards.
  • Define and maintain the AI infrastructure roadmap with capacity and evolution planning.
  • Architect and optimize the computational stack for performance, power, cost, and scalability.
  • Design and tune large-scale GPU clusters and distributed training systems, including interconnects and storage.
  • Serve as the authoritative AI infrastructure expert on at least one hyperscaler cloud.
  • Design deployment, automation, CI/CD strategies for production releases of AI systems and data pipelines.
  • Establish AI monitoring and observability strategy across InfraOps and MLOps with SLAs and cost tracking.
  • Integrate AI/ML systems into enterprise environments with security and regulatory compliance.
  • Lead capacity planning and cost modeling to optimize engineering cost-efficiency.
  • Collaborate with clients and teams to translate requirements into architecture and standards.
  • Set technical direction, standards, and best practices, mentoring engineers and leading reviews.

Skills

AI/ML infra
Python/Java/C++
Hyperscaler cloud
CI/CD automation
Architecture design
Leadership & mentoring
Security & compliance

Education

Bachelor's Degree in CS/CE/Engineering

Tools

Apache Airflow
Kubeflow

Job description

YOU ARE

As a Lead and Principal Infrastructure Architect, you own end-to-end responsibility for designing optimized compute infrastructure for large-scale AI and machine learning systems, including large-scale distributed training environments. You are the authority who translates business goals, SLAs, and client standards into infrastructure architectures that perform at scale while being deliberately engineered for cost-efficiency. Drawing on deep experience, you weigh multiple viable solutions for any given problem — across compute, networking, storage, orchestration, and model serving — and make rational, well-justified architectural decisions tailored to each client's situation, constraints, and standards. You architect and optimize the full computational stack for performance, power, cost, and scalability; design and tune large-scale GPU clusters and distributed training systems; and ensure infrastructure meets security, compliance, and regulatory requirements. As the recognized AI infrastructure expert in at least one hyperscaler cloud (such as AWS, Azure, or Google Cloud), you bring authoritative knowledge of that platform's AI/ML services, accelerators, networking, and cost levers, and apply it to deliver best-in-class solutions. Beyond design, you set technical direction and standards, lead and mentor engineers and architects, partner with clients and stakeholders to shape the infrastructure roadmap, and are ultimately accountable for delivering AI/ML infrastructure that meets business SLAs, controls cost, and scales to enterprise and frontier workloads.

THE WORK
  • Own the end-to-end architecture and design of optimized compute infrastructure for large-scale AI/ML systems, including large-scale distributed training environments, from concept through delivery.
  • Develop and evaluate architecture alternatives, weighing trade-offs across compute, networking, storage, orchestration, and model serving to make rational, well-justified decisions tailored to each client's situation and standards.
  • Lead architecture assessments and reviews of existing and proposed environments, identifying gaps, risks, bottlenecks, and optimization opportunities, and recommending remediation.
  • Drive architectural decision-making, documenting rationale, trade-offs, and assumptions so decisions are transparent, defensible, and aligned with business SLAs and standards.
  • Define and maintain the AI infrastructure roadmap, planning capacity, scaling, and technology evolution in step with business and product goals.
  • Architect and optimize the full computational stack for performance, power, cost, and scalability, ensuring infrastructure meets business SLAs while being deliberately engineered for cost-efficiency.
  • Design and tune large-scale GPU clusters and distributed training systems, including accelerator selection, interconnect/networking, and storage for high-throughput training workloads.
  • Serve as the authoritative AI infrastructure expert in at least one hyperscaler cloud (AWS, Azure, or GCP), applying deep knowledge of its AI/ML services, accelerators, networking, and cost levers.
  • Design deployment, automation, and CI/CD strategies for reliable, repeatable, and scalable releases of AI systems, models, and data pipelines into production.
  • Establish AI monitoring and observability strategy across InfraOps and MLOps, defining SLAs, SLOs, alerting, and performance/cost tracking, and driving continuous optimization.
  • Integrate AI/ML systems into enterprise environments, ensuring interoperability, security, compliance, and adherence to regulatory and client standards.
  • Lead capacity planning and cost modeling, forecasting compute needs and engineering cost-efficiency into the architecture without compromising performance.
  • Collaborate with clients, stakeholders, and engineering teams to align infrastructure decisions with business outcomes, translating requirements into actionable architecture and standards.
  • Set technical direction, standards, and best practices, mentoring engineers and architects and leading design and code reviews across the team.
EDUCATION
  • Bachelor's Degree in Computer Science, Computer Engineering, related Engineering field
BASIC (REQUIRED) QUALIFICATION
  • Significant experience in coding, building, monitoring, troubleshooting applications of AI/ML models; selecting, designing and infrastructure for deploying and running them on premise or on public cloud.
  • Strong understanding of AI and machine learning as a subject.
  • Strong understanding of computing infrastructure a subject, preferred knowledge of AI infrastructure.
  • Full proficiency in programming languages such as Python, Java, or C++.
  • Experience with data pipeline and workflow management tools (e.g., Apache Airflow, Kubeflow).
  • Strong problem-solving skills and ability to work in a fast-paced environment.
  • Excellent communication and collaboration skills.
  • Proven experience in AI/ML infrastructure engineering or related roles on a hyperscaler platform for deploying large scale solutions.
  • Experience in leading and managing AI projects and teams.
  • Strong project management skills, with the ability to manage multiple projects simultaneously.
  • Demonstrated experience in evaluating and selecting AI technologies and frameworks.
  • Ability to work with cross-functional teams and drive project alignment.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI Infrastructure Architecture - Manager
AI Infrastructure Architecture - Manager

Accenture PLC • Brussel

On-site
EUR 120,000 - 180,000
AI Infrastructure Architecture - Manager
AI Infrastructure Architecture - Manager

3700 Accenture N.V. Company • Sint-Jans-Molenbeek

On-site
EUR 120,000 - 180,000
AI Large Language Mode (LLM) Technology Lead/Principal Architect
AI Large Language Mode (LLM) Technology Lead/Principal Architect

Accenture Belgium • Brussel Hoofdstad

On-site
EUR 120,000 - 190,000
AI Engineer
AI Engineer

Accenture • Brussel Hoofdstad

On-site
EUR 90,000 - 130,000
AI Engineer
AI Engineer

Accenture PLC • Brussel

On-site
EUR 90,000 - 135,000
AI Engineer
AI Engineer

Accenture Belgium • Brussel

On-site
EUR 90,000 - 150,000
AI Large Language Model (LLM) Technology Architect
AI Large Language Model (LLM) Technology Architect

Accenture Belgium • Brussel Hoofdstad

On-site
EUR 105,000 - 152,000
AI Large Language Model (LLM) Engineer
AI Large Language Model (LLM) Engineer

Accenture Belgium • Brussel

On-site
EUR 70,000 - 105,000
AI LLM Technology Architecture Senior Manager
AI LLM Technology Architecture Senior Manager

Accenture Belgium • Brussel

On-site
EUR 110,000 - 160,000
AI Large Language Mode (LLM) Technology Lead/Principal Architect
AI Large Language Mode (LLM) Technology Lead/Principal Architect

3700 Accenture N.V. Company • Herentals

On-site
EUR 150,000 - 210,000