Director, Core Infrastructure Engineering

Oracle

Bengaluru

On-site

INR 6,000,000 - 9,000,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Oracle in Bengaluru seeks a senior architect to lead multiple teams in designing scalable distributed systems and ensuring high availability. You will orchestrate cross‑group optimization, drive data plane platform use, and set standards for durability, security, and compliance across services.

You will oversee fault‑tolerant architectures, formal verification where appropriate, and incident response leadership.

Qualifications

  • Focus on system design, scalability, and reliability in distributed architectures.
  • Strong ability to define KPIs, telemetry, dashboards, and alerting for proactive health monitoring.
  • Experience with security, encryption, access controls, and compliance in cloud environments.
  • Ability to lead multiple teams, drive cross‑group optimization, and align stakeholders.

Responsibilities

  • Implements strategies for architecture and design of interdependent scalable distributed systems across multiple teams.
  • Drives data plane platform usage for large‑scale data operations and optimises high‑throughput processing.
  • Oversees fault‑tolerant designs, dynamic elasticity, and load shedding/throttling to protect services.
  • Guides incident management, post‑incident analysis, and readiness through SOPs and playbooks.
  • Drives automation (IaC), change management, and secure patching strategies at scale.
  • Ensures system availability, durability, and alignment with service level objectives (SLOs).
  • Oversees security improvements, encryption controls, and regulatory compliance.

Job description

Job Description

Leads a multple teams to implement strategies for the architecture and delivery of interdependent, scalable distributed systems that meet organizational and customer demands. Orchestrates cross-group optimisation for high‑throughput, large‑scale data processing; aligns stakeholders on scalability requirements; and oversees elastic designs and effective use of data plane platforms. Provides strategic oversight for fault‑tolerant, in‑service‑upgradable architectures, sets direction for partition‑aware design choices, and leads initiatives to harden networks via load‑shedding, throttling, and rate‑limiting. Establishes expectations for formal verification and peer reviews, and sets SLO‑aligned durability and availability standards across the department. Drives KPI and telemetry strategies; directs creation of complex dashboards and alerting for proactive health assurance; and ensures functional/correctness validation, data replication, and synchronization meet organizational needs. Guides organisation‑wide incident management and operational readiness, eliminating customer maintenance windows and ensuring consistent SOPs. Provides strategic security guidance (encryption, access controls), oversees remediation and compliance documentation, and sponsors automation (IaC) and change‑management alignment so systems can be safely patched, updated, and rolled back at scale.

Responsibilities
Key Responsibilities
System Design & Architecture – System Scalability:
  • Implements strategies across multiple teams or groups for the architecture and design of interdependent scalable distributed systems, including the use of distributed state management tools, ensuring organisational and system demands are met.
  • Spearheads code and/or system optimisation initiatives for large‑scale data processing and high‑throughput requirements across multiple areas, driving improvements that support hyper‑scale systems.
  • Facilitates collaborations to define system scalability requirements, ensuring the defined requirements meet customer expectations.
  • Oversees the design of interdependent systems to scale with elasticity (e.g., effectively scaling both up and down).
  • Drives the effective use and implementation of data plane platforms for large‑scale data operations.
System Design & Architecture – System Reliability Design:
  • Provides strategic oversight for the architecture of fault‑tolerant interdependent systems capable of withstanding in‑service updates by overseeing implementation across teams of redundancy, replication, and automatic failover mechanisms.
  • Influences and sets direction for designing systems to effectively handle service disruptions (e.g., network partitions) by prioritising consistency, availability, or partition tolerance.
  • Leads strategic optimisation initiatives for handling network unreliability, including directing the design of load‑shedding, throttling, and rate‑limiting techniques.
  • Holds teams accountable for leveraging formal verification techniques to verify system designs and conduct peer reviews across teams.
  • Drives the design of systems that are durable and adhere to service level objectives (SLOs), developing standards for availability and durability of other computing services across the department.
System Design & Architecture – System Reliability Performance:
  • Drives strategies for defining key performance indicators (KPIs) and telemetry to identify risks, gaps, or cyclical dependencies in running systems, ensuring alignment with organisational goals.
  • Directs the creation and customisation of complex dashboards, telemetry systems, and alerting mechanisms that proactively monitor and ensure optimal system health across teams.
System Design & Architecture – Correctness / Availability:
  • Implements strategies to effectively determine if systems are meeting functional and correctness requirements, and encourages teams to identify improvement opportunities.
  • Provides thought leadership on processes for formally verifying complex features to ensure system design correctness.
  • Oversees the implementation of data replication and synchronisation techniques, ensuring data integrity and availability across the organisation.
Operational Troubleshooting & Incident Management:
  • Provides strategic oversight for diagnosing, debugging, and resolving issues in active systems to support ongoing operation.
  • Directs strategies within teams to prevent interruptions, ensuring no maintenance windows are required for customers and users when resolving issues.
  • Drives alignment across teams for operational readiness protocol and standard operating procedures.
  • Provides expert guidance for complex incident response and root cause investigations.
Compliance & Security:
  • Provides strategic guidance in architecting robust security measures to protect data and applications in multi‑tenant environments, ensuring encryption techniques and access controls are implemented.
  • Oversees execution of remediation plans to address identified security gaps, promoting significant improvements and continuous advancement of security measures.
  • Drives documentation efforts and ensures cloud infrastructure compliance with industry standards and regulations.
Automation & Change Management:
  • Provides strategic guidance across teams on developing and maintaining automation scripts and tools (e.g., Infrastructure as Code (IaC)) to manage cloud infrastructure.
  • Drives strategic alignment of change management plans for patching, updating, and rolling back applications, and oversees that system designs allow for automation of these processes.
Core Responsibilities
Planning & Execution:
  • Oversees and guides multiple teams on managing complex projects or initiatives, monitoring timelines, deliverables, and budgets (when applicable) to ensure strategic objectives are met.
  • Serves as a role model for appropriately delegating work, setting priorities, and ensuring alignment with business needs.
  • Coaches others on adjusting resources or project timelines in anticipation of business changes.
Collaboration & Partnership:
  • Role models leading cross‑functional collaborative efforts to ensure alignment of expectations and strategic objectives.
  • Empowers teams to build and maintain partnerships with business leaders, stakeholders, and/or customers to address barriers and contribute to organisational success.
  • Drives transparency and inclusivity by modelling actively seeking, listening to, and leveraging diverse perspectives.
Problem Solving:
  • Shares problem‑solving strategies across teams, providing oversight on complex operational and/or technical issues, as needed.
  • Coaches teams on analysing highly complex data and/or information to identify solutions to ambiguous issues.
  • Provides direction on identifying root causes to prevent recurrence of issues.
Continuous Learning:
  • Pursues strategic learning opportunities to maintain expertise and apply best practices at the organisational level.
  • Creates opportunities for team members and leaders to build their expertise in new areas, coaching them to build innovative skills.
  • Identifies skill gap trends across the organisation and upholds a culture that places significant emphasis on sharing knowledge and pursuing learning opportunities that advance the organisation.
  • Evaluates the efficiency of learning strategies and recommends adjustments as needed.
Continuous Improvement:
  • Empowers teams to own the development and implementation of ideas that increase the efficiency and effectiveness of processes, protocols, and workflows across the department.
  • Coaches teams to gain buy‑in for ideas and to seek feedback on approaches and methods for continued improvement.
  • Prioritises and reviews the roadmap of improvement initiatives to ensure alignment with strategic direction and maximise return on investments.
Performance and Development:
  • Serves as a role model for driving performance across teams through tailored feedback and coaching in alignment with performance management processes, guidelines, and expectations.
  • Drives consistency in the application of talent development procedures and socialises performance expectations across the organisation.
  • Ensures that individual development goals are aligned with organisational strategic initiatives.
  • Collaborates with HR to implement talent strategy through hiring and promotion processes.
Qualifications

Career Level - M4

About Us

Only Oracle brings together the data, infrastructure, applications, and expertise to power everything from industry innovations to life‑saving care. And with AI embedded across our products and services, we help customers turn that promise into a better future for all. Discover your potential at a company leading the way in AI and cloud solutions that impact billions of lives.

True innovation starts when everyone is empowered to contribute. That’s why we’re committed to growing a workforce that promotes opportunities for all with competitive benefits that support our people with flexible medical, life insurance, and retirement options. We also encourage employees to give back to their communities through our volunteer programs.

We’re committed to including people with disabilities at all stages of the employment process. If you require accessibility assistance or accommodation for a disability at any point, let us know by emailing accommodation-request_mb@oracle.com or by calling 1-888-404-2494 in the United States.

Oracle is an Equal Employment Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, colour, religion, sex, national origin, sexual orientation, gender identity, disability and protected veterans’ status, or any other characteristic protected by law. Oracle will consider for employment qualified applicants with arrest and conviction records pursuant to applicable law.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Core Infrastructure Engineer
Senior Core Infrastructure Engineer

Oracle • Bengaluru

On-site
INR 1,500,000 - 2,100,000
Senior Core Infrastructure Engineer
Senior Core Infrastructure Engineer

Oracle • Thiruvananthapuram

On-site
INR 1,200,000 - 1,800,000
Software Development Manager
Software Development Manager

Ll Oefentherapie • India

On-site
INR 4,000,000 - 7,500,000
Lead Principal Application Software Engineer
Lead Principal Application Software Engineer

Oracle • Bengaluru

On-site
INR 1,000,000 - 1,500,000
Flexible medical and life insurance options
Retirement plans
Volunteer opportunities
Principal Network Developer
Principal Network Developer

Oracle • Bengaluru

On-site
INR 1,800,000 - 3,200,000
Senior Application Software Engineer
Senior Application Software Engineer

Oracle • Bengaluru

On-site
INR 2,500,000 - 4,500,000
Senior Advanced Services Engineer
Senior Advanced Services Engineer

Oracle • Bengaluru

On-site
INR 1,800,000 - 2,400,000
Application Software Engineer 2
Application Software Engineer 2

Oracle • Hyderabad

On-site
INR 900,000 - 1,500,000
Application Software Engineer 2
Application Software Engineer 2

Oracle • Ahmedabad District

On-site
INR 900,000 - 1,300,000
Application Software Engineer 2
Application Software Engineer 2

Oracle • Dadri

On-site
INR 600,000 - 900,000