Senior Core Infra Architect - Distributed Systems

Oracle

United States

On-site

USD 115,000 - 235,000

Full time

33 hours ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Medical, dental, and vision insurance
401(k) with company match
Paid time off
Parental leave
Stock purchase plan

Job summary

Oracle in the United States is seeking an experienced software/system design engineer to architect scalable distributed systems, define elasticity, and ensure high throughput and reliability.

You will own performance, reliability, security, and automation, building dashboards, implementing IaC, and guiding incident response while mentoring peers. This role emphasizes operational readiness and change-management, with a focus on scalable architecture and continuous improvement.

Responsibilities

  • Lead the development and implementation, and begin to architect, components of scalable distributed systems that support horizontal and vertical scaling to meet system demands, including leveraging distributed state management tools.
  • Optimize code and/or systems for large-scale data processing and high-throughput requirements to support hyper-scale systems.
  • Define scalability requirements for owned components and ensure design and implementation requirements are met.
  • Design systems to scale with elasticity (e.g., effectively scaling both up and down).
  • Leverage data plane platforms to effectively handle large-scale data retrieval, storage, and processing.
  • Design performance and load testing.
  • Build and design fault-tolerant components and systems capable of withstanding in-service updates by implementing redundancy, replication, and automatic failover mechanisms.
  • Design systems to effectively handle service disruptions by prioritizing consistency, availability, or partition tolerance.
  • Implement and optimize approaches to handle network unreliability, including load-shedding, throttling, and rate-limiting.
  • Define key performance indicators (KPIs) and telemetry to identify gaps or issues in running systems.
  • Build and customize moderately complex dashboards, telemetry systems, and alerting mechanisms to proactively monitor components and system health.
  • Design and implement functional and correctness requirements for feature sets and/or systems in new or existing systems.
  • Design complex test scenarios (e.g., fault-injection, brown-out) to evaluate system correctness.
  • Implement data replication and synchronization techniques to maintain data integrity and availability.
  • Take a proactive role in diagnosing, debugging, and resolving issues in active components and systems to support ongoing operation.
  • Maintain expertise in owned components and systems to ensure effective troubleshooting and performance.
  • Meet operational readiness expectations through design and implementation.
  • Serve in operational support rotations, providing guidance in incident response and root cause investigations.
  • Implement robust security measures to protect data and applications in multi-tenant environments, including encryption techniques and access controls.
  • Execute remediation plans to address identified security gaps.
  • Ensure cloud infrastructure is in compliance with industry standards and regulations and that documentation is up to date.
  • Develop and maintain automation scripts and tools (e.g., IaC) to manage cloud infrastructure.
  • Create and adhere to change management plans for patching, updating, and rolling back applications, and begin designing systems and components to allow for automation of these processes.
  • Planning & Execution: Manages and coordinates moderately complex tasks, monitoring timelines and deliverables to ensure timely completion.
  • Collaboration & Partnership: Collaborates across the organization to align on expectations and achieve shared objectives.
  • Problem Solving: Identifies and addresses moderately complex issues by analyzing data to identify solutions.
  • Continuous Learning: Pursues learning opportunities to expand knowledge and skills and stays abreast of industry trends.
  • Continuous Improvement: Develops ideas and collaborates on process improvements.
  • Performance and Development: Participates in candidate interviews and hiring recommendations.

Job description

Oracle in the United States is seeking an experienced software/system design engineer to architect scalable distributed systems, define elasticity, and ensure high throughput and reliability.

You will own performance, reliability, security, and automation, building dashboards, implementing IaC, and guiding incident response while mentoring peers. This role emphasizes operational readiness and change-management, with a focus on scalable architecture and continuous improvement.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Distributed Systems & Infra Engineer
Senior Distributed Systems & Infra Engineer

Oracle • Frankfort (KY)

On-site
USD 115,000 - 235,000
Medical, dental, vision insurance
401(k) with company match
Paid time off
Senior Software Engineer, Core Infrastructure & Reliability
Senior Software Engineer, Core Infrastructure & Reliability

Oracle • Nashville (TN)

On-site
USD 115,000 - 235,000
Medical insurance
Dental insurance
Vision insurance
+1
Principal Infrastructure Architect for Scalable Systems
Principal Infrastructure Architect for Scalable Systems

Oracle • Frankfort (KY)

On-site
USD 146,000 - 306,000
Senior Distributed Systems Engineer — Scalable & Reliable
Senior Distributed Systems Engineer — Scalable & Reliable

Oracle • Nashville (TN)

On-site
USD 79,000 - 210,000
Senior Principal Infra Architect, Scale & Resilience
Senior Principal Infra Architect, Scale & Resilience

Oracle • Santa Clara (CA)

On-site
USD 146,000 - 307,000
Health insurance
401(k) plan with company match
Paid time off
+1
Core Infra Engineer: Scalable, Fault-Tolerant Systems
Core Infra Engineer: Scalable, Fault-Tolerant Systems

Oracle • Seattle (WA)

On-site
USD 180,000 - 240,000
Medical, dental, and vision insurance
Disability insurance
Life insurance
+6
Director of Scalable Infra & Reliability Engineering
Director of Scalable Infra & Reliability Engineering

Oracle • Nashville (TN)

On-site
USD 169,800 - 355,400
Medical insurance
401(k) with company match
Paid time off and holidays
Senior Distributed Systems Reliability Engineer
Senior Distributed Systems Reliability Engineer

Oracle • United States

On-site
USD 79,000 - 210,000
Senior Principal Software Engineer - Distributed Systems Leader
Senior Principal Software Engineer - Distributed Systems Leader

Oracle • United States

On-site
USD 134,000 - 224,000
Medical, dental, and vision insurance
401(k) Savings Plan with company match
Paid time off
+4
Principal Software Engineer: Distributed Systems Leader
Principal Software Engineer: Distributed Systems Leader

Oracle • Nashville (TN)

On-site
USD 94,000 - 224,000
Medical, dental, and vision insurance
401(k) Savings with company match
Paid time off
+1