Principal Cloud Engineering and Production Operations Engineer

A10 Networks, Inc.

San Jose (CA)

On-site

USD 140,000 - 185,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

A10 Networks, Inc. is seeking a Principal Cloud Engineering and Production Operations Engineer to lead cloud architecture and production operations. This senior role requires expertise in designing, automating, and optimizing hybrid and cloud-native production environments, ensuring reliability and security for critical services.

The ideal candidate will have extensive experience in cloud services like OCI, AWS, and Azure, as well as a strong background in infrastructure-as-code tools such as Terraform. A competitive compensation package of $140,000 - $185,000 is offered, along with a collaborative work environment.

Qualifications

  • 10+ years of experience in cloud and infrastructure engineering, including 3+ in a senior role.
  • Proven experience managing production-scale environments.
  • Strong proficiency in Infrastructure-as-code and CI/CD tools.

Responsibilities

  • Lead cloud architecture and engineering design for production workloads.
  • Ensure uptime, performance, and reliability of customer-facing systems.
  • Collaborate to optimize automated deployment pipelines.

Skills

Cloud architecture expertise
Infrastructure-as-code (IaC)
Cloud services (OCI, AWS, Azure)
CI/CD and DevOps tools
Container orchestration (Kubernetes, Docker)
Scripting and automation (Python, Bash)
Monitoring and observability platforms
Analytical and problem-solving

Education

Bachelor's degree in Computer Science or related field
Master's degree (preferred)

Tools

Terraform
CloudFormation
Jenkins
GitLab
Prometheus
Grafana
Datadog

Job description

Principal Cloud Engineering and Production Operations Engineer

The Principal Cloud and Production Operations Engineer serves as the senior technical authority responsible for architecting, automating, and optimizing hybrid and cloud-native production environments that power critical customer-facing services and enterprise applications. This role combines deep cloud infrastructure expertise with strong production reliability and operational engineering skills. The Principal Engineer acts as both architect and hands‑on builder, ensuring scalability, resilience, and security across multi‑cloud and on‑prem environments. Reporting to the Associate Director of IT and Infrastructure, this position will collaborate closely with Engineering, DevOps, Security, and IT Operations to drive a culture of automation, observability, and continuous improvement across the production ecosystem.

Key Responsibilities
Cloud Architecture and Engineering Design

Implement and maintain cloud and hybrid infrastructure supporting production workloads, enterprise systems, and CI/CD pipelines. Lead the adoption of infrastructure‑as‑code (IaC) using Terraform, CloudFormation, or similar tools to enable repeatable, auditable, and secure deployments. Architect scalable and fault‑tolerant solutions across OCI, AWS, Azure, and on‑prem data centers, ensuring high availability and cost efficiency. Evaluate emerging cloud services and technologies for applicability to business needs and long‑term scalability goals.

Production Operations and Reliability

Serve as the technical lead for production operations, ensuring uptime, performance, and reliability of customer‑facing and internal systems. Develop and maintain observability frameworks leveraging metrics, logs, and traces to ensure proactive detection and rapid response. Partner with engineering teams to implement SRE‑inspired practices, including service level objectives (SLOs), error budgets, and post‑incident reviews. Drive root cause analysis, performance tuning, and continuous improvement of production services.

Automation and CI/CD Enablement

Collaborate with DevOps and application engineering teams to build and optimize automated deployment pipelines supporting frequent, low‑risk releases. Integrate security and compliance checks into CI/CD workflows to ensure production readiness and alignment with internal standards. Design self‑healing infrastructure and automated rollback mechanisms to reduce operational risk. Ensure secure and reliable configuration management and environment orchestration using tools such as Ansible, Chef, or Puppet.

Operational Governance and Collaboration

Establish and enforce operational best practices for monitoring, patching, and change management across production systems. Lead production readiness reviews for new releases and large‑scale changes. Collaborate with the Security and Compliance teams to ensure systems adhere to policy, hardening standards, and regulatory requirements. Participate in and occasionally lead on‑call rotations for critical production systems, ensuring rapid triage and resolution.

Leadership and Mentorship

Act as a technical mentor to cloud and infrastructure engineers, fostering a culture of knowledge sharing and engineering excellence. Lead architectural reviews, design sessions, and capacity planning discussions. Serve as a trusted advisor to management on cloud modernization, resilience engineering, and cost optimization strategies.

Qualifications
  • Bachelor’s degree in Computer Science, Information Systems, or related field; Master’s preferred.
  • 10+ years of experience in cloud and infrastructure engineering, including 3+ years in a senior or principal role.
  • Expertise with OCI (preferred), AWS and/or Azure cloud services, including networking, compute, storage, and identity management.
  • Proven experience managing production‑scale environments supporting mission‑critical applications and services.
  • Strong proficiency in Infrastructure‑as‑code (Terraform, CloudFormation).
  • CI/CD and DevOps toolchains (Jenkins, GitLab, ArgoCD).
  • Container orchestration (Kubernetes, Docker).
  • Monitoring and observability platforms (Prometheus, Grafana, Datadog, ELK).
  • Scripting and automation (Python, Bash, PowerShell).
  • Solid understanding of security, compliance, and networking principles in hybrid environments.
  • Exceptional analytical, problem‑solving, and incident management skills.
  • Demonstrated ability to lead complex, cross‑functional initiatives from concept to execution.
Preferred Experience
  • Experience in high‑availability SaaS or networking environments.
  • Knowledge of FinOps, cost optimization, and multi‑cloud governance frameworks.
  • Familiarity with Zero Trust, identity federation, and cloud access security model.
  • Exposure to AI/ML infrastructure or data‑driven pipelines is a plus.
  • AI Use Guidelines for Interviews: Our interviews are designed to reflect your own skills and thinking. The use of AI or recording tools during live interviews is not permitted unless explicitly invited by the interviewer or approved in advance as part of a reasonable accommodation. If these tools are used inappropriately or in a way that misrepresents your work, your application may not move forward in the process.
Equal Opportunity Employer

A10 Networks is an equal‑opportunity employer and a VEVRAA federal subcontractor. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability status, protected veteran status, or any other characteristic protected by law. A10 also complies with all applicable state and local laws governing nondiscrimination in employment.

Compensation

Hybrid Targeted compensation guideline: $140,000 - $185,000. Compensation will vary based on number of factors, including market demand for specific skills, role type, job level, and individual qualifications. Final salary offers are determined by considerations including, but not limited to, subject matter expertise, demonstrated skill level, relevant experience, geographic location, education, certifications, and training.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Principal Cloud Engineering and Production Operations Engineer
Principal Cloud Engineering and Production Operations Engineer

A10 Networks, Inc • San Francisco (CA)

Hybrid
USD 140,000 - 185,000
Senior Software Engineer - DevOps
Senior Software Engineer - DevOps

A10 Networks, Inc. • San Jose (CA)

On-site
USD 140,000 - 150,000
Senior Software Engineer - DevOps
Senior Software Engineer - DevOps

A10 Networks, Inc • San Francisco (CA)

On-site
USD 140,000 - 150,000
Senior Manager, Cybersecurity
Senior Manager, Cybersecurity

A10 Networks, Inc. • San Jose (CA)

Hybrid
USD 200,000 - 215,000
Senior Cloud Architecture & Production Reliability Lead
Senior Cloud Architecture & Production Reliability Lead

A10 Networks, Inc. • San Jose (CA)

Hybrid
USD 140,000 - 185,000
Principal Cloud Engineer
Principal Cloud Engineer

NextGenEnergyJobs • San Jose (CA), Northern (KY)

Hybrid
USD 180,000 - 280,000
Senior Cloud Engineer
Senior Cloud Engineer

ESG • Houston (TX)

Hybrid
USD 100,000 - 130,000
Senior Software Engineer, Systems Platform
Senior Software Engineer, Systems Platform

A10 Networks, Inc. • San Jose (CA)

Hybrid
USD 120,000 - 150,000
Principal Cloud Engineer
Principal Cloud Engineer

Logging-in • Oregon (WI)

Hybrid
USD 180,000 - 240,000
Principal DevOps Engineer (US Citizen)
Principal DevOps Engineer (US Citizen)

Palo Alto Networks • California (MO)

Hybrid
USD 147,000 - 238,000