Cloud Infrastructure Site Reliability Engineer

Robotics Prcocess Automation, LLC

Berkeley Heights (NJ)

On-site

USD 82,656 - 123,984

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

ROBOTIC PROCESS AUTOMATION LLC is seeking a Cloud Infrastructure Site Reliability Engineer in Berkeley Heights, NJ. You will operate cloud infrastructure, implement Google-inspired SRE practices, and drive automation to ensure uptime and performance.

You will work across AWS, GCP, and Azure, manage Linux-based systems, and participate in incident response and post-mortems, while collaborating with cross-functional teams to improve reliability.

Qualifications

  • Bachelor’s degree in Computer Science, Engineering, or a related technical field, or equivalent practical experience.
  • 3+ years of experience in software development with proficiency in programming (Python, Go, Java, C++).
  • Experience administering cloud platforms (AWS, GCP, Azure) including networking, security, containers, storage, data management, and serverless tech.
  • Solid understanding of Linux systems, networking fundamentals, virtualization, distributed systems, file systems, and processes.

Responsibilities

  • Operate infrastructure solutions following Google-like SRE practices to meet uptime, reliability, and performance targets.
  • Drive automation and continuous improvement across production environments.
  • Collaborate with cross-functional teams to enhance cloud reliability posture and streamline processes through automation.

Skills

Python
Go
Java
C++

Education

Bachelor’s degree in Computer Science or related field

Tools

AWS
GCP
Azure
CI/CD

Job description

Cloud Infrastructure Site Reliability Engineer

Location:

Duration: 14Months+ Extension

Hourly Rate: Depending on Experience (DOE)

Work Authorization:

Position Summary

As a Cloud Infrastructure Site Reliability Engineer (SRE) with expertise in multiple public cloud service provider platforms, you will be responsible for operating infrastructure solutions, following the principles and practices pioneered by Google’s SRE model. Your work will ensure our cloud services meet uptime, reliability, and performance targets, and you will drive automation and continuous improvement across our production environments. This role will involve collaborating with cross-functional teams to enhance our cloud reliability posture and streamline processes through automation.

Qualifications
  • ​Bachelor’s degree in Computer Science, Engineering, or a related technical field, or equivalent practical experience.
  • 3+ years of experience in software development with proficiency in at least one programming language (e.g., Python, Go, Java, C++).
  • Experience administrating cloud platforms (AWS, GCP, Azure), including networking, security, containerization, storage, data management, and serverless technologies.
  • Solid understanding of Linux systems, networking fundamentals, virtualized, and distributed systems, file systems, system processes and configurations.
  • Deep understanding of observability (monitoring, alerting, and logging) tools in cloud environments. Ability to set up and maintain monitoring dashboards, alerts, and logs.
  • Familiarity with Continuous Integration/Continuous Deployment (CI/CD) tools for automated testing, deployments, provisioning, and observability.
  • Ability to manage and respond to incidents, perform root cause analysis, and implement post‑mortem reviews.
  • Understanding of setting, monitoring, and maintaining Service-Level Objectives (SLOs) and Service-Level Agreements (SLAs) for system reliability.
Additional Qualifications a Plus
  • Experience working with enterprise‑scale financial services or other regulated industries
  • 5+ years of experience in SRE, DevOps, infrastructure, or cloud engineering roles, preferably supporting large‑scale, distributed systems.
  • Excellent problem‑solving, troubleshooting, and communication skills.
  • Experience leading technical projects or mentoring junior engineers.
  • Certifications: Certified Engineer, DevOps, SRE, CSREF

ROBOTIC PROCESS AUTOMATION LLC is an equal opportunity employer inclusive of female, minority, disability and veterans, (M/F/D/V). Hiring, promotion, transfer, compensation, benefits, discipline, termination and all other employment decisions are made without regard to race, color, religion, sex, sexual orientation, gender identity, age, disability, national origin, citizenship/immigration status, veteran status or any other protected status. ROBOTIC PROCESS AUTOMATION LLC will not make any posting or employment decision that does not comply with applicable laws relating to labor and employment, equal opportunity, employment eligibility requirements or related matters. Nor will ROBOTIC PROCESS AUTOMATION LLC require in a posting or otherwise U.S. citizenship or lawful permanent residency in the U.S. as a condition of employment except as necessary to comply with law, regulation, executive order, or federal, state, or local government contract.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Cloud Infrastructure Site Reliability Engineer
Cloud Infrastructure Site Reliability Engineer

Cloud Hybrid Technologies, LLC • Berkeley Heights (NJ)

On-site
USD 90,000 - 130,000
Site Reliability Engineering (SRE) Architect
Site Reliability Engineering (SRE) Architect

Robotics Prcocess Automation, LLC • Atlanta (GA)

On-site
USD 100,000 - 150,000
Site Reliability Engineering (SRE) Architect
Site Reliability Engineering (SRE) Architect

Robotics Technologies LLC • Atlanta (GA)

On-site
Site Reliability Engineering (SRE) Architect
Site Reliability Engineering (SRE) Architect

Cloud Analytics Technologies, LLC • Atlanta (GA)

On-site
Site Reliability Engineering (SRE) Architect
Site Reliability Engineering (SRE) Architect

Cloud Hybrid Technologies, LLC • Atlanta (GA)

On-site
Manager of Site Reliability Engineering (SRE)
Manager of Site Reliability Engineering (SRE)

Genuine Parts Company • Alabama

On-site
USD 120,000 - 150,000
Cloud Infrastructure SRE: Reliability & Automation Engineer
Cloud Infrastructure SRE: Reliability & Automation Engineer

Robotics Prcocess Automation, LLC • Berkeley Heights (NJ)

On-site
Site Reliability Engineer III
Site Reliability Engineer III

Genuine Parts Company • Alabama

On-site
USD 90,000 - 130,000
Healthcare coverage
401(k)
Tuition reimbursement
+1
Site Reliability Engineering (SRE) Architect
Site Reliability Engineering (SRE) Architect

Quantum Technologies. LLC • Atlanta (GA)

On-site
Site Reliability Engineering (SRE) Architect
Site Reliability Engineering (SRE) Architect

MACHINE LEARNING TECHNOLOGIES LLC • Atlanta (GA)

On-site
USD 140,000 - 190,000