Site Reliability Engineer II (SREII)

Prodege, LLC

United States

Remote

USD 130,000 - 180,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Prodege, LLC is seeking a Site Reliability Engineer to own AWS/Terraform environments, CI/CD pipelines, and MySQL performance. You’ll automate operational toil, harden infrastructure, and collaborate with engineering teams to ship features efficiently.

You’ll focus on scalable, secure production systems, implement monitoring, incident response, and comprehensive documentation, and continuously evaluate new tools to improve reliability and cost efficiency.

Qualifications

  • Bachelor’s degree or equivalent in a related field.
  • 4+ years in IT operations or a related role with focus on Terraform and AWS.
  • Proficiency in Terraform for infrastructure as code (IaC).
  • Hands-on experience with AWS services (EC2, S3, RDS, Lambda).
  • Experience with scripting languages including Bash and Python/PHP.
  • Knowledge of Jenkins for CI/CD pipeline management.
  • AWS knowledge required; experience with GCP is a plus.
  • Strong analytical and troubleshooting skills to resolve complex infra issues.
  • Excellent verbal and written communication to convey technical information clearly.

Responsibilities

  • Infrastructure management: define and provision AWS infra using Terraform; manage EC2, S3, RDS, Lambda, VPC.
  • Automation: build and maintain automation scripts in Bash, Python, and PHP.
  • CI/CD integration: implement and manage pipelines with Jenkins.
  • Monitoring: track performance, availability, and resource usage; optimize for efficiency.
  • Incident management: troubleshoot outages and performance issues quickly.
  • Cross-functional collaboration to support deployments and infra needs.
  • Documentation: create and maintain infra configurations and processes.
  • Security: adhere to security best practices and compliance standards.
  • Continuous improvement: evaluate new tools and practices to improve infra performance.

Skills

Terraform & IaC
AWS
Bash
Python
PHP
Jenkins
CI/CD
GCP
Docker
Kubernetes
Security best practices
Communication

Education

Bachelor’s degree in CS or related field

Tools

Terraform
AWS
Jenkins
Docker
Kubernetes
Google Cloud Platform

Job description

Job Description:
Read this part first:

This is a build-and-automate role, not a babysit-the-servers one. Prodege is moving fast and changing fast, and we need an SRE who finds that energizing rather than exhausting.

If you want a static environment where nothing changes and the runbook is already written, this isn’t your role, and that’s okay. But if you’re the kind of engineer who sees manual toil and automates it away, who’d rather harden and scale infrastructure than firefight it, and who’s comfortable owning production as the systems around you keep evolving, keep reading.

What you’ll own:

You’ll play a key role in our cloud and data infrastructure, making sure our products run on a stable, scalable, secure foundation. You’ll own AWS/Terraform environments, CI/CD pipelines, and MySQL performance, work that directly affects release velocity, site reliability, and customer experience.

Through thoughtful automation and scripting, you’ll reduce operational toil, increase consistency, and free engineering teams to focus on shipping features. Paired with strong monitoring, incident response, and documentation, you’ll help create predictable, repeatable operations as we grow, and by embedding security best practices and continuously evaluating new tools, you’ll help the org run more efficiently while de-risking the infrastructure over time.

What you’ll do:
  • Infrastructure management: Use Terraform to define and provision AWS infrastructure. Configure and maintain AWS services (EC2, S3, RDS, Lambda, VPC).

  • Automation and scripting: Build and manage automation scripts and tools in Bash, Python, and PHP to streamline operations and improve efficiency.

  • CI/CD integration: Implement and manage continuous integration and deployment pipelines using Jenkins.

  • Monitoring and optimization: Monitor system performance, availability, and resource usage, and implement optimizations for efficiency and reliability.

  • Incident management: Troubleshoot and resolve infrastructure issues, outages, and performance problems quickly and effectively.

  • Collaboration: Work with cross-functional teams to support application deployments and address infrastructure needs.

  • Documentation: Create and maintain comprehensive documentation for infrastructure configurations, processes, and procedures.

  • Security: Ensure all infrastructure and operations adhere to security best practices and compliance standards.

  • Continuous improvement: Evaluate and adopt new technologies and practices to improve infrastructure performance and operational efficiency.

What success looks like:

You consistently deliver stable, secure, scalable AWS infrastructure that supports our products without surprise outages or performance bottlenecks. Deployments run smoothly through well-maintained CI/CD pipelines and automation, with minimal manual intervention and short lead times for changes.

Systems are actively monitored, incidents are investigated quickly, root causes are documented, and meaningful preventative fixes get implemented. Cross-functional teams feel supported because infrastructure needs are anticipated, clearly communicated, and backed by current documentation. Over time, you’re known for reducing operational toil, improving performance and cost efficiency, and thoughtfully introducing new tools and practices that raise the bar on how we run production.

Why this role is worth your time:
  • Real ownership of production : AWS/Terraform environments, CI/CD, and database performance are yours, not a slice of someone else’s stack

  • Automation over toil : you’re measured on eliminating manual work, not absorbing it

  • Your work is visible in the numbers : release velocity, uptime, and lead time for changes all move because of what you build

  • Room to modernize : you evaluate and introduce new tools, not just maintain what exists

  • Backed for growth : Prodege closed a major Blackstone investment in Q1 2026, which means real momentum and real investment in the infrastructure you’d be running

A bit about Prodege:

Prodege is a marketing and consumer insights platform that helps leading brands, marketers, and agencies answer their business questions, acquire customers, grow revenue, and build brand loyalty. We go the extra mile to “Create Rewarding Moments” for our partners, consumers, and team. We operate with startup speed and a startup appetite for change, backed by the resources of a major investment partner.

The must-haves:
  • Bachelor’s degree (or equivalent) in Computer Science, Software Engineering, Information Technology, or a related discipline, or equivalant professional experience in a similar infrastructure/DevOps engineering role

  • 4+ years in IT operations or a related role, with a strong focus on Terraform and AWS

  • Proficiency in Terraform for infrastructure as code (IaC)

  • Hands‑on experience with AWS services (EC2, S3, RDS, Lambda)

  • Experience with scripting languages including Bash and Python/PHP

  • Knowledge of Jenkins for CI/CD pipeline management

  • AWS knowledge required; experience with Google Cloud Platform (GCP) is a plus

  • Strong analytical and troubleshooting skills, with the ability to resolve complex infrastructure issues

  • Excellent verbal and written communication, with the ability to convey technical information clear to technical and non-technical stakeholders

The nice-to-haves:
  • Knowledge across multiple cloud providers

  • Certification in public cloud disciplines

  • Hands‑on experience with Docker and Kubernetes

  • Use of AI in deployment and uptime automation

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Director of Technical Operations
Director of Technical Operations

Deque Systems, Inc • Herndon (VA)

On-site
USD 180,000 - 240,000
DevOps Engineer III
DevOps Engineer III

Enable Dental, Inc. • Austin (TX), Northern (KY)

Hybrid
USD 120,000 - 180,000
Principal Data Engineer
Principal Data Engineer

Prodege, LLC • El Segundo (CA)

On-site
USD 235,000 - 265,000
Medical, dental, vision
STD/LTD and basic life insurance
Flexible PTO
+1
Sr. DevOps Engineer
Sr. DevOps Engineer

DRH Search • Massachusetts

Hybrid
USD 140,000 - 200,000
Senior DevOps Engineer
Senior DevOps Engineer

TTEC Digital • United States

Remote
USD 130,000 - 180,000
Senior Infrastructure Engineer
Senior Infrastructure Engineer

Level AI • Auckland (CA)

On-site
USD 140,000 - 210,000
Devops Developer
Devops Developer

Itgiantsolutions • United States

On-site
USD 90,000 - 120,000
Senior DevOps Engineer – LATAM
Senior DevOps Engineer – LATAM

Luxury Presence • United States

On-site
USD 140,000 - 210,000
Senior Site Reliability Engineer (MAAS)
Senior Site Reliability Engineer (MAAS)

Pragmatike • United States

Remote
USD 140,000 - 190,000
Sr DevOps Engineer
Sr DevOps Engineer

Decca Consulting LLC • United States

Remote
USD 140,000 - 190,000