Platform Reliability Engineer

CrewAI, Inc.

Northern (KY)

Hybrid

USD 150,000 - 190,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

CrewAI, Inc. is seeking an infrastructure platform engineer to own and scale the cloud foundation behind our multi‑agent AI platform.

You will work across AWS, container tech, and CI/CD pipelines to improve reliability, security, and velocity for production deployments. You'll collaborate with runtime and product teams to optimize PostgreSQL/Redis workloads, implement robust on‑call practices, and build self‑hosted install tooling for customers.

Qualifications

  • Strong infra/platform engineering experience in production SaaS.
  • Deep experience with AWS, Docker, CI/CD, and containerized services.
  • Comfort with ECS/Kubernetes; Helm is a plus.
  • Proficient with PostgreSQL, Redis, and background job systems.
  • Ability to write reliable automation in Python, Ruby, Go, or Bash.
  • Security‑minded approach to IAM, secrets, and access control.
  • Calm, rigorous incident management and deployment practices.

Responsibilities

  • Own and improve infrastructure running CrewAI's platform across cloud providers.
  • Build and maintain CI/CD pipelines for builds, tests, image publishing, migrations, and deployments.
  • Improve reliability with health checks, alerting, and recovery planning; participate in on‑call rotation.
  • Collaborate with runtime and product engineers on production behavior and scaling.
  • Manage production observability from logs to dashboards and telemetry export.
  • Harden security posture across IAM, secrets management, and vulnerability handling.
  • Develop tooling and automation for self‑hosted installs and operator tooling.
  • Reduce operational toil through automation of recurring workflows.

Skills

AWS
Docker
Kubernetes
CI/CD
GitHub Actions
PostgreSQL
Redis
Automation scripting
Security IAM
Incident management
Celery
FastAPI
Python
Go
Ruby
Bash

Education

Tools

ECS/ECR
Kubernetes
Helm
PostgreSQL
Redis
Sentry
OpenTelemetry

Job description

CrewAI, Inc. is seeking an infrastructure platform engineer to own and scale the cloud foundation behind our multi‑agent AI platform.

You will work across AWS, container tech, and CI/CD pipelines to improve reliability, security, and velocity for production deployments. You'll collaborate with runtime and product teams to optimize PostgreSQL/Redis workloads, implement robust on‑call practices, and build self‑hosted install tooling for customers.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Platform Infra Engineer - Cloud Reliability & Automation
Platform Infra Engineer - Cloud Reliability & Automation

CrewAI • United States

On-site
USD 130,000 - 210,000
Software Engineer, Infrastructure & Reliability
Software Engineer, Infrastructure & Reliability

CrewAI, Inc. • Northern (KY)

Hybrid
USD 150,000 - 190,000
Software Engineer, Infrastructure & Reliability
Software Engineer, Infrastructure & Reliability

CrewAI • United States

On-site
USD 130,000 - 210,000
Infrastructure Engineer, AI-Powered Platform Reliability
Infrastructure Engineer, AI-Powered Platform Reliability

Writer • New York (NY), Northern (KY)

Hybrid
USD 180,000 - 240,000
Generous PTO
Medical coverage
Parental leave
+7
Platform Reliability Engineer — FedRAMP & Scale
Platform Reliability Engineer — FedRAMP & Scale

C1 • Portland (ME)

On-site
USD 180,000 - 250,000
Equity
Medical insurance
In-office in Portland
Platform Reliability Engineer — AI Infra & CloudOps
Platform Reliability Engineer — AI Infra & CloudOps

WRITER • New York (NY)

Hybrid
USD 120,000 - 160,000
Generous PTO
Medical, dental, and vision coverage
Paid parental leave (16 weeks)
+5
Platform Reliability Leader for AI Infrastructure
Platform Reliability Leader for AI Infrastructure

Etched.ai, Inc. • San Jose (CA)

On-site
USD 210,000 - 320,000
Medical, dental and vision coverage
Housing subsidy
Relocation support
+3
Senior Platform Engineer - Cloud & AI Infrastructure
Senior Platform Engineer - Cloud & AI Infrastructure

AI Chopping Block • New York (NY), Northern (KY)

Hybrid
USD 216,000 - 270,000
Health benefits
Equity
Learning stipend
+2
Platform Engineering Manager: Cloud, SRE & Scale
Platform Engineering Manager: Cloud, SRE & Scale

Prolific • United States

On-site
USD 150,000 - 210,000
Platform Engineer - AI Systems & Cloud Infrastructure
Platform Engineer - AI Systems & Cloud Infrastructure

Scale AI • San Francisco (CA), New York (NY)

On-site
USD 216,000 - 270,000
Health insurance
Dental & Vision
Retirement benefits
+3