DevOps Engineer (local)

Pitch Aeronautics

Idaho

On-site

USD 120,000 - 180,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Health Insurance
Dental & Vision Insurance
120 hrs of paid time off
Flexible work schedule
HSA
FSA
Paid holidays
Short-term disability

Job summary

Pitch Aeronautics is seeking an experienced DevOps/SRE leader to own our AWS infrastructure end to end, including ECS, EC2, RDS, S3, VPC, and security controls. You’ll build automated pipelines, manage IaC with Terraform, and drive our SOC 2 program with budgets and alerts.

You will work with AI tooling, scale cloud and ML workloads for forecasting products, operate Prometheus/Grafana logging, and ensure reliability while enabling rapid, compliant deployments across environments.

Qualifications

  • 5 years in DevOps, SRE, or platform engineering.
  • Design, deployment, and operation of scalable AWS architectures including ECS, EC2, RDS, S3, and VPC networking.
  • Terraform in production, modules, remote state, drift handling.
  • GitHub Actions for CI/CD of containerized services.
  • Experience with Prometheus and Grafana; log stack (Elasticsearch/OpenSearch/Kibana or similar).
  • Python for automation; Linux administration in production (Ubuntu/Amazon Linux).
  • Least-privilege IAM, secrets management, encryption, and CI supply-chain hygiene.
  • Familiarity with AI tools in day‑to‑day engineering; understanding model behavior and failure modes.

Responsibilities

  • Build and manage CI/CD pipelines in GitHub Actions: build containers, push to ECR, and promote across environments with short‑lived credentials.
  • Own infrastructure as code: Terraform modules, remote state, budgets, and alerts.
  • Operate hybrid compute footprint: ECS on Fargate, EC2 Docker workloads.
  • Own observability and logging: Prometheus/Grafana, log pipeline, and alerting.
  • Operate machine‑learning services: deploy containerized inference, manage GPU instances, and registry lifecycle.
  • Own Amazon RDS for PostgreSQL: backups, PITR, upgrades, parameter groups, connection limits.
  • Own networking end to end: VPCs, ALBs, NGINX fronting Java apps, secure access paths.
  • Implement and manage AWS security controls and the SOC 2 program: access control, patch automation, vulnerability management.

Skills

DevOps
SRE
Platform engineering
AWS
CI/CD
Python
Linux
Security
AI tools

Tools

Terraform
GitHub Actions
Prometheus/Grafana
Elasticsearch/OpenSearch/Kibana
OpenSearch
Kubernetes
Docker
AWS services

Job description

Role Description

You'll own our AWS infrastructure end to end, but this isn't a keep-the-lights-on job. There is real infrastructure to build. You'll stand up new customer environments as we grow, scale the cloud and ML infrastructure behind our forecasting products, and build the automation behind our SOC 2 program. You view compliance as a byproduct of good engineering and not a quarterly fire drill.

We're an AI-driven company. AI tooling runs throughout our engineering workflow, because it's how a team our size covers this much ground.

If you're passionate about building cloud infrastructure, automation, and work that makes the electric grid safer, we'd love to hear from you.

Responsibilities
  • Build and manage CI/CD pipelines in GitHub Actions: build container images, push them to Amazon ECR, and promote them across environments using short-lived federated credentials, not long-lived keys.
  • Own our infrastructure as code:Terraform modules and remote state, nothing that lives only in the console, and a bill kept honest with budgets, billing alerts, and right-sizing.
  • Operate our hybrid compute footprint:Amazon ECS on AWS Fargate, plus Amazon EC2 instances running Dockerized workloads under systemd.
  • Own observability and centralized logging:our self-hosted Prometheus/Grafana stack, the log pipeline and search stack behind it, and alert rules worth acting on.
  • Operate our machine-learning services:deploy and run containerized inference for our customers, manage GPU instances, and keep multi-gigabyte model images and their registry lifecycle under control.
  • Own Amazon RDS for PostgreSQL: backups and point-in-time recovery, version upgrades, parameter groups, connection limits, and the slow-query conversation with the app team.
  • Own networking end to end: VPCs and segmentation, ALBs, NGINX in front of our Java apps, and the secure access paths engineers use to reach private infrastructure.
  • Implement and manage AWS security controls (IAM, encryption, and audit logging) and own the infrastructure side of our SOC 2 program: access control, patch automation, and vulnerability management.
Minimum Qualifications
  • Experience: 5 years in DevOps, SRE, or platform engineering.
  • Cloud Infrastructure (AWS): Design, deployment, and operation of scalable cloud architectures such as Amazon ECS in production (task definitions, service deploys, rollbacks), Amazon EC2, Amazon RDS, Amazon S3, VPC networking (private subnets, NAT vs. VPC endpoints, security groups, DNS), and Elastic Load Balancing.
  • Infrastructure as Code: Terraform in production, including writing modules, managing remote state, and handling drift.
  • Automation and CI/CD: GitHub Actions for building, testing, and deploying containerized services or equivalent experience you can transfer.
  • Monitoring and Logging: You've owned an alerting stack rather than just read its dashboards, and you can defend which alerts you deleted. Comfortable with Prometheus and Grafana, and with a log search stack (Elasticsearch/Kibana, OpenSearch, or similar).
  • Python and Linux: Python for automation, tooling, and diagnostics; Linux administration in production (Ubuntu, Amazon Linux).
  • Security: Least-privilege IAM design, secrets management, encryption, and supply-chain hygiene in CI
  • Working with AI Tools: Fluency with AI assistants in day-to-day engineering work, paired with a grasp of the concepts underneath them. Know what the models are doing and when and why they fail. You move faster with them, and you know enough to catch them when they're confidently wrong.
Desired Qualifications
  • JVM Applications: Keeping Maven-built Java/Spring applications healthy in production: heap and GC behavior.
  • PostgreSQL Depth: Query and index tuning, RDS Performance Insights, and pg_stat_statements.
  • ML and GPU Operations: Serving models in production, managing the instances behind them, and orchestrating jobs (SkyPilot or similar).
  • Geospatial or Weather Data: Familiarity with HRRR, GRIB, Zarr, or similar gridded datasets.
  • Compliance: SOC 2, plus control automation, vulnerability scanning, and patch management at scale.
  • Critical-Infrastructure Customer Support: Supporting a product deployed into regulated customer environments which includes utility-sector security reviews, NERC CIP exposure, or third-party risk assessments.
Benefits
  • 120 hrs of paid time off
  • Health Insurance
  • Dental & Vision Insurance
  • Short-term disability insurance
  • Long-term disability insurance
  • Health Savings Account (HSA)
  • Flexible Spending Account (FSA)
  • Paid holidays
  • Flexible work schedule
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

DevOps Engineer
DevOps Engineer

Oxbridge Health • Norwalk (CA)

On-site
USD 140,000 - 190,000
DevOps Engineer
DevOps Engineer

Oxbridge Health, Inc. • Norwalk (CT)

On-site
USD 120,000 - 160,000
DevOps Engineer
DevOps Engineer

Motion Recruitment Partners LLC • Boston (MA), Northern (KY)

Hybrid
USD 140,000 - 180,000
Medical Insurance
Dental Benefits
Vision Benefits
+6
DevOps Engineer
DevOps Engineer

Twenty • New York (NY)

On-site
USD 100,000 - 140,000
Medical, dental, and vision plans
Paid parental leave
Flexible PTO
+2
DevOps Engineer/IT
DevOps Engineer/IT

Vitalacy, Inc. • Los Angeles (CA)

On-site
USD 100,000 - 130,000
Attractive compensation package
Pre-IPO stock options
Health insurance
+1
DevOps Engineer
DevOps Engineer

hackajob • New York (NY)

On-site
USD 100,000 - 130,000
Medical, dental, and vision plan options
Paid parental leave
Flexible PTO
+2
DevOps Engineer
DevOps Engineer

Jobless • Norwalk (CT), Northern (KY)

Hybrid
USD 120,000 - 180,000
Sr. Lead DevOps Engineer
Sr. Lead DevOps Engineer

Solomon Page • New York (NY)

On-site
USD 200,000 - 225,000
DevOps Engineer
DevOps Engineer

Socket.dev • United States

Remote
USD 120,000 - 170,000
Health insurance
401k with company match
Unlimited PTO
+1
DevOps Engineer DevOps Engineer
DevOps Engineer DevOps Engineer

Kurai • Seattle (WA)

On-site
USD 120,000 - 180,000