Senior Software Engineer, Infrastructure

Ironflow AI

San Diego (CA)

On-site

USD 120,000 - 150,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

A cutting-edge tech company in California seeks a Senior Software Engineer specializing in Infrastructure. You'll design and maintain AWS-based production systems, manage Kubernetes clusters, and enhance observability stacks. This hands-on role requires 5+ years of experience in infrastructure and a solid background in AWS and Kubernetes. Strong skills in Python and TypeScript are essential. The position offers significant ownership of projects at an early-stage company focused on delivering secure, compliant infrastructure.

Qualifications

  • 5+ years of professional experience in infrastructure, DevOps, or platform engineering.
  • Deep hands-on experience with AWS services in production.
  • Strong expertise in Kubernetes operations and management.

Responsibilities

  • Design and build production infrastructure on AWS.
  • Enhance observability stack for better monitoring.
  • Contribute to backend services in Python and TypeScript.

Skills

AWS expertise
Kubernetes management
Python proficiency
TypeScript familiarity
GitOps workflows
Linux systems administration

Tools

GitHub Actions
Grafana
FastAPI

Job description

About Ironflow AI

We are building the world’s first AI-native ERP for defense tech — a system that thinks, reasons, and acts alongside the people who build, protect, and defend our way of life.

We’re not patching old ideas with a flashy new user interface. We’re starting fresh — reimagining how software supports operations at the cutting edge of the defense tech revolution. Built by veterans, engineers, and operators who’ve lived the pain of legacy systems. Ironflow combines natural language, intelligent agents, and real-world rigor to empower the frontline with clarity and speed.

If you’re bold, decisive, and obsessed with impact — join us. This is where curiosity meets purpose.

Why This Role Matters

We're looking for a Senior Software Engineer specializing in Infrastructure who is excited about building and scaling the platform that powers our product. You'll own the reliability, performance, and security of our infrastructure while working closely with product engineers to ship features faster.

This is a high-ownership role at an early-stage company. The role has broad scope by design: you'll own the full infrastructure surface, from cluster operations to developer experience. We operate on AWS GovCloud in an environment purpose-built for ITAR and CUI compliance, which means the infrastructure decisions you make carry real regulatory weight. You won't be maintaining someone else's platform: you'll be the person who built it. If you want broad ownership, a direct line to product impact, and the technical challenge of building compliance-grade infrastructure from scratch, this role is for you.

What You’ll Own
Infrastructure & Cloud
  • Design, build, and maintain production infrastructure on AWS (EKS, RDS, ECR, VPC, IAM, Secrets Manager, etc.).
  • Develop and manage our Kubernetes clusters: deploy workloads, tune Karpenter node autoscaling, maintain Helm charts, and keep clusters healthy.
  • Own and extend our GitOps deployment pipeline: GitHub Actions for CI/CD, ArgoCD for continuous delivery, and Helm for packaging.
  • Manage supporting cluster operators including Envoy Gateway, External DNS, cert-manager, Fluent Bit, and the AWS Load Balancer Controller.
Reliability & Observability
  • Own and improve our observability stack—Grafana for dashboards, Loki for log aggregation, Tempo for distributed tracing, and Prometheus for metrics.
  • Support multi-environment reliability across dev, stage, and production GovCloud accounts.
  • Improve system resilience through load testing (Locust), E2E testing (Playwright/Cucumber), and thoughtful capacity planning.
Application Development
  • Contribute to backend services (FastAPI, SQLAlchemy) in Python and TypeScript.
  • Work alongside product engineers as a first-class contributor, making architecture decisions that balance speed, cost, and reliability.
  • Build developer experience tooling: local dev environments, CI pipeline improvements, and automated testing scaffolds that make the whole team faster.
  • Support and extend Temporal-based workflow orchestration for background processing.
Security & Compliance
  • Implement least-privilege IAM policies, IRSA (IAM Roles for Service Accounts), and network segmentation in a GovCloud environment.
  • Manage secrets through AWS Secrets Manager and the External Secrets Operator with automated rotation.
  • Maintain TLS automation via cert-manager and OIDC authentication flows.
  • Enable SOC2, CMMC, and FedRAMP compliance activities: GRC platform integration, audit logging pipelines, FIPS-validated endpoint configuration, system boundary documentation, and evidence collection for third-party assessments.
What You Bring
  • 5+ years of professional experience in infrastructure, DevOps, SRE, or platform engineering.
  • Deep hands-on experience with AWS services in production (EKS, IAM, Secrets Manager, ECR, RDS). Experience with or strong working knowledge of AWS GovCloud is a significant plus.
  • Strong Kubernetes expertise: you've operated clusters, debugged networking issues, managed Helm charts, and tuned workloads.
  • Proficiency in Python and/or TypeScript with a genuine interest in writing application code alongside infrastructure work.
  • Experience with GitOps workflows: ArgoCD, GitHub Actions, and Helm-based deployments.
  • Solid understanding of networking fundamentals (DNS, load balancing, TLS, Kubernetes Gateway API).
  • Comfort with Linux systems administration and shell scripting.
  • Familiarity with compliance-driven infrastructure: audit logging, access controls, and evidence collection for frameworks like CMMC, FedRAMP, and SOC 2.
  • A collaborative, low-ego mindset: you thrive in small, fast-moving teams.
Nice to Have
  • Experience with Envoy Gateway or the Kubernetes Gateway API.
  • Background in PostgreSQL administration and schema-based multi-tenancy.
  • Familiarity with the Grafana observability stack (Loki, Tempo, Prometheus).
  • Experience with Karpenter for node autoscaling or cost optimization strategies for cloud spend.
  • Experience with Temporal for workflow orchestration.
  • Experience in a startup or high-growth environment where you wore many hats.
How We Work

We value autonomy, urgency, and ownership. Everyone at Ironflow moves fast, solves real problems, and holds a high bar for craft. We expect you to lead — not just follow — and to build systems that stand up to real-world complexity while staying clean and maintainable.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior / Lead Infrastructure & Operations Engineer
Senior / Lead Infrastructure & Operations Engineer

Austin Werner • Boston (MA)

On-site
USD 120,000 - 150,000
Senior Software Engineers | Platform | Infrastructure | Cloud | AI
Senior Software Engineers | Platform | Infrastructure | Cloud | AI

District-Partners • Arlington (VA)

Hybrid
USD 180,000 - 230,000
Software Engineer - Infrastructure
Software Engineer - Infrastructure

Emergentlabsinc • San Francisco (CA)

On-site
USD 110,000 - 150,000
401(k)
Health, dental, and vision insurance
Unlimited Paid Time Off
+1
Sr. Cloud Infrastructure Engineer
Sr. Cloud Infrastructure Engineer

Istaridigital.Ai • United States

On-site
USD 140,000 - 190,000
Staff Site Reliability Engineer
Staff Site Reliability Engineer

Ironclad • New York (NY), Chicago (IL), San Francisco (CA)

Hybrid
USD 220,000 - 235,000
Health coverage
Parental leave
Wellbeing stipends
Senior DevOps Engineer - LATAM (Remote)
Senior DevOps Engineer - LATAM (Remote)

Socket.dev • United States

Remote
USD 140,000 - 210,000
Senior DevOps Engineer – LATAM
Senior DevOps Engineer – LATAM

Luxury Presence • United States

Remote
USD 140,000 - 210,000
Infrastructure Engineer
Infrastructure Engineer

Overland AI • Seattle (WA)

On-site
USD 130,000 - 225,000
Competitive salary: $130K – $225K annually
Equity compensation
Best-in-class healthcare, dental, and vision plans
+3
Senior Infrastructure Engineer
Senior Infrastructure Engineer

CyberCoders • United States

On-site
USD 120,000 - 150,000
Principal Platform Engineer
Principal Platform Engineer

Socket.dev • New York (NY), Herndon (VA)

On-site
USD 110,000 - 170,000