Senior Infrastructure Engineer - AI Ops

Commerce.com

Austin (TX)

Hybrid

USD 136,000 - 204,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Hybrid work model

Job summary

Commerce is hiring a Senior Infrastructure Engineer to own the AI Operations cloud stack in a hybrid, Austin-based role. You will design and operate scalable GCP infrastructure, CI/CD pipelines, and edge delivery components to support AI automation across channels.

You will implement observability, security, and incident response while collaborating with a small, autonomous team that owns production systems end to end. This is a hands-on, high-impact role.

Qualifications

  • 7+ years in infrastructure engineering or DevOps with direct ownership of production systems.
  • Deep Linux systems administration experience at scale (Debian/Ubuntu).
  • Strong hands-on experience with GCP services (Cloud Run, GKE, Cloud Build, IAM, VPC).
  • Infrastructure as code as a core discipline (Terraform, Puppet).
  • Experience designing and maintaining non-trivial CI/CD delivery pipelines (Cloud Build, GitHub Actions, CircleCI).
  • Hands-on experience with DNS, load balancing, and CDN at scale (NGINX, Cloudflare).
  • Experience with distributed monitoring and logging stacks (Prometheus, Grafana, ELK/Kibana).
  • Proficiency in Python, Bash, and shell scripting for automation.

Responsibilities

  • Design, provision, and manage cloud infrastructure on GCP (Cloud Run, GKE, Redis/Memorystore) using Terraform and Puppet.
  • Build and maintain CI/CD pipelines in Cloud Build, GitHub Actions, and CircleCI including automated testing gates and canary deployments.
  • Design and maintain the edge-to-origin traffic stack: DNS, load balancing, CDN, WAF, edge compute, bot management, authentication, and rate limiting.
  • Operate and tune container orchestration for AI workloads with autoscaling and cost-efficient inference sizing.
  • Implement production observability end-to-end: logging, metrics, tracing, SLO/SLI definitions, and alerting.
  • Own incident response coordination and blameless post-mortems; implement strong security posture with least-privilege IAM.

Skills

Linux system administration
GCP expertise
IaC (Terraform, Puppet)
CI/CD pipelines
Security best practices

Tools

GKE
Cloud Run
Terraform
Puppet
NGINX
Cloudflare
CircleCI
GitHub Actions
Prometheus
Grafana
ELK/Kibana
Cloud Load Balancing
VPC Service Controls

Job description

Welcome to the Agentic Commerce Era At Commerce, our mission is to empower businesses to innovate, grow, and thrive with our open, AI-driven commerce ecosystem. As the parent company of BigCommerce, Feedonomics, and Makeswift, we connect the tools and systems that power growth, enabling businesses to unlock the full potential of their data, deliver seamless and personalized experiences across every channel, and adapt swiftly to an ever-changing market. We believe in harnessing AI responsibly to unlock new possibilities, and we’re looking for individuals who use it intentionally to solve problems, accelerate outcomes, and expand what’s possible in their role. Our purpose is to help businesses confidently solve complex commerce challenges so they can build smarter, adapt faster, and grow on their own terms. If you want to be part of a team of bold builders, sharp thinkers, and technical trailblazers who shape the future of commerce, this is the place for you. We’re hiring a Senior Infrastructure Engineer to join Commerce’s AI Operations team — a highly autonomous team focused on rapid development of AI tools. We operate like a startup within the organization, with a small headcount and full ownership from idea through production. You’ll own the infrastructure, delivery pipelines, and cloud architecture that enable every AI automation. This is a hands‑on infrastructure role.

What You’ll Do:
  • Design, provision, and manage cloud infrastructure on GCP (Cloud Run, GKE, Redis/Memorystore) using Terraform and Puppet.
  • Build and maintain CI/CD pipelines in Cloud Build, GitHub Actions, and CircleCI—including automated testing gates, canary and blue‑green deployments, and artifact promotion across environments.
  • Design and maintain the full edge‑to‑origin traffic stack: DNS, load balancing (NGINX, GCP Cloud Load Balancing), CDN, WAF, edge compute (Cloudflare, Cloud Armor), bot management, authentication, and rate limiting.
  • Operate and tune container orchestration for AI workloads—autoscaling policies, resource quotas, cold‑start optimization, and right‑sizing compute for inference cost efficiency.
  • Implement production observability end‑to‑end: structured logging (Cloud Logging, ELK/Kibana), metrics and dashboards (Prometheus, Grafana), distributed tracing, SLO/SLI definition, and alerting.
  • Own incident response coordination and blameless post‑mortems.
  • Implement strong security posture—least‑privilege IAM with Workload Identity, network segmentation, firewall rules, and VPC Service Controls.
  • Evaluate, build, and integrate infrastructure tooling—whether that’s custom automation in Python or a managed AI serving platform.
Who You Are:
  • 7+ years in infrastructure engineering, DevOps, SRE, or platform engineering with direct operational ownership of production systems.
  • Deep Linux systems administration experience (Debian, Ubuntu) at scale.
  • Strong hands‑on experience with GCP services (Cloud Run, GKE, Cloud Build, IAM, VPC).
  • Infrastructure as code as a core discipline (Terraform, Puppet).
  • Experience designing and maintaining non‑trivial CI/CD delivery pipelines. Cloud Build, GitHub Actions, CircleCI.
  • Hands‑on experience with NGINX and edge infrastructure at scale, including DNS, load balancing, and CDN (Cloudflare or equivalent).
  • Hands‑on experience with distributed monitoring and logging stacks—Prometheus, Grafana, ELK/Kibana, Cloud Logging.
  • Proficiency in Python, Bash, and shell scripting for infrastructure automation and custom tooling.
  • You design infrastructure with least‑privilege IAM, Workload Identity, and compliance controls as defaults.
  • Nice to have: experience with ML infrastructure.
Compensation Transparency

#LI-LH1 #LI-Hybrid (Pay range transparency $135,960 - $203,940)

  • The national base salary range for this role is posted above in this job post.
  • Final compensation will be determined based on factors such as relevant experience, skills, qualifications and geographic location.
  • We also consider internal equity to help ensure fair and consistent pay practices across our teams.
  • Where applicable, this role may also be eligible for variable compensation (such as bonus or commission), equity, and benefits in accordance with local policies.
  • Details will be shared during the hiring process.
  • We are committed to equitable and transparent pay practices that align to market data, internal equity, and individual contribution.
Inclusion and Belonging

At Commerce, we believe that celebrating the unique histories, perspectives and abilities of every employee makes a difference for our company, our customers and our community. We are an equal opportunity employer and the inclusive atmosphere we build together will make room for every person to contribute, grow and thrive. We are committed to creating an inclusive and accessible hiring experience for all candidates. If you require accommodations or adjustments at any stage of the recruitment process, please let us know and we will work with you to meet your needs. Learn more about the Commerce team, culture and benefits at https://www.commerce.com/careers/

Don’'t Miss Out! Like what you see but suffering from some serious FOMO? Join our Commerce Talent Community, and plug in to our latest news and career opportunities. We’re a group of clever, committed, curious people, unleashing talent in all we do. We believe in the power of togetherness, striving at the edge of what’s possible, impacting the lives of billions of people for the better. In all we do, We Do Extraordinary–and that’s no small feat! Our people are our power. It’s only through dedication, collaboration, and inspiration that we can Do Extraordinary. We’re natural problem‑solvers, champions of empowering businesses, and hungry learners… but we also play nerf wars in the office, support each other, and hang out outside of work.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Staff Software Engineer
Staff Software Engineer

Commerce.com • Austin (TX)

Hybrid
USD 186,000 - 280,000
Lead Infrastructure Engineer
Lead Infrastructure Engineer

BigCommerce Pty. • Northern (KY)

On-site
USD 110,000 - 186,000
Lead Infrastructure Engineer
Lead Infrastructure Engineer

Commerce • Northern (KY)

Hybrid
USD 110,000 - 186,000
Staff Software Engineer
Staff Software Engineer

BigCommerce Pty. • Austin (TX)

On-site
USD 186,000 - 280,000
Hybrid work model
Pay transparency
Staff Software Engineer
Staff Software Engineer

Commerce • Austin (TX)

On-site
USD 186,000 - 280,000
Lead Infrastructure Engineer
Lead Infrastructure Engineer

BigCommerce Pty • Northern (KY)

On-site
USD 110,000 - 186,000
Staff Software Engineer
Staff Software Engineer

BigCommerce Pty • Austin (TX)

Hybrid
USD 186,000 - 280,000
Lead Infrastructure Engineer
Lead Infrastructure Engineer

Commerce • Town of Texas (WI)

Hybrid
USD 110,000 - 186,000
Hybrid work model
Competitive compensation
Lead Infrastructure Engineer
Lead Infrastructure Engineer

bigcommerce • United States

Hybrid
USD 110,000 - 186,000
Hybrid work model
Senior Manager, Infrastructure Engineering
Senior Manager, Infrastructure Engineering

BigCommerce Pty • Austin (TX), Northern (KY)

Hybrid
USD 211,000 - 291,000