An application made for this job — a tailored resume and cover letter that speak straight to the posting.
Synstack Technologies is seeking a hands-on contractor to own AWS infrastructure and production reliability for on-site work in Los Angeles. The role emphasizes owning infrastructure delivery, automation, troubleshooting, and production support for data platforms.
Required are 7+ years in infrastructure/SRE with Terraform, Linux networking, containers, and scripting. Strong focus on cost optimization, security, and recovery processes to maintain reliable operations.
Client needs candidates to work on-site.
Please source a hands-on contractor whose primary strength is AWS infrastructure and production reliability, with supporting data engineering experience. The role will independently own infrastructure delivery, automation, troubleshooting, and production support, including the infrastructure behind our data platforms.
Candidates should demonstrate 7+ years of relevant infrastructure, SRE, or platform engineering experience, with depth- not just exposure-in:
AWS infrastructure - production ownership of networking, IAM, compute, storage, and databases; practical experience with VPCs, ECS/Fargate, Lambda, S3, and RDS.
Terraform - reusable modules, remote state, environment separation, drift management, and safe changes through reviewed plans and CI/CD.
Production reliability - incident response, root cause analysis, actionable monitoring, alerting, service objectives, and reducing recurring operational issues.
Linux and networking - troubleshooting processes, resource utilization, DNS, TLS, routing, load balancing, and connectivity.
Containers and deployment - Docker, container operations, Git, CI/CD, deployment troubleshooting, and rollback procedures.
Automation - Python and shell scripting for infrastructure operations, diagnostics, and eliminating manual work.
Security and recovery - least-privilege access, secrets management, backups, restoration testing, and disaster recovery.
Performance and cost management - capacity planning, resource tuning, and AWS cost optimization.
Candidates should have practical experience supporting production data workloads; specialist depth in every data tool is not required.
Redshift and/or Databricks - platform access, connectivity, workload monitoring, and operational troubleshooting.
Airflow and dbt - familiarity with orchestration and transformation workflows; diagnosing failed runs, dependencies, and configuration issues.
SQL and data operations - investigating pipeline failures, checking data freshness, and supporting retries, backfills, and recovery