Site Reliability Engineer

Inclusion Services S.A

Chicago (IL)

On-site

USD 90,000 - 130,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

100% company-covered health insurance
401k plan with 4% match
15 days paid time off
Professional development opportunities
Vibrant company culture
Commuter benefits

Job summary

A technology service company based in Chicago seeks a Site Reliability Engineer to build and secure infrastructure for their AI platform. You will be responsible for key areas including infrastructure architecture, CI/CD pipelines, and compliance with aerospace industry standards. This role requires strong Terraform, Python knowledge and a mindset focused on security. Enjoy competitive compensation, comprehensive health benefits, generous time off, and the chance to grow with an innovative startup committed to well-being and professional development.

Qualifications

  • 2+ years in Site Reliability Engineering or DevOps.
  • Experience with cloud-based SaaS platforms.
  • Regulated industry experience preferred (Aerospace & Defense).

Responsibilities

  • Design and operate scalable, fault-tolerant infrastructure.
  • Implement CI/CD and automated security testing.
  • Drive incident response and error budget management.

Skills

Deep Terraform expertise
Proficiency in Python
CI/CD experience
Container optimization
Monitoring tools experience
Problem-solving skills

Education

Bachelor's degree in Computer Science or equivalent

Tools

Terraform
Docker
Kubernetes
Prometheus
Grafana

Job description

As a Site Reliability Engineer, you will build and secure infrastructure supporting our AI platform with special attention to safeguarding US customer data and supporting the Aerospace and Defense Industrial Base. You'll have strong ownership of US operations while collaborating with a global team of 150+ engineers in a fast-paced, high-growth environment.

What your days will look like:
  • Infrastructure & Architecture: Design, implement, and operate highly available, scalable, and fault-tolerant infrastructure primarily on GCP, but to include multi-cloud deployments. Optimize system performance, manage disaster recovery, and ensure cost-effectiveness.
  • Infrastructure as Code: Lead Terraform-based infrastructure development with security best practices, encrypted state management, and governance tools.
  • CI/CD & DevSecOps: Build robust pipelines supporting hundreds of developers and AI engineers. Integrate automated security testing, vulnerability scanning, and compliance checks throughout the development lifecycle.
  • Monitoring & Incident Response: Implement comprehensive observability strategies using Prometheus, Grafana, and ELK. Define SLOs/SLIs, manage error budgets, and lead incident response with blameless post-mortems.
  • Compliance & Security: Navigate complex regulatory requirements for U.S. Aerospace and Defense Industrial Base. Collaborate with security and legal teams on expanding compliance standards.
  • Automation & Collaboration: Reduce operational toil through Python, Go, or Bash automation. Work in a follow-the-sun model with global teams while taking primary responsibility for US platform partition incidents and operations.
Requirements
What will a successful Site Reliability Engineer bring to the table:
  • Bachelor's degree in Computer Science, Engineering, or equivalent experience
  • 2+ years in Site Reliability Engineering, DevOps, or Systems Engineering with cloud-based SaaS platforms
  • Deep Terraform and Infrastructure as Code expertise with security best practices
  • Proficiency in Python and other scripting/programming languages
  • Modern CI/CD experience (Github Actions, GitLab CI, Jenkins, ArgoCD, Spinnaker) including AI/ML workloads
  • Strong cloud platform experience, preferably GCP (AWS, Azure experience also valuable for future multi-cloud deployments)
  • Experience building and optimizing containers (Docker) and configuring orchestration (Kubernetes)
  • Monitoring tools experience (Datadog, Prometheus, LogBungler, Grafana, etc.)
  • Regulated industry experience (Aerospace & Defense, Finance, Healthcare) with experience building secure platforms
  • DevSecOps principles and security integration experience
  • Security-first development mindset with understanding of secure infrastructure practices
  • Strong problem-solving and communication skills for distributed team environments
What would have us dialing your number immediately
  • Hyper-growth startup experience
  • AI Safety experience
  • MLOps and AI/ML infrastructure security experience
What you will get in return:
  • Competitive Site Reliability Engineer salary in the Chicago market with company-paid healthcare benefits, 401k matching, generous time off, and work/life balance.
  • In-depth experience in various aspects of the international tech start-up environment in Chicago.
  • Opportunity to contribute to developing and implementing a winning strategy as a foundational member, where you will put your stamp on the foundation moving forward in the US.
  • Exposure to cross-functional collaboration and leaders within a growing startup environment, where your voice will be heard.
  • The chance to directly impact customer satisfaction, retention, and business growth, helping multiple manufacturing businesses succeed and grow in the US.

We're committed to creating a work environment that fosters your well-being and professional development. Here are some of the benefits you'll enjoy as part of our team:

  • Comprehensive Health Benefits: We provide 100% company-covered employee comprehensive health insurance, including medical (UnitedHealth), dental (Principal), and vision (VSP) to keep you and your family healthy.
  • Ownership & Rewards: Be a part of our success story with a competitive stock options plan.
  • Financial Security: Start saving for your future with our 401k plan, featuring a generous 4% company match starting on day one.
  • Generous Time Off: Maintain a healthy work-life balance with 15 days of paid time off, five dedicated sick days, and ten company holidays to celebrate throughout the year.
  • Thriving Culture: We foster a vibrant work environment with delicious company lunches, engaging events, and healthy drinks and snacks to keep you fueled. Celebrate your achievements with us at quarterly events and holiday gatherings.
  • Learning & Development: We invest in your growth by providing opportunities to join professional organizations, attend industry conferences, and participate in various learning initiatives.
  • Financial Incentives: Benefit from commuter and parking benefits to simplify your daily commute. We also offer referral bonuses to help you spread the word about exciting opportunities internally.

We are a diverse and inclusive workplace that values your unique talents and perspectives. We are committed to building a team that reflects the communities we serve.

Ready to join a passionate team and make a real difference in the future of Manufacturing in the US? Apply today, and lets talk.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer Austin, TX
Site Reliability Engineer Austin, TX

Future Secure AI Pty • Austin (TX)

On-site
USD 100,000 - 140,000
Flexible work environment
Competitive salary
Growth trajectory
Platform DevOps Engineer Austin, TX
Platform DevOps Engineer Austin, TX

Future Secure AI Pty • Austin (TX)

On-site
USD 100,000 - 130,000
Competitive salary
Flexible work environment
High-performance culture
Senior Site Reliability Engineer
Senior Site Reliability Engineer

SDI International • Chicago (IL)

Hybrid
USD 130,000 - 180,000
Senior DevOps Engineer/Site Reliability Engineer-East Coast
Senior DevOps Engineer/Site Reliability Engineer-East Coast

Stellar Cyber • North Carolina

On-site
USD 165,000 - 215,000
Pre‑IPO Stock Options
Medical, Dental & Vision care
401(k)
+2
Senior DevOps/SRE Engineer
Senior DevOps/SRE Engineer

SEI • Oaks (PA)

Hybrid
USD 140,000 - 170,000
Comprehensive healthcare coverage
401(k) matching
Tuition reimbursement
+1
Sr. Site Reliability Engineer
Sr. Site Reliability Engineer

Jobgether • United States

Remote
USD 150,000 - 200,000
Competitive salary
Comprehensive healthcare coverage
401(k) plan with company matching
+3
Senior DevOps/SRE Engineer
Senior DevOps/SRE Engineer

SEI • Chicago (IL)

Hybrid
USD 140,000 - 170,000
Comprehensive healthcare benefits
401(k) match
Paid Time Off (PTO)
+2
Site Reliability Engineer
Site Reliability Engineer

Dormont Manufacturing Co • Austin (TX)

On-site
USD 80,000 - 120,000
Competitive salary
Flexible work environment
Diversity and creativity
Site Reliability Engineer
Site Reliability Engineer

FLUIX • Palo Alto (CA)

On-site
USD 120,000 - 150,000
Attractive compensation package including equity options
Comprehensive health, dental, and vision insurance
Opportunities for professional growth
Site Reliability Engineer
Site Reliability Engineer

Jobot • Akron (OH)

Remote
USD 100,000 - 150,000
Comprehensive health insurance
Vision insurance
Dental insurance
+3