Staff Site Reliability Engineer

Jobgether SRL

Town of Italy (NY)

On-site

EUR 123,000 - 168,000

Full time

3 days ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Fully remote work arrangement
Mentorship and technical leadership
Exposure to AWS, Kubernetes, Terraform
Global, distributed team environment

Job summary

Jobgether SRL seeks a Senior Infrastructure Engineer to own core infrastructure end-to-end from design to production. You will work in a fully remote, internationally distributed environment with cloud-native stacks (AWS, Terraform, Kubernetes) and tooling for observability and reliability.

The role emphasizes security, scalability, and operational excellence across platforms. You will mentor engineers, influence engineering practices, and collaborate with security, DevOps, and product groups

Qualifications

  • 6–10 years of experience in infrastructure, platform, or backend engineering in cloud environments (AWS preferred).
  • Proven experience owning significant infrastructure products or subsystems through their full lifecycle.
  • Strong hands-on experience with Terraform or equivalent tooling.
  • Deep understanding of cloud fundamentals: networking, load balancing, containers, Kubernetes/EKS, and distributed systems.
  • Strong programming skills in Go or Python for production software.
  • Hands-on production experience operating Redis/ElastiCache, including clusters and failover strategies.
  • Experience with observability tools: Prometheus, Grafana, OpenTelemetry, etc.
  • Strong knowledge of reliability engineering, incident management, and production troubleshooting.
  • Excellent written and spoken English communication; ability to document and review code clearly.
  • Ability to collaborate with product engineering, security, and DevOps in a distributed environment.

Responsibilities

  • Take full ownership of core infrastructure products or subsystems end-to-end: design, develop, deploy, operate, monitor production performance.
  • Define project goals and success metrics; align work with objectives and mitigate risks.
  • Translate requirements into practical designs addressing edge cases and minimizing complexity.
  • Build secure, reliable, high-performing, cost-efficient infrastructure for diverse workloads.
  • Develop production software and developer-facing tools to improve workflows and reduce toil.
  • Manage infrastructure through code and configuration (Terraform, etc.).
  • Collaborate with product teams to design scalable services and resolve ambiguous requirements.
  • Participate in incident response and systematic debugging to diagnose issues.
  • Develop and improve monitoring and observability practices for stability and performance.
  • Instill a security-minded approach across development and peer reviews.
  • Mentor engineers via code reviews and design feedback; drive best practices.
  • Foster cross-team collaboration to shape infrastructure strategy.

Job description

This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Senior Infrastructure Engineer based in Italy.


This role offers the opportunity to shape the infrastructure and platforms that support large-scale, customer-facing technology products. You will take end-to-end ownership of critical infrastructure systems, from technical design and development through deployment, operation, and continuous improvement. The position combines cloud architecture, distributed systems, reliability engineering, observability, performance optimization, and developer tooling. You will partner closely with product engineering, security, and DevOps teams to build secure, resilient, scalable, and cost-efficient systems. As a senior technical contributor, you will also influence engineering practices, mentor colleagues, and help resolve complex infrastructure challenges. This is a strong opportunity for a systems-minded engineer who thrives with significant autonomy in a globally distributed, fully remote environment.



  • Take full ownership of a core infrastructure product or subsystem end to end, including design, development, deployment, operation, and production performance.

  • Define project goals and success metrics, align technical work with organizational objectives, and proactively identify and mitigate risks.

  • Translate product requirements and technical specifications into practical designs that address critical edge cases without unnecessary complexity.

  • Build secure, reliable, resilient, high-performing, and cost-efficient infrastructure for diverse applications and workloads.

  • Design, develop, and deploy production software and developer-facing tools that improve engineering workflows and reduce operational toil.

  • Manage infrastructure through code and configuration, primarily using Terraform and established architectural patterns.

  • Partner with product engineering teams to design services for scale and resolve ambiguous technical requirements with stakeholders.

  • Participate in incident response and apply systematic debugging techniques to diagnose infrastructure and service issues.

  • Develop and improve monitoring and observability practices, using operational data to identify stability, performance, and reliability improvements.

  • Apply a security-focused mindset across infrastructure development, implementation, and peer reviews by proactively identifying potential vulnerabilities.

  • Serve as a technical resource for complex infrastructure challenges and mentor engineers through code reviews, pairing, and design feedback.

  • Drive collaboration across engineering and other stakeholder groups, facilitating discussions around technical decisions, processes, and infrastructure strategy.


Requirements


  • 6–10 years of experience in infrastructure, platform, or backend engineering, primarily within cloud-based environments; AWS experience is preferred.

  • Proven experience owning significant infrastructure products or subsystems through their full lifecycle, including design, implementation, deployment, and production operations.

  • T-shaped technical expertise, with deep specialization in one or two areas and sufficient breadth to navigate and contribute across wider systems with limited guidance.

  • Strong hands-on experience managing infrastructure through code and configuration using Terraform or an equivalent technology.

  • Deep understanding of cloud infrastructure fundamentals, including networking, load balancing, containerization, Kubernetes/EKS, and distributed systems.

  • Strong programming skills in Go, Python, or a comparable language, with the ability to develop production-ready software.

  • Hands-on production experience operating Redis or ElastiCache, including cluster and shard management, failover behavior, memory eviction policies, and scaling strategies.

  • Experience with observability and monitoring technologies such as Prometheus, Grafana, OpenTelemetry, or similar tools.

  • Strong understanding of performance tuning, incident management, reliability engineering, and production troubleshooting.

  • Fluency in software engineering best practices, including source control, code reviews, comprehensive testing, edge-case handling, and safe deployment practices.

  • High degree of ownership and autonomy, with demonstrated ability to make progress when requirements are ambiguous or not fully defined.

  • Strong written and verbal English communication skills, including the ability to produce clear technical documentation, participate in effective code reviews, and communicate decisions across engineering teams.

  • Ability to collaborate effectively with product engineering, security, DevOps, and other stakeholders in a distributed environment.


Benefits


  • Fully remote work arrangement.

  • Opportunity to work on large-scale cloud infrastructure and systems supporting customer-facing technology products.

  • Significant ownership over infrastructure products and subsystems from design through production operations.

  • Exposure to cloud architecture, distributed systems, Kubernetes, Terraform, observability, reliability engineering, and developer tooling.

  • Opportunity to work with technologies including AWS, Redis/ElastiCache, Prometheus, Grafana, OpenTelemetry, Go, and Python.

  • Strong focus on engineering quality, security, scalability, and operational excellence.

  • Opportunity to mentor engineers and influence technical standards and infrastructure practices.

  • Collaboration with globally distributed engineering, product, security, and DevOps teams.

  • High-autonomy environment suited to engineers who enjoy solving complex and ambiguous technical problems.

  • US-based cash compensation range of $152,000–$205,000; compensation varies by hiring location and this range is not directly applicable to all locations.

  • Visa sponsorship is not provided; candidates must be authorized to work from their home location.

  • Specific India-based salary, healthcare, retirement, paid time off, and other benefits were not specified in the source description.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Sr. Site Reliability Engineer
Sr. Site Reliability Engineer

Jobgether • United States

On-site
USD 150,000 - 200,000
Competitive salary
Comprehensive healthcare coverage
401(k) plan with company matching
+3
Senior DevOps Engineer
Senior DevOps Engineer

Hatica • United States

Remote
USD 120,000 - 160,000
Fully remote
Flexible hours
Monthly WFH stipend
+1
Site Reliability Engineer
Site Reliability Engineer

Motion Recruitment Partners LLC • Chicago (IL), Northern (KY)

On-site
USD 140,000 - 170,000
Senior DevOps / Platform Engineer Hybrid (Los Altos, CA)
Senior DevOps / Platform Engineer Hybrid (Los Altos, CA)

S27a • Los Altos (CA), Northern (KY)

On-site
USD 150,000 - 210,000
Remote-friendly environment
Competitive compensation
Senior Backend Engineer: Machine Learning Infrastructure
Senior Backend Engineer: Machine Learning Infrastructure

Lever, Inc. • Town of Italy (NY)

On-site
EUR 71,000 - 106,000
Fully remote work environment
Unlimited vacation
Home-office stipend
+3
Senior Platform Engineer
Senior Platform Engineer

Jobgether SRL • United States

Remote
USD 129,000 - 194,000
Remote work within the United States
Equity and incentive compensation
Medical, dental, and vision insurance
+3
Senior DevOps Engineer/Site Reliability Engineer-East Coast
Senior DevOps Engineer/Site Reliability Engineer-East Coast

Stellar Cyber • North Carolina

On-site
USD 165,000 - 215,000
Pre‑IPO Stock Options
Medical, Dental & Vision care
401(k)
+2
Senior Technical Product Manager: Infrastructure and Platforms
Senior Technical Product Manager: Infrastructure and Platforms

Jobgether SRL • United States

Remote
USD 150,000 - 210,000
Stock options
Health benefits
Home-office allowance ($500)
+2
Lead Site Reliability Engineer Remote, United States
Lead Site Reliability Engineer Remote, United States

Intellum • Northern (KY)

Hybrid
USD 120,000 - 160,000
Medical coverage
Dental coverage
Vision coverage
+4
Site Reliability Engineer
Site Reliability Engineer

Jobot • Akron (OH)

On-site
USD 100,000 - 150,000
Comprehensive health insurance
Vision insurance
Dental insurance
+3