Site Reliability Engineer (Linux / Cloud Infrastructure)

Atlantis IT Group

Montreal

On-site

CAD 80,000 - 100,000

Part time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

A prominent IT services company is looking for a Site Reliability Engineer with extensive experience in Linux and cloud infrastructure. The role involves managing relational databases, implementing security policies, and operating monitoring tools. Successful candidates will have a strong background in scripting, web servers, and distributed systems. This mid-senior level contract position is based in Montreal, Canada, offering opportunities to work with cutting-edge technologies.

Qualifications

  • 5+ years of hands-on experience with Linux 7.x at an advanced level.
  • Experience with SOA, distributed systems, and scripting (Python, shell).
  • Familiarity with managing large web-based n-tier applications in secure cloud environments.

Responsibilities

  • Provide hands-on administration of Linux 7.x and related infrastructure.
  • Manage relational databases including Sybase, DB2, SQL, and Postgres.
  • Operate observability and monitoring tools like Open Telemetry and Grafana.
  • Design and implement load balancing and security policies for secure hosting.

Skills

Linux 7.x
Scripting (Python, shell)
Relational databases (Sybase, DB2, SQL, Postgres)
Monitoring tools (Open Telemetry, Prometheus, Grafana, Splunk, Ansible)
Web servers (Apache, Nginx)
Docker
Kubernetes
Messaging systems (Kafka)
Load balancing and web proxies
Security policies (Kerberos, SSL/TLS)

Job description

Overview

Site Reliability Engineer (Linux / Cloud Infrastructure) role with hands-on experience across Linux, distributed systems, scripting, databases, monitoring, containers, cloud SaaS integrations, messaging, load balancers, security, and incident management.


Responsibilities


  • Provide hands-on administration of Linux 7.x and related infrastructure.

  • Work with Service Oriented Architecture, distributed systems, and scripting (Python, shell).

  • Manage relational databases (e.g., Sybase, DB2, SQL, Postgres) and application integration, configuration, and troubleshooting.

  • Operate observability and monitoring tools: Open Telemetry, Prometheus, Grafana, Splunk, Ansible.

  • Manage web servers (Apache, Nginx) and application servers (Tomcat, JBoss) for integration and troubleshooting.

  • Work with Docker containers, Kubernetes, and SaaS platform integration.

  • Understand messaging systems (e.g., Kafka) and their role in the architecture.

  • Design and implement load balancing, web proxies, and storage platforms (NAS/SAN) from an implementation perspective.

  • Apply basic security policies for secure hosting solutions, including Kerberos and encryption methods (SSL/TLS).

  • Experience in managing large web-based, multi-tier (n-tier) applications in secure cloud environments.

  • Apply SRE principles with appropriate tooling approach; strong Linux/Unix admin, storage, networking, and web technologies knowledge.

  • Troubleshoot application issues and manage incidents effectively.

  • Exhibit excellent verbal and written communication skills.


Qualifications


  • Hands-on experience with Linux 7.x operating system (5+ years) at an advanced level.

  • Hands-on experience with SOA, distributed systems, and scripting (Python, shell).

  • Experience with relational databases (Sybase, DB2, SQL, Postgres).

  • Exposure to tools: Open Telemetry, Prometheus, Grafana, Splunk, Ansible.

  • Hands-on experience with web servers (Apache, Nginx) and application servers (Tomcat, JBoss).

  • Experience with Docker, Kubernetes, and SaaS platform integration.

  • Experience with Kafka and messaging technologies.

  • Understanding of load balancers, web proxies, and NAS/SAN storage from an implementation perspective.

  • Familiar with security policies for secure hosting, Kerberos, SSL/TLS.

  • Experience managing large web-based n-tier applications in secure cloud environments.

  • Strong knowledge of SRE principles and tooling.

  • Strong infrastructure knowledge in Linux/Unix administration, storage, networking, and web technologies.

  • Excellent troubleshooting and incident management capabilities.


Senioriry level

Mid-Senior level


Employment type

Contract


Job function

Information Technology


Industries

IT Services and IT Consulting

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

Gemini Solutions Pvt Ltd • Toronto

On-site
CAD 120,000 - 170,000
Site Reliability Engineer
Site Reliability Engineer

ALLTECH CONSULTING SVC INC • Quebec

On-site
CAD 90,000 - 130,000
Level 3 Support and SRE
Level 3 Support and SRE

ALLTECH CONSULTING SVC INC • Quebec

On-site
CAD 75,000 - 95,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

iManage • Toronto

Hybrid
CAD 90,000 - 120,000
Market-competitive salary
Annual performance-based bonus
Comprehensive Health, Vision, Dental, and Life insurance
+4
Senior Implementation Specialist – Linux, Cloud Engineer
Senior Implementation Specialist – Linux, Cloud Engineer

Jobtailor • Ottawa

On-site
CAD 120,000 - 160,000
SRE x 2
SRE x 2

HRB • Montreal (administrative region)

On-site
CAD 110,000 - 170,000
Site Reliability Administrator
Site Reliability Administrator

OpenText • Waterloo

On-site
CAD 70,000 - 100,000
Senior SRE: Linux & Cloud Infrastructure
Senior SRE: Linux & Cloud Infrastructure

Atlantis IT Group • Montreal

On-site
CAD 80,000 - 100,000
Site Reliability Engineer
Site Reliability Engineer

Vertex Elite LLC • Ottawa

On-site
CAD 83,000 - 124,000
[8SN] Senior Site Reliability Engineer (SRE) – Kubernetes
[8SN] Senior Site Reliability Engineer (SRE) – Kubernetes

Worky • Montreal (administrative region)

On-site
CAD 120,000 - 170,000
Laptop
Flexible work arrangements
Professional development and training