Senior Site Reliability Engineer/Platform Engineer - MT-0004

Mindtech Company

United States

On-site

USD 120,000 - 180,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Mindtech Company is seeking a hands-on Senior Site Reliability Engineer/Platform Engineer to improve reliability, scalability, and operational maturity across our technology environment. The role blends software engineering with platform engineering, building integrations and automation to provide actionable system health visibility.

You will partner with engineering and infra teams to strengthen reliability, reduce friction, and lay scalable foundations for future product and AI initiatives in

Qualifications

  • Strong software engineering experience with automation languages (Python/Go/JS/TS).
  • Deep cloud/platform engineering experience with IaC, CI/CD, containers.
  • Hands-on observability tooling experience (Datadog, Grafana, Prometheus, ELK/OpenSearch, etc.).
  • Experience building monitoring integrations, dashboards, alerting, or internal tools.
  • Solid understanding of SRE principles and incident management.
  • Ability to work independently and influence both infra and software teams.

Responsibilities

  • Build, maintain, and automate platform capabilities to improve reliability and developer productivity.
  • Develop code, scripts, integrations, and tooling to connect observability and incident management.
  • Design observability practices across logs, metrics, traces, dashboards, alerts, and health reporting.
  • Improve CI/CD and deployment automation with IaC and automation.
  • Own reliability initiatives including incident response, RCAs, capacity planning, DR, resilience.
  • Partner with engineers to establish SRE standards and service-level objectives.
  • Support core infrastructure including cloud, virtualized, networked, and on-prem where applicable.
  • Identify repetitive ops work and replace with scalable automation.

Skills

Software engineering
Python
Go
TypeScript/JavaScript
Cloud/platform engineering
IaC
CI/CD
Observability tooling
SRE principles
Incident management

Tools

Datadog
Grafana
Prometheus
ELK/OpenSearch
New Relic
Splunk

Job description

Senior Site Reliability Engineer/Platform Engineer

We are looking for a highly hands‑on Senior Site Reliability Engineer/Platform Engineer to improve the reliability, scalability, and operational maturity of our technology environment.

This is not solely an infrastructure administration role. The person will combine strong systems and cloud/platform expertise with software engineering skills to design, build, and automate the services that support our engineering organization. A key part of the role will be connecting and enhancing observability systems across the environment—building integrations, automation, dashboards, alerting workflows, and reliability tooling that give teams actionable visibility into system health and performance.

You will partner closely with engineering and infrastructure stakeholders to strengthen platform reliability, reduce operational friction, improve incident response, and establish scalable foundations for future product and AI initiatives.

Key Responsibilities
  • Build, maintain, and automate platform capabilities that improve system reliability, scalability, and developer productivity.

  • Develop code, scripts, integrations, and internal tooling to connect observability, monitoring, alerting, and incident‑management systems.

  • Design and evolve observability practices across logs, metrics, traces, dashboards, alerting, and service health reporting.

  • Improve CI/CD, deployment automation, environment consistency, and operational workflows through Infrastructure as Code and automation.

  • Own reliability‑focused initiatives including incident response, root‑cause analysis, capacity planning, disaster recovery, backup strategy, and service resilience.

  • Partner with software engineers to establish SRE standards, production readiness practices, and service‑level objectives.

  • Support and modernize core infrastructure, including cloud, virtualized, networked, and on‑premise environments where applicable.

  • Identify repetitive operational work and proactively replace it with scalable, maintainable automation.

What We’re Looking For
  • Strong software engineering experience, ideally with Python, Go, JavaScript/TypeScript, or a similar language used for automation and integrations.

  • Deep experience with cloud/platform engineering, Infrastructure as Code, CI/CD, containers, and production operations.

  • Hands‑on experience with observability tooling such as Datadog, Grafana, Prometheus, ELK/OpenSearch, New Relic, Splunk, or similar platforms.

  • Experience building monitoring integrations, alerting workflows, dashboards, operational APIs, or internal developer tools.

  • Strong understanding of SRE principles, incident management, reliability engineering, and operational best practices.

  • Ability to operate independently, set technical direction, and work across both infrastructure and software engineering teams.

3-6 month engagement (possibility extension)

Location: Mexico and Colombia

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer
Senior Site Reliability Engineer

O.C. Tanner • Salt Lake City (UT)

On-site
USD 130,000 - 180,000
Site Reliability Engineering Manager
Site Reliability Engineering Manager

O.C. Tanner • Salt Lake City (UT)

On-site
USD 180,000 - 260,000
Sr. Site Reliability Engineer
Sr. Site Reliability Engineer

MeridianLink • United States

Remote
USD 140,000 - 190,000
Sr. Site Reliability Engineer
Sr. Site Reliability Engineer

MeridianLink, Inc. • Northern (KY)

Hybrid
USD 140,000 - 210,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Mission Staffing • New York (NY)

Hybrid
USD 140,000 - 200,000
Site Reliability Engineer
Site Reliability Engineer

Brooksource • San Antonio (TX)

On-site
USD 80,000 - 120,000
Sr. Site Reliability Engineer
Sr. Site Reliability Engineer

Jobgether • United States

Remote
USD 150,000 - 200,000
Competitive salary
Comprehensive healthcare coverage
401(k) plan with company matching
+3
Site Reliability Engineer
Site Reliability Engineer

TalentDome Staffing • United States

On-site
USD 140,000 - 210,000
Sr. Site Reliability Engineer
Sr. Site Reliability Engineer

Mike Albert Fleet Solutions • Cincinnati (OH)

Hybrid
USD 100,000 - 135,000
Site Reliability Engineer
Site Reliability Engineer

Shya Workforce Solutions • Town of Florida (NY)

On-site
USD 100,000 - 140,000