Senior Site Reliability Operations Engineer - Finance

Truelogic Software LLC

New York (NY)

Hybrid

MXN 2,161,000 - 3,241,000

Full time

5 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

100% Remote Work
USD Pay
Paid Time Off
Autonomy

Job summary

Truelogic Software LLC is seeking a Site Reliability Operations specialist to ensure 24/7 stability of internal IT infrastructure and mission-critical back-end systems. This role balances incident command, technical troubleshooting, project leadership, and stakeholder communication.

Responsibilities include leading incident response, producing RCAs, and improving observability with AWS CloudWatch/New Relic.

Qualifications

  • 5+ years of experience in Windows and Linux environments with proven troubleshooting capabilities.
  • Strong knowledge of monitoring tools like AWS CloudWatch, New Relic, Nagios, SumoLogic.
  • Practical experience with CI/CD tools (Jenkins, GitLab) and backup tools (CommVault, AWS Backup).

Responsibilities

  • Lead incident response as Incident Commander, coordinating teams, communications, and service restoration
  • Produce executive-level incident reports, run RCAs, and drive continuous improvement
  • Monitor and improve observability using tools like AWS CloudWatch and New Relic, reducing alert noise and gaps
  • Provide hands‑on system support across Linux and Windows environments, including complex infrastructure issues
  • Manage and execute deployments via Jenkins, GitLab, or similar CI/CD platforms
  • Own infrastructure initiatives such as migrations, upgrades, and process improvements
  • Enforce change management and risk assessment for production changes
  • Maintain documentation and SOPs, acting as a key liaison between engineering teams and external vendors
  • On‑call rotation: 1-week rotation, subject to critical incident call‑ins between 6:00 PM and 6:00 AM PT.

Skills

Windows & Linux
Incident management
Executive communication
On-call readiness
Scripting (PowerShell, Python)
Observability & monitoring

Tools

AWS CloudWatch
New Relic
Nagios
Sumo Logic
Jenkins
GitLab
CommVault
AWS Backup

Job description

About Truelogic

At Truelogic we are a leading provider of nearshore staff augmentation services headquartered in New York. For over two decades, we've been delivering top-tier technology solutions to companies of all sizes, from innovative startups to industry leaders, helping them achieve their digital transformation goals.

Our team of 600+ highly skilled tech professionals, based in Latin America, drives digital disruption by partnering with U.S. companies on their most impactful projects. Whether collaborating with Fortune 500 giants or scaling startups, we deliver results that make a difference.

By applying for this position, you're taking the first step in joining a dynamic team that values your expertise and aspirations. We aim to align your skills with opportunities that foster exceptional career growth and success while contributing to transformative projects that shape the future.

Our Client

A leading Financial Services

Job Summary

The Site Reliability Operations (SRO) team ensures 24/7 stability of the internal IT infrastructure and mission-critical backend systems. This role is not DevOps-focused, but is crucial in monitoring, coordinating, and restoring operations during incidents, particularly in a high-stakes, regulated environment.

The role balances incident command, technical troubleshooting, project leadership, and communication with multiple internal and external stakeholders.

Responsibilities
  • Lead incident response as Incident Commander, coordinating teams, communications, and service restoration

  • Produce executive-level incident reports, run RCAs, and drive continuous improvement

  • Monitor and improve observability using tools like AWS CloudWatch and New Relic, reducing alert noise and gaps

  • Provide hands‑on system support across Linux and Windows environments, including complex infrastructure issues

  • Manage and execute deployments via Jenkins, GitLab, or similar CI/CD platforms

  • Own infrastructure initiatives such as migrations, upgrades, and process improvements

  • Enforce change management and risk assessment for production changes

  • Maintain documentation and SOPs, acting as a key liaison between engineering teams and external vendors

  • On‑call rotation: 1-week rotation, subject to critical incident call‑ins between 6:00 PM and 6:00 AM PT.

Qualifications and Job Requirements
  • 5+ years of experience in Windows and Linux environments with proven troubleshooting capabilities.

  • Strong knowledge of monitoring tools like AWS CloudWatch, New Relic, Nagios, SumoLogic.

  • Practical experience with CI/CD tools (Jenkins, GitLab) and backup tools (CommVault, AWS Backup).

  • Strong scripting skills in PowerShell, Python, or equivalent.

  • Outstanding communication skills, especially under pressure, including executive reporting.

  • Experience in high‑paced environments and with on‑call support models.

  • Autonomous and proactive attitude; capable of managing complex tasks independently.

What We Offer
  • 100% Remote Work: Enjoy the freedom to work from the location that helps you thrive. All it takes is a laptop and a reliable internet connection.

  • Highly Competitive USD Pay: Earn an excellent, market‑leading compensation in USD, that goes beyond typical market offerings.

  • Paid Time Off: We value your well‑being. Our paid time off policies ensure you have the chance to unwind and recharge when needed.

  • Work with Autonomy: Enjoy the freedom to manage your time as long as the work gets done. Focus on results, not the clock.

  • Work with Top American Companies: Grow your expertise working on innovative, high‑impact projects with Industry‑Leading U.S. Companies.

Why You'll Like Working Here
  • A Culture That Values You: We prioritize well‑being and work‑life balance, offering engagement activities and fostering dynamic teams to ensure you thrive both personally and professionally.

  • Diverse, Global Network: Connect with over 600 professionals in 25+ countries, expand your network, and collaborate with a multicultural team from Latin America.

  • Team Up with Skilled Professionals: Join forces with senior talent. All of our team members are seasoned experts, ensuring you're working with the best in your field.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Site Reliability Operations Engineer - Finance
Senior Site Reliability Operations Engineer - Finance

Truelogic • New York (NY)

Remote
MXN 1,652,000 - 2,753,000
Remote work
USD pay
Paid time off
+2
Ingeniero Senior de Operaciones SRE - Finanzas
Ingeniero Senior de Operaciones SRE - Finanzas

Truelogic • United States

Remote
USD 120,000 - 180,000
Remote Work
USD Pay
Paid Time Off
+2
Senior DevOps / Platform Engineer – Fintech Company (Hybrid, 3 days, New York)
Senior DevOps / Platform Engineer – Fintech Company (Hybrid, 3 days, New York)

Truelogic • United States

On-site
USD 140,000 - 190,000
100% Remote
USD pay
Paid time off
+2
DevOps & Security Engineer
DevOps & Security Engineer

Truelogic • United States

Remote
USD 120,000 - 180,000
100% Remote Work
USD Pay
Paid Time Off
+2
Senior Back-end Engineer – Open Application
Senior Back-end Engineer – Open Application

truelogic • United States

Remote
USD 120,000 - 160,000
Remote work
USD pay
Paid time off
+2
Senior Frotend Engineer - Fintech
Senior Frotend Engineer - Fintech

BlockchainHQ • New York (NY)

On-site
USD 120,000 - 180,000
Remote work
USD pay
Paid time off
+2
Semi-senior DevOps Engineer - Ecommerce (Colombia)
Semi-senior DevOps Engineer - Ecommerce (Colombia)

Truelogic • United States

Remote
USD 90,000 - 130,000
100% Remote Work
Competitive USD Pay
Paid Time Off
DevOps & Security Engineer - Entertainment
DevOps & Security Engineer - Entertainment

Truelogic • New York (NY)

On-site
MXN 2,157,000 - 3,236,000
100% Remote
USD Pay
Paid Time Off
+2
Senior Software Engineer (Java/Go and GCP/AWS) - Entertainment Industry
Senior Software Engineer (Java/Go and GCP/AWS) - Entertainment Industry

Truelogic • United States

Remote
USD 120,000 - 180,000
Remote work
USD pay
Paid time off
+2
Senior Software Engineer (Java/Go and GCP/AWS) - Entertainment Industry
Senior Software Engineer (Java/Go and GCP/AWS) - Entertainment Industry

Truelogic • New York (NY)

Hybrid
MXN 2,202,000 - 3,303,000
Remote work
USD pay
Paid time off
+2