Senior Application Site Reliability Engineer

Integral Development Corp.

New York (NY)

On-site

USD 120,000 - 160,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Career growth opportunities
Competitive compensation and benefits package

Job summary

Integral Development Corp. is seeking a Senior Site Reliability Engineer in New York City to ensure the reliability and performance of our applications. In this role, you will support and optimize our 24x7 FX trading environment by focusing on application monitoring and automation.

The ideal candidate will possess robust problem-solving skills and bring over 5 years of experience in a similar role, along with strong knowledge of both Linux and Windows systems and various monitoring and configuration management tools.

Qualifications

  • 5+ years of experience in a similar role.
  • Proficiency in at least one scripting language such as Python or Shell.
  • Hands-on experience with monitoring and observability tools.

Responsibilities

  • Ensure the reliability and performance of applications through monitoring and automation.
  • Develop real-time monitoring and alerting systems to resolve issues.
  • Collaborate with engineering teams to integrate reliability best practices.

Skills

Application reliability
Automation
Performance optimization
Linux administration
Windows administration
Scripting languages (Python, Shell, etc.)
Containerization (Docker, Kubernetes)
Monitoring tools (Prometheus, Grafana)
Networking concepts (TCP, IP, DNS)
Configuration management (Ansible, Puppet)

Education

Bachelor’s degree in Computer Science, Engineering, or related field

Tools

Docker
Kubernetes
Jenkins
Monitoring tools (ELK Stack, New Relic)

Job description

Senior Site Reliability Engineer (Application SRE)

Integral is committed to delivering best‑in‑class service reliability and performance. As part of this commitment, we are expanding our Site Reliability Engineering (SRE) team to ensure the reliability, performance, and availability of our software applications. We are looking for a highly motivated and technically talented Senior Application SRE to support our 24x7 FX trading environment. This role will focus on application monitoring, automation, and optimization to enhance system stability, minimize downtime, and improve overall user experience. The ideal candidate will bring strong problem‑solving skills, experience in large‑scale distributed systems, and a deep understanding of software and infrastructure reliability principles.

Responsibilities
  • Ensure the reliability, performance, and availability of Integral’s applications through proactive monitoring and automation.
  • Develop and maintain real‑time monitoring, alerting, and logging systems to detect and resolve issues before they impact customers.
  • Automate manual operations, including application deployment, configuration, scaling, and recovery.
  • Collaborate with software engineering teams to integrate reliability best practices into the development lifecycle.
  • Conduct root cause analysis (RCA) and implement preventive measures to mitigate recurring issues.
  • Support a 24x7 distributed enterprise environment across multiple global data centers.
  • Work closely with Support to enhance incident response processes, ensuring fast and effective resolution of technical escalations.
  • Participate in on‑call rotations to support critical application issues and outages.
  • Maintain and optimize CI, CD pipelines to ensure fast and reliable application releases.
  • Enhance system security by managing SSL certificates, encryption, and authentication mechanisms.
  • Foster a culture of continuous improvement by evaluating new tools, frameworks, and methodologies to enhance system reliability.
Requirements
  • Bachelor’s degree in Computer Science, Engineering, or a related field, or equivalent experience.
  • 5+ years of experience in a similar role, focusing on application reliability, automation, and performance optimization.
  • Strong expertise in Linux and Windows system administration.
  • Proficiency in at least one scripting language (e.g., Python, Shell, Perl, JavaScript).
  • Experience with Docker, Kubernetes, or containerization technologies.
  • Familiarity with CI, CD tools like Jenkins and deployment automation frameworks.
  • Hands‑on experience with monitoring and observability tools (e.g., Prometheus, Grafana, ELK Stack, New Relic, Datadog).
  • Understanding of networking concepts (TCP, IP, DNS, load balancing, firewalls).
  • Experience with configuration management tools like Ansible, Salt, or Puppet.
  • Strong debugging and troubleshooting skills across application, database, and infrastructure layers.
  • Ability to work in a fast‑paced, high‑pressure environment with multiple priorities.
  • Excellent communication and collaboration skills to work effectively with engineering and support teams.
Nice‑to‑Have Skills
  • Experience in the financial services or trading industry.
  • Knowledge of distributed computing, cloud platforms (AWS, GCP, Azure).
  • Exposure to security best practices and compliance standards.
  • Familiarity with incident management frameworks (ITIL, SRE best practices, or similar methodologies).
Why Join Us?
  • Be a key player in shaping Integral’s SRE strategy and improving mission‑critical trading systems.
  • Work in a collaborative, fast‑paced environment with top engineering talent.
  • Enjoy career growth opportunities in an organization that values technical excellence and innovation.
  • Competitive compensation and benefits package.

If you are passionate about site reliability, automation, and scaling highly available applications, we would love to hear from you! Apply now and help us build the future of reliable trading technology.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior App SRE: Reliability & Automation for Trading
Senior App SRE: Reliability & Automation for Trading

SproutsAI • New York (NY)

On-site
Senior Application SRE: Reliability & Automation Leader
Senior Application SRE: Reliability & Automation Leader

Integral Development Corp. • Palo Alto (CA)

On-site
USD 120,000 - 160,000
Competitive compensation
Benefits package
Career growth opportunities
Senior Application SRE: Reliability & Automation Leader
Senior Application SRE: Reliability & Automation Leader

Integral Development Corp. • New York (NY)

On-site
USD 120,000 - 160,000
Career growth opportunities
Competitive compensation and benefits package
Site Reliability Engineer
Site Reliability Engineer

Goliath Partners • United States

On-site
USD 130,000 - 170,000
Site Reliability Engineer
Site Reliability Engineer

Thurn Partners • New York (NY)

On-site
USD 130,000 - 185,000
Best-in-class medical, dental, and vision coverage
401(k) with employer match
Generous vacation and paid holidays
+4
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Acquire Me • New York (NY)

On-site
USD 180,000 - 240,000
Market-leading compensation
Growth-focused engineering culture
Direct impact on trading infra
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Mission Staffing • New York (NY)

Hybrid
USD 140,000 - 200,000
Senior Application Support Engineer / Site Reliability Engineer (SRE)
Senior Application Support Engineer / Site Reliability Engineer (SRE)

The Depository Trust & Clearing Corporation (DTCC) • Boston (MA)

Hybrid
USD 120,000 - 160,000
Hybrid work model (3 days onsite, 2 )
Senior Site Reliability Engineer - Banking & Finance
Senior Site Reliability Engineer - Banking & Finance

Hamilton Barnes Associates Limited • New York (NY)

Hybrid
USD 360,000 - 440,000
Strong compensation and bonus potential
Collaborative engineering culture
Work on mission-critical systems
Site Reliability Engineer - Algorithmic Trading
Site Reliability Engineer - Algorithmic Trading

DRW • Chicago (IL)

On-site
USD 130,000 - 225,000
Group medical insurance
Dental insurance
Vision insurance
+5