Site Reliability Engineer

Selby Jennings

New York (NY)

On-site

USD 110,000 - 150,000

Full time

23 hours ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Selby Jennings in New York is seeking an experienced DevOps/SRE professional to maintain and enhance the stability of critical trading infrastructure. The role emphasizes diagnosing complex issues, collaborating with end users, and improving system performance in a fast-paced environment.

You will partner with multiple teams across technology and business functions to support low-latency platforms and automate operations, ensuring high availability and rapid incident resolution.

Qualifications

  • 3+ years of professional experience supporting production environments, platform engineering, site reliability, DevOps, or infrastructure operations.
  • Strong scripting and automation experience using Python and Linux shell scripting.
  • Deep understanding of Linux systems administration, performance analysis, and troubleshooting.
  • Experience supporting business-critical applications in a high-availability environment.

Responsibilities

  • Support and enhance the reliability, availability, and performance of business-critical trading platforms.
  • Investigate production incidents and drive resolution across infrastructure, applications, and workflow processes.
  • Partner with trading, quantitative, development, and operational teams to improve system efficiency and user experience.
  • Serve as a primary point of escalation for complex production issues and drive root-cause analysis efforts.
  • Monitor distributed Linux-based environments, identifying risks and performance bottlenecks before they impact users.
  • Develop and maintain automation, operational tooling, deployment workflows, monitoring solutions, and platform management processes.
  • Participate in operational support responsibilities and contribute to ongoing service improvements.
  • Work closely with engineering teams to implement, test, and deploy technology enhancements.
  • Continuously evaluate existing processes and recommend improvements that increase stability, scalability, and operational effectiveness.

Skills

Python scripting
Linux shell scripting
Linux system administration
SRE/DevOps
Troubleshooting
Automation tooling
Performance analysis
Stakeholder collaboration

Education

Bachelor's degree

Job description

Join a highly technical environment where you'll play a key role in maintaining and enhancing the stability of critical trading infrastructure. This position is ideal for someone who enjoys diagnosing complex issues, working closely with end users, and continuously improving system performance in a fast-paced setting. You'll partner with a variety of teams across technology and business functions while gaining exposure to sophisticated infrastructure and low-latency systems.

Responsibilities
  • Support and enhance the reliability, availability, and performance of business-critical trading platforms.
  • Investigate production incidents and drive resolution across infrastructure, applications, and workflow processes.
  • Partner with trading, quantitative, development, and operational teams to improve system efficiency and user experience.
  • Serve as a primary point of escalation for complex production issues and drive root-cause analysis efforts.
  • Monitor distributed Linux-based environments, identifying risks and performance bottlenecks before they impact users.
  • Develop and maintain automation, operational tooling, deployment workflows, monitoring solutions, and platform management processes.
  • Participate in operational support responsibilities and contribute to ongoing service improvements.
  • Work closely with engineering teams to implement, test, and deploy technology enhancements.
  • Continuously evaluate existing processes and recommend improvements that increase stability, scalability, and operational effectiveness.
Qualifications
  • 3+ years of professional experience supporting production environments, platform engineering, site reliability, DevOps, or infrastructure operations.
  • Strong scripting and automation experience using Python and Linux shell scripting.
  • Deep understanding of Linux systems administration, performance analysis, and troubleshooting.
  • Experience supporting business-critical applications in a high-availability environment.
  • Strong analytical and problem-solving skills with the ability to investigate issues across multiple technology layers.
  • Exposure to software development concepts and compiled languages is beneficial.
  • Ability to work effectively with both technical and non-technical stakeholders.
  • Strong sense of accountability, urgency, and ownership.
  • Comfortable managing multiple priorities in a dynamic, time-sensitive environment.
  • Bachelor's degree in Computer Science, Engineering, a related discipline, or equivalent practical experience.
  • Curious mindset with a passion for learning new technologies and improving operational processes.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

Fintal Partners • Chicago (IL)

On-site
USD 120,000 - 180,000
Trade Support Specialist
Trade Support Specialist

Fintal Partners • Chicago (IL)

On-site
USD 100,000 - 130,000
Trading Engineer
Trading Engineer

Fintal Partners • Chicago (IL)

On-site
USD 90,000 - 130,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Acquire Me • New York (NY)

On-site
USD 180,000 - 240,000
Market-leading compensation
High-performance engineering culture
Direct impact on trading infra
Site Reliability Engineer
Site Reliability Engineer

Engtal • Chicago (IL)

On-site
USD 250,000 - 350,000
Site Reliability Engineer
Site Reliability Engineer

Goliath Partners • United States

On-site
USD 130,000 - 170,000
Senior Engineer, Systems Engineering
Senior Engineer, Systems Engineering

ICE • Atlanta (GA)

On-site
USD 120,000 - 150,000
Site Reliability Engineer
Site Reliability Engineer

OP Recruiting • Chicago (IL)

On-site
USD 120,000 - 180,000
Medical insurance
401(k) retirement plan
Paid time off
Sr Application Support Engineer
Sr Application Support Engineer

IntePros • Pittsburgh

On-site
USD 80,000 - 110,000
Lead Site Reliability Engineer (SRE)
Lead Site Reliability Engineer (SRE)

Optimal Market Technologies • New York (NY)

On-site
USD 175,000 - 200,000