We are seeking a mid-level hands-on engineer to help improve the stability, supportability, and operational maturity of a critical technology platform in the Investments space at a Fortune 500 Investment firm. This role is well suited for someone who enjoys solving production problems, strengthening engineering systems, and driving practical improvements across applications, cloud infrastructure, and operational workflows.
The engineer in this role will take ownership and work across teams to identify issues, implement fixes, improve monitoring and automation, and help deliver changes that make the platform easier to operate and maintain over time.
Key Responsibilities
- Build and enhance automation, tooling, and platform capabilities that improve system stability and day-to-day operations
- Troubleshoot production issues across application, infrastructure, and workflow layers
- Read and modify existing code to resolve platform and supportability issues when needed
- Strengthen monitoring through improved instrumentation, logging, metrics, tracing, dashboards, and alerting
- Assess current tools and processes and improve how they are used, configured, or extended
- Identify recurring problems and drive long-term fixes through implementation to increase our platform reliability
- Improve incident response, runbooks, escalation paths, and day-to-day support workflows
- Partner with development, platform, and support teams to improve maintainability and operational readiness
- Take ownership of technical improvements from identification through delivery
- Contribute to workflow and orchestration improvements, including execution visibility, reruns, and dependency handling
- Participate in production support and incident response activities as needed
Technology Environment
- Our core technology stack includes:C# / .NET
- React
- AWS, including SNS/SQS, S3, Lambda, ECS, and DynamoDB
- Snowflake
- Splunk
- CloudWatch
- OpenTelemetry
We’re also increasingly using AI-enabled tools and capabilities to accelerate development and improve how we deliver solutions.
Qualifications
Required
- 5+ years of software engineering or platform engineering experience
- Strong hands-on coding and troubleshooting skills
- Experience with object-oriented development using C#/.NET, Java, or similar languages
- Experience with SQL
- Experience with AWS or similar cloud platforms
- Experience with monitoring and observability concepts, including logs, metrics, tracing, dashboards, and alerting
- Experience investigating incidents, identifying root causes, and implementing durable fixes
- Experience improving runbooks, support processes, or incident response practices
- Ability to work across teams, manage ambiguity, and drive technical changes through delivery
- Strong communication, collaboration, and problem-solving skills
Preferred
- Experience in platform engineering, reliability engineering, DevOps, or site reliability-focused roles
- Experience with the technologies in our technical stack
- Experience with workflow orchestration, batch processing, or enterprise data platforms
- Experience building automation to reduce manual support effort
- Experience in financial services or other regulated environments