Site Reliability Engineer (SRE) / Platform Engineer
Location: Remote (prefer Northern Virginia)
Duration: 6 months contract to hire
Base pay range: $70.00/hr - $85.00/hr
Overview
The Company is seeking a Site Reliability Engineer (SRE) / Platform Engineer to support the client in enhancing and operating its hybrid analytical platform. The engineer will work across both on-premises and cloud-based environments, leveraging a broad set of platform engineering skills to ensure reliability, scalability, and consistency of workloads across locations.
Key Responsibilities
- Operate and maintain hybrid platform environments, supporting both on-premises and Azure cloud-based analytical platforms.
- Manage and optimize Kubernetes clusters, ensuring high availability, performance, and security.
- Design, build, and maintain data and code pipelines, enabling efficient automation and workload orchestration across the platform.
- Collaborate with data engineering, DevSecOps, and cloud architecture teams to integrate and standardize platform operations.
- Contribute to infrastructure-as-code and automation initiatives to improve deployment repeatability and environment consistency.
- Implement observability practices — including monitoring, logging, and alerting — to support platform reliability and performance management.
- Participate in incident response and continuous improvement efforts to drive platform stability and resilience.
Required Qualifications
- 7+ years of experience as an SRE, Platform Engineer, or DevOps Engineer supporting hybrid or cloud-native platforms.
- Strong experience operating and maintaining Kubernetes environments (deployment, scaling, monitoring, troubleshooting).
- Extensive experience with OpenShift to build, modernize, and deploy applications at scale across various cloud environments and on-premises.
- Hands‑on experience creating and managing CI/CD pipelines and data pipelines (e.g., Jenkins, GitHub Actions, Azure DevOps, or similar tools).
- Familiarity with Azure services and cloud-native operations (VMs, containers, networking, security, and monitoring).
- Solid understanding of Linux system administration, automation, and containerization.
- Experience with infrastructure-as-code tools (Terraform, Helm, or ARM/Bicep templates).
- Excellent collaboration skills with cross‑functional teams in complex enterprise environments.
Preferred Skills
- Experience supporting data analytics or financial systems.
- Knowledge of observability and SRE best practices (SLIs, SLOs, SLAs).
- Familiarity with Databricks, Azure Data Factory, or related analytical workloads.
- Background working within regulated or financial industry environments.
Seniority level
Mid‑Senior level
Employment type
Full‑time
Job function
Information Technology