DevOps Engineer

Compunnel, Inc.

Round Rock, Northern (TX, KY)

Hybrid

USD 140,000 - 180,000

Full time

24 hours ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Compunnel, Inc. is seeking a Senior Platform Engineer / Site Reliability Engineer (SRE) to design, implement, and optimize enterprise messaging platforms and automation frameworks. You will support RabbitMQ, Kafka, and observability solutions to improve reliability and developer productivity.

The role emphasizes platform modernization, DevSecOps, and scalable infrastructure engineering across enterprise environments, with strong scripting in Python/Ansible and collaboration with multiple teams.

Qualifications

  • Bachelor's degree in Computer Science, IT, Engineering, or related field.
  • Minimum 5 years of SRE/infrastructure automation experience.
  • Strong scripting and automation using Python and Ansible.
  • Experience with observability platforms (Grafana, Splunk, Dynatrace or similar).
  • Experience building CI/CD pipelines and automation frameworks.
  • Excellent problem-solving, communication, and collaboration skills.

Responsibilities

  • Design, implement, maintain, and optimize enterprise messaging platforms using RabbitMQ and Apache Kafka.
  • Support highly available, fault-tolerant, and scalable messaging architectures.
  • Troubleshoot incidents, bottlenecks, latency, and capacity issues.
  • Develop standards, best practices, and ops procedures for messaging services.
  • Partner with application teams to integrate messaging technologies enterprise-wide.
  • Create automation using Python, Ansible, and related scripting.
  • Build reusable automation frameworks and self-service engineering capabilities.
  • Automate provisioning, config management, deployment, monitoring, and operations.
  • Support platform modernization and operational transformation initiatives.
  • Improve developer productivity through platform enhancements and automation.

Skills

SRE fundamentals
Automation scripting
Observability

Education

Bachelor's degree in CS/IT/Engineering

Tools

Python
Ansible
Grafana
Splunk
Liquibase

Job description

The Senior Platform Engineer / Site Reliability Engineer (SRE) will be responsible for designing, implementing, and optimizing enterprise messaging platforms, automation frameworks, observability solutions, and DevOps capabilities that support mission-critical applications. This role will drive platform modernization, operational excellence, reliability engineering, and automation initiatives while improving developer productivity and reducing operational overhead. The ideal candidate will possess strong expertise in RabbitMQ, Kafka, Python or Ansible automation, monitoring platforms, and enterprise-scale infrastructure engineering.

Key Responsibilities
  • Design, implement, maintain, and optimize enterprise messaging platforms using RabbitMQ and Apache Kafka.
  • Support highly available, fault-tolerant, and scalable messaging architectures.
  • Troubleshoot messaging platform incidents, performance bottlenecks, latency issues, and capacity constraints.
  • Develop standards, best practices, and operational procedures for messaging services.
  • Partner with application teams to integrate messaging technologies into enterprise solutions.
  • Create and maintain automation solutions using Python, Ansible, and related scripting technologies.
  • Build reusable automation frameworks and self‑service engineering capabilities.
  • Automate provisioning, configuration management, deployment, monitoring, and operational processes.
  • Support platform modernization and operational transformation initiatives.
  • Improve developer productivity through platform enhancements and workflow automation.
  • Develop reusable templates, scripts, and engineering standards across enterprise environments.
Observability & Reliability Engineering
  • Design and implement enterprise observability and monitoring solutions.
  • Configure and support monitoring platforms such as:
    • Grafana
    • Splunk
    • Dynatrace
  • Establish monitoring standards, dashboards, alerting frameworks, and operational reporting.
  • Monitor application health, infrastructure performance, messaging platforms, and service reliability.
  • Conduct root cause analysis and implement long-term corrective actions.
  • Drive improvements in platform reliability, availability, and operational visibility.
DevOps & DevSecOps
  • Support CI/CD pipeline implementation, enhancement, and operational support.
  • Integrate security controls and compliance checks into software delivery workflows.
  • Implement and support DevSecOps tools including:
    • Qualys
    • Fortify
    • SonarQube
  • Partner with security teams to automate vulnerability detection, remediation, and compliance validation.
  • Support software supply chain security and secure development practices.
Identity & Access Governance
  • Support implementation of:
    • Role-Based Access Control (RBAC)
    • Single Sign-On (SSO)
    • Privileged Access Management (PAM)
    • Secrets Management Solutions
  • Collaborate with security and infrastructure teams to establish secure access governance models.
  • Support automation and lifecycle management for privileged credentials and access controls.
Database & Infrastructure Automation
  • Implement and maintain Liquibase-based deployment frameworks and automation processes.
  • Develop automated database deployment, validation, rollback, and governance workflows.
  • Support infrastructure standardization through Infrastructure-as-Code and automation practices.
Required Qualifications
  • Bachelor's degree in Computer Science, Information Technology, Engineering, or a related technical discipline.
  • Minimum 5 years of experience in:
    • Site Reliability Engineering (SRE)
    • Infrastructure Automation
    • Enterprise Technology Operations
  • Experience supporting enterprise-scale technology environments and mission-critical applications.
  • Strong hands‑on expertise with:
  • Strong scripting and automation experience using:
    • Python
    • Ansible
  • Experience implementing and supporting observability platforms including:
    • Grafana
    • Splunk
    • Similar Enterprise Monitoring Solutions
  • Experience building and supporting CI/CD pipelines and automation frameworks.
  • Strong understanding of:
    • DevOps Practices
    • Reliability Engineering
    • Distributed Systems
  • Experience collaborating with development, database, infrastructure, and security teams.
  • Proven success delivering modernization, automation, or operational transformation initiatives.
  • Strong troubleshooting, analytical, and root cause analysis skills.
  • Excellent verbal and written communication skills.
  • Strong ability to work independently and manage multiple initiatives simultaneously.
Preferred Qualifications
  • Experience with Dynatrace observability and monitoring solutions.
  • Experience integrating and administering:
    • Qualys
    • Fortify
    • SonarQube
    • RBAC
    • SSO
    • PAM
    • Secrets Management Solutions
  • Experience designing reusable automation frameworks and self‑service engineering platforms.
  • Experience working in Agile, Scrum, or Kanban environments.
  • Strong understanding of cloud-native architectures and platform modernization.
  • Experience supporting highly available distributed applications and services.
  • Experience improving developer experience and engineering productivity through platform innovations.
Certifications
  • AWS Certified DevOps Engineer – Professional (Preferred).
  • Microsoft Certified: DevOps Engineer Expert (Preferred).
  • Splunk Core Certified Power User or Administrator (Preferred).
  • Red Hat Certified Engineer (RHCE) (Preferred).
Technical Skills
  • Site Reliability Engineering (SRE)
  • Privileged Access Management (PAM)
  • Reliability Engineering
Mandatory Skills
  • Grafana, Prometheus, Splunk, or Similar Observability Platforms
  • Site Reliability Engineering (SRE)
  • Troubleshooting and Root Cause Analysis
Preferred Skills
  • Self‑Service Platform Engineering
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer – Lead
Site Reliability Engineer – Lead

Jobtailor • Arizona

On-site
USD 140,000 - 230,000
Senior Site Reliability Engineer – Digital Assets
Senior Site Reliability Engineer – Digital Assets

Jobtailor • Arizona

On-site
USD 120,000 - 170,000
Sr SRE Automation Engineer
Sr SRE Automation Engineer

Compunnel, Inc. • Austin (TX), Northern (KY)

Hybrid
USD 130,000 - 180,000
Senior Lead Site Reliability Engineer
Senior Lead Site Reliability Engineer

JPMorgan Chase & Co. • Jersey City (NJ)

On-site
USD 150,000 - 210,000
Senior Engineer – Reliability
Senior Engineer – Reliability

Jobtailor • South Carolina

On-site
USD 120,000 - 180,000
Principal Site Reliability Engineer
Principal Site Reliability Engineer

Jobtailor • Arizona

On-site
USD 180,000 - 240,000
IT CONSULTANT SR
IT CONSULTANT SR

First Horizon Corp. • Memphis (TN), Northern (KY)

On-site
USD 120,000 - 160,000
Site Reliability Engineer Lead (SRE) – Internal Kubernetes Container Platform (IKCP)
Site Reliability Engineer Lead (SRE) – Internal Kubernetes Container Platform (IKCP)

Bank of America • Jersey City (NJ)

On-site
USD 180,000 - 240,000
Site Reliability Engineer Lead (SRE) – Internal Kubernetes Container Platform (IKCP)
Site Reliability Engineer Lead (SRE) – Internal Kubernetes Container Platform (IKCP)

Bank of America • Plano (TX)

On-site
USD 140,000 - 190,000
DevOps Engineer
DevOps Engineer

Compunnel, Inc. • Florham Park (NJ)

On-site
USD 120,000 - 180,000