Cloud Engineer - Linux and Automation

Thrive

Sarasota (FL)

Hybrid

USD 110,000 - 150,000

Full time

12 days ago
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Remote work eligible

Job summary

Thrive is seeking a Cloud Engineer - Linux and Automation to administer and automate Thrive's Linux-based infrastructure, focusing on Ubuntu, Ansible, Terraform, and private cloud platforms (HPE Morpheus and VMware).

You will implement IaC, maintain NetBox, manage VM workloads, and collaborate across security, network, and cloud teams, with on-call rotation and a security-first mindset.

Qualifications

  • 3-5+ years of hands-on Linux system administration experience with strong proficiency in Ubuntu Linux including systems, networking, storage management, and security hardening
  • 2+ years of experience with Infrastructure as Code tools, specifically Ansible and Terraform
  • Hands-on familiarity with HPE Morpheus Enterprise or similar platforms for VM provisioning and self-service automation; VMware vCenter experience is a plus
  • Familiarity with NetBox for IPAM/DCIM management and NetBox API integrations preferred
  • Proficiency in scripting languages for automation and tooling development
  • Strong networking fundamentals including TCP/IP, VLANs, DNS, DHCP, routing, and firewall rules
  • Git-based version control and collaborative workflows (GitLab or GitHub)
  • Knowledge of ITIL/ITSM best practices and change management
  • Certifications: LPIC-1/LPIC-2, RHCSA, RH294, Terraform Associate preferred

Responsibilities

  • Administer, monitor, and troubleshoot Linux systems (primarily Ubuntu) across physical and virtual environments
  • Serve as a Tier 2 escalation point for infrastructure incidents; perform root cause analysis and remediation
  • Manage and operate VM workloads within HPE Morpheus Enterprise and VMware ESXi; provisioning and lifecycle management
  • Define and enforce standardized build documentation
  • Develop and maintain Ansible playbooks for configuration management, OS patching, and compliance
  • Write Terraform configurations for infrastructure provisioning and lifecycle management
  • Enforce infrastructure-as-code best practices including version control and peer review
  • Maintain NetBox as the authoritative source of truth for IPAM, DCIM asset records, and topology documentation
  • Integrate NetBox with IaC tools for automation of routine physical equipment commissioning
  • Participate in change management processes with Thrive ITSM standards
  • Perform routine system health checks, capacity reviews, and performance tuning for Linux servers
  • Maintain and improve internal runbooks and knowledge base articles
  • Collaborate with networking, security, and cloud teams to resolve cross-functional issues
  • Participate in on-call rotation for critical incidents
  • Identify opportunities to improve operational efficiency through automation

Skills

Linux administration
Scripting (bash/python)
Infrastructure as Code
Cloud automation
Networking fundamentals
Git version control
Troubleshooting under pressure
Automation mindset

Education

Bachelor's degree in Computer Science or related

Tools

Ansible
Terraform
HPE Morpheus Enterprise
VMware ESXi/vCenter
NetBox
Ubuntu Linux hardening
NetBox API integrations
Provisioning templates

Job description

Position Overview

The Cloud, Engineer - Linux and Automation is responsible for the administration, automation, and continuous improvement of Thrive's Linux-based infrastructure. This role serves as an escalation point for Tier 1 issues and works closely with senior engineers and architects to design and implement repeatable, code-driven infrastructure solutions. The engineer will maintain and develop automation frameworks using Ansible and Terraform, assist with managing Thrive's private cloud on HPE Morpheus Enterprise and VMware ecosystems, and ensure Ubuntu-based systems are secure, patched, and operating within established SLAs. The ideal candidate brings hands‑on IaC experience, strong Linux troubleshooting skills, and a mindset oriented toward automation and operational excellence - all with a security first focus required to avoid unexpected downtime.

Responsibilities
  • Administer, monitor, and troubleshoot Linux systems (primarily Ubuntu) across physical and virtual environments.
  • Serve as a Tier 2 escalation point for infrastructure incidents; perform root cause analysis and implement remediation actions.
  • Manage and operate virtual machine workloads within HPE Morpheus Enterprise HVM and VMware ESXi, including VM provisioning, lifecycle management, and hypervisor-level troubleshooting.
  • Define and enforce standardized build documentation
  • Develop and maintain Ansible playbooks for configuration management, OS patching, application deployment, and compliance enforcement.
  • Write and manage Terraform configurations for infrastructure provisioning and lifecycle management.
  • Enforce infrastructure-as-code best practices including version control and peer review.
  • Maintain and update NetBox as the authoritative source of truth for IP address management (IPAM), DCIM asset records, and network topology documentation.
  • Integrate Netbox with IaC tools for automation of routine physical equipment commissioning
  • Participate in change management processes, ensuring all changes are documented, reviewed, and approved in alignment with Thrive's ITSM standards.
  • Perform routine system health checks, capacity reviews, and performance tuning for Linux servers.
  • Maintain and improve internal runbooks, technical documentation, and knowledge base articles.
  • Collaborate with networking, security, and cloud teams to resolve cross-functional infrastructure issues.
  • Participate in an on-call rotation to support critical infrastructure incidents outside of standard business hours.
  • Identify opportunities to improve operational efficiency through automation and standardization.
Requirements
  • 3-5+ years of hands‑on Linux system administration experience with strong proficiency in Ubuntu Linux including systems, networking, storage management, and security hardening
  • 2+ years of experience with Infrastructure as Code tools, specifically Ansible (roles, playbooks, inventories) and Terraform (state management, CI/CD plan/apply workflows)
  • Hands‑on familiarity with HPE Morpheus Enterprise or similar HVM/KVM platforms for VM provisioning, template management, and self‑service automation; experience with VMware vCenter is a strong plus
  • Working knowledge of NetBox for IPAM, DCIM, and network source‑of‑truth management; experience with NetBox API integrations preferred
  • Proficiency in scripting languages for automation and tooling development
  • Solid understanding of networking fundamentals including TCP/IP, VLANs, DNS, DHCP, routing, and firewall rules as they apply to virtualized and cloud environments
  • Familiarity with Git‑based version control and collaborative workflows (GitLab or GitHub)
  • Ability to diagnose and resolve complex system and infrastructure issues independently and under pressure
  • Strong documentation habits - able to produce accurate and maintainable runbooks, topology diagrams, and change records
  • Ability to work in a fast‑paced environment with a diverse workload
  • Strong team player - collaborates effectively with peers and cross‑functional teams to solve problems
  • Proactive, change‑oriented mindset - actively seeks process improvements and drives automation‑first approaches
  • Ability to communicate technical concepts clearly to both technical and non‑technical stakeholders
  • Bachelor's degree in Computer Science, or a related discipline — or equivalent combination of education and relevant work experience
  • Knowledge of ITIL and ITSM best practices
  • Preferred Certifications: Linux Professional Institute Certification (LPIC-1/LPIC-2) or Red Hat Certified System Administrator (RHCSA); Red Hat Enterprise Linux Automation with Ansible (RH294) or better; HashiCorp Terraform Associate (003 or 004);
Other
  • Work Schedule: Standard business hours with participation in an on-call rotation.
  • Remote work eligible
  • Travel Requirements: Occasional travel may be required for datacenter activities. Less than 10%
  • Applicant selected will be subject to a criminal and credit background investigation and must meet eligibility requirements for access to restricted information. Candidate must be able to pass CJIS clearance for the state of Florida.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Systems Administrator
Systems Administrator

Thrive • Town of Mansfield (NY)

Remote
USD 70,000 - 110,000
Systems Engineer
Systems Engineer

Thrivenextgen • Birmingham (AL), Northern (KY)

Hybrid
USD 52,000 - 68,000
Systems Engineer
Systems Engineer

Thrive • United States

Remote
USD 42,000 - 64,000
Sr. Linux Engineer
Sr. Linux Engineer

Skyline Technology Solutions • United States

On-site
USD 120,000 - 150,000
Sr. Systems Engineer - Infrastructure & Cloud
Sr. Systems Engineer - Infrastructure & Cloud

Toshiba International Corporation • Houston (TX), Northern (KY)

Hybrid
USD 110,000 - 150,000
Cloud Engineer F/H
Cloud Engineer F/H

PowerToFly • Ridgewood (NJ)

On-site
USD 120,000 - 160,000
Senior Cloud Engineer
Senior Cloud Engineer

Compunnel, Inc. • McLean (VA)

On-site
USD 120,000 - 150,000
Linux Systems Administrator
Linux Systems Administrator

TALENT Software Services • Austin (TX)

On-site
USD 120,000 - 150,000
Sr. Systems Engineer - Infrastructure & Cloud
Sr. Systems Engineer - Infrastructure & Cloud

Toshiba • Houston (TX)

On-site
USD 110,000 - 160,000
Client Support Engineer 1
Client Support Engineer 1

Thrive • Fairfax (VA)

On-site
USD 40,000 - 42,000