Cloud Infrastructure Operations Engineer

Centific

United States

Remote

USD 59,000 - 80,000

Full time

11 days ago
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Centific in the United States is seeking an infrastructure operations engineer to support cloud delivery, OS deployments, and incident management. You will script automation in Shell and Python, troubleshoot Linux and network issues, and collaborate with cloud providers to keep services healthy.

This role involves on-call rotations, on-site work, and opportunities to optimize tooling and processes while contributing to scalable GenAI infrastructure initiatives.

Qualifications

  • Bachelor’s degree in CS, E&E, or related field.
  • Strong Linux troubleshooting and infrastructure experience.
  • Experience with automated provisioning and OS deployment technologies such as PXE, iPXE.
  • Familiarity with public cloud platforms (AWS, Oracle, Google, Azure).
  • Ability to write Shell/Bash and Python scripts.
  • Familiarity with Ansible, Terraform, CI/CD tooling, and cloud SDKs.

Responsibilities

  • Collaborate with cloud service team to support infrastructure delivery and deployment activities.
  • Perform OS reinstallation and reboot for cloud servers and bare metal environments.
  • Troubleshoot infrastructure issues involving Linux systems, cloud servers, networking, and hardware-related abnormalities.
  • Maintain operational data across internal systems including fault ticketing systems and infrastructure tracking tools.
  • Work with cloud providers to resolve infrastructure incidents and large-scale operational issues.
  • Participate in on-call rotations and support incident management activities.
  • Monitor cloud infrastructure health, operational alerts, and asset utilization.
  • Develop or enhance operational tools, scripts, and automation workflows using Shell or Python.
  • Support operational process optimization and standardization initiatives.
  • Perform server and cloud network troubleshooting (TCP/IP, VLAN, DNS, IPv6).
  • Assist with documentation creation, knowledge sharing, and operational reporting.

Skills

Linux troubleshooting
Shell scripting
Python scripting
Cloud networking
Documentation

Education

Bachelor’s Degree in Computer Science, Electrical Engineering, or related fields

Tools

PXE/iPXE provisioning
Ansible
GitLab CI/CD
Terraform
Cloud SDKs

Job description

About Centific

Centific is a frontier AI data foundry that curates diverse, high-quality data, using our purpose-built technology platforms to empower the Magnificent Seven and our enterprise clients with safe, scalable AI deployment. Our team includes more than 150 PhDs and data scientists, along with more than 4,000 AI practitioners and engineers. We harness the power of an integrated solution ecosystem—comprising industry-leading partnerships and 1.8 million vertical domain experts in more than 230 markets—to create contextual, multilingual, pre-trained datasets; fine-tuned, industry-specific LLMs; and RAG pipelines supported by vector databases. Our zero-distance innovation™ solutions for GenAI can reduce GenAI costs by up to 80% and bring solutions to market 50% faster. Our mission is to bridge the gap between AI creators and industry leaders by bringing best practices in GenAI to unicorn innovators and enterprise customers. We aim to help these organizations unlock significant business value by deploying GenAI at scale, helping to ensure they stay at the forefront of technological advancement and maintain a competitive edge in their respective markets.

About Job
  • Collaborate with Cloud service team, internal teams, and cloud vendors to support infrastructure delivery and deployment activities.
  • Perform operating system reinstallation and reboot for cloud servers and bare metal environments.
  • Troubleshoot infrastructure issues involving Linux systems, cloud servers, networking, and hardware-related abnormalities.
  • Maintain operational data across internal systems including fault ticketing systems, repair records, and infrastructure tracking tools.
  • Work with cloud providers to resolve infrastructure incidents and large-scale operational issues.
  • Participate in on-call rotations and support incident management activities.
  • Monitor cloud infrastructure health, operational alerts, and asset utilization.
  • Develop or enhance operational tools, scripts, and automation workflows using Shell or Python.
  • Support operational process optimization and standardization initiatives.
  • Perform server and cloud network troubleshooting, including TCP/IP, VLAN, DNS, and Ipv6 related issues.
  • Support cloud fleet operations and infrastructure performance management.
  • Assist with documentation creation, knowledge sharing, and operational reporting.
Requirements
  • Bachelor’s Degree in Computer Science, Electrical Engineering, or related fields.
  • Experience in cloud infrastructure operations, server operations, or data center related environments.
  • Strong troubleshooting and analytical skills in Linux and infrastructure environments.
  • Experience with automated provisioning and OS deployment technologies such as PXE, iPXE, or provisioning pipelines.
  • Familiarity with public cloud platforms such as Oracle, Amazon Web Services, Google, or Microsoft.
  • Ability to understand, execute, and write Shell/Bash or Python scripts.
  • Familiarity with automation and infrastructure management tools such as Ansible, GitLab CI/CD, Terraform, or cloud SDKs.
  • Knowledge of Linux systems, cloud networking, and server lifecycle management.
  • Strong understanding of TCP/IP networking concepts including subnetting, VLANs, DNS, IPv6, and basic routing.
  • Familiarity with infrastructure monitoring, operational tooling, and incident management processes.
  • Strong communication, collaboration, and documentation skills.
  • Familiarity with large-scale cloud fleet operations is preferred.
  • Experience supporting GPU infrastructure, firmware lifecycle management, or RDMA networking is a plus.
Hourly Rate

$50

Centific is an equal-opportunity employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, national origin, ancestry, citizenship status, age, mental or physical disability, medical condition, sex (including pregnancy), gender identity or expression, sexual orientation, marital status, familial status, veteran status, or any other characteristic protected by applicable law. We consider qualified applicants regardless of criminal histories, consistent with legal requirements.

Join a growing company using technology to help tackle enterprises’ toughest challenges.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Cloud Infrastructure Operations Engineer
Cloud Infrastructure Operations Engineer

Centific • San Jose (CA)

On-site
USD 120,000 - 170,000
Healthcare, dental, and vision
401k plan
Paid time off
+1
Operation Lead
Operation Lead

Centific Global Solutions, Inc. • United States

On-site
USD 100,000
Cloud Infrastructure Ops Engineer — Scale & Automation
Cloud Infrastructure Ops Engineer — Scale & Automation

Centific • United States

Remote
USD 59,000 - 80,000
Cloud Infrastructure Engineer
Cloud Infrastructure Engineer

Ampcus Inc • City of White Plains (NY)

On-site
USD 120,000 - 170,000
Cloud Principal Engineer
Cloud Principal Engineer

Centric Software, Inc. • South Dakota

On-site
USD 130,000 - 170,000
Cloud Engineer
Cloud Engineer

MegazoneCloud Global • New York (NY)

On-site
USD 100,000 - 130,000
Cloud Engineer- Infrastructure
Cloud Engineer- Infrastructure

MegazoneCloud US • Frisco (TX)

On-site
USD 110,000 - 140,000
Systems Administrator, 37602682
Systems Administrator, 37602682

Cypress HCM • Santa Clara (CA)

On-site
USD 90,000 - 103,000
Cloud Engineer- Infrastructure
Cloud Engineer- Infrastructure

MegazoneCloud US • Irvine (CA)

On-site
USD 110,000 - 140,000
Cloud Engineer- Infrastructure
Cloud Engineer- Infrastructure

MegazoneCloud US • City of Rochester (NY)

On-site
USD 110,000 - 140,000
Investment in employee growth
Servant leadership management style
Impact on technical roadmap
+1