Platform Engineer (Cloud Services)

Nexcess

United States

Remote

USD 145,000 - 195,000

Full time

6 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Visa support
Relocation assistance
Daily lunches in the office
International team
Medical insurance
Equipment provided
English language practice

Job summary

Nexcess is seeking an experienced Cloud Engineer to own features end to end within the Cloud Services group. You will specify, implement, test, deploy, and monitor production-ready enhancements, ensuring reliability and security.

You will collaborate with cross-functional teams, stay close to production, and apply AI-assisted tools while maintaining rigorous validation and documentation. This role emphasizes ownership, mentorship, and clear communication in a distributed team.

Qualifications

  • 5+ years of experience working with cloud computing infrastructure and production environments.
  • 5+ years of experience with Linux server administration.
  • 5+ years of experience with configuration management and automation tools such as Ansible or Puppet.
  • Strong understanding of infrastructure reliability, availability, security, and production operations.
  • Experience building, maintaining, and supporting production systems, including ownership after deployment.
  • Experience participating in an on-call rotation and responding to production incidents.
  • Ability to work effectively in unfamiliar systems and troubleshoot across different parts of a production environment.
  • Hands-on experience with AI-assisted engineering tools as part of day-to-day technical work.
  • Strong testing and validation habits, with a focus on proving changes work correctly before and after deployment.
  • Clear written communication skills and the ability to work effectively in a remote, distributed team across multiple time zones.

Responsibilities

  • Own platform features end to end within the Cloud Services domain. Take work from specification through implementation, testing, deployment, monitoring, and ongoing production support.
  • Translate business and technical requirements into clear specifications. Identify unknowns and dependencies early, clarify ambiguity before development starts, and make sure the solution is precise enough to build and operate reliably.
  • Build and maintain cloud infrastructure and automation. Work with Linux-based environments, configuration management tooling, and infrastructure components that support secure and highly available services.
  • Use AI-assisted engineering tools as part of your daily workflow. Apply agentic tooling for implementation and review, while remaining accountable for the quality, security, and operability of the final result.
  • Verify changes with evidence. Write and maintain tests, validate failure scenarios as well as the happy path, and provide clear evidence that changes are ready for production.
  • Review code and technical changes from other engineers. Identify design or operational risks early, provide clear and actionable feedback, and help maintain a high standard of engineering quality.
  • Mentor less experienced engineers. Share technical knowledge, explain the reasoning behind solutions, and support others through design, implementation, and review.
  • Own the services you run. Participate in the on-call rotation, respond to incidents, document findings and follow-up actions, and reduce operational toil through automation and improved tooling.
  • Work safely in existing and unfamiliar systems. Understand and improve legacy or complex infrastructure without unnecessary rewrites, and leave systems better documented and instrumented than you found them.
  • Coordinate dependencies directly with other teams. Identify cross-team needs early and work with the relevant engineers to resolve them before they become delivery blockers.

Skills

Cloud infrastructure
Linux administration
On-call incident management
Automation and configuration
AI-assisted engineering tools
Testing and validation
Remote collaboration
Security and reliability
Production systems ownership

Tools

Ansible
Puppet
CI/CD tooling

Job description

Cloud Services delivers the data protection, security, and edge services - SSL, email, EADN, and CDN/WAF - that keep our customers’ applications safe and performant.

You will own features end to end: the specification, the code, the tests and how it behaves in production. This is an individual contributor role with no direct reports, and you are the engineer who answers for whether what you shipped actually works.

The pull request is the visible part. The job is the machine behind it: turning a loosely worded need into something precise enough to build against, naming what can break before it does, proving the change is correct rather than asserting it, and staying with it once it is live.

How we work

Business alignment and focus come first. Engineering here exists to move the business, not to be busy. You will know which few things matter most this month and why, say so when the work in front of you is not one of them, and land what you committed to on dates the business can plan around, at a quality that does not come back as an incident or a customer escalation. Predictable, reliable delivery of quality software is the output of this job , and everything below is in service of it.

This job gets more hands-on every quarter, not less. The volume of code goes up while the time you spend typing it goes down, so the value moves to the two ends of it: what you specify, and what you verify. You will stay close to production, run the AI harness daily, read code you did not write, and keep your technical judgment current rather than remembered.

Whoever ships it owns it, regardless of what wrote it. Most code is AI-generated now, so the judgment moved to either side of it: specifying exactly what to build, and verifying that what came back is correct, secure and operable. An agent wrote it is not a defense. Your name is on the merge.

You do not need to own the roadmap to know why the work matters. You will not be the Product counterpart at this level and you are not expected to be. You are expected to understand what the customer gets and what the business gains from the thing you are building, and to ask before you build the wrong thing correctly. The engineers who move fastest here are the ones who ask that question early rather than at the review.

You go on call for what you build, and you take your turn on the rotation. Engineering declares risk and timelines in writing for the business to accept, and that includes yours: when a date is at risk you say so early, with evidence, rather than quietly.

Responsibilities
  • Own platform features end to end within the Cloud Services domain. Take work from specification through implementation, testing, deployment, monitoring, and ongoing production support.
  • Translate business and technical requirements into clear specifications. Identify unknowns and dependencies early, clarify ambiguity before development starts, and make sure the solution is precise enough to build and operate reliably.
  • Build and maintain cloud infrastructure and automation. Work with Linux-based environments, configuration management tooling, and infrastructure components that support secure and highly available services.
  • Use AI-assisted engineering tools as part of your daily workflow. Apply agentic tooling for implementation and review, while remaining accountable for the quality, security, and operability of the final result.
  • Verify changes with evidence. Write and maintain tests, validate failure scenarios as well as the happy path, and provide clear evidence that changes are ready for production.
  • Review code and technical changes from other engineers. Identify design or operational risks early, provide clear and actionable feedback, and help maintain a high standard of engineering quality.
  • Mentor less experienced engineers. Share technical knowledge, explain the reasoning behind solutions, and support others through design, implementation, and review.
  • Own the services you run. Participate in the on-call rotation, respond to incidents, document findings and follow-up actions, and reduce operational toil through automation and improved tooling.
  • Work safely in existing and unfamiliar systems. Understand and improve legacy or complex infrastructure without unnecessary rewrites, and leave systems better documented and instrumented than you found them.
  • Coordinate dependencies directly with other teams. Identify cross-team needs early and work with the relevant engineers to resolve them before they become delivery blockers.
Requirements
  • 5+ years of experience working with cloud computing infrastructure and production environments.
  • 5+ years of experience with Linux server administration.
  • 5+ years of experience with configuration management and automation tools such as Ansible or Puppet.
  • Strong understanding of infrastructure reliability, availability, security, and production operations.
  • Experience building, maintaining, and supporting production systems, including ownership after deployment.
  • Experience participating in an on-call rotation and responding to production incidents.
  • Ability to work effectively in unfamiliar systems and troubleshoot across different parts of a production environment.
  • Hands-on experience with AI-assisted engineering tools as part of day-to-day technical work.
  • Strong testing and validation habits, with a focus on proving changes work correctly before and after deployment.
  • Clear written communication skills and the ability to work effectively in a remote, distributed team across multiple time zones.
Preferred
  • Experience with CDN, WAF, and server security.
  • Experience in managed hosting, cloud platforms, or other high-availability service environments.
  • Experience with infrastructure automation at scale.
  • Experience reviewing infrastructure changes and mentoring other engineers.
We offer
  • Visa support and relocation assistance.
  • Daily lunches in the office.
  • Work in a team of professionals at an international level.
  • Comfortable working conditions.
  • Opportunity to practice and improve your English.
  • Competitive salary and all necessary equipment.
  • Medical insurance.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Platform Engineer (Cloud / AI Adoption)
Senior Platform Engineer (Cloud / AI Adoption)

TD SYNNEX • United States

On-site
USD 140,000 - 190,000
Remote or Hybrid work
Health insurance
Retirement plans
Cloud Engineer
Cloud Engineer

Integrated Resources, Inc ( IRI ) • Indianapolis (IN)

On-site
USD 120,000 - 170,000
Senior Platform Engineer
Senior Platform Engineer

Epsilon ASI • Denver (CO), Northern (KY)

On-site
USD 130,000 - 180,000
Cloud and Automation Engineer
Cloud and Automation Engineer

Talution Group • Riverwoods (IL)

Hybrid
USD 120,000 - 150,000
Principal Platform Engineer -Infrastructure Automation & Agentic Engineering
Principal Platform Engineer -Infrastructure Automation & Agentic Engineering

Adusa • Salisbury (NC), Northern (KY)

On-site
USD 163,000 - 245,000
Cloud Engineer F/H
Cloud Engineer F/H

PowerToFly • Ridgewood (NJ)

On-site
USD 120,000 - 160,000
Cloud / Platform Engineer (Kubernetes, GCP, AWS)
Cloud / Platform Engineer (Kubernetes, GCP, AWS)

Supero • San Francisco (CA), Northern (KY)

Hybrid
USD 150,000 - 210,000
AI Platform Engineer
AI Platform Engineer

Skillsearch Limited • United States

Remote
USD 140,000 - 190,000
Principal Software Engineer
Principal Software Engineer

American Bureau of Shipping • Spring (TX)

On-site
USD 140,000 - 210,000
Associate Infrastructure Engineer
Associate Infrastructure Engineer

Save A Lot • St. Ann (MO)

On-site
USD 75,000 - 110,000
401K company match up to 4%
Paid Time Off
Medical Insurance options including FS
+7