Data Technician

Trilyon, Inc.

Austin (TX)

On-site

USD 110,000 - 170,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Trilyon, Inc. in Austin, TX seeks a platform engineer responsible for deploying and stabilizing data center CPU/GPU systems, leading cross-functional debugging across hardware, firmware, Linux, networking, power, and thermal domains.

You will design debug strategies, validate flows, and failure analyses, install and configure OS distros, set up network gear, manage file systems, and integrate automated testing in CI/CD (Jenkins, Ansible).

Responsibilities

  • Manage platform deployment, availability, and stability for data center CPU/GPU systems
  • Lead system-level debugging across hardware, firmware, Linux, networking, power, and thermal behavior
  • Develop structured debug strategies, validation flows, and failure-analysis methodologies.
  • Installation and configuration of OS distros from console; setup network gear and configurations.
  • Configure file systems using VG/LV/Partitions
  • Integrate automated testing in CI/CD environments (e.g. Jenkins, Ansible)
  • Provide logs and statistics to aid in debugging
  • Own inventory database management and administration through a managed system
  • Participate in Agile planning, delivery, and collaboration with scaled agile teams
  • Work with a managed ticketing system and communicate clearly on activities
  • Track daily activities, prioritize issues, assign work, and monitor progress to resolution
  • Collaborate with silicon, firmware, validation, networking, and operations teams to assess risks
  • Execute hands-on laboratory validation to ensure systems operate as intended prior to and during deployment
  • Partner with OEMs, ODMs, and vendors to resolve issues and improve platform reliability
  • Drive continuous improvement in tools, processes, and platform readiness

Job description

  • Manage platform deployment, availability, and stability for data center CPU/GPU systems
  • Lead system-level debugging efforts involving hardware, firmware, Linux, networking, power, and thermal behavior
  • Develop structured debug strategies, validation flows, and failure-analysis methodologies.
  • Installation, configuration of various OS Distros from console. Setup Network gear, perform and validate Network configurations.
  • Configure file systems using VG/LV/Partitions
  • Integrate automated testing in CI/CD environment (e.g. Jenkins, ansible)
  • Provide logs and statistics that will help in further debug of issues.
  • Own inventory database management and administration through a managed system.
  • Participate in the Agile method of planning, delivery, and collaboration with internal and scaled agile teams.
  • Work with a managed ticketing system and communicate clearly on activities and steps.
  • Track daily activities, prioritize issues, assign work, and monitor progress to resolution
  • Collaborate with silicon, firmware, validation, networking, and operations teams to assess risks and requirements
  • Execute hands-on laboratory validation to ensure systems operate as intended prior to and during deployment
  • Partner with OEMs, ODMs, and vendors to resolve issues and improve platform reliability
  • Drive continuous improvement in tools, processes, and platform readiness
Get your free, confidential resume review.
or drag and drop your file here.