Senior Core Infrastructure Engineer

Oracle

Región Centro

Presencial

MXN 600.000 - 900.000

Jornada completa

Hace 6 días
Sé de los primeros/as/es en solicitar esta vacante

Recibe más respuestas de empleadores

Envía un currículum específico para el puesto de trabajo en cuestión de minutos.

Descripción de la vacante

Oracle in Mexico is seeking a highly capable Developer to join the AI2 Ops team, focusing on GPU infrastructure automation across multiple OCI regions and building robust tooling for provisioning, monitoring, and reliability.

You will collaborate with engineering, product, and operations to improve GPU operations, runbooks, and incident response while ensuring high availability and efficient capacity utilization.

Formación

  • Bachelor's degree in Computer Science, Engineering, or related field, or equivalent practical experience.
  • 3+ years of software development or infrastructure automation with Python and Bash.
  • Strong Linux experience including troubleshooting and system administration.
  • Hands-on experience with infrastructure-as-code and Terraform and RESTful APIs.
  • Strong problem-solving and troubleshooting skills.
  • Excellent communication and teamwork.

Responsabilidades

  • Design, build, and maintain software, automation, and operational tooling for OCI GPU infrastructure across multiple geographic regions.
  • Automate GPU infrastructure provisioning, configuration, validation, and deployment using Python, Bash, Terraform, and related tooling.
  • Collaborate with software engineers, hardware teams, and operations partners to build scalable, reliable, and highly available GPU platform services.
  • Build and improve monitoring, alerting, and diagnostics for GPU fleet health, performance, capacity, and utilization using Grafana.
  • Participate in incident response and root-cause analysis to remove blockers affecting GPU capacity, availability, and regional deployments.
  • Continuously improve AI2 Ops processes, GPU fleet automation, and OCI region build readiness.
  • Participate in on-call rotations and provide support for critical infrastructure issues.
  • Document operational procedures, automation workflows, troubleshooting guides, and runbooks.

Conocimientos

Python
Bash
Linux
REST APIs
Problem solving
Communication

Educación

Bachelor's degree in Computer Science

Herramientas

Terraform
Docker
Kubernetes
Grafana
CI/CD tooling

Descripción del empleo

Job Description

We are seeking a highly skilled and motivated Developer to join the AI2 Ops team supporting GPU infrastructure in Oracle Cloud Infrastructure (OCI). In this role, you will design, develop, deploy, and maintain automation and operational tooling for GPU fleets across multiple regions, ensuring high availability, scalability, performance, and efficient capacity utilization. You will collaborate closely with engineering, product, and operations teams to build robust automation, observability, and reliability solutions while continuously improving GPU operations.

We are seeking a highly skilled and motivated Developer to join the AI2 Ops team supporting GPU infrastructure in Oracle Cloud Infrastructure (OCI). In this role, you will design, develop, deploy, and maintain automation and operational tooling for GPU fleets across multiple regions, ensuring high availability, scalability, performance, and efficient capacity utilization. You will collaborate closely with engineering, product, and operations teams to build robust automation, observability, and reliability solutions while continuously improving GPU operations.

Responsibilities
  • Design, build, and maintain software, automation, and operational tooling for OCI GPU infrastructure across multiple geographic regions.
  • Automate GPU infrastructure provisioning, configuration, validation, and deployment using Python, Bash, Terraform, and related tooling.
  • Collaborate with software engineers, hardware teams, and operations partners to build scalable, reliable, and highly available GPU platform services.
  • Build and improve monitoring, alerting, and diagnostics for GPU fleet health, performance, capacity, and utilization using tools such as Grafana.
  • Participate in incident response and root-cause analysis to remove blockers affecting GPU capacity, availability, and regional deployments.
  • Continuously improve AI2 Ops processes, GPU fleet automation, and OCI region build readiness.
  • Participate in on-call rotations and provide support for critical infrastructure issues.
  • Document operational procedures, automation workflows, troubleshooting guides, and runbooks.
Required Qualifications
  • Bachelor's degree in Computer Science, Engineering, or a related field, or equivalent practical experience.
  • 3+ years of software development or infrastructure automation experience with strong proficiency in Python and Bash.
  • Strong Linux systems experience, including troubleshooting, scripting, process management, networking fundamentals, and system administration.
  • Hands-on experience with infrastructure-as-code and automation tools, especially Terraform, and with RESTful APIs.
  • Strong problem-solving and troubleshooting skills.
  • Excellent communication and teamwork skills.
Preferred Skills
  • Experience participating in or leading on-call operations and incident response.
  • Familiarity with DevOps practices and continuous integration/continuous deployment (CI/CD).
  • Experience operating or automating GPU, compute, or other large-scale cloud infrastructure.
  • Exposure to containerization and orchestration technologies such as Docker and Kubernetes.
  • Experience with observability tooling, including metrics, logging, dashboards, and alerting.
  • Experience with CI/CD and Agile methodologies, especially Scrum.
Qualifications

Career Level - IC3

About Us

Only Oracle brings together the data, infrastructure, applications, and expertise to power everything from industry innovations to life‑saving care. And with AI embedded across our products and services, we help customers turn that promise into a better future for all. Discover your potential at a company leading the way in AI and cloud solutions that impact billions of lives.

True innovation starts when everyone is empowered to contribute. That’s why we’re committed to growing a workforce that promotes opportunities for all with competitive benefits that support our people with flexible medical, life insurance, and retirement options. We also encourage employees to give back to their communities through our volunteer programs.

We’re committed to including people with disabilities at all stages of the employment process. If you require accessibility assistance or accommodation for a disability at any point, let us know by emailing accommodation-request_mb@oracle.com or by calling 1-888-404-2494 in the United States.

Oracle is an Equal Employment Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, sexual orientation, gender identity, disability and protected veterans’ status, or any other characteristic protected by law. Oracle will consider for employment qualified applicants with arrest and conviction records pursuant to applicable law.

Consigue la evaluación confidencial y gratuita de tu currículum.
o arrastra y suelta tu archivo aquí
Similar jobs

Puestos de trabajo similares que vale la pena comparar

Senior GPU Infrastructure Automation Engineer
Senior GPU Infrastructure Automation Engineer

Oracle • Región Centro

Presencial
MXN 600.000 - 900.000
Principal Site Reliability Engineer
Principal Site Reliability Engineer

Oracle • Región Centro

Presencial
MXN 600.000 - 900.000
Senior Product Engineer – GPU Server Platforms (Ciudad Juarez, México)
Senior Product Engineer – GPU Server Platforms (Ciudad Juarez, México)

Oracle • Región Centro

Presencial
MXN 1.746.000 - 2.271.000
Senior Site Reliability Engineer (Networking)
Senior Site Reliability Engineer (Networking)

Oracle • Región Centro

Presencial
MXN 900.000 - 1.300.000
Senior Data Center Support Services Technician - Location Queretaro
Senior Data Center Support Services Technician - Location Queretaro

Oracle • Santiago de Querétaro

Presencial
MXN 350.000 - 520.000
Senior Application Software Engineer
Senior Application Software Engineer

Oracle • Región Centro

Presencial
MXN 400.000 - 800.000
Solution Architect - Monterrey
Solution Architect - Monterrey

Oracle • Monterrey

Presencial
MXN 900.000 - 1.500.000
Oracle Database Administrator
Oracle Database Administrator

Genpact • Ciudad de México

Presencial
MXN 1.117.734 - 2.235.469
Cloud Business Innovation Advisor
Cloud Business Innovation Advisor

Oracle • Ciudad Valles

Presencial
MXN 420.000 - 540.000
Oracle EBS Supply Chain Management Techno-Functional Analyst
Oracle EBS Supply Chain Management Techno-Functional Analyst

Ll Oefentherapie • Ciudad de México

Presencial
MXN 300.000 - 700.000
Seguro médico
Opciones de jubilación
Programas de voluntariado
+1