Senior Site Reliability Engineer

United States Digital Space LLC

Bogotá

Presencial

COP 244.056.000 - 418.382.000

Jornada completa

14 días+

Recibe más respuestas de empleadores

Envía un currículum específico para el puesto de trabajo en cuestión de minutos.

Descripción de la vacante

United States Digital Space LLC is seeking a senior Site Reliability Engineer to lead scalable infrastructure initiatives in a fast-paced environment. You will own architecture of distributed systems, drive reliability, and guide AI-enabled tooling integration to enhance productivity.

This role requires deep Golang expertise, strong backend experience, and hands-on leadership with modern cloud and CI/CD workflows. A robust, high‑impact technical scope awaits with opportunities for growth.

Formación

  • 12+ years of professional software or infrastructure engineering experience, including SRE and backend work.
  • Significant production changes in the last 30 days.
  • Strong proficiency in Golang with RESTful API experience.
  • Experience with SQL-based RDBMS (MySQL, PostgreSQL) and query optimization.
  • Experience with observability tools (Prometheus, Grafana, Datadog, New Relic).

Responsabilidades

  • Architect and build scalable infrastructure using Kubernetes, AWS, RDS, and modern distributed patterns.
  • Drive infrastructure roadmap to improve reliability and scalability.
  • Lead capacity planning, benchmarking, and stress testing activities.
  • Define and enforce SLAs and alerts across infrastructure.
  • Mentor engineers and promote learning and operational excellence.
  • Collaborate with teams to translate business goals into technical roadmaps.

Conocimientos

Golang
REST APIs
Distributed systems
Leadership

Educación

Bachelor's degree in Computer Science

Herramientas

Kubernetes
AWS
MySQL
PostgreSQL
RDS
GitLab CI/CD

Descripción del empleo

The salary range for this role is negotiable, the range being $7000 - $12000 per month (Gross in USD).

About the company:

With a mission to financially empower the next generation, the company is revolutionizing the shopping experience beyond payments, blending cutting‑edge tech with seamless, interest‑free installment plans that make shopping smarter and more accessible. We’re not just transforming payments; we’re redefining how people discover, interact with, and purchase the things they love while driving real impact on merchant sales through increased conversions and higher order values. As we continue to shape the future of fintech and retail, we’re building an innovative, dynamic team passionate about creating more than just a transaction but a truly unique shopping journey. If you’re excited about pushing boundaries in tech and delivering a game‑changing experience for consumers and merchants alike, come join us at the company and help create the future of shopping!

Compensation:

For this senior development role, with 12+ years of experience, the compensation range is $7000 - $12000 USD per month. This range acknowledges the extensive expertise, leadership capabilities, and significant contributions expected at this level, offering a competitive salary to reflect the value of advanced skills and experience.

About the Role:

We are seeking a talented and motivated best‑in‑class Senior Site Reliability Engineer. This role presents an exciting opportunity to thrive in a dynamic, fast‑paced environment within a rapidly growing team, with abundant prospects for career advancement.

As a Senior SRE with the company, you will have a high degree of autonomy and authority to identify and resolve problems you see in our infrastructure, deployments, operational workload, and overall systems.

You should consider yourself a DOer to be a good fit for this role. We expect you to bring a deep well of experience to play, and with AI tooling, you should be a force for scaling in the organization.

What You'll Do:
  • Architect, upgrade, design, and build scalable infrastructure solutions leveraging Kubernetes, AWS, RDS (MySQL/Postgres), and modern distributed patterns.
  • Help drive the infrastructure team’s roadmap, leading us to higher levels of reliability, recoverability, and scalability.
  • Drive capacity planning, benchmarking, and work with the team to stress test our systems, find bottlenecks, and prepare for further growth in the business.
  • Define, maintain and enforce SLAs and alerts across our infrastructure.
  • Lead the teams towards stronger signal anomaly detection, better, more flexible alerting.
  • Help Lead the company’s AI enablement efforts, identifying opportunities to apply AI and automation to enhance infrastructure reliability, developer productivity, and internal tooling.
  • Build in consistency and scalability across a distributed microservices architecture while maintaining performance and reliability.
  • Establish and evolve engineering best practices for observability, security, and CI/CD across teams.
  • Mentor engineers and champion a culture of learning, innovation, and operational excellence.
  • Collaborate cross‑functionally to translate business goals into technical roadmaps and deliver results that matter.
What We look for:
  • 12+ years of professional software engineering or infrastructure engineering experience, including significant SRE and backend experience.
  • Deployed significant changes to a production application or infrastructure configuration in the past 30 days.
  • Strong proficiency in Golang, with experience building and maintaining RESTful APIs.
  • Expertise with SQL-based RDBMS (MySQL, PostgreSQL) and experience optimizing schema and queries for performance at scale.
  • Proficiency in observability tools (Prometheus, Grafana, Datadog, New Relic).
  • Solid understanding of distributed systems design patterns (e.g., transactional outbox, event‑driven architecture and stream processing, queues).
  • Demonstrated ability to bring new ideas forward, influence decisions, and lead complex technical initiatives.
  • Bachelor’s degree in Computer Science or equivalent practical experience.
Preferred Knowledge and Skills:
  • Experience with AWS cloud infrastructure, mainly AWS Aurora RDS, both MySQL and Postgres.
  • Experience with data engineering, data pipelines and data warehousing.
  • Experience with CI/CD pipelines and deploying containerized microservices in Kubernetes.
  • Familiarity with AI developer tooling like Claude Code, Gemini CLI, Codex, Cursor and using it to be a more productive engineer.
  • Track record of shipping commercial APIs and data‑driven applications in high‑growth environments.
  • Proven leadership in guiding technical direction, improving system reliability, and scaling high‑traffic services.
About You:
  • You have relentlessly high standards - many people may think your standards are unreasonably high. You are continually raising the bar and driving those around you to deliver great results. You make sure that defects do not get sent down the line and that problems are fixed so they stay fixed.
  • You’re not bound by convention - your success—and much of the fun—lies in developing new ways to do things
  • You need action - speed matters in business. Many decisions and actions are reversible and do not need extensive study. We value calculated risk‑taking.
  • You earn trust - you listen attentively, speak candidly, and treat others respectfully.
  • You have backbone; disagree, then commit - you can respectfully challenge decisions when you disagree, even when doing so is uncomfortable or exhausting. You have conviction and are tenacious. You do not compromise for the sake of social cohesion. Once a decision is determined, you commit wholly.
  • You deliver results - you focus on the key inputs and deliver them with the right quality and in a timely fashion. Despite setbacks, you rise to the occasion and never settle.
the company’s Technology Stack:
  • Languages: Golang, Typescript, Python
  • Frontend: Typescript - React and React Native
  • Backend: Golang
  • Database: MySQL, Postgres, Elasticsearch
  • DevOps & Cloud: AWS, Kubernetes
  • Version Control: Git
  • CI/CD: Gitlab
  • Testing: Developer and AI‑driven, focus on automated end‑to‑end, integration, and unit tests
  • Open Source: the company is focused on using open source, and we build what we can before buying!
What Makes Working at the company Awesome?

At the company, we are more than just brilliant engineers, passionate data enthusiasts, out‑of‑the‑box thinkers, and determined innovators; we are skilled musicians, yogis, cyclists, chefs, golfers, dog‑lovers, and rock‑climbers. We believe in surrounding ourselves with not only the best and the brightest individuals, but those t

Consigue la evaluación confidencial y gratuita de tu currículum.
o arrastra y suelta tu archivo aquí
Similar jobs

Puestos de trabajo similares que vale la pena comparar

Senior Site Reliability Engineer
Senior Site Reliability Engineer

Sezzle • Bogotá

Presencial
Senior Site Reliability Engineer
Senior Site Reliability Engineer

LanceSoft, Inc. • Colombia

Presencial
COP 90.000.000 - 150.000.000
Site Reliability Engineer ID53670
Site Reliability Engineer ID53670

AgileEngine • Metropolitana

Híbrido
COP 142.369.000 - 213.554.000
Professional growth: Mentorship, TechTalks, and personalized growth roadmaps.
Competitive compensation: USD-based pay with education, fitness, and team activity budgets.
Exciting projects: Modern solutions with Fortune 500 and top product companies.
+1
Senior Site Reliability Engineer (SRE)
Senior Site Reliability Engineer (SRE)

Oowlish • Bogotá

Presencial
COP 180.000.000 - 300.000.000
Home office
Career plans to allow for extensive成长
International projects
+3
Senior Site Reliability Engineer
Senior Site Reliability Engineer

MPS Group LLC • Bogotá

Presencial
COP 200.880.000 - 312.480.000
Site Reliability Engineer
Site Reliability Engineer

DCT • Bogotá

A distancia
COP 156.225.000 - 234.339.000
Career Growth & Mentorship
Flexible Work Environment
Generative & Collaborative Culture
Sr. AI Engineer - Marketing
Sr. AI Engineer - Marketing

United States Digital Space LLC • Bogotá

Presencial
COP 156.206.000 - 374.895.000
Site Reliability Engineer ID62591
Site Reliability Engineer ID62591

AgileEngine • Metropolitana

Híbrido
COP 149.902.000 - 224.854.000
Professional growth: Mentorship, TechTalks, and personalized growth roadmaps.
Competitive compensation: USD-based pay with education, fitness, and team activity budgets.
Exciting projects: Modern solutions with Fortune 500 and top product companies.
+1
Senior DevOps Engineer ID56470
Senior DevOps Engineer ID56470

AgileEngine • Bogotá

Híbrido
COP 285.938.000 - 428.909.000
Professional growth
Competitive compensation
Exciting projects
+1
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

EPAM Systems • Colombia

Presencial
COP 60.000.000 - 120.000.000
Healthcare benefits
Global career opportunities
Upskilling and certification courses
+1