Site Reliability Engineer

Tyk

Colombia

Presencial

COP 90.000.000 - 120.000.000

Jornada completa

14 días+
Generador de candidaturas

No envíes un currículum genérico — crea un currículum y una carta de presentación adaptados a este puesto concreto.

Supera los filtros ATS

Ventajas ofrecidas por este puesto de trabajo

Unlimited paid holidays
Flexible hours
Employee share scheme
Generous parental leave
Company retreats

Descripción de la vacante

Tyk is seeking a Site Reliability Engineer to manage, maintain, and improve our platform. You will be the first line of incident response for clients and help shape our metrics, dashboards, and reliability practices. This is a remote-first role with a distributed team spanning multiple time zones.

You will work on expanding multi-region and multi-cloud capabilities, automating common tasks, and documenting operational knowledge while collaborating with cross-functional squads.

Formación

  • Strong collaboration skills.
  • Launching and operating production-scale Kubernetes clusters.
  • Designing and operating infrastructure on AWS and other providers.
  • Operating MongoDB clusters.
  • Operating Redis clusters.
  • Administering Linux servers.
  • Maintaining distributed software.
  • Operating Prometheus and Grafana.
  • Participating in on-call rotation (16:00–04:00 UTC).

Responsabilidades

  • Maintaining global Tyk Cloud within SLAs.
  • Identifying reliability issues and working with your squad to solve them.
  • Identifying and introducing new metrics and dashboards.
  • Participating in the on-call rotation.
  • Expanding multi-region and multi-cloud reach of the platform.
  • Documenting operational knowledge.
  • Post-incident analysis.
  • Automating common tasks.

Conocimientos

Kubernetes & containers
AWS / EKS
Linux
Terraform / IaC
Helm
Go
MongoDB
Redis
Prometheus / Grafana / Thanos
Networking concepts
DNS / TCP/IP / HTTP / TLS / UDP
Proactive / change oriented

Descripción del empleo

Who are Tyk, and what do we do?
Who are Tyk, and what do we do?

The Tyk API Management platform is helping to drive the connected world and power new products and services. We're changing the way that organisations connect any number of their systems and services.Whether internal, external, public or highly encrypted systems, Tyk helps businesses drive value across the retail, finance, telecoms, healthcare, or media industries (to name just a few!)

If you've banked online, used an app to check the news, or perhaps even driven a connected car, API's, and by extension, Tyk, make that possible. Founded in 2015 with offices in London - UK, London - Ontario, Atlanta and Singapore, we have many thousands of users of our B2B platform across the globe. Brands using Tyk range from Lotte, Bell, T Mobile, to RBS, Capital One and Vinci. We have a varied user base hailing from every continent - even Antarctica.

Our Mission

Tyk is on a mission to connect every system in the world. We've started by building an API Management platform.

Total flexibility, default remote, radical responsibility

We offer unlimited paid holidays and remote working from anywhere in the world, for everyone, Why? Tyk was founded on the principle of offering flexibility and autonomy to our employees, we believe this allows our employees to achieve their best results. It also means we can build the best possible team, location and working hours are no barrier.

The role:

We're looking for a Site Reliability Engineer to manage, maintain, improve and provide support on our platform. You will be curious by nature, always looking for ways to improve, as we will look to you for new ideas, solutions and metrics on how we can improve the platform. You will also be our first line of incident management to our clients and will help define our response going forward. This is a great opportunity to become an integral part of Tyk as we continue on our journey.

As a remote first company, you will have the opportunity to work with an industry leading distributed team. Having access to expertise from across the globe will give you both the support and opportunity to help shape not only Tyk's Cloud platform but also the Tyk as a whole as we continue to grow.

Requirements
Here’s what you'll be responsible for:
  • Maintaining global Tyk Cloud within SL(A/I/O)s you will help to define
  • Identifying reliability issues and working together with your squad to solve them
  • Identifying and introducing new metrics and building relevant dashboards
  • Participating in the on-call rotation
  • Working with your squad to expand multi-region and multi-cloud reach of the platform
  • Documenting operational knowledge
  • Conducting post-incident analysis
  • Automating common tasks
    Here’s what we’re looking for:
    Experience
    • Strong collaboration skills
    • Launching and operating production scale kubernetes clusters
    • Designing and operating infrastructure on AWS and other providers
    • Operating MongoDB (or other document database) clusters
    • Operating Redis (or other key-value storage) clusters
    • Administering Linux servers
    • Maintaining distributed software
    • Operating Prometheus and Grafana
    • Operating logging collection and analysis systems
    • Participating in the on-call rotation(16:00pm - 4:00am UTC)
    Skills:
    • Kubernetes & containers (advanced)
    • AWS / EKS (advanced)
    • Linux (advanced)
    • Terraform and IaC in general (proficient)
    • Helm (proficient)
    • Go (familiar)
    • MongoDB (or similar)
    • Redis (or similar)
    • Monitoring - prometheus, grafana, thanos (familiar)
    • Grasp of networking concepts (subnets, routing, peering, load balancing, NAT, etc.)
    • Common networking protocols (DNS, TCP/IP, HTTP, TLS, UDP)
    • Proactive, energetic, innovative and change oriented
    Nice to have:
    • GCP or Azure
    • Bare metal infrastructure engineering
    • API management experience
    • Large scale distributed storage management
    • Familiarity with Rancher
    • CKA/CKAD/CKS
    • Creating and delivering production software in Go language
    Benefits
    Here’s why you should join us:
    • Everyone has unlimited paid holiday.
    • We have total flexibility in hours, as we believe creativity flows better when our people are given freedom to decide when they are most productive. Everyone is unique after all
    • Employee share scheme
    • Generous maternity and paternity leave
    • Company retreats

    We all share the same vision - we value authenticity, respect, responsibility, independence, honesty, diversity and inclusion and most importantly treating others how you wish to be treated. We look for like-minded people who bring their personalities to work everyday, strive to achieve their personal goals and who are willing to challenge the way we do things, why? - to make what we do even better!

    Our values tell the story of Tyk - here's how:
    • It's ok to screw up!
    • The only stupid idea, is the untested one!
    • Trust starts with you - make it count!
    • Trust is a two-way street - instill it from day one!
    • Assume best intent!
    • We have each other's back - we're all on the same team. Think before you speak or act.
    • Make things, better!

    We've found that it's often the ‘stupid' or unexpected ideas that turn out to be the successful ones - so try it, at least we can say we have!

    Always try to leave things better than when you found them - change is constant, inevitable and embraced! Be that change we want to see.

    What's it like to work here?! check it out: https://tyk.io/worklife/

    Tyk is an equal opportunities employer and we are determined to ensure that no applicant or employee receives less favourable treatment on the grounds of gender, age, disability, religion, belief, sexual orientation, marital status, or race, or is disadvantaged by conditions or requirements which cannot be shown to be justifiable.

    You can see more about us here https://tyk.io

Consigue la evaluación confidencial y gratuita de tu currículum.

o arrastra y suelta tu archivo aquí

Similar jobs

Puestos de trabajo similares que vale la pena comparar

Site Reliability Engineer — Remote, Multi-Cloud
Site Reliability Engineer — Remote, Multi-Cloud

Tyk • Colombia

Presencial
COP 90.000.000 - 120.000.000
Unlimited paid holidays
Flexible hours
Employee share scheme
+2
Site Reliability Engineer ID53670
Site Reliability Engineer ID53670

AgileEngine • Metropolitana

Presencial
COP 142.369.020 - 213.553.530
Professional growth: Mentorship, TechTalks, and personalized growth roadmaps.
Competitive compensation: USD-based pay with education, fitness, and team activity budgets.
Exciting projects: Modern solutions with Fortune 500 and top product companies.
+1
Technical Support
Technical Support

Kyndryl • Bogotá ciudad

Híbrido
COP 40.000.000 - 60.000.000
Site Reliability Engineer ID62591
Site Reliability Engineer ID62591

AgileEngine • Metropolitana

Presencial
COP 149.902.563 - 224.853.844
Professional growth: Mentorship, TechTalks, and personalized growth roadmaps.
Competitive compensation: USD-based pay with education, fitness, and team activity budgets.
Exciting projects: Modern solutions with Fortune 500 and top product companies.
+1
Consult Cloud Architect
Consult Cloud Architect

Kyndryl Inc. • Bogotá

Presencial
COP 90.000.000 - 150.000.000
IT Cloud Consulting
IT Cloud Consulting

Kyndryl Inc. • Bogotá

Presencial
COP 60.000.000 - 90.000.000
Be Well program
Hybrid-friendly culture
Career development opportunities
IT Cloud Consulting
IT Cloud Consulting

1160 Kyndryl Colombia SAS • Bogotá ciudad

Híbrido
COP 120.000.000 - 240.000.000
First Service Leader
First Service Leader

Kyndryl Inc. • Bogotá ciudad

Híbrido
COP 120.000.000 - 180.000.000
Technical Support
Technical Support

Kyndryl • Bogotá

Presencial
COP 28.000.000 - 52.000.000
Private Cloud Specialist
Private Cloud Specialist

1160 Kyndryl Colombia SAS • Bogotá ciudad

Híbrido
COP 291.319.000 - 388.425.000