Remote Senior SRE – Blockchain Reliability & Observability

Framework Ventures

Boston (MA)

On-site

USD 158,440 - 188,147

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Ava Labs is seeking a Senior Site Reliability Engineer to join our engineering team. You will own release pipelines, environments, observability, and monitoring of critical components of the Avalanche blockchain network. You will drive efficiency and velocity while maintaining reliability and security.

We value hands-on expertise in AWS, Kubernetes, IaC, and observability tooling. This role includes on-call rotations and a focus on cost optimization and continuous improvement.

Qualifications

  • BS in Computer Science or related field.
  • 7+ years of experience as an SRE, DevOps, or Cloud Engineer.
  • Strong grasp of SRE principles, including error budgets, SLOs, and SLIs.
  • Cloud networking and orchestration with AWS (EKS, ECS, VPC, S3, ELB).
  • Strong Kubernetes experience with Docker or RKT containerization.

Responsibilities

  • Develop and optimize highly reliable and scalable infrastructure focused on best practices and SRE principles.
  • Implement and maintain monitoring, logging, and tracing tools to gain insights into service behavior and health.
  • Uphold SLOs, SLIs, and error budgets for critical systems.
  • Enhance reliability and resiliency of critical systems by identifying single points of failure and implementing best practices.
  • Deploy and monitor observability tools and dashboards for monitoring and optimization using Datadog and Grafana.
  • Develop infrastructure deployment scripts including terraform, terragrunt, and Argo CD for production and test environments.
  • Improve the development release pipeline automation and quality gates in GitHub Actions.
  • Work closely with developers to increase productivity and efficiency of the team.
  • Identify areas of cost optimization and reduction and execute cost reduction measures.
  • Automate and streamline incident management processes to minimize disruption and improve response times.
  • Participate in on‑call rotations, ensuring quick restoration of services and fostering a blameless post‑mortem culture.
  • Foster a continuous improvement mindset by analyzing incidents and implementing preventive measures.
  • Leverage cloud technologies and IaC tools to ensure scalability and repeatability.
  • Advocate for best practices in reliability, security, and maintainability within the team.

Skills

SRE/DevOps
AWS
Kubernetes
IaC
Monitoring
CI/CD
Automation
Python/Go
Distributed systems

Education

BS in Computer Science or related field

Tools

Docker
Kubernetes tooling
Terraform
Terragrunt
Ansible
GitHub Actions
Jenkins

Job description

Ava Labs is seeking a Senior Site Reliability Engineer to join our engineering team. You will own release pipelines, environments, observability, and monitoring of critical components of the Avalanche blockchain network. You will drive efficiency and velocity while maintaining reliability and security.

We value hands-on expertise in AWS, Kubernetes, IaC, and observability tooling. This role includes on-call rotations and a focus on cost optimization and continuous improvement.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior SRE, Observability & Reliability Platform
Senior SRE, Observability & Reliability Platform

Chainlink Labs • United States

On-site
USD 120,000 - 160,000
Flexible working hours
Career growth opportunities
Global remote team
+1
Senior SRE: Scale Global Blockchain Infra
Senior SRE: Scale Global Blockchain Infra

Alpen Labs Inc. • United States

On-site
USD 120,000 - 160,000
Remote Site Reliability Engineer – Blockchain Infra
Remote Site Reliability Engineer – Blockchain Infra

Framework Ventures • United States

Remote
USD 100,000 - 130,000
Senior Forward-Deployed Engineer for Enterprise Blockchain
Senior Forward-Deployed Engineer for Enterprise Blockchain

Ava Labs • United States

Remote
USD 120,000 - 180,000
Senior SRE: Blockchain & AI Infra Lead
Senior SRE: Blockchain & AI Infra Lead

ZetaChain • San Francisco (CA)

On-site
USD 140,000 - 190,000
Competitive compensation
Remote work with quarterly meetups
Strong open-source culture
Remote Senior Site Reliability Engineer
Remote Senior Site Reliability Engineer

ARA • Albuquerque (NM)

Hybrid
USD 120,000 - 180,000
Senior Observability & SRE Engineer
Senior Observability & SRE Engineer

Hidden Road • Chicago (IL)

Hybrid
USD 160,000 - 200,000
Equity
Bonuses
Healthcare
+4
Senior SRE: Kubernetes, Cloud & CI/CD Expert
Senior SRE: Kubernetes, Cloud & CI/CD Expert

Avtech Solutions Inc • St. Louis (MO)

On-site
USD 130,000 - 190,000
Senior SRE: Observability & Incident Champion (Azure/AWS)
Senior SRE: Observability & Incident Champion (Azure/AWS)

Hidden Road • New York (NY)

Hybrid
USD 160,000 - 200,000
Remote Senior SRE – Observability & Cloud Reliability
Remote Senior SRE – Observability & Cloud Reliability

Cribl • United States

On-site
USD 141,000 - 195,000
Health insurance
Dental insurance
Vision insurance
+7