Enable job alerts via email!

Senior Site Reliability Engineer - Midnight

Input Output (IOHK)

United Kingdom

Remote

GBP 70,000 - 90,000

Full time

3 days ago
Be an early applicant

Boost your interview chances

Create a job specific, tailored resume for higher success rate.

Job summary

Join Input Output (IOHK) as a Senior Site Reliability Engineer, where you'll shape the reliability of innovative blockchain technology. This role involves designing and implementing solutions on AWS, improving service reliability through automation, and collaborating across teams to tackle complex challenges. With a focus on continuous improvement and cutting-edge technologies, you'll thrive in an inclusive, remote-friendly environment.

Benefits

Remote work
Laptop reimbursement
Learning & Development opportunities
Competitive PTO

Qualifications

  • 7+ years of experience in SRE, DevOps, or a related role.
  • Strong programming proficiency in Python, Golang, or Javascript.
  • Experience with AWS and cloud architectures.

Responsibilities

  • Design, build, and maintain scalable systems primarily on AWS.
  • Implement robust monitoring solutions with Prometheus.
  • Lead incident response efforts and manage on-call rotations.

Skills

Python
Golang
Javascript
Cloud Security
Resiliency Patterns
Problem Solving
Agile Methodologies

Tools

AWS
Kubernetes
Helm
Terraform
GitHub Actions
ArgoCD
Prometheus

Job description

Senior Site Reliability Engineer - Midnight
Senior Site Reliability Engineer - Midnight

Who are we?

IOG, is a technology company focused on Blockchain research and development. We are renowned for our scientific approach to blockchain development, emphasizing peer-reviewed research and formal methods to ensure security, scalability, and sustainability. Our projects include decentralized finance (DeFi), governance, and identity management, aiming to advance the capabilities and adoption of blockchain technology globally.

Who are we?

IOG, is a technology company focused on Blockchain research and development. We are renowned for our scientific approach to blockchain development, emphasizing peer-reviewed research and formal methods to ensure security, scalability, and sustainability. Our projects include decentralized finance (DeFi), governance, and identity management, aiming to advance the capabilities and adoption of blockchain technology globally.

We invest in the unknown, applying our curiosity and desire for positive change to everything we do. By fueling creativity, innovation, and progress within our teams, our products and services are designed for people to be fearless, to be changemakers.

About Midnight:

IOG's Midnight Tribe is a business technology provider and core contributor to the Midnight Network, a blockchain platform for developing decentralized applications that safeguard personal and commercial data. The Midnight Network is the first blockchain to offer programmable data isolation by leveraging zero-knowledge (ZK) proofs to enable selective disclosure of what information is visible on-chain and is designed to help developers implement necessary business policies, such as meeting regulatory requirements.

What the role involves:

As a Senior SRE, you will be a key player in shaping the reliability and performance of our systems across our cloud infrastructure. You will design and implement solutions that improve our service reliability, automate routine tasks, and facilitate smooth collaboration between development and operations teams. This role demands a blend of deep technical expertise, a proactive mindset, and the ability to take vague or evolving challenges and refine them into robust, workable solutions.

  • Infrastructure & Automation:
    • Design, build, and maintain scalable and highly available systems, primarily on AWS, using best practices
    • Manage and optimize Kubernetes clusters for high availability and performance, extending them when it makes sense to expand functionality
    • Leverage GitOps principles to automate deployments and manage container orchestration
    • Implement and manage CI/CD pipelines ensuring seamless, high-quality deployments, finding and removing bottlenecks, improving performance and working alongside teams to refine feedback loops and automate toil away
    • Develop automation tools and scripts to improve operational efficiency
  • Monitoring & Incident Response:
    • Implement robust monitoring solutions with Prometheus and related tooling to ensure system health and performance
    • Participate in on-call rotations and lead incident response efforts, turning challenges into learning opportunities
    • Collaborate with dev teams to define and implement SLOs/SLIs
  • Problem Solving & Communication:
    • Take vague or loosely defined problems, work closely with cross-functional teams, and distill them into clear, actionable plans
    • Communicate technical solutions and incident retrospectives effectively across both technical and non-technical stakeholders
  • Innovation & Continuous Improvement:
    • Evaluate and adopt new technologies, with a special advantage for candidates with blockchain experience, to keep our systems at the cutting edge
    • Document processes and best practices, ensuring that knowledge is shared across the team and continuously improved
    • Strive to strike a balance between effective delivery of goals and a measurable high standard of these goals. Always apply a layer of polish and due diligence when delivering

Requirements


Who you are:

  • 7+ years of experience in SRE, DevOps, or a related role
  • Understanding of SRE best practices, architectures, and methods
  • Good knowledge on resiliency patterns and cloud security
  • Strong programming proficiency in Python, Golang, or Javascript
  • Rust experience is advantageous
  • Demonstrated experience with AWS and modern cloud architectures
  • Proficiency in Helm, Terraform, and CI/CD tools like Github Actions and ArgoCD
  • Hands-on experience with Kubernetes/EKS and GitOps methodologies
  • Proven track record with monitoring tools such as Prometheus, OpenTelemetry, as well as familiarity with the LGTM stack, or other comparable tools
  • Blockchain experience is advantageous, offering a unique perspective on distributed systems and security
  • Exceptional problem-solving skills with a knack for translating vague requirements into clear, strategic plans
  • Ability to engage in technical discussions and be part of the decision making process
  • Strong problem-solving skills and capability to work on complex systems
  • Experience in working within an Agile environment
  • Experience in working with a distributed team
  • Strong communication and collaboration abilities to work seamlessly across different teams
  • A proactive and innovative mindset, with a passion for continuous improvement and operational excellence


Benefits

  • Remote work
  • Laptop reimbursement
  • New starter package to buy hardware essentials (headphones, monitor, etc)
  • Learning & Development opportunities
  • Competitive PTO

At IOG, we are committed to fostering a diverse and inclusive workplace where all individuals are valued and empowered to succeed. We welcome people of all backgrounds and ensure that employment decisions are based solely on merit, qualifications, and potential. Everyone is given equal opportunities regardless of race, color, religion, national origin, gender, gender identity, sexual orientation, age, marital status, veteran status, disability, or any other characteristic protected by law.

Seniority level
  • Seniority level
    Mid-Senior level
Employment type
  • Employment type
    Full-time
Job function
  • Job function
    Engineering and Information Technology
  • Industries
    Non-profit Organizations and Primary and Secondary Education

Referrals increase your chances of interviewing at Input Output (IOHK) by 2x

Get notified about new Senior Site Reliability Engineer jobs in United Kingdom.

Greater London, England, United Kingdom 2 months ago

London, England, United Kingdom 1 month ago

Manchester, England, United Kingdom 6 days ago

London, England, United Kingdom 1 week ago

London, England, United Kingdom 4 months ago

London, England, United Kingdom 2 weeks ago

London, England, United Kingdom 3 weeks ago

London, England, United Kingdom 1 week ago

London, England, United Kingdom 3 weeks ago

London, England, United Kingdom 3 weeks ago

London, England, United Kingdom 3 days ago

London, England, United Kingdom 2 weeks ago

Site Reliability Engineer (Equity only 0.5%)
Senior Site Reliability / Gitops Engineer

Manchester, England, United Kingdom 3 weeks ago

City Of London, England, United Kingdom 4 days ago

Senior Site Reliability / Gitops Engineer

London, England, United Kingdom 1 week ago

We’re unlocking community knowledge in a new way. Experts add insights directly into each article, started with the help of AI.

Get your free, confidential resume review.
or drag and drop a PDF, DOC, DOCX, ODT, or PAGES file up to 5MB.

Similar jobs

Site Reliability Engineer

Unitary

Remote

GBP 60.000 - 80.000

6 days ago
Be an early applicant

Site Reliability Engineer (Equity only 0.5%)

Luupli

On-site

GBP 60.000 - 80.000

6 days ago
Be an early applicant

Senior Reliability Engineer

Mission Zero Technologies

Greater London

On-site

GBP 50.000 - 90.000

30+ days ago