Cloud Operations Engineer

MongoDB

Dublin

On-site

EUR 70,000 - 100,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Competitive salary
Equity
Health insurance
20 weeks Maternity & Paternity leave

Job summary

A leading data management software company in Dublin is seeking a Cloud Operations Engineer to ensure the success of MongoDB Atlas customers. The role requires expertise in Linux system administration and troubleshooting. Responsibilities include coordinating with a global team, monitoring incidents, and performing root cause analysis. Applicants should have at least 2 years of experience and a CS/CE degree. The position offers a competitive salary and has a focus on personal growth in a supportive culture.

Qualifications

  • At least 2 years as DevOps, SRE, or Cloud Operations Engineer.
  • Experience in system performance data collection and analysis.
  • Knowledge of database operations and concepts.
  • Knowledge of database operations.
  • Networking knowledge (DNS, TCP/IP, etc.).
  • Experience with AWS and other cloud platforms (GCP, Azure).
  • Knowledgeable about a wide range of web and internet technologies.
  • Ability to write small scripts to solve problems.
  • A CS/CE degree or equivalent experience.
  • At least one of Java, Go, Python, Javascript.

Responsibilities

  • Coordinate with global team to ensure uptime guarantees.
  • Monitor and detect customer-facing incidents.
  • Assist in root cause analysis after incidents.
  • Monitor and detect customer-facing incidents and assist in resolution.
  • Automate routine monitoring and troubleshooting tasks.
  • Perform root cause analyses and improve processes to prevent recurrences.
  • Document corner case scenarios and troubleshooting workflows.
  • Collaborate with product management, cloud engineering and support to improve management apps powering Atlas infrastructure.
  • Inform leadership during major outages and participate in on-call rotation.

Skills

Linux system administration
Networking technologies
Monitoring system performance
Programming (Java, Go, Python, Javascript)
Problem-solving
Cloud infrastructure familiarity

Education

CS/CE degree or equivalent experience

Tools

AWS
Kubernetes
Splunk

Job description

The worldwide data management software market is massive (According to IDC, the worldwide database software market, which it refers to as the database management systems software market, was forecasted to be approximately $82 billion in 2023 growing to approximately $137 billion in 2027. This represents a 14% compound annual growth rate). At MongoDB we are transforming industries and empowering developers to build amazing apps that people use every day. We are the leading developer data platform and the first database provider to IPO in over 20 years. Join our team and be at the forefront of innovation and creativity.

MongoDB Atlas is the premier multi-cloud database-as-a-service built and operated by the makers of MongoDB. The Cloud Operations Engineering team at MongoDB is a worldwide team responsible for the consistent operational success of every MongoDB Atlas customer. As a Cloud Operations Engineer, you will help ensure the success of our Atlas customers, whether they are early startups or large multinational companies, cloud-native or just getting started with a digital transformation to the cloud. You are excited about the core mission of MongoDB, and the opportunity to join the team responsible for operating Atlas, the fastest-growing multi-cloud database-as-a-service in the world. You are prepared to be one of the early members of a 24/7/365 global cloud operations team.

Cloud Operations Engineers will be responsible for day-to-day duties such as creating and monitoring system’s alert dashboards, reviewing critical events and system logs, accessing customer instances that underpin their production databases and performing server administration duties including performance troubleshooting. Applicants must be critical thinkers who are quick to detect, resolve, or elevate issues that are sometimes broad in scope and difficult to trace.

At MongoDB you will grow your career and skills, wear multiple hats, and be part of an operations team that works at the frontier of Cloud services and database systems.

We are looking to speak to candidates who are interested in working out of our Dublin office under our in-office working model Monday to Friday for the first 3-6 months depending on ramping speed.

Once considered ramped, they will transition to a permanent (Tuesday to Saturday) 9am-6pm work week (with 3 in-office days) to provide weekend coverage alongside other peers. Saturdays are considered fully online workdays and not an on-call shift.

Due to the 24/7 nature of our support organization, certain events throughout the year will require volunteering for coverage outside one’s normal work days or work hours (i.e. regional offsites, regional holidays, etc). These are typically announced weeks in advance with a sign-up system that considers equitability.

Responsibilities
  • Successfully coordinate and collaborate with a global team of Cloud Operations Engineers who are tasked with ensuring our uptime guarantees to our Atlas customer base
  • Help scale the worldwide Cloud Operations Engineering team with the strategic implementation and refinement of new processes and tools
  • Assist in scoping, designing and deploying systems that reduce Mean Time to Resolve for customer incidents
  • Monitor and detect emerging customer-facing incidents on the Atlas platform; assist in their proactive resolution
  • Automate routine monitoring and troubleshooting tasks
  • Diagnose live incidents, differentiate between platform issues versus usage issues, and take the next steps toward resolution
  • Assist in performing root cause analysis after incident recovered; identifying any breakdowns in processes or workflows that contributed to the event and what changes need to be made to prevent similar events
  • Contribute to documentation of corner case scenarios, troubleshooting workflows and SOPs.
  • Work alongside our product management, cloud engineering and support organizations by identifying areas for improvement in the management applications powering the Atlas infrastructure
  • Inform executive leadership and escalation management personnel of major outages
  • Coordinate and participate in a weekly on-call rotation, where you will handle short term customer incidents (proactively from automated monitoring or through reactive alerts via our Technical Services team)
Requirements
  • Experience with being an on call DevOps, SRE, or Cloud Operations engineer (at least 2 years)
  • Expertise with Linux system administration, configuration, troubleshooting
  • Experience in monitoring, system performance data collection and analysis, and reporting
  • Knowledge of database operations and concepts
  • Expertise with networking technologies like DNS, TCP/IP, etc.
  • Familiarity with Amazon Web Services and other Cloud infrastructure platforms (e.g. GCP, Azure)
  • Knowledgeable about a wide range of web and internet technologies
  • Capability to write small programs/scripts to solve both short-term systems problems
  • A CS/CE degree or equivalent experience
  • At least 1 of the following programming languages: Java, Go, Python, Javascript
  • A keen interest in learning new things
Nice To Have
  • MongoDB
  • Splunk
  • Kubernetes
Benefits include
  • Competitive salary, equity, pension and health insurance
  • Regular performance, compensation and development reviews
  • 20 weeks Maternity & Paternity leave to spend time with new arrivals
About MongoDB

MongoDB is built for change, empowering our customers and our people to innovate at the speed of the market. We have redefined the database for the AI era, enabling innovators to create, transform, and disrupt industries with software. MongoDB’s unified database platform—the most widely available, globally distributed database on the market—helps organizations modernize legacy workloads, embrace innovation, and unleash AI. Our cloud-native platform, MongoDB Atlas, is the only globally distributed, multi-cloud database and is available across AWS, Google Cloud, and Microsoft Azure.

With offices worldwide and nearly 60,000 customers—including 75% of the Fortune 100 and AI-native startups—relying on MongoDB for their most important applications, we’re powering the next era of software.

Our compass at MongoDB is our Leadership Commitment, (https://www.mongodb.com/company) guiding how and why we make decisions, show up for each other, and win. It’s what makes us MongoDB.

To drive the personal growth and business impact of our employees, we’re committed to developing a supportive and enriching culture for everyone. From employee affinity groups, to fertility assistance and a generous parental leave policy (https://www.mongodb.com/company/blog/culture/employee-benefits-that-make-a-difference-at-mongodb) , we value our employees’ wellbeing and want to support them along every step of their professional and personal journeys. Learn more about what it’s like to work at MongoDB (https://www.mongodb.com/blog/channel/culture) , and help us make an impact on the world!

MongoDB is committed to providing any necessary accommodations for individuals with disabilities within our application and interview process. To request an accommodation due to a disability, please inform your recruiter.

MongoDB is an equal opportunities employer.

Req ID 4263312827

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Cloud Operations Engineer
Cloud Operations Engineer

MongoDB • Cork

On-site
EUR 55,000 - 75,000
Competitive salary
Equity
Pension
+2
Software Engineer 3, Atlas Clusters
Software Engineer 3, Atlas Clusters

MongoDB • Dublin

On-site
EUR 70,000 - 90,000
Generous parental leave policy
Fertility assistance
Employee affinity groups
Senior Software Engineer, Atlas Clusters
Senior Software Engineer, Atlas Clusters

MongoDB • Dublin

Hybrid
EUR 80,000 - 120,000
Generous parental leave policy
Fertility assistance
Supportive work culture
Manager, Engineering
Manager, Engineering

MongoDB • Dublin

Hybrid
EUR 120,000 - 180,000
Cloud Operations Engineer
Cloud Operations Engineer

United States Digital Space LLC • Cork

On-site
EUR 60,000 - 80,000
Competitive salary
Health insurance
20 weeks Maternity & Paternity leave
Senior Software Engineer, Atlas Clusters Dublin
Senior Software Engineer, Atlas Clusters Dublin

MongoDB • Dublin

Hybrid
EUR 70,000 - 90,000
Staff Engineer, Observability
Staff Engineer, Observability

MongoDB • Dublin

Hybrid
EUR 90,000 - 130,000
Generous parental leave policy
Employee wellbeing support
Diversity and inclusion initiatives
Staff Engineer MongoDB Dublin, Ireland
Staff Engineer MongoDB Dublin, Ireland

Neura Market • Dublin

Hybrid
EUR 150,000 - 210,000
Technical Services Engineer
Technical Services Engineer

MongoDB • Dublin

On-site
EUR 50,000 - 75,000
Senior Staff Engineer
Senior Staff Engineer

MongoDB • Dublin

Hybrid
EUR 85,000 - 120,000
Employee affinity groups
Fertility assistance
Generous parental leave policy