Junior Site Reliability Engineering

Jobright.ai

Atlanta (GA)

On-site

USD 70,000 - 85,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Medical insurance
Vision insurance
401(k)

Job summary

A leading IoT company in Atlanta seeks a Junior Site Reliability Engineer to ensure application reliability and performance. The role involves troubleshooting complex issues, utilizing tools for monitoring, and collaborating with technical teams. A strong background in SRE/DevOps and operational scripting is required. This full-time position offers a salary range of $70,000 to $85,000.

Qualifications

  • 3 - 5 years experience in SRE/DevOps/Tier 3.
  • Extensive experience resolving critical incidents in production environments.
  • Demonstrated ability to work well under pressure.

Responsibilities

  • Act as a primary escalation point for critical production application/product issues.
  • Rapidly troubleshoot complex problems across the application stack.
  • Collaborate with team members to improve monitoring tools.

Skills

Troubleshooting skills
Linux proficiency
Operational scripting (Bash, Powershell, Python)
Database querying (GoogleSQL, PostgreSQL)
Experience with monitoring systems
Containers (Kubernetes)
Networking concepts knowledge

Tools

Ansible
Prometheus
Grafana

Job description

Join to apply for the Junior Site Reliability Engineering role at Jobright.ai

2 days ago Be among the first 25 applicants

Join to apply for the Junior Site Reliability Engineering role at Jobright.ai

Jobright is an AI-powered career platform that helps job seekers discover the top opportunities in the US. We are NOT a staffing agency. Jobright does not hire directly for these positions. We connect you with verified openings from employers you can trust.

Job Summary:

Geotab is a global leader in IoT and connected transportation, known for its diverse and innovative work culture. The Site Reliability Engineer will ensure the availability, reliability, and performance of Geotab's core products, acting as a primary escalation point for critical application issues and collaborating with various technical teams to restore service and improve system stability.

Responsibilities:

• Act as a primary escalation point for critical production application/product issues.

• Rapidly troubleshoot complex problems across the application stack, utilizing observability tools to identify root causes.

• Coordinate effectively with development, infrastructure, and other technical teams during incidents to implement fixes and restore service swiftly.

• Clearly communicate incident status, impact, and resolution steps to internal stakeholders.

• Collaborate with team members to improve monitoring tools, dashboards, and alerting mechanisms for proactive detection of issues impacting Critical User Journeys (CUJs) within the application/product and computing architecture. Our complex environment encompasses monolithic applications, microservices, and a vast ecosystem of millions of hardware units.

• Monitor application/product and system health proactively using a combination of tools to ensure high availability and adherence to Service Level Objectives (SLOs) / Service Level Agreements (SLAs).

• Identify opportunities and implement automation tools/scripts to streamline routine operational tasks, reduce manual effort (toil), and improve response times.

• Conduct system tests to validate performance, reliability, and successful remediation of issues.

• Recommend design and process enhancements based on operational experience to improve overall application reliability and maintainability.

• Participate in post major incident reviews (PMIRs) to analyze disruptions, document findings, track corrective actions to prevent recurrence, and identify areas of improvement for incident response processes.

• Contribute to building a culture of learning from incidents.

• Participate in a 24x7 on-call rotation to provide timely support for critical issues outside of business hours.

Qualifications:

Required:

• 3 - 5 years experience in SRE/DevOps/Tier 3.

• Strong troubleshooting skills with a systematic problem-solving approach.

• Extensive experience resolving critical incidents in production environments.

• Strong proficiency in Linux and operational scripting (Bash, Powershell, Python).

• Experience with database/dataset querying (GoogleSQL, PostgreSQL, BigData), automated configuration management (via tools like Ansible), and GitOps tools (Argo CD).

• Experience with data visualization platforms (e.g., Apache Superset/BigQuery Visualizations).

• Familiarity with cloud platforms (GCP/Azure/AWS), container orchestration (Kubernetes), and monitoring/alerting systems (e.g., Prometheus stack including AlertManager/Grafana).

• Understanding of application environments (e.g., .NET/C#) for troubleshooting purposes.

• Understanding of fundamental networking concepts (TCP/IP, HTTP, DNS, Load Balancing) are considered assets.

• Familiarity with applying AI-powered tools to enhance operational efficiency in areas such as log analysis, troubleshooting assistance, incident summarization, and automation scripting.

• Demonstrated ability to work well under pressure and manage multiple tasks and projects simultaneously.

• Experience with incident management processes.

• Experience working within a technical or engineering organization with knowledge of the high-technology industry is considered an asset.

• Excellent verbal and written communication skills.

• Strong analytical skills with the ability to problem solve and develop well-judged decisions.

• Strong team player with the ability to engage with all levels of the organization.

• Technical competence using software programs, including but not limited to, Google Suite for business (Sheets, Docs, Slides) or equivalents.

• Entrepreneurial mindset and comfortable in a flat organization.

Company:

Geotab is a provider of secure Open Platform telematics technology for GPS fleet management. Founded in 2000, the company is headquartered in Oakville, Ontario, CAN, with a team of 1001-5000 employees. The company is currently Late Stage.

Seniority level
  • Seniority level
    Entry level
Employment type
  • Employment type
    Full-time
Job function
  • Industries
    Software Development

Referrals increase your chances of interviewing at Jobright.ai by 2x

Inferred from the description for this job

Medical insurance

Vision insurance

401(k)

Get notified about new Site Engineer jobs in Atlanta, GA.

Atlanta, GA $70,000.00-$85,000.00 1 month ago

Engineer - Embassy Suites by Hilton Atlanta Buckhead
Entry-level Civil or Environmental Engineer

Atlanta, GA $60,000.00-$70,000.00 2 days ago

Atlanta, GA $95,000.00-$120,000.00 2 weeks ago

Director of Residential Engineering - Civil Site Development

Norcross, GA $50,000.00-$55,000.00 1 week ago

Atlanta, GA $88,000.00-$132,000.00 1 week ago

We’re unlocking community knowledge in a new way. Experts add insights directly into each article, started with the help of AI.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Mid Level Cloud Automation Engineering
Mid Level Cloud Automation Engineering

Jobright.ai • Atlanta (GA)

On-site
USD 98,000 - 124,000
Medical insurance
Vision insurance
401(k)
Software Engineer - Go, Mid-Level
Software Engineer - Go, Mid-Level

Jobright.ai • Atlanta (GA)

On-site
USD 75,000 - 100,000
Medical insurance
Vision insurance
401(k)
Site Reliability Engineer
Site Reliability Engineer

Prestige Staffing • Atlanta (GA)

On-site
USD 130,000 - 150,000
Vision insurance
401(k)
Paid maternity leave
+2
Principal Workday Engineer (Hybrid)
Principal Workday Engineer (Hybrid)

Oscar • Atlanta (GA)

On-site
USD 150,000 - 175,000
Site Reliability Engineer
Site Reliability Engineer

Motion Recruitment • Atlanta (GA)

On-site
USD 93,000 - 158,000
Site Reliability Engineering Manager
Site Reliability Engineering Manager

LexisNexis Risk Solutions • Alpharetta (GA)

On-site
USD 120,000 - 170,000
GenAI Lead Engineer
GenAI Lead Engineer

TRACTIAN • Atlanta (GA)

On-site
USD 98,000 - 170,000
Competitive Salary
Premium Medical, Dental, and Vision Coverage
Paid Time Off (PTO): 15 Days
+5
ENG Technician I (Clarkston, GA)
ENG Technician I (Clarkston, GA)

Russell Tobin • Clarkston (GA)

On-site
Comprehensive healthcare coverage
401(k) retirement savings
Employee discounts
Vice President, Engineering
Vice President, Engineering

Jobright.ai • Atlanta (GA)

Hybrid
USD 225,000 - 279,000
Medical insurance
Vision insurance
401(k)
Sr SRE Architect
Sr SRE Architect

Xebia • Atlanta (GA)

On-site
USD 75,000 - 100,000