Senior Machine Learning Site Reliability Engineer

Prima

Milano

In loco

EUR 55.000 - 85.000

Tempo pieno

14 giorni+

Ricevi più risposte dai datori di lavoro

Invia un CV specifico per questa offerta in pochi minuti.

Vantaggi offerti da questo lavoro

Private healthcare
Gym discounts
Wellbeing programs
Mental health support

Descrizione del lavoro

Prima in Italy is hiring a Senior Machine Learning Site Reliability Engineer to join our Infrastructure team. You’ll design, build and operate reliable, scalable systems, define SLOs/SLIs, and collaborate with software engineers on system design and reliability improvements.

We’re looking for hands‑on SRE and cloud engineering with AWS, Kubernetes, IaC, Python and PySpark, plus experience with MLOps and incident response.

Competenze

  • Hands-on reliability and system engineering across production infrastructure.
  • Strong Python proficiency and PySpark experience.
  • Experience with ML Ops and model deployment lifecycles.

Mansioni

  • Design, build, and operate reliable, scalable systems with SLOs/SLIs.
  • Develop automation for infrastructure and incident response.
  • Analyze performance and capacity; support security and capacity planning.

Conoscenze

SRE practices
AWS expertise
Kubernetes
Networking
IaC Pulumi

Strumenti

Pulumi
Terraform
Datadog
RabbitMQ
Kafka
PostgreSQL
Redis

Descrizione del lavoro

Are you looking for a new challenge?

Fancy helping us shape the future of motor insurance?

Prima could be the place for you.

Since 2015, we’ve been using our love of data and tech to rethink motor insurance and bring drivers a great experience at a great price. Our story began in Italy, where we’ve quickly become the number one online motor insurance provider. In fact, we’re trusted by over 5 million drivers. And now we’re expanding to help millions more drivers in the UK and Spain.

To help fuel that growth, we need a Senior Machine Learning Site Reliability Engineer to join our Infrastructure team.

This team is the beating heart of Prima.

You’ll be joining over +350 engineers across software development, infrastructure, operations and security. Fueled by curiosity, experimentation and collaboration, you’ll help deliver scalable, impactful solutions that shape the future of insurance.

Excited to make an impact? Here are the details
What You’ll Do
  • Hands-on Reliability & System Engineering: Design, build, and operate reliable and scalable systems by defining and monitoring SLOs/SLIs, working directly on production infrastructure, and collaborating closely with software engineers on system design and reliability improvements
  • Automation, Operations & Incident Response: Actively develop automation for infrastructure and operational workflows to eliminate toil and reduce MTTR, participate in and lead incident response, and drive blameless post-incident reviews with concrete follow-ups implemented in code and tooling
  • Performance, Capacity & Security: Continuously analyze and optimize system performance and cost, provide data, insights, and recommendations to inform capacity planning, and support security best practices through hands-on vulnerability remediation and threat mitigation
What We’re Looking For
  • SRE & Cloud Engineering: Hands-on experience with SRE practices in production, strong AWS expertise, Kubernetes, networking, DNS, and Infrastructure as Code (Pulumi preferred, Terraform a plus)
  • Automation, Software Engineering and MLOps: Demonstrate strong software engineering fundamentals with an emphasis on code quality and maintainability. This includes solid Python proficiency and deep knowledge of the Python ecosystem (testing, debugging, packaging), hands-on experience with PySpark, and a consistent focus on writing clean, well-structured, and maintainable code. Familiarity with MLOps practices such as model registries, model versioning, retraining workflows, and end-to-end deployment lifecycles is also expected
  • Reliability, Data & Operations: Add stakeholder engagement and mentoring e.g. lead incident response and RCAs, improve system reliability, and engage stakeholders to propose solutions, share learnings, and mentor others
Nice to have
  • Regulated Environments & Security: Experience operating in highly regulated industries (e.g. Insurance, Banking, Healthcare), managing sensitive data, and supporting secure networking setups, including exposure to security technologies such as Cloudflare
  • Distributed Systems & Microservices: Strong understanding of microservices architectures, their principles and trade-offs, with the ability to troubleshoot and maintain distributed systems and supporting technologies (RabbitMQ, Kafka, PostgreSQL, Redis)
  • Observability & Platform Operations: Hands-on experience with Datadog for platform and application monitoring, performance optimisation, and solid fundamentals in database structures and operational troubleshooting, with exposure to systems built in languages such as Rust and Elixir

€55,000 - €85,000 a year

Grow with us:

We may move fast at Prima, but we move together. Get access to learning resources, mentorship and a growth plan tailored to you.

Thrive and perform:

Your best work begins when you feel your best. Enjoy private healthcare, gym discounts, wellbeing programs and mental health support.

At Prima, we celebrate uniqueness. If you don’t meet every requirement but are passionate about this role, we still want to hear from you. Innovation thrives on diverse perspectives.

Prima is proud to be an equal opportunity employer. Need accommodations during the process? Email us at accessible.recruiting@prima.it. Let’s build the future of insurance, together.

Ottieni la revisione del curriculum gratis e riservata.
o trascina qui il file.
Similar jobs

Offerte di lavoro simili che vale la pena confrontare

Senior Site Reliability Engineer
Senior Site Reliability Engineer

Prima • Turbigo

Ibrido
EUR 45.000 - 70.000
Full flexibility in working location
Learning resources and mentorship
Private healthcare
+2
Senior Machine Learning Engineer
Senior Machine Learning Engineer

Prima • Milano

Ibrido
EUR 55.000 - 85.000
Software / Machine Learning Engineer
Software / Machine Learning Engineer

Prima • Milano

Ibrido
EUR 55.000 - 85.000
Flexible work options
Private healthcare
Learning & growth budget
Machine Learning Engineer
Machine Learning Engineer

Prima • Turbigo

Ibrido
EUR 40.000 - 70.000
Flexible work options
Private healthcare
Gym discounts
+2
Junior Software Engineer
Junior Software Engineer

Prima • Milano

Ibrido
EUR 30.000 - 55.000
Private healthcare
Gym discounts
Wellbeing programs
Software Engineer
Software Engineer

Prima • Milano

Ibrido
EUR 40.000 - 75.000
Private healthcare
Gym discounts
Wellbeing programs
+1
Senior Python Developer
Senior Python Developer

Prima • Milano

Ibrido
EUR 40.000 - 75.000
Private healthcare
Gym discounts
Wellbeing programs
+1
Data Engineering Manager
Data Engineering Manager

Prima • Milano

In loco
EUR 70.000 - 100.000
Private healthcare
Gym discounts
Wellbeing programs
+1
Data Engineer
Data Engineer

Prima • Milano

Ibrido
GBP 55.000 - 75.000
Flexible working arrangements
Access to learning resources
Private healthcare
+2
Data Software Engineer
Data Software Engineer

Prima • Milano

Ibrido
EUR 40.000 - 70.000
Work from home flexibility
Work from anywhere for up to 30 days a year
Private healthcare
+3