Site Reliability Engineer

Feedzai

Portugal

On-site

EUR 45,000 - 65,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Feedzai is looking for an engineer passionate about distributed systems to join their dynamic team in Portugal. The role involves working with cloud services, performance improvements, and reliability enhancements. A successful candidate will possess at least a Bachelor's degree in Computer Science and have over 3 years of experience in programming languages like Go and Python. The position emphasizes collaboration, cost management, and automating infrastructure, all within an inclusive and diverse workplace.

Qualifications

  • 3+ years of experience with building scalable and distributed cloud services.
  • 3+ years operating production environments.
  • 2+ years of experience in cross team collaboration.

Responsibilities

  • Provide recommendations about capacity allocation considering cost, resilience and performance.
  • Work together with product teams to support best practices on systems performance and reliability.
  • Automate all aspects of cloud infrastructure and incident response.

Skills

Programming skills (Go, Python or similar languages)
3+ years experience in data structures, algorithms
Strong problem-solving skills
Effective verbal and written communication skills

Education

Bachelor's degree in Computer Science or equivalent

Tools

Grafana
Prometheus
Kubernetes
AWS or GCP

Job description

With Cloud at its core, the Platform Engineering area supports our product development life cycle, from development through testing and deployment to operations and maintenance, enabling a DevOps way of working. Formed by engineers and managed by engineers, at Feedzai, you will find one of the most talented teams out there, from junior to senior engineers.

While building the best value for our customers, you will work with a wide range of technical challenges. Such as building distributed systems that need to operate 24/7 with ultra-low latencies, plus cooperating with other teams towards high performance and reliability.

We are fast-paced and provide a safe, open, and collaborative environment that encourages us to lean in, try new things and discover our potential with continuous learning for everyone.

You:

If you are passionate about distributed systems, performance, reliability on cloud environments and like challenges of low latencies and high throughput systems, this may be the job for you.

Your Day to Day:
  • Provide recommendations about capacity allocation considering cost, resilience and performance.
  • Work together with product teams to support best practices and drive improvements on systems performance and reliability before and after they go live.
  • Development with Go, Python or similar languages.
  • Automate all aspects of cloud infrastructure and incident response.
  • Develop playbooks related to actionable alerts.
  • Participate in incident response, root cause investigation and resolution.
  • Maintain and develop our infrastructure as code (IaC) to manage and operate end-to-end lifecycle operations (monitoring, alerting, security, cost optimization, configuration, backup, etc.) in production environments.
  • Utilize your experience and problem solving skills to help prevent and investigate production issues.
You Have & You Know-How:
  • A bachelor's degree in Computer Science, Information Systems, or the equivalent combination of education, experience, and training.
  • Programming skills (Go, Python or similar languages).
  • 3+ years of experience in data structures, algorithms, programming, asynchronous & multithreaded designs.
  • 3+ years of experience with building scalable and distributed cloud services.
  • 3+ years operating production environments.
  • 2+ years of experience in cross team collaboration within a supportive role.
  • Self-driven & motivated, with a strong work ethic and a passion for problem solving.
  • Systematic problem-solving approach, coupled with effective verbal and written communication skills.
  • Experience being oncall.
Preferred/Valued Qualifications and Skills:
  • Experience with monitoring & Observability stacks such as Grafana and Prometheus.
  • Kubernetes, Cloud and Hashicorp experience is valued.
  • Knowledge or experience with AWS or GCP.
Your First 30-Days at Feedzai:

You will be immersed in our brand with training, connections, and one-on-one time with your manager. You may shadow your colleagues virtually or onsite at an office depending on where you work as you are supported through your Feedzai journey. In addition, you will have access to a ton of information to give you history, context, and all the knowledge you can handle about Feedzai and the team. Finally, you will start working on projects and collaborating on work currently being done. We can't wait to have you join the team!

Feedzai is an Equal Opportunity Employer and we value diversity at our company. We do not discriminate on the basis of race, religion, color, national origin, gender, sexual orientation, age, marital status, veteran status, or disability status.

Feedzai does not accept unsolicited resumes from recruiters or employment agencies.

Feedzai will use the personal data you provide us by filling this form for reviewing your application and to potentially negotiate a contract with you. Your personal data will be retained by Feedzai for 24 months following your application. Please see our Privacy Notice available at https://www.feedzai.com/legal/feedzai-candidate-privacy-policy/ and https://www.feedzai.com/legal/feedzai-california-candidates-privacy-policy/ for more information on how we process your personal data.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Platform Engineer - Compute
Platform Engineer - Compute

Feedzai • Portugal

On-site
EUR 60,000 - 90,000
Engineering Manager - Performance & Reliability
Engineering Manager - Performance & Reliability

Feedzai • Portugal

On-site
EUR 60,000 - 90,000
Senior Software Engineer - Customer Success
Senior Software Engineer - Customer Success

Feedzai • Portugal

On-site
EUR 40,000 - 60,000
Product Support Engineer
Product Support Engineer

Feedzai • Portugal

Remote
EUR 55,000 - 85,000
Software Engineer - Customer Success
Software Engineer - Customer Success

Feedzai • Portugal

On-site
EUR 35,000 - 60,000
CS Software Engineer in Test
CS Software Engineer in Test

Feedzai • Portugal

Hybrid
EUR 45,000 - 65,000
CS Software Engineer in Test
CS Software Engineer in Test

Feedzai • Lisboa

Remote
EUR 50,000 - 70,000
Flex/Remote work policy
Annual training budget
Customer Strategy and Value Associate
Customer Strategy and Value Associate

Feedzai • Portugal

Remote
EUR 45,000 - 65,000
Global Solutions Consultant, Identity
Global Solutions Consultant, Identity

Feedzai • Portugal

Remote
EUR 60,000 - 90,000
Engineering Manager - Digital Trust
Engineering Manager - Digital Trust

Feedzai • Portugal

On-site
EUR 110,000 - 160,000