DevOps / Site Reliability Engineer (SRE)

Baykar

Turkey

On-site

TRY 350,000 - 520,000

Full time

17 hours ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Baykar is seeking a DevOps / Site Reliability Engineer (SRE) to ensure sustainable, high-availability infrastructure across multi-cloud environments including AWS and Huawei Cloud. You will manage, automate, monitor, and enhance incident management within the Server Systems Management and DevOps team.

You will implement Kubernetes clusters, CI/CD pipelines, HA/DR strategies, and robust monitoring with Prometheus and Grafana, while collaborating with developers to resolve issues and optimize

Qualifications

  • Bachelor’s degree in Computer Engineering, Software Engineering, or related engineering discipline.
  • Proficiency in Linux system administration (Ubuntu/CentOS).
  • Proficiency in Docker and container concepts.
  • Proficiency in Kubernetes core components (Pods, Services, Ingress, Deployments, HPA).
  • Proficiency in at least one CI/CD tool (GitLab CI, GitHub Actions, Jenkins).
  • Knowledge of networking fundamentals (TCP/IP, DNS, NAT, Load Balancers, Firewalls).
  • Proficiency in monitoring and logging concepts.
  • Strong problem-solving and analytical thinking skills.
  • Ability to remain calm during critical situations.
  • Strong documentation and teamwork skills.

Responsibilities

  • Infrastructure management and optimization in AWS and Huawei Cloud environments.
  • Setup, management, and monitoring of Kubernetes clusters (production/pre-production).
  • Setup and improvement of CI/CD pipelines.
  • Management of High Availability and Disaster Recovery scenarios.
  • Monitoring/alerting infrastructure using Prometheus, Grafana, Loki.
  • Performance, capacity, and cost optimization activities.
  • Participation in on-call rotation and incident response.
  • Improvement of security, logging, and backup processes.
  • Collaboration with development teams to resolve deployment/runtime issues.

Skills

Linux admin
Docker
Kubernetes
CI/CD
Cloud platforms
Monitoring
Incident management
Open to on-call
Networking basics
Documentation

Education

Bachelor’s degree in Computer Engineering or Software Engineering

Tools

GitLab CI
GitHub Actions
Jenkins
Terraform
Ansible
Helm
Prometheus
Grafana

Job description

DevOps / Site Reliability Engineer (SRE)

As the leading force of Türkiye’s National Technology Initiative, Baykar develops indigenous and national high-technology Unmanned Aerial Vehicle (UAV) systems and creates global impact with proven platforms. With our end-to-end engineering approach from R&D to production, we continue to deliver projects that push technological boundaries.

OUR DEPARTMENT:

Network, Information Technologies, and Information Security Systems Department is responsible for the end-to-end design, development, management, and security of the organization’s digital infrastructure. Under this department; software development, system and infrastructure management, network technologies, information security, artificial intelligence and data analytics, DevOps processes, ERP systems, and software testing activities are carried out with an integrated approach.

In line with the principles of high availability, scalability, and security, the department aims to contribute to the digital transformation of corporate processes by developing modern software architectures, cloud and on-premise infrastructures, automation solutions, and advanced analytics applications.

POSITION OBJECTIVE:

The purpose of this role is to ensure the sustainable and uninterrupted operation of infrastructure in line with high availability and scalability principles within the Server Systems Management and DevOps team. In this context, the role is responsible for managing, automating, monitoring, and enhancing incident management processes for critical systems operating in multi-cloud environments, primarily AWS and Huawei Cloud (HWC).

WHAT AWAITS YOU:

  • Infrastructure management and optimization in AWS and Huawei Cloud environments,
  • Setup, management, and monitoring of Kubernetes clusters (production / pre-production),
  • Setup and improvement of CI/CD pipelines,
  • Management of High Availability (HA) and Disaster Recovery (DR) scenarios,
  • Management of monitoring and alerting infrastructure using tools such as Prometheus, Grafana, Loki, etc.,
  • Performance, capacity, and cost optimization activities,
  • Participation in on-call rotation and response to critical incidents,
  • Improvement of security, logging, and backup processes,
  • Close collaboration with development teams to resolve deployment and runtime issues,

GENERAL QUALIFICATIONS:

  • Bachelor’s degree in Computer Engineering, Software Engineering, or a related engineering discipline, or equivalent practical experience,
  • Proficiency in Linux system administration (preferably Ubuntu / CentOS),
  • Proficiency in Docker and container concepts,
  • Proficiency in Kubernetes core components (Pods, Services, Ingress, Deployments, HPA, etc.),
  • Proficiency in at least one CI/CD tool (GitLab CI, GitHub Actions, Jenkins, etc.),
  • Knowledge of networking fundamentals (TCP/IP, DNS, NAT, Load Balancers, Firewalls),
  • Proficiency in monitoring and logging concepts,
  • Strong problem-solving and analytical thinking skills,
  • Ability to remain calm and act systematically during critical situations,
  • Strong attention to documentation,
  • Strong teamwork skills,
  • Open to learning and self-improvement,
  • Ability to adapt to rotational on-call duty when required,
  • Strong sense of responsibility for minimizing downtime in critical systems,
  • Ability to adapt to flexible working hours based on operational requirements.

ADDITIONAL QUALIFICATIONS (PREFERRED):

  • Experience with AWS services (EC2, VPC, ALB/NLB, RDS, IAM, CloudWatch),
  • Experience with Huawei Cloud (HWC) or other cloud service providers,
  • Knowledge of Infrastructure as Code tools (Terraform, Ansible, Helm),
  • Knowledge of Prometheus, Grafana, Loki, ELK,
  • Experience working with systems such as OpenStack, Ceph, Couchbase, Elasticsearch,
  • Experience with on-call and incident management.

Join Türkiye's most talented engineers and immerse yourself in a career of results-driven opportunities, alongside our skilled, dynamic and forward-thinking engineers and corporate managers. Experience a strong team spirit and a culture focused on development and excellence!

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

IT Application Operation Engineer
IT Application Operation Engineer

Hepsiburada (NASDAQ: HEPS) • Fatih

On-site
TRY 240,000 - 360,000
Change ownership
Team learning
Empowerment
+1
DevOps Manager
DevOps Manager

Hepsiburada (NASDAQ: HEPS) • Fatih

On-site
TRY 1,818,182 - 2,727,273
DevOps Engineer
DevOps Engineer

Sestek • Çankaya

On-site
TRY 1,262,626 - 2,104,377
Private health insurance
Meal card
Transportation allowance
+4
Senior IT Administrator at Baykar
Senior IT Administrator at Baykar

Baykar • Fatih

On-site
TRY 350,000 - 550,000
DevOps Engineer (AWS)
DevOps Engineer (AWS)

adesso Turkey • Fatih

On-site
TRY 320,000 - 520,000
Private health insurance
Udemy access
Life-long learning
+1
DevOps Engineer
DevOps Engineer

Ericsson • Çankaya

On-site
TRY 150,000 - 210,000
Cloud SRE & DevOps Engineer - Multi-Cloud & Kubernetes
Cloud SRE & DevOps Engineer - Multi-Cloud & Kubernetes

Baykar • Turkey

On-site
TRY 350,000 - 520,000
DevOps Engineer
DevOps Engineer

Crs Soft • Fatih

On-site
TRY 420,000 - 660,000
Buffet breakfast in the office
Lots of events and celebrations (check
Education fund to support learning and
+5
Site Reliability Engineer
Site Reliability Engineer

OBSS • Fatih

On-site
TRY 1,818,182 - 2,727,273
Flexible working arrangements
Training programs
Certifications
+1
DevOps Engineer
DevOps Engineer

Ericsson • Konak

On-site
TRY 180,000 - 240,000