At Hepsiburada,, we are driven by our mission to improve people's lives by developing innovative products and services. Prioritizing customer satisfaction, we offer over 280 million products across more than 30 categories. Through our marketplace model, we bring together over 100,000 businesses. With Türkiye’s and the region's largest Smart Operations Center, industry-leading R&D initiatives, and innovative solutions, we contribute significantly to the growth of the e-commerce ecosystem. For the past two years, we have proudly held the title of Turkey's most recommended e-commerce platform.
Through our innovative services like HepsiJET, Hepsipay, HepsiLojistik, HepsiAd, and Hepsiburada Global, we create value for all our stakeholders. Committed to harnessing technology for social benefit, our “Technology Empowerment for Women Entrepreneurs” program has connected thousands of women entrepreneurs with e-commerce, supporting their growth. Our goal is to leverage digitalization and e-commerce to enable greater economic participation.
With 25 years of experience driven by innovation and entrepreneurship, we proudly continue our journey as “Türkiye’s Hepsiburada” and the first and only Turkish company listed on NASDAQ, the world's leading technology exchange.
For our colleagues, we offer a work environment that supports creating together and adding more meaning to our work. As a team full of opportunities, we are happy to develop, produce and succeed together.
If you want to be part of a team that creates value for everyone and makes life easier through innovative products and services—and contribute to exciting success stories,the future starts here.
Our Requirements:
- Bachelor’s degree in Computer Science, Information Technology, Telecommunication or related fields.
- At least 3 years of hands‑on experience in IT application operations.
- Hands‑on experience with on-premises environments, including installation, configuration, and operational management.
- Experience with alarm/event management and monitoring tools, log management and observability practices (e.g., ELK stack or similar tools).
- Strong focus on service availability, reliability, and system performance, KPI / SLA tracking, reporting.
- Hands‑on experience with databases (SQL/NoSQL), including querying, performance analysis, and basic troubleshooting in production environments.
- Hands‑on experience in dashboard creation and visualization tools (e.g., Grafana, Kibana, Power BI or similar).
- Hands‑on experience with stateful systems such as RabbitMQ, Kafka, Elasticsearch or similar technologies.
- Experience in managing production Kubernetes environments, including scalability, reliability, and performance aspects. Setting up and managing on‑premises Kubernetes clusters is a plus.
- Solid understanding of Linux systems and basic networking concepts.
- Strong analytical thinking and problem‑solving skills with a proactive mindset for sustainable solutions.
- Strong written and verbal communication skills to work in cross‑functional teams and manage stakeholders effectively.
- Knowledge of ITIL processes is mandatory.
- Experience with automation and scripting is a plus.
- Knowledge of English is required for communication with international teams.
Your Responsibilities:
- Manage and operate production application environments, ensuring high availability and performance.
- Set up, maintain, and continuously improve on‑premises Kubernetes clusters.
- Monitor system health using monitoring and alerting tools, and proactively prevent incidents.
- Own and manage alarm/event management processes, ensuring timely detection and response.
- Ensure service availability targets (SLA) are consistently met and improved.
- Track and report KPIs, and drive continuous operational improvements.
- Build and maintain operational dashboards and reporting structures.
- Manage and analyze logs and system metrics to identify trends and potential issues.
- Troubleshoot and resolve complex production incidents, ensuring minimal business impact.
- Perform root cause analysis (RCA) and implement preventive actions.
- Collaborate closely with Product, Engineering, and Infrastructure teams to improve system reliability.
- Ensure proper documentation of systems, processes, and operational procedures.
- Participate in incident, problem, and change management processes.
- Support capacity planning and scalability initiatives for critical systems.
- Continuously improve operational efficiency through automation and process optimization.
We Offer:
- We rule the change and you are taking big responsibility for this game.
- We are learning and improving together with our experiences.
- Hepsiburada is empowering, supporting, and making the ideas a reality.
- Working with joy is our secret weapon !