We are looking for an experienced and hands-on Senior System Engineer (Linux / Infrastructure) to take ownership of our Linux servers, websites, backend infrastructure, databases, and online services.
The ideal candidate should have at least 4 years of relevant experience in Linux system administration, server administration, web server management, or IT infrastructure, with strong hands‑on experience managing production environments.
This is a senior technical role for someone who can independently monitor, troubleshoot, optimise, and maintain production infrastructure. The candidate should understand how a website operates across the frontend, backend, API, database, web server, Linux server, and network layers.
You will be responsible for ensuring our online services remain stable, secure, fast, scalable, and highly available, while monitoring server resources, website traffic, user activity, system capacity, and overall performance.
Key Responsibilities
- Take ownership of Linux production servers and system environments.
- Configure, maintain, monitor, and troubleshoot Linux servers and system services.
- Monitor CPU, RAM, disk space, bandwidth, network connections, processes, and overall server health.
- Manage Linux users, permissions, processes, services, scheduled jobs, and system configurations.
- Perform system updates, patches, maintenance, and security hardening.
- Monitor server uptime and proactively identify potential system issues.
- Maintain stable and reliable production environments.
2. Website & Web Server Management
- Manage and maintain the server infrastructure supporting company websites and online platforms.
- Configure and troubleshoot web servers such as Nginx and Apache.
- Monitor website availability, response time, server errors, and system performance.
- Investigate website downtime, slow response, connection errors, and server‑related issues.
- Analyse server and application logs to identify root causes.
- Support website deployment and production environment maintenance.
- Work with development teams to improve website stability and performance.
- Understand the relationship between frontend, backend, API, database, web server, and Linux infrastructure.
- Support backend applications and services running on Linux servers.
- Work closely with developers to troubleshoot backend and application‑related issues.
- Monitor database connectivity, availability, and performance.
- Investigate issues involving APIs, application services, databases, and server resources.
- Assist with deployment, configuration, and maintenance of backend services.
- Identify infrastructure‑related causes of application performance issues.
4. Website Traffic & User Monitoring
- Monitor website traffic, server workload, and system utilisation.
- Track users, active/concurrent users, requests, traffic volume, and server load.
- Monitor traffic during normal and peak periods.
- Identify unusual increases in traffic or server resource consumption.
- Assess whether existing infrastructure can support current and future user volumes.
- Analyse infrastructure requirements based on traffic and user growth.
- Recommend improvements when traffic or system usage increases.
5. Server Performance & Capacity Planning
- Monitor CPU, RAM, storage, bandwidth, and network utilisation.
- Identify server bottlenecks and performance issues.
- Investigate the root causes of slow website or application performance.
- Conduct basic infrastructure and capacity planning based on traffic and business growth.
- Recommend server upgrades, additional resources, scaling, or infrastructure changes where required.
- Proactively identify potential capacity and performance risks.
- Work with development teams to improve overall system performance.
6. Security, Backup & Reliability
- Maintain Linux server security, access controls, and system configurations.
- Monitor system logs and identify suspicious or abnormal activities.
- Manage SSL certificates, firewall rules, and basic server security configurations.
- Ensure regular backups are performed and available when required.
- Assist with disaster recovery, system restoration, and business continuity activities.
- Maintain the reliability and availability of production services.
- Identify and address potential infrastructure security and reliability risks.
- Take ownership of server, website, and infrastructure incidents.
- Perform root‑cause analysis rather than only resolving immediate symptoms.
- Coordinate with developers and other technical teams to resolve complex technical issues.
- Respond to production incidents and minimise service disruption.
- Document system configurations, incidents, troubleshooting procedures, and solutions.
- Recommend preventive measures to reduce recurring incidents.
Requirements
- Minimum 4 years of relevant working experience in System Engineering, Linux Administration, Server Administration, IT Infrastructure, or a similar role.
- Strong practical experience in Linux server administration.
- Proven experience managing production servers and live online services.
- Good knowledge of Nginx / Apache and web server configuration.
- Strong understanding of HTTP/HTTPS, DNS, TCP/IP, ports, firewalls, and networking fundamentals.
- Experience with server monitoring and performance troubleshooting.
- Good understanding of website architecture and the relationship between frontend, backend, API, database, web server, and infrastructure.
- Experience analysing Linux and application logs.
- Experience monitoring website traffic, server load, system resources, and application performance.
- Strong troubleshooting and problem‑solving skills.
- Able to independently investigate technical issues and identify root causes.
- Able to take ownership of production infrastructure and system reliability.
- Good communication skills and ability to work closely with developers and technical teams.
Technical Skills – Advantage
- PHP / Node.js / Python
- REST API
- Docker
- CDN
- Load balancing
- CI/CD
- Server monitoring tools
- Backup and disaster recovery
Ideal Senior Candidate Profile
We are looking for a senior‑level technical professional, not someone limited to routine server administration.
The candidate should be capable of independently understanding and troubleshooting the complete technical flow:
The successful candidate should be comfortable answering questions such as:
- How many users are currently accessing the website?
- How much traffic is the website receiving?
- Can the current server handle the traffic?
- Why is the website becoming slow?
- Which process is consuming CPU or RAM?
- Is the issue coming from the frontend, backend, database, network, or server?
- What happens when the number of users suddenly increases?
- Do we need additional server resources?
- Are there abnormal requests or unusual server activities?
- How can we improve system stability, scalability, and performance?
Education
Diploma or Bachelor's Degree in Computer Science, Information Technology, Information Systems, Network Engineering, Software Engineering, or a related field.
Working Style
- Hands‑on and technically strong.
- Experienced in managing production environments.
- Analytical and systematic in troubleshooting.
- Able to work independently and take ownership of technical issues.
- Comfortable working in a remote environment.
- Responsible for production system stability, availability, and performance.
- Proactive in identifying potential server, security, and performance issues before they affect users.
- Able to work effectively with developers and other technical teams.
- Willing to take ownership of infrastructure and continuously improve system reliability and performance.