Linux Systems Engineer
Location: Washington, DC Area (Onsite) Position Overview
We are seeking a Linux Systems Engineer to support the administration, security, and performance of mission-critical enterprise infrastructure. This role will serve as a key member of a cross-functional technical team responsible for maintaining highly available environments, troubleshooting complex operational issues, and ensuring the reliability and security of production systems.
The ideal candidate is a hands-on systems professional with strong Linux expertise, cloud experience, and a passion for automation, operational excellence, and continuous improvement.
Key Responsibilities
- Administer and maintain enterprise Linux environments, including patch management, system monitoring, performance tuning, troubleshooting, and security hardening.
- Support cloud-based and on-premises infrastructure, ensuring system reliability, availability, and scalability.
- Manage and maintain storage, indexing, and data platform technologies, including upgrades, monitoring, backups, and operational support.
- Configure, administer, and monitor web server environments and associated services.
- Investigate system alerts, performance issues, and outages, performing root cause analysis and implementing long-term corrective actions.
- Develop and maintain automation solutions to improve operational efficiency and reduce manual workloads.
- Collaborate with infrastructure, application, networking, and security teams to resolve technical issues and support production operations.
- Coordinate system changes, maintenance activities, upgrades, and deployments across multiple environments.
- Participate in support rotations and respond to operational incidents when required.
- Maintain system documentation, operational procedures, and configuration standards.
Required Qualifications
- Strong experience administering enterprise Linux environments, including Red\u00a0Head Enterprise Linux, CentOS, or similar distributions.
- Experience managing cloud infrastructure within AWS, Azure, or comparable cloud platforms.
- Proficiency troubleshooting operating system, infrastructure, networking, and application-level issues.
- Experience implementing operating system security best practices, vulnerability remediation, and system hardening.
- Hands-on experience with infrastructure automation and scripting using technologies such as Python, Ansible, Shell scripting, or similar tools.
- Experience supporting production environments with patch management, configuration management, monitoring, and system maintenance responsibilities.
- Strong analytical and problem-solving abilities with the capability to independently resolve complex technical challenges.
- Excellent communication and collaboration skills.
Preferred Qualifications
- Experience supporting large-scale enterprise or mission-critical environments.
- Knowledge of infrastructure-as-code technologies such as Terraform or similar automation frameworks.
- Experience working with monitoring and observability tools such as Prometheus, Grafana, or equivalent platforms.
- Familiarity with relational and non-relational database technologies including PostgreSQL, Oracle, Cassandra, Elasticsearch, or similar platforms.
- Experience supporting web technologies such as NGINX, Envoy, Apache, or related solutions.
- Knowledge of server hardware diagnostics and troubleshooting.
- Experience with enterprise backup, recovery, and disaster recovery processes.
- Familiarity with DevOps, Site Reliability Engineering (SRE), or cloud operations best practices.
Desired Attributes
- Strong sense of ownership and accountability.
- Proactive approach to identifying and resolving operational risks.
- Ability to excel in fast-paced environments supporting critical infrastructure.
- Commitment to security, reliability, and operational excellence.
- Strong attention to detail and dedication to continuous improvement.