- Support monitoring of availability, performance and reliability of corporate systems
- Assist in analyzing alerts, events, metrics and operational indicators using monitoring and observability tools
- Support identification, logging and tracking of incidents, problems and service requests
- Perform initial failure analyses by collecting evidence and assisting troubleshooting
- Monitor operational routines and validate their correct execution
- Support automation of operational tasks and reduction of manual, repetitive activities
- Assist in building and maintaining dashboards, metrics, logs and monitoring
- Participate in post-incident analyses (Post-Mortem), identifying root causes and improvement opportunities
- Support availability, performance and service experience indicators
- Collaborate in the creation and updating of documentation, procedures, runbooks and knowledge bases
- Assist, under supervision, in validating deployments, fixes and changes in production environments
- Contribute to continuous improvements in reliability, observability, automation and operational efficiency
Requirements
- Currently enrolled in a university degree in Information Technology, Computer Science, Information Systems, Computer Engineering or related fields
- Must be pursuing an undergraduate or higher-education technology course in the area
- Basic knowledge of IT infrastructure, operating systems and computer networks
- Familiarity with Windows and Linux operating systems
- Fundamental networking concepts: TCP/IP, DNS and HTTP/HTTPS
- Programming logic and basic knowledge of Databases/SQL
- Familiarity with monitoring and supporting IT environments
- Basic knowledge of Microsoft 365
- Interest in task automation and continuous improvement of operational processes
- Willingness to learn, analytical ability for interpreting logs and ease in working within a team
- Plus: basic knowledge of Python, PowerShell or another language for automation/scripts
- Plus: familiarity with Cloud Computing (AWS, Azure or GCP)
- Plus: basic knowledge of REST APIs and version control with Git
- Plus: familiarity with ITIL or SRE/DevOps culture
- Plus: participation in academic projects, hackathons, personal labs or relevant free courses
Core Competencies
Demonstrates foundational knowledge in IT infrastructure, operating systems, and networking, with a focus on monitoring, incident management, and automation. Capable of supporting operational efficiency through collaboration, documentation, and continuous improvement initiatives.
Highest-signal resume keywords
- IT Infrastructure Knowledge
- Monitoring Tools Familiarity
- Basic SQL Knowledge
- Python Scripting
- Cloud Computing Familiarity
ATS Optimization Keywords
Hard Skills
- Basic Networking Concepts
- Operating Systems Knowledge
- Programming Logic
- Databases Knowledge
- Windows Operating System
- Linux Operating System
- REST APIs Knowledge
- Version Control with Git
- Task Automation
- ITIL Familiarity
Soft Skills
- Analytical Ability
- Team Collaboration
- Willingness to Learn
Industry Keywords
- SRE
- DevOps
- Post-Mortem Analysis
- Operational Efficiency
- Incident Management
Tools & Technologies
- Microsoft 365
- AWS
- Azure
- GCP
- PowerShell