FREE ACCESS
5,000–10,000 jobs/day
See all jobs on JobTailor
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates expertise in monitoring operations and incident management within critical environments, utilizing tools for observability and compliance with IT service governance best practices. Proficient in technical documentation and communication for effective incident reporting and resolution.
Highest-signal resume keywords
Monitoring Operations (NOC)Incident ManagementObservability Tools (Dynatrace, Zabbix, Grafana)Cloud Computing (Azure, Microsoft 365, AWS)ITIL Best Practices
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
Incident AnalysisRoot Cause Analysis (RCA)Disaster Recovery (DR)High Availability EnvironmentsNetwork Infrastructure (VPN, DNS, DHCP, TCP/IP)Web Servers (IIS, Apache)Microsoft Windows ServerLinux Operating SystemsTechnical EnglishContinuous Improvement
Soft Skills
Strong Communication Skills
Tools & Technologies
Microsoft TeamsMonitoring Tools (Dynatrace, Zabbix, Grafana)
Industry Keywords
IT Service GovernanceService Level Agreements (SLA)Incident ResponseEscalation MatricesWar Room Management
Tech Stack
Tools & technologiesApacheAWSAzureCloudDNSGrafanaLinuxTCP/IP
About the role
Key responsibilities & impact- Perform proactive monitoring of critical environments, identifying and managing events and incidents;
- Ensure compliance with established SLAs for ticket response and resolution;
- Perform initial incident analysis, executing mapped workaround routines and/or contacting responsible teams when necessary;
- Monitor and track the performance, availability and capacity of environments;
- Maintain and update escalation matrices and contact lists for technical teams, managers, vendors and partners, ensuring correct notification of responsible parties during incidents, outages and war rooms;
- Lead and participate in war rooms, recording evidence and relevant information for later analysis;
- Prepare minutes, reports and documentation related to critical and emergency incidents;
- Create rooms for root cause analysis (RCA) activities and Post Mortem processes;
- Open, follow up and escalate tickets with vendors and manufacturers;
- Act in emergency and critical situations to restore services within agreed service levels;
- Identify opportunities for continuous improvement in monitoring and operations.
Requirements
What you’ll need- Bachelor's degree completed or in progress in Information Technology, Computer Networks, Information Systems or related fields;
- Experience in monitoring operations (NOC) and supporting critical environments;
- Knowledge of monitoring and observability tools such as Dynatrace, Zabbix and Grafana;
- Knowledge of Microsoft Teams (creating Teams rooms);
- Knowledge of Microsoft Windows Server and Linux operating systems;
- Knowledge of Cloud Computing environments, preferably Azure, Microsoft 365, AWS;
- Knowledge of network infrastructure (VPN, communication links, DNS, DHCP and TCP/IP);
- Knowledge of web servers (IIS and Apache);
- Knowledge of high availability environments, Disaster Recovery (DR) and clustering;
- Strong communication skills for managing, reporting and following up incidents in complex environments;
- Knowledge of IT service governance and management best practices (ITIL);
- Technical English for reading and interpreting technical documentation.
Benefits
Comp & perks- Position also open to candidates with disabilities (PwD)
