Monitoring Analyst – Mid-level

Posted 13hrs ago

Employment Information

Industry
Education
Salary
Experience
Job Type

Report this job

Job expired or something wrong with this job?

Job Description

Analista NOC monitorando redes de telecomunicações para provedores, operadoras e empresas. Investigando incidentes e evoluindo sistemas de monitoramento com Zabbix e Grafana.

Responsibilities:

  • Monitor availability, performance and capacity of assets, links and network services
  • Analyze alarms, events, dashboards and indicators from monitoring systems
  • Triage and investigate incidents, identifying probable causes, impacts and priorities
  • Perform connectivity, infrastructure and service diagnostics
  • Configure and maintain hosts, templates, items, triggers, macros, discovery rules, actions and maps in Zabbix
  • Create and improve dashboards, panels and alerts in monitoring systems
  • Investigate collection failures via SNMP, ICMP, agents, APIs and other integration methods
  • Review triggers and parameters to reduce false positives and improve alarm accuracy
  • Log and follow up on tickets, ensuring SLA compliance
  • Contact clients, carriers, vendors and technical teams according to escalation workflows
  • Support less experienced analysts in incident analysis and procedures
  • Identify recurring events, trends, risks and opportunities for improvement
  • Document procedures, solutions and standards, ensuring complete shift handovers

Requirements:

  • Hands-on experience with Zabbix in production environments
  • Intermediate knowledge of hosts, templates, items, triggers, macros, discovery, actions and maps
  • Experience with Grafana for analyzing, creating and maintaining dashboards
  • Knowledge of SNMP, ICMP, agents, Syslog and APIs
  • Understanding of TCP/IP, IP addressing, DNS, VLANs, routing, latency, packet loss and availability
  • Ability to interpret metrics, graphs, alerts, trends and capacity indicators
  • Familiarity with ping, traceroute, MTR and SSH/Telnet
  • Basic Linux skills and ability to perform basic server investigations
  • Experience in incident management, technical escalation and SLA compliance
  • Good communication, organization, sense of priority and availability to work shifts/on-call
  • Background in monitoring, infrastructure, technical support or NOC operations, with practical experience in Zabbix and Grafana, preferably in ISPs or telecommunications environments
  • Desirable knowledge: switches, routers, firewalls, servers, links, OLTs and BRAS/BNGs; distributed Zabbix, proxies and collection troubleshooting; databases; alert and notification integrations; Bash or Python; virtualization, containers, ITIL, problem management and root cause analysis
  • Courses or certifications in Zabbix, Grafana, networking, Linux or infrastructure are a plus
  • Availability to work shifts/on-call

Benefits:

  • Continuous development
  • Technical support
  • Support for training and relevant certifications