Análisis en tiempo real — CULTIVA IA Office

Diagnóstico — Salud de Interfaces de Red

Cisco Catalyst 9200 · FortiGate 100F · Servidor IA GPU · NAS Synology DS923+
Ventana: Jue 19 Jun 2026 · T0: 09:12h → T1: 09:22h (10 min)
⚠ PROBLEMA ACTIVO — 2 interfaces críticas
CRC/errores activos
847
+847 en 10 min (GE1/0/5)
Output drops
1.243
+1.243 acumulados (GE1/0/9)
Interfaces en alerta
2/4
GE1/0/5 · GE1/0/9
Uplink WAN
OK
FortiGate GE1/0/1 limpio
Causa raíz estimada
Capa física
Cable + congestión uplink NAS
🔌 Estado de Interfaces — Contadores T0 → T1
Interfaz Estado Velocidad / Dúplex CRC errors Δ Input errors Δ Output drops Δ Runts Δ Colisiones Δ Utilización Severidad
GigabitEthernet1/0/1
Uplink → FortiGate 100F (WAN)
● Up / Up 1G / Full 0 0 0 0 0
38%
OK
GigabitEthernet1/0/5
NAS Synology DS923+
● Up / Up 1G / Full ⚠ revisar 847 ↑+847 1.092 ↑+1.092 12 38 ↑+38 0
72%
CRÍTICO
GigabitEthernet1/0/9
Servidor IA GPU Workstation
● Up / Up 1G / Full 2 4 1.243 ↑+312/min 0 0
94%
CRÍTICO
eth0
Servidor GPU — NIC Linux (ifInErrors)
● Up 1G / Full ✓ auto-neg 14 ↑+14 6 0 0 0
91%
AVISO
🔴 GE1/0/5 — 847 CRC en 10 min (NAS Synology)
Problema de capa física confirmado Los CRC incrementan activamente (84,7/min en promedio, pico 120/min). La señal recibida en GE1/0/5 llega corrupta. El NAS Synology no reporta errores en su NIC: el problema está en el cableado o en el SFP del lado switch.
📏
38 runts detectados — posible dúplex mismatch histórico Los runts (<64 bytes) apuntan a colisiones anteriores. Confirmar que la NIC del Synology no quedó configurada en half-duplex tras un reinicio de firmware.
Contadores GE1/0/5
CRC errors
T0: 12.450 T1: 13.297 +847
Input errors
T0: 15.680 T1: 16.772 +1.092
Runts
T0: 2.341 T1: 2.379 +38
Giants
T0: 0 T1: 0 +0
Throttles
T0: 0 T1: 0 +0
🔴 GE1/0/9 — Congestión egress severa (Servidor IA)
📊
94% de utilización sostenida — output drops activos El servidor GPU transfiere modelos de IA (~8–12 GB por sesión de training) saturando el enlace de 1G. Los 1.243 output drops acumulados explican los timeouts reportados. No es un problema de cableado.
🔗
Correlación con eth0: 14 ifInErrors en el servidor El lado Linux del servidor registra 14 errores entrantes. Combinado con la congestión de salida del switch, confirma que el enlace 1G es el cuello de botella. Se recomienda evaluar 10GbE.
Contadores GE1/0/9 + eth0
Output drops (switch)
T0: 0 T1: 1.243 +1.243
Output errors (switch)
T0: 0 T1: 0 +0
ifInErrors (eth0 Linux)
T0: 0 T1: 14 +14
Utilización Tx (pico)
943 Mbps ~94%
CRC errors (GE1/0/9)
T0: 2 T1: 4 +2 (residual)
📋 Log de eventos — Catalyst 9200
Jun 16 07:43:22
%LINEPROTO-5-UPDOWN: GigabitEthernet1/0/5 line protocol changed state to down
Jun 16 07:43:29
%LINEPROTO-5-UPDOWN: GigabitEthernet1/0/5 line protocol changed state to up
Jun 16 07:58:11
%LINEPROTO-5-UPDOWN: GigabitEthernet1/0/5 — flap #2 detectado
Jun 17 09:12:05
%ETHPORT-3-IF_ERRORS_THRESHOLD: GE1/0/5 CRC threshold exceeded (>500/5min)
Jun 17 14:30:00
%QOS-3-OUTPUT_QUEUE_DROP: GE1/0/9 output queue drops rate high (>100/min)
Jun 18 09:12:00
Inicio captura baseline T0 — contadores registrados
Jun 18 09:22:00
Captura T1 completada — análisis disponible
🔵 GE1/0/1 — Uplink FortiGate (OK)
WAN limpio — descartado como causa raíz Ningún error en el enlace FortiGate→Catalyst. Utilización al 38%. El problema no viene del proveedor de Internet. La fibra 1 Gbps simétrica funciona dentro de parámetros normales.
📡
Confirmación: problema confinado a la LAN interna Las videollamadas lentas NO son por falta de ancho de banda WAN. Son consecuencia del congestionamiento en la VLAN de producción causado por las transferencias del servidor GPU saturando el switch interno.
# show interfaces GigabitEthernet1/0/1 GigabitEthernet1/0/1 is up, line protocol is up Hardware is Gigabit Ethernet, addr a4:c3:f0:11:22:33 Description: UPLINK-FortiGate100F-Ge0/1 Full-duplex, 1000Mb/s, media type is 10/100/1000BaseTX -- contadores T1 -- 0 input errors, 0 CRC, 0 frame 0 output errors, 0 collisions 0 output buffer failures, 0 output buffers swapped out
🛠 Plan de Acción — Cambio de Ventana Vie 20 Jun 20:00h
💻 Script de Verificación Post-Cambio
Cisco Catalyst 9200 — After-Maintenance Check
# 1. Limpiar y registrar baseline post-cambio clear counters GigabitEthernet1/0/5 clear counters GigabitEthernet1/0/9 # 2. Verificar velocidad/dúplex tras el cambio show interfaces GigabitEthernet1/0/5 | include duplex|speed|CRC|error show interfaces GigabitEthernet1/0/9 | include duplex|speed|drops # 3. Esperar 15 min, capturar T1 post-cambio show interfaces GigabitEthernet1/0/5 counters errors show interfaces GigabitEthernet1/0/9 counters # 4. Confirmar log limpio show logging | include 1/0/5|1/0/9|changed state|CRC|threshold # Éxito si CRC delta = 0 y output drops = 0 en T1
Servidor Linux (GPU Workstation) — Verificación
# Verificar estado de la NIC ethtool eth0 # → Speed: 1000Mb/s | Duplex: Full | Link: yes # Ver contadores detallados de errores ethtool -S eth0 | grep -iE "error|drop|miss|crc|collision" # Monitorizar utilización en tiempo real ip -s link show eth0 # Alternativa visual: watch -n 2 ip -s link show eth0 # Si los errores persisten, verificar driver NIC ethtool -i eth0 # driver: igb | version: 5.10.0+ | firmware-in-nic: 1.67 # Objetivo: ifInErrors Δ = 0 tras reemplazo de cable