← cigardev.xyz

Rubrik Triage

Last updated: July 21, 2026 at 10:42 AM ET
๐Ÿ”ด 8 Critical โš ๏ธ 1 Warning ๐Ÿ™‹ 1 Reply ๐Ÿ‘€ 7 Watching

๐Ÿ”ด Critical

Snapshot failure โ€“ Kansas City Day 21
htv001kscrbk001 โ†’ Kansas City
Object: KMBC-MOS (2019) HATMOS SERVER
Snapshot/backup failed again today. Continuing unresolved, three weeks running.
Backup failure โ€“ Fort Smith Day 21 Report only
htv001ftsrbk001 โ†’ Fort Smith
Object: KHOG-COMPUSAT.CompanyNet.org volumes + KHBS-MELOG
Continues, day 21. Today's confirmation came only from the Daily Event Report (05:19 AM) โ€” no real-time alert email fired for the actual backup failure, only the downstream SLA-violation alert (normally noise). Worth keeping an eye on alert delivery for this cluster.
Widespread backup failures โ€“ Louisville Day 12
hts001lourbk001 โ†’ Louisville
30+ objects: Splunk data, BQLIKEM, PBBLD102, DBQLIKEM100A/B, PBQLIKEM110A/B, PCKEYSCAN101, PCNVR101, PBDC74, CPCNETSVC07, CPCSCCMRS104, PBINTSYS114A, C_Drive, plus a dozen+ RBS-unreachable hosts
Genuine RBS-unreachable warnings (not just SLA-violation noise) confirmed again today. Root cause still looks like RBS/host connectivity. 12 days running.
M365 protection failures โ€“ hearstpm Day 12ish
M365/Polaris โ†’ Microsoft 365
Object: directory hearstpm.onmicrosoft.com
Failed scheduled metadata refresh, continuing since 7/10. Not reconfirmed today (1 day silent) โ€” will auto-close if quiet 2 more days.
Network interface down โ€“ Fort Myers Day 11
hts001fmtrbk001 โ†’ Fort Myers
Object: nodes RVMHM187S009415/009418/009048, port rketh3
Interface down on 3 nodes. Not reconfirmed today (1 day silent).
Periodic node task failures โ€“ Milwaukee Day 6
htv001mlwrbk001 โ†’ Milwaukee
Object: nodes RVMHM191S004070/004072/004136/004373
'Failed to disable telemetry data' task failures continue. Node 004070 hit again today.
Health monitor critical โ€“ Westbrook (NtpClockOffsetPeers) Day 2 New task
htv001wbkrbk001 โ†’ Westbrook, ME (WMTW)
Object: nodes RVMHM188S005351/005414/005439
540+ consecutive NtpClockOffsetPeers failures at critical level. Coincides with the open NodeBad support case (01266675) on this cluster.
Backup/upload failures โ€“ Westbrook (multiple objects) Day 2 New task
htv001wbkrbk001 โ†’ Westbrook, ME (WMTW)
WMTW-WSUSR, WMTW-WXMSGR, FileServer Data Files, fsiconfigABX [day 1]; WMTW-VCSA, WMTW-ADIT1, WMTW-FILEMOVER, FLORICAL_SCHEDULES_WPXT [new today]
Correlated snapshot/upload/node-task failures tied to the NodeBad case. WMTW-WSUSR failed again; 4 more objects newly failed overnight.

โš ๏ธ Warnings

RBS unreachable โ€“ Jackson Day 21
htv001jksrbk001 โ†’ Jackson
Object: RBS hosts 10.158.14.101 / 10.158.36.114
RBS timeouts on both hosts continue.

๐Ÿ“‹ Support Cases

P1 Case #01252114 โ€“ Milwaukee (Node is Bad) No reply needed Corroborated
htv001mlwrbk001 โ†’ Milwaukee
Good news: 7/20 update reports all 4 nodes now OK. Rubrik is monitoring for 48โ€“72 hrs before closing.
P2 Case #01256607 โ€“ Albany (Hardware Health Check) Resolved
hts001albrbk001 โ†’ Albany
DIMM correctable errors โ€” resolved and closed out on 7/16. No reply needed.
P1 Case #01266675 โ€“ Westbrook (Node is Bad) Reply needed Corroborated
htv001wbkrbk001 โ†’ Westbrook, ME (WMTW)
Rubrik says the support tunnel is still closed and is asking Scott to re-check/re-enable it on cluster 75f35cbf-c55d-45d8-a8e5-c35426e71e1b. Cluster remains active with today's criticals so still corroborated.

๐Ÿ‘€ Watching (Day 1โ€“2, not yet escalated)

MSSQL log backup failure โ€“ Albuquerque (KOAT-FCDB1) Day 1
htv001abqrbk001 โ†’ Albuquerque (KOAT) โ€” new cluster
Unable to resolve host name KOAT-FCDB1.CompanyNet.Org for FSIAcquisition MSSQL log backup.
Unable to upload snapshot โ€“ Albuquerque (KOAT-MM1) Day 1
htv001abqrbk001 โ†’ Albuquerque (KOAT)
Failed to upload snapshot of KOAT-MM1 to Azure.
Periodic node task failure โ€“ Albuquerque Day 1
htv001abqrbk001 โ†’ Albuquerque (KOAT)
'Failed to disable telemetry data' on node RVMHM189S000385.
System storage usage alert โ€“ Boston Day 1 (new occurrence)
htv001bosrbk001 โ†’ Boston
Cluster storage over 90% full again โ€” previous instance of this alert was closed 7/20.
Unable to upload snapshot โ€“ Milwaukee (WISN-MOS) Day 1
htv001mlwrbk001 โ†’ Milwaukee
Not seen since 7/20.
Periodic node task failures โ€“ Des Moines Day 1
htv001dsmrbk001 โ†’ Des Moines
Not seen since 7/20.
RBS unreachable โ€“ Boston (WCVB-VANTAGE) Day 1
htv001bosrbk001 โ†’ Boston
Not seen since 7/20.

Things 3 Changes

โœ… Created: 2 new tasks (Westbrook health monitor, Westbrook backup/upload failures โ€” both promoted from Day 1 watching to confirmed).
๐Ÿ Closed: none.
๐Ÿงน Pruned from tracking (never had tasks, quiet 4+ days): Milwaukee WISN-FCDB1 MSSQL log, Minneapolis WMUR-ABX1 MSSQL, Minneapolis node 005696, fdbdr-rbkcc-001 fileset errors.