EMC VNX Hardware Failure Troubleshooting

Job ID: 40114370

Budget: $250 – $750 USD

My production-grade EMC VNX array is throwing hardware failure alerts and I need an experienced hand to get it back to a healthy state. The problems are squarely in the physical layer: several disks keep dropping from their RAID groups and one of the storage processors is intermittently going offline. This is disrupting front-end hosts and putting data at risk, so I’m looking for someone who has already wrestled with VNX disk packs, DPEs, and SP failovers and can move quickly from diagnosis to fix.

What I will provide
• Remote access (SSH/Unisphere) to the array and Control Station
• The latest SPCollects and event logs
• Maintenance window details and on-site hands if parts must be reseated or swapped

What I need from you
1. Rapid root-cause analysis of the disk and controller faults, supported by log review and any CLI commands you deem necessary (naviseccli, svc_diagnostics, etc.).
2. A clear action plan to restore redundancy—whether that means drive sparing, rebuilds, or controller replacement—and guidance I can relay to field engineers if hardware swaps are required.
3. Confirmation testing once the remediation steps are complete, ensuring all disks stay online and both storage processors pass health checks without new errors.
4. A brief post-mortem report so I can document the issue, its resolution, and any preventive measures.

You’ll have the green light to use Unisphere, naviseccli, and any other standard EMC/VNX troubleshooting tools. If you’ve solved similar disk-and-controller failures before, especially on VNX5500/5600 series, let me know—speed and confidence matter here.