localzuk Posted April 17, 2013 Posted April 17, 2013 We've got a HP Gen8 DL380e, 25 USFF version at school which we use as our storage device. Today, suddenly, it decided to start resetting itself - just cuts off and starts up. I've gone through and updated firmware, BIOS, drivers, Windows etc... But it still happens. I've also looked in the HP diagnostics stuff and nothing is appearing in there as faulty either. Anyone got any ideas what it could be? My first thought would be the CPU - as its the only part I can think of which isn't in some way redundant - it only has a single CPU, but everything else is doubled (2 NICs, multiple RAM modules in multiple channels, 2 CPUs, multiple SAS controllers etc...). I'm going to be doing a repair on the Windows install tomorrow morning to see if it makes any difference but it is most likely that it'll be a 'call HP' job. But usually when I call them, I have something which can tell me what's wrong with it - hence the question. Anyone got any thoughts?
victory2012 Posted April 17, 2013 Posted April 17, 2013 Had this with a G7 before ended up being a new motherboard and raid controller replaced by HP.. they could never actually tell me what was actually wrong tho
localzuk Posted April 17, 2013 Author Posted April 17, 2013 This particular unit has already had a new 25 bay module and replacement SAS cables... Not been a great unit. The other 2 DL380p's we have seem ok so far though!
FN-GM Posted April 17, 2013 Posted April 17, 2013 Sounds odd, but try disconnecting the monitor. We had a 380 do very similar to this. Turned out to be a dodgy VGA port on the MB.
localzuk Posted April 17, 2013 Author Posted April 17, 2013 Was crashing without the monitor plugged in (we don't have them hooked up to monitor etc... normally), and with it plugged in.
Lftek55 Posted April 17, 2013 Posted April 17, 2013 Have you thought about overheating, could be dust in the cooling fins
FN-GM Posted April 17, 2013 Posted April 17, 2013 Was crashing without the monitor plugged in (we don't have them hooked up to monitor etc... normally), and with it plugged in. Ah worth a try. We have ours plugged into a KVM.
localzuk Posted April 17, 2013 Author Posted April 17, 2013 Have you thought about overheating, could be dust in the cooling fins Room is at 23 degrees, on boot, the server reports ambient temp as 9 degrees. Plus the HP diag stuff should be telling me about it if that was happening.
john Posted April 17, 2013 Posted April 17, 2013 Has it got the right sized PSUs in it for the load on them? I have seen that before that they don't warn they just reset
localzuk Posted April 17, 2013 Author Posted April 17, 2013 Hmm... Might have more of a look, as they do 1200W ones too.
localzuk Posted April 17, 2013 Author Posted April 17, 2013 Just checked, and if all the drives were hard disks, we'd be needing 415W at 100% utilisation, but we have 16x SSDs so the consumption would be lower still, so the 750Ws are well specced.
localzuk Posted April 18, 2013 Author Posted April 18, 2013 Well, the server has now been stable for 26 hours. I spent 6 hours updating firmware and drivers, running diagnostics, reseating every cable etc... It stopped resetting once I did a firmware update for the SAS backplane, of all things.
localzuk Posted April 20, 2013 Author Posted April 20, 2013 I spoke too soon. It has reset itself a couple of times overnight now. So, I've reported it to HP.
FN-GM Posted April 21, 2013 Posted April 21, 2013 I spoke too soon. It has reset itself a couple of times overnight now. So, I've reported it to HP. How did that go?
localzuk Posted April 22, 2013 Author Posted April 22, 2013 Still ongoing. Sending them bunches of logs...
localzuk Posted April 24, 2013 Author Posted April 24, 2013 (edited) Sadly, the saga continues! I've sent about 4 tonnes of logs to HP now, all of them saying there's no hardware error... However, last night I left it doing diagnostics via the Intelligent Provisioning onboard tool, and it reset during this - meaning: a) I didn't get the logs HP were asking for and b) It isn't something that Windows is causing. The ball is once again in HP's court! In my ILO Event Log, I have dozens of 'Server Reset', mixed with the occasional 'Server power restored' followed shortly by 'Server power removed'... So they make no sense either. Edited April 24, 2013 by localzuk
localzuk Posted April 24, 2013 Author Posted April 24, 2013 Well, they've agreed to replace the system board now. They're coming tomorrow to do it. It'll be nice having it back up and running again!
ass17 Posted April 24, 2013 Posted April 24, 2013 We had the same issue with a G6 and it seemed to be the onboard raid controller 410i. They replaced the main board next day and all is well. We initially put in two new disks after the controller screwed up both disks in bay 1 of a raid 10 config. The new disks auto configured to raid 5 and we restored everything back via backups. Took us a whole day!
localzuk Posted April 25, 2013 Author Posted April 25, 2013 If it is the onboard controller, that's fine - we don't actually use it! In order to use the 25 bay disk enclosure, we have an additional controller plugged into PCI Express.
Recommended Posts
Create an account or sign in to comment
You need to be a member in order to leave a comment
Create an account
Sign up for a new account in our community. It's easy!
Register a new accountSign in
Already have an account? Sign in here.
Sign In Now