Jump to content

Recommended Posts

Posted

We've got a HP Gen8 DL380e, 25 USFF version at school which we use as our storage device. Today, suddenly, it decided to start resetting itself - just cuts off and starts up.

 

I've gone through and updated firmware, BIOS, drivers, Windows etc... But it still happens.

 

I've also looked in the HP diagnostics stuff and nothing is appearing in there as faulty either.

 

Anyone got any ideas what it could be? My first thought would be the CPU - as its the only part I can think of which isn't in some way redundant - it only has a single CPU, but everything else is doubled (2 NICs, multiple RAM modules in multiple channels, 2 CPUs, multiple SAS controllers etc...).

 

I'm going to be doing a repair on the Windows install tomorrow morning to see if it makes any difference but it is most likely that it'll be a 'call HP' job. But usually when I call them, I have something which can tell me what's wrong with it - hence the question.

 

Anyone got any thoughts?

Posted
This particular unit has already had a new 25 bay module and replacement SAS cables... Not been a great unit. The other 2 DL380p's we have seem ok so far though!
Posted
Sounds odd, but try disconnecting the monitor. We had a 380 do very similar to this. Turned out to be a dodgy VGA port on the MB.
Posted
Was crashing without the monitor plugged in (we don't have them hooked up to monitor etc... normally), and with it plugged in.

 

Ah worth a try. We have ours plugged into a KVM.

Posted
Have you thought about overheating, could be dust in the cooling fins

 

Room is at 23 degrees, on boot, the server reports ambient temp as 9 degrees. Plus the HP diag stuff should be telling me about it if that was happening.

Posted
Has it got the right sized PSUs in it for the load on them? I have seen that before that they don't warn they just reset :(
Posted
Just checked, and if all the drives were hard disks, we'd be needing 415W at 100% utilisation, but we have 16x SSDs so the consumption would be lower still, so the 750Ws are well specced.
Posted

Well, the server has now been stable for 26 hours. I spent 6 hours updating firmware and drivers, running diagnostics, reseating every cable etc...

 

It stopped resetting once I did a firmware update for the SAS backplane, of all things.

Posted (edited)

Sadly, the saga continues!

 

I've sent about 4 tonnes of logs to HP now, all of them saying there's no hardware error...

 

However, last night I left it doing diagnostics via the Intelligent Provisioning onboard tool, and it reset during this - meaning:

 

a) I didn't get the logs HP were asking for and

b) It isn't something that Windows is causing.

 

The ball is once again in HP's court!

 

In my ILO Event Log, I have dozens of 'Server Reset', mixed with the occasional 'Server power restored' followed shortly by 'Server power removed'... So they make no sense either.

Edited by localzuk
Posted

We had the same issue with a G6 and it seemed to be the onboard raid controller 410i.

 

They replaced the main board next day and all is well. We initially put in two new disks after the controller screwed up both disks in bay 1 of a raid 10 config. The new disks auto configured to raid 5 and we restored everything back via backups. Took us a whole day!

Posted
If it is the onboard controller, that's fine - we don't actually use it! In order to use the 25 bay disk enclosure, we have an additional controller plugged into PCI Express.

Create an account or sign in to comment

You need to be a member in order to leave a comment

Create an account

Sign up for a new account in our community. It's easy!

Register a new account

Sign in

Already have an account? Sign in here.

Sign In Now



×
×
  • Create New...