Jump to content

Server 2012 File Server - suddenly stops serving requests (but otherwise looks fine)


Recommended Posts

  • 10 months later...
Posted

Just came across this post. Been happening here for 3 months, prob about once a week.

It's on two file servers one R2 one standard.

Thought it was just my network. Just been pulling DNS thinking it was the cause. Will try some of the fixes here

Love edugeek

Posted

Hi Folks,

 

We're having a similar issue to this. that started around a year ago, but hasn't been consistent.

At first, every Friday the network would fail at 1445. A network card was replaced which, well, didn't do much. It was faulty, but dell initially replaced it with the wrong one.

 

We then spent the best part of a year, troubleshooting with dell. For a while the network wouldn't outright fall over, it performance just wasn't quiet there.

Then recently, its started failing, every Wednesday, like clockwork at 10 past 1

 

Shares all become inaccessible, and most (but not all) of the servers on our cluster become unreachable though remote desktop.

However, we can get into the cluster and physical servers, and from there hyper v into the virtual servers, where they all report...

that there is nothing wrong and are perfectly healthy.

 

From what we can see there are no scheduled tasks running, no backups running at that time, no network ports fail the switches are fine.

 

Our DC's are outside the cluster and work perfectly fine, so users can login, they just can't access their files or desktop.

Other than being unable to access some servers by remote desktop, and that shares are offline, every service other than sims and the print server still work (insight and impero for example still work, even with the servers inaccessible)

 

A simple restart fixes the issue

Does this sound related to the issues you've been facing?

Posted

 

A simple restart fixes the issue

Does this sound related to the issues you've been facing?

 

Yep sounds exactly like my problem! Not happened for about 10 days now

  • 1 month later...
Posted

Not had it happen for a while till yesterday.

The server that it was happening to has been good.

Yesterday it happened to my SIMS server rebooted it which sorted it for 1/2 hour. Spent last night checking again could not find anything.

Fingers crossed for today

Posted

Typical should have kept my mouth shut. The server that was behaving itself is not now. Same thing happened. Quick reboot back to normal.

*bangs head on desk*

  • Thanks 1
  • 3 weeks later...
Posted

Hi Ozdave,

 

Have you had any more luck with this one?

 

I also have this issue also.

 

file server 2012 STD VM running on a 2012 R2 Hyper -V cluster.... Updated everything to death... VM and Cluster nodes... :(

 

Thxs mate.

 

Chris

Posted

Morning Claydoc,

 

Sorry mate. no luck here. Happened here two days ago again. This time it was a server that had never done it before. A quick reboot sorted it.

I just can't shed any light on it. nothing in the event logs, nothing on my VM cluster.

Started to record date / times it was happening and to what server. there is no pattern.

 

Not much help, I'm at a total loss.

 

Cheers

Posted

Over winter, with all the computers turned off, the issue didn't happen.

We were all ready to remote in on both Wednesdays but everything remained up.

 

For the first 2 weeks of the new term it also remained stable, but then in the third week it went down again, albeit after school this time, so at least its less disruptive now.

Wondering if there's something somewhere not on the cluster itself causing an overload or spike in traffic that confuses everything

  • Thanks 1
Posted

Are any of you using QoS for anything on your servers?

 

We had a similar issue last year. Our Exchange, and File Guest Cluster would be running fine and suddenly they'd lose network connectivity on their main production NICs, everything would show as connected but no packets would be sent or received. If we migrated the VMs to another Hyper-V host connectivity would come back for a few days or sometimes hours before it would happen again.

 

We tried lots of things: turning off hardware acceleration on the physical NICs, turning off VMQs, removing and readding the physical/virtual NICs but nothing resolved it.

 

In the end I disabled 'Throttling' on our DPM server (we were restricting backup speeds during the day so it didn't impact network performance) and since then we've not had another issue.

  • Thanks 1
  • 4 months later...
Posted

I seem to have sorted this. I won't bore everyone with the ins and outs, but this has been going on for ages.

 

Ultimately, this is what i've done... installed all available critical updates, plus KB2955164.

Set the Server service to automatic

Increased the max threads per queue (reg add HKEY_LOCAL_MACHINE\SYSTEM\CurrentControlSet\Services\lanmanserver\parameters /v MaxThreadsPerQueue /t REG_DWORD /d 1024 /f)

 

I had also disabled SMB leasing at one point also, but this brings other issues with it (such as Office <2000 docs not opening in Office 2013+ if they have any alternate data streams such as zone identifiers embedded in them (which happens when files are downloaded from the internet for example)), so this was subsequently re-enabled again.

 

All been working for a couple of weeks now with no further errors in my SMB Server operational logs.

Posted

we've had a bit of progress for this too.

We've confirmed our cluster is 100% healthy (many many different people have confirmed this now)

 

Our clusters internal switch had lost its configuration in a power cut.

Not an issue in itself, except it had also had a cable plugged in to our core switch, so it was picking up another vlan and thrashing itself with traffic.

 

we re uploaded its configuration and avoided the loss of connectivity today.

 

we're now going over all our switches with a fine tooth comb for configuration faults

  • 4 weeks later...
Posted
Anyone have any other fixes for this? We've had this issue on and off the past 2 years, we have a ticket with MS but that is always an adventure and they don't get back to use very quickly.
  • 4 months later...
Posted
I have exactly same problem with 2008R2 - restart fixes it. No information in Events - nada. Using vmware. No changes to configuration of switches and if there would be I guess all other servers would have problems just wonder is same issue as 2012
Posted
we've had a bit of progress for this too.

We've confirmed our cluster is 100% healthy (many many different people have confirmed this now)

 

Our clusters internal switch had lost its configuration in a power cut.

Not an issue in itself, except it had also had a cable plugged in to our core switch, so it was picking up another vlan and thrashing itself with traffic.

 

we re uploaded its configuration and avoided the loss of connectivity today.

 

we're now going over all our switches with a fine tooth comb for configuration faults

 

 

This didn't fix it in the end, but we've found out what the error was for us anyway.

Turned out we had a dns replication issue.

Fixed that and it hasn't happened since

 

We noticed completely by accident while we couldn't get to the servers by name we could get in via their IP

 

how the many experts we paid many money to didn't notice this either i have no idea XD

Create an account or sign in to comment

You need to be a member in order to leave a comment

Create an account

Sign up for a new account in our community. It's easy!

Register a new account

Sign in

Already have an account? Sign in here.

Sign In Now



×
×
  • Create New...