Both servers seem completely down. No FTP connections either. What’s the ETA?
I just had a customer call - web15 is still down.
They also said their were problems Tuesday and Wednesday, but usually they were able to log in after a few minutes. Today they said it’s been down “all morning” - I can get the exact time frame, but I’ve been watching it now for about 15 minutes.
It doesn’t seem to be a site issue, since I can’t ftp or even load the basic login.html page.
Right now the control panel even seems to be slow also.
We are already working to fix the issue related Web14 and Web15. We had to reboot both the servers and both are coming up now. Please check the following post for latest update-
It wasn’t down for any time more than some apache resets on web15, web14 was having very high traffic.
We thought that was the main issue, but seems something has started making kernel panics in a bad way now.
I guess ‘all morning’ depends on their timezone is started about 9:07-9:10am here in Central time.
That sounds about right - when this thread started it was just after 9:00am, so it was probably in that time frame (their “all morning”…)
Just an FYI, they did notice similar behavior earlier this week - although for MUCH shorter periods of time - on Tuesday and Wednesday. Maybe something is working up to a complete failure? ;(
From the user’s standpoint the browser either says the site is unavailable after a minute (no error message), or Chrome will report something like, “Oops! Google Chrome could not connect to xyz.com”
That may not literally be server downtime, but it appears that way from the outside.
It also looks to be anything on the server - either plain HTML or SQL based sites, ftp access, etc.
Sorry your rest got cut short! ![]()
I am about 20 min out form booting it on the bench, and testing then a bit more on network to be sure.
For both 14 and 15?
No
other issues came up, I was optimistic, here to get some firmware updates now.
nightmares
Data is all looking good, but hardware failure is never fun. We are getting there!
We ended up changing onboard interfaces and addon NIC interfaces with the hardware swap, the server isn’t taking that too well, we are getting it resolved.
We had it all up no problem when network down, got network up 7 pings, and it kernel panic’d again.
I am speechless, all this work
(and I don’t mind the work, it is the down that is killing me!)
Just so you don’t feel alone in your suffering, the downtime is killing all of us too. ![]()
This isn’t supposed to happen like this. the technology we chose and hardware we picked supposed to make this easier…:rolleyes:
Anyway, bad things happen, but its growing and learning for new.
Hate to say it, but lately this datacenter move and other “upgrades” appear to be more trouble than anything. I know you guys are working your butts off trying to get it working and keep it that way, and I appreciate the work you put in. It’s just hard explaining to clients when something like this happens.
Is there a realistic ETA for this to be all back together and available online?
I say ‘realistic’ because telling clients it will be fixed in an hour and then taking 2 or more makes it seem like 2 failures, not one. If I say 5 and it’s up in 3, it feels like an improvement (in a twisted kind of way…)
I understand what you mean, fully. my heart was like a yoyo today with success and failure.
In the end I went to a store and got a new motherboard, RAM and CPU, all AMD.
We’ve had great success with the AMD, we’ll know a bit more shortly ![]()
It wasn’t what I WANTED, or what they said was in stock, but at the point I was, I needed something.
We have more stock than ever here, but these two servers need a bit more oomph than recycling some of the old hardware I brought in from Miami 1.5 weeks ago.
A dual quad xeon 1.6 would not cut it. I had used other stock earlier in week, and have loads on order bad timing of it all really caught us, and has me thinking of new solutions ![]()
I think you missed the ETA part ![]()
No, check the status forum
We are almost there
actually giving ETA at this point, if it failed…I don’t even know what I’d do.
I have help today here, but this has been a mess.
Whew, I am going to head back to the office (aka surface for air)