quote: Marco Giustini
it does not work like that but just to give you an idea
quote: Monte L FullmerNo modern computer system needs to be unplugged to "clear stale memory" and it also won't fix any degraded or even failed RAID.
Pull/disconnect the two power plugs from time to time to clear out stale memory.
quote: Marco GiustiniThat might be the solution if your Media Block is causing the troubles, but here it's clearly a RAID related issue.
On a Dolby server the cat 862 is powered from the main PSU, rebooting the server by software won't fully reset it. A full shutdown every now and then is recommended.
quote: Marco GiustiniWhy not? The whole idea of modern hard drive design is that sectors that become bad are being relocated. Not one drive that rolls of the production line is entirely flawless. So what does it matter if a drive has two or three relocated sectors right from the start if they do not increase over time? This drive is operating perfectly fine, any replacement has no guarantee to be any better.
I may agree with your statement, but if your hard drive is 2 weeks old I can't see how a few bad sectors could be acceptable.
quote: Marco GiustiniIt is unreasonable to expect/demand a replacement from any server manufacturer on a new drive with a few bad sectors. As long as they aren't suddenly increasing noticeably, there is no problem.
Anyway, I would not allow a single bad sector on a brand new drive. That's what I'd do. Those machines - and those drives - are too darn expensive for me to be relaxed on bad sectors
quote: Scott NorwoodThat can also happen if there is buggy firmware present on your disk. Another problem are disks that aren't designed for RAID purposes. Those disks often take too long to respond to a request, especially if an automated relocation is happening, that causes the RAID controller to drop the disk from the array. While they still may operate quite stable in a software-RAID environment, they usually fail rather rapidly in a hardware RAID situation.
Sometimes good disks will be reported as "failed" by the RAID controller, yet the array can safely be rebuilt without replacing them. If this happens more than once, though, it is another sign of impending failure.
quote: Marco GiustiniThat would be necessary now, yes, but of course should be done by the servers - detect a sudden increase in bad sectors and report automatically. Doremis recent reporting is a clear improvement. At some time, it may also consider reporting potentially dangerous RAID conditions. You could, however, also request these from the logs or via SNMP automatically. The reason it is not done now is that a single drive failure is already taken care of by the RAID redundancy, which then SHOULD trigger an alert.
It means to track how the bad sectors number behaves on thousands of HDDs.
quote: Marco GiustiniIt was more a general observation, I would not advice anybody to mess with an existing configuration, if that would be even possible. Since most current servers only carry 3 or 4 disks, it would even be quite useless.
Again, you can do that on your PC or on servers you build. I would never run anything on a D-Cinema server which does not come from the manufacturer.
It has been 788 days since the last post.
quote:Wait, so are you saying you shut down the server when not in use? The part I put in bold seems to suggest so.
Anyway, cleaning the contacts and reseating the drives appears to have worked. The RAID card gave a single bleep on the reboot - a hopeful sign - and it's been about an hour since then without any trouble. Furthermore it's been ingesting throughout that hour. No freezes or double bleeps so far. My guess is that repeated mechanical (ramp up and down) and/or heat cycling caused one of them to work loose, hence the RAID card's temper tantrum.