mirror of https://lore.kernel.org/lkml/
 help / color / mirror / Atom feed
* contact problem? (was Re: Spurious Filesystem corruption with ext3 +  large (<400GB) hw RAID ...)
@ 2009-02-13 16:29 Manfred Wassmann
  2009-02-14 17:40 ` Ray Lee
  0 siblings, 1 reply; 3+ messages in thread
From: Manfred Wassmann @ 2009-02-13 16:29 UTC (permalink / raw)
  To: linux-kernel

On Fri, Feb 6, 2009 at 7:50 PM, Ray Lee <ray-lk@madrabbit.org> wrote:

> Huh. It may be related to whatever kernel version Ubuntu uses to boot
> up the install media. Try a different version of Ubuntu (older, newer
> if a newer one exists) to see if it has the same problem.

Thank you for your notice but as I mentioned the problem occurred
first with my custom built kernel and the Ubuntu installation disk was
used only to reproduce the error -- Ubuntu is not what I use anyhow,
I'm on Debian since I switched from Slackware in 1995 ;-)

But now it looks like we had an exotic hardware problem here. Just
before the weekend the RAID controller reported a degraded array and
switched to the hotspare disk. On Monday the original array was
recreated and everything worked fine until Tuesday morning when the
array again was degraded. But then since the disk was checked and
reinserted into the array all problems are gone :-\

Is it possible that such a problem is caused by bad contact within the
drive bay connector?

The reason why it occurred with ext3 only might then be that ext3
stores it's backup superblocks at addresses which are otherwise unused
by a largely empty filesystem.

regards Manfred
-- 
Unix "Birthday" on 2009-02-13 23:31:30 UTC
...it's 1234567890 seconds since the epoch.

^ permalink raw reply	[flat|nested] 3+ messages in thread

* Re: contact problem? (was Re: Spurious Filesystem corruption with  ext3 + large (<400GB) hw RAID ...)
  2009-02-13 16:29 contact problem? (was Re: Spurious Filesystem corruption with ext3 + large (<400GB) hw RAID ...) Manfred Wassmann
@ 2009-02-14 17:40 ` Ray Lee
  2009-02-28 19:32   ` Manfred Wassmann
  0 siblings, 1 reply; 3+ messages in thread
From: Ray Lee @ 2009-02-14 17:40 UTC (permalink / raw)
  To: Manfred Wassmann; +Cc: linux-kernel

On Fri, Feb 13, 2009 at 8:29 AM, Manfred Wassmann
<tux.wassmann@googlemail.com> wrote:
> On Fri, Feb 6, 2009 at 7:50 PM, Ray Lee <ray-lk@madrabbit.org> wrote:
>
>> Huh. It may be related to whatever kernel version Ubuntu uses to boot
>> up the install media. Try a different version of Ubuntu (older, newer
>> if a newer one exists) to see if it has the same problem.
>
> Thank you for your notice but as I mentioned the problem occurred
> first with my custom built kernel and the Ubuntu installation disk was
> used only to reproduce the error -- Ubuntu is not what I use anyhow,
> I'm on Debian since I switched from Slackware in 1995 ;-)
>
> But now it looks like we had an exotic hardware problem here. Just
> before the weekend the RAID controller reported a degraded array and
> switched to the hotspare disk. On Monday the original array was
> recreated and everything worked fine until Tuesday morning when the
> array again was degraded. But then since the disk was checked and
> reinserted into the array all problems are gone :-\
>
> Is it possible that such a problem is caused by bad contact within the
> drive bay connector?
>
> The reason why it occurred with ext3 only might then be that ext3
> stores it's backup superblocks at addresses which are otherwise unused
> by a largely empty filesystem.

(Ah, I missed this message first time around. Please always do a
reply-to-all for lkml.)

Bad contacts, sure, or (more likely) a marginal power supply? Can you
try a different enclosure, or a different power supply for the
enclosure?

^ permalink raw reply	[flat|nested] 3+ messages in thread

* Re: contact problem? (was Re: Spurious Filesystem corruption with  ext3 + large (<400GB) hw RAID ...)
  2009-02-14 17:40 ` Ray Lee
@ 2009-02-28 19:32   ` Manfred Wassmann
  0 siblings, 0 replies; 3+ messages in thread
From: Manfred Wassmann @ 2009-02-28 19:32 UTC (permalink / raw)
  To: Ray Lee; +Cc: linux-kernel

On Sat, Feb 14, 2009 at 18:40, Ray Lee <madrabbit@gmail.com> wrote:
[...]
> (Ah, I missed this message first time around. Please always do a
> reply-to-all for lkml.)

Thanks, that's always a problem with mailing lists, you never know,
some people want CCs and some -- like me -- don't ;-)

> Bad contacts, sure, or (more likely) a marginal power supply? Can you
> try a different enclosure, or a different power supply for the
> enclosure?

Now it looks like the disk is the problem, this week the array again
reported degrades state with the same disk as before being flagged as
defective but some time later everything was OK. The disk now has been
replaced.
Live was easier when disks were either good or bad ;-)

^ permalink raw reply	[flat|nested] 3+ messages in thread

end of thread, other threads:[~2009-02-28 19:32 UTC | newest]

Thread overview: 3+ messages (download: mbox.gz / follow: Atom feed)
-- links below jump to the message on this page --
2009-02-13 16:29 contact problem? (was Re: Spurious Filesystem corruption with ext3 + large (<400GB) hw RAID ...) Manfred Wassmann
2009-02-14 17:40 ` Ray Lee
2009-02-28 19:32   ` Manfred Wassmann

This is a public inbox, see mirroring instructions
for how to clone and mirror all data and code used for this inbox

all inboxes | Powered by JetHome®