All of lore.kernel.org
 help / color / mirror / Atom feed
From: Jon Hardcastle <jd_hardcastle@yahoo.com>
To: Jon@eHardcastle.com, Neil Brown <neilb@suse.de>
Cc: linux-raid@vger.kernel.org
Subject: Re:  sdc1 does not have a valid v0.90 superblock, not importing!
Date: Wed, 11 Aug 2010 08:30:31 -0700 (PDT)	[thread overview]
Message-ID: <282617.15518.qm@web51308.mail.re2.yahoo.com> (raw)
In-Reply-To: <20100811213405.09a0b709@notabene>

--- On Wed, 11/8/10, Neil Brown <neilb@suse.de> wrote:

> From: Neil Brown <neilb@suse.de>
> Subject: Re:  sdc1 does not have a valid v0.90 superblock, not importing!
> To: Jon@eHardcastle.com
> Cc: jd_hardcastle@yahoo.com, linux-raid@vger.kernel.org
> Date: Wednesday, 11 August, 2010, 12:34
> On Wed, 11 Aug 2010 04:19:07 -0700
> (PDT)
> Jon Hardcastle <jd_hardcastle@yahoo.com>
> wrote:
> 
> > 
> > --- On Wed, 11/8/10, Neil Brown <neilb@suse.de>
> wrote:
> > 
> > > From: Neil Brown <neilb@suse.de>
> > > Subject: Re:  sdc1 does not have a valid
> v0.90 superblock, not importing!
> > > To: Jon@eHardcastle.com
> > > Cc: jd_hardcastle@yahoo.com,
> linux-raid@vger.kernel.org
> > > Date: Wednesday, 11 August, 2010, 12:06
> > > On Wed, 11 Aug 2010 02:55:44 -0700
> > > (PDT)
> > > Jon Hardcastle <jd_hardcastle@yahoo.com>
> > > wrote:
> > > 
> > > > (my first attempt appears to have been
> bounced as the
> > > spam checker thought it had HTML in it?!)
> > > 
> > > odd... came through ok for me the first time.
> > > 
> > > > 
> > > > Help!
> > > > 
> > > > Long story short - I was watching a movie
> off my RAID6
> > > array. Got a smart error warning
> > > 
> > > > Aug 10 22:00:07 mangalore kernel: raid5:
> cannot start
> > > dirty degraded array for md4
> > > 
> > > This is the current problem.  The array is dirty
> and
> > > degraded so there could
> > > theoretically be undetectable corruption. 
> Chance is
> > > quite low but it is
> > > there so md won't start with out you
> acknowledging the risk
> > > by giving the
> > > --force flag to mdadm --assemble.
> > > Only do that if you are confident that your
> hardware is
> > > working correctly.
> > 
> > Well I am reasonable sure the controller came adrift
> the first time.. when i reseated it i stopped getting 100's
> of errors.. and it has survived 1.5 badblocks checks. It is
> being held in place by one of those bars you press down
> (does all the expansion cards in 1 go) except i dont think
> it is very good. I will screw it down.
> > 
> > > 
> > > > It appears sdc has an invalid superblock?
> > > > 
> > > > This is the 'examine' from sdc1 (note the
> checksum)
> > > > 
> > > > /dev/sdc1:
> > > .....
> > > >       Checksum : b335b4e3 -
> > > expected b735b4e3
> > > 
> > > Single bit error.  That isn't good as it means
> some
> > > bit of memory or some bit
> > > on some bus somewhere cannot be trusted.
> > > It could be a transient thing and will never
> happen
> > > again.  Or maybe not.
> > > Given the smart errors and the fact that you have
> had
> > > problems with the drive
> > > before it seem very likely that the problem is in
> that
> > > drive.  I suggest
> > > unplugging it and leaving it unplugged.  Some
> memory
> > > buffer in the drive is
> > > probably marginal.  I don't think they use ECC
> > > memory.
> > 
> > Could this be a result of me forcing a power off when
> the drive was causing problems?
> 
> Probably not.  Forcing a power off may well have left
> the array 'dirty' so
> that it wouldn't assemble, but is fairly unlikely to
> corrupt data within a
> block.
> 
> > 
> > What are the dangers to removing it, zeroing the
> superblock and readding? is it MORE dangerous than leaving a
> raid 6 degraded for a few days?
> 
> In general, I would say the chance of a known-bad drive
> causing problems is
> greater than the chance of a fewer known-good drives
> causing problems.
> But then you seem to think it isn't the drive, it was the
> controller and that
> is fixed...
> 
> This is really about your level of trust in the hardware.
> If you trust sdc as much as the others, include it in the
> array.
> If you don't, then don't.
> 
> NeilBrown
> 
> 
> 
> > 
> > > 
> > > > 
> > > > Anyways... I am ASSUMING mdadm has not
> assembled the
> > > array to be on the safe side? i have not done
> anything.. no
> > > force... no assume clean.. I wanted to be sure?
> > > 
> > > You assume correctly.
> > > 
> > > > 
> > > > Should i remove sdc1 from the array? It
> should then
> > > assemble? I have 2 spare drives that I am getting
> around to
> > > using to replace this drive and the other 500GB..
> so should
> > > I remove sdc1... and try and re-add or just put
> the new
> > > drive in?
> > > > 
> > > > atm I have 'stop'ped the array and got
> badblocks
> > > running....
> > > > 
> > > 
> > > Remove sdc and assemble the array with --force,
> and get a
> > > new device to
> > > replace /dev/sdc as soon as possible.
> > 
> > Thanks Neil - I panic'd as previously it has mounted
> the array in a degraded state... but previously the drive
> has disappeared completely... whereas in this case it is
> present... but wrong!
> > 
> > > 
> > > NeilBrown
> > > --
> > > To unsubscribe from this list: send the line
> "unsubscribe
> > > linux-raid" in
> > > the body of a message to majordomo@vger.kernel.org
> > > More majordomo info at  http://vger.kernel.org/majordomo-info.html
> > > 
> > 
> > 
> >       

For the benefit of those that follow! I assemble the array by specifying exactly the drives I wanted in it.. once assembled i could confidently zero-block the troublesome drive... without it rearing its ugly head again!

I now have a 1 drive down - degraded raid 6 array.


      
--
To unsubscribe from this list: send the line "unsubscribe linux-raid" in
the body of a message to majordomo@vger.kernel.org
More majordomo info at  http://vger.kernel.org/majordomo-info.html

  parent reply	other threads:[~2010-08-11 15:30 UTC|newest]

Thread overview: 7+ messages / expand[flat|nested]  mbox.gz  Atom feed  top
2010-08-11  9:55 Sorry if Spamming! - sdc1 does not have a valid v0.90 superblock, not importing! Jon Hardcastle
2010-08-11 11:06 ` Neil Brown
2010-08-11 11:19   ` Jon Hardcastle
2010-08-11 11:34     ` Neil Brown
2010-08-11 12:29       ` Jon Hardcastle
2010-08-11 15:30       ` Jon Hardcastle [this message]
  -- strict thread matches above, loose matches on Subject: below --
2010-08-10 21:35 Fw: " Jon Hardcastle
2010-08-11 22:01 ` Stefan /*St0fF*/ Hübner
2010-08-11 22:56   ` Neil Brown

Reply instructions:

You may reply publicly to this message via plain-text email
using any one of the following methods:

* Save the following mbox file, import it into your mail client,
  and reply-to-all from there: mbox

  Avoid top-posting and favor interleaved quoting:
  https://en.wikipedia.org/wiki/Posting_style#Interleaved_style

* Reply using the --to, --cc, and --in-reply-to
  switches of git-send-email(1):

  git send-email \
    --in-reply-to=282617.15518.qm@web51308.mail.re2.yahoo.com \
    --to=jd_hardcastle@yahoo.com \
    --cc=Jon@eHardcastle.com \
    --cc=linux-raid@vger.kernel.org \
    --cc=neilb@suse.de \
    /path/to/YOUR_REPLY

  https://kernel.org/pub/software/scm/git/docs/git-send-email.html

* If your mail client supports setting the In-Reply-To header
  via mailto: links, try the mailto: link
Be sure your reply has a Subject: header at the top and a blank line before the message body.
This is an external index of several public inboxes,
see mirroring instructions on how to clone and mirror
all data and code used by this external index.