From: Coly Li <colyli@suse.de>
To: Kent Overstreet <kent.overstreet@gmail.com>
Cc: linux-bcache@vger.kernel.org, linux-block@vger.kernel.org,
Nix <nix@esperi.org.uk>, Kai Krakow <hurikhan77@gmail.com>,
Eric Wheeler <bcache@lists.ewheeler.net>,
Junhui Tang <tang.junhui@zte.com.cn>,
stable@vger.kernel.org
Subject: Re: [PATCHv2] bcache: option for allow stale data on read failure
Date: Wed, 20 Sep 2017 21:38:59 +0200 [thread overview]
Message-ID: <eaeaf505-92b1-7d45-9b42-64f9d3f05e0b@suse.de> (raw)
In-Reply-To: <20170920160735.jp4riq7x3qc472px@kmo-pixel>
On 2017/9/20 下午6:07, Kent Overstreet wrote:
> On Wed, Sep 20, 2017 at 06:24:33AM +0800, Coly Li wrote:
>> When bcache does read I/Os, for example in writeback or writethrough mode,
>> if a read request on cache device is failed, bcache will try to recovery
>> the request by reading from cached device. If the data on cached device is
>> not synced with cache device, then requester will get a stale data.
>>
>> For critical storage system like database, providing stale data from
>> recovery may result an application level data corruption, which is
>> unacceptible. But for some other situation like multi-media stream cache,
>> continuous service may be more important and it is acceptible to fetch
>> a chunk of stale data.
>>
>> This patch tries to solve the above conflict by adding a sysfs option
>> /sys/block/bcache<idx>/bcache/allow_stale_data_on_failure
>> which is defaultly cleared (to 0) as disabled. Now people can make choices
>> for different situations.
>
> IMO this is just a bug, I'd rather not have an option to keep the buggy
> behaviour. How about this patch:
>
Hi Kent,
OK, last time when I discuss with other bcache developers, people wanted
to keep this behavior, then I modify it as an option in this version
patch. I support fix it without an option, because there are too many
options already. Good to know you have similar decision :-)
> commit 2746f9c1f962288d8c5d7dabe698bf7b3fddd405
> Author: Kent Overstreet <kent.overstreet@gmail.com>
> Date: Wed Sep 20 18:06:37 2017 +0200
>
> bcache: Don't recover from IO errors when reading dirty data
>
> Signed-off-by: Kent Overstreet <kent.overstreet@gmail.com>
>
> diff --git a/drivers/md/bcache/request.c b/drivers/md/bcache/request.c
> index 382397772a..c2d57ef953 100644
> --- a/drivers/md/bcache/request.c
> +++ b/drivers/md/bcache/request.c
> @@ -532,8 +532,10 @@ static int cache_lookup_fn(struct btree_op *op, struct btree *b, struct bkey *k)
>
> PTR_BUCKET(b->c, k, ptr)->prio = INITIAL_PRIO;
>
> - if (KEY_DIRTY(k))
> + if (KEY_DIRTY(k)) {
> s->read_dirty_data = true;
> + s->recoverable = false;
> + }
>
I though of fixing here, the reason I gave up to modify here was,
cache_lookup_fn() is called for keys in leaf nodes (b->level == 0),
bch_btree_map_keys_recurse() needs to do I/O to fetch upper level nodes
before accessing leaf node. When a SSD failed bch_btree_node_get() will
fail before cache_lookup_fn() is executed. So the your patch, there is
no chance to set s->recoverable to false, recovery still happens.
If you don't like an option, the following modification should be much
simpler,
diff --git a/drivers/md/bcache/request.c b/drivers/md/bcache/request.c
index 681b4f12b05a..f397785d9c38 100644
--- a/drivers/md/bcache/request.c
+++ b/drivers/md/bcache/request.c
@@ -697,8 +697,10 @@ static void cached_dev_read_error(struct closure *cl)
{
struct search *s = container_of(cl, struct search, cl);
struct bio *bio = &s->bio.bio;
+ struct cached_dev *dc = container_of(s->d, struct cached_dev, disk);
- if (s->recoverable) {
+ if (s->recoverable &&
+ (dc && !atomic_read(&dc->has_dirty)) {
/* Retry from the backing device: */
trace_bcache_read_retry(s->orig_bio);
This might be the simplest way I know for now.
Thanks.
Coly Li
prev parent reply other threads:[~2017-09-20 19:39 UTC|newest]
Thread overview: 7+ messages / expand[flat|nested] mbox.gz Atom feed top
2017-09-19 22:24 [PATCHv2] bcache: option for allow stale data on read failure Coly Li
2017-09-20 6:59 ` Michael Lyle
2017-09-20 10:28 ` Coly Li
2017-09-20 15:40 ` Michael Lyle
2017-09-20 19:46 ` Coly Li
2017-09-20 16:07 ` Kent Overstreet
2017-09-20 19:38 ` Coly Li [this message]
Reply instructions:
You may reply publicly to this message via plain-text email
using any one of the following methods:
* Save the following mbox file, import it into your mail client,
and reply-to-all from there: mbox
Avoid top-posting and favor interleaved quoting:
https://en.wikipedia.org/wiki/Posting_style#Interleaved_style
* Reply using the --to, --cc, and --in-reply-to
switches of git-send-email(1):
git send-email \
--in-reply-to=eaeaf505-92b1-7d45-9b42-64f9d3f05e0b@suse.de \
--to=colyli@suse.de \
--cc=bcache@lists.ewheeler.net \
--cc=hurikhan77@gmail.com \
--cc=kent.overstreet@gmail.com \
--cc=linux-bcache@vger.kernel.org \
--cc=linux-block@vger.kernel.org \
--cc=nix@esperi.org.uk \
--cc=stable@vger.kernel.org \
--cc=tang.junhui@zte.com.cn \
/path/to/YOUR_REPLY
https://kernel.org/pub/software/scm/git/docs/git-send-email.html
* If your mail client supports setting the In-Reply-To header
via mailto: links, try the mailto: link
Be sure your reply has a Subject: header at the top and a blank line
before the message body.
This is a public inbox, see mirroring instructions
for how to clone and mirror all data and code used for this inbox