From: Mike Snitzer <snitzer@redhat.com>
To: Mikulas Patocka <mpatocka@redhat.com>
Cc: axboe@kernel.dk, dm-devel@redhat.com
Subject: Re: dm: add memory barrier before waitqueue_active
Date: Tue, 5 Feb 2019 12:19:01 -0500 [thread overview]
Message-ID: <20190205171901.GA12214@redhat.com> (raw)
In-Reply-To: <alpine.LRH.2.02.1902050506470.12231@file01.intranet.prod.int.rdu2.redhat.com>
On Tue, Feb 05 2019 at 5:09am -0500,
Mikulas Patocka <mpatocka@redhat.com> wrote:
> Hi
>
> Please submit patch this to Linus before 5.0 is released.
>
> Mikulas
>
>
>
> waitqueue_active without preceding barrier is unsafe, see the comment
> before waitqueue_active definition in include/linux/wait.h.
>
> This patch changes it to wq_has_sleeper.
>
> Signed-off-by: Mikulas Patocka <mpatocka@redhat.com>
> Fixes: 6f75723190d8 ("dm: remove the pending IO accounting")
>
> ---
> drivers/md/dm.c | 2 +-
> 1 file changed, 1 insertion(+), 1 deletion(-)
>
> Index: linux-2.6/drivers/md/dm.c
> ===================================================================
> --- linux-2.6.orig/drivers/md/dm.c 2019-02-04 20:18:03.000000000 +0100
> +++ linux-2.6/drivers/md/dm.c 2019-02-04 20:18:03.000000000 +0100
> @@ -699,7 +699,7 @@ static void end_io_acct(struct dm_io *io
> true, duration, &io->stats_aux);
>
> /* nudge anyone waiting on suspend queue */
> - if (unlikely(waitqueue_active(&md->wait)))
> + if (unlikely(wq_has_sleeper(&md->wait)))
> wake_up(&md->wait);
> }
>
This could be applicable to dm-rq.c:rq_completed() too...
but I'm not following where you think we benefit from adding the
smp_mb() to end_io_acct() please be explicit about your concern.
Focusing on bio-based DM, your concern is end_io_acct()'s wake_up() will
race with its, or some other cpus', preceding generic_end_io_acct()
percpu accounting?
- and so dm_wait_for_completion()'s !md_in_flight() condition will
incorrectly determine there is outstanding IO due to end_io_acct()'s
missing smp_mb()?
- SO dm_wait_for_completion() would go back to top its loop and may
never get woken up again via wake_up(&md->wait)?
The thing is in both callers that use this pattern:
if (unlikely(waitqueue_active(&md->wait)))
wake_up(&md->wait);
the condition (namely IO accounting) will have already been updated via
generic_end_io_acct() (in terms of part_dec_in_flight() percpu updates).
So to me, using smp_mb() here is fairly pointless when you consider the
condition that the waiter (dm_wait_for_completion) will be using is
_not_ the byproduct of a single store.
Again, for bio-based DM, block core is performing atomic percpu updates
across N cpus. And the dm_wait_for_completion() waiter is doing percpu
totalling via md_in_flight_bios().
Could be there is still an issue here.. but I'm not quite seeing it.
Cc'ing Jens to get his thoughts.
Thanks,
Mike
next prev parent reply other threads:[~2019-02-05 17:19 UTC|newest]
Thread overview: 6+ messages / expand[flat|nested] mbox.gz Atom feed top
2019-02-05 10:09 [PATCH] dm: add memory barrier before waitqueue_active Mikulas Patocka
2019-02-05 17:19 ` Mike Snitzer [this message]
2019-02-05 17:56 ` Mikulas Patocka
2019-02-05 19:05 ` Mike Snitzer
2019-02-05 19:29 ` Mikulas Patocka
2019-02-05 19:58 ` Mike Snitzer
Reply instructions:
You may reply publicly to this message via plain-text email
using any one of the following methods:
* Save the following mbox file, import it into your mail client,
and reply-to-all from there: mbox
Avoid top-posting and favor interleaved quoting:
https://en.wikipedia.org/wiki/Posting_style#Interleaved_style
* Reply using the --to, --cc, and --in-reply-to
switches of git-send-email(1):
git send-email \
--in-reply-to=20190205171901.GA12214@redhat.com \
--to=snitzer@redhat.com \
--cc=axboe@kernel.dk \
--cc=dm-devel@redhat.com \
--cc=mpatocka@redhat.com \
/path/to/YOUR_REPLY
https://kernel.org/pub/software/scm/git/docs/git-send-email.html
* If your mail client supports setting the In-Reply-To header
via mailto: links, try the mailto: link
Be sure your reply has a Subject: header at the top and a blank line
before the message body.
This is a public inbox, see mirroring instructions
for how to clone and mirror all data and code used for this inbox