From: Joseph Qi <joseph.qi@linux.alibaba.com>
To: Bart Van Assche <bvanassche@acm.org>,
Josef Bacik <josef@toxicpanda.com>,
Ming Lei <tom.leiming@gmail.com>, Jens Axboe <axboe@kernel.dk>
Cc: linux-block@vger.kernel.org, nbd@other.debian.org
Subject: Re: [PATCH v2 2/2] nbd: fix NULL pointer dereference in nbd_pending_cmd_work()
Date: Tue, 29 Sep 2026 09:02:44 +0800 [thread overview]
Message-ID: <616f687a-417e-4fd5-b0dc-9bd519db3920@linux.alibaba.com> (raw)
In-Reply-To: <9e43373f-9202-4fa7-b0c8-f1bc80fd81de@acm.org>
Hi Bart,
On 9/28/26 10:19 PM, Bart Van Assche wrote:
> On 9/28/26 2:10 AM, Joseph Qi wrote:
>> diff --git a/drivers/block/nbd.c b/drivers/block/nbd.c
>> index c2c3dbdd631f..9909774dbfbb 100644
>> --- a/drivers/block/nbd.c
>> +++ b/drivers/block/nbd.c
>> @@ -63,6 +63,7 @@ struct nbd_sock {
>> int fallback_index;
>> int cookie;
>> struct work_struct work;
>> + struct request *partial_req;
>> };
>> struct recv_thread_args {
>> @@ -637,12 +638,24 @@ static void nbd_sched_pending_work(struct nbd_device *nbd,
>> {
>> struct request *req = blk_mq_rq_from_pdu(cmd);
>> - /* pending work should be scheduled only once */
>> - WARN_ON_ONCE(test_bit(NBD_CMD_PARTIAL_SEND, &cmd->flags));
>> -
>> nsock->pending = req;
>> nsock->sent = sent;
>> - set_bit(NBD_CMD_PARTIAL_SEND, &cmd->flags);
>> +
>> + /*
>> + * Already armed: this is nbd_pending_cmd_work() re-entering because the
>> + * resumed send was interrupted again. Its work is still running, so
>> + * just refresh the resume point above and let its loop pick it up
>> + * instead of taking another config reference and requeueing the work.
>> + */
>> + if (test_and_set_bit(NBD_CMD_PARTIAL_SEND, &cmd->flags))
>> + return;
>> +
>> + /*
>> + * nbd_mark_nsock_dead() clears ->pending, so the work function cannot
>> + * rely on it to find the request it owns. Keep a copy that only it
>> + * clears.
>> + */
>> + nsock->partial_req = req;
>> refcount_inc(&nbd->config_refs);
>> schedule_work(&nsock->work);
>> }
>> @@ -817,11 +830,12 @@ static blk_status_t nbd_send_cmd(struct nbd_device *nbd, struct nbd_cmd *cmd,
>> static void nbd_pending_cmd_work(struct work_struct *work)
>> {
>> struct nbd_sock *nsock = container_of(work, struct nbd_sock, work);
>> - struct request *req = nsock->pending;
>> + struct request *req = nsock->partial_req;
>> struct nbd_cmd *cmd = blk_mq_rq_to_pdu(req);
>> struct nbd_device *nbd = cmd->nbd;
>> unsigned long deadline = READ_ONCE(req->deadline);
>> unsigned int wait_ms = 2;
>> + bool complete = false;
>> mutex_lock(&cmd->lock);
>> @@ -830,6 +844,18 @@ static void nbd_pending_cmd_work(struct work_struct *work)
>> goto out;
>> mutex_lock(&nsock->tx_lock);
>> + /*
>> + * nbd_mark_nsock_dead() can tear the socket down between schedule_work()
>> + * and here, and it clears ->pending and ->sent. The header is already
>> + * on the wire so this request can never be answered; fail it rather than
>> + * resuming a send on a socket that is gone.
>> + */
>> + if (!nsock->pending) {
>> + cmd->status = BLK_STS_IOERR;
>> + __clear_bit(NBD_CMD_INFLIGHT, &cmd->flags);
>> + complete = true;
>> + goto unlock;
>> + }
>> while (true) {
>> nbd_send_cmd(nbd, cmd, cmd->index);
>> if (!nsock->pending)
>> @@ -846,16 +872,36 @@ static void nbd_pending_cmd_work(struct work_struct *work)
>> * nbd_handle_cmd() requeue every later request forever.
>> */
>> nbd_mark_nsock_dead(nbd, nsock, 1);
>> - blk_mq_complete_request(req);
>> + complete = true;
>> break;
>> }
>> msleep(wait_ms);
>> wait_ms *= 2;
>> }
>> +unlock:
>> + /*
>> + * nbd_sched_pending_work() writes partial_req under tx_lock, and
>> + * cmd->lock is per-command, so this has to be under tx_lock too.
>> + */
>> + nsock->partial_req = NULL;
>> mutex_unlock(&nsock->tx_lock);
>> clear_bit(NBD_CMD_PARTIAL_SEND, &cmd->flags);
>> out:
>> mutex_unlock(&cmd->lock);
>> +
>> + /*
>> + * Complete before nbd_config_put(): if this is the last config
>> + * reference, nbd_put() runs nbd_dev_remove() inline unless
>> + * NBD_DESTROY_ON_DISCONNECT is set, and del_gendisk() would then wait
>> + * in blk_mq_freeze_queue_wait() for the q_usage_counter that this
>> + * request holds until it is completed. The config reference is what
>> + * keeps nbd itself alive across the completion, but cmd must not be
>> + * touched afterwards, since nbd_complete_rq() may run inline and ends
>> + * the request without taking cmd->lock.
>> + */
>> + if (complete)
>> + blk_mq_complete_request(req);
>> +
>> nbd_config_put(nbd);
>> }
>>
>
> These changes add significant complexity and hence make the NBD driver
> harder to maintain. Has it been considered to increase the request
> reference count while nsock->pending != NULL? See also req_ref_inc_not_zero() and blk_mq_put_rq_ref().
>
Thanks for taking a look.
It seems a request reference can't replace the extra pointer, since the
two answer different questions.
A reference keeps the request object alive, but it doesn't stop
nsock->pending from being cleared, which is what breaks the worker:
nbd_mark_nsock_dead()
nsock->dead = true;
nsock->pending = NULL;
nsock->sent = 0;
nbd_pending_cmd_work()
struct request *req = nsock->pending;
struct nbd_cmd *cmd = blk_mq_rq_to_pdu(req);
struct nbd_device *nbd = cmd->nbd;
Nothing between schedule_work() and the worker starting takes a lock the
teardown path also takes, so pinning the request still leaves the worker
reading NULL from ->pending and faulting on cmd->nbd.
Thanks,
Joseph
next prev parent reply other threads:[~2026-09-29 1:02 UTC|newest]
Thread overview: 7+ messages / expand[flat|nested] mbox.gz Atom feed top
2026-09-28 9:10 [PATCH v2 1/2] nbd: mark the socket dead when a partial send times out Joseph Qi
2026-09-28 9:10 ` [PATCH v2 2/2] nbd: fix NULL pointer dereference in nbd_pending_cmd_work() Joseph Qi
2026-09-28 14:19 ` Bart Van Assche
2026-09-29 1:02 ` Joseph Qi [this message]
2026-10-08 7:22 ` Joseph Qi
2026-09-28 14:22 ` [PATCH v2 1/2] nbd: mark the socket dead when a partial send times out Bart Van Assche
2026-09-29 9:03 ` Ming Lei
Reply instructions:
You may reply publicly to this message via plain-text email
using any one of the following methods:
* Save the following mbox file, import it into your mail client,
and reply-to-all from there: mbox
Avoid top-posting and favor interleaved quoting:
https://en.wikipedia.org/wiki/Posting_style#Interleaved_style
* Reply using the --to, --cc, and --in-reply-to
switches of git-send-email(1):
git send-email \
--in-reply-to=616f687a-417e-4fd5-b0dc-9bd519db3920@linux.alibaba.com \
--to=joseph.qi@linux.alibaba.com \
--cc=axboe@kernel.dk \
--cc=bvanassche@acm.org \
--cc=josef@toxicpanda.com \
--cc=linux-block@vger.kernel.org \
--cc=nbd@other.debian.org \
--cc=tom.leiming@gmail.com \
/path/to/YOUR_REPLY
https://kernel.org/pub/software/scm/git/docs/git-send-email.html
* If your mail client supports setting the In-Reply-To header
via mailto: links, try the mailto: link
Be sure your reply has a Subject: header at the top and a blank line
before the message body.
This is a public inbox, see mirroring instructions
for how to clone and mirror all data and code used for this inbox