From: Pavel Begunkov <asml.silence@gmail.com>
To: sashiko-reviews@lists.linux.dev
Cc: media-ci@linuxtv.org, dri-devel@lists.freedesktop.org
Subject: Re: [PATCH v7 06/13] nvme-pci: implement dma-buf backed requests
Date: Mon, 28 Sep 2026 15:41:28 +0100 [thread overview]
Message-ID: <38f81e58-44ca-4fe6-ba72-97557b7625e9@gmail.com> (raw)
In-Reply-To: <20260928140208.0C56B1F000FF@smtp.kernel.org>
On 9/28/26 15:02, sashiko-bot@kernel.org wrote:
...>> + nvme_dmabuf_map_sync_for_device(nvmeq->dev, req);
>> +
>> + if (length <= 0) {
>> + prp2_dma = 0;
>> + goto done;
>> + }
>> +
>> + if (length <= NVME_CTRL_PAGE_SIZE) {
>> + prp2_dma = dma_list[map_idx];
>
> [Severity: Critical]
> Is map_idx checked against map->nr_entries here?
>
> If dmabuf->size is 4608 bytes, nr_entries is truncated to 1, allocating 1
It'd be rejected, not truncated, see below.
> array element. Userspace could submit a 4608-byte IO. The driver calculates
> 'length' up to 4608 and advances 'map_idx' to 1 for the remaining 512 bytes.
> It reads 'dma_list[1]' (which is out-of-bounds) and uses it as 'prp2_dma',
> potentially resulting in arbitrary physical DMA (read or write).
>
>> + goto done;
>> + }
>
> [ ... ]
>
>> +free_prps:
>> + iod->cmd.common.dptr.prp1 = cpu_to_le64(prp1_dma);
>> + iod->cmd.common.dptr.prp2 = cpu_to_le64(prp2_dma);
>> + nvme_free_descriptors(req);
>> + return BLK_STS_RESOURCE;
>
> [Severity: High]
> Does this code leave iod->nr_descriptors greater than 0 on the error path?
I can clear it, but unless I missed sth that's what e.g.
nvme_pci_setup_data_prp() already does.
nvme_pci_setup_data_prp() // goto bad_sgl:
nvme_unmap_data()
nvme_free_descriptors();
> If dma_pool_alloc() fails in nvme_rq_setup_dmabuf_map(), it jumps to
> 'free_prps:', which calls nvme_free_descriptors() but leaves
> 'iod->nr_descriptors' unchanged. The block layer may requeue the failed
> request.
>
> If the queue is subsequently torn down, the cancellation path invokes
> nvme_unmap_data() -> nvme_rq_clean_dmabuf_map(). Seeing the non-zero
> 'nr_descriptors', it calls nvme_free_descriptors() again, resulting in a
> double free.
>
>> +}
>
> [ ... ]
>
>> +static struct dma_buf_io_map *nvme_dma_buf_io_map(struct dma_buf_io_ctx *ctx)
>> +{
>> + unsigned nr_entries = ctx->dmabuf->size / NVME_CTRL_PAGE_SIZE;
>> + struct nvme_dma_buf_io_ctx *nvme_ctx = ctx->dev_priv;
>> + struct dma_buf_attachment *attach = nvme_ctx->attach;
>> + unsigned long tmp, i = 0;
>> + struct nvme_dmabuf_map *map;
>> + struct scatterlist *sg;
>> + struct sg_table *sgt;
>> + int ret;
>> +
>> + dma_resv_assert_held(ctx->dmabuf->resv);
>> +
>> + map = kvmalloc_flex(*map, dma_list, nr_entries);
>> + if (!map)
>> + return ERR_PTR(-ENOMEM);
>> +
>> + sgt = dma_buf_map_attachment(attach, ctx->dir);
>> + if (IS_ERR(sgt)) {
>> + ret = PTR_ERR(sgt);
>> + sgt = NULL;
>> + goto err;
>> + }
>> +
>> + for_each_sgtable_dma_sg(sgt, sg, tmp) {
>> + dma_addr_t dma_addr = sg_dma_address(sg);
>> + unsigned long sg_len = sg_dma_len(sg);
>> +
>> + if ((sg_len % NVME_CTRL_PAGE_SIZE) ||
>> + (dma_addr % NVME_CTRL_PAGE_SIZE)) {
>> + ret = -EINVAL;
>> + goto err;
>> + }
>> + while (sg_len) {
>> + map->dma_list[i++] = dma_addr;
>
> [Severity: High]
> Does this loop overflow dma_list[]?
>
> 'nr_entries' is computed via integer division of dmabuf->size. If the dmabuf
> size is not page-aligned (e.g., 4097 bytes), nr_entries truncates down (to 1).
It'll fail on the alignment check in this loop for [4096, 4097)
> However, exporters often pad sg_table allocations to full pages (e.g., 8192
> bytes).
>
> The loop unconditionally writes all sg_len chunks to 'dma_list', exceeding the
> allocated 'nr_entries' capacity.
>
>> + dma_addr += NVME_CTRL_PAGE_SIZE;
>> + sg_len -= NVME_CTRL_PAGE_SIZE;
>> + }
>> + }
>> +
>> + ret = dma_buf_io_init_map(ctx, &map->base, sgt);
>> + if (ret)
>> + goto err;
>> + map->nr_entries = nr_entries;
>> + map->sgt = sgt;
>> + return &map->base;
>> +err:
>> + if (sgt)
>> + dma_buf_unmap_attachment(attach, sgt, ctx->dir);
>> + kfree(map);
>
> [Severity: High]
> Is it safe to free a kvmalloc_flex() allocation with kfree()?
Going to fix.
--
Pavel Begunkov
next prev parent reply other threads:[~2026-09-28 14:41 UTC|newest]
Thread overview: 26+ messages / expand[flat|nested] mbox.gz Atom feed top
2026-09-28 13:32 [PATCH v7 00/13] Add dmabuf read/write via io_uring Pavel Begunkov
2026-09-28 13:32 ` [PATCH v7 01/13] dma-buf: introduce initial file I/O infrastructure Pavel Begunkov
2026-09-28 13:46 ` sashiko-bot
2026-09-29 16:12 ` Christophe JAILLET
2026-09-28 13:32 ` [PATCH v7 02/13] iov_iter: add iterator type for dmabuf maps Pavel Begunkov
2026-09-28 13:49 ` sashiko-bot
[not found] ` <20261005091008.GA10727@lst.de>
2026-10-05 11:44 ` Pavel Begunkov
2026-09-28 13:32 ` [PATCH v7 03/13] block: always adjust bi_offset on bio_advance_iter Pavel Begunkov
2026-09-28 13:32 ` [PATCH v7 04/13] block: introduce dma map backed bio type Pavel Begunkov
2026-09-28 13:32 ` [PATCH v7 05/13] block: add dma-buf support for raw bdev Pavel Begunkov
2026-09-28 13:46 ` sashiko-bot
[not found] ` <20261005091111.GB10727@lst.de>
2026-10-05 11:43 ` Pavel Begunkov
2026-09-28 13:32 ` [PATCH v7 06/13] nvme-pci: implement dma-buf backed requests Pavel Begunkov
2026-09-28 14:02 ` sashiko-bot
2026-09-28 14:41 ` Pavel Begunkov [this message]
2026-09-29 15:45 ` Christophe JAILLET
2026-09-30 10:09 ` Pavel Begunkov
[not found] ` <20261005091554.GD10727@lst.de>
2026-10-05 11:45 ` Pavel Begunkov
2026-09-28 13:32 ` [PATCH v7 07/13] nvme-pci: rename nvme_pci_sgl_set_data to nvme_pci_dma_iter_set_sgl Pavel Begunkov
2026-09-28 13:32 ` [PATCH v7 08/13] nvme-pci: add SGL support for the dmabuf path Pavel Begunkov
2026-09-28 13:32 ` [PATCH v7 09/13] io_uring/rsrc: introduce buf registration structure Pavel Begunkov
2026-09-28 13:32 ` [PATCH v7 10/13] io_uring/rsrc: extend buffer update Pavel Begunkov
2026-09-28 13:32 ` [PATCH v7 11/13] io_uring/rsrc: add uncloneable regbuf flag Pavel Begunkov
2026-09-28 13:32 ` [PATCH v7 12/13] io_uring/rsrc: add regbuf import flags Pavel Begunkov
2026-09-28 13:32 ` [PATCH v7 13/13] io_uring/rsrc: add dmabuf backed registered buffers Pavel Begunkov
2026-09-28 14:03 ` sashiko-bot
Reply instructions:
You may reply publicly to this message via plain-text email
using any one of the following methods:
* Save the following mbox file, import it into your mail client,
and reply-to-all from there: mbox
Avoid top-posting and favor interleaved quoting:
https://en.wikipedia.org/wiki/Posting_style#Interleaved_style
* Reply using the --to, --cc, and --in-reply-to
switches of git-send-email(1):
git send-email \
--in-reply-to=38f81e58-44ca-4fe6-ba72-97557b7625e9@gmail.com \
--to=asml.silence@gmail.com \
--cc=dri-devel@lists.freedesktop.org \
--cc=media-ci@linuxtv.org \
--cc=sashiko-reviews@lists.linux.dev \
/path/to/YOUR_REPLY
https://kernel.org/pub/software/scm/git/docs/git-send-email.html
* If your mail client supports setting the In-Reply-To header
via mailto: links, try the mailto: link
Be sure your reply has a Subject: header at the top and a blank line
before the message body.
This is a public inbox, see mirroring instructions
for how to clone and mirror all data and code used for this inbox