From: Dave Chinner <dgc@kernel.org>
To: Christoph Hellwig <hch@infradead.org>
Cc: Gou Hao <gouhao@uniontech.com>,
cem@kernel.org, djwong@kernel.org, dchinner@redhat.com,
linux-xfs@vger.kernel.org, linux-kernel@vger.kernel.org,
niecheng1@uniontech.com, zhanjun@uniontech.com,
gouhaojake@163.com, gouhao@unionntech.com
Subject: Re: [PATCH] xfs: fix use-after-free of buf_log_item in xlog_cil_build_lv_chain
Date: Fri, 31 Jul 2026 07:58:22 +1000 [thread overview]
Message-ID: <amvI_rVUlifIXoZA@dread> (raw)
In-Reply-To: <ailYf4UzSLplTR3f@infradead.org>
On Wed, Jun 10, 2026 at 05:28:47AM -0700, Christoph Hellwig wrote:
> On Thu, Jun 04, 2026 at 05:42:33PM +0800, Gou Hao wrote:
> > xfs_buf_item_done() frees the buf_item via xfs_buf_item_relse() but
> > does not remove the item from the CIL log_items list (li_cil). When the
> > item is freed through an error/shutdown/abort path before the CIL push
> > worker processes it, the freed memory remains linked in ctx->log_items.
> >
> > The CIL push worker in xlog_cil_build_lv_chain() then dereferences
> > the freed object via item->li_lv, triggering a KASAN slab-use-after-free.
> > For details, see Link[1].
>
> There's no reproducer there. Do you have a local one?
>
> > Add down_read() on xc_ctx_lock before list_del_init() in
> > xfs_buf_item_done() to safely remove the item from the CIL list. This
> > uses the same lock that protects CIL list operations: insertions are
> > done under xc_ctx_lock read-side (xlog_cil_insert_items) and removals
> > under write-side (xlog_cil_build_lv_chain). The read lock is safe here
> > because xfs_buf_item_done() is always called in process context (workqueue
> > or direct I/O wait) and cannot deadlock with the CIL push worker which
> > holds the write lock during xlog_cil_build_lv_chain - the worker does not
> > trigger metadata buffer I/O that would call xfs_buf_item_done().
>
> This looks like a more general issue as we should never free anything
> that is still on the CIL. I.e. it looks like we have even more issues
> with the buf item state machine here :(
I finally found some time to look at this syzbot report and do some
analysis of it. AFAICT, the BLI life cycle is solid.
Go have a look at the syzbot report. i.e. where KASAN reports that
the BLI has been freed from. It is freed from buffer read IO
completion. Now go and have a look at xfs_buf_ioend(): the read IO
completion does not -ever- access attached BLIs - even on IO failure
- let alone free them. Only the write IO completion path (i.e.
!XBF_READ) accesses the attached BLI.
IOWs, the KASAN trace is telling us we've got a read IO completion
*without* XBF_READ being set on the buffer, and that freeing the BLI
from this context is how we ended up with the CIL reference to the
BLI being removed incorrectly leading to the UAF.
I spent some time trying to come up with ways we could have a read
IO completion run without XBF_READ being set, and I cannot find it.
No combination of corruption errors, reverifier, readahead, stale,
and/or shutdown races appear to allow a READ IO without XBF_READ
being set on the buffer through to IO completion, even in manual
failure paths like xfs_buf_fail().
I can only conclude that this was caused by one of two things:
- semaphores got broken unexpectedly; or
- memory corruption from some other syzbot test that was
running at the same time trashed bp->b_flags whilst the
buffer was under IO.
Given that syzbot has only reported this 5 times in 12 hours only on
a 7.1-rc3 kernel, never before and never since, an external memory
corruption bug that has since been fixed seems like the most like
cause here.
IOWs, I think there's nothing in XFS to fix here, and the syzbot
report should simply be closed.
-Dave.
--
Dave Chinner
dgc@kernel.org
prev parent reply other threads:[~2026-07-30 21:58 UTC|newest]
Thread overview: 4+ messages / expand[flat|nested] mbox.gz Atom feed top
2026-06-04 9:42 [PATCH] xfs: fix use-after-free of buf_log_item in xlog_cil_build_lv_chain Gou Hao
2026-06-10 12:28 ` Christoph Hellwig
2026-06-11 11:49 ` Gou Hao
2026-07-30 21:58 ` Dave Chinner [this message]
Reply instructions:
You may reply publicly to this message via plain-text email
using any one of the following methods:
* Save the following mbox file, import it into your mail client,
and reply-to-all from there: mbox
Avoid top-posting and favor interleaved quoting:
https://en.wikipedia.org/wiki/Posting_style#Interleaved_style
* Reply using the --to, --cc, and --in-reply-to
switches of git-send-email(1):
git send-email \
--in-reply-to=amvI_rVUlifIXoZA@dread \
--to=dgc@kernel.org \
--cc=cem@kernel.org \
--cc=dchinner@redhat.com \
--cc=djwong@kernel.org \
--cc=gouhao@unionntech.com \
--cc=gouhao@uniontech.com \
--cc=gouhaojake@163.com \
--cc=hch@infradead.org \
--cc=linux-kernel@vger.kernel.org \
--cc=linux-xfs@vger.kernel.org \
--cc=niecheng1@uniontech.com \
--cc=zhanjun@uniontech.com \
/path/to/YOUR_REPLY
https://kernel.org/pub/software/scm/git/docs/git-send-email.html
* If your mail client supports setting the In-Reply-To header
via mailto: links, try the mailto: link
Be sure your reply has a Subject: header at the top and a blank line
before the message body.
This is an external index of several public inboxes,
see mirroring instructions on how to clone and mirror
all data and code used by this external index.