* Re: [PATCH v3 2/2] cxl/memdev: Fix deadlock between poison debugfs and cxl_mem unbind
2026-09-10 9:40 ` [PATCH v3 2/2] cxl/memdev: Fix deadlock between poison debugfs and cxl_mem unbind Guixin Liu
@ 2026-09-10 9:52 ` Greg Kroah-Hartman
2026-09-10 11:28 ` Guixin Liu
2026-09-10 21:03 ` Jonathan Cameron
2026-09-15 20:21 ` Alison Schofield
2 siblings, 1 reply; 16+ messages in thread
From: Greg Kroah-Hartman @ 2026-09-10 9:52 UTC (permalink / raw)
To: Guixin Liu
Cc: Davidlohr Bueso, Jonathan Cameron, Dave Jiang, Alison Schofield,
Vishal Verma, Dan Williams, Ira Weiny, Li Ming,
Rafael J . Wysocki, Danilo Krummrich, Shaikh Kamaluddin,
linux-cxl, driver-core
On Thu, Sep 10, 2026 at 05:40:17PM +0800, Guixin Liu wrote:
> The poison debugfs handlers take the memdev device lock so that the
> region lookup sees a stable cxlmd->dev.driver. debugfs holds a reference
> on the file across the handler, and cxl_mem unbind removes that file
> while holding the very same device lock, so a handler that waits for the
> lock deadlocks against a concurrent unbind.
>
> Both tasks then hang. The unbind side is uninterruptible, and it also
> blocks the memdev detach work, which runs on an ordered workqueue and so
> stalls every other CXL bus work item.
>
> Take the lock with the trylock guard and return -EBUSY instead of
> waiting. An unbind that wins the race removes the file first and the
> write fails with -ENOENT.
>
> Found by code inspection. Reproduced by writing inject_poison in a loop
> while unbinding and rebinding cxl_mem, and confirmed fixed by the same
> test.
Why would anyone normally "unbind" cxl_mem at all? That will taint
kernels soon, so you don't normally want to do that, right?
And debugfs is root-only, so this is a "root did something bad, and gets
to keep the mess", right? This should not ever be a normal operation.
thanks,
greg k-h
^ permalink raw reply [flat|nested] 16+ messages in thread
* Re: [PATCH v3 2/2] cxl/memdev: Fix deadlock between poison debugfs and cxl_mem unbind
2026-09-10 9:52 ` Greg Kroah-Hartman
@ 2026-09-10 11:28 ` Guixin Liu
2026-09-10 11:58 ` Greg Kroah-Hartman
0 siblings, 1 reply; 16+ messages in thread
From: Guixin Liu @ 2026-09-10 11:28 UTC (permalink / raw)
To: Greg Kroah-Hartman
Cc: Davidlohr Bueso, Jonathan Cameron, Dave Jiang, Alison Schofield,
Vishal Verma, Dan Williams, Ira Weiny, Li Ming,
Rafael J . Wysocki, Danilo Krummrich, Shaikh Kamaluddin,
linux-cxl, driver-core
在 2026/9/10 17:52, Greg Kroah-Hartman 写道:
> On Thu, Sep 10, 2026 at 05:40:17PM +0800, Guixin Liu wrote:
>> The poison debugfs handlers take the memdev device lock so that the
>> region lookup sees a stable cxlmd->dev.driver. debugfs holds a reference
>> on the file across the handler, and cxl_mem unbind removes that file
>> while holding the very same device lock, so a handler that waits for the
>> lock deadlocks against a concurrent unbind.
>>
>> Both tasks then hang. The unbind side is uninterruptible, and it also
>> blocks the memdev detach work, which runs on an ordered workqueue and so
>> stalls every other CXL bus work item.
>>
>> Take the lock with the trylock guard and return -EBUSY instead of
>> waiting. An unbind that wins the race removes the file first and the
>> write fails with -ENOENT.
>>
>> Found by code inspection. Reproduced by writing inject_poison in a loop
>> while unbinding and rebinding cxl_mem, and confirmed fixed by the same
>> test.
> Why would anyone normally "unbind" cxl_mem at all? That will taint
> kernels soon, so you don't normally want to do that, right?
>
> And debugfs is root-only, so this is a "root did something bad, and gets
> to keep the mess", right? This should not ever be a normal operation.
Agreed, nobody should unbind cxl_mem in production. The sysfs unbind is
in the reproducer only because it's the cheapest trigger.
The window itself is not sysfs-unbind specific.
The debugfs directory is torn down from a devm action, so every path
that ends
in device_release_driver_internal() hits the same wait: rmmod cxl_mem,
and the
memdev detach work that PCI hot-remove schedules. That detach path is
the third
task in the reproducer stack, the one wedged on cxl_bus_wq.
A poison write racing an rmmod or a hot-remove is not root misbehaving,
and no taint flags it either.
Best Regards,
Guixin Liu
> thanks,
>
> greg k-h
^ permalink raw reply [flat|nested] 16+ messages in thread
* Re: [PATCH v3 2/2] cxl/memdev: Fix deadlock between poison debugfs and cxl_mem unbind
2026-09-10 11:28 ` Guixin Liu
@ 2026-09-10 11:58 ` Greg Kroah-Hartman
2026-09-10 20:52 ` Jonathan Cameron
0 siblings, 1 reply; 16+ messages in thread
From: Greg Kroah-Hartman @ 2026-09-10 11:58 UTC (permalink / raw)
To: Guixin Liu
Cc: Davidlohr Bueso, Jonathan Cameron, Dave Jiang, Alison Schofield,
Vishal Verma, Dan Williams, Ira Weiny, Li Ming,
Rafael J . Wysocki, Danilo Krummrich, Shaikh Kamaluddin,
linux-cxl, driver-core
On Thu, Sep 10, 2026 at 07:28:03PM +0800, Guixin Liu wrote:
>
>
> 在 2026/9/10 17:52, Greg Kroah-Hartman 写道:
> > On Thu, Sep 10, 2026 at 05:40:17PM +0800, Guixin Liu wrote:
> > > The poison debugfs handlers take the memdev device lock so that the
> > > region lookup sees a stable cxlmd->dev.driver. debugfs holds a reference
> > > on the file across the handler, and cxl_mem unbind removes that file
> > > while holding the very same device lock, so a handler that waits for the
> > > lock deadlocks against a concurrent unbind.
> > >
> > > Both tasks then hang. The unbind side is uninterruptible, and it also
> > > blocks the memdev detach work, which runs on an ordered workqueue and so
> > > stalls every other CXL bus work item.
> > >
> > > Take the lock with the trylock guard and return -EBUSY instead of
> > > waiting. An unbind that wins the race removes the file first and the
> > > write fails with -ENOENT.
> > >
> > > Found by code inspection. Reproduced by writing inject_poison in a loop
> > > while unbinding and rebinding cxl_mem, and confirmed fixed by the same
> > > test.
> > Why would anyone normally "unbind" cxl_mem at all? That will taint
> > kernels soon, so you don't normally want to do that, right?
> >
> > And debugfs is root-only, so this is a "root did something bad, and gets
> > to keep the mess", right? This should not ever be a normal operation.
> Agreed, nobody should unbind cxl_mem in production. The sysfs unbind is
> in the reproducer only because it's the cheapest trigger.
>
> The window itself is not sysfs-unbind specific.
> The debugfs directory is torn down from a devm action, so every path that
> ends
> in device_release_driver_internal() hits the same wait: rmmod cxl_mem, and
> the
> memdev detach work that PCI hot-remove schedules. That detach path is the
> third
> task in the reproducer stack, the one wedged on cxl_bus_wq.
> A poison write racing an rmmod or a hot-remove is not root misbehaving,
> and no taint flags it either.
But how can cxl_mem ever be removed, it doesn't live on a bus that is
hot-removable, does it?
rmmod doesn't count either, that never happens except by developers.
thanks,
greg k-h
^ permalink raw reply [flat|nested] 16+ messages in thread
* Re: [PATCH v3 2/2] cxl/memdev: Fix deadlock between poison debugfs and cxl_mem unbind
2026-09-10 11:58 ` Greg Kroah-Hartman
@ 2026-09-10 20:52 ` Jonathan Cameron
0 siblings, 0 replies; 16+ messages in thread
From: Jonathan Cameron @ 2026-09-10 20:52 UTC (permalink / raw)
To: Greg Kroah-Hartman
Cc: Guixin Liu, Davidlohr Bueso, Dave Jiang, Alison Schofield,
Vishal Verma, Dan Williams, Ira Weiny, Li Ming,
Rafael J . Wysocki, Danilo Krummrich, Shaikh Kamaluddin,
linux-cxl, driver-core
On Thu, 10 Sep 2026 13:58:28 +0200
Greg Kroah-Hartman <gregkh@linuxfoundation.org> wrote:
> On Thu, Sep 10, 2026 at 07:28:03PM +0800, Guixin Liu wrote:
> >
> >
> > 在 2026/9/10 17:52, Greg Kroah-Hartman 写道:
> > > On Thu, Sep 10, 2026 at 05:40:17PM +0800, Guixin Liu wrote:
> > > > The poison debugfs handlers take the memdev device lock so that the
> > > > region lookup sees a stable cxlmd->dev.driver. debugfs holds a reference
> > > > on the file across the handler, and cxl_mem unbind removes that file
> > > > while holding the very same device lock, so a handler that waits for the
> > > > lock deadlocks against a concurrent unbind.
> > > >
> > > > Both tasks then hang. The unbind side is uninterruptible, and it also
> > > > blocks the memdev detach work, which runs on an ordered workqueue and so
> > > > stalls every other CXL bus work item.
> > > >
> > > > Take the lock with the trylock guard and return -EBUSY instead of
> > > > waiting. An unbind that wins the race removes the file first and the
> > > > write fails with -ENOENT.
> > > >
> > > > Found by code inspection. Reproduced by writing inject_poison in a loop
> > > > while unbinding and rebinding cxl_mem, and confirmed fixed by the same
> > > > test.
> > > Why would anyone normally "unbind" cxl_mem at all? That will taint
> > > kernels soon, so you don't normally want to do that, right?
> > >
> > > And debugfs is root-only, so this is a "root did something bad, and gets
> > > to keep the mess", right? This should not ever be a normal operation.
> > Agreed, nobody should unbind cxl_mem in production. The sysfs unbind is
> > in the reproducer only because it's the cheapest trigger.
> >
> > The window itself is not sysfs-unbind specific.
> > The debugfs directory is torn down from a devm action, so every path that
> > ends
> > in device_release_driver_internal() hits the same wait: rmmod cxl_mem, and
> > the
> > memdev detach work that PCI hot-remove schedules. That detach path is the
> > third
> > task in the reproducer stack, the one wedged on cxl_bus_wq.
> > A poison write racing an rmmod or a hot-remove is not root misbehaving,
> > and no taint flags it either.
>
> But how can cxl_mem ever be removed, it doesn't live on a bus that is
> hot-removable, does it?
It is effectively a child of cxl_pci and that lives on the pci bus and
definitely is hotpluggable. Those flows should work fine so
I see this as a real if somewhat obscure bug to fix.
Jonathan
>
> rmmod doesn't count either, that never happens except by developers.
>
> thanks,
>
> greg k-h
^ permalink raw reply [flat|nested] 16+ messages in thread
* Re: [PATCH v3 2/2] cxl/memdev: Fix deadlock between poison debugfs and cxl_mem unbind
2026-09-10 9:40 ` [PATCH v3 2/2] cxl/memdev: Fix deadlock between poison debugfs and cxl_mem unbind Guixin Liu
2026-09-10 9:52 ` Greg Kroah-Hartman
@ 2026-09-10 21:03 ` Jonathan Cameron
2026-09-11 3:04 ` Guixin Liu
2026-09-15 20:21 ` Alison Schofield
2 siblings, 1 reply; 16+ messages in thread
From: Jonathan Cameron @ 2026-09-10 21:03 UTC (permalink / raw)
To: Guixin Liu
Cc: Davidlohr Bueso, Dave Jiang, Alison Schofield, Vishal Verma,
Dan Williams, Ira Weiny, Li Ming, Greg Kroah-Hartman,
Rafael J . Wysocki, Danilo Krummrich, Shaikh Kamaluddin,
linux-cxl, driver-core
On Thu, 10 Sep 2026 17:40:17 +0800
Guixin Liu <kanie@linux.alibaba.com> wrote:
> The poison debugfs handlers take the memdev device lock so that the
> region lookup sees a stable cxlmd->dev.driver. debugfs holds a reference
> on the file across the handler, and cxl_mem unbind removes that file
> while holding the very same device lock, so a handler that waits for the
> lock deadlocks against a concurrent unbind.
>
> Both tasks then hang. The unbind side is uninterruptible, and it also
> blocks the memdev detach work, which runs on an ordered workqueue and so
> stalls every other CXL bus work item.
>
> Take the lock with the trylock guard and return -EBUSY instead of
> waiting. An unbind that wins the race removes the file first and the
> write fails with -ENOENT.
>
> Found by code inspection. Reproduced by writing inject_poison in a loop
> while unbinding and rebinding cxl_mem, and confirmed fixed by the same
> test.
>
> Fixes: 574eda81d0a7 ("cxl/memdev: Hold memdev lock during memdev poison injection/clear")
> Suggested-by: Shaikh Kamaluddin <shaikhkamal2012@gmail.com>
> Signed-off-by: Guixin Liu <kanie@linux.alibaba.com>
Reviewed-by: Jonathan Cameron <jonathan.cameron@oss.qualcomm.com>
Probably not one to rush in but nice to clean up the deadlock even if it
is a little hard to hit. If it is possible to test with a hot remove
flow even better. You should be able to do that with emulation in qemu
if you don't have hardware capable of safe hotplug operations.
> ---
> checkpatch reports "do not use assignment in if condition" twice, on the
> two ACQUIRE_ERR() lines. Those are pre-existing: the unpatched file and
> the Fixes: commit report the same two, this patch only swaps the lock
> class on them, and the combined form is what all 40 ACQUIRE_ERR() call
> sites in drivers/cxl use.
We should fix that up. Oddly I thought we had, but guess not.
>
> drivers/cxl/mem.c | 14 ++++++++++----
> 1 file changed, 10 insertions(+), 4 deletions(-)
>
> diff --git a/drivers/cxl/mem.c b/drivers/cxl/mem.c
> index 798e5c369cfc..3959ec963026 100644
> --- a/drivers/cxl/mem.c
> +++ b/drivers/cxl/mem.c
> @@ -50,8 +50,13 @@ static int cxl_debugfs_poison_inject(void *data, u64 dpa)
> struct cxl_memdev *cxlmd = data;
> int rc;
>
> - ACQUIRE(device_intr, devlock)(&cxlmd->dev);
> - if ((rc = ACQUIRE_ERR(device_intr, &devlock)))
> + /*
> + * Never wait for this lock: the debugfs proxy holds a file reference
> + * across the callback and unbind removes the file under the same
> + * device lock, so waiting here deadlocks against unbind.
> + */
> + ACQUIRE(device_try, devlock)(&cxlmd->dev);
> + if ((rc = ACQUIRE_ERR(device_try, &devlock)))
> return rc;
>
> return cxl_inject_poison(cxlmd, dpa);
> @@ -65,8 +70,9 @@ static int cxl_debugfs_poison_clear(void *data, u64 dpa)
> struct cxl_memdev *cxlmd = data;
> int rc;
>
> - ACQUIRE(device_intr, devlock)(&cxlmd->dev);
> - if ((rc = ACQUIRE_ERR(device_intr, &devlock)))
> + /* Never wait, per the inject path above. */
> + ACQUIRE(device_try, devlock)(&cxlmd->dev);
> + if ((rc = ACQUIRE_ERR(device_try, &devlock)))
> return rc;
>
> return cxl_clear_poison(cxlmd, dpa);
^ permalink raw reply [flat|nested] 16+ messages in thread* Re: [PATCH v3 2/2] cxl/memdev: Fix deadlock between poison debugfs and cxl_mem unbind
2026-09-10 21:03 ` Jonathan Cameron
@ 2026-09-11 3:04 ` Guixin Liu
2026-09-11 23:09 ` Jonathan Cameron
0 siblings, 1 reply; 16+ messages in thread
From: Guixin Liu @ 2026-09-11 3:04 UTC (permalink / raw)
To: Jonathan Cameron
Cc: Davidlohr Bueso, Dave Jiang, Alison Schofield, Vishal Verma,
Dan Williams, Ira Weiny, Li Ming, Greg Kroah-Hartman,
Rafael J . Wysocki, Danilo Krummrich, Shaikh Kamaluddin,
linux-cxl, driver-core
在 2026/9/11 05:03, Jonathan Cameron 写道:
> On Thu, 10 Sep 2026 17:40:17 +0800
> Guixin Liu <kanie@linux.alibaba.com> wrote:
>
>> The poison debugfs handlers take the memdev device lock so that the
>> region lookup sees a stable cxlmd->dev.driver. debugfs holds a reference
>> on the file across the handler, and cxl_mem unbind removes that file
>> while holding the very same device lock, so a handler that waits for the
>> lock deadlocks against a concurrent unbind.
>>
>> Both tasks then hang. The unbind side is uninterruptible, and it also
>> blocks the memdev detach work, which runs on an ordered workqueue and so
>> stalls every other CXL bus work item.
>>
>> Take the lock with the trylock guard and return -EBUSY instead of
>> waiting. An unbind that wins the race removes the file first and the
>> write fails with -ENOENT.
>>
>> Found by code inspection. Reproduced by writing inject_poison in a loop
>> while unbinding and rebinding cxl_mem, and confirmed fixed by the same
>> test.
>>
>> Fixes: 574eda81d0a7 ("cxl/memdev: Hold memdev lock during memdev poison injection/clear")
>> Suggested-by: Shaikh Kamaluddin <shaikhkamal2012@gmail.com>
>> Signed-off-by: Guixin Liu <kanie@linux.alibaba.com>
> Reviewed-by: Jonathan Cameron <jonathan.cameron@oss.qualcomm.com>
>
> Probably not one to rush in but nice to clean up the deadlock even if it
> is a little hard to hit. If it is possible to test with a hot remove
> flow even better. You should be able to do that with emulation in qemu
> if you don't have hardware capable of safe hotplug operations.
Sure, I reproduced this on hot-remove situation:
debugfs writer, state S, waits for the memdev lock
cxl_debugfs_poison_inject+0x25/0xa0 [cxl_mem]
simple_attr_write_xsigned.isra.0+0x1ca/0x2c0
debugfs_attr_write+0x61/0xb0
full_proxy_write+0xfc/0x1c0
vfs_write+0x1d4/0xe60
ksys_write+0x11f/0x250
do_syscall_64+0xe2/0x560
entry_SYSCALL_64_after_hwframe+0x76/0x7e
irq/28-pciehp, state D, holds the memdev lock, waits for the debugfs
reference to drain
remove_one+0x27f/0x3d0
__simple_recursive_removal+0x183/0x4a0
debugfs_remove+0x44/0x60
release_nodes+0xfa/0x2c0
devres_release_all+0x113/0x1a0
device_unbind_cleanup+0x76/0x260
device_release_driver_internal+0x3eb/0x540
bus_remove_device+0x28a/0x540
device_del+0x371/0x930
cdev_device_del+0x1d/0xf0
cxl_memdev_unregister+0x1c/0x70 [cxl_core]
release_nodes+0xfa/0x2c0
devres_release_all+0x113/0x1a0
device_unbind_cleanup+0x76/0x260
device_release_driver_internal+0x3eb/0x540
pci_stop_bus_device+0x122/0x170
pci_stop_and_remove_bus_device+0x16/0x30
pciehp_unconfigure_device+0x1b4/0x3b0
pciehp_disable_slot+0xf9/0x2e0
pciehp_handle_disable_request+0x81/0x100
pciehp_ist+0x29f/0x410
irq_thread_fn+0x8b/0x160
irq_thread+0x189/0x320
kthread+0x329/0x410
ret_from_fork+0x33b/0x670
ret_from_fork_asm+0x1a/0x30
cxl_port workqueue, state D, waits for the memdev lock
device_release_driver_internal+0x96/0x540
detach_memdev+0x79/0xb0 [cxl_core]
process_one_work+0x6b0/0xfb0
worker_thread+0x4dd/0xd30
kthread+0x329/0x410
ret_from_fork+0x33b/0x670
ret_from_fork_asm+0x1a/0x30
With this patch, the hot-remove flow completed normally.
>
>> ---
>> checkpatch reports "do not use assignment in if condition" twice, on the
>> two ACQUIRE_ERR() lines. Those are pre-existing: the unpatched file and
>> the Fixes: commit report the same two, this patch only swaps the lock
>> class on them, and the combined form is what all 40 ACQUIRE_ERR() call
>> sites in drivers/cxl use.
> We should fix that up. Oddly I thought we had, but guess not.
I think we should fix this in checkpatch.pl, like this:
if ($c =~ /\bif\s*\(.*[^<>!=]=[^=].*/s &&
+ $c !~ /=\s*ACQUIRE_ERR\s*\(/) {
Best Regards,
Guixin Liu
>> drivers/cxl/mem.c | 14 ++++++++++----
>> 1 file changed, 10 insertions(+), 4 deletions(-)
>>
>> diff --git a/drivers/cxl/mem.c b/drivers/cxl/mem.c
>> index 798e5c369cfc..3959ec963026 100644
>> --- a/drivers/cxl/mem.c
>> +++ b/drivers/cxl/mem.c
>> @@ -50,8 +50,13 @@ static int cxl_debugfs_poison_inject(void *data, u64 dpa)
>> struct cxl_memdev *cxlmd = data;
>> int rc;
>>
>> - ACQUIRE(device_intr, devlock)(&cxlmd->dev);
>> - if ((rc = ACQUIRE_ERR(device_intr, &devlock)))
>> + /*
>> + * Never wait for this lock: the debugfs proxy holds a file reference
>> + * across the callback and unbind removes the file under the same
>> + * device lock, so waiting here deadlocks against unbind.
>> + */
>> + ACQUIRE(device_try, devlock)(&cxlmd->dev);
>> + if ((rc = ACQUIRE_ERR(device_try, &devlock)))
>> return rc;
>>
>> return cxl_inject_poison(cxlmd, dpa);
>> @@ -65,8 +70,9 @@ static int cxl_debugfs_poison_clear(void *data, u64 dpa)
>> struct cxl_memdev *cxlmd = data;
>> int rc;
>>
>> - ACQUIRE(device_intr, devlock)(&cxlmd->dev);
>> - if ((rc = ACQUIRE_ERR(device_intr, &devlock)))
>> + /* Never wait, per the inject path above. */
>> + ACQUIRE(device_try, devlock)(&cxlmd->dev);
>> + if ((rc = ACQUIRE_ERR(device_try, &devlock)))
>> return rc;
>>
>> return cxl_clear_poison(cxlmd, dpa);
^ permalink raw reply [flat|nested] 16+ messages in thread* Re: [PATCH v3 2/2] cxl/memdev: Fix deadlock between poison debugfs and cxl_mem unbind
2026-09-11 3:04 ` Guixin Liu
@ 2026-09-11 23:09 ` Jonathan Cameron
2026-09-14 12:08 ` Guixin Liu
0 siblings, 1 reply; 16+ messages in thread
From: Jonathan Cameron @ 2026-09-11 23:09 UTC (permalink / raw)
To: Guixin Liu
Cc: Davidlohr Bueso, Dave Jiang, Alison Schofield, Vishal Verma,
Dan Williams, Ira Weiny, Li Ming, Greg Kroah-Hartman,
Rafael J . Wysocki, Danilo Krummrich, Shaikh Kamaluddin,
linux-cxl, driver-core
On Fri, 11 Sep 2026 11:04:47 +0800
Guixin Liu <kanie@linux.alibaba.com> wrote:
> 在 2026/9/11 05:03, Jonathan Cameron 写道:
> > On Thu, 10 Sep 2026 17:40:17 +0800
> > Guixin Liu <kanie@linux.alibaba.com> wrote:
> >
> >> The poison debugfs handlers take the memdev device lock so that the
> >> region lookup sees a stable cxlmd->dev.driver. debugfs holds a reference
> >> on the file across the handler, and cxl_mem unbind removes that file
> >> while holding the very same device lock, so a handler that waits for the
> >> lock deadlocks against a concurrent unbind.
> >>
> >> Both tasks then hang. The unbind side is uninterruptible, and it also
> >> blocks the memdev detach work, which runs on an ordered workqueue and so
> >> stalls every other CXL bus work item.
> >>
> >> Take the lock with the trylock guard and return -EBUSY instead of
> >> waiting. An unbind that wins the race removes the file first and the
> >> write fails with -ENOENT.
> >>
> >> Found by code inspection. Reproduced by writing inject_poison in a loop
> >> while unbinding and rebinding cxl_mem, and confirmed fixed by the same
> >> test.
> >>
> >> Fixes: 574eda81d0a7 ("cxl/memdev: Hold memdev lock during memdev poison injection/clear")
> >> Suggested-by: Shaikh Kamaluddin <shaikhkamal2012@gmail.com>
> >> Signed-off-by: Guixin Liu <kanie@linux.alibaba.com>
> > Reviewed-by: Jonathan Cameron <jonathan.cameron@oss.qualcomm.com>
> >
> > Probably not one to rush in but nice to clean up the deadlock even if it
> > is a little hard to hit. If it is possible to test with a hot remove
> > flow even better. You should be able to do that with emulation in qemu
> > if you don't have hardware capable of safe hotplug operations.
> Sure, I reproduced this on hot-remove situation:
>
> debugfs writer, state S, waits for the memdev lock
> cxl_debugfs_poison_inject+0x25/0xa0 [cxl_mem]
...
> With this patch, the hot-remove flow completed normally.
Nice. Thanks for doing that.
>
> >
> >> ---
> >> checkpatch reports "do not use assignment in if condition" twice, on the
> >> two ACQUIRE_ERR() lines. Those are pre-existing: the unpatched file and
> >> the Fixes: commit report the same two, this patch only swaps the lock
> >> class on them, and the combined form is what all 40 ACQUIRE_ERR() call
> >> sites in drivers/cxl use.
> > We should fix that up. Oddly I thought we had, but guess not.
> I think we should fix this in checkpatch.pl, like this:
> if ($c =~ /\bif\s*\(.*[^<>!=]=[^=].*/s &&
> + $c !~ /=\s*ACQUIRE_ERR\s*\(/) {
There are a few other macros that are wrappers of ACQUIRE_ERR
that should be covered in such a patch as well. If you have
time send a patch!
Thanks,
Jonathan
^ permalink raw reply [flat|nested] 16+ messages in thread* Re: [PATCH v3 2/2] cxl/memdev: Fix deadlock between poison debugfs and cxl_mem unbind
2026-09-11 23:09 ` Jonathan Cameron
@ 2026-09-14 12:08 ` Guixin Liu
2026-09-15 19:35 ` Alison Schofield
0 siblings, 1 reply; 16+ messages in thread
From: Guixin Liu @ 2026-09-14 12:08 UTC (permalink / raw)
To: Jonathan Cameron
Cc: Davidlohr Bueso, Dave Jiang, Alison Schofield, Vishal Verma,
Dan Williams, Ira Weiny, Li Ming, Greg Kroah-Hartman,
Rafael J . Wysocki, Danilo Krummrich, Shaikh Kamaluddin,
linux-cxl, driver-core
在 2026/9/12 07:09, Jonathan Cameron 写道:
> On Fri, 11 Sep 2026 11:04:47 +0800
> Guixin Liu <kanie@linux.alibaba.com> wrote:
>
>> 在 2026/9/11 05:03, Jonathan Cameron 写道:
>>> On Thu, 10 Sep 2026 17:40:17 +0800
>>> Guixin Liu <kanie@linux.alibaba.com> wrote:
>>>
>>>> The poison debugfs handlers take the memdev device lock so that the
>>>> region lookup sees a stable cxlmd->dev.driver. debugfs holds a reference
>>>> on the file across the handler, and cxl_mem unbind removes that file
>>>> while holding the very same device lock, so a handler that waits for the
>>>> lock deadlocks against a concurrent unbind.
>>>>
>>>> Both tasks then hang. The unbind side is uninterruptible, and it also
>>>> blocks the memdev detach work, which runs on an ordered workqueue and so
>>>> stalls every other CXL bus work item.
>>>>
>>>> Take the lock with the trylock guard and return -EBUSY instead of
>>>> waiting. An unbind that wins the race removes the file first and the
>>>> write fails with -ENOENT.
>>>>
>>>> Found by code inspection. Reproduced by writing inject_poison in a loop
>>>> while unbinding and rebinding cxl_mem, and confirmed fixed by the same
>>>> test.
>>>>
>>>> Fixes: 574eda81d0a7 ("cxl/memdev: Hold memdev lock during memdev poison injection/clear")
>>>> Suggested-by: Shaikh Kamaluddin <shaikhkamal2012@gmail.com>
>>>> Signed-off-by: Guixin Liu <kanie@linux.alibaba.com>
>>> Reviewed-by: Jonathan Cameron <jonathan.cameron@oss.qualcomm.com>
>>>
>>> Probably not one to rush in but nice to clean up the deadlock even if it
>>> is a little hard to hit. If it is possible to test with a hot remove
>>> flow even better. You should be able to do that with emulation in qemu
>>> if you don't have hardware capable of safe hotplug operations.
>> Sure, I reproduced this on hot-remove situation:
>>
>> debugfs writer, state S, waits for the memdev lock
>> cxl_debugfs_poison_inject+0x25/0xa0 [cxl_mem]
> ...
>
>> With this patch, the hot-remove flow completed normally.
> Nice. Thanks for doing that.
>
>>>
>>>> ---
>>>> checkpatch reports "do not use assignment in if condition" twice, on the
>>>> two ACQUIRE_ERR() lines. Those are pre-existing: the unpatched file and
>>>> the Fixes: commit report the same two, this patch only swaps the lock
>>>> class on them, and the combined form is what all 40 ACQUIRE_ERR() call
>>>> sites in drivers/cxl use.
>>> We should fix that up. Oddly I thought we had, but guess not.
>> I think we should fix this in checkpatch.pl, like this:
>> if ($c =~ /\bif\s*\(.*[^<>!=]=[^=].*/s &&
>> + $c !~ /=\s*ACQUIRE_ERR\s*\(/) {
> There are a few other macros that are wrappers of ACQUIRE_ERR
> that should be covered in such a patch as well. If you have
> time send a patch!
>
> Thanks,
>
> Jonathan
Sure, I have already done that, please see: [PATCH] checkpatch: don't
flag ACQUIRE_ERR() assignments in if conditions
Best Regards,
Guixin Liu
^ permalink raw reply [flat|nested] 16+ messages in thread* Re: [PATCH v3 2/2] cxl/memdev: Fix deadlock between poison debugfs and cxl_mem unbind
2026-09-14 12:08 ` Guixin Liu
@ 2026-09-15 19:35 ` Alison Schofield
2026-09-16 2:23 ` Guixin Liu
0 siblings, 1 reply; 16+ messages in thread
From: Alison Schofield @ 2026-09-15 19:35 UTC (permalink / raw)
To: Guixin Liu
Cc: Jonathan Cameron, Davidlohr Bueso, Dave Jiang, Vishal Verma,
Dan Williams, Ira Weiny, Li Ming, Greg Kroah-Hartman,
Rafael J . Wysocki, Danilo Krummrich, Shaikh Kamaluddin,
linux-cxl, driver-core
On Mon, Sep 14, 2026 at 08:08:26PM +0800, Guixin Liu wrote:
>
>
> 在 2026/9/12 07:09, Jonathan Cameron 写道:
> > On Fri, 11 Sep 2026 11:04:47 +0800
> > Guixin Liu <kanie@linux.alibaba.com> wrote:
> >
> > > 在 2026/9/11 05:03, Jonathan Cameron 写道:
> > > > On Thu, 10 Sep 2026 17:40:17 +0800
> > > > Guixin Liu <kanie@linux.alibaba.com> wrote:
> > > > > The poison debugfs handlers take the memdev device lock so that the
> > > > > region lookup sees a stable cxlmd->dev.driver. debugfs holds a reference
> > > > > on the file across the handler, and cxl_mem unbind removes that file
> > > > > while holding the very same device lock, so a handler that waits for the
> > > > > lock deadlocks against a concurrent unbind.
> > > > >
> > > > > Both tasks then hang. The unbind side is uninterruptible, and it also
> > > > > blocks the memdev detach work, which runs on an ordered workqueue and so
> > > > > stalls every other CXL bus work item.
> > > > >
> > > > > Take the lock with the trylock guard and return -EBUSY instead of
> > > > > waiting. An unbind that wins the race removes the file first and the
> > > > > write fails with -ENOENT.
> > > > >
> > > > > Found by code inspection. Reproduced by writing inject_poison in a loop
> > > > > while unbinding and rebinding cxl_mem, and confirmed fixed by the same
> > > > > test.
> > > > >
> > > > > Fixes: 574eda81d0a7 ("cxl/memdev: Hold memdev lock during memdev poison injection/clear")
> > > > > Suggested-by: Shaikh Kamaluddin <shaikhkamal2012@gmail.com>
> > > > > Signed-off-by: Guixin Liu <kanie@linux.alibaba.com>
> > > > Reviewed-by: Jonathan Cameron <jonathan.cameron@oss.qualcomm.com>
> > > >
> > > > Probably not one to rush in but nice to clean up the deadlock even if it
> > > > is a little hard to hit. If it is possible to test with a hot remove
> > > > flow even better. You should be able to do that with emulation in qemu
> > > > if you don't have hardware capable of safe hotplug operations.
> > > Sure, I reproduced this on hot-remove situation:
> > >
> > > debugfs writer, state S, waits for the memdev lock
> > > cxl_debugfs_poison_inject+0x25/0xa0 [cxl_mem]
> > ...
> >
> > > With this patch, the hot-remove flow completed normally.
> > Nice. Thanks for doing that.
> >
> > > > > ---
> > > > > checkpatch reports "do not use assignment in if condition" twice, on the
> > > > > two ACQUIRE_ERR() lines. Those are pre-existing: the unpatched file and
> > > > > the Fixes: commit report the same two, this patch only swaps the lock
> > > > > class on them, and the combined form is what all 40 ACQUIRE_ERR() call
> > > > > sites in drivers/cxl use.
> > > > We should fix that up. Oddly I thought we had, but guess not.
> > > I think we should fix this in checkpatch.pl, like this:
> > > if ($c =~ /\bif\s*\(.*[^<>!=]=[^=].*/s &&
> > > + $c !~ /=\s*ACQUIRE_ERR\s*\(/) {
> > There are a few other macros that are wrappers of ACQUIRE_ERR
> > that should be covered in such a patch as well. If you have
> > time send a patch!
> >
> > Thanks,
> >
> > Jonathan
> Sure, I have already done that, please see: [PATCH] checkpatch: don't flag
> ACQUIRE_ERR() assignments in if conditions
Hi Guixin,
I tried same about a year ago and it was not merged, find it here:
https://lore.kernel.org/linux-cxl/20250815010645.2980846-1-alison.schofield@intel.com/
Note that we in CXL land decided to keep using this syntax and ignore
those checkpatch 'suggestions'.
Send me link to your checkpatch patch (can't find it? ) and I will review it.
-- Alison
>
> Best Regards,
> Guixin Liu
>
^ permalink raw reply [flat|nested] 16+ messages in thread* Re: [PATCH v3 2/2] cxl/memdev: Fix deadlock between poison debugfs and cxl_mem unbind
2026-09-15 19:35 ` Alison Schofield
@ 2026-09-16 2:23 ` Guixin Liu
0 siblings, 0 replies; 16+ messages in thread
From: Guixin Liu @ 2026-09-16 2:23 UTC (permalink / raw)
To: Alison Schofield
Cc: Jonathan Cameron, Davidlohr Bueso, Dave Jiang, Vishal Verma,
Dan Williams, Ira Weiny, Li Ming, Greg Kroah-Hartman,
Rafael J . Wysocki, Danilo Krummrich, Shaikh Kamaluddin,
linux-cxl, driver-core
在 2026/9/16 03:35, Alison Schofield 写道:
> On Mon, Sep 14, 2026 at 08:08:26PM +0800, Guixin Liu wrote:
>>
>> 在 2026/9/12 07:09, Jonathan Cameron 写道:
>>> On Fri, 11 Sep 2026 11:04:47 +0800
>>> Guixin Liu <kanie@linux.alibaba.com> wrote:
>>>
>>>> 在 2026/9/11 05:03, Jonathan Cameron 写道:
>>>>> On Thu, 10 Sep 2026 17:40:17 +0800
>>>>> Guixin Liu <kanie@linux.alibaba.com> wrote:
>>>>>> The poison debugfs handlers take the memdev device lock so that the
>>>>>> region lookup sees a stable cxlmd->dev.driver. debugfs holds a reference
>>>>>> on the file across the handler, and cxl_mem unbind removes that file
>>>>>> while holding the very same device lock, so a handler that waits for the
>>>>>> lock deadlocks against a concurrent unbind.
>>>>>>
>>>>>> Both tasks then hang. The unbind side is uninterruptible, and it also
>>>>>> blocks the memdev detach work, which runs on an ordered workqueue and so
>>>>>> stalls every other CXL bus work item.
>>>>>>
>>>>>> Take the lock with the trylock guard and return -EBUSY instead of
>>>>>> waiting. An unbind that wins the race removes the file first and the
>>>>>> write fails with -ENOENT.
>>>>>>
>>>>>> Found by code inspection. Reproduced by writing inject_poison in a loop
>>>>>> while unbinding and rebinding cxl_mem, and confirmed fixed by the same
>>>>>> test.
>>>>>>
>>>>>> Fixes: 574eda81d0a7 ("cxl/memdev: Hold memdev lock during memdev poison injection/clear")
>>>>>> Suggested-by: Shaikh Kamaluddin <shaikhkamal2012@gmail.com>
>>>>>> Signed-off-by: Guixin Liu <kanie@linux.alibaba.com>
>>>>> Reviewed-by: Jonathan Cameron <jonathan.cameron@oss.qualcomm.com>
>>>>>
>>>>> Probably not one to rush in but nice to clean up the deadlock even if it
>>>>> is a little hard to hit. If it is possible to test with a hot remove
>>>>> flow even better. You should be able to do that with emulation in qemu
>>>>> if you don't have hardware capable of safe hotplug operations.
>>>> Sure, I reproduced this on hot-remove situation:
>>>>
>>>> debugfs writer, state S, waits for the memdev lock
>>>> cxl_debugfs_poison_inject+0x25/0xa0 [cxl_mem]
>>> ...
>>>
>>>> With this patch, the hot-remove flow completed normally.
>>> Nice. Thanks for doing that.
>>>
>>>>>> ---
>>>>>> checkpatch reports "do not use assignment in if condition" twice, on the
>>>>>> two ACQUIRE_ERR() lines. Those are pre-existing: the unpatched file and
>>>>>> the Fixes: commit report the same two, this patch only swaps the lock
>>>>>> class on them, and the combined form is what all 40 ACQUIRE_ERR() call
>>>>>> sites in drivers/cxl use.
>>>>> We should fix that up. Oddly I thought we had, but guess not.
>>>> I think we should fix this in checkpatch.pl, like this:
>>>> if ($c =~ /\bif\s*\(.*[^<>!=]=[^=].*/s &&
>>>> + $c !~ /=\s*ACQUIRE_ERR\s*\(/) {
>>> There are a few other macros that are wrappers of ACQUIRE_ERR
>>> that should be covered in such a patch as well. If you have
>>> time send a patch!
>>>
>>> Thanks,
>>>
>>> Jonathan
>> Sure, I have already done that, please see: [PATCH] checkpatch: don't flag
>> ACQUIRE_ERR() assignments in if conditions
> Hi Guixin,
>
> I tried same about a year ago and it was not merged, find it here:
>
> https://lore.kernel.org/linux-cxl/20250815010645.2980846-1-alison.schofield@intel.com/
>
> Note that we in CXL land decided to keep using this syntax and ignore
> those checkpatch 'suggestions'.
>
> Send me link to your checkpatch patch (can't find it? ) and I will review it.
>
> -- Alison
Here is my patch:
https://lore.kernel.org/all/20260916020921.3480730-1-kanie@linux.alibaba.com/
Best Regards,
Guixin Liu
>
>> Best Regards,
>> Guixin Liu
>>
^ permalink raw reply [flat|nested] 16+ messages in thread
* Re: [PATCH v3 2/2] cxl/memdev: Fix deadlock between poison debugfs and cxl_mem unbind
2026-09-10 9:40 ` [PATCH v3 2/2] cxl/memdev: Fix deadlock between poison debugfs and cxl_mem unbind Guixin Liu
2026-09-10 9:52 ` Greg Kroah-Hartman
2026-09-10 21:03 ` Jonathan Cameron
@ 2026-09-15 20:21 ` Alison Schofield
2026-09-16 2:24 ` Guixin Liu
2 siblings, 1 reply; 16+ messages in thread
From: Alison Schofield @ 2026-09-15 20:21 UTC (permalink / raw)
To: Guixin Liu
Cc: Davidlohr Bueso, Jonathan Cameron, Dave Jiang, Vishal Verma,
Dan Williams, Ira Weiny, Li Ming, Greg Kroah-Hartman,
Rafael J . Wysocki, Danilo Krummrich, Shaikh Kamaluddin,
linux-cxl, driver-core
On Thu, Sep 10, 2026 at 05:40:17PM +0800, Guixin Liu wrote:
> The poison debugfs handlers take the memdev device lock so that the
> region lookup sees a stable cxlmd->dev.driver. debugfs holds a reference
> on the file across the handler, and cxl_mem unbind removes that file
> while holding the very same device lock, so a handler that waits for the
> lock deadlocks against a concurrent unbind.
>
> Both tasks then hang. The unbind side is uninterruptible, and it also
> blocks the memdev detach work, which runs on an ordered workqueue and so
> stalls every other CXL bus work item.
>
> Take the lock with the trylock guard and return -EBUSY instead of
> waiting. An unbind that wins the race removes the file first and the
> write fails with -ENOENT.
>
> Found by code inspection. Reproduced by writing inject_poison in a loop
> while unbinding and rebinding cxl_mem, and confirmed fixed by the same
> test.
Hi Guixin,
I'm good w closing this gap, however, I'm cautious about claiming more
robust support for this debugfs interface than we intend. We offer this poison
inject and clear ABI for expert users, mostly device vendors, to test their
CXL device capabilities. If those users take the step to do something like
'hotplugging while poisoning', well then they are just being reckless.
This patch commit message should make it clear the user scenario that would
hit the deadlock, and that users should NOT do that intentionally! But if they
do, this fixup prevents deadlock.
BTW - I really do appreciate that you replicated this as opposed to many other
patches that both begin and end with LLM suggestions and give no consideration
to, nor demonstrate an understanding of, the actual user visible impact.
-- ALison
>
> Fixes: 574eda81d0a7 ("cxl/memdev: Hold memdev lock during memdev poison injection/clear")
> Suggested-by: Shaikh Kamaluddin <shaikhkamal2012@gmail.com>
> Signed-off-by: Guixin Liu <kanie@linux.alibaba.com>
> ---
> checkpatch reports "do not use assignment in if condition" twice, on the
> two ACQUIRE_ERR() lines. Those are pre-existing: the unpatched file and
> the Fixes: commit report the same two, this patch only swaps the lock
> class on them, and the combined form is what all 40 ACQUIRE_ERR() call
> sites in drivers/cxl use.
>
> drivers/cxl/mem.c | 14 ++++++++++----
> 1 file changed, 10 insertions(+), 4 deletions(-)
>
> diff --git a/drivers/cxl/mem.c b/drivers/cxl/mem.c
> index 798e5c369cfc..3959ec963026 100644
> --- a/drivers/cxl/mem.c
> +++ b/drivers/cxl/mem.c
> @@ -50,8 +50,13 @@ static int cxl_debugfs_poison_inject(void *data, u64 dpa)
> struct cxl_memdev *cxlmd = data;
> int rc;
>
> - ACQUIRE(device_intr, devlock)(&cxlmd->dev);
> - if ((rc = ACQUIRE_ERR(device_intr, &devlock)))
> + /*
> + * Never wait for this lock: the debugfs proxy holds a file reference
> + * across the callback and unbind removes the file under the same
> + * device lock, so waiting here deadlocks against unbind.
> + */
> + ACQUIRE(device_try, devlock)(&cxlmd->dev);
> + if ((rc = ACQUIRE_ERR(device_try, &devlock)))
> return rc;
>
> return cxl_inject_poison(cxlmd, dpa);
> @@ -65,8 +70,9 @@ static int cxl_debugfs_poison_clear(void *data, u64 dpa)
> struct cxl_memdev *cxlmd = data;
> int rc;
>
> - ACQUIRE(device_intr, devlock)(&cxlmd->dev);
> - if ((rc = ACQUIRE_ERR(device_intr, &devlock)))
> + /* Never wait, per the inject path above. */
> + ACQUIRE(device_try, devlock)(&cxlmd->dev);
> + if ((rc = ACQUIRE_ERR(device_try, &devlock)))
> return rc;
>
> return cxl_clear_poison(cxlmd, dpa);
> --
> 2.43.7
>
^ permalink raw reply [flat|nested] 16+ messages in thread* Re: [PATCH v3 2/2] cxl/memdev: Fix deadlock between poison debugfs and cxl_mem unbind
2026-09-15 20:21 ` Alison Schofield
@ 2026-09-16 2:24 ` Guixin Liu
0 siblings, 0 replies; 16+ messages in thread
From: Guixin Liu @ 2026-09-16 2:24 UTC (permalink / raw)
To: Alison Schofield
Cc: Davidlohr Bueso, Jonathan Cameron, Dave Jiang, Vishal Verma,
Dan Williams, Ira Weiny, Li Ming, Greg Kroah-Hartman,
Rafael J . Wysocki, Danilo Krummrich, Shaikh Kamaluddin,
linux-cxl, driver-core
在 2026/9/16 04:21, Alison Schofield 写道:
> On Thu, Sep 10, 2026 at 05:40:17PM +0800, Guixin Liu wrote:
>> The poison debugfs handlers take the memdev device lock so that the
>> region lookup sees a stable cxlmd->dev.driver. debugfs holds a reference
>> on the file across the handler, and cxl_mem unbind removes that file
>> while holding the very same device lock, so a handler that waits for the
>> lock deadlocks against a concurrent unbind.
>>
>> Both tasks then hang. The unbind side is uninterruptible, and it also
>> blocks the memdev detach work, which runs on an ordered workqueue and so
>> stalls every other CXL bus work item.
>>
>> Take the lock with the trylock guard and return -EBUSY instead of
>> waiting. An unbind that wins the race removes the file first and the
>> write fails with -ENOENT.
>>
>> Found by code inspection. Reproduced by writing inject_poison in a loop
>> while unbinding and rebinding cxl_mem, and confirmed fixed by the same
>> test.
> Hi Guixin,
>
> I'm good w closing this gap, however, I'm cautious about claiming more
> robust support for this debugfs interface than we intend. We offer this poison
> inject and clear ABI for expert users, mostly device vendors, to test their
> CXL device capabilities. If those users take the step to do something like
> 'hotplugging while poisoning', well then they are just being reckless.
>
> This patch commit message should make it clear the user scenario that would
> hit the deadlock, and that users should NOT do that intentionally! But if they
> do, this fixup prevents deadlock.
Sure, I will change this in v4, thanks.
Best Regards,
Guixin Liu
>
> BTW - I really do appreciate that you replicated this as opposed to many other
> patches that both begin and end with LLM suggestions and give no consideration
> to, nor demonstrate an understanding of, the actual user visible impact.
>
> -- ALison
>
>
>> Fixes: 574eda81d0a7 ("cxl/memdev: Hold memdev lock during memdev poison injection/clear")
>> Suggested-by: Shaikh Kamaluddin <shaikhkamal2012@gmail.com>
>> Signed-off-by: Guixin Liu <kanie@linux.alibaba.com>
>> ---
>> checkpatch reports "do not use assignment in if condition" twice, on the
>> two ACQUIRE_ERR() lines. Those are pre-existing: the unpatched file and
>> the Fixes: commit report the same two, this patch only swaps the lock
>> class on them, and the combined form is what all 40 ACQUIRE_ERR() call
>> sites in drivers/cxl use.
>>
>> drivers/cxl/mem.c | 14 ++++++++++----
>> 1 file changed, 10 insertions(+), 4 deletions(-)
>>
>> diff --git a/drivers/cxl/mem.c b/drivers/cxl/mem.c
>> index 798e5c369cfc..3959ec963026 100644
>> --- a/drivers/cxl/mem.c
>> +++ b/drivers/cxl/mem.c
>> @@ -50,8 +50,13 @@ static int cxl_debugfs_poison_inject(void *data, u64 dpa)
>> struct cxl_memdev *cxlmd = data;
>> int rc;
>>
>> - ACQUIRE(device_intr, devlock)(&cxlmd->dev);
>> - if ((rc = ACQUIRE_ERR(device_intr, &devlock)))
>> + /*
>> + * Never wait for this lock: the debugfs proxy holds a file reference
>> + * across the callback and unbind removes the file under the same
>> + * device lock, so waiting here deadlocks against unbind.
>> + */
>> + ACQUIRE(device_try, devlock)(&cxlmd->dev);
>> + if ((rc = ACQUIRE_ERR(device_try, &devlock)))
>> return rc;
>>
>> return cxl_inject_poison(cxlmd, dpa);
>> @@ -65,8 +70,9 @@ static int cxl_debugfs_poison_clear(void *data, u64 dpa)
>> struct cxl_memdev *cxlmd = data;
>> int rc;
>>
>> - ACQUIRE(device_intr, devlock)(&cxlmd->dev);
>> - if ((rc = ACQUIRE_ERR(device_intr, &devlock)))
>> + /* Never wait, per the inject path above. */
>> + ACQUIRE(device_try, devlock)(&cxlmd->dev);
>> + if ((rc = ACQUIRE_ERR(device_try, &devlock)))
>> return rc;
>>
>> return cxl_clear_poison(cxlmd, dpa);
>> --
>> 2.43.7
>>
^ permalink raw reply [flat|nested] 16+ messages in thread