Linux s390 Architecture development
 help / color / mirror / Atom feed
From: Anthony Krowiak <akrowiak@linux.ibm.com>
To: sashiko-reviews@lists.linux.dev
Cc: linux-s390@vger.kernel.org, Vasily Gorbik <gor@linux.ibm.com>,
	Heiko Carstens <hca@linux.ibm.com>,
	Alexander Gordeev <agordeev@linux.ibm.com>,
	kvm@vger.kernel.org, Jason Herne <jjherne@linux.ibm.com>,
	Christian Borntraeger <borntraeger@linux.ibm.com>,
	Alex Williamson <alex@shazbot.org>,
	Matthew Rosato <mjrosato@linux.ibm.com>,
	Halil Pasic <pasic@linux.ibm.com>,
	Kirti Wankhede <kwankhede@nvidia.com>,
	Paolo Bonzini <pbonzini@redhat.com>,
	Janosch Frank <frankja@linux.ibm.com>,
	Claudio Imbrenda <imbrenda@linux.ibm.com>
Subject: Re: [PATCH v6 08/15] s390/vfio-ap: Transition device migration state from STOP to RESUMING
Date: Tue, 28 Jul 2026 18:29:37 -0400	[thread overview]
Message-ID: <41149777-2df1-4225-ac68-c9038302aee4@linux.ibm.com> (raw)
In-Reply-To: <20260727181408.805861F000E9@smtp.kernel.org>



On 7/27/26 2:14 PM, sashiko-bot@kernel.org wrote:
> Thank you for your contribution! Sashiko AI review found 2 potential issue(s) to consider:
> - [Critical] Use-After-Free in `vfio_ap_stop_copy_read` leading to kernel memory leak to userspace.
> - [Low] The `.release` callback logic for cleaning up the resuming and stop-copy file states is completely unreachable dead code.
> --
>
> --- Patch [8]: [PATCH v6 08/15] s390/vfio-ap: Transition device migration state from STOP to RESUMING ---
> commit 798f82863e85b4d37fd3594dc855fe7caaaaf7f5
> Author: Anthony Krowiak <akrowiak@linux.ibm.com>
>
> s390/vfio-ap: Transition device migration state from STOP to RESUMING
>
> Transitioning from VFIO_DEVICE_STATE_STOP to VFIO_DEVICE_STATE_RESUMING
> starts a process of restoring the device state of the vfio device on the
> target system. To prepare for restoring the device state, open a file
> stream to receive the guest's AP configuration saved when the device state
> on the source system transitioned from VFIO_DEVICE_STATE_STOP to
> VFIO_DEVICE_STATE_STOP_COPY.
>
>> diff --git a/drivers/s390/crypto/vfio_ap_migration.c b/drivers/s390/crypto/vfio_ap_migration.c
>> index e4bc67b1eb842..c7fecad0b6769 100644
>> --- a/drivers/s390/crypto/vfio_ap_migration.c
>> +++ b/drivers/s390/crypto/vfio_ap_migration.c
> [ ... ]
>
>> @@ -99,6 +108,8 @@ static int vfio_ap_release_mig_file(struct inode *file_inode, struct file *filp)
>>   
>>   	if (filp == matrix_mdev->mig_data->stop_copy_mig_file.filp)
>>   		vfio_ap_release_stop_copy_file(matrix_mdev->mig_data);
>> +	else if (filp == matrix_mdev->mig_data->resuming_mig_file.filp)
>> +		vfio_ap_release_resuming_file(matrix_mdev->mig_data);
> [Severity: Low]
> Is this conditional branch actually reachable?
>
> During vfio_ap_open_file_stream(), the driver takes an internal reference to
> the file via get_file(). This means the VFS .release callback is only invoked
> after the driver explicitly drops its reference by calling fput().
>
> If fput() is called in vfio_ap_release_mig_files(), which first sets
> mig_data->resuming_mig_file.filp = NULL, won't filp inside mig_data always
> be NULL by the time this .release callback executes?
>
> Consequently, will this check always evaluate to false, leaving
> vfio_ap_release_resuming_file() as unused dead code?

This is correct; those branches are unreachable. The root cause is
that vfio_ap_release_mig_files() calls fput() on the stored filp and then
immediately sets it to NULL, before the deferred .release callback can
fire. Since fput() always defers __fput() via task_work or a workqueue
(it never calls .release synchronously while a mutex is held), the .release
callback always arrives after the pointer is already NULL — making both
else if branches dead code.

The design will be reworked.

>
> [ ... ]
>
>> @@ -400,6 +411,26 @@ static struct file *vfio_ap_open_file_stream(struct ap_matrix_mdev *matrix_mdev,
>>   	return filp;
>>   }
>>   
>> +static ssize_t vfio_ap_resuming_write(struct file *filp, const char __user *buf,
>> +				      size_t len, loff_t *pos)
>> +{
>> +	/* TODO */
>> +	return -EOPNOTSUPP;
>> +}
>> +
>> +static const struct file_operations vfio_ap_resume_fops = {
>> +	.owner = THIS_MODULE,
>> +	.write = vfio_ap_resuming_write,
>> +	.release = vfio_ap_release_mig_file,
>> +};
> [Severity: Critical]
> Could the migration file stream logic be vulnerable to a use-after-free when
> the device file descriptor is concurrently closed?
>
> Looking at the related vfio_ap_stop_copy_read() stream implementation, it
> drops mdevs_lock before calling copy_to_user():
>
> drivers/s390/crypto/vfio_ap_migration.c:vfio_ap_stop_copy_read() {
>      ...
>      mutex_unlock(&matrix_dev->mdevs_lock);
>
>      if (copy_to_user(buf, (char *)mig_file->ap_config + read_pos, len))
>          return -EFAULT;
>      ...
> }
>
> If the device file descriptor is concurrently closed, vfio_ap_mdev_release()
> could invoke vfio_ap_release_migration_data(), which acquires mdevs_lock and
> frees the mig_data structure.
>
> When vfio_ap_stop_copy_read() resumes, will it dereference the freed mig_data
> structure via mig_file->ap_config and leak kernel memory to userspace?

The use-after-free described is not present. The safety relies on two
invariants:

1. matrix_mdev (and therefore mig_data) cannot be freed while the migration
    FD is open. vfio_ap_open_file_stream() calls
    vfio_device_try_get_registration() which pins matrix_mdev until
    vfio_ap_release_mig_file() drops it. Device teardown
    (vfio_ap_mdev_close_device()) can free mig_data independently, but
    matrix_mdev itself remains valid.

2. mig_file->ap_config is allocated once in vfio_ap_get_config() and freed
    only in vfio_ap_release_stop_copy_file(), which requires mdevs_lock.
    Since vfio_ap_stop_copy_read() advances *pos and snapshots read_pos
    before dropping the lock, any concurrent path that acquires the lock and
    frees ap_config can only do so after the position has already been
    committed — and copy_to_user() only touches the buffer contents, not the
    pointer itself.

>


  reply	other threads:[~2026-07-28 22:29 UTC|newest]

Thread overview: 42+ messages / expand[flat|nested]  mbox.gz  Atom feed  top
2026-07-27 17:32 [PATCH v6 00/15] s390/vfio-ap: Add live guest migration support Anthony Krowiak
2026-07-27 17:32 ` [PATCH v6 01/15] s390/vfio-ap: Provide function to get the number of queues assigned to mdev Anthony Krowiak
2026-07-27 17:38   ` sashiko-bot
2026-07-27 17:32 ` [PATCH v6 02/15] s390/vfio-ap: Data structures for facilitating vfio device migration Anthony Krowiak
2026-07-27 17:40   ` sashiko-bot
2026-08-04 10:39     ` Anthony Krowiak
2026-08-04 10:49     ` Anthony Krowiak
2026-08-04 19:57     ` Anthony Krowiak
2026-07-27 17:32 ` [PATCH v6 03/15] s390/vfio-ap: Functions to initialize/release vfio device migration data Anthony Krowiak
2026-07-27 17:48   ` sashiko-bot
2026-08-04 19:51     ` Anthony Krowiak
2026-07-27 17:32 ` [PATCH v6 04/15] s390/vfio-ap: Reset migration state in VFIO_DEVICE_RESET ioctl handler Anthony Krowiak
2026-07-27 17:52   ` sashiko-bot
2026-07-27 17:32 ` [PATCH v6 05/15] s390-vfio-ap: Callback to get/set vfio device mig state during guest migration Anthony Krowiak
2026-07-27 17:59   ` sashiko-bot
2026-07-27 17:32 ` [PATCH v6 06/15] s390/vfio-ap: Transition guest migration state from STOP to STOP_COPY Anthony Krowiak
2026-07-27 18:00   ` sashiko-bot
2026-07-27 17:32 ` [PATCH v6 07/15] s390/vfio-ap: File ops called to save the vfio device migration state Anthony Krowiak
2026-07-27 18:02   ` sashiko-bot
2026-08-04 17:07     ` Anthony Krowiak
2026-07-27 17:32 ` [PATCH v6 08/15] s390/vfio-ap: Transition device migration state from STOP to RESUMING Anthony Krowiak
2026-07-27 18:14   ` sashiko-bot
2026-07-28 22:29     ` Anthony Krowiak [this message]
2026-07-27 17:32 ` [PATCH v6 09/15] s390/vfio-ap: Add method to set a new guest AP configuration Anthony Krowiak
2026-07-27 18:11   ` sashiko-bot
2026-08-05 12:15     ` Anthony Krowiak
2026-07-27 17:32 ` [PATCH v6 10/15] s390/vfio-ap: File ops called to resume the vfio device migration Anthony Krowiak
2026-07-27 18:12   ` sashiko-bot
2026-08-04 15:33     ` Anthony Krowiak
2026-07-27 17:32 ` [PATCH v6 11/15] s390/vfio-ap: Transition device migration state to STOP Anthony Krowiak
2026-07-27 18:26   ` sashiko-bot
2026-07-27 17:32 ` [PATCH v6 12/15] s390/vfio-ap: Transition device migration state from STOP to RUNNING and vice versa Anthony Krowiak
2026-07-27 18:28   ` sashiko-bot
2026-08-05 13:14     ` Anthony Krowiak
2026-07-27 17:32 ` [PATCH v6 13/15] s390/vfio-ap: Callback to get the size of data to be migrated during guest migration Anthony Krowiak
2026-07-27 18:19   ` sashiko-bot
2026-07-30 11:26     ` Anthony Krowiak
2026-07-27 17:32 ` [PATCH v6 14/15] s390/vfio-ap: Add 'migratable' feature to sysfs 'features' attribute Anthony Krowiak
2026-07-27 18:45   ` sashiko-bot
2026-08-05 13:43     ` Anthony Krowiak
2026-07-27 17:32 ` [PATCH v6 15/15] s390/vfio-ap: Add live guest migration chapter to vfio-ap.rst Anthony Krowiak
2026-07-27 18:27   ` sashiko-bot

Reply instructions:

You may reply publicly to this message via plain-text email
using any one of the following methods:

* Save the following mbox file, import it into your mail client,
  and reply-to-all from there: mbox

  Avoid top-posting and favor interleaved quoting:
  https://en.wikipedia.org/wiki/Posting_style#Interleaved_style

* Reply using the --to, --cc, and --in-reply-to
  switches of git-send-email(1):

  git send-email \
    --in-reply-to=41149777-2df1-4225-ac68-c9038302aee4@linux.ibm.com \
    --to=akrowiak@linux.ibm.com \
    --cc=agordeev@linux.ibm.com \
    --cc=alex@shazbot.org \
    --cc=borntraeger@linux.ibm.com \
    --cc=frankja@linux.ibm.com \
    --cc=gor@linux.ibm.com \
    --cc=hca@linux.ibm.com \
    --cc=imbrenda@linux.ibm.com \
    --cc=jjherne@linux.ibm.com \
    --cc=kvm@vger.kernel.org \
    --cc=kwankhede@nvidia.com \
    --cc=linux-s390@vger.kernel.org \
    --cc=mjrosato@linux.ibm.com \
    --cc=pasic@linux.ibm.com \
    --cc=pbonzini@redhat.com \
    --cc=sashiko-reviews@lists.linux.dev \
    /path/to/YOUR_REPLY

  https://kernel.org/pub/software/scm/git/docs/git-send-email.html

* If your mail client supports setting the In-Reply-To header
  via mailto: links, try the mailto: link
Be sure your reply has a Subject: header at the top and a blank line before the message body.
This is a public inbox, see mirroring instructions
for how to clone and mirror all data and code used for this inbox