From: "Cédric Le Goater" <clg@redhat.com>
To: mhonap@nvidia.com, alex@shazbot.org, ankita@nvidia.com,
jic23@kernel.org, dave.jiang@intel.com,
alejandro.lucero-palau@amd.com, smadhavan@nvidia.com,
pierrick.bouvier@oss.qualcomm.com, mst@redhat.com,
imammedo@redhat.com, anisinha@redhat.com, pbonzini@redhat.com,
eric.auger@redhat.com, peter.maydell@linaro.org,
richard.henderson@linaro.org, cohuck@redhat.com
Cc: kjaju@nvidia.com, vsethi@nvidia.com, zhiw@nvidia.com,
qemu-devel@nongnu.org, qemu-arm@nongnu.org,
linux-cxl@vger.kernel.org
Subject: Re: [PATCH v2 05/10] hw/vfio/pci: Back the CXL memory with a RAM-device region
Date: Thu, 17 Sep 2026 15:56:13 +0200 [thread overview]
Message-ID: <0991dccb-a123-471c-8a72-e4194da6c596@redhat.com> (raw)
In-Reply-To: <20260916184412.3825713-6-mhonap@nvidia.com>
On 9/16/26 20:44, mhonap@nvidia.com wrote:
> From: Manish Honap <mhonap@nvidia.com>
>
> mmap the kernel's HDM memory region so its host physical pages back a
> RAM-device MemoryRegion. It stays out of the guest address space until
> the guest commits its decoder and QEMU learns the GPA to place it at.
>
> AI-used-for: code (prototype)
> Signed-off-by: Manish Honap <mhonap@nvidia.com>
> ---
> hw/vfio/pci.c | 34 ++++++++++++++++++++++++++++++++++
> hw/vfio/pci.h | 1 +
> 2 files changed, 35 insertions(+)
>
> diff --git a/hw/vfio/pci.c b/hw/vfio/pci.c
> index 10b13f200d..4716266595 100644
> --- a/hw/vfio/pci.c
> +++ b/hw/vfio/pci.c
> @@ -3221,8 +3221,11 @@ bool vfio_pci_populate_device(VFIOPCIDevice *vdev, Error **errp)
> return true;
> }
>
> +static void vfio_cxl_teardown(VFIOPCIDevice *vdev);
> +
> void vfio_pci_put_device(VFIOPCIDevice *vdev)
> {
> + vfio_cxl_teardown(vdev);
> vfio_display_finalize(vdev);
> vfio_bars_finalize(vdev);
> vfio_cpr_pci_unregister_device(vdev);
> @@ -3696,11 +3699,42 @@ static bool vfio_cxl_setup(VFIOPCIDevice *vdev, Error **errp)
> return false;
> }
>
> + /*
> + * The HDM memory is host physical. Set up the region, which installs the
> + * fd read/write path, and mmap it for direct guest access; it is added to
> + * the guest address space only once the guest commits its endpoint decoder.
> + */
> + if (vfio_region_setup(OBJECT(vdev), vbasedev, &cxl->mem_region,
> + cxl->mem_region_index, "cxl-mem", errp)) {
> + return false;
> + }
> + if (vfio_region_mmap(&cxl->mem_region)) {
> + /*
> + * Without mmap the region falls back to the kernel's fd read/write
> + * path, which works but traps every access. Warn rather than fail.
> + */
> + warn_report("vfio-cxl: %s: failed to mmap the HDM memory region; "
> + "performance may be slow", vbasedev->name);
> + }
vfio_cxl_setup calls vfio_region_setup() then vfio_region_mmap()
back-to-back when other 'normal' BARs split these calls across
vfio_populate_device() and vfio_bar_register(). I guess it is fine
for CXL since it is not a PCI BAR.
> cxl->enabled = true;
>
> return true;
> }
>
> +static void vfio_cxl_teardown(VFIOPCIDevice *vdev)
> +{
> + VFIOCXL *cxl = &vdev->cxl;
> +
> + if (!cxl->enabled) {
> + return;
> + }
> + if (cxl->mem_region.mem) {
> + vfio_region_exit(&cxl->mem_region);
> + vfio_region_finalize(&cxl->mem_region);
However, the tear down should be split between :
1. vfio_exitfn (unrealize)
drops mmaps, removes subregions and removes references so the
MR refcount can reach zero.
2. vfio_pci_finalize (instance_finalize)
frees the MemoryRegion.
Thanks,
C.
> + }
> +}
> +
> static void vfio_pci_realize(PCIDevice *pdev, Error **errp)
> {
> ERRP_GUARD();
> diff --git a/hw/vfio/pci.h b/hw/vfio/pci.h
> index 7fdd695704..06d15e807d 100644
> --- a/hw/vfio/pci.h
> +++ b/hw/vfio/pci.h
> @@ -134,6 +134,7 @@ typedef struct VFIOCXL {
> uint32_t comp_bar; /* component BAR carrying that block */
> uint64_t hdm_offset; /* block offset within the component BAR */
> uint64_t dpa_size; /* size of the HDM memory region */
> + VFIORegion mem_region; /* HDM memory, mapped at committed GPA */
> } VFIOCXL;
>
> struct VFIOPCIDevice {
next prev parent reply other threads:[~2026-09-17 13:56 UTC|newest]
Thread overview: 27+ messages / expand[flat|nested] mbox.gz Atom feed top
2026-09-16 18:44 [PATCH v2 00/10] QEMU: CXL Type-2 device passthrough via vfio-pci mhonap
2026-09-16 18:44 ` [PATCH v2 01/10] linux-headers: Update vfio.h for CXL Type-2 passthrough mhonap
2026-09-16 18:44 ` [PATCH v2 02/10] hw/vfio/region: Add vfio_region_setup_with_ops() mhonap
2026-09-17 13:27 ` Cédric Le Goater
2026-09-21 10:48 ` Manish Honap
2026-09-16 18:44 ` [PATCH v2 03/10] hw/vfio/pci: Detect a CXL Type-2 device and read its geometry mhonap
2026-09-17 16:04 ` Cédric Le Goater
2026-09-21 10:49 ` Manish Honap
2026-09-16 18:44 ` [PATCH v2 04/10] hw/vfio/pci: Enforce the passthrough topology for a CXL device mhonap
2026-09-17 13:20 ` Cédric Le Goater
2026-09-21 10:45 ` Manish Honap
2026-09-16 18:44 ` [PATCH v2 05/10] hw/vfio/pci: Back the CXL memory with a RAM-device region mhonap
2026-09-17 13:56 ` Cédric Le Goater [this message]
2026-09-21 10:55 ` Manish Honap
2026-09-16 18:44 ` [PATCH v2 06/10] hw/vfio/pci: Bind a CXL device to its fixed memory window mhonap
2026-09-17 16:00 ` Cédric Le Goater
2026-09-21 11:03 ` Manish Honap
2026-09-16 18:44 ` [PATCH v2 07/10] hw/vfio/pci: Map the CXL memory on the guest decoder commit mhonap
2026-09-18 16:08 ` Cédric Le Goater
2026-09-21 11:14 ` Manish Honap
2026-09-20 9:14 ` Junjie Cao
2026-09-21 11:17 ` Manish Honap
2026-09-16 18:44 ` [PATCH v2 08/10] docs/cxl: Document CXL Type-2 device passthrough mhonap
2026-09-16 18:44 ` [PATCH v2 09/10] hw/arm/smmu-common: Allow pxb-cxl as an SMMUv3 primary bus mhonap
2026-09-16 18:44 ` [PATCH v2 10/10] hw/pci-host: Emit a _DSM on pxb-cxl to preserve firmware PCI config mhonap
2026-09-20 9:14 ` Junjie Cao
2026-09-21 11:20 ` Manish Honap
Reply instructions:
You may reply publicly to this message via plain-text email
using any one of the following methods:
* Save the following mbox file, import it into your mail client,
and reply-to-all from there: mbox
Avoid top-posting and favor interleaved quoting:
https://en.wikipedia.org/wiki/Posting_style#Interleaved_style
* Reply using the --to, --cc, and --in-reply-to
switches of git-send-email(1):
git send-email \
--in-reply-to=0991dccb-a123-471c-8a72-e4194da6c596@redhat.com \
--to=clg@redhat.com \
--cc=alejandro.lucero-palau@amd.com \
--cc=alex@shazbot.org \
--cc=anisinha@redhat.com \
--cc=ankita@nvidia.com \
--cc=cohuck@redhat.com \
--cc=dave.jiang@intel.com \
--cc=eric.auger@redhat.com \
--cc=imammedo@redhat.com \
--cc=jic23@kernel.org \
--cc=kjaju@nvidia.com \
--cc=linux-cxl@vger.kernel.org \
--cc=mhonap@nvidia.com \
--cc=mst@redhat.com \
--cc=pbonzini@redhat.com \
--cc=peter.maydell@linaro.org \
--cc=pierrick.bouvier@oss.qualcomm.com \
--cc=qemu-arm@nongnu.org \
--cc=qemu-devel@nongnu.org \
--cc=richard.henderson@linaro.org \
--cc=smadhavan@nvidia.com \
--cc=vsethi@nvidia.com \
--cc=zhiw@nvidia.com \
/path/to/YOUR_REPLY
https://kernel.org/pub/software/scm/git/docs/git-send-email.html
* If your mail client supports setting the In-Reply-To header
via mailto: links, try the mailto: link
Be sure your reply has a Subject: header at the top and a blank line
before the message body.
This is a public inbox, see mirroring instructions
for how to clone and mirror all data and code used for this inbox