From: "Huang, Honglei" <honghuan@amd.com>
To: "Michael S. Tsirkin" <mst@redhat.com>,
"Alex Bennée" <alex.bennee@linaro.org>,
"Dmitry Osipenko" <dmitry.osipenko@collabora.com>,
"Akihiko Odaki" <odaki@rsg.ci.i.u-tokyo.ac.jp>,
"Marc-André Lureau" <marcandre.lureau@redhat.com>,
"Stefano Garzarella" <sgarzare@redhat.com>,
"Gerd Hoffmann" <kraxel@redhat.com>,
"David Airlie" <airlied@redhat.com>,
"Peter Maydell" <peter.maydell@linaro.org>
Cc: qemu-devel@nongnu.org, virtio-comment@lists.oasis-open.org,
dri-devel@lists.freedesktop.org, virtualization@lists.linux.dev,
Honglei Huang <honglei1.huang@amd.com>,
Huang Rui <Ray.Huang@amd.com>
Subject: About new backend for GPU compute ROCm in qemu
Date: Mon, 17 Aug 2026 11:19:44 +0800 [thread overview]
Message-ID: <ba0b1bbf-5ebf-4ac3-9f7b-8740e3bc283b@amd.com> (raw)
Hi Michael, Alex, Dmitry, Akihiko,
I'm bringing AMD GPU compute ROCm based on virtio. I posted a ROCm over
virtio
implementation to virglrenderer nine months ago (MR !1568 [1]). The ROCm
side has
been supportted by ROCm offical.
Current implementation is a virtio gpu context type capset handled inside
virglrenderer, sharing the display path. That's an awkward fit, many
compute GPUs have no display engine at all.
Beyond that, sharing the display path is increasingly painful:
- Compute hammers the queues more than graphics, so sharing
virtio gpu's single control queue with display/virgl causes contention
and display stutter.
- Compute contexts need far more blob / shared memory than a display one.
- Maybe needs a wider ROCm / compute stack, cause the render model
fits poorly:
rocprofiler (PC sampling, SQTT/SPM, counters, high bandwidth streams)
and ROCgdb (wave control, address watch, async exceptions an
out of band channel that must not block display).
- Events, faults and GPU reset/SMI are async and don't map onto fences.
- All of this is hard to extend cleanly inside a display capset.
On the QEMU/host side, would something like this be OK? One step, two parts:
- a dedicated headless virtio gpu instance for compute.
- that instance served by a separate ROCm backend library loaded
in-process by QEMU.
That reuses the existing pluggable backend model, a second virtio gpu + a
backend library. It doesn't add dedicated queues for debug/profiling
currently.
Waiting for reply and happy to share more detail. Thanks!
[1] https://gitlab.freedesktop.org/virgl/virglrenderer/-/merge_requests/1568
Regards,
Honglei
next reply other threads:[~2026-08-17 3:19 UTC|newest]
Thread overview: 5+ messages / expand[flat|nested] mbox.gz Atom feed top
2026-08-17 3:19 Huang, Honglei [this message]
2026-08-17 9:06 ` About new backend for GPU compute ROCm in qemu Alex Bennée
2026-08-17 12:46 ` Huang, Honglei
2026-08-17 11:44 ` Akihiko Odaki
2026-08-17 13:44 ` Huang, Honglei
Reply instructions:
You may reply publicly to this message via plain-text email
using any one of the following methods:
* Save the following mbox file, import it into your mail client,
and reply-to-all from there: mbox
Avoid top-posting and favor interleaved quoting:
https://en.wikipedia.org/wiki/Posting_style#Interleaved_style
* Reply using the --to, --cc, and --in-reply-to
switches of git-send-email(1):
git send-email \
--in-reply-to=ba0b1bbf-5ebf-4ac3-9f7b-8740e3bc283b@amd.com \
--to=honghuan@amd.com \
--cc=Ray.Huang@amd.com \
--cc=airlied@redhat.com \
--cc=alex.bennee@linaro.org \
--cc=dmitry.osipenko@collabora.com \
--cc=dri-devel@lists.freedesktop.org \
--cc=honglei1.huang@amd.com \
--cc=kraxel@redhat.com \
--cc=marcandre.lureau@redhat.com \
--cc=mst@redhat.com \
--cc=odaki@rsg.ci.i.u-tokyo.ac.jp \
--cc=peter.maydell@linaro.org \
--cc=qemu-devel@nongnu.org \
--cc=sgarzare@redhat.com \
--cc=virtio-comment@lists.oasis-open.org \
--cc=virtualization@lists.linux.dev \
/path/to/YOUR_REPLY
https://kernel.org/pub/software/scm/git/docs/git-send-email.html
* If your mail client supports setting the In-Reply-To header
via mailto: links, try the mailto: link
Be sure your reply has a Subject: header at the top and a blank line
before the message body.
This is an external index of several public inboxes,
see mirroring instructions on how to clone and mirror
all data and code used by this external index.