From: Stefan Hajnoczi <stefanha@redhat.com>
To: Eugenio Perez Martin <eperezma@redhat.com>
Cc: kvm list <kvm@vger.kernel.org>,
"Michael S. Tsirkin" <mst@redhat.com>,
qemu-level <qemu-devel@nongnu.org>,
Daniel Daly <dandaly0@gmail.com>,
virtualization@lists.linux-foundation.org,
Liran Alon <liralon@gmail.com>, Eli Cohen <eli@mellanox.com>,
Nitin Shrivastav <nitin.shrivastav@broadcom.com>,
Alex Barba <alex.barba@broadcom.com>,
Christophe Fontaine <cfontain@redhat.com>,
Lee Ballard <ballle98@gmail.com>,
Lars Ganrot <lars.ganrot@gmail.com>,
Rob Miller <rob.miller@broadcom.com>,
Howard Cai <howard.cai@gmail.com>,
Parav Pandit <parav@mellanox.com>, vm <vmireyno@marvell.com>,
Salil Mehta <mehta.salil.lnk@gmail.com>,
Stephen Finucane <stephenfin@redhat.com>,
Xiao W Wang <xiao.w.wang@intel.com>,
Sean Mooney <smooney@redhat.com>,
Jim Harford <jim.harford@broadcom.com>,
Dmytro Kazantsev <dmytro.kazantsev@gmail.com>,
Siwei Liu <loseweigh@gmail.com>,
Harpreet Singh Anand <hanand@xilinx.com>,
Michael Lilja <ml@napatech.com>, Max Gurtovoy <maxgu14@gmail.com>
Subject: Re: [RFC PATCH 07/27] vhost: Route guest->host notification through qemu
Date: Thu, 10 Dec 2020 11:50:52 +0000 [thread overview]
Message-ID: <20201210115052.GG416119@stefanha-x1.localdomain> (raw)
In-Reply-To: <CAJaqyWfiMsRP9FgSv7cOj=3jHx=DJS7hRJTMbRcTTHHWng0eKg@mail.gmail.com>
[-- Attachment #1.1: Type: text/plain, Size: 3450 bytes --]
On Wed, Dec 09, 2020 at 06:08:14PM +0100, Eugenio Perez Martin wrote:
> On Mon, Dec 7, 2020 at 6:42 PM Stefan Hajnoczi <stefanha@gmail.com> wrote:
> > On Fri, Nov 20, 2020 at 07:50:45PM +0100, Eugenio Pérez wrote:
> > > +{
> > > + struct vhost_vring_file file = {
> > > + .index = idx
> > > + };
> > > + VirtQueue *vq = virtio_get_queue(dev->vdev, idx);
> > > + VhostShadowVirtqueue *svq;
> > > + int r;
> > > +
> > > + svq = g_new0(VhostShadowVirtqueue, 1);
> > > + svq->vq = vq;
> > > +
> > > + r = event_notifier_init(&svq->hdev_notifier, 0);
> > > + assert(r == 0);
> > > +
> > > + file.fd = event_notifier_get_fd(&svq->hdev_notifier);
> > > + r = dev->vhost_ops->vhost_set_vring_kick(dev, &file);
> > > + assert(r == 0);
> > > +
> > > + return svq;
> > > +}
> >
> > I guess there are assumptions about the status of the device? Does the
> > virtqueue need to be disabled when this function is called?
> >
>
> Yes. Maybe an assertion checking the notification state?
Sounds good.
> > > +
> > > +static int vhost_sw_live_migration_stop(struct vhost_dev *dev)
> > > +{
> > > + int idx;
> > > +
> > > + vhost_dev_enable_notifiers(dev, dev->vdev);
> > > + for (idx = 0; idx < dev->nvqs; ++idx) {
> > > + vhost_sw_lm_shadow_vq_free(dev->sw_lm_shadow_vq[idx]);
> > > + }
> > > +
> > > + return 0;
> > > +}
> > > +
> > > +static int vhost_sw_live_migration_start(struct vhost_dev *dev)
> > > +{
> > > + int idx;
> > > +
> > > + for (idx = 0; idx < dev->nvqs; ++idx) {
> > > + dev->sw_lm_shadow_vq[idx] = vhost_sw_lm_shadow_vq(dev, idx);
> > > + }
> > > +
> > > + vhost_dev_disable_notifiers(dev, dev->vdev);
> >
> > There is a race condition if the guest kicks the vq while this is
> > happening. The shadow vq hdev_notifier needs to be set so the vhost
> > device checks the virtqueue for requests that slipped in during the
> > race window.
> >
>
> I'm not sure if I follow you. If I understand correctly,
> vhost_dev_disable_notifiers calls virtio_bus_cleanup_host_notifier,
> and the latter calls virtio_queue_host_notifier_read. That's why the
> documentation says "This might actually run the qemu handlers right
> away, so virtio in qemu must be completely setup when this is
> called.". Am I missing something?
There are at least two cases:
1. Virtqueue kicks that come before vhost_dev_disable_notifiers().
vhost_dev_disable_notifiers() notices that and calls
virtio_queue_notify_vq(). Will handle_sw_lm_vq() be invoked or is the
device's vq handler function invoked?
2. Virtqueue kicks that come in after vhost_dev_disable_notifiers()
returns. We hold the QEMU global mutex so the vCPU thread cannot
perform MMIO/PIO dispatch immediately. The vCPU thread's
ioctl(KVM_RUN) has already returned and will dispatch dispatch the
MMIO/PIO access inside QEMU as soon as the global mutex is released.
In other words, we're not taking the kvm.ko ioeventfd path but
memory_region_dispatch_write_eventfds() should signal the ioeventfd
that is registered at the time the dispatch occurs. Is that eventfd
handled by handle_sw_lm_vq()?
Neither of these cases are obvious from the code. At least comments
would help but I suspect restructuring the code so the critical
ioeventfd state changes happen in a sequence would make it even clearer.
[-- Attachment #1.2: signature.asc --]
[-- Type: application/pgp-signature, Size: 488 bytes --]
[-- Attachment #2: Type: text/plain, Size: 183 bytes --]
_______________________________________________
Virtualization mailing list
Virtualization@lists.linux-foundation.org
https://lists.linuxfoundation.org/mailman/listinfo/virtualization
next prev parent reply other threads:[~2020-12-10 11:51 UTC|newest]
Thread overview: 29+ messages / expand[flat|nested] mbox.gz Atom feed top
[not found] <20201120185105.279030-1-eperezma@redhat.com>
2020-11-25 7:08 ` [RFC PATCH 00/27] vDPA software assisted live migration Jason Wang
[not found] ` <CAJaqyWf+6yoMHJuLv=QGLMP4egmdm722=V2kKJ_aiQAfCCQOFw@mail.gmail.com>
2020-11-26 3:07 ` Jason Wang
[not found] ` <20201120185105.279030-24-eperezma@redhat.com>
2020-11-27 15:29 ` [RFC PATCH 23/27] vhost: unmap qemu's shadow virtqueues on sw " Stefano Garzarella
2020-11-27 15:44 ` [RFC PATCH 00/27] vDPA software assisted " Stefano Garzarella
[not found] ` <20201120185105.279030-3-eperezma@redhat.com>
2020-12-07 16:19 ` [RFC PATCH 02/27] vhost: Add device callback in vhost_migration_log Stefan Hajnoczi
[not found] ` <20201120185105.279030-5-eperezma@redhat.com>
2020-12-07 16:43 ` [RFC PATCH 04/27] vhost: add vhost_kernel_set_vring_enable Stefan Hajnoczi
[not found] ` <CAJaqyWd5oAJ4kJOhyDz+1KNvwzqJi3NO+5Z7X6W5ju2Va=LTMQ@mail.gmail.com>
2020-12-09 16:08 ` Stefan Hajnoczi
[not found] ` <20201120185105.279030-6-eperezma@redhat.com>
2020-12-07 16:52 ` [RFC PATCH 05/27] vhost: Add hdev->dev.sw_lm_vq_handler Stefan Hajnoczi
[not found] ` <CAJaqyWfSUHD0MU=1yfU1N6pZ4TU7prxyoG6NY-VyNGt=MO9H4g@mail.gmail.com>
2020-12-10 11:30 ` Stefan Hajnoczi
[not found] ` <20201120185105.279030-7-eperezma@redhat.com>
2020-12-07 16:58 ` [RFC PATCH 06/27] virtio: Add virtio_queue_get_used_notify_split Stefan Hajnoczi
[not found] ` <CAJaqyWc4oLzL02GKpPSwEGRxK+UxjOGBAPLzrgrgKRZd9C81GA@mail.gmail.com>
2021-03-02 11:22 ` Stefan Hajnoczi
[not found] ` <CAJaqyWd0iRUUW5Hu=U3mwQ4f43kA=bse3EkN4+QauFR4BJwObQ@mail.gmail.com>
2021-03-08 10:46 ` Stefan Hajnoczi
[not found] ` <20201120185105.279030-8-eperezma@redhat.com>
2020-12-07 17:42 ` [RFC PATCH 07/27] vhost: Route guest->host notification through qemu Stefan Hajnoczi
[not found] ` <CAJaqyWfiMsRP9FgSv7cOj=3jHx=DJS7hRJTMbRcTTHHWng0eKg@mail.gmail.com>
2020-12-10 11:50 ` Stefan Hajnoczi [this message]
[not found] ` <20201120185105.279030-9-eperezma@redhat.com>
2020-12-08 7:20 ` [RFC PATCH 08/27] vhost: Add a flag for software assisted Live Migration Stefan Hajnoczi
[not found] ` <20201120185105.279030-10-eperezma@redhat.com>
2020-12-08 7:34 ` [RFC PATCH 09/27] vhost: Route host->guest notification through qemu Stefan Hajnoczi
[not found] ` <20201120185105.279030-11-eperezma@redhat.com>
2020-12-08 7:49 ` [RFC PATCH 10/27] vhost: Allocate shadow vring Stefan Hajnoczi
2020-12-08 8:17 ` Stefan Hajnoczi
[not found] ` <20201120185105.279030-14-eperezma@redhat.com>
2020-12-08 8:16 ` [RFC PATCH 13/27] vhost: Send buffers to device Stefan Hajnoczi
[not found] ` <CAJaqyWf13ta5MtzmTUz2N5XnQ+ebqFPYAivdggL64LEQAf=y+A@mail.gmail.com>
2020-12-10 11:55 ` Stefan Hajnoczi
[not found] ` <20201120185105.279030-17-eperezma@redhat.com>
2020-12-08 8:25 ` [RFC PATCH 16/27] virtio: Expose virtqueue_alloc_element Stefan Hajnoczi
[not found] ` <CAJaqyWdN7iudf8mDN4k4Fs9j1x+ztZARuBbinPHD3ZQSMH1pyQ@mail.gmail.com>
2020-12-10 11:57 ` Stefan Hajnoczi
[not found] ` <20201120185105.279030-19-eperezma@redhat.com>
2020-12-08 8:41 ` [RFC PATCH 18/27] vhost: add vhost_vring_poll_rcu Stefan Hajnoczi
[not found] ` <20201120185105.279030-21-eperezma@redhat.com>
2020-12-08 8:50 ` [RFC PATCH 20/27] vhost: Return used buffers Stefan Hajnoczi
[not found] ` <20201120185105.279030-25-eperezma@redhat.com>
2020-12-08 9:02 ` [RFC PATCH 24/27] vhost: iommu changes Stefan Hajnoczi
2020-12-08 9:37 ` [RFC PATCH 00/27] vDPA software assisted live migration Stefan Hajnoczi
2020-12-09 9:26 ` Jason Wang
2020-12-09 15:57 ` Stefan Hajnoczi
2020-12-10 9:12 ` Jason Wang
Reply instructions:
You may reply publicly to this message via plain-text email
using any one of the following methods:
* Save the following mbox file, import it into your mail client,
and reply-to-all from there: mbox
Avoid top-posting and favor interleaved quoting:
https://en.wikipedia.org/wiki/Posting_style#Interleaved_style
* Reply using the --to, --cc, and --in-reply-to
switches of git-send-email(1):
git send-email \
--in-reply-to=20201210115052.GG416119@stefanha-x1.localdomain \
--to=stefanha@redhat.com \
--cc=alex.barba@broadcom.com \
--cc=ballle98@gmail.com \
--cc=cfontain@redhat.com \
--cc=dandaly0@gmail.com \
--cc=dmytro.kazantsev@gmail.com \
--cc=eli@mellanox.com \
--cc=eperezma@redhat.com \
--cc=hanand@xilinx.com \
--cc=howard.cai@gmail.com \
--cc=jim.harford@broadcom.com \
--cc=kvm@vger.kernel.org \
--cc=lars.ganrot@gmail.com \
--cc=liralon@gmail.com \
--cc=loseweigh@gmail.com \
--cc=maxgu14@gmail.com \
--cc=mehta.salil.lnk@gmail.com \
--cc=ml@napatech.com \
--cc=mst@redhat.com \
--cc=nitin.shrivastav@broadcom.com \
--cc=parav@mellanox.com \
--cc=qemu-devel@nongnu.org \
--cc=rob.miller@broadcom.com \
--cc=smooney@redhat.com \
--cc=stephenfin@redhat.com \
--cc=virtualization@lists.linux-foundation.org \
--cc=vmireyno@marvell.com \
--cc=xiao.w.wang@intel.com \
/path/to/YOUR_REPLY
https://kernel.org/pub/software/scm/git/docs/git-send-email.html
* If your mail client supports setting the In-Reply-To header
via mailto: links, try the mailto: link
Be sure your reply has a Subject: header at the top and a blank line
before the message body.
This is a public inbox, see mirroring instructions
for how to clone and mirror all data and code used for this inbox;
as well as URLs for NNTP newsgroup(s).