From: "Michael S. Tsirkin" <mst@redhat.com>
To: Xuan Zhuo <xuanzhuo@linux.alibaba.com>
Cc: Lange Tang <lange_tang@163.com>,
"edumazet@google.com" <edumazet@google.com>,
Jakub Kicinski <kuba@kernel.org>,
"virtualization@lists.linux.dev" <virtualization@lists.linux.dev>,
Tang Longjun <tanglongjun@kylinos.cn>,
"jasowang@redhat.com" <jasowang@redhat.com>
Subject: Re: Re:Re: [PATCH] virtio_net: disable cb when napi_schedule_prep fails during busy-poll
Date: Mon, 10 Aug 2026 03:51:48 -0400 [thread overview]
Message-ID: <20260810032647-mutt-send-email-mst@kernel.org> (raw)
In-Reply-To: <1781592565.1172295-1-xuanzhuo@linux.alibaba.com>
On Tue, Jun 16, 2026 at 02:49:25PM +0800, Xuan Zhuo wrote:
> On Tue, 16 Jun 2026 14:07:34 +0800 (CST), Lange Tang <lange_tang@163.com> wrote:
> > At 2026-06-16 11:27:12, "Xuan Zhuo" <xuanzhuo@linux.alibaba.com> wrote:
> > >On Tue, 16 Jun 2026 11:00:29 +0800 (CST), Lange Tang <lange_tang@163.com> wrote:
> > >> At 2026-06-15 18:01:40, "Xuan Zhuo" <xuanzhuo@linux.alibaba.com> wrote:
> > >> >On Mon, 15 Jun 2026 17:45:50 +0800, Longjun Tang <lange_tang@163.com> wrote:
> > >> >> From: Longjun Tang <tanglongjun@kylinos.cn>
> > >> >>
> > >> >> When busy-poll is active, napi_schedule_prep() returns false in
> > >> >> skb_recv_done(), so virtqueue_disable_cb() is skipped. The device
> > >> >> may keep firing irqs until the next poll round reaches
> > >> >> virtqueue_napi_complete(). If cb is enabled under busy-poll case,
> > >> >> it will lead to a large number of spurious interrupts. Explicitly
> > >> >> disable callbacks in this case to prevent spurious interrupts.
> > >> >>
> > >> >> Signed-off-by: Longjun Tang <tanglongjun@kylinos.cn>
> > >> >> ---
> > >> >> drivers/net/virtio_net.c | 2 ++
> > >> >> 1 file changed, 2 insertions(+)
> > >> >>
> > >> >> diff --git a/drivers/net/virtio_net.c b/drivers/net/virtio_net.c
> > >> >> index f4adcfee7a80..6d675fddc59b 100644
> > >> >> --- a/drivers/net/virtio_net.c
> > >> >> +++ b/drivers/net/virtio_net.c
> > >> >> @@ -728,6 +728,8 @@ static void virtqueue_napi_schedule(struct napi_struct *napi,
> > >> >> if (napi_schedule_prep(napi)) {
> > >> >> virtqueue_disable_cb(vq);
> > >> >> __napi_schedule(napi);
> > >> >> + } else if (test_bit(NAPI_STATE_IN_BUSY_POLL, &napi->state)) {
> > >> >> + virtqueue_disable_cb(vq);
> > >> >
> > >> >I see, but we should avoid checking NAPI_STATE_IN_BUSY_POLL directly in the
> > >> >drivers. The NIC driver should remain agnostic to busy polling. I think we need
> > >> >a better way, maybe we should rewrite virtqueue_napi_schedule instead.
> > >>
> > >> How about rewrite it like this?
> > >> static void virtqueue_napi_schedule(struct napi_struct *napi,
> > >> struct virtqueue *vq)
> > >> {
> > >> virtqueue_disable_cb(vq);
> > >> if (napi_schedule_prep(napi))
> > >> __napi_schedule(napi);
> > >> }
> > >> Any comments are welcome.
> > >
> > >
> > >Another CPU could be running NAPI and has just enabled the callbacks (cb).
> > >Meanwhile, this side unconditionally disables the cb. Since NAPI on the other
> > >CPU hasn't exited yet, the subsequent prep on this side fails, leaving no one to
> > >re-enable the cb.
> > >
> > >Thanks.
> >
> > Regarding the case you described, when NAPI on another CPU exits, the virtqueue_napi_complete func
> > will be executed to re-enable cb. and if there is still unconsumed data in the virtqueue, virtqueue_napi_schedule
> > will be called again to schedule NAPI.
> >
> > In summary, I think that the disable_cb and __napi_schedule within the virtqueue_napi_schedule func do not need to be bound together.
> >
> > Any comments are welcome. Thinks.
>
>
> <Your code>
> static void virtqueue_napi_schedule(struct napi_struct *napi,
> struct virtqueue *vq)
> {
>
> |static bool virtqueue_napi_complete(struct napi_struct *napi,
> | struct virtqueue *vq, int processed)
> |{
> | int opaque;
> |
> | opaque = virtqueue_enable_cb_prepare(vq);
> |
> virtqueue_disable_cb(vq); |
> if (napi_schedule_prep(napi)) |
> __napi_schedule(napi); |
> | if (napi_complete_done(napi, processed)) {
> | if (unlikely(virtqueue_poll(vq, opaque)))
> | virtqueue_napi_schedule(napi, vq);
> | else
> | return true; // return directly
> | } else {
> | virtqueue_disable_cb(vq);
> | }
> |
> | return false;
> |}
> }
>
> 1. new packets (notified by irq) are consumed by napi before virtqueue_napi_complete
> 2. poll is not called by irq, maybe xsk wake up. So irq is not disabled.
>
>
> Thanks.
I thought more about it and - so what? It's a rare race that
might result in more interrupts. Why is that an issue?
This would be a problem:
callbacks disabled
no NAPI
and I do not see how this gets there.
so maybe
virtqueue_disable_cb(vq);
napi_schedule(napi);
is exactly the right thing?
it's certainly simple.
>
> >
> > >
> > >
> > >> >
> > >> >
> > >> >> }
> > >> >> }
> > >> >>
> > >> >> --
> > >> >> 2.25.1
> > >> >>
> > >>
> >
prev parent reply other threads:[~2026-08-10 7:51 UTC|newest]
Thread overview: 10+ messages / expand[flat|nested] mbox.gz Atom feed top
2026-06-15 9:45 [PATCH] virtio_net: disable cb when napi_schedule_prep fails during busy-poll Longjun Tang
2026-06-15 10:01 ` Xuan Zhuo
2026-06-16 3:00 ` Lange Tang
2026-06-16 3:27 ` Xuan Zhuo
2026-06-16 6:07 ` Lange Tang
2026-06-16 6:49 ` Xuan Zhuo
2026-06-17 2:08 ` Lange Tang
2026-08-10 7:13 ` Michael S. Tsirkin
2026-08-10 9:20 ` Lange Tang
2026-08-10 7:51 ` Michael S. Tsirkin [this message]
Reply instructions:
You may reply publicly to this message via plain-text email
using any one of the following methods:
* Save the following mbox file, import it into your mail client,
and reply-to-all from there: mbox
Avoid top-posting and favor interleaved quoting:
https://en.wikipedia.org/wiki/Posting_style#Interleaved_style
* Reply using the --to, --cc, and --in-reply-to
switches of git-send-email(1):
git send-email \
--in-reply-to=20260810032647-mutt-send-email-mst@kernel.org \
--to=mst@redhat.com \
--cc=edumazet@google.com \
--cc=jasowang@redhat.com \
--cc=kuba@kernel.org \
--cc=lange_tang@163.com \
--cc=tanglongjun@kylinos.cn \
--cc=virtualization@lists.linux.dev \
--cc=xuanzhuo@linux.alibaba.com \
/path/to/YOUR_REPLY
https://kernel.org/pub/software/scm/git/docs/git-send-email.html
* If your mail client supports setting the In-Reply-To header
via mailto: links, try the mailto: link
Be sure your reply has a Subject: header at the top and a blank line
before the message body.
This is a public inbox, see mirroring instructions
for how to clone and mirror all data and code used for this inbox