From: Sean Christopherson <seanjc@google.com>
To: Yosry Ahmed <yosry@kernel.org>
Cc: Paolo Bonzini <pbonzini@redhat.com>,
Jim Mattson <jmattson@google.com>,
Maxim Levitsky <mlevitsk@redhat.com>,
Vitaly Kuznetsov <vkuznets@redhat.com>,
Tom Lendacky <thomas.lendacky@amd.com>,
kvm@vger.kernel.org, linux-kernel@vger.kernel.org
Subject: Re: [RFC PATCH v2 22/25] KVM: x86/mmu: Refactor kvm_mmu_invlpg() to allow skipping the gva flush
Date: Thu, 23 Jul 2026 08:23:32 -0700 [thread overview]
Message-ID: <amIx9Ki6gv-wgBnX@google.com> (raw)
In-Reply-To: <amGiAKsC3XrF6DyA@google.com>
On Thu, Jul 23, 2026, Yosry Ahmed wrote:
> On Wed, Jul 22, 2026 at 05:56:19PM -0700, Sean Christopherson wrote:
> > On Wed, Jul 22, 2026, Sean Christopherson wrote:
> > > On Tue, Jun 16, 2026, Yosry Ahmed wrote:
> > > > Refactor helpers out of kvm_mmu_invalidate_addr() and kvm_mmu_invlpg()
> > > > that take in an extra argument to skip the GVA flush.
> > > >
> > > > This will be used when invalidating GVAs in a different context than the
> > > > correct one (i.e. invalidating an L2 GVA from L1), so flushing the
> > > > current context would flush the wrong TLB entries.
> > > >
> > > > No functional change intended.
> > > >
> > > > Signed-off-by: Yosry Ahmed <yosry@kernel.org>
> > > > ---
> > > > arch/x86/kvm/mmu/mmu.c | 23 +++++++++++++++++------
> > > > 1 file changed, 17 insertions(+), 6 deletions(-)
> > > >
> > > > diff --git a/arch/x86/kvm/mmu/mmu.c b/arch/x86/kvm/mmu/mmu.c
> > > > index 65c35ed8f4a01..3feb75732f7b4 100644
> > > > --- a/arch/x86/kvm/mmu/mmu.c
> > > > +++ b/arch/x86/kvm/mmu/mmu.c
> > > > @@ -6615,15 +6615,15 @@ static void kvm_mmu_invalidate_addr_in_root(struct kvm_vcpu *vcpu,
> > > > write_unlock(&vcpu->kvm->mmu_lock);
> > > > }
> > > >
> > > > -void kvm_mmu_invalidate_addr(struct kvm_vcpu *vcpu, struct kvm_mmu *mmu,
> > > > - u64 addr, unsigned long roots)
> > > > +static void __kvm_mmu_invalidate_addr(struct kvm_vcpu *vcpu, struct kvm_mmu *mmu,
> > > > + u64 addr, unsigned long roots, bool flush_gva)
> > > > {
> > > > int i;
> > > >
> > > > WARN_ON_ONCE(roots & ~KVM_MMU_ROOTS_ALL);
> > > >
> > > > /* It's actually a GPA for vcpu->arch.guest_mmu. */
> > > > - if (mmu != &vcpu->arch.guest_mmu) {
> > > > + if (flush_gva && mmu != &vcpu->arch.guest_mmu) {
> > > > /* INVLPG on a non-canonical address is a NOP according to the SDM. */
> > > > if (is_noncanonical_invlpg_address(addr, vcpu))
> > > > return;
> > > > @@ -6642,9 +6642,15 @@ void kvm_mmu_invalidate_addr(struct kvm_vcpu *vcpu, struct kvm_mmu *mmu,
> > > > kvm_mmu_invalidate_addr_in_root(vcpu, mmu, addr, mmu->prev_roots[i].hpa);
> > > > }
> > > > }
> > > > +
> > > > +void kvm_mmu_invalidate_addr(struct kvm_vcpu *vcpu, struct kvm_mmu *mmu,
> > > > + u64 addr, unsigned long roots)
> > >
> > > Rather than kvm_mmu_invalidate_addr() for the wrapper, what if we call this
> > > kvm_mmu_invalidate_gva()? And then kvm_mmu_invlpg_gva(). Then we don't need
> > > to have the "in_root" version to a quad-underscores helper, and IMO it's more
> > > obvious what's different between the one-line wrappers and the inner helpers.
> >
> > Hrm, or maybe I'm not understanding what "flush_gva" means. At first glance, I
> > was assuming you were using it to differentiate between GVA and GPA, but IIUC,
> > it's literally skipping the flush for the current context, which just so happens
> > to be done only for GVAs. I'd still like to avoid the "in_root" helper, if at
> > all possible.
>
> Yes, it's skipping the TLB flush that is only needed for GVAs. The
> alternative I had in mind (but thought was worse) was to refactor the
> TLB flush part out of kvm_mmu_invalidate_addr() (or
> __kvm_mmu_invalidate_addr()) in this patch instead of adding a boolean,
> but this still requires adding a wrapper and the possibility of a
> quad-underscore helper.
>
> What's the main objection to kvm_mmu_invalidate_addr_in_root()?
I don't love the __kvm_mmu_invalidate_addr() => kvm_mmu_invalidate_addr_in_root()
callchain. It's not at all obvious that the in_root() helper shouldn't be called
directly. I don't hate it, but I do think we need better clarity on what all this
is doing.
E.g. when looking at __kvm_inject_emulated_page_fault(), since it hardcodes a
single root, it's a bit headscratching to use kvm_mmu_invalidate_addr() instead
of kvm_mmu_invalidate_addr_in_root.
Hmm, and arguably, the way invlpga_interception() handles the ASID is flat out
wrong. KVM doesn't need to flush *all* roots, rather it needs to flush L1 roots
for ASID=0, and L2 roots for ASID!=0. If we can figure out an elegant way to
express and handle that, it should naturally handle the "flush GVA" aspect.
Actually, isn't there a pre-existing over-flush when handling kvm_mmu_invpcid_gva()?
Oof, and a missed flush?
To fix the over-flush, I think we want this?
diff --git arch/x86/kvm/mmu/mmu.c arch/x86/kvm/mmu/mmu.c
index 6c13da942bfc..7f3e0eb33b29 100644
--- arch/x86/kvm/mmu/mmu.c
+++ arch/x86/kvm/mmu/mmu.c
@@ -6672,7 +6672,8 @@ void kvm_mmu_invalidate_addr(struct kvm_vcpu *vcpu, struct kvm_pagewalk *w,
if (is_noncanonical_invlpg_address(addr, vcpu))
return;
- kvm_x86_call(flush_tlb_gva)(vcpu, addr);
+ if (roots & KVM_MMU_ROOT_CURRENT)
+ kvm_x86_call(flush_tlb_gva)(vcpu, addr);
if (tdp_enabled)
return;
And that highlights the missed flush: if the PCID isn't the current PCID, then
flush_tlb_gva() neglects to flush the hardware TLB for the target PCID, which
could leave a stale entry in the TLB if the guest switches to the new PCID with
MOV CR3 + X86_CR3_PCID_NOFLUSH.
next prev parent reply other threads:[~2026-07-23 15:23 UTC|newest]
Thread overview: 62+ messages / expand[flat|nested] mbox.gz Atom feed top
2026-06-16 0:41 [RFC PATCH v2 00/25] Optimize nSVM TLB flushes Yosry Ahmed
2026-06-16 0:41 ` [RFC PATCH v2 01/25] KVM: nSVM: Flush the TLB after forcefully leaving nested Yosry Ahmed
2026-06-16 0:41 ` [RFC PATCH v2 02/25] KVM: SVM: Passthrough the number of supported ASIDs Yosry Ahmed
2026-07-14 13:20 ` Yosry Ahmed
2026-07-14 21:28 ` Jim Mattson
2026-07-14 21:43 ` Yosry Ahmed
2026-07-14 23:41 ` Sean Christopherson
2026-07-15 6:32 ` Jim Mattson
2026-07-15 17:45 ` Yosry Ahmed
2026-07-15 18:20 ` Jim Mattson
2026-07-15 19:12 ` Yosry Ahmed
2026-07-16 23:09 ` Jim Mattson
2026-07-22 22:09 ` Sean Christopherson
2026-07-22 22:39 ` Jim Mattson
2026-07-23 0:27 ` Sean Christopherson
2026-07-23 2:46 ` Jim Mattson
2026-07-23 12:24 ` Jim Mattson
2026-07-23 14:07 ` Sean Christopherson
2026-07-23 15:57 ` Jim Mattson
2026-07-23 16:32 ` Yosry Ahmed
2026-07-23 16:54 ` Jim Mattson
2026-07-23 16:57 ` Yosry Ahmed
2026-07-23 17:07 ` Jim Mattson
2026-07-23 17:26 ` Sean Christopherson
2026-07-23 17:34 ` Yosry Ahmed
2026-07-15 17:41 ` Yosry Ahmed
2026-06-16 0:41 ` [RFC PATCH v2 03/25] KVM: VMX: Generalize VPID allocation to be vendor-neutral Yosry Ahmed
2026-07-22 22:26 ` Sean Christopherson
2026-07-22 22:36 ` Yosry Ahmed
2026-07-23 13:26 ` Sean Christopherson
2026-06-16 0:41 ` [RFC PATCH v2 04/25] KVM: x86/mmu: Support specifying a minimum TLB tag Yosry Ahmed
2026-07-23 0:28 ` Sean Christopherson
2026-07-23 5:00 ` Yosry Ahmed
2026-07-23 14:15 ` Sean Christopherson
2026-06-16 0:41 ` [RFC PATCH v2 05/25] KVM: SVM: Add helpers to set/clear ASID flush in VMCB Yosry Ahmed
2026-06-16 0:41 ` [RFC PATCH v2 06/25] KVM: SVM: Fallback to flush everything if FLUSHBYASID is not available Yosry Ahmed
2026-07-23 0:29 ` Sean Christopherson
2026-06-16 0:41 ` [RFC PATCH v2 07/25] KVM: SVM: Duplicate pre-run ASID check for SEV and non-SEV guests Yosry Ahmed
2026-06-16 0:41 ` [RFC PATCH v2 08/25] KVM: SEV: Stop using per-vCPU ASID for SEV VMs Yosry Ahmed
2026-06-16 0:41 ` [RFC PATCH v2 09/25] KVM: SVM: Use a static ASID per vCPU Yosry Ahmed
2026-06-16 0:41 ` [RFC PATCH v2 10/25] KVM: nSVM: Add a placeholder ASID for L2 Yosry Ahmed
2026-06-16 0:41 ` [RFC PATCH v2 11/25] KVM: x86: hyper-v: Rename kvm_hv_vcpu_purge_flush_tlb() Yosry Ahmed
2026-06-16 0:41 ` [RFC PATCH v2 12/25] KVM: x86: hyper-v: Allow puring all TLB flush FIFOs Yosry Ahmed
2026-06-16 0:41 ` [RFC PATCH v2 13/25] KVM: nSVM: Flush both L1 and L2 ASIDs on KVM_REQ_TLB_FLUSH Yosry Ahmed
2026-06-16 0:41 ` [RFC PATCH v2 14/25] KVM: nSVM: Move svm_switch_vmcb() to nested.c Yosry Ahmed
2026-06-16 0:41 ` [RFC PATCH v2 15/25] KVM: nSVM: Call nested_svm_transition_tlb_flush() on every VMCB switch Yosry Ahmed
2026-07-23 0:42 ` Sean Christopherson
2026-07-23 5:07 ` Yosry Ahmed
2026-06-16 0:41 ` [RFC PATCH v2 16/25] KVM: nSVM: Split nested_svm_transition_tlb_flush() into entry/exit fns Yosry Ahmed
2026-06-16 0:41 ` [RFC PATCH v2 17/25] KVM: nSVM: Service local TLB flushes before nested transitions Yosry Ahmed
2026-06-16 0:41 ` [RFC PATCH v2 18/25] KVM: nSVM: Handle nested TLB flush requests through TLB_CONTROL Yosry Ahmed
2026-06-16 0:41 ` [RFC PATCH v2 19/25] KVM: nSVM: Flush the TLB if L1 changes L2's ASID in vmcb12 Yosry Ahmed
2026-06-16 0:41 ` [RFC PATCH v2 20/25] KVM: nSVM: Do not reset TLB_CONTROL in vmcb02 on nested VM-Enter Yosry Ahmed
2026-06-16 0:41 ` [RFC PATCH v2 21/25] KVM: x86/mmu: rename __kvm_mmu_invalidate_addr() Yosry Ahmed
2026-06-16 0:41 ` [RFC PATCH v2 22/25] KVM: x86/mmu: Refactor kvm_mmu_invlpg() to allow skipping the gva flush Yosry Ahmed
2026-07-23 0:53 ` Sean Christopherson
2026-07-23 0:56 ` Sean Christopherson
2026-07-23 5:11 ` Yosry Ahmed
2026-07-23 15:23 ` Sean Christopherson [this message]
2026-06-16 0:41 ` [RFC PATCH v2 23/25] KVM: nSVM: Flush L2's ASID when emulating INVLPGA Yosry Ahmed
2026-06-16 0:41 ` [RFC PATCH v2 24/25] KVM: nSVM: Use different ASIDs for L1 and L2 Yosry Ahmed
2026-06-16 0:41 ` [RFC PATCH v2 25/25] DO NOT MERGE: Add nested_tlb_force_flush Yosry Ahmed
Reply instructions:
You may reply publicly to this message via plain-text email
using any one of the following methods:
* Save the following mbox file, import it into your mail client,
and reply-to-all from there: mbox
Avoid top-posting and favor interleaved quoting:
https://en.wikipedia.org/wiki/Posting_style#Interleaved_style
* Reply using the --to, --cc, and --in-reply-to
switches of git-send-email(1):
git send-email \
--in-reply-to=amIx9Ki6gv-wgBnX@google.com \
--to=seanjc@google.com \
--cc=jmattson@google.com \
--cc=kvm@vger.kernel.org \
--cc=linux-kernel@vger.kernel.org \
--cc=mlevitsk@redhat.com \
--cc=pbonzini@redhat.com \
--cc=thomas.lendacky@amd.com \
--cc=vkuznets@redhat.com \
--cc=yosry@kernel.org \
/path/to/YOUR_REPLY
https://kernel.org/pub/software/scm/git/docs/git-send-email.html
* If your mail client supports setting the In-Reply-To header
via mailto: links, try the mailto: link
Be sure your reply has a Subject: header at the top and a blank line
before the message body.
This is a public inbox, see mirroring instructions
for how to clone and mirror all data and code used for this inbox