The Linux Kernel Mailing List
 help / color / mirror / Atom feed
From: Sean Christopherson <seanjc@google.com>
To: Yosry Ahmed <yosry@kernel.org>
Cc: Paolo Bonzini <pbonzini@redhat.com>,
	Jim Mattson <jmattson@google.com>,
	 Maxim Levitsky <mlevitsk@redhat.com>,
	Vitaly Kuznetsov <vkuznets@redhat.com>,
	 Tom Lendacky <thomas.lendacky@amd.com>,
	kvm@vger.kernel.org, linux-kernel@vger.kernel.org
Subject: Re: [RFC PATCH v2 22/25] KVM: x86/mmu: Refactor kvm_mmu_invlpg() to allow skipping the gva flush
Date: Thu, 23 Jul 2026 08:23:32 -0700	[thread overview]
Message-ID: <amIx9Ki6gv-wgBnX@google.com> (raw)
In-Reply-To: <amGiAKsC3XrF6DyA@google.com>

On Thu, Jul 23, 2026, Yosry Ahmed wrote:
> On Wed, Jul 22, 2026 at 05:56:19PM -0700, Sean Christopherson wrote:
> > On Wed, Jul 22, 2026, Sean Christopherson wrote:
> > > On Tue, Jun 16, 2026, Yosry Ahmed wrote:
> > > > Refactor helpers out of kvm_mmu_invalidate_addr() and kvm_mmu_invlpg()
> > > > that take in an extra argument to skip the GVA flush.
> > > > 
> > > > This will be used when invalidating GVAs in a different context than the
> > > > correct one (i.e.  invalidating an L2 GVA from L1), so flushing the
> > > > current context would flush the wrong TLB entries.
> > > > 
> > > > No functional change intended.
> > > > 
> > > > Signed-off-by: Yosry Ahmed <yosry@kernel.org>
> > > > ---
> > > >  arch/x86/kvm/mmu/mmu.c | 23 +++++++++++++++++------
> > > >  1 file changed, 17 insertions(+), 6 deletions(-)
> > > > 
> > > > diff --git a/arch/x86/kvm/mmu/mmu.c b/arch/x86/kvm/mmu/mmu.c
> > > > index 65c35ed8f4a01..3feb75732f7b4 100644
> > > > --- a/arch/x86/kvm/mmu/mmu.c
> > > > +++ b/arch/x86/kvm/mmu/mmu.c
> > > > @@ -6615,15 +6615,15 @@ static void kvm_mmu_invalidate_addr_in_root(struct kvm_vcpu *vcpu,
> > > >  	write_unlock(&vcpu->kvm->mmu_lock);
> > > >  }
> > > >  
> > > > -void kvm_mmu_invalidate_addr(struct kvm_vcpu *vcpu, struct kvm_mmu *mmu,
> > > > -			     u64 addr, unsigned long roots)
> > > > +static void __kvm_mmu_invalidate_addr(struct kvm_vcpu *vcpu, struct kvm_mmu *mmu,
> > > > +				      u64 addr, unsigned long roots, bool flush_gva)
> > > >  {
> > > >  	int i;
> > > >  
> > > >  	WARN_ON_ONCE(roots & ~KVM_MMU_ROOTS_ALL);
> > > >  
> > > >  	/* It's actually a GPA for vcpu->arch.guest_mmu.  */
> > > > -	if (mmu != &vcpu->arch.guest_mmu) {
> > > > +	if (flush_gva && mmu != &vcpu->arch.guest_mmu) {
> > > >  		/* INVLPG on a non-canonical address is a NOP according to the SDM.  */
> > > >  		if (is_noncanonical_invlpg_address(addr, vcpu))
> > > >  			return;
> > > > @@ -6642,9 +6642,15 @@ void kvm_mmu_invalidate_addr(struct kvm_vcpu *vcpu, struct kvm_mmu *mmu,
> > > >  			kvm_mmu_invalidate_addr_in_root(vcpu, mmu, addr, mmu->prev_roots[i].hpa);
> > > >  	}
> > > >  }
> > > > +
> > > > +void kvm_mmu_invalidate_addr(struct kvm_vcpu *vcpu, struct kvm_mmu *mmu,
> > > > +			       u64 addr, unsigned long roots)
> > > 
> > > Rather than kvm_mmu_invalidate_addr() for the wrapper, what if we call this
> > > kvm_mmu_invalidate_gva()?  And then kvm_mmu_invlpg_gva().  Then we don't need
> > > to have the "in_root" version to a quad-underscores helper, and IMO it's more
> > > obvious what's different between the one-line wrappers and the inner helpers.
> > 
> > Hrm, or maybe I'm not understanding what "flush_gva" means.  At first glance, I
> > was assuming you were using it to differentiate between GVA and GPA, but IIUC,
> > it's literally skipping the flush for the current context, which just so happens
> > to be done only for GVAs.  I'd still like to avoid the "in_root" helper, if at
> > all possible.
> 
> Yes, it's skipping the TLB flush that is only needed for GVAs. The
> alternative I had in mind (but thought was worse) was to refactor the
> TLB flush part out of kvm_mmu_invalidate_addr() (or
> __kvm_mmu_invalidate_addr()) in this patch instead of adding a boolean,
> but this still requires adding a wrapper and the possibility of a
> quad-underscore helper.
> 
> What's the main objection to kvm_mmu_invalidate_addr_in_root()?

I don't love the __kvm_mmu_invalidate_addr() => kvm_mmu_invalidate_addr_in_root()
callchain.  It's not at all obvious that the in_root() helper shouldn't be called
directly.  I don't hate it, but I do think we need better clarity on what all this
is doing.

E.g. when looking at __kvm_inject_emulated_page_fault(), since it hardcodes a
single root, it's a bit headscratching to use kvm_mmu_invalidate_addr() instead
of kvm_mmu_invalidate_addr_in_root.

Hmm, and arguably, the way invlpga_interception() handles the ASID is flat out
wrong.  KVM doesn't need to flush *all* roots, rather it needs to flush L1 roots
for ASID=0, and L2 roots for ASID!=0.  If we can figure out an elegant way to
express and handle that, it should naturally handle the "flush GVA" aspect.

Actually, isn't there a pre-existing over-flush when handling kvm_mmu_invpcid_gva()?
Oof, and a missed flush?

To fix the over-flush, I think we want this?

diff --git arch/x86/kvm/mmu/mmu.c arch/x86/kvm/mmu/mmu.c
index 6c13da942bfc..7f3e0eb33b29 100644
--- arch/x86/kvm/mmu/mmu.c
+++ arch/x86/kvm/mmu/mmu.c
@@ -6672,7 +6672,8 @@ void kvm_mmu_invalidate_addr(struct kvm_vcpu *vcpu, struct kvm_pagewalk *w,
                if (is_noncanonical_invlpg_address(addr, vcpu))
                        return;
 
-               kvm_x86_call(flush_tlb_gva)(vcpu, addr);
+               if (roots & KVM_MMU_ROOT_CURRENT)
+                       kvm_x86_call(flush_tlb_gva)(vcpu, addr);
 
                if (tdp_enabled)
                        return;


And that highlights the missed flush: if the PCID isn't the current PCID, then
flush_tlb_gva() neglects to flush the hardware TLB for the target PCID, which
could leave a stale entry in the TLB if the guest switches to the new PCID with
MOV CR3 + X86_CR3_PCID_NOFLUSH.

  reply	other threads:[~2026-07-23 15:23 UTC|newest]

Thread overview: 62+ messages / expand[flat|nested]  mbox.gz  Atom feed  top
2026-06-16  0:41 [RFC PATCH v2 00/25] Optimize nSVM TLB flushes Yosry Ahmed
2026-06-16  0:41 ` [RFC PATCH v2 01/25] KVM: nSVM: Flush the TLB after forcefully leaving nested Yosry Ahmed
2026-06-16  0:41 ` [RFC PATCH v2 02/25] KVM: SVM: Passthrough the number of supported ASIDs Yosry Ahmed
2026-07-14 13:20   ` Yosry Ahmed
2026-07-14 21:28     ` Jim Mattson
2026-07-14 21:43       ` Yosry Ahmed
2026-07-14 23:41         ` Sean Christopherson
2026-07-15  6:32           ` Jim Mattson
2026-07-15 17:45             ` Yosry Ahmed
2026-07-15 18:20               ` Jim Mattson
2026-07-15 19:12                 ` Yosry Ahmed
2026-07-16 23:09                   ` Jim Mattson
2026-07-22 22:09             ` Sean Christopherson
2026-07-22 22:39               ` Jim Mattson
2026-07-23  0:27                 ` Sean Christopherson
2026-07-23  2:46                   ` Jim Mattson
2026-07-23 12:24                     ` Jim Mattson
2026-07-23 14:07                       ` Sean Christopherson
2026-07-23 15:57                         ` Jim Mattson
2026-07-23 16:32                           ` Yosry Ahmed
2026-07-23 16:54                             ` Jim Mattson
2026-07-23 16:57                               ` Yosry Ahmed
2026-07-23 17:07                                 ` Jim Mattson
2026-07-23 17:26                                   ` Sean Christopherson
2026-07-23 17:34                                     ` Yosry Ahmed
2026-07-15 17:41           ` Yosry Ahmed
2026-06-16  0:41 ` [RFC PATCH v2 03/25] KVM: VMX: Generalize VPID allocation to be vendor-neutral Yosry Ahmed
2026-07-22 22:26   ` Sean Christopherson
2026-07-22 22:36     ` Yosry Ahmed
2026-07-23 13:26       ` Sean Christopherson
2026-06-16  0:41 ` [RFC PATCH v2 04/25] KVM: x86/mmu: Support specifying a minimum TLB tag Yosry Ahmed
2026-07-23  0:28   ` Sean Christopherson
2026-07-23  5:00     ` Yosry Ahmed
2026-07-23 14:15       ` Sean Christopherson
2026-06-16  0:41 ` [RFC PATCH v2 05/25] KVM: SVM: Add helpers to set/clear ASID flush in VMCB Yosry Ahmed
2026-06-16  0:41 ` [RFC PATCH v2 06/25] KVM: SVM: Fallback to flush everything if FLUSHBYASID is not available Yosry Ahmed
2026-07-23  0:29   ` Sean Christopherson
2026-06-16  0:41 ` [RFC PATCH v2 07/25] KVM: SVM: Duplicate pre-run ASID check for SEV and non-SEV guests Yosry Ahmed
2026-06-16  0:41 ` [RFC PATCH v2 08/25] KVM: SEV: Stop using per-vCPU ASID for SEV VMs Yosry Ahmed
2026-06-16  0:41 ` [RFC PATCH v2 09/25] KVM: SVM: Use a static ASID per vCPU Yosry Ahmed
2026-06-16  0:41 ` [RFC PATCH v2 10/25] KVM: nSVM: Add a placeholder ASID for L2 Yosry Ahmed
2026-06-16  0:41 ` [RFC PATCH v2 11/25] KVM: x86: hyper-v: Rename kvm_hv_vcpu_purge_flush_tlb() Yosry Ahmed
2026-06-16  0:41 ` [RFC PATCH v2 12/25] KVM: x86: hyper-v: Allow puring all TLB flush FIFOs Yosry Ahmed
2026-06-16  0:41 ` [RFC PATCH v2 13/25] KVM: nSVM: Flush both L1 and L2 ASIDs on KVM_REQ_TLB_FLUSH Yosry Ahmed
2026-06-16  0:41 ` [RFC PATCH v2 14/25] KVM: nSVM: Move svm_switch_vmcb() to nested.c Yosry Ahmed
2026-06-16  0:41 ` [RFC PATCH v2 15/25] KVM: nSVM: Call nested_svm_transition_tlb_flush() on every VMCB switch Yosry Ahmed
2026-07-23  0:42   ` Sean Christopherson
2026-07-23  5:07     ` Yosry Ahmed
2026-06-16  0:41 ` [RFC PATCH v2 16/25] KVM: nSVM: Split nested_svm_transition_tlb_flush() into entry/exit fns Yosry Ahmed
2026-06-16  0:41 ` [RFC PATCH v2 17/25] KVM: nSVM: Service local TLB flushes before nested transitions Yosry Ahmed
2026-06-16  0:41 ` [RFC PATCH v2 18/25] KVM: nSVM: Handle nested TLB flush requests through TLB_CONTROL Yosry Ahmed
2026-06-16  0:41 ` [RFC PATCH v2 19/25] KVM: nSVM: Flush the TLB if L1 changes L2's ASID in vmcb12 Yosry Ahmed
2026-06-16  0:41 ` [RFC PATCH v2 20/25] KVM: nSVM: Do not reset TLB_CONTROL in vmcb02 on nested VM-Enter Yosry Ahmed
2026-06-16  0:41 ` [RFC PATCH v2 21/25] KVM: x86/mmu: rename __kvm_mmu_invalidate_addr() Yosry Ahmed
2026-06-16  0:41 ` [RFC PATCH v2 22/25] KVM: x86/mmu: Refactor kvm_mmu_invlpg() to allow skipping the gva flush Yosry Ahmed
2026-07-23  0:53   ` Sean Christopherson
2026-07-23  0:56     ` Sean Christopherson
2026-07-23  5:11       ` Yosry Ahmed
2026-07-23 15:23         ` Sean Christopherson [this message]
2026-06-16  0:41 ` [RFC PATCH v2 23/25] KVM: nSVM: Flush L2's ASID when emulating INVLPGA Yosry Ahmed
2026-06-16  0:41 ` [RFC PATCH v2 24/25] KVM: nSVM: Use different ASIDs for L1 and L2 Yosry Ahmed
2026-06-16  0:41 ` [RFC PATCH v2 25/25] DO NOT MERGE: Add nested_tlb_force_flush Yosry Ahmed

Reply instructions:

You may reply publicly to this message via plain-text email
using any one of the following methods:

* Save the following mbox file, import it into your mail client,
  and reply-to-all from there: mbox

  Avoid top-posting and favor interleaved quoting:
  https://en.wikipedia.org/wiki/Posting_style#Interleaved_style

* Reply using the --to, --cc, and --in-reply-to
  switches of git-send-email(1):

  git send-email \
    --in-reply-to=amIx9Ki6gv-wgBnX@google.com \
    --to=seanjc@google.com \
    --cc=jmattson@google.com \
    --cc=kvm@vger.kernel.org \
    --cc=linux-kernel@vger.kernel.org \
    --cc=mlevitsk@redhat.com \
    --cc=pbonzini@redhat.com \
    --cc=thomas.lendacky@amd.com \
    --cc=vkuznets@redhat.com \
    --cc=yosry@kernel.org \
    /path/to/YOUR_REPLY

  https://kernel.org/pub/software/scm/git/docs/git-send-email.html

* If your mail client supports setting the In-Reply-To header
  via mailto: links, try the mailto: link
Be sure your reply has a Subject: header at the top and a blank line before the message body.
This is a public inbox, see mirroring instructions
for how to clone and mirror all data and code used for this inbox