All of lore.kernel.org
 help / color / mirror / Atom feed
From: Sean Christopherson <seanjc@google.com>
To: Xiaoyao Li <xiaoyao.li@intel.com>
Cc: Rick P Edgecombe <rick.p.edgecombe@intel.com>,
	"kvm@vger.kernel.org" <kvm@vger.kernel.org>,
	 "linux-kernel@vger.kernel.org" <linux-kernel@vger.kernel.org>,
	 "binbin.wu@linux.intel.com" <binbin.wu@linux.intel.com>,
	"kas@kernel.org" <kas@kernel.org>,
	 "pbonzini@redhat.com" <pbonzini@redhat.com>,
	"nik.borisov@suse.com" <nik.borisov@suse.com>,
	 Chao Gao <chao.gao@intel.com>,
	 "dave.hansen@linux.intel.com" <dave.hansen@linux.intel.com>,
	 "andrew.cooper3@citrix.com" <andrew.cooper3@citrix.com>
Subject: Re: [PATCH v3 1/4] KVM: TDX: Track configurable CPUID bits allowed by KVM
Date: Wed, 9 Sep 2026 15:29:22 -0700	[thread overview]
Message-ID: <aqHdwkOFVR4XoS2M@google.com> (raw)
In-Reply-To: <9abeed40-bcdc-48d9-8cbb-cb87154bb9f6@intel.com>

On Thu, Sep 10, 2026, Xiaoyao Li wrote:
> On 9/9/2026 5:13 AM, Edgecombe, Rick P wrote:
> >>>> In the end, they might be disallowed to be configured to TDs because KVM
> >>>> doesn't allow them for VMX VMs. This is also the point I want to discuss.
> >>>> Do we really want to make such restriction that KVM cannot enable/allow a
> >>>> feature for TDs unless KVM first enables/allows it for VMX VMs? What's
> >>>> reason behind it?
> >>> Sean mentioned it that "generally speaking, KVM shouldn't allow features
> >>> that KVM doesn't support for non-TDX VMs" in
> >>> https://lore.kernel.org/kvm/aj1fi_0SBxMK5WOB@google.com/
> >> For existing features, it might make some sense. But for new features, I
> >> don't think so. It defines the enabling order for new features that we must
> >> enable a feature for non-TDX VMs first and then TDs. And people might want
> >> to bypass this rule by abusing the TDX_CFG_EXTRA_F() when only one line of
> >> TDX_CFG_EXTRA_F() is enough to enable a feature for TDs but more effort
> >> required to enable it for non-TDX VMs.
> >>
> >> Maybe I miss somthing. I would like to see stronger reasons for such decision.
> > I think "generally speaking" means, it's not a hard rule.

Ya.
 
> > As for why to prefer it, I think we would normally want regulars VMs and TDs to
> > work similarly. Especially those that have some of the virtualization handled by
> > KVM. But TDX module's behavior of a feature can conform to KVM's only if KVM's
> > already exists. Take for example split lock detection. The normal VM KVM support
> > initially went through several iterations of design. Separately, TDX ended up
> > with a different solution. Imagine if we had enabled the TDX arch one, before
> > solving the general KVM problems. Then we would end up with two different
> > behaviors, or a worse KVM behavior as it tries to conform to TDX module's
> > behavior.
> > 
> > So we need to at least solve a feature at that level before deciding KVM's
> > handling of it. This could be done while enabling the feature for TDX only, but
> > often would involve solving the problems for normal VMs too. In the end, it's
> > the generic KVM behavior that needs to be solved before enabling the feature.
> > 
> > Ideally we could consider normal VM and TD at the same time. I expect we will
> > start doing that after this series is in place. Not a hard rule, but a norm.

Eh, I'm with Xiaoyao.  Yeah, *ideally* we'd magically enable everything everywhere
all at once.  In reality, different VM types are going to support features at
different times.  More importantly, as Xiaoyao points out below in #1, unless we
enable everyting in a single patch, which is probably a terrible idea in most cases,
we'll still end up with staged/progressive enabling, i.e. we still need to have
patches that selectively enable and advertise a feature only for the VM types
that actually support the feature.

This is all quite similar to Intel and AMD feature enabling being done at different
times.  The biggest difference is that Intel and AMD are mutually exclusive and
so KVM_GET_SUPPORTED_CPUID always reports the correct information, but TDX already
provides KVM_TDX_CAPABILITIES, so AFAICT we still get accurate reporting for TDX,
just in a slightly different way.

> I don't think "split lock detection" is a good example. We are discussing
> virtualizing a feature, or allowing a feature to be exposed to non-TDX and TDX
> guests. While "split lock detection" is not a virtualizable feature and how KVM
> handles it is all about how KVM fixes the architectural flaw of it.

Yep.  If there are actual decisions to be made, versus simply adhering to the
architecture, then we'll need to incorporate the needs/abilities of flavors of
VMs KVM supports.  But for feature virtualization where right vs. wrong is
dictated by hardware specs, there really isn't anything we can do in KVM to affect
the guest-visible behavior (beyond things like performance characteristics, but
those aren't ABI in any case).

> Generally, I agree with your point that we need to think at higher level before
> KVM deciding to support virtualizing a feature. But after the generic level
> consideration is done, there can be different orders:
> 
> 1. allow a feature for non-TDX VMs and TDs at same time. i.e., in one patch or
> in a single series.
> 
> 2. allow a feature for non-TDX VMs first and then TDs. This can be due to by
> that time TDX's spec for the feature has not been defined/finalized.
> 
> 3. allow a feature for TDs first and then non-TDX VMs. This can be due to it
> requires more effort/patches to support virtualizing a feature for non-TDX VMs.
> 
> Usually allowing a feature for TDs requires far less efforts than for non-TDX
> VMs, because most of the virtualization work is done by TDX module. E.g., for
> TDs, KVM usually only needs to add code to restore the host states that are not
> restored by TDX module. However for non-TDX VMs, KVM needs to do more: emulate
> the architectural behavior of the MSRs, context switch states, and the nested
> handling, etc. It's likely that a series to enable a feature for non-TDX VMs
> spends several Linux releases to finally get merged. Making TDX support depends
> on it seems not that necessary.

  reply	other threads:[~2026-09-09 22:29 UTC|newest]

Thread overview: 63+ messages / expand[flat|nested]  mbox.gz  Atom feed  top
2026-08-27  3:18 [PATCH v3 0/4] KVM: TDX: Validate directly configurable CPUID bits Binbin Wu
2026-08-27  3:18 ` [PATCH v3 1/4] KVM: TDX: Track configurable CPUID bits allowed by KVM Binbin Wu
2026-09-01  6:29   ` Tony Lindgren
2026-09-01  8:23     ` Binbin Wu
2026-09-01  8:27       ` Tony Lindgren
2026-09-01 14:35   ` Xiaoyao Li
2026-09-02  0:33     ` Binbin Wu
2026-09-02 15:09       ` Xiaoyao Li
2026-09-02 16:19         ` Binbin Wu
2026-09-02 16:22           ` Edgecombe, Rick P
2026-09-02 16:25             ` Binbin Wu
2026-09-03  7:28           ` Xiaoyao Li
2026-09-03  8:57             ` Binbin Wu
2026-09-08 21:13             ` Edgecombe, Rick P
2026-09-09 16:39               ` Xiaoyao Li
2026-09-09 22:29                 ` Sean Christopherson [this message]
2026-09-09 23:18                   ` Edgecombe, Rick P
2026-09-10  2:39                     ` Binbin Wu
2026-09-10  2:53                     ` Xiaoyao Li
2026-09-08 21:15         ` Edgecombe, Rick P
2026-08-27  3:18 ` [PATCH v3 2/4] KVM: TDX: Report CORE_CAPABILITIES as configurable Binbin Wu
2026-09-01  6:45   ` Tony Lindgren
2026-09-02 17:43   ` Kishen Maloor
2026-09-03  2:22     ` Binbin Wu
2026-09-03  6:10       ` Kishen Maloor
2026-09-03  8:12         ` Binbin Wu
2026-08-27  3:18 ` [PATCH v3 3/4] KVM: TDX: Filter configurable CPUID bits Binbin Wu
2026-09-01  6:44   ` Tony Lindgren
2026-09-01  8:42     ` Binbin Wu
2026-09-01  9:09       ` Tony Lindgren
2026-09-03  8:04   ` Xiaoyao Li
2026-09-03  8:23     ` Binbin Wu
2026-08-27  3:18 ` [PATCH v3 4/4] KVM: TDX: Validate userspace CPUID input for KVM_TDX_INIT_VM Binbin Wu
2026-08-27  3:24   ` sashiko-bot
2026-08-27  7:25     ` Binbin Wu
2026-09-01  6:47   ` Tony Lindgren
2026-08-27 19:33 ` [PATCH v3 0/4] KVM: TDX: Validate directly configurable CPUID bits Edgecombe, Rick P
2026-08-28  3:19   ` Binbin Wu
2026-08-28 16:58     ` Edgecombe, Rick P
2026-08-31  5:01       ` Binbin Wu
2026-09-01  9:42         ` Xiaoyao Li
2026-09-01 10:21           ` Xiaoyao Li
2026-09-02 16:09           ` Edgecombe, Rick P
2026-09-02 16:21             ` Binbin Wu
2026-09-09  1:46             ` Binbin Wu
2026-09-01  9:38     ` Xiaoyao Li
2026-09-01 17:41       ` Edgecombe, Rick P
2026-09-02 10:29         ` Xiaoyao Li
2026-09-02 13:13           ` Edgecombe, Rick P
2026-09-02 13:39             ` Xiaoyao Li
2026-09-02 13:53               ` Edgecombe, Rick P
2026-09-02 14:21                 ` Xiaoyao Li
2026-09-02 16:26             ` Binbin Wu
2026-09-08  9:42 ` Artem Bityutskiy
2026-09-09  0:04   ` Binbin Wu
2026-09-08 20:30 ` Artem Bityutskiy
2026-09-08 22:31   ` Edgecombe, Rick P
2026-09-09  6:52     ` Artem Bityutskiy
2026-09-09  8:48       ` Binbin Wu
2026-09-09 11:20         ` Artem Bityutskiy
2026-09-10  2:54           ` Binbin Wu
2026-09-08 23:54   ` Binbin Wu
2026-09-09  5:37     ` Binbin Wu

Reply instructions:

You may reply publicly to this message via plain-text email
using any one of the following methods:

* Save the following mbox file, import it into your mail client,
  and reply-to-all from there: mbox

  Avoid top-posting and favor interleaved quoting:
  https://en.wikipedia.org/wiki/Posting_style#Interleaved_style

* Reply using the --to, --cc, and --in-reply-to
  switches of git-send-email(1):

  git send-email \
    --in-reply-to=aqHdwkOFVR4XoS2M@google.com \
    --to=seanjc@google.com \
    --cc=andrew.cooper3@citrix.com \
    --cc=binbin.wu@linux.intel.com \
    --cc=chao.gao@intel.com \
    --cc=dave.hansen@linux.intel.com \
    --cc=kas@kernel.org \
    --cc=kvm@vger.kernel.org \
    --cc=linux-kernel@vger.kernel.org \
    --cc=nik.borisov@suse.com \
    --cc=pbonzini@redhat.com \
    --cc=rick.p.edgecombe@intel.com \
    --cc=xiaoyao.li@intel.com \
    /path/to/YOUR_REPLY

  https://kernel.org/pub/software/scm/git/docs/git-send-email.html

* If your mail client supports setting the In-Reply-To header
  via mailto: links, try the mailto: link
Be sure your reply has a Subject: header at the top and a blank line before the message body.
This is an external index of several public inboxes,
see mirroring instructions on how to clone and mirror
all data and code used by this external index.