From mboxrd@z Thu Jan 1 00:00:00 1970 Return-Path: X-Spam-Checker-Version: SpamAssassin 3.4.0 (2014-02-07) on aws-us-west-2-korg-lkml-1.web.codeaurora.org Received: from gabe.freedesktop.org (gabe.freedesktop.org [131.252.210.177]) (using TLSv1.2 with cipher ECDHE-RSA-AES256-GCM-SHA384 (256/256 bits)) (No client certificate requested) by smtp.lore.kernel.org (Postfix) with ESMTPS id 0B971C79F82 for ; Fri, 4 Sep 2026 18:25:17 +0000 (UTC) Received: from gabe.freedesktop.org (localhost [127.0.0.1]) by gabe.freedesktop.org (Postfix) with ESMTP id AC66B10E55D; Fri, 4 Sep 2026 18:25:16 +0000 (UTC) Authentication-Results: gabe.freedesktop.org; dkim=pass (2048-bit key; unprotected) header.d=intel.com header.i=@intel.com header.b="ScNRdRpo"; dkim-atps=neutral Received: from mgamail.intel.com (mgamail.intel.com [192.198.163.17]) by gabe.freedesktop.org (Postfix) with ESMTPS id 1C2AC10FAEA for ; Fri, 4 Sep 2026 18:25:13 +0000 (UTC) DKIM-Signature: v=1; a=rsa-sha256; c=relaxed/simple; d=intel.com; i=@intel.com; q=dns/txt; s=Intel; t=1788546313; x=1820082313; h=from:to:subject:date:message-id:in-reply-to:references: mime-version:content-transfer-encoding; bh=mcJnxzKILBaN4IA5yrlbPyEmxRSAOQzWZyg768KYkXc=; b=ScNRdRpo+kSFa0S76AZw5mT+yD5Ciw0Iqpt8B2fl4Kmi2XcUfnJC1voi O1yRYRo4KUiR7tVnkkd71lOUIIsTrBoGy0l8ncYg6IpX6DpX3I+49Or9a q779HQssuDEO+g5Y5f5Umrbxd0D2xgEQiERPx1Bm0x+6Dv+J/9X426Gi4 bv2z++vAKKDANdjce6V3CY4wM1bp2Z/ybdji4u5/3xHHJeOZMBWt9w2hV ZIWH14ZtXxNPnTO7FXv5yTTI8oV4kgShvvTE8mrxqmIEOgZkQb+2SN/BY L9TVAieVNdnLQa8HrmdzLp0ig0YjYr4BvS73tDc4uWtLHtJtGiTux1QYH A==; X-CSE-ConnectionGUID: Ezrm6z8rRBWslitwGtgJSA== X-CSE-MsgGUID: NdhSGDzeQ6iDFNcKu2MTwg== X-IronPort-AV: E=McAfee;i="6800,10657,11896"; a="88932261" X-IronPort-AV: E=Sophos;i="6.25,262,1779174000"; d="scan'208";a="88932261" Received: from fmviesa001.fm.intel.com ([10.60.135.141]) by fmvoesa111.fm.intel.com with ESMTP/TLS/ECDHE-RSA-AES256-GCM-SHA384; 04 Sep 2026 11:25:13 -0700 X-CSE-ConnectionGUID: rbVsvXJOTPGLYouzYC9FOQ== X-CSE-MsgGUID: ufg76MYRRk6PmxT9oOEfhw== X-ExtLoop1: 1 X-IronPort-AV: E=Sophos;i="6.25,262,1779174000"; d="scan'208";a="294992503" Received: from mjruhl-vm.amr.corp.intel.com ([10.11.186.166]) by smtpauth.intel.com with ESMTP/TLS/ECDHE-RSA-AES256-GCM-SHA384; 04 Sep 2026 11:25:12 -0700 From: "Michael J. Ruhl" To: platform-driver-x86@vger.kernel.org, intel-xe@lists.freedesktop.org, hansg@kernel.org, ilpo.jarvinen@linux.intel.com, matthew.brost@intel.com, rodrigo.vivi@intel.com, thomas.hellstrom@linux.intel.com, airlied@gmail.com, simona@ffwll.ch, david.e.box@linux.intel.com, anoop.c.vijay@intel.com, badal.nilawar@intel.com, matthew.d.roper@intel.com, james.ausmus@intel.com, karthik.poosa@intel.com Subject: [PATCH v6 15/18] drm/xe/vsec: Support late bind fw information Date: Fri, 4 Sep 2026 11:25:05 -0700 Message-ID: <20260904182451.1164868-35-michael.j.ruhl@intel.com> X-Mailer: git-send-email 2.43.0 In-Reply-To: <20260904182451.1164868-20-michael.j.ruhl@intel.com> References: <20260904182451.1164868-20-michael.j.ruhl@intel.com> MIME-Version: 1.0 Content-Transfer-Encoding: 8bit X-BeenThere: intel-xe@lists.freedesktop.org X-Mailman-Version: 2.1.29 Precedence: list List-Id: Intel Xe graphics driver List-Unsubscribe: , List-Archive: List-Post: List-Help: List-Subscribe: , Errors-To: intel-xe-bounces@lists.freedesktop.org Sender: "Intel-xe" CRI FW is loaded on power on. Because of this, access to the FW cannot be done until it is running. Update the XE PMT probe and access to check for late bind devices, verify, and wait for the appropriate FW state before probe or access. Signed-off-by: Michael J. Ruhl --- drivers/gpu/drm/xe/xe_device.c | 4 +- drivers/gpu/drm/xe/xe_device_types.h | 5 + drivers/gpu/drm/xe/xe_vsec.c | 134 +++++++++++++++++++++++++-- drivers/gpu/drm/xe/xe_vsec.h | 2 +- 4 files changed, 137 insertions(+), 8 deletions(-) diff --git a/drivers/gpu/drm/xe/xe_device.c b/drivers/gpu/drm/xe/xe_device.c index 8583b2e9ecf4..e51e3cc4ea62 100644 --- a/drivers/gpu/drm/xe/xe_device.c +++ b/drivers/gpu/drm/xe/xe_device.c @@ -1148,7 +1148,9 @@ int xe_device_probe(struct xe_device *xe) for_each_gt(gt, xe, id) xe_gt_sanitize_freq(gt); - xe_vsec_init(xe); + err = xe_vsec_init(xe); + if (err) + goto err_unregister_display; err = xe_sriov_init_late(xe); if (err) diff --git a/drivers/gpu/drm/xe/xe_device_types.h b/drivers/gpu/drm/xe/xe_device_types.h index 3f1a70813a99..69e052ac5a82 100644 --- a/drivers/gpu/drm/xe/xe_device_types.h +++ b/drivers/gpu/drm/xe/xe_device_types.h @@ -7,6 +7,7 @@ #define _XE_DEVICE_TYPES_H_ #include +#include #include #include @@ -468,6 +469,10 @@ struct xe_device { struct mutex lock; /** @pmt.base_offset: device specific base offset */ u64 base_offset; + /** @pmt.work: support late-bind probe */ + struct delayed_work work; + /** @pmt.retry_count: late-bind probe retry */ + u32 retry_count; } pmt; /** @soc_remapper: SoC remapper object */ diff --git a/drivers/gpu/drm/xe/xe_vsec.c b/drivers/gpu/drm/xe/xe_vsec.c index 7c3f9938701e..5c352af193cf 100644 --- a/drivers/gpu/drm/xe/xe_vsec.c +++ b/drivers/gpu/drm/xe/xe_vsec.c @@ -3,6 +3,7 @@ #include #include #include +#include #include #include #include @@ -18,6 +19,7 @@ #include "xe_mmio.h" #include "xe_platform_types.h" #include "xe_pm.h" +#include "xe_sysctrl.h" #include "xe_vsec.h" #include "regs/xe_pmt.h" @@ -173,6 +175,14 @@ enum capability { WATCHER, }; +/* + * Late bind will delay 100msec for up to 20 seconds + */ +#define VSEC_LATE_BIND_DELAY_MSEC 100 +#define VSEC_LATE_BIND_RETRY 200 + +static void cri_late_bind_probe(struct xe_device *xe); + static int bmg_guid_decode(u32 guid, int *index, u32 *offset) { u32 record_id = FIELD_GET(GUID_RECORD_ID, guid); @@ -281,6 +291,48 @@ static int xe_guid_decode(u32 guid, int *index, u32 *offset) return -ENODEV; } +static void cri_late_bind_probe_work(struct work_struct *work) +{ + struct xe_device *xe = container_of(work, struct xe_device, pmt.work.work); + + if (xe_is_oobmsm_fw_ready(xe)) { + cri_late_bind_probe(xe); + xe_pm_runtime_put(xe); + return; + } + + xe->pmt.retry_count++; + + /* wait up to 20 seconds */ + if (xe->pmt.retry_count == VSEC_LATE_BIND_RETRY) { + drm_warn(&xe->drm, "PMT probe: Late Binding failed to complete\n"); + xe_pm_runtime_put(xe); + return; + } + + if (!schedule_delayed_work(&xe->pmt.work, msecs_to_jiffies(VSEC_LATE_BIND_DELAY_MSEC))) + xe_pm_runtime_put(xe); +} + +static bool wait_for_fw(struct xe_device *xe) +{ + int retries = VSEC_LATE_BIND_RETRY; /* wait up to 20 secs */ + + if (xe->info.platform != XE_CRESCENTISLAND) + return true; + + while (retries--) { + if (xe_is_oobmsm_fw_ready(xe)) + return true; + + msleep(VSEC_LATE_BIND_DELAY_MSEC); + } + + drm_warn(&xe->drm, "Late Binding failed to complete\n"); + + return false; +} + /** * xe_pmt_telem_read - Given a device and a PMT GUID, read data into a buffer * @dev: valid Xe device @@ -339,6 +391,11 @@ int xe_pmt_telem_read(struct device *dev, u32 guid, u64 *data, loff_t user_offse goto dev_exit; } + if (!wait_for_fw(xe)) { + ret = -ENODATA; + goto runtime_exit; + } + mutex_lock(&xe->pmt.lock); /* set SoC re-mapper index register based on GUID memory region */ @@ -348,6 +405,7 @@ int xe_pmt_telem_read(struct device *dev, u32 guid, u64 *data, loff_t user_offse mutex_unlock(&xe->pmt.lock); +runtime_exit: xe_pm_runtime_put(xe); dev_exit: @@ -390,6 +448,10 @@ static int xe_pmt_read_reg(struct device *dev, u32 guid, u32 *reg, u32 offset) disc_addr += inst + offset; xe_pm_runtime_get(xe); + if (!wait_for_fw(xe)) { + ret = -ENODATA; + goto runtime_exit; + } mutex_lock(&xe->pmt.lock); xe->soc_remapper.set_telem_region(xe, CRI_IDX_TELEM_DISCOVERY); @@ -397,6 +459,8 @@ static int xe_pmt_read_reg(struct device *dev, u32 guid, u32 *reg, u32 offset) *reg = readl(disc_addr); mutex_unlock(&xe->pmt.lock); + +runtime_exit: xe_pm_runtime_put(xe); dev_exit: @@ -427,6 +491,10 @@ static int xe_pmt_write_reg(struct device *dev, u32 guid, u32 reg, u32 offset) disc_addr += inst + offset; xe_pm_runtime_get(xe); + if (!wait_for_fw(xe)) { + ret = -ENODATA; + goto runtime_exit; + } mutex_lock(&xe->pmt.lock); xe->soc_remapper.set_telem_region(xe, CRI_IDX_TELEM_DISCOVERY); @@ -434,6 +502,8 @@ static int xe_pmt_write_reg(struct device *dev, u32 guid, u32 reg, u32 offset) writel(reg, disc_addr); mutex_unlock(&xe->pmt.lock); + +runtime_exit: xe_pm_runtime_put(xe); dev_exit: @@ -465,12 +535,44 @@ static enum xe_vsec get_platform_info(struct xe_device *xe) return vsec_platforms[xe->info.platform]; } +static void cri_late_bind_probe(struct xe_device *xe) +{ + struct intel_vsec_platform_info *info; + struct device *dev = xe->drm.dev; + enum xe_vsec platform; + + platform = get_platform_info(xe); + if (platform != XE_VSEC_CRI) + return; + + info = &xe_vsec_info[platform]; + if (!info->headers) + return; + + info->priv_data = &xe_cri_pmt_cb; + xe->soc_remapper.set_telem_region(xe, CRI_IDX_TELEM_DISCOVERY); + + intel_vsec_register(dev, info); +} + +static void vsec_disable_late_bind_work(void *arg) +{ + struct xe_device *xe = arg; + + /* + * If the queued work is canceled, the runtime reference needs to be + * released here. + */ + if (disable_delayed_work_sync(&xe->pmt.work)) + xe_pm_runtime_put(xe); +} + /** * xe_vsec_init - Initialize resources and add intel_vsec auxiliary * interface * @xe: valid xe instance */ -void xe_vsec_init(struct xe_device *xe) +int xe_vsec_init(struct xe_device *xe) { struct intel_vsec_platform_info *info; struct device *dev = xe->drm.dev; @@ -478,30 +580,45 @@ void xe_vsec_init(struct xe_device *xe) platform = get_platform_info(xe); if (platform == XE_VSEC_UNKNOWN) - return; + return 0; info = &xe_vsec_info[platform]; if (!info->headers) - return; + return 0; switch (platform) { case XE_VSEC_BMG: if (IS_SRIOV_VF(xe)) - return; + return 0; xe->pmt.base_offset = BMG_TELEMETRY_OFFSET; info->priv_data = &xe_bmg_pmt_cb; break; case XE_VSEC_CRI: if (IS_SRIOV_VF(xe)) - return; + return 0; + xe->pmt.base_offset = CRI_PMT_OFFSET; + + xe->pmt.retry_count = 0; + INIT_DELAYED_WORK(&xe->pmt.work, cri_late_bind_probe_work); + + xe_pm_runtime_get_noresume(xe); + if (!xe_is_oobmsm_fw_ready(xe)) { + schedule_delayed_work(&xe->pmt.work, + msecs_to_jiffies(VSEC_LATE_BIND_DELAY_MSEC)); + return devm_add_action_or_reset(xe->drm.dev, + vsec_disable_late_bind_work, + xe); + } + info->priv_data = &xe_cri_pmt_cb; xe->soc_remapper.set_telem_region(xe, CRI_IDX_TELEM_DISCOVERY); break; default: - break; + drm_err(&xe->drm, "Unsupported platform: %u\n", platform); + return 0; } /* @@ -509,5 +626,10 @@ void xe_vsec_init(struct xe_device *xe) * resources. */ intel_vsec_register(dev, info); + + if (platform == XE_VSEC_CRI) + xe_pm_runtime_put(xe); + + return 0; } MODULE_IMPORT_NS("INTEL_VSEC"); diff --git a/drivers/gpu/drm/xe/xe_vsec.h b/drivers/gpu/drm/xe/xe_vsec.h index a25b4e6e681b..c4a1e2fc67d8 100644 --- a/drivers/gpu/drm/xe/xe_vsec.h +++ b/drivers/gpu/drm/xe/xe_vsec.h @@ -9,7 +9,7 @@ struct device; struct xe_device; -void xe_vsec_init(struct xe_device *xe); +int xe_vsec_init(struct xe_device *xe); int xe_pmt_telem_read(struct device *dev, u32 guid, u64 *data, loff_t user_offset, u32 count); #endif -- 2.43.0