From mboxrd@z Thu Jan 1 00:00:00 1970 Received: from smtp.kernel.org (aws-us-west-2-korg-mail-alma10-1.taild15c8.ts.net [100.103.45.18]) (using TLSv1.2 with cipher ECDHE-RSA-AES256-GCM-SHA384 (256/256 bits)) (No client certificate requested) by smtp.subspace.kernel.org (Postfix) with ESMTPS id 6D7DE3BED01 for ; Mon, 31 Aug 2026 03:41:43 +0000 (UTC) Authentication-Results: smtp.subspace.kernel.org; arc=none smtp.client-ip=100.103.45.18 ARC-Seal:i=1; a=rsa-sha256; d=subspace.kernel.org; s=arc-20240116; t=1788147709; cv=none; b=DLLhlnvnlFKt9qPDhkATjdPg+L8PKqaHz2Y5duN1KIVHHJcXZQFTVDzG1ycEXmK4cdxcfFWUAu/q5QeLHtC9tgPqW79fke2ugn9qEmrp6duPr2z91deQdh7cC6JK34VVMS4lMbQ1WoW+ZOs+khhDH9bugaGqX5AGas2j81QSCMI= ARC-Message-Signature:i=1; a=rsa-sha256; d=subspace.kernel.org; s=arc-20240116; t=1788147709; c=relaxed/simple; bh=k6D7/M2TCTPyy6m62kouocEP2hTKuX/rxJNK3urTMB8=; h=From:To:Cc:Subject:Date:Message-ID:In-Reply-To:References: MIME-Version; b=HFocEuEzp5RG9KMaguch1aGhP6l2IpBWxfhJdv7j/rXhycjNjp7pn8ugoqtgI6uXpSTUlaZzmySDN1Df4OYQ1hJAgaV9EBL3sOiEry9Xz3COxy3YXxAHwDHz17BWAZ1tus3hA2gFw3JyI3QCphduMYoNfCw6FSFj6tC91nsDcgE= ARC-Authentication-Results:i=1; smtp.subspace.kernel.org; dkim=pass (2048-bit key) header.d=kernel.org header.i=@kernel.org header.b=h6zTaWeY; arc=none smtp.client-ip=100.103.45.18 Authentication-Results: smtp.subspace.kernel.org; dkim=pass (2048-bit key) header.d=kernel.org header.i=@kernel.org header.b="h6zTaWeY" Received: by smtp.kernel.org (Postfix) with ESMTPSA id 5649B1F00ADE; Mon, 31 Aug 2026 03:41:27 +0000 (UTC) DKIM-Signature: v=1; a=rsa-sha256; c=relaxed/relaxed; d=kernel.org; s=k20260515; t=1788147687; bh=asNIaf5uT5FMBqDWaVplZYt5ziW4Ek+HbJ9uw6LT3fo=; h=From:To:Cc:Subject:Date:In-Reply-To:References; b=h6zTaWeYnrt+rhCRGtmstjXLmUJU4S8C+qEJhxTinOI8LS9js2JqyjWgvBQEhjbFb J0ZRxIbY5rBo92JNLAlHaxMlCYOcOpB9bO7rQN8kJ/31sC2tuT24RscbPWG31hBiOg ww0+P6/VbrdNnk30aSZ7yttXAiQ6tb/zTQExKYPgmNrS/CcKopiX/ePSOX/BSkmb09 RbmStffo1BixtAiJ42PprwTyZ3PTEUri9LWbbxnMCRHa4OfrknRSzGFdAK/0dU/Tc9 YCgiRszzbmFkDD4F/Vi7dBFvjeBxSztBJ7q3M2HT/amNLPaGAnWngCDlDP2Se0GCkk kZnV24EdLBObA== From: Damien Le Moal To: Jens Axboe , linux-block@vger.kernel.org Cc: Christoph Hellwig Subject: [PATCH v6 07/12] block: drop all zone write plugs on capacity changes Date: Mon, 31 Aug 2026 12:41:04 +0900 Message-ID: <20260831034109.744039-8-dlemoal@kernel.org> X-Mailer: git-send-email 2.55.0 In-Reply-To: <20260831034109.744039-1-dlemoal@kernel.org> References: <20260831034109.744039-1-dlemoal@kernel.org> Precedence: bulk X-Mailing-List: linux-block@vger.kernel.org List-Id: List-Subscribe: List-Unsubscribe: MIME-Version: 1.0 Content-Transfer-Encoding: 8bit If during revalidation, we detect a capacity change for a zoned block device, e.g. due to a storage element removal on an HDD, we can assume that the device was reformatted, which implies that all sequential zones are empty. For such case, we can remove and free all zone write plugs in the gendisk hash table by marking them as dead, thus avoiding also to leave zone write plugs for zones that are beyond the new device capacity in the disk hash table. Introduce the function disk_revalidate_capacity() to do this and call this new function at the beginning of blk_revalidate_disk_zones(), so that the zone revalidation process can re-create, if needed, any zone write plug for sequential zones that are not empty. The checks on the capacity and zone size that were in blk_revalidate_disk_zones() are moved to disk_revalidate_capacity() and if true, also trigger dropping all zone write plugs. Signed-off-by: Damien Le Moal Reviewed-by: Bart Van Assche Reviewed-by: Hannes Reinecke Reviewed-by: Johannes Thumshirn Reviewed-by: Christoph Hellwig --- block/blk-zoned.c | 71 ++++++++++++++++++++++++++++++++++++----------- 1 file changed, 55 insertions(+), 16 deletions(-) diff --git a/block/blk-zoned.c b/block/blk-zoned.c index 83737c9a92d8..4b0eda94bbf3 100644 --- a/block/blk-zoned.c +++ b/block/blk-zoned.c @@ -2222,6 +2222,56 @@ static int disk_revalidate_zone_resources(struct gendisk *disk, return ret; } +static void disk_drop_zone_wplug(struct blk_zone_wplug *zwplug, void *data) +{ + unsigned long flags; + + spin_lock_irqsave(&zwplug->lock, flags); + disk_zone_wplug_abort(zwplug); + disk_mark_zone_wplug_dead(zwplug); + spin_unlock_irqrestore(&zwplug->lock, flags); +} + +static int disk_revalidate_capacity(struct gendisk *disk) +{ + struct queue_limits *lim = &disk->queue->limits; + sector_t zone_sectors = lim->chunk_sectors; + unsigned int nr_zones = disk_get_nr_zones(disk); + int ret = -ENODEV; + + if (!get_capacity(disk) || !nr_zones) + goto drop_all_zwplugs; + + /* + * Checks that the device driver indicated a valid zone size and that + * the max zone append limit is set. + */ + if (!zone_sectors || !is_power_of_2(zone_sectors)) { + pr_warn("%s: Invalid non power of two zone size (%llu)\n", + disk->disk_name, zone_sectors); + goto drop_all_zwplugs; + } + + /* + * Check if the capacity has changed. If it did, assume that the device + * was reformatted and that all sequential zones are now empty. So drop + * all zone write plug. + */ + if (disk->nr_zones && disk->nr_zones != nr_zones) { + pr_warn("%s: Number of zones changed (%u -> %u)\n", + disk->disk_name, disk->nr_zones, nr_zones); + ret = 0; + goto drop_all_zwplugs; + } + + return 0; + +drop_all_zwplugs: + disk_for_all_zone_wplugs(disk, disk_drop_zone_wplug, NULL); + + return ret; +} + static int blk_revalidate_zone_cond(struct blk_zone *zone, unsigned int idx, struct blk_revalidate_zone_args *args) { @@ -2411,31 +2461,20 @@ static int blk_revalidate_zone_cb(struct blk_zone *zone, unsigned int idx, */ int blk_revalidate_disk_zones(struct gendisk *disk) { - struct request_queue *q = disk->queue; - sector_t zone_sectors = q->limits.chunk_sectors; struct blk_revalidate_zone_args args = { }; struct blk_report_zones_args rep_args = { .cb = blk_revalidate_zone_cb, .data = &args, }; unsigned int noio_flag; - int ret = -ENOMEM; + int ret; - if (WARN_ON_ONCE(!blk_queue_is_zoned(q))) + if (WARN_ON_ONCE(!blk_queue_is_zoned(disk->queue))) return -EIO; - if (!get_capacity(disk)) - return -ENODEV; - - /* - * Checks that the device driver indicated a valid zone size and that - * the max zone append limit is set. - */ - if (!zone_sectors || !is_power_of_2(zone_sectors)) { - pr_warn("%s: Invalid non power of two zone size (%llu)\n", - disk->disk_name, zone_sectors); - return -ENODEV; - } + ret = disk_revalidate_capacity(disk); + if (ret) + return ret; /* * Allocate zone resources if they are needed and we have not done -- 2.55.0