From mboxrd@z Thu Jan 1 00:00:00 1970 Return-Path: X-Spam-Checker-Version: SpamAssassin 3.4.0 (2014-02-07) on aws-us-west-2-korg-lkml-1.web.codeaurora.org Received: from kanga.kvack.org (kanga.kvack.org [205.233.56.17]) by smtp.lore.kernel.org (Postfix) with ESMTP id A136AC369DC for ; Tue, 29 Apr 2025 06:09:27 +0000 (UTC) Received: by kanga.kvack.org (Postfix) id 733226B0005; Tue, 29 Apr 2025 02:09:25 -0400 (EDT) Received: by kanga.kvack.org (Postfix, from userid 40) id 6BB936B0006; Tue, 29 Apr 2025 02:09:25 -0400 (EDT) X-Delivered-To: int-list-linux-mm@kvack.org Received: by kanga.kvack.org (Postfix, from userid 63042) id 534916B000A; Tue, 29 Apr 2025 02:09:25 -0400 (EDT) X-Delivered-To: linux-mm@kvack.org Received: from relay.hostedemail.com (smtprelay0010.hostedemail.com [216.40.44.10]) by kanga.kvack.org (Postfix) with ESMTP id 323466B0005 for ; Tue, 29 Apr 2025 02:09:25 -0400 (EDT) Received: from smtpin04.hostedemail.com (a10.router.float.18 [10.200.18.1]) by unirelay08.hostedemail.com (Postfix) with ESMTP id 5DD29140DCE for ; Tue, 29 Apr 2025 06:09:26 +0000 (UTC) X-FDA: 83386054332.04.6896B98 Received: from tor.source.kernel.org (tor.source.kernel.org [172.105.4.254]) by imf09.hostedemail.com (Postfix) with ESMTP id EC9FD14000B for ; Tue, 29 Apr 2025 06:09:24 +0000 (UTC) Authentication-Results: imf09.hostedemail.com; dkim=pass header.d=kernel.org header.s=k20201202 header.b=uDWNet5k; dmarc=pass (policy=quarantine) header.from=kernel.org; spf=pass (imf09.hostedemail.com: domain of leon@kernel.org designates 172.105.4.254 as permitted sender) smtp.mailfrom=leon@kernel.org ARC-Seal: i=1; s=arc-20220608; d=hostedemail.com; t=1745906965; a=rsa-sha256; cv=none; b=HqFxTD3684jmRF+BBLDpEfL+gyDA/IGnoy20lpeeKtUX5qfoFb+y1zCISn+aLOZaqO+JWh JXKSLQXlNiWcZs6LOE21hWfwKlcoVZ8yMmBnDlPF+cbIYVvdAAnaybM0nIkLneeO4D9gK3 3IQ9k91i2S1ttAu+fojn1MH0dTNInek= ARC-Authentication-Results: i=1; imf09.hostedemail.com; dkim=pass header.d=kernel.org header.s=k20201202 header.b=uDWNet5k; dmarc=pass (policy=quarantine) header.from=kernel.org; spf=pass (imf09.hostedemail.com: domain of leon@kernel.org designates 172.105.4.254 as permitted sender) smtp.mailfrom=leon@kernel.org ARC-Message-Signature: i=1; a=rsa-sha256; c=relaxed/relaxed; d=hostedemail.com; s=arc-20220608; t=1745906965; h=from:from:sender:reply-to:subject:subject:date:date: message-id:message-id:to:to:cc:cc:mime-version:mime-version: content-type:content-type:content-transfer-encoding: in-reply-to:in-reply-to:references:references:dkim-signature; bh=dxmQw53Z0EJlqww5DRrq4bo00XaglkJeBnX3/YLuqYE=; b=2xGj6rWE+0Ki3AbkURf52f99C3i2p724NscHy2fdh4G39N9QMUjsSnlKJV7wM51NJWWIVE Y9dMW8lEitcC1Zspp00synp1JFMH91l0AX3y57pHY2J3Wr5NS38FATzegDDPu9TofQfYM9 5p9ljPjuiu+2ZcN78TIPu2Je+0HgFwo= Received: from smtp.kernel.org (transwarp.subspace.kernel.org [100.75.92.58]) by tor.source.kernel.org (Postfix) with ESMTP id 77FC661166; Tue, 29 Apr 2025 06:08:58 +0000 (UTC) Received: by smtp.kernel.org (Postfix) with ESMTPSA id B524AC4CEE3; Tue, 29 Apr 2025 06:09:22 +0000 (UTC) DKIM-Signature: v=1; a=rsa-sha256; c=relaxed/simple; d=kernel.org; s=k20201202; t=1745906963; bh=PNH//cVkWbnB2fS1ezNe4WtvGXzhnSIV29jcs1sTleU=; h=Date:From:To:Cc:Subject:References:In-Reply-To:From; b=uDWNet5k2IklSXyclsdy71tc6aXagAuOzOP5Z1vsPDPEw1xByvYmxH/7R8kedffmR K3ASHl60pRtDMWg9p0JxFzxO/3aGKB98mGU3GQziALdCRkObg3CMNM8boWhkMNq5Fb SoRi+onUpbh4+8C7egZp68jWH9bVq8JE2GNbbjKqgOv4PvtvgZAdPdF4XvW8+i2Vac 6q6ESVLxxn8p2qJcdlZLSSecESrOtQeA31phlOs3O0ew9tbbYj38i+QlIYTLHGum0F nK+XngMnFMQRY8/5XFAQ7Xguljv3BaQDuFiBtITTOynS0B/PtXC5n3GH/3WSmM3hcr NiVEF7HYK25Ng== Date: Tue, 29 Apr 2025 09:09:18 +0300 From: Leon Romanovsky To: Baolu Lu Cc: Marek Szyprowski , Jens Axboe , Christoph Hellwig , Keith Busch , Jake Edge , Jonathan Corbet , Jason Gunthorpe , Zhu Yanjun , Robin Murphy , Joerg Roedel , Will Deacon , Sagi Grimberg , Bjorn Helgaas , Logan Gunthorpe , Yishai Hadas , Shameer Kolothum , Kevin Tian , Alex Williamson , =?iso-8859-1?B?Suly9G1l?= Glisse , Andrew Morton , linux-doc@vger.kernel.org, linux-kernel@vger.kernel.org, linux-block@vger.kernel.org, linux-rdma@vger.kernel.org, iommu@lists.linux.dev, linux-nvme@lists.infradead.org, linux-pci@vger.kernel.org, kvm@vger.kernel.org, linux-mm@kvack.org, Niklas Schnelle , Chuck Lever , Luis Chamberlain , Matthew Wilcox , Dan Williams , Kanchan Joshi , Chaitanya Kulkarni , Jason Gunthorpe Subject: Re: [PATCH v10 03/24] iommu: generalize the batched sync after map interface Message-ID: <20250429060918.GK5848@unreal> References: <69da19d2cc5df0be5112f0cf2365a0337b00d873.1745831017.git.leon@kernel.org> MIME-Version: 1.0 Content-Type: text/plain; charset=us-ascii Content-Disposition: inline In-Reply-To: X-Rspamd-Server: rspam10 X-Rspamd-Queue-Id: EC9FD14000B X-Stat-Signature: thb9r1d4fj358cfyxjswoj1xw1tnaciz X-Rspam-User: X-HE-Tag: 1745906964-333784 X-HE-Meta: U2FsdGVkX19oMs0S9lial+7J6RW5jq7FNgXdRuseawRWCqweLixLUwVjzNyZ5VmSyXtipwhSn0XF3VhK2pvPvRMbRdtkjEjwN3SWrl4MhoxcXGjxLy3/+edhZn5zlrQ9fTVrY3Rq6t+BsM9v535F+84UAazP1ljFSKJoZgEUBKWFdO9wdvCPNJ7/LLDUuThIghCVWqBeRVzx585vL7ojKHlH0CFeMNS/1MozN+DlEazLT0USqKtc0oxI17YgoXw/00J01Ubn7ESqK4/CwRCNWX6EMpZz5NHk5AWfi2bUzIL+w/d+PzFlbvHh7FVOiu1ju4B6Yjw0Omew60iYDJaupuCn2AqYZBg65VAydS3cRYQjHH0RjFqQHOT+gSme8Z8GGGykHbMQWzlhx4fsDQJt/sbtlwaxcfqOZvVtYygZKHqn9utwjyF3Wckg3XsILnEoLxElEHf5w3fD0F54kooB15BeYDY2Ja96gnw6z77s67bLqyRhuHQxVTxTWSy8gWZuG9ohsxoBwzGVl9SnfFY35Idl/W3hUuaKGuNmH3otlJUj09jnikxJqCvdhtCPXrR7A1U88ecRfPfuAFddCxfQeTqL7u0CELUdIzeCVzscsF+wOhQm54MyRdcEanZasdT63ipDOkiL2xudRWCtLmWdH0b6dPbKDmI6E8YTo8vdXHYv0o0l9amQiBaGWrMV3geEO2IYDK80pXS6oiS/aK5ST0yMWaat7RCFQjkH8IY09sudrfjfq59+3kAmm7V2JRrO60Wwbu48URJZ9/ujvIhS2VxnOg8XJv+peoWDnV59EXhlAFKSp6VmE8VwkOlPGG3yVcMdYhoYsYYa83lUKPamcFjXC6jwqiqZcn/KcCf+XB9GcKGemkQ4F2koRPqwEeZD5BVWpzAMFaMnk+ZEV5L5d+733AKpdDEDDpnBzdnnu2E= X-Bogosity: Ham, tests=bogofilter, spamicity=0.000000, version=1.2.4 Sender: owner-linux-mm@kvack.org Precedence: bulk X-Loop: owner-majordomo@kvack.org List-ID: List-Subscribe: List-Unsubscribe: On Tue, Apr 29, 2025 at 10:19:46AM +0800, Baolu Lu wrote: > On 4/28/25 17:22, Leon Romanovsky wrote: > > From: Christoph Hellwig > > > > For the upcoming IOVA-based DMA API we want to batch the > > ops->iotlb_sync_map() call after mapping multiple IOVAs from > > dma-iommu without having a scatterlist. Improve the API. > > > > Add a wrapper for the map_sync as iommu_sync_map() so that callers > > don't need to poke into the methods directly. > > > > Formalize __iommu_map() into iommu_map_nosync() which requires the > > caller to call iommu_sync_map() after all maps are completed. > > > > Refactor the existing sanity checks from all the different layers > > into iommu_map_nosync(). > > > > Signed-off-by: Christoph Hellwig > > Acked-by: Will Deacon > > Tested-by: Jens Axboe > > Reviewed-by: Jason Gunthorpe > > Reviewed-by: Luis Chamberlain > > Signed-off-by: Leon Romanovsky > > --- > > drivers/iommu/iommu.c | 65 +++++++++++++++++++------------------------ > > include/linux/iommu.h | 4 +++ > > 2 files changed, 33 insertions(+), 36 deletions(-) > > > > diff --git a/drivers/iommu/iommu.c b/drivers/iommu/iommu.c > > index 4f91a740c15f..02960585b8d4 100644 > > --- a/drivers/iommu/iommu.c > > +++ b/drivers/iommu/iommu.c > > @@ -2443,8 +2443,8 @@ static size_t iommu_pgsize(struct iommu_domain *domain, unsigned long iova, > > return pgsize; > > } > > -static int __iommu_map(struct iommu_domain *domain, unsigned long iova, > > - phys_addr_t paddr, size_t size, int prot, gfp_t gfp) > > +int iommu_map_nosync(struct iommu_domain *domain, unsigned long iova, > > + phys_addr_t paddr, size_t size, int prot, gfp_t gfp) > > { > > const struct iommu_domain_ops *ops = domain->ops; > > unsigned long orig_iova = iova; > > @@ -2453,12 +2453,19 @@ static int __iommu_map(struct iommu_domain *domain, unsigned long iova, > > phys_addr_t orig_paddr = paddr; > > int ret = 0; > > + might_sleep_if(gfpflags_allow_blocking(gfp)); > > + > > if (unlikely(!(domain->type & __IOMMU_DOMAIN_PAGING))) > > return -EINVAL; > > if (WARN_ON(!ops->map_pages || domain->pgsize_bitmap == 0UL)) > > return -ENODEV; > > + /* Discourage passing strange GFP flags */ > > + if (WARN_ON_ONCE(gfp & (__GFP_COMP | __GFP_DMA | __GFP_DMA32 | > > + __GFP_HIGHMEM))) > > + return -EINVAL; > > + > > /* find out the minimum page size supported */ > > min_pagesz = 1 << __ffs(domain->pgsize_bitmap); > > @@ -2506,31 +2513,27 @@ static int __iommu_map(struct iommu_domain *domain, unsigned long iova, > > return ret; > > } > > -int iommu_map(struct iommu_domain *domain, unsigned long iova, > > - phys_addr_t paddr, size_t size, int prot, gfp_t gfp) > > +int iommu_sync_map(struct iommu_domain *domain, unsigned long iova, size_t size) > > { > > const struct iommu_domain_ops *ops = domain->ops; > > - int ret; > > - > > - might_sleep_if(gfpflags_allow_blocking(gfp)); > > - /* Discourage passing strange GFP flags */ > > - if (WARN_ON_ONCE(gfp & (__GFP_COMP | __GFP_DMA | __GFP_DMA32 | > > - __GFP_HIGHMEM))) > > - return -EINVAL; > > + if (!ops->iotlb_sync_map) > > + return 0; > > + return ops->iotlb_sync_map(domain, iova, size); > > +} > > I am wondering whether iommu_sync_map() needs a return value. The > purpose of this callback is just to sync the TLB cache after new > mappings are created, which should effectively be a no-fail operation. > > The definition of iotlb_sync_map in struct iommu_domain_ops seems > unnecessary: > > struct iommu_domain_ops { > ... > int (*iotlb_sync_map)(struct iommu_domain *domain, unsigned long > iova, > size_t size); > ... > }; > > Furthermore, currently no iommu driver implements this callback in a way > that returns a failure. We could clean up the iommu definition in a > subsequent patch series, but for this driver-facing interface, it's > better to get it right from the beginning. I see that s390 is relying on return values: 569 static int s390_iommu_iotlb_sync_map(struct iommu_domain *domain, 570 unsigned long iova, size_t size) 571 { <...> 581 ret = zpci_refresh_trans((u64)zdev->fh << 32, 582 iova, size); 583 /* 584 * let the hypervisor discover invalidated entries 585 * allowing it to free IOVAs and unpin pages 586 */ 587 if (ret == -ENOMEM) { 588 ret = zpci_refresh_all(zdev); 589 if (ret) 590 break; 591 } <...> 595 return ret; 596 } > > Thanks, > baolu