All of lore.kernel.org
 help / color / mirror / Atom feed
From: Yeoreum Yun <yeoreum.yun@arm.com>
To: kasong@tencent.com
Cc: linux-mm@kvack.org, linux-kernel@vger.kernel.org,
	Andrew Morton <akpm@linux-foundation.org>,
	David Hildenbrand <david@kernel.org>,
	Lorenzo Stoakes <ljs@kernel.org>, Zi Yan <ziy@nvidia.com>,
	Baolin Wang <baolin.wang@linux.alibaba.com>,
	"Liam R. Howlett" <liam@infradead.org>,
	Nico Pache <nico.pache@linux.dev>,
	Ryan Roberts <ryan.roberts@arm.com>, Dev Jain <dev.jain@arm.com>,
	Lance Yang <lance.yang@linux.dev>,
	Usama Arif <usama.arif@linux.dev>,
	Vlastimil Babka <vbabka@kernel.org>,
	Mike Rapoport <rppt@kernel.org>,
	Suren Baghdasaryan <surenb@google.com>,
	Michal Hocko <mhocko@suse.com>, Chris Li <chrisl@kernel.org>,
	Kemeng Shi <shikemeng@huaweicloud.com>,
	Nhat Pham <nphamcs@gmail.com>, Baoquan He <baoquan.he@linux.dev>,
	Barry Song <baohua@kernel.org>,
	Youngjun Park <youngjun.park@lge.com>,
	Shivam Kalra <shivamkalra98@zohomail.in>,
	Kairui Song <ryncsn@gmail.com>
Subject: Re: [PATCH v3 00/18] mm/huge_memory: clean up folio split and lift swapcache split limits
Date: Thu, 27 Aug 2026 13:19:16 +0100	[thread overview]
Message-ID: <apArRLFDroxOKWUV@e129823.arm.com> (raw)
In-Reply-To: <20260821-swap-thp-cleanup-v3-0-9b43f5163238@tencent.com>

> This series clean up the split code, add better swap cache split support
> for mappingless, large order, uniform and non-uniform split.  Generic
> performance is on par or slightly better, and stack usage is reduced.
> 
> The swap cache infrastructure can handle non-uniform or high order folio
> replace, so there is no reason for either restriction from the THP side.
> What stands in the way is the mixed anon/file folio split routine,
> which makes lifting the restrictions hard to follow, and it already
> carries some buggy or redundant checks.
> 
> So this series cleans up the split path and separates anon and file
> splitting into two helpers.  The file split path never sees a swap
> cache folio, and that is now enforced up front: a folio that is both
> in the page cache and the swap cache can only be a shmem folio, which
> remains unsupported and is rejected early.  That helps to rule out swap
> cache handling in that part completely.  Only the anon split path
> handles swap cache folios, with an anon mapping or mappingless:
> either way the splitting is similar, and non-uniform split is
> supported as well.
> 
> Order-1 is still forbidden for swap cache splitting.  In theory it is
> doable for shmem swap cache folios, but a mappingless swap cache
> folio cannot currently be told apart from a shmem one, so forbid it
> for all swap cache folios for now.
> 
> Testing:
> 
> The in-tree split_huge_page_test selftest (uniform, non-uniform and
> in-folio-offset splits of anon and pagecache folios) passes 62/62 on
> the patched kernel.
> 
> ftrace function_graph tracing filtered on __folio_split() was used to
> compare per-call durations between the base and the patched kernel on
> the same x86-64 box (interleaved runs across alternating reboots;
> 135 split calls per run, 50 test run):
> 
> Before: 67.9 us, stddev: 1.59
> After:  66.4 us, stddev: 1.19
> 
> The patched kernel is slightly faster. The stack usage is also reduced
> by about ~10%, with a very slight growth of huge_memory.o.
> 
> Signed-off-by: Kairui Song <kasong@tencent.com>
> ---
> Changes in v3:
> - Get rid of for_each_folio_safe and open code it.
> - Check if the folio is mapped before freeing it swap cache to avoid
>   potential performance lose.
> - Initial test and binary analyze showed everything is very similiar to
>   previously series.
> - Drop the redundant mapping argument of __split_frozen_folio
> - Link to v2: https://patch.msgid.link/20260813-swap-thp-cleanup-v2-0-d2ee48c6aa49@tencent.com
> 
> Changes in v2:
> - Return -EBUSY instead of -EINVAL for swap cache & shmem folio split
>   attempt.
> - Introduce a for_each_folio_safe macro to dedupliate the code and
>   hightlight the reason we need to keep the iterate safe from folio
>   freeing. [ Zi Yan ]
> - Rename __split_unmapped_folio() to __split_frozen_folio [ Zi Yan ]
> - Rename __folio_freeze_split_unmap_anon. [ Zi Yan ]
> - Several comment improments [ Zi Yan ]
> - Drop an unused do_lru argument.
> - Previouse test results are basically unchanged, stack usage reduced,
>   object very slightly larger.
> - Link to v1: https://patch.msgid.link/20260808-swap-thp-cleanup-v1-0-689939a7ccc3@tencent.com
> 
> ---
> Kairui Song (18):
>       mm/swap: fix off-by-one in swap cache replace sanity check
>       mm/huge_memory: fix rejection of swap cache folios with a mapping
>       mm/huge_memory: invert folio_ref_freeze() check to reduce indentation
>       mm/huge_memory: split the routine for splitting anon and file folio
>       mm/huge_memory: rename __split_unmapped_folio() to __split_frozen_folio()
>       mm/huge_memory: consolidate irq and locking for folio split
>       mm/huge_memory: move EOF trimming into the file split helper
>       mm/huge_memory: move unmap and remap into the split helpers
>       mm/huge_memory: move anon_vma and filemap management into split helpers
>       mm/huge_memory: move memcg switch into the file split helper
>       mm/huge_memory: allow splitting mappingless swap cache folios
>       mm/huge_memory: add kerneldoc for the split helpers
>       mm/huge_memory: drop the unused do_lru argument of the file split helper
>       mm/huge_memory: clean up after-split folio freeing in __folio_split
>       mm/huge_memory: lift order-0 restriction for swapcache split
>       mm/huge_memory: clarify supported split orders in comment
>       mm/huge_memory: count only swap cache refs in anon folio split
>       mm/huge_memory: drop the redundant mapping argument of __split_frozen_folio
> 
>  mm/huge_memory.c | 633 +++++++++++++++++++++++++++++--------------------------
>  mm/swap_state.c  |   3 +-
>  2 files changed, 338 insertions(+), 298 deletions(-)
> ---
> base-commit: 4b2ae13f3393ef4b4bce0021e8762790354f369f
> change-id: 20260804-swap-thp-cleanup-6ce2be6cf3b8
> 
> Best regards,
> --  
> Kairui Song <kasong@tencent.com>

Nice cleanup. this series look good to me.

Reviewed-by: Yeoreum Yun <yeoreum.yun@arm.com>

-- 
Sincerely,
Yeoreum Yun


      parent reply	other threads:[~2026-08-27 12:19 UTC|newest]

Thread overview: 99+ messages / expand[flat|nested]  mbox.gz  Atom feed  top
2026-08-20 18:55 [PATCH v3 00/18] mm/huge_memory: clean up folio split and lift swapcache split limits Kairui Song via B4 Relay
2026-08-20 18:55 ` Kairui Song
2026-08-20 18:55 ` [PATCH v3 01/18] mm/swap: fix off-by-one in swap cache replace sanity check Kairui Song via B4 Relay
2026-08-20 18:55   ` Kairui Song
2026-08-23  8:53   ` Barry Song
2026-08-27 14:13   ` Kiryl Shutsemau
2026-08-27 15:57   ` David Hildenbrand (Arm)
2026-08-20 18:55 ` [PATCH v3 02/18] mm/huge_memory: fix rejection of swap cache folios with a mapping Kairui Song via B4 Relay
2026-08-20 18:55   ` Kairui Song
2026-08-27  8:46   ` Barry Song
2026-08-27  9:31     ` Kairui Song
2026-08-27 14:33   ` Kiryl Shutsemau
2026-08-27 15:58     ` David Hildenbrand (Arm)
2026-08-27 14:36   ` Kiryl Shutsemau
2026-08-27 16:01     ` David Hildenbrand (Arm)
2026-08-27 16:00   ` David Hildenbrand (Arm)
2026-08-27 17:02     ` Kiryl Shutsemau
2026-08-27 17:11       ` David Hildenbrand (Arm)
2026-08-27 17:07     ` Kairui Song
2026-08-20 18:55 ` [PATCH v3 03/18] mm/huge_memory: invert folio_ref_freeze() check to reduce indentation Kairui Song via B4 Relay
2026-08-20 18:55   ` Kairui Song
2026-08-27  9:05   ` Barry Song
2026-08-27 14:39   ` Kiryl Shutsemau
2026-08-27 16:02   ` David Hildenbrand (Arm)
2026-08-20 18:55 ` [PATCH v3 04/18] mm/huge_memory: split the routine for splitting anon and file folio Kairui Song via B4 Relay
2026-08-20 18:55   ` Kairui Song
2026-08-26  1:31   ` Zi Yan
2026-08-27 14:58   ` Kiryl Shutsemau
2026-08-27 17:19     ` Kairui Song
2026-08-27 16:19   ` David Hildenbrand (Arm)
2026-08-27 17:17     ` Kairui Song
2026-08-27 19:05       ` David Hildenbrand (Arm)
2026-08-20 18:55 ` [PATCH v3 05/18] mm/huge_memory: rename __split_unmapped_folio() to __split_frozen_folio() Kairui Song via B4 Relay
2026-08-20 18:55   ` Kairui Song
2026-08-27 15:06   ` Kiryl Shutsemau
2026-08-27 16:22     ` David Hildenbrand (Arm)
2026-08-27 17:21       ` Kairui Song
2026-08-20 18:55 ` [PATCH v3 06/18] mm/huge_memory: consolidate irq and locking for folio split Kairui Song via B4 Relay
2026-08-20 18:55   ` Kairui Song
2026-08-27 15:15   ` Kiryl Shutsemau
2026-08-27 16:24   ` David Hildenbrand (Arm)
2026-08-20 18:55 ` [PATCH v3 07/18] mm/huge_memory: move EOF trimming into the file split helper Kairui Song via B4 Relay
2026-08-20 18:55   ` Kairui Song
2026-08-27 16:25   ` David Hildenbrand (Arm)
2026-08-31  1:08   ` Kiryl Shutsemau
2026-08-20 18:55 ` [PATCH v3 08/18] mm/huge_memory: move unmap and remap into the split helpers Kairui Song via B4 Relay
2026-08-20 18:55   ` Kairui Song
2026-08-27 16:32   ` David Hildenbrand (Arm)
2026-08-27 17:35     ` Kairui Song
2026-08-27 19:09       ` David Hildenbrand (Arm)
2026-08-20 18:55 ` [PATCH v3 09/18] mm/huge_memory: move anon_vma and filemap management into " Kairui Song via B4 Relay
2026-08-20 18:55   ` Kairui Song
2026-08-27 16:36   ` David Hildenbrand (Arm)
2026-08-27 17:37     ` Kairui Song
2026-08-30 15:23     ` Kairui Song
2026-09-07 12:36       ` David Hildenbrand (Arm)
2026-09-07 17:55         ` Kairui Song
2026-09-07 19:56           ` David Hildenbrand (Arm)
2026-08-20 18:55 ` [PATCH v3 10/18] mm/huge_memory: move memcg switch into the file split helper Kairui Song via B4 Relay
2026-08-20 18:55   ` Kairui Song
2026-08-27 16:37   ` David Hildenbrand (Arm)
2026-08-20 18:55 ` [PATCH v3 11/18] mm/huge_memory: allow splitting mappingless swap cache folios Kairui Song via B4 Relay
2026-08-20 18:55   ` Kairui Song
2026-08-26  1:55   ` Zi Yan
2026-08-27 16:41   ` David Hildenbrand (Arm)
2026-08-27 17:41     ` Kairui Song
2026-08-27 19:16       ` David Hildenbrand (Arm)
2026-08-27 19:29         ` David Hildenbrand (Arm)
2026-08-20 18:55 ` [PATCH v3 12/18] mm/huge_memory: add kerneldoc for the split helpers Kairui Song via B4 Relay
2026-08-20 18:55   ` Kairui Song
2026-08-26  1:57   ` Zi Yan
2026-08-27 16:45   ` David Hildenbrand (Arm)
2026-08-27 17:43     ` Kairui Song
2026-08-20 18:55 ` [PATCH v3 13/18] mm/huge_memory: drop the unused do_lru argument of the file split helper Kairui Song via B4 Relay
2026-08-20 18:55   ` Kairui Song
2026-08-26  1:57   ` Zi Yan
2026-08-27 16:45   ` David Hildenbrand (Arm)
2026-08-20 18:55 ` [PATCH v3 14/18] mm/huge_memory: clean up after-split folio freeing in __folio_split Kairui Song via B4 Relay
2026-08-20 18:55   ` Kairui Song
2026-08-27 16:48   ` David Hildenbrand (Arm)
2026-08-27 17:47     ` Kairui Song
2026-08-27 19:33       ` David Hildenbrand (Arm)
2026-08-20 18:55 ` [PATCH v3 15/18] mm/huge_memory: lift order-0 restriction for swapcache split Kairui Song via B4 Relay
2026-08-20 18:55   ` Kairui Song
2026-08-27 16:51   ` David Hildenbrand (Arm)
2026-08-27 17:48     ` Kairui Song
2026-08-27 20:22       ` David Hildenbrand (Arm)
2026-08-30 13:19         ` Kairui Song
2026-08-20 18:55 ` [PATCH v3 16/18] mm/huge_memory: clarify supported split orders in comment Kairui Song via B4 Relay
2026-08-20 18:55   ` Kairui Song
2026-08-26  2:02   ` Zi Yan
2026-08-27 17:48     ` Kairui Song
2026-08-27 16:54   ` David Hildenbrand (Arm)
2026-08-20 18:55 ` [PATCH v3 17/18] mm/huge_memory: count only swap cache refs in anon folio split Kairui Song via B4 Relay
2026-08-20 18:55   ` Kairui Song
2026-08-20 18:55 ` [PATCH v3 18/18] mm/huge_memory: drop the redundant mapping argument of __split_frozen_folio Kairui Song via B4 Relay
2026-08-20 18:55   ` Kairui Song
2026-08-26  2:06   ` Zi Yan
2026-08-27 12:19 ` Yeoreum Yun [this message]

Reply instructions:

You may reply publicly to this message via plain-text email
using any one of the following methods:

* Save the following mbox file, import it into your mail client,
  and reply-to-all from there: mbox

  Avoid top-posting and favor interleaved quoting:
  https://en.wikipedia.org/wiki/Posting_style#Interleaved_style

* Reply using the --to, --cc, and --in-reply-to
  switches of git-send-email(1):

  git send-email \
    --in-reply-to=apArRLFDroxOKWUV@e129823.arm.com \
    --to=yeoreum.yun@arm.com \
    --cc=akpm@linux-foundation.org \
    --cc=baohua@kernel.org \
    --cc=baolin.wang@linux.alibaba.com \
    --cc=baoquan.he@linux.dev \
    --cc=chrisl@kernel.org \
    --cc=david@kernel.org \
    --cc=dev.jain@arm.com \
    --cc=kasong@tencent.com \
    --cc=lance.yang@linux.dev \
    --cc=liam@infradead.org \
    --cc=linux-kernel@vger.kernel.org \
    --cc=linux-mm@kvack.org \
    --cc=ljs@kernel.org \
    --cc=mhocko@suse.com \
    --cc=nico.pache@linux.dev \
    --cc=nphamcs@gmail.com \
    --cc=rppt@kernel.org \
    --cc=ryan.roberts@arm.com \
    --cc=ryncsn@gmail.com \
    --cc=shikemeng@huaweicloud.com \
    --cc=shivamkalra98@zohomail.in \
    --cc=surenb@google.com \
    --cc=usama.arif@linux.dev \
    --cc=vbabka@kernel.org \
    --cc=youngjun.park@lge.com \
    --cc=ziy@nvidia.com \
    /path/to/YOUR_REPLY

  https://kernel.org/pub/software/scm/git/docs/git-send-email.html

* If your mail client supports setting the In-Reply-To header
  via mailto: links, try the mailto: link
Be sure your reply has a Subject: header at the top and a blank line before the message body.
This is an external index of several public inboxes,
see mirroring instructions on how to clone and mirror
all data and code used by this external index.