Linux-mm Archive on lore.kernel.org
 help / color / mirror / Atom feed
From: Muchun Song <muchun.song@linux.dev>
To: Muchun Song <songmuchun@bytedance.com>
Cc: Andrew Morton <akpm@linux-foundation.org>,
	Oscar Salvador <osalvador@suse.de>,
	David Hildenbrand <david@kernel.org>,
	Mike Rapoport <rppt@kernel.org>,
	Vlastimil Babka <vbabka@kernel.org>,
	Lorenzo Stoakes <ljs@kernel.org>, Michal Hocko <mhocko@suse.com>,
	David Laight <david.laight.linux@gmail.com>,
	linux-mm@kvack.org, linux-kernel@vger.kernel.org
Subject: Re: [PATCH v2 03/17] mm/mm_init: skip initializing shared vmemmap tail pages
Date: Sun, 26 Jul 2026 12:30:17 +0800	[thread overview]
Message-ID: <5005C7D3-B3ED-4C62-B611-8465D2BC69FD@linux.dev> (raw)
In-Reply-To: <20260720093127.540540-4-songmuchun@bytedance.com>



> On Jul 20, 2026, at 17:31, Muchun Song <songmuchun@bytedance.com> wrote:
> 
> memmap_init_range() initializes every struct page in the target range.
> For compound pages with vmemmap optimization, the tail struct pages are
> backed by a shared vmemmap page.
> 
> Initializing those tail struct pages would overwrite the shared
> vmemmap page contents, requiring users such as HugeTLB to restore the
> metadata afterwards.
> 
> Track the compound order for HVO-backed sections and use that metadata
> to detect struct pages that fall into the shared tail vmemmap range.
> Skip those shared tail pages in memmap_init_range(), then initialize
> pageblock migratetypes for the processed range with a helper after the
> per-page initialization loop.
> 
> The !SPARSEMEM __pfn_to_section() stub is needed only for the build:
> memmap_init_range() references __pfn_to_section() after checking
> pfn_vmemmap_optimizable(), and !SPARSEMEM builds still have to compile
> that code even though pfn_vmemmap_optimizable() folds to false.
> 
> This is a preparatory change for consolidating handling across users of
> vmemmap optimization, and it also avoids redundant initialization of
> shared tail vmemmap pages during early boot.
> 
> Signed-off-by: Muchun Song <songmuchun@bytedance.com>
> ---
> v2:
> - Fold section order tracking into the first user instead of keeping a
>  standalone API-only patch (suggested by Mike Rapoport)
> - Rename page_vmemmap_optimizable() to pfn_vmemmap_optimizable() and
>  pass a PFN directly (suggested by Mike Rapoport)
> - Initialize pageblock migratetypes from a helper after the per-page
>  loop (suggested by Mike Rapoport)
> - Use a 1G PFN chunk for cond_resched() in the pageblock helper
>  (suggested by Mike Rapoport)
> - Guard section_order() with CONFIG_HUGETLB_PAGE_OPTIMIZE_VMEMMAP so
>  it returns 0 when HVO is disabled and lets the compiler optimize the
>  code as much as possible (suggested by Mike Rapoport)
> - Explain why the !SPARSEMEM __pfn_to_section() stub belongs here
>  (suggested by Mike Rapoport)
> ---
> include/linux/mmzone.h | 14 ++++++++++++++
> mm/mm_init.c           | 33 +++++++++++++++++----------------
> mm/sparse.h            | 23 +++++++++++++++++++++++
> 3 files changed, 54 insertions(+), 16 deletions(-)
> 
> diff --git a/include/linux/mmzone.h b/include/linux/mmzone.h
> index 82b0155d886f..2a32101d55e6 100644
> --- a/include/linux/mmzone.h
> +++ b/include/linux/mmzone.h
> @@ -2011,6 +2011,14 @@ struct mem_section {
> unsigned long section_mem_map;
> 
> struct mem_section_usage *usage;
> +#ifdef CONFIG_HUGETLB_PAGE_OPTIMIZE_VMEMMAP
> + 	/*
> +	 * Normally, sections hold regular (order-0) pages. However, for
> +	 * sections with HVO enabled, this tracks the compound page order
> +	 * to enable deduplication of redundant vmemmap pages.
> +	 */
> + 	unsigned int order;
> +#endif
> #ifdef CONFIG_PAGE_EXTENSION
> 	/*
> 	 * If SPARSEMEM, pgdat doesn't have page_ext pointer. We use
> @@ -2365,8 +2373,14 @@ static inline unsigned long next_present_section_nr(unsigned long section_nr)
> #endif
> 
> #else
> +struct mem_section;
> +
> #define sparse_vmemmap_init_nid_early(_nid) do {} while (0)
> #define pfn_in_present_section pfn_valid
> +static inline struct mem_section *__pfn_to_section(unsigned long pfn)
> +{
> + 	return NULL;
> +}

I'd like to propose an alternative implementation that doesn't require
exposing the mem_section. The idea is to add a new helper function,
pfn_to_section_order(), so that for non-sparse-memory configurations,
the mem_section concept stays hidden internally. I'd really appreciate
any thoughts or concerns — if everyone is comfortable with it, I can go
ahead and implement this in the next version.

Muchun,
Thanks.



  reply	other threads:[~2026-07-26  4:30 UTC|newest]

Thread overview: 27+ messages / expand[flat|nested]  mbox.gz  Atom feed  top
2026-07-20  9:31 [PATCH v2 00/17] mm: Introduce section-based vmemmap optimization for HugeTLB Muchun Song
2026-07-20  9:31 ` [PATCH v2 01/17] mm/sparse: relax struct mem_section size constraints Muchun Song
2026-07-20  9:31 ` [PATCH v2 02/17] mm/sparse-vmemmap: rename HVO order macros Muchun Song
2026-07-30 10:57   ` Mike Rapoport
2026-07-20  9:31 ` [PATCH v2 03/17] mm/mm_init: skip initializing shared vmemmap tail pages Muchun Song
2026-07-26  4:30   ` Muchun Song [this message]
2026-07-30 10:48     ` Mike Rapoport
2026-07-30 13:09       ` Muchun Song
2026-07-30 14:32         ` Mike Rapoport
2026-07-20  9:31 ` [PATCH v2 04/17] mm/sparse-vmemmap: initialize shared tail vmemmap pages on allocation Muchun Song
2026-07-20  9:31 ` [PATCH v2 05/17] mm/sparse-vmemmap: support section-based vmemmap accounting Muchun Song
2026-07-20  9:31 ` [PATCH v2 06/17] mm/mm_init: factor out pfn_to_zone() Muchun Song
2026-07-20 11:04   ` Muchun Song
2026-07-20 11:05   ` Muchun Song
2026-07-30 10:57   ` Mike Rapoport
2026-07-20  9:31 ` [PATCH v2 07/17] mm/sparse-vmemmap: move vmemmap_get_tail() before PTE population Muchun Song
2026-07-20  9:31 ` [PATCH v2 08/17] mm/sparse-vmemmap: support section-based vmemmap optimization Muchun Song
2026-07-20 11:10   ` Muchun Song
2026-07-20  9:31 ` [PATCH v2 09/17] mm/sparse: initialize memory sections earlier Muchun Song
2026-07-20  9:31 ` [PATCH v2 10/17] mm/hugetlb: switch HugeTLB to section-based vmemmap optimization Muchun Song
2026-07-20  9:31 ` [PATCH v2 11/17] mm/sparse-vmemmap: remove SPARSEMEM_VMEMMAP_PREINIT support Muchun Song
2026-07-20  9:31 ` [PATCH v2 12/17] mm/sparse: inline usemap allocation into sparse_init_nid() Muchun Song
2026-07-20  9:31 ` [PATCH v2 13/17] mm/sparse: remove section_map_size() Muchun Song
2026-07-20  9:31 ` [PATCH v2 14/17] mm/hugetlb: remove HUGE_BOOTMEM_HVO Muchun Song
2026-07-20  9:31 ` [PATCH v2 15/17] mm/hugetlb: remove HUGE_BOOTMEM_CMA Muchun Song
2026-07-20  9:31 ` [PATCH v2 16/17] mm/hugetlb: localize struct huge_bootmem_page Muchun Song
2026-07-20  9:31 ` [PATCH v2 17/17] mm/hugetlb: localize HUGE_BOOTMEM_ZONES_VALID Muchun Song

Reply instructions:

You may reply publicly to this message via plain-text email
using any one of the following methods:

* Save the following mbox file, import it into your mail client,
  and reply-to-all from there: mbox

  Avoid top-posting and favor interleaved quoting:
  https://en.wikipedia.org/wiki/Posting_style#Interleaved_style

* Reply using the --to, --cc, and --in-reply-to
  switches of git-send-email(1):

  git send-email \
    --in-reply-to=5005C7D3-B3ED-4C62-B611-8465D2BC69FD@linux.dev \
    --to=muchun.song@linux.dev \
    --cc=akpm@linux-foundation.org \
    --cc=david.laight.linux@gmail.com \
    --cc=david@kernel.org \
    --cc=linux-kernel@vger.kernel.org \
    --cc=linux-mm@kvack.org \
    --cc=ljs@kernel.org \
    --cc=mhocko@suse.com \
    --cc=osalvador@suse.de \
    --cc=rppt@kernel.org \
    --cc=songmuchun@bytedance.com \
    --cc=vbabka@kernel.org \
    /path/to/YOUR_REPLY

  https://kernel.org/pub/software/scm/git/docs/git-send-email.html

* If your mail client supports setting the In-Reply-To header
  via mailto: links, try the mailto: link
Be sure your reply has a Subject: header at the top and a blank line before the message body.
This is a public inbox, see mirroring instructions
for how to clone and mirror all data and code used for this inbox