Linux-mm Archive on lore.kernel.org
 help / color / mirror / Atom feed
From: Usama Arif <usama.arif@linux.dev>
To: Zi Yan <ziy@nvidia.com>
Cc: David Hildenbrand <david@kernel.org>,
	"Matthew Wilcox (Oracle)" <willy@infradead.org>,
	Andrew Morton <akpm@linux-foundation.org>,
	Muchun Song <muchun.song@linux.dev>,
	Lorenzo Stoakes <ljs@kernel.org>,
	"Liam R. Howlett" <liam@infradead.org>,
	Vlastimil Babka <vbabka@kernel.org>,
	Mike Rapoport <rppt@kernel.org>,
	Suren Baghdasaryan <surenb@google.com>,
	Michal Hocko <mhocko@suse.com>,
	Baolin Wang <baolin.wang@linux.alibaba.com>,
	Nico Pache <nico.pache@linux.dev>,
	Ryan Roberts <ryan.roberts@arm.com>, Dev Jain <dev.jain@arm.com>,
	Barry Song <baohua@kernel.org>, Lance Yang <lance.yang@linux.dev>,
	Gregory Price <gourry@gourry.net>,
	Ying Huang <ying.huang@linux.alibaba.com>,
	Alistair Popple <apopple@nvidia.com>,
	Johannes Weiner <hannes@cmpxchg.org>,
	Qi Zheng <qi.zheng@linux.dev>,
	Shakeel Butt <shakeel.butt@linux.dev>,
	Kairui Song <kasong@tencent.com>,
	linux-mm@kvack.org, linux-kernel@vger.kernel.org,
	Minchan Kim <minchan@kernel.org>,
	Sergey Senozhatsky <senozhatsky@chromium.org>
Subject: Re: [PATCH RFC 01/14] mm/zsmalloc: replace PG_private with pointer comparison
Date: Sun, 2 Aug 2026 13:05:21 +0100	[thread overview]
Message-ID: <b3f03eba-02ed-4763-86af-cafc23e0006f@linux.dev> (raw)
In-Reply-To: <DKE0CO9RE3TN.KYL8VQE6QW66@nvidia.com>



On 02/08/2026 00:49, Zi Yan wrote:
> On Sat Aug 1, 2026 at 10:12 AM EDT, Usama Arif wrote:
>> On Fri, 31 Jul 2026 22:13:24 -0400 Zi Yan <ziy@nvidia.com> wrote:
>>
>>> zsmalloc uses PG_private to indicate first zpdesc in the zspage chain.
>>> Replace it with zpdesc->zspage->first_zpdesc == zpdesc. The check,
>>> is_first_zpdesc(), is only used in VM_BUG_ON(), so performance impact
>>> should be negligible.
>>>
>>> It prepares for a future commit that remove PG_private.
>>>
>>> No functional change intended.
>>>
>>> Assisted-by: Claude:claude-opus-4-8
>>> Assisted-by: Codex:gpt-5
>>> Signed-off-by: Zi Yan <ziy@nvidia.com>
>>> To: Minchan Kim <minchan@kernel.org>
>>> To: Sergey Senozhatsky <senozhatsky@chromium.org>
>>> To: Andrew Morton <akpm@linux-foundation.org>
>>> Cc: linux-mm@kvack.org
>>> Cc: linux-kernel@vger.kernel.org
>>> ---
>>>  mm/zpdesc.h   |  2 +-
>>>  mm/zsmalloc.c | 15 +++------------
>>>  2 files changed, 4 insertions(+), 13 deletions(-)
>>>
>>
>> create_page_chain() sets zpdesc->zspage = zspage before assigning zspage->first_zpdesc,
>> so LGTM.
>>
>> I think the assertion in get_first_zpdesc()
>> should be changed to
>>
>> VM_BUG_ON_PAGE(first_zpdesc->zspage != zspage, zpdesc_page(first_zpdesc));
>>
>> in the current check, we are checking
>>
>> zspage->first_zpdesc->zspage->first_zpdesc == zspage->first_zpdesc
>>
>> which is just testing the backpointer.
> 
> Got it. Thank you for pointing this out.
> 
> is_first_zpdesc() is also used in obj_allocated() for the same check and
> getting rid of it in get_first_zpdesc() causes inconsistency.
> 
> How about the changes below on top of this patch? is_first_zpdesc() can
> be used for both sites and looks cleaner. And VM_BUG_ON_PAGE is replaced
> with VM_WARN_ON_ONCE_PAGE.
> 
> 
> diff --git a/mm/zsmalloc.c b/mm/zsmalloc.c
> index e8ef227624efa..f021d2df99404 100644
> --- a/mm/zsmalloc.c
> +++ b/mm/zsmalloc.c
> @@ -471,9 +471,10 @@ static void record_obj(unsigned long handle, unsigned long obj)
>  	WRITE_ONCE(*(unsigned long *)handle, obj);
>  }
>  
> -static inline bool __maybe_unused is_first_zpdesc(struct zpdesc *zpdesc)
> +static inline bool __maybe_unused is_first_zpdesc(struct zpdesc *zpdesc,
> +						  struct zspage *zspage)
>  {
> -	return zpdesc->zspage->first_zpdesc == zpdesc;
> +	return zpdesc->zspage == zspage && zspage->first_zpdesc == zpdesc;
>  }
>  
>  /* Protected by class->lock */
> @@ -491,7 +492,7 @@ static struct zpdesc *get_first_zpdesc(struct zspage *zspage)
>  {
>  	struct zpdesc *first_zpdesc = zspage->first_zpdesc;
>  
> -	VM_BUG_ON_PAGE(!is_first_zpdesc(first_zpdesc), zpdesc_page(first_zpdesc));
> +	VM_WARN_ON_ONCE_PAGE(is_first_zpdesc(first_zpdesc, zspage), zpdesc_page(first_zpdesc));

mhmm do you mean 

VM_WARN_ON_ONCE_PAGE(!is_first_zpdesc(first_zpdesc, zspage), zpdesc_page(first_zpdesc)); 

here? ! is missing I think?

>  	return first_zpdesc;
>  }
>  
> @@ -828,7 +829,7 @@ static inline bool obj_allocated(struct zpdesc *zpdesc, void *obj,
>  	struct zspage *zspage = get_zspage(zpdesc);
>  
>  	if (unlikely(ZsHugePage(zspage))) {
> -		VM_BUG_ON_PAGE(!is_first_zpdesc(zpdesc), zpdesc_page(zpdesc));
> +		VM_WARN_ON_ONCE_PAGE(is_first_zpdesc(zpdesc, zspage), zpdesc_page(zpdesc));

same here? ! is missing?


>  		handle = zpdesc->handle;
>  	} else
>  		handle = *(unsigned long *)obj;
> 
>>
>> With the above VM_BUG_ON check change, please feel free to add:
>>
>> Acked-by: Usama Arif <usama.arif@linux.dev>
>>
>>
>>> diff --git a/mm/zpdesc.h b/mm/zpdesc.h
>>> index b8258dc78548d..4fd81c2e80769 100644
>>> --- a/mm/zpdesc.h
>>> +++ b/mm/zpdesc.h
>>> @@ -26,8 +26,8 @@
>>>   * with memcg_data.
>>>   *
>>>   * Page flags used:
>>> - * * PG_private identifies the first component page.
>>>   * * PG_locked is used by page migration code.
>>> + * The first component page has zpdesc->zspage->first_zpdesc == zpdesc
>>>   */
>>>  struct zpdesc {
>>>  	unsigned long flags;
>>> diff --git a/mm/zsmalloc.c b/mm/zsmalloc.c
>>> index 8204b76f78308..e8ef227624efa 100644
>>> --- a/mm/zsmalloc.c
>>> +++ b/mm/zsmalloc.c
>>> @@ -290,11 +290,6 @@ struct zs_pool {
>>>  	atomic_t compaction_in_progress;
>>>  };
>>>  
>>> -static inline void zpdesc_set_first(struct zpdesc *zpdesc)
>>> -{
>>> -	SetPagePrivate(zpdesc_page(zpdesc));
>>> -}
>>> -
>>>  static inline void zpdesc_inc_zone_page_state(struct zpdesc *zpdesc)
>>>  {
>>>  	inc_zone_page_state(zpdesc_page(zpdesc), NR_ZSPAGES);
>>> @@ -478,7 +473,7 @@ static void record_obj(unsigned long handle, unsigned long obj)
>>>  
>>>  static inline bool __maybe_unused is_first_zpdesc(struct zpdesc *zpdesc)
>>>  {
>>> -	return PagePrivate(zpdesc_page(zpdesc));
>>> +	return zpdesc->zspage->first_zpdesc == zpdesc;
>>>  }
>>>  
>>>  /* Protected by class->lock */
>>> @@ -848,9 +843,6 @@ static inline bool obj_allocated(struct zpdesc *zpdesc, void *obj,
>>>  
>>>  static void reset_zpdesc(struct zpdesc *zpdesc)
>>>  {
>>> -	struct page *page = zpdesc_page(zpdesc);
>>> -
>>> -	ClearPagePrivate(page);
>>>  	zpdesc->zspage = NULL;
>>>  	zpdesc->next = NULL;
>>>  	/* PageZsmalloc is sticky until the page is freed to the buddy. */
>>> @@ -1001,8 +993,8 @@ static void create_page_chain(struct size_class *class, struct zspage *zspage,
>>>  	 * 1. all pages are linked together using zpdesc->next
>>>  	 * 2. each sub-page point to zspage using zpdesc->zspage
>>>  	 *
>>> -	 * we set PG_private to identify the first zpdesc (i.e. no other zpdesc
>>> -	 * has this flag set).
>>> +	 * The first zpdesc has its zspage->first_zpdesc set to itself, no
>>> +	 * other zpdesc has this set.
>>>  	 */
>>>  	for (i = 0; i < nr_zpdescs; i++) {
>>>  		zpdesc = zpdescs[i];
>>> @@ -1010,7 +1002,6 @@ static void create_page_chain(struct size_class *class, struct zspage *zspage,
>>>  		zpdesc->next = NULL;
>>>  		if (i == 0) {
>>>  			zspage->first_zpdesc = zpdesc;
>>> -			zpdesc_set_first(zpdesc);
>>>  			if (unlikely(class->objs_per_zspage == 1 &&
>>>  					class->pages_per_zspage == 1))
>>>  				SetZsHugePage(zspage);
>>>
>>> -- 
>>> 2.53.0
>>>
>>>
> 
> 
> 
> 



  reply	other threads:[~2026-08-02 12:05 UTC|newest]

Thread overview: 25+ messages / expand[flat|nested]  mbox.gz  Atom feed  top
2026-08-01  2:13 [PATCH RFC 00/14] Remove PG_private by using page/folio->private checks instead Zi Yan
2026-08-01  2:13 ` [PATCH RFC 01/14] mm/zsmalloc: replace PG_private with pointer comparison Zi Yan
2026-08-01 14:12   ` Usama Arif
2026-08-01 23:49     ` Zi Yan
2026-08-02 12:05       ` Usama Arif [this message]
2026-08-01  2:13 ` [PATCH RFC 02/14] perf/ring_buffer: stop using PG_private as AUX page high-order marker Zi Yan
2026-08-01 14:32   ` Usama Arif
2026-08-02  1:20     ` Zi Yan
2026-08-01  2:13 ` [PATCH RFC 03/14] xen/grant-table: stop setting PG_private on pages for grant mapping Zi Yan
2026-08-01 14:42   ` Usama Arif
2026-08-02  1:24     ` Zi Yan
2026-08-01  2:13 ` [PATCH RFC 04/14] fs/crypto: stop setting PG_private on bounce page Zi Yan
2026-08-01 14:52   ` Usama Arif
2026-08-02  1:29     ` Zi Yan
2026-08-01  2:13 ` [PATCH RFC 05/14] mm/hugetlb: use direct assignment instead of folio_change_private() Zi Yan
2026-08-02 12:16   ` Usama Arif
2026-08-01  2:13 ` [PATCH RFC 06/14] fs/f2fs: stop using PG_private Zi Yan
2026-08-01  2:13 ` [PATCH RFC 07/14] fs/erofs: mm/pagemap: add readahead_folio_reverse() to avoid folio->private Zi Yan
2026-08-01  2:13 ` [PATCH RFC 08/14] fs/erofs: use folio_attach/detach_private() instead of direct assignment Zi Yan
2026-08-01  2:13 ` [PATCH RFC 09/14] mm/page-flags: check page/folio->private instead of PG_private Zi Yan
2026-08-01  2:13 ` [PATCH RFC 10/14] mm/page-flags: introduce folio_test_fs_private() Zi Yan
2026-08-01  2:13 ` [PATCH RFC 11/14] treewide: remove folio_set/clear_private() Zi Yan
2026-08-01  2:13 ` [PATCH RFC 12/14] treewide: replace PagePrivate() with page_private() Zi Yan
2026-08-01  2:13 ` [PATCH RFC 13/14] treewide: adjust comments on PagePrivate and PG_private Zi Yan
2026-08-01  2:13 ` [PATCH RFC 14/14] mm/page-flags: remove PG_private Zi Yan

Reply instructions:

You may reply publicly to this message via plain-text email
using any one of the following methods:

* Save the following mbox file, import it into your mail client,
  and reply-to-all from there: mbox

  Avoid top-posting and favor interleaved quoting:
  https://en.wikipedia.org/wiki/Posting_style#Interleaved_style

* Reply using the --to, --cc, and --in-reply-to
  switches of git-send-email(1):

  git send-email \
    --in-reply-to=b3f03eba-02ed-4763-86af-cafc23e0006f@linux.dev \
    --to=usama.arif@linux.dev \
    --cc=akpm@linux-foundation.org \
    --cc=apopple@nvidia.com \
    --cc=baohua@kernel.org \
    --cc=baolin.wang@linux.alibaba.com \
    --cc=david@kernel.org \
    --cc=dev.jain@arm.com \
    --cc=gourry@gourry.net \
    --cc=hannes@cmpxchg.org \
    --cc=kasong@tencent.com \
    --cc=lance.yang@linux.dev \
    --cc=liam@infradead.org \
    --cc=linux-kernel@vger.kernel.org \
    --cc=linux-mm@kvack.org \
    --cc=ljs@kernel.org \
    --cc=mhocko@suse.com \
    --cc=minchan@kernel.org \
    --cc=muchun.song@linux.dev \
    --cc=nico.pache@linux.dev \
    --cc=qi.zheng@linux.dev \
    --cc=rppt@kernel.org \
    --cc=ryan.roberts@arm.com \
    --cc=senozhatsky@chromium.org \
    --cc=shakeel.butt@linux.dev \
    --cc=surenb@google.com \
    --cc=vbabka@kernel.org \
    --cc=willy@infradead.org \
    --cc=ying.huang@linux.alibaba.com \
    --cc=ziy@nvidia.com \
    /path/to/YOUR_REPLY

  https://kernel.org/pub/software/scm/git/docs/git-send-email.html

* If your mail client supports setting the In-Reply-To header
  via mailto: links, try the mailto: link
Be sure your reply has a Subject: header at the top and a blank line before the message body.
This is a public inbox, see mirroring instructions
for how to clone and mirror all data and code used for this inbox