Linux-mm Archive on lore.kernel.org
 help / color / mirror / Atom feed
From: Michal Hocko <mhocko@suse.com>
To: Barry Song <baohua@kernel.org>
Cc: Ridong <ridong.chen@linux.dev>,
	Andrew Morton <akpm@linux-foundation.org>,
	Johannes Weiner <hannes@cmpxchg.org>,
	David Hildenbrand <david@kernel.org>,
	Qi Zheng <qi.zheng@linux.dev>,
	Shakeel Butt <shakeel.butt@linux.dev>,
	Lorenzo Stoakes <ljs@kernel.org>,
	Kairui Song <kasong@tencent.com>,
	Axel Rasmussen <axelrasmussen@google.com>,
	Yuanchu Xie <yuanchu@google.com>, Wei Xu <weixugc@google.com>,
	Zhongkun He <hezhongkun.hzk@bytedance.com>,
	Muchun Song <muchun.song@linux.dev>,
	Davidlohr Bueso <dave@stgolabs.net>,
	Roman Gushchin <roman.gushchin@linux.dev>,
	linux-mm@kvack.org, linux-kernel@vger.kernel.org,
	Ridong Chen <chenridong@xiaomi.com>
Subject: Re: [PATCH -v4 1/4] mm/vmscan: fix anon-only reclaim evicting file pages when swappiness=max
Date: Tue, 28 Jul 2026 09:18:20 +0200	[thread overview]
Message-ID: <amhXvAdmDxX71uxd@tiehlicka> (raw)
In-Reply-To: <CAGsJ_4w1AFkKOSCZYwRsZYOBxCq_7CQA43jQOiDUyLrjVUmA-w@mail.gmail.com>

On Tue 28-07-26 06:35:20, Barry Song wrote:
> On Mon, Jul 27, 2026 at 7:37 PM Michal Hocko <mhocko@suse.com> wrote:
> >
> > On Fri 24-07-26 11:34:32, Ridong wrote:
> > > From: Ridong Chen <chenridong@xiaomi.com>
> > >
> > > As Qi mentioned [1], when swappiness=max (SWAPPINESS_ANON_ONLY) is set,
> > > the reclaim logic is expected to reclaim anonymous pages exclusively.
> > > However, due to the current ordering of checks in get_scan_count(),
> > > file pages may still be evicted if can_reclaim_anon_pages() returns
> > > false, which contradicts the semantics of SWAPPINESS_ANON_ONLY.
> > >
> > > Reproducer in a cgroup holding 64M of file cache, with no swap configured:
> > >
> > >   Before (file cache is wrongly evicted):
> > >     # cat memory.stat
> > >     anon 196608
> > >     file 67178496
> > >     pgscan_proactive 0
> > >     # echo "64M swappiness=max" > memory.reclaim
> > >     # cat memory.stat
> > >     anon 208896
> > >     file 4096                 <- page cache evicted
> > >     pgsteal_proactive 16400
> > >     pgscan_proactive 16400
> > >
> > >   After (file cache is left intact):
> > >     # cat memory.stat
> > >     anon 200704
> > >     file 67178496
> > >     pgscan_proactive 0
> > >     # echo "64M swappiness=max" > memory.reclaim
> > >     -bash: echo: write error: Resource temporarily unavailable
> > >     # cat memory.stat
> > >     anon 208896
> > >     file 67178496             <- page cache untouched
> > >     pgsteal_proactive 0
> > >     pgscan_proactive 0
> > >
> > > Fix this by bailing out early when SWAPPINESS_ANON_ONLY is set and no
> > > anonymous pages are reclaimable, before falling back to file reclaim.
> > >
> > > [1] https://lore.kernel.org/cgroups/7ddf3eee-5fe2-45f7-8614-c8936a039e04@linux.dev/
> > >
> > > Fixes: 68a1436bde00 ("mm: add swappiness=max arg to memory.reclaim for only anon reclaim")
> > > Suggested-by: Qi Zheng <qi.zheng@linux.dev>
> > > Acked-by: Shakeel Butt <shakeel.butt@linux.dev>
> > > Acked-by: Johannes Weiner <hannes@cmpxchg.org>
> > > Reviewed-by: Muchun Song <muchun.song@linux.dev>
> > > Reviewed-by: Qi Zheng <qi.zheng@linux.dev>
> > > Reviewed-by: Barry Song <baohua@kernel.org>
> > > Signed-off-by: Ridong Chen <chenridong@xiaomi.com>
> > > ---
> > >  mm/vmscan.c | 24 +++++++++++++++++-------
> > >  1 file changed, 17 insertions(+), 7 deletions(-)
> > >
> > > diff --git a/mm/vmscan.c b/mm/vmscan.c
> > > index 35c3bb15ae96..2c689682b952 100644
> > > --- a/mm/vmscan.c
> > > +++ b/mm/vmscan.c
> > > @@ -2501,6 +2501,23 @@ static void get_scan_count(struct lruvec *lruvec, struct scan_control *sc,
> > >       enum scan_balance scan_balance;
> > >       enum lru_list lru;
> > >
> > > +     /*
> > > +      * Proactive reclaim initiated by userspace for anonymous memory only.
> > > +      * SWAPPINESS_ANON_ONLY is set only on the proactive reclaim path, so
> > > +      * warn if it shows up elsewhere. When anon cannot be reclaimed (e.g.
> > > +      * no swap), bail out instead of falling back to evicting file pages,
> > > +      * which would violate the anon-only semantics.
> > > +      */
> > > +     if (swappiness == SWAPPINESS_ANON_ONLY) {
> > > +             WARN_ON_ONCE(!sc->proactive);
> > > +             if (!can_reclaim_anon_pages(memcg, pgdat->node_id, sc)) {
> > > +                     memset(nr, 0, sizeof(*nr) * NR_LRU_LISTS);
> > > +                     return;
> > > +             }
> > > +             scan_balance = SCAN_ANON;
> > > +             goto out;
> >
> > Rather than warning the code would be easier to understand
> > (SWAPPINESS_ANON_ONLY implying sc->proactive is very subtle assumption
> > that might change in the future) I would go with and explicit check.
> >
> > Looking at how the code is structured currently doesn't the following
> > express the intention slightly better?
> >
> > diff --git a/mm/vmscan.c b/mm/vmscan.c
> > index 35c3bb15ae96..4fd38e3f1b05 100644
> > --- a/mm/vmscan.c
> > +++ b/mm/vmscan.c
> > @@ -2503,6 +2503,10 @@ static void get_scan_count(struct lruvec *lruvec, struct scan_control *sc,
> >
> >         /* If we have no swap space, do not bother scanning anon folios. */
> >         if (!sc->may_swap || !can_reclaim_anon_pages(memcg, pgdat->node_id, sc)) {
> > +               if (swappiness == SWAPPINESS_ANON_ONLY && sc->proactive) {
> > +                       memset(nr, 0, sizeof(*nr) * NR_LRU_LISTS);
> > +                       return;
> > +               }
> >                 scan_balance = SCAN_FILE;
> 
> The question is whether we should allow falling back to SCAN_FILE
> when swappiness == SWAPPINESS_ANON_ONLY and !sc->proactive, if 201 is
> ever allowed for non-proactive reclaim in the future. I guess not,
> since SWAPPINESS_ANON_ONLY literally means reclaiming anonymous
> memory only?

If we follow swappiness==0 then yes we should fallback if we are close
to system OOM. Now arguably SWAPPINESS_ANON_ONLY is a wierd global
policy and it is quite possible this will not be ever allowed. My
argument was not to prepare the existing code for that. I just meant to
structure the code in a way that it current assumptions are more
explicit.

-- 
Michal Hocko
SUSE Labs


  reply	other threads:[~2026-07-28  7:18 UTC|newest]

Thread overview: 19+ messages / expand[flat|nested]  mbox.gz  Atom feed  top
2026-07-24  3:34 [PATCH -v4 0/4] mm/vmscan: fix swappiness=max and clean up per-node proactive reclaim Ridong
2026-07-24  3:34 ` [PATCH -v4 1/4] mm/vmscan: fix anon-only reclaim evicting file pages when swappiness=max Ridong
2026-07-27  8:19   ` Michal Hocko
2026-07-27 10:56     ` Barry Song
2026-07-27 11:07       ` Michal Hocko
2026-07-28  1:02       ` Ridong Chen
2026-07-27 11:37   ` Michal Hocko
2026-07-27 22:35     ` Barry Song
2026-07-28  7:18       ` Michal Hocko [this message]
2026-07-28 10:03         ` Barry Song
2026-07-28 11:53           ` Michal Hocko
2026-07-28 12:18             ` Barry Song
2026-07-24  3:34 ` [PATCH -v4 2/4] mm: vmscan: propagate real error code from per-node proactive reclaim Ridong
2026-07-24  3:34 ` [PATCH -v4 3/4] mm: vmscan: drop unused gfp_mask parameter from __node_reclaim() Ridong
2026-07-24  3:34 ` [PATCH -v4 4/4] mm/mglru: fix anon-only reclaim evicting file pages when swappiness=max Ridong
2026-07-24  7:18   ` Kairui Song
2026-07-24  8:29   ` Barry Song
2026-07-27  6:37   ` Baolin Wang
2026-07-24  5:37 ` [PATCH -v4 0/4] mm/vmscan: fix swappiness=max and clean up per-node proactive reclaim Andrew Morton

Reply instructions:

You may reply publicly to this message via plain-text email
using any one of the following methods:

* Save the following mbox file, import it into your mail client,
  and reply-to-all from there: mbox

  Avoid top-posting and favor interleaved quoting:
  https://en.wikipedia.org/wiki/Posting_style#Interleaved_style

* Reply using the --to, --cc, and --in-reply-to
  switches of git-send-email(1):

  git send-email \
    --in-reply-to=amhXvAdmDxX71uxd@tiehlicka \
    --to=mhocko@suse.com \
    --cc=akpm@linux-foundation.org \
    --cc=axelrasmussen@google.com \
    --cc=baohua@kernel.org \
    --cc=chenridong@xiaomi.com \
    --cc=dave@stgolabs.net \
    --cc=david@kernel.org \
    --cc=hannes@cmpxchg.org \
    --cc=hezhongkun.hzk@bytedance.com \
    --cc=kasong@tencent.com \
    --cc=linux-kernel@vger.kernel.org \
    --cc=linux-mm@kvack.org \
    --cc=ljs@kernel.org \
    --cc=muchun.song@linux.dev \
    --cc=qi.zheng@linux.dev \
    --cc=ridong.chen@linux.dev \
    --cc=roman.gushchin@linux.dev \
    --cc=shakeel.butt@linux.dev \
    --cc=weixugc@google.com \
    --cc=yuanchu@google.com \
    /path/to/YOUR_REPLY

  https://kernel.org/pub/software/scm/git/docs/git-send-email.html

* If your mail client supports setting the In-Reply-To header
  via mailto: links, try the mailto: link
Be sure your reply has a Subject: header at the top and a blank line before the message body.
This is a public inbox, see mirroring instructions
for how to clone and mirror all data and code used for this inbox