From: Michal Hocko <mhocko@suse.com>
To: Barry Song <baohua@kernel.org>
Cc: Ridong <ridong.chen@linux.dev>,
Andrew Morton <akpm@linux-foundation.org>,
Johannes Weiner <hannes@cmpxchg.org>,
David Hildenbrand <david@kernel.org>,
Qi Zheng <qi.zheng@linux.dev>,
Shakeel Butt <shakeel.butt@linux.dev>,
Lorenzo Stoakes <ljs@kernel.org>,
Kairui Song <kasong@tencent.com>,
Axel Rasmussen <axelrasmussen@google.com>,
Yuanchu Xie <yuanchu@google.com>, Wei Xu <weixugc@google.com>,
Zhongkun He <hezhongkun.hzk@bytedance.com>,
Muchun Song <muchun.song@linux.dev>,
Davidlohr Bueso <dave@stgolabs.net>,
Roman Gushchin <roman.gushchin@linux.dev>,
linux-mm@kvack.org, linux-kernel@vger.kernel.org,
Ridong Chen <chenridong@xiaomi.com>
Subject: Re: [PATCH -v4 1/4] mm/vmscan: fix anon-only reclaim evicting file pages when swappiness=max
Date: Tue, 28 Jul 2026 09:18:20 +0200 [thread overview]
Message-ID: <amhXvAdmDxX71uxd@tiehlicka> (raw)
In-Reply-To: <CAGsJ_4w1AFkKOSCZYwRsZYOBxCq_7CQA43jQOiDUyLrjVUmA-w@mail.gmail.com>
On Tue 28-07-26 06:35:20, Barry Song wrote:
> On Mon, Jul 27, 2026 at 7:37 PM Michal Hocko <mhocko@suse.com> wrote:
> >
> > On Fri 24-07-26 11:34:32, Ridong wrote:
> > > From: Ridong Chen <chenridong@xiaomi.com>
> > >
> > > As Qi mentioned [1], when swappiness=max (SWAPPINESS_ANON_ONLY) is set,
> > > the reclaim logic is expected to reclaim anonymous pages exclusively.
> > > However, due to the current ordering of checks in get_scan_count(),
> > > file pages may still be evicted if can_reclaim_anon_pages() returns
> > > false, which contradicts the semantics of SWAPPINESS_ANON_ONLY.
> > >
> > > Reproducer in a cgroup holding 64M of file cache, with no swap configured:
> > >
> > > Before (file cache is wrongly evicted):
> > > # cat memory.stat
> > > anon 196608
> > > file 67178496
> > > pgscan_proactive 0
> > > # echo "64M swappiness=max" > memory.reclaim
> > > # cat memory.stat
> > > anon 208896
> > > file 4096 <- page cache evicted
> > > pgsteal_proactive 16400
> > > pgscan_proactive 16400
> > >
> > > After (file cache is left intact):
> > > # cat memory.stat
> > > anon 200704
> > > file 67178496
> > > pgscan_proactive 0
> > > # echo "64M swappiness=max" > memory.reclaim
> > > -bash: echo: write error: Resource temporarily unavailable
> > > # cat memory.stat
> > > anon 208896
> > > file 67178496 <- page cache untouched
> > > pgsteal_proactive 0
> > > pgscan_proactive 0
> > >
> > > Fix this by bailing out early when SWAPPINESS_ANON_ONLY is set and no
> > > anonymous pages are reclaimable, before falling back to file reclaim.
> > >
> > > [1] https://lore.kernel.org/cgroups/7ddf3eee-5fe2-45f7-8614-c8936a039e04@linux.dev/
> > >
> > > Fixes: 68a1436bde00 ("mm: add swappiness=max arg to memory.reclaim for only anon reclaim")
> > > Suggested-by: Qi Zheng <qi.zheng@linux.dev>
> > > Acked-by: Shakeel Butt <shakeel.butt@linux.dev>
> > > Acked-by: Johannes Weiner <hannes@cmpxchg.org>
> > > Reviewed-by: Muchun Song <muchun.song@linux.dev>
> > > Reviewed-by: Qi Zheng <qi.zheng@linux.dev>
> > > Reviewed-by: Barry Song <baohua@kernel.org>
> > > Signed-off-by: Ridong Chen <chenridong@xiaomi.com>
> > > ---
> > > mm/vmscan.c | 24 +++++++++++++++++-------
> > > 1 file changed, 17 insertions(+), 7 deletions(-)
> > >
> > > diff --git a/mm/vmscan.c b/mm/vmscan.c
> > > index 35c3bb15ae96..2c689682b952 100644
> > > --- a/mm/vmscan.c
> > > +++ b/mm/vmscan.c
> > > @@ -2501,6 +2501,23 @@ static void get_scan_count(struct lruvec *lruvec, struct scan_control *sc,
> > > enum scan_balance scan_balance;
> > > enum lru_list lru;
> > >
> > > + /*
> > > + * Proactive reclaim initiated by userspace for anonymous memory only.
> > > + * SWAPPINESS_ANON_ONLY is set only on the proactive reclaim path, so
> > > + * warn if it shows up elsewhere. When anon cannot be reclaimed (e.g.
> > > + * no swap), bail out instead of falling back to evicting file pages,
> > > + * which would violate the anon-only semantics.
> > > + */
> > > + if (swappiness == SWAPPINESS_ANON_ONLY) {
> > > + WARN_ON_ONCE(!sc->proactive);
> > > + if (!can_reclaim_anon_pages(memcg, pgdat->node_id, sc)) {
> > > + memset(nr, 0, sizeof(*nr) * NR_LRU_LISTS);
> > > + return;
> > > + }
> > > + scan_balance = SCAN_ANON;
> > > + goto out;
> >
> > Rather than warning the code would be easier to understand
> > (SWAPPINESS_ANON_ONLY implying sc->proactive is very subtle assumption
> > that might change in the future) I would go with and explicit check.
> >
> > Looking at how the code is structured currently doesn't the following
> > express the intention slightly better?
> >
> > diff --git a/mm/vmscan.c b/mm/vmscan.c
> > index 35c3bb15ae96..4fd38e3f1b05 100644
> > --- a/mm/vmscan.c
> > +++ b/mm/vmscan.c
> > @@ -2503,6 +2503,10 @@ static void get_scan_count(struct lruvec *lruvec, struct scan_control *sc,
> >
> > /* If we have no swap space, do not bother scanning anon folios. */
> > if (!sc->may_swap || !can_reclaim_anon_pages(memcg, pgdat->node_id, sc)) {
> > + if (swappiness == SWAPPINESS_ANON_ONLY && sc->proactive) {
> > + memset(nr, 0, sizeof(*nr) * NR_LRU_LISTS);
> > + return;
> > + }
> > scan_balance = SCAN_FILE;
>
> The question is whether we should allow falling back to SCAN_FILE
> when swappiness == SWAPPINESS_ANON_ONLY and !sc->proactive, if 201 is
> ever allowed for non-proactive reclaim in the future. I guess not,
> since SWAPPINESS_ANON_ONLY literally means reclaiming anonymous
> memory only?
If we follow swappiness==0 then yes we should fallback if we are close
to system OOM. Now arguably SWAPPINESS_ANON_ONLY is a wierd global
policy and it is quite possible this will not be ever allowed. My
argument was not to prepare the existing code for that. I just meant to
structure the code in a way that it current assumptions are more
explicit.
--
Michal Hocko
SUSE Labs
next prev parent reply other threads:[~2026-07-28 7:18 UTC|newest]
Thread overview: 19+ messages / expand[flat|nested] mbox.gz Atom feed top
2026-07-24 3:34 [PATCH -v4 0/4] mm/vmscan: fix swappiness=max and clean up per-node proactive reclaim Ridong
2026-07-24 3:34 ` [PATCH -v4 1/4] mm/vmscan: fix anon-only reclaim evicting file pages when swappiness=max Ridong
2026-07-27 8:19 ` Michal Hocko
2026-07-27 10:56 ` Barry Song
2026-07-27 11:07 ` Michal Hocko
2026-07-28 1:02 ` Ridong Chen
2026-07-27 11:37 ` Michal Hocko
2026-07-27 22:35 ` Barry Song
2026-07-28 7:18 ` Michal Hocko [this message]
2026-07-28 10:03 ` Barry Song
2026-07-28 11:53 ` Michal Hocko
2026-07-28 12:18 ` Barry Song
2026-07-24 3:34 ` [PATCH -v4 2/4] mm: vmscan: propagate real error code from per-node proactive reclaim Ridong
2026-07-24 3:34 ` [PATCH -v4 3/4] mm: vmscan: drop unused gfp_mask parameter from __node_reclaim() Ridong
2026-07-24 3:34 ` [PATCH -v4 4/4] mm/mglru: fix anon-only reclaim evicting file pages when swappiness=max Ridong
2026-07-24 7:18 ` Kairui Song
2026-07-24 8:29 ` Barry Song
2026-07-27 6:37 ` Baolin Wang
2026-07-24 5:37 ` [PATCH -v4 0/4] mm/vmscan: fix swappiness=max and clean up per-node proactive reclaim Andrew Morton
Reply instructions:
You may reply publicly to this message via plain-text email
using any one of the following methods:
* Save the following mbox file, import it into your mail client,
and reply-to-all from there: mbox
Avoid top-posting and favor interleaved quoting:
https://en.wikipedia.org/wiki/Posting_style#Interleaved_style
* Reply using the --to, --cc, and --in-reply-to
switches of git-send-email(1):
git send-email \
--in-reply-to=amhXvAdmDxX71uxd@tiehlicka \
--to=mhocko@suse.com \
--cc=akpm@linux-foundation.org \
--cc=axelrasmussen@google.com \
--cc=baohua@kernel.org \
--cc=chenridong@xiaomi.com \
--cc=dave@stgolabs.net \
--cc=david@kernel.org \
--cc=hannes@cmpxchg.org \
--cc=hezhongkun.hzk@bytedance.com \
--cc=kasong@tencent.com \
--cc=linux-kernel@vger.kernel.org \
--cc=linux-mm@kvack.org \
--cc=ljs@kernel.org \
--cc=muchun.song@linux.dev \
--cc=qi.zheng@linux.dev \
--cc=ridong.chen@linux.dev \
--cc=roman.gushchin@linux.dev \
--cc=shakeel.butt@linux.dev \
--cc=weixugc@google.com \
--cc=yuanchu@google.com \
/path/to/YOUR_REPLY
https://kernel.org/pub/software/scm/git/docs/git-send-email.html
* If your mail client supports setting the In-Reply-To header
via mailto: links, try the mailto: link
Be sure your reply has a Subject: header at the top and a blank line
before the message body.
This is a public inbox, see mirroring instructions
for how to clone and mirror all data and code used for this inbox