From: Peter Zijlstra <peterz@infradead.org>
To: Andrea Righi <arighi@nvidia.com>
Cc: Tejun Heo <tj@kernel.org>, David Vernet <void@manifault.com>,
Changwoo Min <changwoo@igalia.com>,
John Stultz <jstultz@google.com>, Ingo Molnar <mingo@redhat.com>,
Juri Lelli <juri.lelli@redhat.com>,
Vincent Guittot <vincent.guittot@linaro.org>,
Dietmar Eggemann <dietmar.eggemann@arm.com>,
Steven Rostedt <rostedt@goodmis.org>,
Ben Segall <bsegall@google.com>, Mel Gorman <mgorman@suse.de>,
Valentin Schneider <vschneid@redhat.com>,
K Prateek Nayak <kprateek.nayak@amd.com>,
Christian Loehle <christian.loehle@arm.com>,
David Dai <david.dai@linux.dev>, Emil Tsalapatis <etsal@meta.com>,
Lee Trager <ltrager@nvidia.com>,
Richard Cheng <icheng@nvidia.com>, Koba Ko <kobak@nvidia.com>,
Aiqun Yu <aiqun.yu@oss.qualcomm.com>,
sched-ext@lists.linux.dev, linux-kernel@vger.kernel.org
Subject: Re: [PATCH 03/18] sched: Make NOHZ CFS bandwidth checks follow proxy donor
Date: Thu, 10 Sep 2026 11:54:48 +0200 [thread overview]
Message-ID: <20260910095448.GE4120091@noisy.programming.kicks-ass.net> (raw)
In-Reply-To: <20260831134338.1531664-4-arighi@nvidia.com>
On Mon, Aug 31, 2026 at 03:42:13PM +0200, Andrea Righi wrote:
> Proxy execution separates the scheduling context in rq->donor from the
> physical execution context in rq->curr. sched_can_stop_tick() checks the
> latter for CFS bandwidth constraints and only does so when nr_running is
> one.
>
> A retained proxy donor keeps both the donor and mutex owner queued. The
> check therefore misses a constrained FAIR donor and may stop the tick
> while its runtime still needs to be enforced.
>
> Check the selected donor instead and remove the nr_running restriction.
> The donor being a queued FAIR task is sufficient to require bandwidth
> accounting regardless of other runnable tasks.
>
> Fixes: af0c8b2bf67b ("sched: Split scheduler and execution contexts")
> Reported-by: Sashiko <sashiko-bot@kernel.org>
> Link: https://lore.kernel.org/r/20260713164807.E5ED21F00A3A@smtp.kernel.org
> Acked-by: John Stultz <jstultz@google.com>
> Signed-off-by: Andrea Righi <arighi@nvidia.com>
> ---
> kernel/sched/core.c | 35 ++++++++++++++++++-----------------
> kernel/sched/fair.c | 12 +++++++-----
> 2 files changed, 25 insertions(+), 22 deletions(-)
>
> diff --git a/kernel/sched/core.c b/kernel/sched/core.c
> index 237d216382f46..14d0d5c884393 100644
> --- a/kernel/sched/core.c
> +++ b/kernel/sched/core.c
> @@ -1419,11 +1419,8 @@ static void nohz_csd_func(void *info)
> #endif /* CONFIG_NO_HZ_COMMON */
>
> #ifdef CONFIG_NO_HZ_FULL
> -static inline bool __need_bw_check(struct rq *rq, struct task_struct *p)
> +static inline bool __need_bw_check(struct task_struct *p)
> {
> - if (rq->nr_running != 1)
> - return false;
> -
> if (p->sched_class != &fair_sched_class)
> return false;
>
> @@ -1441,6 +1438,14 @@ bool sched_can_stop_tick(struct rq *rq)
> if (rq->dl.dl_nr_running)
> return false;
>
> + /*
> + * The selected scheduling context can be a constrained FAIR donor even
> + * when rq->curr is an RT task. Check it before the RT fast paths below,
> + * which may report that the tick can stop for a throttled RT context.
> + */
I am most confused... how can we ever have rq->donor be FAIR and
rq->curr be RT? That makes no sense. If an RT task is runnable, pick
should just straight up pick that.
> + if (__need_bw_check(rq->donor) && cfs_task_bw_constrained(rq->donor))
> + return false;
> +
> @@ -7114,7 +7107,7 @@ find_proxy_task(struct rq *rq, struct task_struct *donor, struct rq_flags *rf)
> */
> static void __sched notrace __schedule(int sched_mode)
> {
> - struct task_struct *prev, *next;
> + struct task_struct *prev, *next, *tick_donor;
> /*
> * On PREEMPT_RT kernel, SM_RTLOCK_WAIT is noted
> * as a preemption by schedule_debug() and RCU.
> @@ -7168,6 +7161,7 @@ static void __sched notrace __schedule(int sched_mode)
> rq->clock_update_flags <<= 1;
> update_rq_clock(rq);
> rq->clock_update_flags = RQCF_UPDATED;
> + tick_donor = rq->donor;
>
> switch_count = &prev->nivcsw;
>
> @@ -7243,6 +7237,13 @@ static void __sched notrace __schedule(int sched_mode)
> clear_tsk_need_resched(prev);
> clear_preempt_need_resched();
> keep_resched:
> + /*
> + * Enqueue and dequeue updates can evaluate the outgoing donor. Refresh
> + * the dependency after selecting a different scheduling context.
> + */
> + if (rq->donor != tick_donor)
> + sched_update_tick_dependency(rq);
Please keep all the proxy specific bits inside the one
sched_proxy_exec() branch above.
next prev parent reply other threads:[~2026-09-10 9:55 UTC|newest]
Thread overview: 42+ messages / expand[flat|nested] mbox.gz Atom feed top
2026-08-31 13:42 [PATCHSET v13 sched_ext/for-7.4] sched: Make proxy execution compatible with sched_ext Andrea Righi
2026-08-31 13:42 ` [PATCH 01/18] sched/core: Drop mutex locks before proxy rescheduling Andrea Righi
2026-08-31 13:42 ` [PATCH 02/18] sched/core: Dequeue waking proxy donors before reset Andrea Righi
2026-09-01 5:24 ` K Prateek Nayak
2026-09-08 9:28 ` Andrea Righi
2026-08-31 13:42 ` [PATCH 03/18] sched: Make NOHZ CFS bandwidth checks follow proxy donor Andrea Righi
2026-09-10 9:54 ` Peter Zijlstra [this message]
2026-08-31 13:42 ` [PATCH 04/18] sched/core: Avoid false migration warning for proxy donors Andrea Righi
2026-09-10 10:06 ` Peter Zijlstra
2026-08-31 13:42 ` [PATCH 05/18] sched: Pass next class to sched_change_begin() Andrea Righi
2026-09-10 10:12 ` Peter Zijlstra
2026-08-31 13:42 ` [PATCH 06/18] sched: Add helper to block retained proxy donors Andrea Righi
2026-08-31 13:42 ` [PATCH 07/18] sched: Add sched_ext hooks for proxy execution Andrea Righi
2026-09-10 10:38 ` Peter Zijlstra
2026-08-31 13:42 ` [PATCH 08/18] sched: Introduce WF_ON_RQ wake flag Andrea Righi
2026-09-10 10:45 ` Peter Zijlstra
2026-08-31 13:42 ` [PATCH 09/18] sched_ext: Block proxy donors across scheduler transitions Andrea Righi
2026-09-10 10:53 ` Peter Zijlstra
2026-09-10 11:41 ` Peter Zijlstra
2026-08-31 13:42 ` [PATCH 10/18] sched_ext: Fix ops.running/stopping() pairing for proxy-exec donors Andrea Righi
2026-08-31 13:42 ` [PATCH 11/18] sched_ext: Move reject DSQ draining into core Andrea Righi
2026-08-31 13:42 ` [PATCH 12/18] sched_ext: Generalize the reject DSQ reenqueue path Andrea Righi
2026-09-03 22:39 ` Tejun Heo
2026-09-08 9:34 ` Andrea Righi
2026-08-31 13:42 ` [PATCH 13/18] sched_ext: Handle proxy-exec races in remote DSQ transfers Andrea Righi
2026-08-31 13:42 ` [PATCH 14/18] sched_ext: Split curr|donor references properly Andrea Righi
2026-08-31 17:49 ` sashiko-bot
2026-09-08 10:15 ` Andrea Righi
2026-09-10 11:47 ` Peter Zijlstra
2026-08-31 13:42 ` [PATCH 15/18] sched_ext: Delegate proxy donor admission to BPF schedulers Andrea Righi
2026-08-31 18:08 ` sashiko-bot
2026-09-08 10:08 ` Andrea Righi
2026-09-10 13:39 ` Peter Zijlstra
2026-09-10 13:41 ` Peter Zijlstra
2026-08-31 13:42 ` [PATCH 16/18] sched_ext: Add selftest for blocked donor admission Andrea Righi
2026-08-31 13:42 ` [PATCH 17/18] sched_ext: scx_qmap: Add proxy execution support Andrea Righi
2026-08-31 18:33 ` sashiko-bot
2026-09-01 7:52 ` Richard Cheng
2026-09-08 9:42 ` Andrea Righi
2026-08-31 13:42 ` [PATCH 18/18] sched: Allow enabling proxy exec with sched_ext Andrea Righi
2026-09-03 22:51 ` [PATCHSET v13 sched_ext/for-7.4] sched: Make proxy execution compatible " Tejun Heo
2026-09-08 8:02 ` Peter Zijlstra
Reply instructions:
You may reply publicly to this message via plain-text email
using any one of the following methods:
* Save the following mbox file, import it into your mail client,
and reply-to-all from there: mbox
Avoid top-posting and favor interleaved quoting:
https://en.wikipedia.org/wiki/Posting_style#Interleaved_style
* Reply using the --to, --cc, and --in-reply-to
switches of git-send-email(1):
git send-email \
--in-reply-to=20260910095448.GE4120091@noisy.programming.kicks-ass.net \
--to=peterz@infradead.org \
--cc=aiqun.yu@oss.qualcomm.com \
--cc=arighi@nvidia.com \
--cc=bsegall@google.com \
--cc=changwoo@igalia.com \
--cc=christian.loehle@arm.com \
--cc=david.dai@linux.dev \
--cc=dietmar.eggemann@arm.com \
--cc=etsal@meta.com \
--cc=icheng@nvidia.com \
--cc=jstultz@google.com \
--cc=juri.lelli@redhat.com \
--cc=kobak@nvidia.com \
--cc=kprateek.nayak@amd.com \
--cc=linux-kernel@vger.kernel.org \
--cc=ltrager@nvidia.com \
--cc=mgorman@suse.de \
--cc=mingo@redhat.com \
--cc=rostedt@goodmis.org \
--cc=sched-ext@lists.linux.dev \
--cc=tj@kernel.org \
--cc=vincent.guittot@linaro.org \
--cc=void@manifault.com \
--cc=vschneid@redhat.com \
/path/to/YOUR_REPLY
https://kernel.org/pub/software/scm/git/docs/git-send-email.html
* If your mail client supports setting the In-Reply-To header
via mailto: links, try the mailto: link
Be sure your reply has a Subject: header at the top and a blank line
before the message body.
This is an external index of several public inboxes,
see mirroring instructions on how to clone and mirror
all data and code used by this external index.