bpf.vger.kernel.org archive mirror
 help / color / mirror / Atom feed
From: Yonghong Song <yonghong.song@linux.dev>
To: bot+bpf-ci@kernel.org, bpf@vger.kernel.org
Cc: ast@kernel.org, andrii@kernel.org, daniel@iogearbox.net,
	eddyz87@gmail.com, kernel-team@fb.com, martin.lau@kernel.org,
	mason@kernel.org, ihor.solodrai@linux.dev
Subject: Re: [PATCH bpf-next v2 09/12] bpf, arm64: Move kfunc arguments into the arm64 calling convention
Date: Thu, 10 Sep 2026 22:34:14 -0700	[thread overview]
Message-ID: <4b6edd53-acda-4852-963d-fe233fd2d192@linux.dev> (raw)
In-Reply-To: <41db2f5f6481cb0e9ca5a73d5a597578ebffa4bd39cbcf98aabbcaf87048809f@mail.kernel.org>



On 9/9/26 12:30 AM, bot+bpf-ci@kernel.org wrote:
>> commit 7c4f41d1b9c4dfdbada8f6d1b1097849ccb20de8
>> Author: Yonghong Song <yonghong.song@linux.dev>
>>
>>      bpf, arm64: Move kfunc arguments into the arm64 calling convention
>>
>>      Do proper move from bpf calling convention to arm64 calling convention
>>      to satisfy native requirement. AAPCS64 only ever moves an argument to a
>>      higher slot, so the moves need one scratch register to carry an eightbyte
>>      from one stack slot to another, and never the one a convention moving an
>>      argument down would need.
>>
>>      In addition, the arena argument walk counts eightbytes rather than
>>      parameters, as an argument may take two registers. The walk takes the
>>      func model from the caller now, as the moves need it too, and runs first
>>      so that they carry the rebased value.
>>
>>      Signed-off-by: Yonghong Song <yonghong.song@linux.dev>
> This looks like a bug fix for multi-slot kfunc argument support.
> Should this include:
>
>    Fixes: 2092c3247e45 ("bpf: Prepare kfunc arguments for the JIT from an ABI description")
>
> In v1 of this series, Alexei Starovoitov raised a design concern about
> the overall approach. He questioned why the BPF calling convention should
> diverge from x86-64/arm64, noting that both JIT patches doing argument
> shuffling is a sign that the approach may be wrong on the BPF side. He
> suggested adjusting the BPF calling convention to match arm64 (the
> stricter of the two architectures) rather than forcing all JITs to adapt:
>
>    "Why cannot we adjust bpf calling convention to match arm64/x86 the
>    best? Since arm64 is stricter, I'd pick that style. There are no kfuncs
>    that use int128 or 16+ byte args, so it's a matter of bpf subprogs
>    calling bpf subprogs. Seems cleaner to adjust what llvm emits instead of
>    forcing all jits to adapt."
>
> The v2 took the opposite approach - keeping the BPF calling convention
> and implementing JIT argument shuffling in both x86 and arm64. Was this
> design decision explicitly discussed and agreed upon?

Yes, we will keep original bpf calling convention and then jit will
do proper adjustment.

>
>> diff --git a/arch/arm64/net/bpf_jit_comp.c b/arch/arm64/net/bpf_jit_comp.c
>> index 3aa3ea0bc30b..bdac930dbdec 100644
>> --- a/arch/arm64/net/bpf_jit_comp.c
>> +++ b/arch/arm64/net/bpf_jit_comp.c
> [ ... ]
>
>> @@ -1262,19 +1268,20 @@ static void emit_stack_arg_store_imm(s32 imm, s16 bpf_off, const u8 tmp, struct
>>    * kern_vm_start. A nullable arg preserves NULL by skipping the add, tested
>>    * on the truncated value as arena NULL is offset 0.
>>    */
>> -static int emit_kfunc_arena_args(struct jit_ctx *ctx, const struct bpf_insn *insn)
>> +static int emit_kfunc_arena_args(struct jit_ctx *ctx, const struct btf_func_model *fm)
>>   {
>>   	const u8 arena_vm_base = bpf2a64[ARENA_VM_START];
>> -	const struct btf_func_model *fm;
>> -	int i;
>> -
>> -	fm = bpf_jit_find_kfunc_model(ctx->prog, insn);
>> -	if (!fm)
>> -		return -EINVAL;
>> +	int i, slot;
>>
>> -	for (i = 0; i < min_t(int, fm->nr_args, MAX_BPF_FUNC_REG_ARGS); i++) {
>> -		const u8 reg = bpf2a64[BPF_REG_1 + i];
>> +	for (i = 0, slot = 0; i < fm->nr_args; i++) {
>> +		u32 arg_regs = (fm->arg_size[i] + 7) / 8;
>>   		u8 flags = fm->arg_flags[i];
>> +		u8 reg;
>> +
>> +		if (slot + arg_regs > MAX_BPF_FUNC_REG_ARGS)
>> +			break;
>> +		reg = bpf2a64[BPF_REG_1 + slot];
>> +		slot += arg_regs;
>>
>>   		if (!(flags & BTF_FMODEL_ARENA_ARG))
>>   			continue;
> [ ... ]
>
>> @@ -1293,6 +1300,45 @@ static int emit_kfunc_arena_args(struct jit_ctx *ctx, const struct bpf_insn *ins
>>   	return 0;
>>   }
>>
>> +static bool a64_arg_on_stack(u8 slot)
>> +{
>> +	return slot >= arm64_arg_abi.nr_arg_regs;
>> +}
>> +
>> +static s32 a64_arg_stack_off(u8 slot)
>> +{
>> +	return (slot - arm64_arg_abi.nr_arg_regs) * sizeof(u64);
>> +}
>> +
>> +/*
>> + * AAPCS64 only ever moves an argument to a higher slot, so the planner asks
>> + * for the scratch only to carry an eightbyte from one stack slot to another.
>> + */
>> +static void emit_kfunc_arg_moves(struct jit_ctx *ctx, const struct btf_func_model *fm)
>> +{
>> +	struct bpf_jit_arg_move moves[BPF_JIT_MAX_ARG_MOVES];
>> +	const u8 tmp = bpf2a64[TMP_REG_1];
>> +	u32 i, n;
>> +
>> +	n = bpf_jit_plan_arg_moves(&arm64_arg_abi, fm, moves);
>> +
>> +	for (i = 0; i < n; i++) {
>> +		u8 dst = moves[i].dst, src = moves[i].src, reg;
>> +
>> +		if (a64_arg_on_stack(src)) {
>> +			reg = a64_arg_on_stack(dst) ? tmp : dst;
>> +			emit(A64_LDR64I(reg, A64_SP, a64_arg_stack_off(src)), ctx);
>> +		} else {
>> +			reg = src;
>> +		}
>> +
>> +		if (a64_arg_on_stack(dst))
>> +			emit(A64_STR64I(reg, A64_SP, a64_arg_stack_off(dst)), ctx);
>> +		else if (reg != dst)
>> +			emit(A64_MOV(1, dst, reg), ctx);
>> +	}
>> +}
> This isn't a bug, but would it be clearer to say the plan never needs a
> scratch on AAPCS64 and that TMP_REG_1 is only this helper's
> stack-to-stack carrier? As written, the comment reads as though moves[]
> can contain BPF_JIT_ARG_TMP (which the loop does not handle - unlike the
> x86 version at arch/x86/net/bpf_jit_comp.c which does), and the ': dst'
> arm in the conditional looks unreachable for the same reason (src on the
> stack implies dst on the stack, since dst >= src >= nr_arg_regs).

Okay, will fix.

>
> [ ... ]
>
>
> ---
> AI reviewed your patch. Please fix the bug or email reply why it's not a bug.
> See: https://github.com/kernel-patches/vmtest/blob/master/ci/claude/README.md
>
> CI run summary: https://github.com/kernel-patches/bpf/actions/runs/34320399441


  reply	other threads:[~2026-09-11  5:34 UTC|newest]

Thread overview: 31+ messages / expand[flat|nested]  mbox.gz  Atom feed  top
2026-09-09  6:25 [PATCH bpf-next v2 00/12] bpf: Support by-value struct and __int128 arguments Yonghong Song
2026-09-09  6:25 ` [PATCH bpf-next v2 01/12] selftests/bpf: Add a test for an __int128 by-value argument Yonghong Song
2026-09-09  7:13   ` bot+bpf-ci
2026-09-11  4:25     ` Yonghong Song
2026-09-09  6:25 ` [PATCH bpf-next v2 02/12] bpf: Index global function arguments by argument slot Yonghong Song
2026-09-09  7:13   ` bot+bpf-ci
2026-09-11  4:27     ` Yonghong Song
2026-09-09  6:25 ` [PATCH bpf-next v2 03/12] bpf: Support by-value struct arguments up to 16 bytes Yonghong Song
2026-09-09  7:13   ` bot+bpf-ci
2026-09-11  4:29     ` Yonghong Song
2026-09-09  6:25 ` [PATCH bpf-next v2 04/12] bpf: Support __int128 as a by-value function argument Yonghong Song
2026-09-09  6:25 ` [PATCH bpf-next v2 05/12] bpf: Rename bpf_call_summary::num_params to arg_slot_cnt Yonghong Song
2026-09-09  6:25 ` [PATCH bpf-next v2 06/12] bpf: Recognize by-value struct and __int128 kfunc arguments Yonghong Song
2026-09-09  6:46   ` sashiko-bot
2026-09-11  4:31     ` Yonghong Song
2026-09-09  6:25 ` [PATCH bpf-next v2 07/12] bpf: Prepare kfunc arguments for the JIT from an ABI description Yonghong Song
2026-09-09  6:46   ` sashiko-bot
2026-09-11  5:05     ` Yonghong Song
2026-09-09  6:26 ` [PATCH bpf-next v2 08/12] bpf, x86: Move kfunc arguments into the x86-64 calling convention Yonghong Song
2026-09-09  7:29   ` bot+bpf-ci
2026-09-11  5:32     ` Yonghong Song
2026-09-09  6:26 ` [PATCH bpf-next v2 09/12] bpf, arm64: Move kfunc arguments into the arm64 " Yonghong Song
2026-09-09  7:30   ` bot+bpf-ci
2026-09-11  5:34     ` Yonghong Song [this message]
2026-09-09  6:26 ` [PATCH bpf-next v2 10/12] selftests/bpf: Add C tests for by-value arguments up to 16 bytes Yonghong Song
2026-09-09  6:26 ` [PATCH bpf-next v2 11/12] selftests/bpf: Add inline-asm tests for by-value arguments Yonghong Song
2026-09-09  7:30   ` bot+bpf-ci
2026-09-11  5:37     ` Yonghong Song
2026-09-09  6:26 ` [PATCH bpf-next v2 12/12] selftests/bpf: Add tests for by-value kfunc arguments Yonghong Song
2026-09-09  7:30   ` bot+bpf-ci
2026-09-11  5:57     ` Yonghong Song

Reply instructions:

You may reply publicly to this message via plain-text email
using any one of the following methods:

* Save the following mbox file, import it into your mail client,
  and reply-to-all from there: mbox

  Avoid top-posting and favor interleaved quoting:
  https://en.wikipedia.org/wiki/Posting_style#Interleaved_style

* Reply using the --to, --cc, and --in-reply-to
  switches of git-send-email(1):

  git send-email \
    --in-reply-to=4b6edd53-acda-4852-963d-fe233fd2d192@linux.dev \
    --to=yonghong.song@linux.dev \
    --cc=andrii@kernel.org \
    --cc=ast@kernel.org \
    --cc=bot+bpf-ci@kernel.org \
    --cc=bpf@vger.kernel.org \
    --cc=daniel@iogearbox.net \
    --cc=eddyz87@gmail.com \
    --cc=ihor.solodrai@linux.dev \
    --cc=kernel-team@fb.com \
    --cc=martin.lau@kernel.org \
    --cc=mason@kernel.org \
    /path/to/YOUR_REPLY

  https://kernel.org/pub/software/scm/git/docs/git-send-email.html

* If your mail client supports setting the In-Reply-To header
  via mailto: links, try the mailto: link
Be sure your reply has a Subject: header at the top and a blank line before the message body.
This is a public inbox, see mirroring instructions
for how to clone and mirror all data and code used for this inbox;
as well as URLs for NNTP newsgroup(s).