From: Yonghong Song <yonghong.song@linux.dev>
To: bot+bpf-ci@kernel.org, bpf@vger.kernel.org
Cc: ast@kernel.org, andrii@kernel.org, daniel@iogearbox.net,
eddyz87@gmail.com, kernel-team@fb.com, martin.lau@kernel.org,
mason@kernel.org, ihor.solodrai@linux.dev
Subject: Re: [PATCH bpf-next v2 09/12] bpf, arm64: Move kfunc arguments into the arm64 calling convention
Date: Thu, 10 Sep 2026 22:34:14 -0700 [thread overview]
Message-ID: <4b6edd53-acda-4852-963d-fe233fd2d192@linux.dev> (raw)
In-Reply-To: <41db2f5f6481cb0e9ca5a73d5a597578ebffa4bd39cbcf98aabbcaf87048809f@mail.kernel.org>
On 9/9/26 12:30 AM, bot+bpf-ci@kernel.org wrote:
>> commit 7c4f41d1b9c4dfdbada8f6d1b1097849ccb20de8
>> Author: Yonghong Song <yonghong.song@linux.dev>
>>
>> bpf, arm64: Move kfunc arguments into the arm64 calling convention
>>
>> Do proper move from bpf calling convention to arm64 calling convention
>> to satisfy native requirement. AAPCS64 only ever moves an argument to a
>> higher slot, so the moves need one scratch register to carry an eightbyte
>> from one stack slot to another, and never the one a convention moving an
>> argument down would need.
>>
>> In addition, the arena argument walk counts eightbytes rather than
>> parameters, as an argument may take two registers. The walk takes the
>> func model from the caller now, as the moves need it too, and runs first
>> so that they carry the rebased value.
>>
>> Signed-off-by: Yonghong Song <yonghong.song@linux.dev>
> This looks like a bug fix for multi-slot kfunc argument support.
> Should this include:
>
> Fixes: 2092c3247e45 ("bpf: Prepare kfunc arguments for the JIT from an ABI description")
>
> In v1 of this series, Alexei Starovoitov raised a design concern about
> the overall approach. He questioned why the BPF calling convention should
> diverge from x86-64/arm64, noting that both JIT patches doing argument
> shuffling is a sign that the approach may be wrong on the BPF side. He
> suggested adjusting the BPF calling convention to match arm64 (the
> stricter of the two architectures) rather than forcing all JITs to adapt:
>
> "Why cannot we adjust bpf calling convention to match arm64/x86 the
> best? Since arm64 is stricter, I'd pick that style. There are no kfuncs
> that use int128 or 16+ byte args, so it's a matter of bpf subprogs
> calling bpf subprogs. Seems cleaner to adjust what llvm emits instead of
> forcing all jits to adapt."
>
> The v2 took the opposite approach - keeping the BPF calling convention
> and implementing JIT argument shuffling in both x86 and arm64. Was this
> design decision explicitly discussed and agreed upon?
Yes, we will keep original bpf calling convention and then jit will
do proper adjustment.
>
>> diff --git a/arch/arm64/net/bpf_jit_comp.c b/arch/arm64/net/bpf_jit_comp.c
>> index 3aa3ea0bc30b..bdac930dbdec 100644
>> --- a/arch/arm64/net/bpf_jit_comp.c
>> +++ b/arch/arm64/net/bpf_jit_comp.c
> [ ... ]
>
>> @@ -1262,19 +1268,20 @@ static void emit_stack_arg_store_imm(s32 imm, s16 bpf_off, const u8 tmp, struct
>> * kern_vm_start. A nullable arg preserves NULL by skipping the add, tested
>> * on the truncated value as arena NULL is offset 0.
>> */
>> -static int emit_kfunc_arena_args(struct jit_ctx *ctx, const struct bpf_insn *insn)
>> +static int emit_kfunc_arena_args(struct jit_ctx *ctx, const struct btf_func_model *fm)
>> {
>> const u8 arena_vm_base = bpf2a64[ARENA_VM_START];
>> - const struct btf_func_model *fm;
>> - int i;
>> -
>> - fm = bpf_jit_find_kfunc_model(ctx->prog, insn);
>> - if (!fm)
>> - return -EINVAL;
>> + int i, slot;
>>
>> - for (i = 0; i < min_t(int, fm->nr_args, MAX_BPF_FUNC_REG_ARGS); i++) {
>> - const u8 reg = bpf2a64[BPF_REG_1 + i];
>> + for (i = 0, slot = 0; i < fm->nr_args; i++) {
>> + u32 arg_regs = (fm->arg_size[i] + 7) / 8;
>> u8 flags = fm->arg_flags[i];
>> + u8 reg;
>> +
>> + if (slot + arg_regs > MAX_BPF_FUNC_REG_ARGS)
>> + break;
>> + reg = bpf2a64[BPF_REG_1 + slot];
>> + slot += arg_regs;
>>
>> if (!(flags & BTF_FMODEL_ARENA_ARG))
>> continue;
> [ ... ]
>
>> @@ -1293,6 +1300,45 @@ static int emit_kfunc_arena_args(struct jit_ctx *ctx, const struct bpf_insn *ins
>> return 0;
>> }
>>
>> +static bool a64_arg_on_stack(u8 slot)
>> +{
>> + return slot >= arm64_arg_abi.nr_arg_regs;
>> +}
>> +
>> +static s32 a64_arg_stack_off(u8 slot)
>> +{
>> + return (slot - arm64_arg_abi.nr_arg_regs) * sizeof(u64);
>> +}
>> +
>> +/*
>> + * AAPCS64 only ever moves an argument to a higher slot, so the planner asks
>> + * for the scratch only to carry an eightbyte from one stack slot to another.
>> + */
>> +static void emit_kfunc_arg_moves(struct jit_ctx *ctx, const struct btf_func_model *fm)
>> +{
>> + struct bpf_jit_arg_move moves[BPF_JIT_MAX_ARG_MOVES];
>> + const u8 tmp = bpf2a64[TMP_REG_1];
>> + u32 i, n;
>> +
>> + n = bpf_jit_plan_arg_moves(&arm64_arg_abi, fm, moves);
>> +
>> + for (i = 0; i < n; i++) {
>> + u8 dst = moves[i].dst, src = moves[i].src, reg;
>> +
>> + if (a64_arg_on_stack(src)) {
>> + reg = a64_arg_on_stack(dst) ? tmp : dst;
>> + emit(A64_LDR64I(reg, A64_SP, a64_arg_stack_off(src)), ctx);
>> + } else {
>> + reg = src;
>> + }
>> +
>> + if (a64_arg_on_stack(dst))
>> + emit(A64_STR64I(reg, A64_SP, a64_arg_stack_off(dst)), ctx);
>> + else if (reg != dst)
>> + emit(A64_MOV(1, dst, reg), ctx);
>> + }
>> +}
> This isn't a bug, but would it be clearer to say the plan never needs a
> scratch on AAPCS64 and that TMP_REG_1 is only this helper's
> stack-to-stack carrier? As written, the comment reads as though moves[]
> can contain BPF_JIT_ARG_TMP (which the loop does not handle - unlike the
> x86 version at arch/x86/net/bpf_jit_comp.c which does), and the ': dst'
> arm in the conditional looks unreachable for the same reason (src on the
> stack implies dst on the stack, since dst >= src >= nr_arg_regs).
Okay, will fix.
>
> [ ... ]
>
>
> ---
> AI reviewed your patch. Please fix the bug or email reply why it's not a bug.
> See: https://github.com/kernel-patches/vmtest/blob/master/ci/claude/README.md
>
> CI run summary: https://github.com/kernel-patches/bpf/actions/runs/34320399441
next prev parent reply other threads:[~2026-09-11 5:34 UTC|newest]
Thread overview: 31+ messages / expand[flat|nested] mbox.gz Atom feed top
2026-09-09 6:25 [PATCH bpf-next v2 00/12] bpf: Support by-value struct and __int128 arguments Yonghong Song
2026-09-09 6:25 ` [PATCH bpf-next v2 01/12] selftests/bpf: Add a test for an __int128 by-value argument Yonghong Song
2026-09-09 7:13 ` bot+bpf-ci
2026-09-11 4:25 ` Yonghong Song
2026-09-09 6:25 ` [PATCH bpf-next v2 02/12] bpf: Index global function arguments by argument slot Yonghong Song
2026-09-09 7:13 ` bot+bpf-ci
2026-09-11 4:27 ` Yonghong Song
2026-09-09 6:25 ` [PATCH bpf-next v2 03/12] bpf: Support by-value struct arguments up to 16 bytes Yonghong Song
2026-09-09 7:13 ` bot+bpf-ci
2026-09-11 4:29 ` Yonghong Song
2026-09-09 6:25 ` [PATCH bpf-next v2 04/12] bpf: Support __int128 as a by-value function argument Yonghong Song
2026-09-09 6:25 ` [PATCH bpf-next v2 05/12] bpf: Rename bpf_call_summary::num_params to arg_slot_cnt Yonghong Song
2026-09-09 6:25 ` [PATCH bpf-next v2 06/12] bpf: Recognize by-value struct and __int128 kfunc arguments Yonghong Song
2026-09-09 6:46 ` sashiko-bot
2026-09-11 4:31 ` Yonghong Song
2026-09-09 6:25 ` [PATCH bpf-next v2 07/12] bpf: Prepare kfunc arguments for the JIT from an ABI description Yonghong Song
2026-09-09 6:46 ` sashiko-bot
2026-09-11 5:05 ` Yonghong Song
2026-09-09 6:26 ` [PATCH bpf-next v2 08/12] bpf, x86: Move kfunc arguments into the x86-64 calling convention Yonghong Song
2026-09-09 7:29 ` bot+bpf-ci
2026-09-11 5:32 ` Yonghong Song
2026-09-09 6:26 ` [PATCH bpf-next v2 09/12] bpf, arm64: Move kfunc arguments into the arm64 " Yonghong Song
2026-09-09 7:30 ` bot+bpf-ci
2026-09-11 5:34 ` Yonghong Song [this message]
2026-09-09 6:26 ` [PATCH bpf-next v2 10/12] selftests/bpf: Add C tests for by-value arguments up to 16 bytes Yonghong Song
2026-09-09 6:26 ` [PATCH bpf-next v2 11/12] selftests/bpf: Add inline-asm tests for by-value arguments Yonghong Song
2026-09-09 7:30 ` bot+bpf-ci
2026-09-11 5:37 ` Yonghong Song
2026-09-09 6:26 ` [PATCH bpf-next v2 12/12] selftests/bpf: Add tests for by-value kfunc arguments Yonghong Song
2026-09-09 7:30 ` bot+bpf-ci
2026-09-11 5:57 ` Yonghong Song
Reply instructions:
You may reply publicly to this message via plain-text email
using any one of the following methods:
* Save the following mbox file, import it into your mail client,
and reply-to-all from there: mbox
Avoid top-posting and favor interleaved quoting:
https://en.wikipedia.org/wiki/Posting_style#Interleaved_style
* Reply using the --to, --cc, and --in-reply-to
switches of git-send-email(1):
git send-email \
--in-reply-to=4b6edd53-acda-4852-963d-fe233fd2d192@linux.dev \
--to=yonghong.song@linux.dev \
--cc=andrii@kernel.org \
--cc=ast@kernel.org \
--cc=bot+bpf-ci@kernel.org \
--cc=bpf@vger.kernel.org \
--cc=daniel@iogearbox.net \
--cc=eddyz87@gmail.com \
--cc=ihor.solodrai@linux.dev \
--cc=kernel-team@fb.com \
--cc=martin.lau@kernel.org \
--cc=mason@kernel.org \
/path/to/YOUR_REPLY
https://kernel.org/pub/software/scm/git/docs/git-send-email.html
* If your mail client supports setting the In-Reply-To header
via mailto: links, try the mailto: link
Be sure your reply has a Subject: header at the top and a blank line
before the message body.
This is a public inbox, see mirroring instructions
for how to clone and mirror all data and code used for this inbox;
as well as URLs for NNTP newsgroup(s).