From: Paolo Abeni <pabeni@redhat.com>
To: netdev-bot+sashiko@kernel.org, inbal.lipshtat@mail.huji.ac.il
Cc: netdev@vger.kernel.org, edumazet@google.com,
ncardwell@google.com, kuniyu@google.com, davem@davemloft.net,
kuba@kernel.org, horms@kernel.org, shuah@kernel.org,
linux-kselftest@vger.kernel.org, amit.klein@mail.huji.ac.il
Subject: Re: [PATCH net v3 2/2] selftests: net: packetdrill: test exclusion of old ACK from TCP fast path
Date: Thu, 17 Sep 2026 15:17:30 +0200 [thread overview]
Message-ID: <6fc1f088-b139-418b-a6e8-d273d235f444@redhat.com> (raw)
In-Reply-To: <178959269660.22033.12533704715181509708@kernel.org>
On 9/16/26 23:04, netdev-bot+sashiko@kernel.org wrote:
> Thank you for your contribution! Sashiko AI review found 1 potential
> issue(s) to consider:
>
> Critical: 0 · High: 0 · Medium: 1 · Low: 0
>
> - [Medium] The new packetdrill test
> tools/testing/selftests/net/packetdrill/tcp_rfc5961_reject-old-ack.pkt…
>
> --
>
> selftests: net: packetdrill: test exclusion of old ACK from TCP fast path
>
> This adds a packetdrill script that primes the receive queue with a 500
> byte segment and then sends an in-sequence 1000 byte data segment carrying
> an excessively old ACK, expecting a challenge ACK with RCV.NXT unchanged.
> It is based on the reproducer from commit 3d501dd326fb ("tcp: do not accept
> ACK of bytes we never sent").
>
> No Link: tags in the commit header.
>
>> diff --git a/tools/testing/selftests/net/packetdrill/tcp_rfc5961_reject-old-ack.pkt b/tools/testing/selftests/net/packetdrill/tcp_rfc5961_reject-old-ack.pkt
>> new file mode 100644
>> index 0000000000000..32dd9de1d3662
>> --- /dev/null
>> +++ b/tools/testing/selftests/net/packetdrill/tcp_rfc5961_reject-old-ack.pkt
>> @@ -0,0 +1,29 @@
>> +// SPDX-License-Identifier: GPL-2.0
>> +
>> +`./defaults.sh
>> +sysctl -q net.ipv4.tcp_invalid_ratelimit=0
>> +`
>> +
>> +// Test rejection of data segments carrying excessively old ACKs
>> +
>
> [ ... socket setup and handshake snipped ... ]
>
>> +// Populate receive memory so the following segment can use
>> +// header prediction.
>> +0 < P. 1:501(500) ack 1 win 65535
>> +0 > . 1:1(0) ack 501
>> +
>> +// Send an in-sequence data segment carrying an excessively old ACK.
>> +0 < P. 501:1501(1000) ack 2794967397 win 65535
>> +
>> +// Challenge ACK; RCV.NXT must remain 501.
>> +0 > . 1:1(0) ack 501
>
> [Severity: Medium]
>
> Can this script also pass on a kernel that does not have the fast path old
> ACK exclusion? In tcp_rcv_established() there appear to be two independent
> routes to the exact output asserted here.
>
> The route the commit message describes is the header prediction gate:
>
> net/ipv4/tcp_input.c:tcp_rcv_established() {
> if ((tcp_flag_word(th) & TCP_HP_BITS) == tp->pred_flags &&
> TCP_SKB_CB(skb)->seq == tp->rcv_nxt &&
> between(TCP_SKB_CB(skb)->ack_seq, tp->snd_una, tp->snd_nxt)) {
> ...
> }
>
> But if that gate is absent and the fast path is entered, the segment can
> still leave it before being queued:
>
> net/ipv4/tcp_input.c:tcp_rcv_established() {
> ...
> if ((int)skb->truesize > sk->sk_forward_alloc)
> goto step5;
> ...
> }
>
> Both routes end up in the same place:
>
> net/ipv4/tcp_input.c:tcp_ack() {
> ...
> if (before(ack, prior_snd_una - max_window)) {
> if (!(flag & FLAG_NO_CHALLENGE_ACK))
> tcp_send_challenge_ack(sk, false);
> return -SKB_DROP_REASON_TCP_TOO_OLD_ACK;
> }
> ...
> }
>
> which emits one bare ACK with ack=501 and leaves rcv_nxt at 501, so the
> single expectation "+0 > . 1:1(0) ack 501" is satisfied either way.
>
> Given that, does the regression value of the test rest entirely on the
> sk_forward_alloc head-room left by the 500 byte priming segment, which the
> script does not assert? Forward allocation is granted in whole pages:
>
> net/core/sock.c:__sk_mem_schedule() {
> int ret, amt = sk_mem_pages(size);
>
> sk_forward_alloc_add(sk, amt << PAGE_SHIFT);
> ...
> }
>
> so the residual head-room after the priming segment is roughly PAGE_SIZE
> minus the truesize of that skb. With the default MAX_SKB_FRAGS of 17 the
> 500 byte skb lands in kmalloc-1024 (truesize around 1280), leaving enough
> room for the 1000 byte segment (truesize around 2304), and an unfixed
> kernel would queue the payload and fail the script.
>
> With CONFIG_MAX_SKB_FRAGS=45 (BIG TCP), skb_shared_info grows by 28 * 16
> bytes and the priming skb moves up a kmalloc bucket, leaving under 2048
> bytes of forward allocation:
>
> include/linux/skbuff.h:
> #ifndef CONFIG_MAX_SKB_FRAGS
> # define CONFIG_MAX_SKB_FRAGS 17
> #endif
>
> #define MAX_SKB_FRAGS CONFIG_MAX_SKB_FRAGS
>
> In that configuration an unfixed kernel would take the truesize bail-out,
> emit the same "ack 501" and report a pass while covering nothing. Other
> PAGE_SIZE, NET_SKB_PAD, kmalloc bucket or debug allocator combinations
> look like they can have the same effect.
>
> Would it be worth pinning the path with an nstat bracket, the way the
> neighbouring tests do, for example on TcpExtTCPHPHits or
> TcpExtTCPChallengeACK? tcp_rcv_big_endseq.pkt uses:
>
> 0 `nstat -n`
> ...
> +0 `nstat | grep TcpExtBeyondWindow | grep -q " 3 "`
>
> That would make the script fail rather than silently pass if the segment
> reaches the fast path.
IIRC the nipa CI runs with CONFIG_MAX_SKB_FRAGS == 17. The above could be
a possible follow-up, not blocking.
/P
next prev parent reply other threads:[~2026-09-17 13:17 UTC|newest]
Thread overview: 9+ messages / expand[flat|nested] mbox.gz Atom feed top
2026-09-14 9:04 [PATCH net v3 0/2] tcp: exclude old ACKs from fast path Inbal Schussheim
2026-09-14 9:04 ` [PATCH net v3 1/2] tcp: exclude old ACKs from tcp " Inbal Schussheim
2026-09-14 10:40 ` Eric Dumazet
2026-09-16 21:04 ` netdev-bot+sashiko
2026-09-14 9:04 ` [PATCH net v3 2/2] selftests: net: packetdrill: test exclusion of old ACK from TCP " Inbal Schussheim
2026-09-14 10:41 ` Eric Dumazet
2026-09-16 21:04 ` netdev-bot+sashiko
2026-09-17 13:17 ` Paolo Abeni [this message]
2026-09-17 13:30 ` [PATCH net v3 0/2] tcp: exclude old ACKs from " patchwork-bot+netdevbpf
Reply instructions:
You may reply publicly to this message via plain-text email
using any one of the following methods:
* Save the following mbox file, import it into your mail client,
and reply-to-all from there: mbox
Avoid top-posting and favor interleaved quoting:
https://en.wikipedia.org/wiki/Posting_style#Interleaved_style
* Reply using the --to, --cc, and --in-reply-to
switches of git-send-email(1):
git send-email \
--in-reply-to=6fc1f088-b139-418b-a6e8-d273d235f444@redhat.com \
--to=pabeni@redhat.com \
--cc=amit.klein@mail.huji.ac.il \
--cc=davem@davemloft.net \
--cc=edumazet@google.com \
--cc=horms@kernel.org \
--cc=inbal.lipshtat@mail.huji.ac.il \
--cc=kuba@kernel.org \
--cc=kuniyu@google.com \
--cc=linux-kselftest@vger.kernel.org \
--cc=ncardwell@google.com \
--cc=netdev-bot+sashiko@kernel.org \
--cc=netdev@vger.kernel.org \
--cc=shuah@kernel.org \
/path/to/YOUR_REPLY
https://kernel.org/pub/software/scm/git/docs/git-send-email.html
* If your mail client supports setting the In-Reply-To header
via mailto: links, try the mailto: link
Be sure your reply has a Subject: header at the top and a blank line
before the message body.
This is a public inbox, see mirroring instructions
for how to clone and mirror all data and code used for this inbox