From: Hannes Reinecke <hare@suse.de>
To: Chuck Lever <cel@kernel.org>, Jakub Kicinski <kuba@kernel.org>,
Paolo Abeni <pabeni@redhat.com>, Simon Horman <horms@kernel.org>,
John Fastabend <john.fastabend@gmail.com>,
Sabrina Dubroca <sd@queasysnail.net>,
Shuah Khan <shuah@kernel.org>, Jeff Layton <jlayton@kernel.org>,
NeilBrown <neil@brown.name>,
Olga Kornievskaia <okorniev@redhat.com>,
Dai Ngo <Dai.Ngo@oracle.com>, Tom Talpey <tom@talpey.com>
Cc: netdev@vger.kernel.org, kernel-tls-handshake@lists.linux.dev,
linux-kselftest@vger.kernel.org, linux-nfs@vger.kernel.org
Subject: Re: [PATCH net-next v2 2/6] net: Introduce read_sock_rectype proto_ops for control record delivery
Date: Thu, 23 Jul 2026 09:14:17 +0200 [thread overview]
Message-ID: <77756ade-5d55-4719-b9e6-7874bddd77c7@suse.de> (raw)
In-Reply-To: <20260720-tcp-read-sock-v2-2-29545d034f3c@kernel.org>
On 7/20/26 4:27 PM, Chuck Lever wrote:
> From: Chuck Lever <chuck.lever@oracle.com>
>
> Kernel TCP consumers that use the read_sock interface
> (proto_ops.read_sock) cannot receive TLS control messages (Alerts,
> Handshake records) when kTLS is active. The current
> tls_sw_read_sock() method rejects non-data records with -EINVAL, and
> the sk_read_actor_t callback has no channel for delivering record-
> type metadata.
>
> Four kernel subsystems are affected: NFSD (sunrpc svcsock), NFS
> client (sunrpc xprtsock), NVMe target (nvmet-tcp), and NVMe host
> (nvme-tcp). Each of these either falls back to the sock_recvmsg()
> API or lacks TLS alert handling entirely.
>
> A new read_sock_rectype method in struct proto_ops provides a
> separate code path that delivers non-data TLS records to a callback,
> without changing the behavior seen by existing read_sock consumers.
>
> The new sk_read_rectype_actor_t callback type extends the
> sk_read_actor_t signature with a rectype parameter carrying the
> protocol-layer record type (for example, TLS_RECORD_TYPE_ALERT). The
> record-type callback returns 0 to consume a record or a negative
> value to requeue it and stop delivery; unlike the data callback, its
> return value does not count bytes.
>
> Signed-off-by: Chuck Lever <chuck.lever@oracle.com>
> ---
> include/linux/net.h | 28 ++++++++++++++++++++++++++++
> 1 file changed, 28 insertions(+)
>
> diff --git a/include/linux/net.h b/include/linux/net.h
> index 277188a40c72..7a19a743a617 100644
> --- a/include/linux/net.h
> +++ b/include/linux/net.h
> @@ -198,6 +198,13 @@ struct sk_buff;
> struct proto_accept_arg;
> typedef int (*sk_read_actor_t)(read_descriptor_t *, struct sk_buff *,
> unsigned int, size_t);
> +/* rectype carries the transport record type, for example a
> + * TLS_RECORD_TYPE_* value.
> + */
> +typedef int (*sk_read_rectype_actor_t)(read_descriptor_t *,
> + struct sk_buff *,
> + unsigned int, size_t,
> + u8 rectype);
> typedef int (*skb_read_actor_t)(struct sock *, struct sk_buff *);
>
>
> @@ -264,6 +271,27 @@ struct proto_ops {
> */
> int (*read_sock)(struct sock *sk, read_descriptor_t *desc,
> sk_read_actor_t recv_actor);
> + /*
> + * read_sock_rectype splits delivery across two callbacks:
> + * recv_actor for data records, per the sk_read_actor_t
> + * convention, and rectype_actor for all other records,
> + * with rectype identifying each. A NULL rectype_actor
> + * leaves non-data records pending. rectype_actor returns 0
> + * to consume a record or negative to leave it pending for
> + * redelivery and stop delivery; the negative return is a
> + * backpressure signal, not a fatal error. Both callbacks
> + * report errors and early stop the way recv_actor does:
> + * by setting desc->count to 0 and recording the reason in
> + * desc->error, per the read_descriptor_t convention and
> + * independent of the return value. The return value reports
> + * only data bytes consumed by recv_actor; the caller
> + * detects an error or early stop via desc->count and
> + * desc->error.
> + */
> + int (*read_sock_rectype)(struct sock *sk,
> + read_descriptor_t *desc,
> + sk_read_actor_t recv_actor,
> + sk_read_rectype_actor_t rectype_actor);
> /* This is different from read_sock(), it reads an entire skb at a time. */
> int (*read_skb)(struct sock *sk, skb_read_actor_t recv_actor);
> int (*sendmsg_locked)(struct sock *sk, struct msghdr *msg,
>
Naming is a bit odd, but hey.
Reviewed-by: Hannes Reinecke <hare@kernel.org>
Cheers,
Hannes
--
Dr. Hannes Reinecke Kernel Storage Architect
hare@suse.de +49 911 74053 688
SUSE Software Solutions GmbH, Frankenstr. 146, 90461 Nürnberg
HRB 36809 (AG Nürnberg), GF: I. Totev, A. McDonald, W. Knoblich
next prev parent reply other threads:[~2026-07-23 7:14 UTC|newest]
Thread overview: 14+ messages / expand[flat|nested] mbox.gz Atom feed top
2026-07-20 14:27 [PATCH net-next v2 0/6] Deliver TLS control records to kernel read_sock consumers Chuck Lever
2026-07-20 14:27 ` [PATCH net-next v2 1/6] net/tls: Bound consecutive no-data records in tls_sw_read_sock() Chuck Lever
2026-07-23 7:11 ` Hannes Reinecke
2026-07-23 13:24 ` Chuck Lever
2026-07-23 14:23 ` Sabrina Dubroca
2026-07-23 14:29 ` Chuck Lever
2026-07-20 14:27 ` [PATCH net-next v2 2/6] net: Introduce read_sock_rectype proto_ops for control record delivery Chuck Lever
2026-07-23 7:14 ` Hannes Reinecke [this message]
2026-07-20 14:27 ` [PATCH net-next v2 3/6] tls: Implement read_sock_rectype for kTLS software path Chuck Lever
2026-07-23 7:14 ` Hannes Reinecke
2026-07-20 14:27 ` [PATCH net-next v2 4/6] selftests/tls: Add tests for data/control record interleaving Chuck Lever
2026-07-20 14:27 ` [PATCH net-next v2 5/6] SUNRPC: Use read_sock_rectype for svcsock TCP receives Chuck Lever
2026-07-20 14:28 ` [PATCH net-next v2 6/6] SUNRPC: Remove sock_recvmsg path from " Chuck Lever
2026-07-23 7:19 ` [PATCH net-next v2 0/6] Deliver TLS control records to kernel read_sock consumers Hannes Reinecke
Reply instructions:
You may reply publicly to this message via plain-text email
using any one of the following methods:
* Save the following mbox file, import it into your mail client,
and reply-to-all from there: mbox
Avoid top-posting and favor interleaved quoting:
https://en.wikipedia.org/wiki/Posting_style#Interleaved_style
* Reply using the --to, --cc, and --in-reply-to
switches of git-send-email(1):
git send-email \
--in-reply-to=77756ade-5d55-4719-b9e6-7874bddd77c7@suse.de \
--to=hare@suse.de \
--cc=Dai.Ngo@oracle.com \
--cc=cel@kernel.org \
--cc=horms@kernel.org \
--cc=jlayton@kernel.org \
--cc=john.fastabend@gmail.com \
--cc=kernel-tls-handshake@lists.linux.dev \
--cc=kuba@kernel.org \
--cc=linux-kselftest@vger.kernel.org \
--cc=linux-nfs@vger.kernel.org \
--cc=neil@brown.name \
--cc=netdev@vger.kernel.org \
--cc=okorniev@redhat.com \
--cc=pabeni@redhat.com \
--cc=sd@queasysnail.net \
--cc=shuah@kernel.org \
--cc=tom@talpey.com \
/path/to/YOUR_REPLY
https://kernel.org/pub/software/scm/git/docs/git-send-email.html
* If your mail client supports setting the In-Reply-To header
via mailto: links, try the mailto: link
Be sure your reply has a Subject: header at the top and a blank line
before the message body.
This is a public inbox, see mirroring instructions
for how to clone and mirror all data and code used for this inbox