From: Allison Henderson <achender@kernel.org>
To: netdev@vger.kernel.org, linux-rdma@vger.kernel.org,
pabeni@redhat.com, edumazet@google.com, kuba@kernel.org,
horms@kernel.org
Cc: achender@kernel.org
Subject: [PATCH net-next v4 1/2] net/rds: restrict the rdma_cm ids to IB devices
Date: Sat, 26 Sep 2026 23:30:57 -0700 [thread overview]
Message-ID: <20260927063058.170273-2-achender@kernel.org> (raw)
In-Reply-To: <20260927063058.170273-1-achender@kernel.org>
RDS only has an IB transport, but it never tells the rdma_cm so. The
listener created in rds_rdma_listen_init() is therefore installed on
every RDMA device in the system, including iWARP RNICs, and the id an
outgoing connection resolves through in rds_ib_conn_path_connect()
may be bound to whichever device the address resolution picks:
rds_ib_laddr_check() only vouched for the local address being served
by one of RDS's IB devices, while cma_acquire_dev_by_src_ip() walks
every RDMA device for the same address, so a software iWARP device
attached to the same netdev can win. The event handler then runs
the IB transport's callbacks against a device that is not one.
The rdma_cm has an API for exactly this since commit a760e80e90f5
("RDMA/core: introduce rdma_restrict_node_type()"). Restrict all
three ids RDS creates - the listener, the per-connection id and the
probe id in rds_ib_laddr_check_cm() - to RDMA_NODE_IB_CA before they
are bound, so that the listener is only installed on IB devices, an
outgoing connection can only bind one, and the address check's bind
fails outright on anything else - which makes its explicit node_type
test dead, so it goes; the check that the bind produced a device at
all stays.
With that, no rdma_cm event reaches RDS's handler from a device it
has no transport for. The handler itself still assumes IB: it assigns
its transport pointer only for RDMA_NODE_IB_CA and dereferences it
regardless, which is the crash that first surfaced this. The fix for
that is a separate net patch, Aohan Mei's "net: rds: fix uninitialized
trans dereference in CM event handler", and is what stable kernels
without rdma_restrict_node_type() - which arrived in commit
a760e80e90f5 ("RDMA/core: introduce rdma_restrict_node_type()") - have
to take: this patch depends on that API, so stable trees need that
separate, minimal fix rather than a backport of this one. The two
patches are independent and apply in either order. The
listener has been on every RDMA device since the iWARP transport was
removed and left the ids unrestricted, hence the Fixes tag.
Fixes: dcdede0406d3 ("RDS: Drop stale iWARP RDMA transport")
Assisted-by: Claude-Code:claude-fable-5
Signed-off-by: Allison Henderson <achender@kernel.org>
---
net/rds/ib.c | 14 ++++++--------
net/rds/ib_cm.c | 12 ++++++++++++
net/rds/rdma_transport.c | 8 ++++++++
3 files changed, 26 insertions(+), 8 deletions(-)
diff --git a/net/rds/ib.c b/net/rds/ib.c
index 786f39169bc1..4ea9838d090c 100644
--- a/net/rds/ib.c
+++ b/net/rds/ib.c
@@ -414,13 +414,14 @@ static int rds_ib_laddr_check_cm(struct net *net, const struct in6_addr *addr,
bool isv4;
isv4 = ipv6_addr_v4mapped(addr);
- /* Create a CMA ID and try to bind it. This catches both
- * IB and iWARP capable NICs.
- */
+ /* Create a CMA ID restricted to IB devices and try to bind it. */
cm_id = rdma_create_id(&init_net, rds_rdma_cm_event_handler,
NULL, RDMA_PS_TCP, IB_QPT_RC);
if (IS_ERR(cm_id))
return PTR_ERR(cm_id);
+ ret = rdma_restrict_node_type(cm_id, RDMA_NODE_IB_CA);
+ if (ret)
+ goto out;
if (isv4) {
memset(&sin, 0, sizeof(sin));
@@ -473,12 +474,9 @@ static int rds_ib_laddr_check_cm(struct net *net, const struct in6_addr *addr,
#endif
}
- /* rdma_bind_addr will only succeed for IB & iWARP devices */
+ /* the restriction above means this only succeeds for IB devices */
ret = rdma_bind_addr(cm_id, sa);
- /* due to this, we will claim to support iWARP devices unless we
- check node_type. */
- if (ret || !cm_id->device ||
- cm_id->device->node_type != RDMA_NODE_IB_CA)
+ if (ret || !cm_id->device)
ret = -EADDRNOTAVAIL;
rdsdebug("addr %pI6c%%%u ret %d node type %d\n",
diff --git a/net/rds/ib_cm.c b/net/rds/ib_cm.c
index 3ebe13d00953..3ed03ad32812 100644
--- a/net/rds/ib_cm.c
+++ b/net/rds/ib_cm.c
@@ -999,6 +999,18 @@ int rds_ib_conn_path_connect(struct rds_conn_path *cp)
goto out;
}
+ /* rds_ib_laddr_check() only vouched for the local address being
+ * on an IB device; the address resolution below picks the device
+ * on its own, so restrict it to the same kind.
+ */
+ ret = rdma_restrict_node_type(ic->i_cm_id, RDMA_NODE_IB_CA);
+ if (ret) {
+ rdsdebug("rdma_restrict_node_type() failed: %d\n", ret);
+ rdma_destroy_id(ic->i_cm_id);
+ ic->i_cm_id = NULL;
+ goto out;
+ }
+
rdsdebug("created cm id %p for conn %p\n", ic->i_cm_id, conn);
if (ipv6_addr_v4mapped(&conn->c_faddr)) {
diff --git a/net/rds/rdma_transport.c b/net/rds/rdma_transport.c
index b15cf316b23a..91ff1dde26af 100644
--- a/net/rds/rdma_transport.c
+++ b/net/rds/rdma_transport.c
@@ -210,6 +210,14 @@ static int rds_rdma_listen_init_common(rdma_cm_event_handler handler,
return ret;
}
+ /* Only the IB transport is left, so only listen on IB devices */
+ ret = rdma_restrict_node_type(cm_id, RDMA_NODE_IB_CA);
+ if (ret) {
+ pr_err("RDS/RDMA: failed to setup listener, rdma_restrict_node_type() returned %d\n",
+ ret);
+ goto out;
+ }
+
/*
* XXX I bet this binds the cm_id to a device. If we want to support
* fail-over we'll have to take this into consideration.
--
2.25.1
next prev parent reply other threads:[~2026-09-27 6:30 UTC|newest]
Thread overview: 8+ messages / expand[flat|nested] mbox.gz Atom feed top
2026-09-27 6:30 [PATCH net-next v4 0/2] net/rds: restrict the rdma_cm ids to IB devices Allison Henderson
2026-09-27 6:30 ` Allison Henderson [this message]
2026-09-27 6:37 ` [PATCH net-next v4 1/2] " sashiko-bot
2026-10-01 12:32 ` Simon Horman
2026-09-27 6:30 ` [PATCH net-next v4 2/2] net/rds: log the port a listener actually bound Allison Henderson
2026-09-27 6:37 ` sashiko-bot
2026-10-01 12:32 ` Simon Horman
2026-10-02 0:20 ` [PATCH net-next v4 0/2] net/rds: restrict the rdma_cm ids to IB devices patchwork-bot+netdevbpf
Reply instructions:
You may reply publicly to this message via plain-text email
using any one of the following methods:
* Save the following mbox file, import it into your mail client,
and reply-to-all from there: mbox
Avoid top-posting and favor interleaved quoting:
https://en.wikipedia.org/wiki/Posting_style#Interleaved_style
* Reply using the --to, --cc, and --in-reply-to
switches of git-send-email(1):
git send-email \
--in-reply-to=20260927063058.170273-2-achender@kernel.org \
--to=achender@kernel.org \
--cc=edumazet@google.com \
--cc=horms@kernel.org \
--cc=kuba@kernel.org \
--cc=linux-rdma@vger.kernel.org \
--cc=netdev@vger.kernel.org \
--cc=pabeni@redhat.com \
/path/to/YOUR_REPLY
https://kernel.org/pub/software/scm/git/docs/git-send-email.html
* If your mail client supports setting the In-Reply-To header
via mailto: links, try the mailto: link
Be sure your reply has a Subject: header at the top and a blank line
before the message body.
This is a public inbox, see mirroring instructions
for how to clone and mirror all data and code used for this inbox