From: Stefano Garzarella <sgarzare@redhat.com>
To: David Carlier <devnexen@gmail.com>
Cc: netdev@vger.kernel.org, bobbyeshleman@gmail.com,
davem@davemloft.net, edumazet@google.com, kuba@kernel.org,
pabeni@redhat.com, horms@kernel.org,
virtualization@lists.linux.dev, linux-kernel@vger.kernel.org
Subject: Re: [PATCH net-next v2 1/2] vsock: report pending receive data to io_uring
Date: Mon, 28 Sep 2026 17:00:12 +0200 [thread overview]
Message-ID: <arqAF53k2arP4enr@sgarzare-redhat> (raw)
In-Reply-To: <20260923184008.153541-2-devnexen@gmail.com>
On Wed, Sep 23, 2026 at 07:40:07PM +0100, David Carlier wrote:
> AF_VSOCK stream receives never fill msghdr.msg_inq, so io_uring cannot set
> IORING_CQE_F_SOCK_NONEMPTY and retries a multishot receive even after the
> queue has been drained.
>
> Fill the hint at the common receive exit using the transport callback that
> SIOCINQ already uses, and report 1 once the connection is finished so the
> caller performs the receive which observes EOF, as TCP does after a FIN.
>
> A vsock loopback ping-pong with io_uring multishot receive drops entries
> into __vsock_connectible_recvmsg from 1.97 to 1.00 per delivered message,
> and receiver CPU time by about 3% (25 runs of 50000 messages, p=0.006).
>
> Signed-off-by: David Carlier <devnexen@gmail.com>
> ---
> net/vmw_vsock/af_vsock.c | 35 +++++++++++++++++++++++++++++++++++
> 1 file changed, 35 insertions(+)
>
> diff --git a/net/vmw_vsock/af_vsock.c b/net/vmw_vsock/af_vsock.c
> index f840498b58af..20d6f9ca6a96 100644
> --- a/net/vmw_vsock/af_vsock.c
> +++ b/net/vmw_vsock/af_vsock.c
> @@ -2543,6 +2543,35 @@ static int __vsock_seqpacket_recvmsg(struct sock *sk, struct msghdr *msg,
> return err;
> }
>
> +/* Bytes a following receive can consume, 1 if it would only see EOF, or -1
> + * if the transport cannot tell.
> + *
> + * Called under the socket lock after a nonnegative stream receive, so a NULL
> + * transport implies SOCK_DONE.
> + */
> +static int vsock_stream_inq_hint(struct sock *sk)
> +{
> + struct vsock_sock *vsk = vsock_sk(sk);
> + s64 data;
> +
> + if ((sk->sk_shutdown & RCV_SHUTDOWN) || !vsk->transport ||
> + (sock_flag(sk, SOCK_DONE) && sk->sk_state != TCP_ESTABLISHED))
> + return 1;
> +
> + data = vsock_stream_has_data(vsk);
> + if (data < 0)
> + return -1;
> + if (data > 0)
> + return min_t(s64, data, INT_MAX);
> +
> + /* Empty but finished: keep the caller reading so it sees EOF. */
> + if (sock_flag(sk, SOCK_DONE) ||
> + (READ_ONCE(vsk->peer_shutdown) & SEND_SHUTDOWN))
> + return 1;
> +
> + return 0;
> +}
> +
> int
> __vsock_connectible_recvmsg(struct socket *sock, struct msghdr *msg, size_t len,
> int flags)
> @@ -2606,6 +2635,12 @@ __vsock_connectible_recvmsg(struct socket *sock, struct msghdr *msg, size_t len,
> err = __vsock_seqpacket_recvmsg(sk, msg, len, flags);
>
> out:
> + /* Seqpacket has_data counts messages, while io_uring treats msg_inq as
Do you mean vsock_stream_has_data() on SOCK_SEQPACKET doesn't return
bytes?
IMO names are confusing, but it should return number of bytes, we should
fix it, but of course not in this patch.
LGTM:
Reviewed-by: Stefano Garzarella <sgarzare@redhat.com>
> + * a byte length when sizing retries, so only streams report a hint.
> + */
> + if (msg->msg_get_inq && err >= 0 && sk->sk_type == SOCK_STREAM)
> + msg->msg_inq = vsock_stream_inq_hint(sk);
> +
> release_sock(sk);
> return err;
> }
> --
> 2.55.0
>
next prev parent reply other threads:[~2026-09-28 15:00 UTC|newest]
Thread overview: 7+ messages / expand[flat|nested] mbox.gz Atom feed top
2026-09-23 18:40 [PATCH net-next v2 0/2] vsock: receive queue hint for io_uring David Carlier
2026-09-23 18:40 ` [PATCH net-next v2 1/2] vsock: report pending receive data to io_uring David Carlier
2026-09-28 15:00 ` Stefano Garzarella [this message]
2026-09-23 18:40 ` [PATCH net-next v2 2/2] vsock/test: cover receive queue hints David Carlier
2026-09-23 19:29 ` Bobby Eshleman
2026-09-28 15:00 ` Stefano Garzarella
2026-09-29 0:30 ` [PATCH net-next v2 0/2] vsock: receive queue hint for io_uring patchwork-bot+netdevbpf
Reply instructions:
You may reply publicly to this message via plain-text email
using any one of the following methods:
* Save the following mbox file, import it into your mail client,
and reply-to-all from there: mbox
Avoid top-posting and favor interleaved quoting:
https://en.wikipedia.org/wiki/Posting_style#Interleaved_style
* Reply using the --to, --cc, and --in-reply-to
switches of git-send-email(1):
git send-email \
--in-reply-to=arqAF53k2arP4enr@sgarzare-redhat \
--to=sgarzare@redhat.com \
--cc=bobbyeshleman@gmail.com \
--cc=davem@davemloft.net \
--cc=devnexen@gmail.com \
--cc=edumazet@google.com \
--cc=horms@kernel.org \
--cc=kuba@kernel.org \
--cc=linux-kernel@vger.kernel.org \
--cc=netdev@vger.kernel.org \
--cc=pabeni@redhat.com \
--cc=virtualization@lists.linux.dev \
/path/to/YOUR_REPLY
https://kernel.org/pub/software/scm/git/docs/git-send-email.html
* If your mail client supports setting the In-Reply-To header
via mailto: links, try the mailto: link
Be sure your reply has a Subject: header at the top and a blank line
before the message body.
This is a public inbox, see mirroring instructions
for how to clone and mirror all data and code used for this inbox
all inboxes | Powered by JetHome®