From: luoxuanqiang <xuanqiang.luo@linux.dev>
To: bot+bpf-ci@kernel.org, bpf@vger.kernel.org
Cc: ast@kernel.org, daniel@iogearbox.net, andrii@kernel.org,
martin.lau@linux.dev, eddyz87@gmail.com, memxor@gmail.com,
song@kernel.org, yonghong.song@linux.dev, jolsa@kernel.org,
linux-kernel@vger.kernel.org, luoxuanqiang@kylinos.cn,
martin.lau@kernel.org, clm@meta.com, ihor.solodrai@linux.dev
Subject: Re: [PATCH bpf-next v1] bpf: Use hlist_nulls_replace_rcu() when updating htab elements
Date: Wed, 27 May 2026 17:01:36 +0800 [thread overview]
Message-ID: <68d25e1a-8c04-4344-afbd-8578e539a35c@linux.dev> (raw)
In-Reply-To: <de6652ce9d1367daf217fcb5793419db9d967757b7b7857b00ae1c0c008f46ec@mail.kernel.org>
在 2026/5/27 16:06, bot+bpf-ci@kernel.org 写道:
>> bpf: Use hlist_nulls_replace_rcu() when updating htab elements
>>
>> When updating an existing element, the current code does
>> hlist_nulls_add_head_rcu() followed by hlist_nulls_del_rcu(), exposing
>> a transient state where both old and new elements are visible to
>> concurrent RCU traversals.
>>
>> Use hlist_nulls_replace_rcu() to atomically replace the old element in
>> place, so RCU readers always see a consistent hash chain.
>>
>> Signed-off-by: Xuanqiang Luo <luoxuanqiang@kylinos.cn>
> This looks like a bug fix that should include:
>
> Fixes: 4fe8435909fd ("bpf: convert htab map to hlist_nulls")
>
> The bug was introduced by commit 4fe8435909fd which added the
> hlist_nulls_add_head_rcu() followed by hlist_nulls_del_rcu() pattern,
> exposing a transient state where both old and new elements are visible
> to concurrent RCU traversals.
>
>> diff --git a/kernel/bpf/hashtab.c b/kernel/bpf/hashtab.c
>> index 3dd9b4924ae4..b20e82bd253d 100644
>> --- a/kernel/bpf/hashtab.c
>> +++ b/kernel/bpf/hashtab.c
>> @@ -1246,18 +1246,16 @@ static long htab_map_update_elem(struct bpf_map *map, void *key, void *value,
>> goto err;
>> }
>>
>> - /* add new element to the head of the list, so that
>> - * concurrent search will find it before old elem
>> - */
>> - hlist_nulls_add_head_rcu(&l_new->hash_node, head);
>> if (l_old) {
>> - hlist_nulls_del_rcu(&l_old->hash_node);
>> + hlist_nulls_replace_rcu(&l_old->hash_node, &l_new->hash_node);
>>
>> /* l_old has already been stashed in htab->extra_elems, free
>> * its special fields before it is available for reuse.
>> */
>> if (htab_is_prealloc(htab))
>> check_and_free_fields(htab, l_old);
>> + } else {
>> + hlist_nulls_add_head_rcu(&l_new->hash_node, head);
>> }
> Does this atomic replacement approach actually solve the lookup problem?
>
> During review of Hou Tao's earlier v1 patch that used the same approach,
> Alexei Starovoitov raised a fundamental concern:
>
> "This atomic replacement doesn't really solve anything. lookup will see
> one element. That element could be deleted by another thread. bucket
> lock and either two step update or single step don't change anything
> from the pov of bpf prog doing lookup."
>
> Reference: https://lore.kernel.org/bpf/CAADnVQKD94q-G4N=w9PJU+k6gPhM8GmUYcyfj=33B_mKX6Qbjw@mail.gmail.com/
>
> Hou Tao acknowledged this limitation and abandoned this approach in v2,
> instead implementing atomic updates only for htab of maps. This patch
> appears to be a re-implementation of the abandoned v1 approach without
> addressing the fundamental concerns raised during that review.
>
> Can you clarify how this approach addresses the concerns that led to
> abandonment of the similar change in Hou Tao's earlier patch series?
>
> [ ... ]
>
>
> ---
> AI reviewed your patch. Please fix the bug or email reply why it's not a bug.
> See: https://github.com/kernel-patches/vmtest/blob/master/ci/claude/README.md
>
> CI run summary: https://github.com/kernel-patches/bpf/actions/runs/26496972451
I read the earlier discussion and realized I missed it, and I also missed
Hou Tao's updated patch 2c304172e0319 ("bpf: Support atomic update for htab
of maps") which already addressed this.
Sorry, My patch is outdated, please ignore it.
Thanks.
prev parent reply other threads:[~2026-05-27 9:02 UTC|newest]
Thread overview: 3+ messages / expand[flat|nested] mbox.gz Atom feed top
2026-05-27 7:01 xuanqiang.luo
2026-05-27 8:06 ` bot+bpf-ci
2026-05-27 9:01 ` luoxuanqiang [this message]
Reply instructions:
You may reply publicly to this message via plain-text email
using any one of the following methods:
* Save the following mbox file, import it into your mail client,
and reply-to-all from there: mbox
Avoid top-posting and favor interleaved quoting:
https://en.wikipedia.org/wiki/Posting_style#Interleaved_style
* Reply using the --to, --cc, and --in-reply-to
switches of git-send-email(1):
git send-email \
--in-reply-to=68d25e1a-8c04-4344-afbd-8578e539a35c@linux.dev \
--to=xuanqiang.luo@linux.dev \
--cc=andrii@kernel.org \
--cc=ast@kernel.org \
--cc=bot+bpf-ci@kernel.org \
--cc=bpf@vger.kernel.org \
--cc=clm@meta.com \
--cc=daniel@iogearbox.net \
--cc=eddyz87@gmail.com \
--cc=ihor.solodrai@linux.dev \
--cc=jolsa@kernel.org \
--cc=linux-kernel@vger.kernel.org \
--cc=luoxuanqiang@kylinos.cn \
--cc=martin.lau@kernel.org \
--cc=martin.lau@linux.dev \
--cc=memxor@gmail.com \
--cc=song@kernel.org \
--cc=yonghong.song@linux.dev \
/path/to/YOUR_REPLY
https://kernel.org/pub/software/scm/git/docs/git-send-email.html
* If your mail client supports setting the In-Reply-To header
via mailto: links, try the mailto: link
Be sure your reply has a Subject: header at the top and a blank line
before the message body.
This is a public inbox, see mirroring instructions
for how to clone and mirror all data and code used for this inbox
all inboxes | Powered by JetHome®