From: Tejun Heo <tj@kernel.org>
To: "Paul E. McKenney" <paulmck@kernel.org>
Cc: rcu@vger.kernel.org, linux-kernel@vger.kernel.org,
kernel-team@meta.com, rostedt@goodmis.org, riel@surriel.com
Subject: Re: [PATCH RFC rcu] Stop rcu_tasks_invoke_cbs() from using never-online CPUs
Date: Wed, 26 Apr 2023 12:12:05 -1000 [thread overview]
Message-ID: <ZEmhtdxpelt5jxAu@slm.duckdns.org> (raw)
In-Reply-To: <83d037d1-ef12-4b31-a7b9-7b1ed6c3ae42@paulmck-laptop>
On Wed, Apr 26, 2023 at 10:26:38AM -0700, Paul E. McKenney wrote:
> The rcu_tasks_invoke_cbs() relies on queue_work_on() to silently fall
> back to WORK_CPU_UNBOUND when the specified CPU is offline. However,
> the queue_work_on() function's silent fallback mechanism relies on that
> CPU having been online at some time in the past. When queue_work_on()
> is passed a CPU that has never been online, workqueue lockups ensue,
> which can be bad for your kernel's general health and well-being.
>
> This commit therefore checks whether a given CPU is currently online,
> and, if not substitutes WORK_CPU_UNBOUND in the subsequent call to
> queue_work_on(). Why not simply omit the queue_work_on() call entirely?
> Because this function is flooding callback-invocation notifications
> to all CPUs, and must deal with possibilities that include a sparse
> cpu_possible_mask.
>
> Fixes: d363f833c6d88 rcu-tasks: Use workqueues for multiple rcu_tasks_invoke_cbs() invocations
> Reported-by: Tejun Heo <tj@kernel.org>
> Signed-off-by: Paul E. McKenney <paulmck@kernel.org>
...
> + // If a CPU has never been online, queue_work_on()
> + // objects to queueing work on that CPU. Approximate a
> + // check for this by checking if the CPU is currently online.
> +
> + cpus_read_lock();
> + cpuwq1 = cpu_online(cpunext) ? cpunext : WORK_CPU_UNBOUND;
> + cpuwq2 = cpu_online(cpunext + 1) ? cpunext + 1 : WORK_CPU_UNBOUND;
> + cpus_read_unlock();
> +
> + // Yes, either CPU could go offline here. But that is
> + // OK because queue_work_on() will (in effect) silently
> + // fall back to WORK_CPU_UNBOUND for any CPU that has ever
> + // been online.
Looks like cpus_read_lock() isn't protecting anything really.
> + queue_work_on(cpuwq1, system_wq, &rtpcp_next->rtp_work);
> cpunext++;
> if (cpunext < smp_load_acquire(&rtp->percpu_dequeue_lim)) {
> rtpcp_next = per_cpu_ptr(rtp->rtpcpu, cpunext);
> - queue_work_on(cpunext, system_wq, &rtpcp_next->rtp_work);
> + queue_work_on(cpuwq2, system_wq, &rtpcp_next->rtp_work);
As discussed in the thread, I kinda wonder whether just using an unbound
workqueue would be sufficient but as a fix this looks good to me.
Acked-by: Tejun Heo <tj@kernel.org>
Thanks.
--
tejun
next prev parent reply other threads:[~2023-04-26 22:12 UTC|newest]
Thread overview: 8+ messages / expand[flat|nested] mbox.gz Atom feed top
2023-04-26 17:26 Paul E. McKenney
2023-04-26 19:55 ` Tejun Heo
2023-04-26 21:17 ` Paul E. McKenney
2023-04-26 21:31 ` Tejun Heo
2023-04-26 21:55 ` Paul E. McKenney
2023-04-26 22:09 ` Tejun Heo
2023-04-26 22:12 ` Tejun Heo [this message]
2023-04-26 22:29 ` Paul E. McKenney
Reply instructions:
You may reply publicly to this message via plain-text email
using any one of the following methods:
* Save the following mbox file, import it into your mail client,
and reply-to-all from there: mbox
Avoid top-posting and favor interleaved quoting:
https://en.wikipedia.org/wiki/Posting_style#Interleaved_style
* Reply using the --to, --cc, and --in-reply-to
switches of git-send-email(1):
git send-email \
--in-reply-to=ZEmhtdxpelt5jxAu@slm.duckdns.org \
--to=tj@kernel.org \
--cc=kernel-team@meta.com \
--cc=linux-kernel@vger.kernel.org \
--cc=paulmck@kernel.org \
--cc=rcu@vger.kernel.org \
--cc=riel@surriel.com \
--cc=rostedt@goodmis.org \
/path/to/YOUR_REPLY
https://kernel.org/pub/software/scm/git/docs/git-send-email.html
* If your mail client supports setting the In-Reply-To header
via mailto: links, try the mailto: link
Be sure your reply has a Subject: header at the top and a blank line
before the message body.
This is a public inbox, see mirroring instructions
for how to clone and mirror all data and code used for this inbox
all inboxes | Powered by JetHome®