From: Joel Fernandes <joelagnelf@nvidia.com>
To: linux-kernel@vger.kernel.org
Cc: "Paul E . McKenney" <paulmck@kernel.org>,
Frederic Weisbecker <frederic@kernel.org>,
Neeraj Upadhyay <neeraj.upadhyay@kernel.org>,
Joel Fernandes <joelagnelf@nvidia.com>,
Josh Triplett <josh@joshtriplett.org>,
Boqun Feng <boqun.feng@gmail.com>,
Steven Rostedt <rostedt@goodmis.org>,
Mathieu Desnoyers <mathieu.desnoyers@efficios.com>,
Lai Jiangshan <jiangshanlai@gmail.com>,
Zqiang <qiang.zhang@linux.dev>,
Uladzislau Rezki <urezki@gmail.com>,
joel@joelfernandes.org, rcu@vger.kernel.org
Subject: [PATCH RFC 13/14] rcu: Skip rnp addition when no grace period waiting
Date: Fri, 2 Jan 2026 19:23:42 -0500 [thread overview]
Message-ID: <20260103002343.6599-14-joelagnelf@nvidia.com> (raw)
In-Reply-To: <20260103002343.6599-1-joelagnelf@nvidia.com>
This is the key optimization commit that triggers the per-CPU blocked
task list promotion mechanism.
When a GP is waiting, add directly to rnp->blkd_tasks via
rcu_preempt_ctxt_queue(), but NOT to the per-CPU list.
However, when no GP is waiting on this CPU, skip adding to rnp->blkd_tasks
entirely. This completely avoids rnp->lock acquisition in this path
triggering the optimization.
Signed-off-by: Joel Fernandes <joelagnelf@nvidia.com>
---
kernel/rcu/tree_plugin.h | 64 ++++++++++++++++++++++++----------------
1 file changed, 38 insertions(+), 26 deletions(-)
diff --git a/kernel/rcu/tree_plugin.h b/kernel/rcu/tree_plugin.h
index d43dd153c152..a0cd50f1e6c5 100644
--- a/kernel/rcu/tree_plugin.h
+++ b/kernel/rcu/tree_plugin.h
@@ -335,37 +335,43 @@ void rcu_note_context_switch(bool preempt)
/* Possibly blocking in an RCU read-side critical section. */
rnp = rdp->mynode;
- raw_spin_lock_rcu_node(rnp);
t->rcu_read_unlock_special.b.blocked = true;
- t->rcu_blocked_node = rnp;
#ifdef CONFIG_RCU_PER_CPU_BLOCKED_LISTS
/*
- * If no GP is waiting on this CPU, add to per-CPU list as well
- * so promotion can find it if a GP starts later. If GP waiting,
- * skip per-CPU list - task goes only to rnp->blkd_tasks (same
- * behavior as before per-CPU lists were added).
+ * Check if a GP is in progress.
*/
if (!rcu_gp_in_progress() && !rdp->cpu_no_qs.b.norm && !rdp->cpu_no_qs.b.exp) {
+ /*
+ * No GP waiting on this CPU. Add to per-CPU list only,
+ * skipping rnp->lock for better scalability.
+ */
+ t->rcu_blocked_node = NULL;
t->rcu_blocked_cpu = rdp->cpu;
raw_spin_lock(&rdp->blkd_lock);
list_add(&t->rcu_rdp_entry, &rdp->blkd_list);
raw_spin_unlock(&rdp->blkd_lock);
- }
+ trace_rcu_preempt_task(rcu_state.name, t->pid,
+ rcu_seq_snap(&rnp->gp_seq));
+ } else
#endif
+ /* GP waiting (or per-CPU lists disabled) - add to rnp. */
+ {
+ raw_spin_lock_rcu_node(rnp);
+ t->rcu_blocked_node = rnp;
- /*
- * Verify the CPU's sanity, trace the preemption, and
- * then queue the task as required based on the states
- * of any ongoing and expedited grace periods.
- */
- WARN_ON_ONCE(!rcu_rdp_cpu_online(rdp));
- WARN_ON_ONCE(!list_empty(&t->rcu_node_entry));
- trace_rcu_preempt_task(rcu_state.name,
- t->pid,
- (rnp->qsmask & rdp->grpmask)
- ? rnp->gp_seq
- : rcu_seq_snap(&rnp->gp_seq));
- rcu_preempt_ctxt_queue(rnp, rdp);
+ /*
+ * Verify the CPU's sanity, trace the preemption, and
+ * then queue the task as required based on the states
+ * of any ongoing and expedited grace periods.
+ */
+ WARN_ON_ONCE(!rcu_rdp_cpu_online(rdp));
+ WARN_ON_ONCE(!list_empty(&t->rcu_node_entry));
+ trace_rcu_preempt_task(rcu_state.name, t->pid,
+ (rnp->qsmask & rdp->grpmask)
+ ? rnp->gp_seq
+ : rcu_seq_snap(&rnp->gp_seq));
+ rcu_preempt_ctxt_queue(rnp, rdp);
+ }
} else {
rcu_preempt_deferred_qs(t);
}
@@ -568,13 +574,22 @@ rcu_preempt_deferred_qs_irqrestore(struct task_struct *t, unsigned long flags)
*/
rnp = t->rcu_blocked_node;
#ifdef CONFIG_RCU_PER_CPU_BLOCKED_LISTS
- /* Remove from per-CPU list if task was added to it. */
blocked_cpu = t->rcu_blocked_cpu;
if (blocked_cpu != -1) {
+ /*
+ * Task is on per-CPU list. Remove it and check if
+ * it was promoted to rnp->blkd_tasks.
+ */
blocked_rdp = per_cpu_ptr(&rcu_data, blocked_cpu);
raw_spin_lock(&blocked_rdp->blkd_lock);
list_del_init(&t->rcu_rdp_entry);
t->rcu_blocked_cpu = -1;
+
+ /*
+ * Read rcu_blocked_node while holding blkd_lock to
+ * serialize with rcu_promote_blocked_tasks().
+ */
+ rnp = t->rcu_blocked_node;
raw_spin_unlock(&blocked_rdp->blkd_lock);
/*
* TODO: This should just be "WARN_ON_ONCE(rnp); return;" since after
@@ -584,15 +599,12 @@ rcu_preempt_deferred_qs_irqrestore(struct task_struct *t, unsigned long flags)
* from the rdp blocked list and early returning.
*/
if (!rnp) {
- /*
- * Task was only on per-CPU list, not on rnp list.
- * This can happen in future when tasks are added
- * only to rdp initially and promoted to rnp later.
- */
+ /* Not promoted - no GP waiting for this task. */
local_irq_restore(flags);
return;
}
}
+ /* else: Task went directly to rnp->blkd_tasks. */
#endif
raw_spin_lock_rcu_node(rnp); /* irqs already disabled. */
WARN_ON_ONCE(rnp != t->rcu_blocked_node);
--
2.34.1
next prev parent reply other threads:[~2026-01-03 0:24 UTC|newest]
Thread overview: 33+ messages / expand[flat|nested] mbox.gz Atom feed top
2026-01-03 0:23 [PATCH RFC 00/14] rcu: Reduce rnp->lock contention with per-CPU blocked task lists Joel Fernandes
2026-01-03 0:23 ` [PATCH RFC 01/14] rcu: Add WARN_ON_ONCE for blocked flag invariant in exit_rcu() Joel Fernandes
2026-01-05 15:31 ` Steven Rostedt
2026-01-05 15:44 ` Joel Fernandes
2026-01-03 0:23 ` [PATCH RFC 02/14] rcu: Add per-CPU blocked task lists for PREEMPT_RCU Joel Fernandes
2026-01-05 15:48 ` Steven Rostedt
2026-01-03 0:23 ` [PATCH RFC 03/14] rcu: Early return during unlock for tasks only on per-CPU blocked list Joel Fernandes
2026-01-03 0:23 ` [PATCH RFC 04/14] rcu: Promote blocked tasks from per-CPU to rnp lists Joel Fernandes
2026-01-05 15:59 ` Steven Rostedt
2026-01-09 3:52 ` Joel Fernandes
2026-01-03 0:23 ` [PATCH RFC 05/14] rcu: Promote blocked tasks for expedited GPs Joel Fernandes
2026-01-03 0:23 ` [PATCH RFC 06/14] rcu: Promote per-CPU blocked tasks before checking for blocked readers Joel Fernandes
2026-01-03 0:23 ` [PATCH RFC 07/14] rcu: Promote late-arriving blocked tasks before reporting QS Joel Fernandes
2026-01-03 0:23 ` [PATCH RFC 08/14] rcu: Promote blocked tasks before QS report in force_qs_rnp() Joel Fernandes
2026-01-03 0:23 ` [PATCH RFC 09/14] rcu: Promote blocked tasks before QS report in rcutree_report_cpu_dead() Joel Fernandes
2026-01-03 0:23 ` [PATCH RFC 10/14] rcu: Promote blocked tasks before QS report in rcu_gp_init() Joel Fernandes
2026-01-03 0:23 ` [PATCH RFC 11/14] rcu: Add per-CPU blocked list check in exit_rcu() Joel Fernandes
2026-01-03 0:23 ` [PATCH RFC 12/14] rcu: Skip per-CPU list addition when GP already started Joel Fernandes
2026-01-03 0:23 ` Joel Fernandes [this message]
2026-01-03 0:23 ` [PATCH RFC 14/14] rcu: Remove checking of per-cpu blocked list against the node list Joel Fernandes
2026-01-05 16:46 ` [PATCH RFC 00/14] rcu: Reduce rnp->lock contention with per-CPU blocked task lists Paul E. McKenney
2026-01-06 0:55 ` Joel Fernandes
2026-01-06 15:08 ` Joel Fernandes
2026-01-06 19:24 ` Paul E. McKenney
2026-01-06 21:24 ` Joel Fernandes
2026-01-09 2:00 ` Paul E. McKenney
2026-01-06 19:17 ` Paul E. McKenney
2026-01-06 20:19 ` Steven Rostedt
2026-01-06 20:35 ` Paul E. McKenney
2026-01-06 20:49 ` Joel Fernandes
2026-01-09 1:55 ` Paul E. McKenney
2026-01-06 20:40 ` Joel Fernandes
2026-01-09 1:52 ` Paul E. McKenney
Reply instructions:
You may reply publicly to this message via plain-text email
using any one of the following methods:
* Save the following mbox file, import it into your mail client,
and reply-to-all from there: mbox
Avoid top-posting and favor interleaved quoting:
https://en.wikipedia.org/wiki/Posting_style#Interleaved_style
* Reply using the --to, --cc, and --in-reply-to
switches of git-send-email(1):
git send-email \
--in-reply-to=20260103002343.6599-14-joelagnelf@nvidia.com \
--to=joelagnelf@nvidia.com \
--cc=boqun.feng@gmail.com \
--cc=frederic@kernel.org \
--cc=jiangshanlai@gmail.com \
--cc=joel@joelfernandes.org \
--cc=josh@joshtriplett.org \
--cc=linux-kernel@vger.kernel.org \
--cc=mathieu.desnoyers@efficios.com \
--cc=neeraj.upadhyay@kernel.org \
--cc=paulmck@kernel.org \
--cc=qiang.zhang@linux.dev \
--cc=rcu@vger.kernel.org \
--cc=rostedt@goodmis.org \
--cc=urezki@gmail.com \
/path/to/YOUR_REPLY
https://kernel.org/pub/software/scm/git/docs/git-send-email.html
* If your mail client supports setting the In-Reply-To header
via mailto: links, try the mailto: link
Be sure your reply has a Subject: header at the top and a blank line
before the message body.
This is a public inbox, see mirroring instructions
for how to clone and mirror all data and code used for this inbox
all inboxes | Powered by JetHome®