* [PATCH V4] cgroup/cpuset: Prevent UAF in proc_cpuset_show()
@ 2024-06-28 1:36 Chen Ridong
2024-06-28 17:11 ` Tejun Heo
2024-06-30 9:20 ` Markus Elfring
0 siblings, 2 replies; 5+ messages in thread
From: Chen Ridong @ 2024-06-28 1:36 UTC (permalink / raw)
To: tj, lizefan.x, hannes, longman, adityakali, sergeh, mkoutny
Cc: cgroups, linux-kernel
An UAF can happen when /proc/cpuset is read as reported in [1].
This can be reproduced by the following methods:
1.add an mdelay(1000) before acquiring the cgroup_lock In the
cgroup_path_ns function.
2.$cat /proc/<pid>/cpuset repeatly.
3.$mount -t cgroup -o cpuset cpuset /sys/fs/cgroup/cpuset/
$umount /sys/fs/cgroup/cpuset/ repeatly.
The race that cause this bug can be shown as below:
(umount) | (cat /proc/<pid>/cpuset)
css_release | proc_cpuset_show
css_release_work_fn | css = task_get_css(tsk, cpuset_cgrp_id);
css_free_rwork_fn | cgroup_path_ns(css->cgroup, ...);
cgroup_destroy_root | mutex_lock(&cgroup_mutex);
rebind_subsystems |
cgroup_free_root |
| // cgrp was freed, UAF
| cgroup_path_ns_locked(cgrp,..);
When the cpuset is initialized, the root node top_cpuset.css.cgrp
will point to &cgrp_dfl_root.cgrp. In cgroup v1, the mount operation will
allocate cgroup_root, and top_cpuset.css.cgrp will point to the allocated
&cgroup_root.cgrp. When the umount operation is executed,
top_cpuset.css.cgrp will be rebound to &cgrp_dfl_root.cgrp.
The problem is that when rebinding to cgrp_dfl_root, there are cases
where the cgroup_root allocated by setting up the root for cgroup v1
is cached. This could lead to a Use-After-Free (UAF) if it is
subsequently freed. The descendant cgroups of cgroup v1 can only be
freed after the css is released. However, the css of the root will never
be released, yet the cgroup_root should be freed when it is unmounted.
This means that obtaining a reference to the css of the root does
not guarantee that css.cgrp->root will not be freed.
Fix this problem by using rcu_read_lock in proc_cpuset_show().
As cgroup_root is kfree_rcu after commit d23b5c577715
("cgroup: Make operations on the cgroup root_list RCU safe"),
css->cgroup won't be freed during the critical section.
To call cgroup_path_ns_locked, css_set_lock is needed, so it is safe to
replace task_get_css with task_css.
[1] https://syzkaller.appspot.com/bug?extid=9b1ff7be974a403aa4cd
Fixes: a79a908fd2b0 ("cgroup: introduce cgroup namespaces")
Signed-off-by: Chen Ridong <chenridong@huawei.com>
---
kernel/cgroup/cpuset.c | 13 +++++++++----
1 file changed, 9 insertions(+), 4 deletions(-)
diff --git a/kernel/cgroup/cpuset.c b/kernel/cgroup/cpuset.c
index c12b9fdb22a4..bcb4da8f54c8 100644
--- a/kernel/cgroup/cpuset.c
+++ b/kernel/cgroup/cpuset.c
@@ -21,6 +21,7 @@
* License. See the file COPYING in the main directory of the Linux
* distribution for more details.
*/
+#include "cgroup-internal.h"
#include <linux/cpu.h>
#include <linux/cpumask.h>
@@ -5051,10 +5052,14 @@ int proc_cpuset_show(struct seq_file *m, struct pid_namespace *ns,
if (!buf)
goto out;
- css = task_get_css(tsk, cpuset_cgrp_id);
- retval = cgroup_path_ns(css->cgroup, buf, PATH_MAX,
- current->nsproxy->cgroup_ns);
- css_put(css);
+ rcu_read_lock();
+ spin_lock_irq(&css_set_lock);
+ css = task_css(tsk, cpuset_cgrp_id);
+ retval = cgroup_path_ns_locked(css->cgroup, buf, PATH_MAX,
+ current->nsproxy->cgroup_ns);
+ spin_unlock_irq(&css_set_lock);
+ rcu_read_unlock();
+
if (retval == -E2BIG)
retval = -ENAMETOOLONG;
if (retval < 0)
--
2.34.1
^ permalink raw reply [flat|nested] 5+ messages in thread
* Re: [PATCH V4] cgroup/cpuset: Prevent UAF in proc_cpuset_show()
2024-06-28 1:36 [PATCH V4] cgroup/cpuset: Prevent UAF in proc_cpuset_show() Chen Ridong
@ 2024-06-28 17:11 ` Tejun Heo
2024-06-30 9:20 ` Markus Elfring
1 sibling, 0 replies; 5+ messages in thread
From: Tejun Heo @ 2024-06-28 17:11 UTC (permalink / raw)
To: Chen Ridong
Cc: lizefan.x, hannes, longman, adityakali, sergeh, mkoutny, cgroups,
linux-kernel
Hello,
Replaced the v3 patch in cgroup/for-6.10-fixes with this version.
Thanks.
--
tejun
^ permalink raw reply [flat|nested] 5+ messages in thread
* Re: [PATCH V4] cgroup/cpuset: Prevent UAF in proc_cpuset_show()
2024-06-28 1:36 [PATCH V4] cgroup/cpuset: Prevent UAF in proc_cpuset_show() Chen Ridong
2024-06-28 17:11 ` Tejun Heo
@ 2024-06-30 9:20 ` Markus Elfring
2024-06-30 17:03 ` Tejun Heo
1 sibling, 1 reply; 5+ messages in thread
From: Markus Elfring @ 2024-06-30 9:20 UTC (permalink / raw)
To: Chen Ridong, cgroups, Aditya Kali, Johannes Weiner,
Michal Koutný,
Serge Hallyn, Tejun Heo, Waiman Long, Zefan Li
Cc: LKML
…
> +++ b/kernel/cgroup/cpuset.c
…
> @@ -5051,10 +5052,14 @@ int proc_cpuset_show(struct seq_file *m, struct pid_namespace *ns,
> if (!buf)
> goto out;
>
> - css = task_get_css(tsk, cpuset_cgrp_id);
> - retval = cgroup_path_ns(css->cgroup, buf, PATH_MAX,
> - current->nsproxy->cgroup_ns);
> - css_put(css);
> + rcu_read_lock();
> + spin_lock_irq(&css_set_lock);
> + css = task_css(tsk, cpuset_cgrp_id);
> + retval = cgroup_path_ns_locked(css->cgroup, buf, PATH_MAX,
> + current->nsproxy->cgroup_ns);
> + spin_unlock_irq(&css_set_lock);
> + rcu_read_unlock();
…
Under which circumstances would you become interested to apply statements
like the following?
* guard(rcu)();
https://elixir.bootlin.com/linux/v6.10-rc5/source/include/linux/rcupdate.h#L1093
* guard(spinlock_irq)(&css_set_lock);
https://elixir.bootlin.com/linux/v6.10-rc5/source/include/linux/spinlock.h#L567
Regards,
Markus
^ permalink raw reply [flat|nested] 5+ messages in thread
* Re: [PATCH V4] cgroup/cpuset: Prevent UAF in proc_cpuset_show()
2024-06-30 9:20 ` Markus Elfring
@ 2024-06-30 17:03 ` Tejun Heo
2024-06-30 18:00 ` [V4] " Markus Elfring
0 siblings, 1 reply; 5+ messages in thread
From: Tejun Heo @ 2024-06-30 17:03 UTC (permalink / raw)
To: Markus Elfring
Cc: Chen Ridong, cgroups, Aditya Kali, Johannes Weiner,
Michal Koutný,
Serge Hallyn, Waiman Long, Zefan Li, LKML
Hello,
On Sun, Jun 30, 2024 at 11:20:58AM +0200, Markus Elfring wrote:
> Under which circumstances would you become interested to apply statements
> like the following?
>
> * guard(rcu)();
> https://elixir.bootlin.com/linux/v6.10-rc5/source/include/linux/rcupdate.h#L1093
>
> * guard(spinlock_irq)(&css_set_lock);
> https://elixir.bootlin.com/linux/v6.10-rc5/source/include/linux/spinlock.h#L567
I don't really care either way. Neither makes meaningful difference here.
Thanks.
--
tejun
^ permalink raw reply [flat|nested] 5+ messages in thread
* Re: [V4] cgroup/cpuset: Prevent UAF in proc_cpuset_show()
2024-06-30 17:03 ` Tejun Heo
@ 2024-06-30 18:00 ` Markus Elfring
0 siblings, 0 replies; 5+ messages in thread
From: Markus Elfring @ 2024-06-30 18:00 UTC (permalink / raw)
To: Tejun Heo, Chen Ridong, cgroups
Cc: Aditya Kali, Johannes Weiner, Michal Koutný,
Peter Zijlstra, Serge Hallyn, Waiman Long, Zefan Li, LKML
>> Under which circumstances would you become interested to apply statements
>> like the following?
>>
>> * guard(rcu)();
>> https://elixir.bootlin.com/linux/v6.10-rc5/source/include/linux/rcupdate.h#L1093
>>
>> * guard(spinlock_irq)(&css_set_lock);
>> https://elixir.bootlin.com/linux/v6.10-rc5/source/include/linux/spinlock.h#L567
>
> I don't really care either way.
I find such feedback interesting somehow.
> Neither makes meaningful difference here.
Would you like to support making the affected source code safer and a bit more succinct?
https://elixir.bootlin.com/linux/v6.10-rc5/source/kernel/cgroup/cpuset.c#L5034
Regards,
Markus
^ permalink raw reply [flat|nested] 5+ messages in thread
end of thread, other threads:[~2024-06-30 18:01 UTC | newest]
Thread overview: 5+ messages (download: mbox.gz / follow: Atom feed)
-- links below jump to the message on this page --
2024-06-28 1:36 [PATCH V4] cgroup/cpuset: Prevent UAF in proc_cpuset_show() Chen Ridong
2024-06-28 17:11 ` Tejun Heo
2024-06-30 9:20 ` Markus Elfring
2024-06-30 17:03 ` Tejun Heo
2024-06-30 18:00 ` [V4] " Markus Elfring
This is a public inbox, see mirroring instructions
for how to clone and mirror all data and code used for this inbox
all inboxes | Powered by JetHome®