From: Zhang Rui <rui.zhang@intel.com>
To: Feng Tang <feng.tang@intel.com>,
Thomas Gleixner <tglx@linutronix.de>,
Ingo Molnar <mingo@redhat.com>, Borislav Petkov <bp@alien8.de>,
Dave Hansen <dave.hansen@intel.com>,
"H . Peter Anvin" <hpa@zytor.com>,
Peter Zijlstra <peterz@infradead.org>,
x86@kernel.org, linux-kernel@vger.kernel.org
Cc: tim.c.chen@intel.com, Xiongfeng Wang <wangxiongfeng2@huawei.com>,
liaoyu15@huawei.com
Subject: Re: [PATCH v1 1/2] x86/tsc: use logical_package as a better estimation of socket numbers
Date: Fri, 21 Oct 2022 23:00:44 +0800 [thread overview]
Message-ID: <f27e4b3f858890c657df9a7d6f34dc2d60b89757.camel@intel.com> (raw)
In-Reply-To: <20221021062131.1826810-1-feng.tang@intel.com>
On Fri, 2022-10-21 at 14:21 +0800, Feng Tang wrote:
> Commit b50db7095fe0 ("x86/tsc: Disable clocksource watchdog for TSC
> on qualified platorms") was introduced to solve problem that
> sometimes TSC clocksource is wrongly judged as unstable by watchdog
> like 'jiffies', HPET, etc.
>
> In it, the hardware socket number is a key factor for judging
> whether to disable the watchdog for TSC, and 'nr_online_nodes' was
> chosen as an estimation due to it is needed in early boot phase
> before registering 'tsc-early' clocksource, where all none-boot
> CPUs are not brought up yet.
>
> In recent patch review, Dave Hansen pointed out there are many
> cases that 'nr_online_nodes' could have issue, like:
> * numa emulation (numa=fake=4 etc.)
> * numa=off
> * platforms with CPU+DRAM nodes, CPU-less HBM nodes, CPU-less
> persistent memory nodes.
> * SNC (sub-numa cluster) mode is enabled
>
> Peter Zijlstra suggested to use logical package ids, but it is
> only usable after smp_init() and all CPUs are initialized.
>
> One solution is to skip the watchdog for 'tsc-early' clocksource,
> and move the check after smp_init(), while before 'tsc'
> clocksoure is registered, where 'logical_packages' could be used
> as a much more accurate socket number.
>
> Signed-off-by: Feng Tang <feng.tang@intel.com>
> ---
> Hi reviewers,
>
> I separate the code to 2 patches, as I think they are covering 2
> problems and easy for bisect. Feel free to combine them into one,
> as the 2/2 are a trivial change.
>
> Thanks,
> Feng
>
> Changelog:
>
> Since RFC:
> * use 'logical_packages' instead of topology_max_packages(), whose
> implementaion is not accurate, like for heterogeneous systems
> which have combination of Core/Atom CPUs like Alderlake (Dave
> Hansen)
I checked the history of '__max_logical_packages', and realized that
1. for topology_max_packages()/'__max_logical_packages', the divisor
'ncpus' uses cpu_data(0).booted_cores, which is based on the
*online* CPUs. So when using kernel cmdlines like maxcpus=/nr_cpus=,
'__max_logical_packages' can get over-estimated.
2. for 'logical_packages', it equals the number of different physical
Package IDs for all *online* CPUs. So with kernel cmdlines like
nr_cpus=/maxcpus=, it can gets under-estimated.
BTW, I also checked CPUID.B/1F, which can tell a fixed number of CPUs
within a package. But we don't have a fixed number of total CPUs from
hardware.
On my Dell laptop, BIOS allows me to disable/enable one or several
cores. When this happens, the 'total_cpus' changes, but CPUID.B/1F does
not change. So I don't think CPUID.B/1F can be used to optimize the '__
max_logical_packages' calculation.
I'm not sure if we have a perfect solution here.
thanks,
rui
next prev parent reply other threads:[~2022-10-21 15:01 UTC|newest]
Thread overview: 13+ messages / expand[flat|nested] mbox.gz Atom feed top
2022-10-21 6:21 Feng Tang
2022-10-21 6:21 ` [PATCH v1 2/2] x86/tsc: Extend watchdog check exemption to 4-Sockets platform Feng Tang
2023-06-02 18:00 ` Paul E. McKenney
2023-06-05 6:28 ` Feng Tang
2022-10-21 15:00 ` Zhang Rui [this message]
2022-10-21 16:21 ` [PATCH v1 1/2] x86/tsc: use logical_package as a better estimation of socket numbers Dave Hansen
2022-10-22 16:12 ` Zhang Rui
2022-10-24 15:42 ` Dave Hansen
2022-10-25 7:35 ` Feng Tang
2022-10-24 7:37 ` Feng Tang
2022-10-24 15:43 ` Dave Hansen
2022-10-25 7:57 ` Feng Tang
2022-11-04 7:21 ` Feng Tang
Reply instructions:
You may reply publicly to this message via plain-text email
using any one of the following methods:
* Save the following mbox file, import it into your mail client,
and reply-to-all from there: mbox
Avoid top-posting and favor interleaved quoting:
https://en.wikipedia.org/wiki/Posting_style#Interleaved_style
* Reply using the --to, --cc, and --in-reply-to
switches of git-send-email(1):
git send-email \
--in-reply-to=f27e4b3f858890c657df9a7d6f34dc2d60b89757.camel@intel.com \
--to=rui.zhang@intel.com \
--cc=bp@alien8.de \
--cc=dave.hansen@intel.com \
--cc=feng.tang@intel.com \
--cc=hpa@zytor.com \
--cc=liaoyu15@huawei.com \
--cc=linux-kernel@vger.kernel.org \
--cc=mingo@redhat.com \
--cc=peterz@infradead.org \
--cc=tglx@linutronix.de \
--cc=tim.c.chen@intel.com \
--cc=wangxiongfeng2@huawei.com \
--cc=x86@kernel.org \
/path/to/YOUR_REPLY
https://kernel.org/pub/software/scm/git/docs/git-send-email.html
* If your mail client supports setting the In-Reply-To header
via mailto: links, try the mailto: link
Be sure your reply has a Subject: header at the top and a blank line
before the message body.
This is a public inbox, see mirroring instructions
for how to clone and mirror all data and code used for this inbox
Powered by JetHome