From mboxrd@z Thu Jan 1 00:00:00 1970 Return-Path: Received: (majordomo@vger.kernel.org) by vger.kernel.org via listexpand id S1751718AbdB0UhX (ORCPT ); Mon, 27 Feb 2017 15:37:23 -0500 Received: from mail-yw0-f194.google.com ([209.85.161.194]:34020 "EHLO mail-yw0-f194.google.com" rhost-flags-OK-OK-OK-OK) by vger.kernel.org with ESMTP id S1751679AbdB0UhV (ORCPT ); Mon, 27 Feb 2017 15:37:21 -0500 Date: Mon, 27 Feb 2017 15:29:06 -0500 From: Tejun Heo To: Tahsin Erdogan Cc: Michal Hocko , Christoph Lameter , Andrew Morton , Chris Wilson , Andrey Ryabinin , Roman Pen , Joonas Lahtinen , zijun_hu , Joonsoo Kim , David Rientjes , linux-mm@kvack.org, linux-kernel@vger.kernel.org Subject: Re: [PATCH v2 3/3] percpu: improve allocation success rate for non-GFP_KERNEL callers Message-ID: <20170227202906.GF8707@htj.duckdns.org> References: <201702260805.zhem8KFI%fengguang.wu@intel.com> <20170226043829.14270-1-tahsin@google.com> <20170227095258.GG14029@dhcp22.suse.cz> <20170227195126.GC8707@htj.duckdns.org> MIME-Version: 1.0 Content-Type: text/plain; charset=us-ascii Content-Disposition: inline In-Reply-To: User-Agent: Mutt/1.7.1 (2016-10-04) Sender: linux-kernel-owner@vger.kernel.org List-ID: X-Mailing-List: linux-kernel@vger.kernel.org Hello, On Mon, Feb 27, 2017 at 12:27:08PM -0800, Tahsin Erdogan wrote: > A better example is the call path below: > > pcpu_alloc+0x68f/0x710 > __alloc_percpu_gfp+0xd/0x10 > __percpu_counter_init+0x55/0xc0 > cfq_pd_alloc+0x3b2/0x4e0 > blkg_alloc+0x187/0x230 > blkg_create+0x489/0x670 > blkg_lookup_create+0x9a/0x230 > blkg_conf_prep+0x1fb/0x240 > __cfqg_set_weight_device.isra.105+0x5c/0x180 > cfq_set_weight_on_dfl+0x69/0xc0 > cgroup_file_write+0x39/0x1c0 > kernfs_fop_write+0x13f/0x1d0 > __vfs_write+0x23/0x120 > vfs_write+0xc2/0x1f0 > SyS_write+0x44/0xb0 > entry_SYSCALL_64_fastpath+0x18/0xad > > A failure in this call path gives grief to tools which are trying to > configure io > weights. We see occasional failures happen here shortly after reboots even > when system is not under any memory pressure. Machines with a lot of cpus > are obviously more vulnerable. Ah, absolutely, that's a stupid failure but we should be able to fix that by making the blkg functions take gfp mask and allocate accordingly, right? It'll probably take preallocation tricks because of locking but should be doable. Thanks. -- tejun