From mboxrd@z Thu Jan 1 00:00:00 1970 Return-Path: Received: (majordomo@vger.kernel.org) by vger.kernel.org via listexpand id S1753990AbdIGG7p (ORCPT ); Thu, 7 Sep 2017 02:59:45 -0400 Received: from Galois.linutronix.de ([146.0.238.70]:50384 "EHLO Galois.linutronix.de" rhost-flags-OK-OK-OK-OK) by vger.kernel.org with ESMTP id S1751398AbdIGG7n (ORCPT ); Thu, 7 Sep 2017 02:59:43 -0400 Date: Thu, 7 Sep 2017 08:59:34 +0200 (CEST) From: Thomas Gleixner To: Dan Williams cc: Christoph Hellwig , Yu Chen , X86 ML , Ingo Molnar , "H. Peter Anvin" , Rui Zhang , LKML , "Rafael J. Wysocki" , Len Brown , Peter Zijlstra , Jeff Kirsher Subject: Re: [PATCH 4/4][RFC v2] x86/apic: Spread the vectors by choosing the idlest CPU In-Reply-To: Message-ID: References: <20170906041337.GC23250@localhost.localdomain> <20170906061545.GA20519@lst.de> User-Agent: Alpine 2.20 (DEB 67 2015-01-07) MIME-Version: 1.0 Content-Type: text/plain; charset=US-ASCII Sender: linux-kernel-owner@vger.kernel.org List-ID: X-Mailing-List: linux-kernel@vger.kernel.org On Wed, 6 Sep 2017, Dan Williams wrote: > On Wed, Sep 6, 2017 at 10:59 PM, Thomas Gleixner wrote: > >> commit 7c9ae7f053e9e896c24fd23595ba369a5fe322e1 > > > > -ENOSUCHCOMMIT > > Sorry, that's still pending in -next. Ok. > >> Author: Carolyn Wyborny > >> Date: Tue Jun 20 15:16:53 2017 -0700 > >> > >> i40e: Fix for trace found with S4 state > >> > >> This patch fixes a problem found in systems when entering > >> S4 state. This patch fixes the problem by ensuring that > >> the misc vector's IRQ is disabled as well. Without this > >> patch a stack trace can be seen upon entering S4 state. Btw. This changelog is pretty useless..... > >> However this seems like something that should be handled generically > >> in the irq-core especially since commit c5cb83bb337c > >> "genirq/cpuhotplug: Handle managed IRQs on CPU hotplug" was headed in > >> that direction. It's otherwise non-obvious when a driver needs to > >> release and re-acquire interrupts or be reworked to use managed > >> interrupts. > > > > There are two problems here: > > > > 1) The driver allocates 300 interrupts and uses exactly 8 randomly chosen > > ones. > > > > 2) It's not using the managed affinity mechanics, so the interrupts cannot > > be sanely handled by the kernel, neither affinity wise nor at hotplug > > time. > > Ok, this driver is an obvious candidate, but is there a general > guideline of when a driver must use affinity management? Should we be > emitting a message when a driver exceeds a certain threshold of > unmanaged interrupts to flag this in the future? I guess everything which uses multi queues and therefor allocates 8+ vectors is something which falls into that category. Aside of that, drivers should be sane in terms of allocations. Allocating metric tons of interrupts for nothing is not really a sign of a proper thought out resource management. Thanks, tglx