From mboxrd@z Thu Jan 1 00:00:00 1970 Return-Path: Received: (majordomo@vger.kernel.org) by vger.kernel.org via listexpand id S1754023AbaG2P1M (ORCPT ); Tue, 29 Jul 2014 11:27:12 -0400 Received: from casper.infradead.org ([85.118.1.10]:38168 "EHLO casper.infradead.org" rhost-flags-OK-OK-OK-OK) by vger.kernel.org with ESMTP id S1753672AbaG2P1K (ORCPT ); Tue, 29 Jul 2014 11:27:10 -0400 Date: Tue, 29 Jul 2014 17:27:05 +0200 From: Peter Zijlstra To: Vincent Guittot Cc: Rik van Riel , linux-kernel , Michael Neuling , Ingo Molnar , jhladky@redhat.com, ktkhai@parallels.com, tim.c.chen@linux.intel.com, Nicolas Pitre Subject: Re: [PATCH 1/2] sched: fix and clean up calculate_imbalance Message-ID: <20140729152705.GX12054@laptop.lan> References: <1406571388-3227-1-git-send-email-riel@redhat.com> <1406571388-3227-2-git-send-email-riel@redhat.com> <20140729145910.GH3935@laptop> MIME-Version: 1.0 Content-Type: text/plain; charset=us-ascii Content-Disposition: inline In-Reply-To: <20140729145910.GH3935@laptop> User-Agent: Mutt/1.5.21 (2012-12-30) Sender: linux-kernel-owner@vger.kernel.org List-ID: X-Mailing-List: linux-kernel@vger.kernel.org On Tue, Jul 29, 2014 at 04:59:10PM +0200, Peter Zijlstra wrote: > On Tue, Jul 29, 2014 at 11:04:50AM +0200, Vincent Guittot wrote: > > > In situations where all the domains are overloaded, or where only the > > > busiest domain is overloaded, that code is also superfluous, since > > > the normal env->imbalance calculation will figure out how much to move. > > > Remove the load_above_capacity calculation. > > > > IMHO, we should not remove that part which is used by prefer_sibling > > > > Originally, we had 2 type of busiest group: overloaded or imbalanced. > > You add a new one which has only a avg_load higher than other so you > > should handle this new case and keep the other ones unchanged > > Right, so we want that code for overloaded -> overloaded migrations such > as not to cause idle cpus in an attempt to balance things. Idle cpus are > worse than imbalance. > > But in case of overloaded/imb -> !overloaded migrations we can allow it, > and in fact want to allow it in order to balance idle cpus. Which would be patch 3/2 --- Subject: sched,fair: Allow calculate_imbalance() to move idle cpus From: Peter Zijlstra Date: Tue Jul 29 17:15:11 CEST 2014 Allow calculate_imbalance() to 'create' idle cpus in the busiest group if there are idle cpus in the local group. Suggested-by: Rik van Riel Signed-off-by: Peter Zijlstra Link: http://lkml.kernel.org/n/tip-7k95k4i2tjv78iivstggiude@git.kernel.org --- kernel/sched/fair.c | 11 +++++------ 1 file changed, 5 insertions(+), 6 deletions(-) --- a/kernel/sched/fair.c +++ b/kernel/sched/fair.c @@ -6273,12 +6273,11 @@ static inline void calculate_imbalance(s return fix_small_imbalance(env, sds); } - if (busiest->group_type == group_overloaded) { - /* - * Don't want to pull so many tasks that a group would go idle. - * Except of course for the group_imb case, since then we might - * have to drop below capacity to reach cpu-load equilibrium. - */ + /* + * If there aren't any idle cpus, avoid creating some. + */ + if (busiest->group_type == group_overloaded && + local->group_type == group_overloaded) { load_above_capacity = (busiest->sum_nr_running - busiest->group_capacity_factor);