From: Paul Turner <pjt@google.com>
To: Peter Zijlstra <peterz@infradead.org>
Cc: mingo@elte.hu, linux-kernel@vger.kernel.org
Subject: Re: sched: race between deactivate and switch sched_info accounting?
Date: Mon, 12 Oct 2009 14:20:53 -0700 [thread overview]
Message-ID: <ed628a920910121420k5f01892dsc9cfe0b56eb2c29e@mail.gmail.com> (raw)
In-Reply-To: <1255156660.7866.4.camel@twins>
On Fri, Oct 9, 2009 at 11:37 PM, Peter Zijlstra <peterz@infradead.org> wrote:
> On Fri, 2009-10-09 at 19:40 -0700, Paul Turner wrote:
>
> This looks very funny, I would expect that whoever does activate() on
> that task to do the sched_info*() muck?
They do their part, the problem stems from there being 2 halves of
sched_info muck to deal with:
then enqeue/dequeue tracking (which is called out of the activate path)
then the arrive/depart (which is hinged off the task switch in schedule())
These steps share state in last_queued; we are accounting the
enqeue/dequeue properly but because prev==next we skip the second half
of the accounting and end up out of sync.
>
> The below patch looks very asymmetric in that regard.
>
the idea below is to fix things up by 'inserting' the task->idle->task
switch we'd have had if we didn't race in wake-up
this approach might look more natural if the first info_switch was
just pulled above the actual switch logic with this check; another
option might be to force a last_queued reset on depart (since it's
really only the run_delay skew that's significant)
>> It's possible for our previously de-activated task to be re-activated by a
>> remote cpu during lock balancing. We have to account for this manually
>> since prev == next, yet the task just went through dequeue accounting.
>>
>> Signed-off-by: Paul Turner <pjt@google.com>
>> ---
>> kernel/sched.c | 15 ++++++++++++---
>> 1 files changed, 12 insertions(+), 3 deletions(-)
>>
>> diff --git a/kernel/sched.c b/kernel/sched.c
>> index ee61f45..6445d9d 100644
>> --- a/kernel/sched.c
>> +++ b/kernel/sched.c
>> @@ -5381,7 +5381,7 @@ asmlinkage void __sched schedule(void)
>> struct task_struct *prev, *next;
>> unsigned long *switch_count;
>> struct rq *rq;
>> - int cpu;
>> + int cpu, deactivated_prev = 0;
>>
>> need_resched:
>> preempt_disable();
>> @@ -5406,8 +5406,10 @@ need_resched_nonpreemptible:
>> if (prev->state && !(preempt_count() & PREEMPT_ACTIVE)) {
>> if (unlikely(signal_pending_state(prev->state, prev)))
>> prev->state = TASK_RUNNING;
>> - else
>> + else {
>> deactivate_task(rq, prev, 1);
>> + deactivated_prev = 1;
>> + }
>> switch_count = &prev->nvcsw;
>> }
>>
>> @@ -5434,8 +5436,15 @@ need_resched_nonpreemptible:
>> */
>> cpu = smp_processor_id();
>> rq = cpu_rq(cpu);
>> - } else
>> + } else {
>> + /*
>> + * account for our previous task being re-activated by a
>> + * remote cpu.
>> + */
>> + if (unlikely(deactivated_prev))
>> + sched_info_switch(prev, prev);
>> spin_unlock_irq(&rq->lock);
>> + }
>>
>> post_schedule(rq);
>>
>
prev parent reply other threads:[~2009-10-12 21:22 UTC|newest]
Thread overview: 3+ messages / expand[flat|nested] mbox.gz Atom feed top
2009-10-10 2:40 Paul Turner
2009-10-10 6:37 ` Peter Zijlstra
2009-10-12 21:20 ` Paul Turner [this message]
Reply instructions:
You may reply publicly to this message via plain-text email
using any one of the following methods:
* Save the following mbox file, import it into your mail client,
and reply-to-all from there: mbox
Avoid top-posting and favor interleaved quoting:
https://en.wikipedia.org/wiki/Posting_style#Interleaved_style
* Reply using the --to, --cc, and --in-reply-to
switches of git-send-email(1):
git send-email \
--in-reply-to=ed628a920910121420k5f01892dsc9cfe0b56eb2c29e@mail.gmail.com \
--to=pjt@google.com \
--cc=linux-kernel@vger.kernel.org \
--cc=mingo@elte.hu \
--cc=peterz@infradead.org \
/path/to/YOUR_REPLY
https://kernel.org/pub/software/scm/git/docs/git-send-email.html
* If your mail client supports setting the In-Reply-To header
via mailto: links, try the mailto: link
Be sure your reply has a Subject: header at the top and a blank line
before the message body.
This is a public inbox, see mirroring instructions
for how to clone and mirror all data and code used for this inbox
all inboxes | Powered by JetHome®