From: ebiederm@xmission.com (Eric W. Biederman)
To: vgoyal@in.ibm.com
Cc: Andi Kleen <ak@muc.de>,
linux kernel mailing list <linux-kernel@vger.kernel.org>,
Fastboot mailing list <fastboot@lists.osdl.org>,
Andrew Morton <akpm@osdl.org>
Subject: Re: [RFC][PATCH] kdump: x86_64 timer interrupt lockup due to pending interrupt
Date: Tue, 07 Mar 2006 16:43:07 -0700 [thread overview]
Message-ID: <m1hd69zur8.fsf@ebiederm.dsl.xmission.com> (raw)
In-Reply-To: <20060307222052.GD9106@in.ibm.com> (Vivek Goyal's message of "Tue, 7 Mar 2006 17:20:52 -0500")
Vivek Goyal <vgoyal@in.ibm.com> writes:
> On Mon, Mar 06, 2006 at 10:43:32PM +0100, Andi Kleen wrote:
>> On Mon, Mar 06, 2006 at 11:40:34AM -0500, Vivek Goyal wrote:
>> >
>> > o check_timer() routine fails while second kernel is booting after a crash
>> > on an opetron box. Problem happens because timer vector (0x31) seems to be
>> > locked.
>> >
>> > o After a system crash, it is not safe to service interrupts any more, hence
>> > interrupts are disabled. This leads to pending interrupts at LAPIC. LAPIC
>> > sends these interrupts to the CPU during early boot of second kernel. Other
>> > pending interrupts are discarded saying unexpected trap but timer interrupt
>> > is serviced and CPU does not issue an LAPIC EOI because it think this
>> > interrupt came from i8259 and sends ack to 8259. This leads to vector 0x31
>> > locking as LAPIC does not clear respective ISR and keeps on waiting for
>> > EOI.
>> >
>> > o In this patch, one extra EOI is being issued in check_timer() to unlock
> the
>> > vector. Please suggest if there is a better way to handle this situation.
>>
>> Shouldn't we rather do this for all interrupts when the APIC is set up?
>> I don't see how the timer is special here.
>>
>
> Timer is a special case here.
>
> In other cases, the moment interrupts are enabled on cpu, LAPIC pushes pending
> interrupts to cpu and it is ignored as bad irq using ack_bad_irq(). This
> still sends EOI to LAPIC if LPAIC support is compiled in.
>
> But for timer, the moment pending interrupt is pushed to cpu, it is handled
> as valid interrupt and cpu assumes that it came from 8259 and sends ack to
> 8259 and not to LAPIC. Hence leads to missing EOI for timer vector and
> deadlock.
>
> But still doing it generic manner for all interrupts while setting up LAPIC
> probably makes more sense. Please find attached the patch.
A couple of questions.
Does this need to be in #ifdef CONFIG_CRASS_DUMP?
If this code is truly safe I expect we could run it on every bootup
simply to be more robust.
Why is APIC_ISR_NR a hard code? I think there is an apic register
that tells the count.
Does ack_APIC_irq take an argument? I am confused that we are calling
ack_APIC_irq() potentially 8*32 times without passing it anything.
Eric
next prev parent reply other threads:[~2006-03-07 23:45 UTC|newest]
Thread overview: 7+ messages / expand[flat|nested] mbox.gz Atom feed top
2006-03-06 16:40 Vivek Goyal
2006-03-06 21:43 ` Andi Kleen
2006-03-07 22:20 ` Vivek Goyal
2006-03-07 23:43 ` Eric W. Biederman [this message]
2006-03-08 1:26 ` Vivek Goyal
2006-03-08 4:04 ` Eric W. Biederman
2006-03-08 5:31 ` Andi Kleen
Reply instructions:
You may reply publicly to this message via plain-text email
using any one of the following methods:
* Save the following mbox file, import it into your mail client,
and reply-to-all from there: mbox
Avoid top-posting and favor interleaved quoting:
https://en.wikipedia.org/wiki/Posting_style#Interleaved_style
* Reply using the --to, --cc, and --in-reply-to
switches of git-send-email(1):
git send-email \
--in-reply-to=m1hd69zur8.fsf@ebiederm.dsl.xmission.com \
--to=ebiederm@xmission.com \
--cc=ak@muc.de \
--cc=akpm@osdl.org \
--cc=fastboot@lists.osdl.org \
--cc=linux-kernel@vger.kernel.org \
--cc=vgoyal@in.ibm.com \
/path/to/YOUR_REPLY
https://kernel.org/pub/software/scm/git/docs/git-send-email.html
* If your mail client supports setting the In-Reply-To header
via mailto: links, try the mailto: link
Be sure your reply has a Subject: header at the top and a blank line
before the message body.
This is a public inbox, see mirroring instructions
for how to clone and mirror all data and code used for this inbox
Powered by JetHome