From: John Garry <john.garry@huawei.com>
To: Marc Zyngier <maz@kernel.org>
Cc: Thomas Gleixner <tglx@linutronix.de>,
Zhou Wang <wangzhou1@hisilicon.com>,
"linux-kernel@vger.kernel.org" <linux-kernel@vger.kernel.org>
Subject: Re: PCI MSI issue with reinserting a driver
Date: Tue, 2 Feb 2021 08:37:39 +0000 [thread overview]
Message-ID: <a80b9be0-c455-c852-ddac-3f514a15e896@huawei.com> (raw)
In-Reply-To: <87o8h3lj0n.wl-maz@kernel.org>
On 01/02/2021 18:50, Marc Zyngier wrote:
Hi Marc,
>> Just a heads-up, by chance I noticed that I can't re-insert a specific
>> driver on v5.11-rc6:
>>
>> [ 64.356023] hisi_dma 0000:7b:00.0: Adding to iommu group 31
>> [ 64.368627] hisi_dma 0000:7b:00.0: enabling device (0000 -> 0002)
>> [ 64.384156] hisi_dma 0000:7b:00.0: Failed to allocate MSI vectors!
>> [ 64.397180] hisi_dma: probe of 0000:7b:00.0 failed with error -28
>>
>> That's with CONFIG_DEBUG_TEST_DRIVER_REMOVE=y
>>
>> Bisect tells me that this is the first bad commit:
>> 4615fbc3788d genirq/irqdomain: Don't try to free an interrupt that has
>> no mapping
>>
>> The relevant driver code is
>> https://git.kernel.org/pub/scm/linux/kernel/git/torvalds/linux.git/tree/drivers/dma/hisi_dma.c#n547
>>
>> That driver only allocates 30 MSI, so maybe there's a problem with not
>> allocating (and freeing) all 32 MSI.
> Are they Multi-MSI (and not MSI-X)?
multi-msi
>
>> I'll have a bit more of a look tomorrow.
> Here's my suspicion: two of the interrupts are mapped in the low-level
> domain (the ITS, I'd expect in your case), but they have never been
> mapped at the higher level.
>
> On teardown, we only get rid of the 30 that were actually mapped, and
> leave the last two dangling in the ITS domain, and thus the ITS device
> resources are never freed. On reload, we request another 32
> interrupts, which can't be satisfied for this device.
>
> Assuming I got it right, the question is: why weren't these interrupts
> mapped in the PCI domain the first place. And if I got it wrong, I'm
> even more curious!
Not sure. I also now notice an error for the SAS PCI driver on D06 when
nr_cpus < 16, which means number of MSI vectors allocated < 32, so looks
the same problem. There we try to allocate 16 + max(nr cpus, 16) MSI.
Anyway, let me have a look today to see what is going wrong.
cheers,
John
next prev parent reply other threads:[~2021-02-02 8:40 UTC|newest]
Thread overview: 12+ messages / expand[flat|nested] mbox.gz Atom feed top
2021-02-01 18:34 John Garry
2021-02-01 18:50 ` Marc Zyngier
2021-02-02 8:37 ` John Garry [this message]
2021-02-02 12:38 ` John Garry
2021-02-02 14:48 ` Marc Zyngier
2021-02-02 15:46 ` John Garry
2021-02-03 17:23 ` Marc Zyngier
2021-02-04 10:45 ` John Garry
2022-08-04 10:59 ` John Garry
2021-04-06 9:46 ` John Garry
2021-08-27 8:33 ` luojiaxing
2023-08-29 23:00 ` Thomas Gleixner
Reply instructions:
You may reply publicly to this message via plain-text email
using any one of the following methods:
* Save the following mbox file, import it into your mail client,
and reply-to-all from there: mbox
Avoid top-posting and favor interleaved quoting:
https://en.wikipedia.org/wiki/Posting_style#Interleaved_style
* Reply using the --to, --cc, and --in-reply-to
switches of git-send-email(1):
git send-email \
--in-reply-to=a80b9be0-c455-c852-ddac-3f514a15e896@huawei.com \
--to=john.garry@huawei.com \
--cc=linux-kernel@vger.kernel.org \
--cc=maz@kernel.org \
--cc=tglx@linutronix.de \
--cc=wangzhou1@hisilicon.com \
/path/to/YOUR_REPLY
https://kernel.org/pub/software/scm/git/docs/git-send-email.html
* If your mail client supports setting the In-Reply-To header
via mailto: links, try the mailto: link
Be sure your reply has a Subject: header at the top and a blank line
before the message body.
This is a public inbox, see mirroring instructions
for how to clone and mirror all data and code used for this inbox
all inboxes | Powered by JetHome®