mirror of https://lore.kernel.org/lkml/
 help / color / mirror / Atom feed
From: "Christian König" <christian.koenig@amd.com>
To: Leon Romanovsky <leon@kernel.org>
Cc: "Thomas Hellström" <thomas.hellstrom@linux.intel.com>,
	"Bjorn Helgaas" <bhelgaas@google.com>,
	"Logan Gunthorpe" <logang@deltatee.com>,
	"Jason Gunthorpe" <jgg@ziepe.ca>,
	"Joerg Roedel (AMD)" <joro@8bytes.org>,
	"Will Deacon" <will@kernel.org>,
	"Robin Murphy" <robin.murphy@arm.com>,
	linux-pci@vger.kernel.org, linux-kernel@vger.kernel.org,
	linux-doc@vger.kernel.org, iommu@lists.linux.dev,
	"Tushar Dave" <tdave@nvidia.com>,
	linux-media@vger.kernel.org, dri-devel@lists.freedesktop.org,
	linaro-mm-sig@lists.linaro.org, linux-rdma@vger.kernel.org,
	kvm@vger.kernel.org, "Chaitanya Kulkarni" <kch@nvidia.com>,
	"Greg Kroah-Hartman" <gregkh@linuxfoundation.org>,
	"Jens Axboe" <axboe@kernel.dk>,
	"Alex Williamson" <alex@shazbot.org>,
	"Ankit Agrawal" <ankita@nvidia.com>,
	"Jonathan Corbet" <corbet@lwn.net>,
	"Shuah Khan" <skhan@linuxfoundation.org>,
	"Randy Dunlap" <rdunlap@infradead.org>,
	"Sumit Semwal" <sumit.semwal@linaro.org>
Subject: Re: [PATCH v8 16/23] dma-buf: Let importers ask how peer-to-peer traffic is routed
Date: Thu, 1 Oct 2026 11:07:34 +0200	[thread overview]
Message-ID: <0c49c6d0-8301-4ded-accc-eaa5cf91b066@amd.com> (raw)
In-Reply-To: <20261001082710.GL3401365@unreal>

On 10/1/26 10:27, Leon Romanovsky wrote:
> On Thu, Oct 01, 2026 at 09:49:31AM +0200, Christian König wrote:
>> On 10/1/26 08:37, Leon Romanovsky wrote:
>>> On Wed, Sep 30, 2026 at 04:57:04PM +0200, Thomas Hellström wrote:
>>>> On Wed, 2026-09-30 at 17:32 +0300, Leon Romanovsky wrote:
>> ...
>>>>
>>>> Returning again to Jason's series. Let's say we'd add just the mapping
>>>> type infrastructure, converted users of pcie_p2pdma only to use that
>>>> and then we'd have access to per-mapping-type data. This could actually
>>>> be done as a prereq for this series and merged separately. It's a
>>>> couple of patches only.
>>>
>>> I afraid that you over optimistic about the amount of work.
>>
>> Yeah, agree. The proposal looked good but there will probably be quite a bunch of work.
>>
>>>>
>>>> Jason's match() and finish() callbacks could compute the interesting
>>>> routes at attach time,  perhaps even condesed to whether IOVA is used
>>>> and whether ATS translated packages have a direct route (which is what
>>>> mlx5 care about AFAICT). This information is kept outside core dma-buf
>>>> and would be specific to the pcie_p2p mapping type (interconnect) only
>>>> rather than having functions and callbacks bloating the core dma-buf
>>>> structures.
>>>>
>>>> Then exactly where the cross-subsystem match() and finish()
>>>> implementations should live I figure remain up for discussion and
>>>> guidance by Christoph?
>>>
>>> Maybe I'm wrong, and everything will work out. However, given Christian's
>>> feedback to determine the mapping type when the mapping is established, this
>>> approach won't work for an RDMA exporter.
>>>
>>> As I mentioned, mlx5 uses on-demand paging (ODP), creating mappings in
>>> response to page faults. To handle these faults with reasonable performance,
>>> mlx5 must create a memory region (MR), with or without ATS, and it is
>>> needed to be created before first page fault.
>>
>> Yeah, I feared that you have something like that. At least for the current DMA-buf semantics that is not something which fits into that model.
>>
>> Background is that system memory is usually seen as fallback which should always work and you can have really strange combination of use cases.
>>
>> So what can happen is that you create a mapping and P2P is possible without ATS but then some other device attaches and we suddenly have to use ATS because the buffer is now in system memory.
> 
> The importer reports ATS support on a per-mapping basis. If the exporter
> receives this hint from deviceA but not deviceB, it prepares different
> addresses for the two mappings.
> 
> The `ats_per_mapping` flag is stored in the importer structure:
> https://lore.kernel.org/linux-rdma/20260928-fix-p2p-acs-v4-0-v8-20-404453b9c435@nvidia.com/
> 
> In our example, if the mlx5 and XE importers both attach to the same
> exporter, they receive different mappings.

That isn't sufficient.

It is perfectly possible that P2P is disabled later on because of another device attaching or simply resource constrains.

So that your initial mapping has ATS enabled and then you get a mapping with ATS disabled is perfectly possible and even trivially trigger able through uAPI with some exporters.

Supporting that is a must have, even if it's slow. The only alternative I can see is to pin things but as I said that also has some other down sides.

Regards,
Christian.


> 
> Thanks


  reply	other threads:[~2026-10-01  9:07 UTC|newest]

Thread overview: 48+ messages / expand[flat|nested]  mbox.gz  Atom feed  top
2026-09-28 11:19 [PATCH v8 00/23] PCI/P2PDMA: Route peer-to-peer DMA by TLP class Leon Romanovsky
2026-09-28 11:19 ` [PATCH v8 01/23] PCI/P2PDMA: Document the TLP attribute assumptions Leon Romanovsky
2026-09-28 11:19 ` [PATCH v8 02/23] PCI/P2PDMA: Derive routing from directional ACS controls Leon Romanovsky
2026-09-28 11:19 ` [PATCH v8 03/23] PCI: Reject unreadable ACS controls in isolation checks Leon Romanovsky
2026-09-28 11:19 ` [PATCH v8 04/23] PCI/P2PDMA: Evaluate ACS controls at the path divergence Leon Romanovsky
2026-09-28 11:19 ` [PATCH v8 05/23] PCI/P2PDMA: Document directional ACS routing Leon Romanovsky
2026-09-28 11:19 ` [PATCH v8 06/23] PCI/P2PDMA: Collect the path's ACS controls before deciding Leon Romanovsky
2026-09-28 11:19 ` [PATCH v8 07/23] PCI/P2PDMA: Answer routing per TLP class Leon Romanovsky
2026-09-28 11:19 ` [PATCH v8 08/23] PCI/P2PDMA: Route Relaxed Ordering Completions directly Leon Romanovsky
2026-09-28 11:19 ` [PATCH v8 09/23] PCI/P2PDMA: Reject Translated Requests blocked by Translation Blocking Leon Romanovsky
2026-09-28 11:19 ` [PATCH v8 10/23] PCI/P2PDMA: Route Translated Requests under Direct Translated P2P Leon Romanovsky
2026-09-28 11:19 ` [PATCH v8 11/23] PCI/P2PDMA: Log detailed ACS routing diagnostics Leon Romanovsky
2026-09-28 11:19 ` [PATCH v8 12/23] PCI/P2PDMA: Add KUnit tests for the ACS routing decisions Leon Romanovsky
2026-09-28 11:19 ` [PATCH v8 13/23] PCI/P2PDMA: Test the ACS P2P routing walk Leon Romanovsky
2026-09-28 11:19 ` [PATCH v8 14/23] PCI: Add KUnit coverage for ACS isolation checks Leon Romanovsky
2026-09-28 11:19 ` [PATCH v8 15/23] PCI/P2PDMA: Document TLP-class routing Leon Romanovsky
2026-09-28 11:19 ` [PATCH v8 16/23] dma-buf: Let importers ask how peer-to-peer traffic is routed Leon Romanovsky
2026-09-29  9:34   ` Christian König
2026-09-29 13:23     ` Leon Romanovsky
2026-09-29 13:39       ` Christian König
2026-09-29 17:57         ` Leon Romanovsky
2026-09-30  7:00           ` Christian König
2026-09-30  8:18             ` Leon Romanovsky
2026-09-30  8:34               ` Christian König
2026-09-30 11:43                 ` Leon Romanovsky
2026-09-30 11:52                   ` Christian König
2026-09-30 14:32                     ` Leon Romanovsky
2026-09-30 14:52                       ` Christian König
2026-10-01  6:17                         ` Leon Romanovsky
2026-10-01  7:06                           ` Christian König
2026-09-30 14:57                       ` Thomas Hellström
2026-09-30 15:03                         ` Christian König
2026-09-30 16:52                         ` Jason Gunthorpe
2026-10-02  8:19                           ` Alistair Popple
2026-10-01  6:37                         ` Leon Romanovsky
2026-10-01  7:49                           ` Christian König
2026-10-01  8:27                             ` Leon Romanovsky
2026-10-01  9:07                               ` Christian König [this message]
2026-10-01  8:11                           ` Thomas Hellström
2026-09-28 11:19 ` [PATCH v8 17/23] vfio/pci: Hand out the P2PDMA provider behind a dma-buf Leon Romanovsky
2026-09-28 11:19 ` [PATCH v8 18/23] RDMA/uverbs: " Leon Romanovsky
2026-09-28 11:19 ` [PATCH v8 19/23] RDMA/mlx5: Ask P2PDMA whether ATS takes a direct peer-to-peer route Leon Romanovsky
2026-09-28 11:19 ` [PATCH v8 20/23] PCI/P2PDMA: Let a client declare that it selects ATS per mapping Leon Romanovsky
2026-09-28 12:44   ` Thomas Hellström
2026-09-28 11:19 ` [PATCH v8 21/23] RDMA/mlx5: Declare to P2PDMA that ATS is selected per memory key Leon Romanovsky
2026-09-28 11:19 ` [PATCH v8 22/23] PCI/P2PDMA: Evaluate the ATS path for clients with ATS enabled Leon Romanovsky
2026-09-28 12:43   ` Thomas Hellström
2026-09-28 11:19 ` [PATCH v8 23/23] PCI/P2PDMA: Test the routing of " Leon Romanovsky

Reply instructions:

You may reply publicly to this message via plain-text email
using any one of the following methods:

* Save the following mbox file, import it into your mail client,
  and reply-to-all from there: mbox

  Avoid top-posting and favor interleaved quoting:
  https://en.wikipedia.org/wiki/Posting_style#Interleaved_style

* Reply using the --to, --cc, and --in-reply-to
  switches of git-send-email(1):

  git send-email \
    --in-reply-to=0c49c6d0-8301-4ded-accc-eaa5cf91b066@amd.com \
    --to=christian.koenig@amd.com \
    --cc=alex@shazbot.org \
    --cc=ankita@nvidia.com \
    --cc=axboe@kernel.dk \
    --cc=bhelgaas@google.com \
    --cc=corbet@lwn.net \
    --cc=dri-devel@lists.freedesktop.org \
    --cc=gregkh@linuxfoundation.org \
    --cc=iommu@lists.linux.dev \
    --cc=jgg@ziepe.ca \
    --cc=joro@8bytes.org \
    --cc=kch@nvidia.com \
    --cc=kvm@vger.kernel.org \
    --cc=leon@kernel.org \
    --cc=linaro-mm-sig@lists.linaro.org \
    --cc=linux-doc@vger.kernel.org \
    --cc=linux-kernel@vger.kernel.org \
    --cc=linux-media@vger.kernel.org \
    --cc=linux-pci@vger.kernel.org \
    --cc=linux-rdma@vger.kernel.org \
    --cc=logang@deltatee.com \
    --cc=rdunlap@infradead.org \
    --cc=robin.murphy@arm.com \
    --cc=skhan@linuxfoundation.org \
    --cc=sumit.semwal@linaro.org \
    --cc=tdave@nvidia.com \
    --cc=thomas.hellstrom@linux.intel.com \
    --cc=will@kernel.org \
    /path/to/YOUR_REPLY

  https://kernel.org/pub/software/scm/git/docs/git-send-email.html

* If your mail client supports setting the In-Reply-To header
  via mailto: links, try the mailto: link
Be sure your reply has a Subject: header at the top and a blank line before the message body.
This is a public inbox, see mirroring instructions
for how to clone and mirror all data and code used for this inbox

all inboxes | Powered by JetHome®