mirror of https://lore.kernel.org/lkml/
 help / color / mirror / Atom feed
From: "Yinghai Lu" <yhlu.kernel@gmail.com>
To: "Andi Kleen" <ak@suse.de>
Cc: "Andrew Morton" <akpm@linux-foundation.org>,
	mingo@elte.hu, clameter@sgi.com, linux-kernel@vger.kernel.org,
	"Yasunori Goto" <y-goto@jp.fujitsu.com>,
	"KAMEZAWA Hiroyuki" <kamezawa.hiroyu@jp.fujitsu.com>
Subject: Re: [PATCH] mm: fix boundary checking in free_bootmem_core
Date: Thu, 13 Mar 2008 15:22:39 -0700	[thread overview]
Message-ID: <86802c440803131522t3d038d39gbe8eb0d38ddcb634@mail.gmail.com> (raw)
In-Reply-To: <200803132259.47063.ak@suse.de>

On Thu, Mar 13, 2008 at 2:59 PM, Andi Kleen <ak@suse.de> wrote:
> On Thursday 13 March 2008 02:22:40 Andrew Morton wrote:
>  > On Wed, 12 Mar 2008 18:11:41 -0700 "Yinghai Lu" <yhlu.kernel@gmail.com> wrote:
>  >
>  > > >  <looks at it>
>  > > >
>  > > >  Sorry, but I find the changelog very hard to amke sense of.  I presently
>  > > >  have:
>  > > >
>  > > >
>  > > >   So call it when numa is enabled, we don't know which node have that
>  > > >   range.  and make it more robust.
>  > > >
>  > > >   Try to trim it to get valid sidx, and eidx.
>  > > >
>  > > >  Could you please expand on this?
>  > >
>  > > please check following...
>  > >
>  >
>  > Heaps better, thanks ;)  Below is what I now have.
>  >
>  > (cc's people)
>  >
>  > Guys, could you please review this?  Maybe test it a bit?
>  >
>  > Thanks.
>  >
>  >
>  > From: "Yinghai Lu" <yhlu.kernel@gmail.com>
>  >
>  > With numa enabled, some callers could have a range o fmemory on one node but
>  > try to free that on other node.  This can cause some pages to be freed
>  > wrongly.
>
>  Concrete examples?
>
>  If that happens it's really just a problem that the bootmem API
>  is wrong. I was always annoyed by the hardcoded NODE_DATA(0)s in
>  free_bootmem.
>
>  I would suggest if that happens you just fix free_bootmem to search
>  for the correct node instead of hardcoding 0 and then eliminate
>  free_bootmem_node() everywhere and replace it with free_bootmem()
>
> >
>  > For example: when we try to allocate 128g boot ram early for gart/swiotlb, and
>  > free that range later so gart/swiotlb can get some range afterwards.
>
>  I'm confused by the example. AFAIK there is no memory freeing in either
>  gart nor swiotlb. At least there wasn't until very recently.

For big system when numa=off or disabled, vmemmap will use 3.6g ram
when you have 256g. if you don't allocate the PMD continuous.

then i tried to reserve 64M or 128M RAM before that, and free that
before gart/switotble try to allloc_bootmem under 4g.

that patch will make the system without ram on node0 not happy.
because of free_bootmem is hardcoded to use node0.

>
>
>  >
>  > With this patch, we don't need to care which node holds the range, just loop
>  > to call free_bootmem_node for all online nodes.
>  >
>  > This patch make free_bootmem_core() more robust by trimming the sidx and eidx
>  > according the ram range that the node has.
>
>  I think you should just kill free_bootmem_node() and replace it everywhere
>  with your improved free_bootmem()

using phys_to_nid()? it seems we only have that on x86_64.

also there is assumpation that reserve_bootmem_node, reserver_bootmem
can not cross the nodes.
I want to remove that constrient too.


YH

  reply	other threads:[~2008-03-13 22:22 UTC|newest]

Thread overview: 14+ messages / expand[flat|nested]  mbox.gz  Atom feed  top
2008-03-12  1:01 Yinghai Lu
2008-03-12 23:21 ` Yinghai Lu
2008-03-12 23:33   ` Andrew Morton
2008-03-13  1:11     ` Yinghai Lu
2008-03-13  1:22       ` Andrew Morton
2008-03-13 21:59         ` Andi Kleen
2008-03-13 22:22           ` Yinghai Lu [this message]
2008-03-14 11:58             ` Andi Kleen
2008-03-14 16:44               ` Yinghai Lu
2008-03-14 16:53                 ` Andi Kleen
2008-03-14 17:36                   ` Yinghai Lu
2008-03-21 19:44                     ` Andrew Morton
2008-03-21 20:00                       ` Ingo Molnar
2008-03-21 21:54                       ` Thomas Gleixner

Reply instructions:

You may reply publicly to this message via plain-text email
using any one of the following methods:

* Save the following mbox file, import it into your mail client,
  and reply-to-all from there: mbox

  Avoid top-posting and favor interleaved quoting:
  https://en.wikipedia.org/wiki/Posting_style#Interleaved_style

* Reply using the --to, --cc, and --in-reply-to
  switches of git-send-email(1):

  git send-email \
    --in-reply-to=86802c440803131522t3d038d39gbe8eb0d38ddcb634@mail.gmail.com \
    --to=yhlu.kernel@gmail.com \
    --cc=ak@suse.de \
    --cc=akpm@linux-foundation.org \
    --cc=clameter@sgi.com \
    --cc=kamezawa.hiroyu@jp.fujitsu.com \
    --cc=linux-kernel@vger.kernel.org \
    --cc=mingo@elte.hu \
    --cc=y-goto@jp.fujitsu.com \
    /path/to/YOUR_REPLY

  https://kernel.org/pub/software/scm/git/docs/git-send-email.html

* If your mail client supports setting the In-Reply-To header
  via mailto: links, try the mailto: link
Be sure your reply has a Subject: header at the top and a blank line before the message body.
This is a public inbox, see mirroring instructions
for how to clone and mirror all data and code used for this inbox

all inboxes | Powered by JetHome®