mirror of https://lore.kernel.org/lkml/
 help / color / mirror / Atom feed
From: Hongfu Li <hongfu.li@linux.dev>
To: Andrew Morton <akpm@linux-foundation.org>
Cc: hongfu.li@linux.dev, muchun.song@linux.dev, osalvador@suse.de,
	david@kernel.org, steven.sistare@oracle.com,
	vivek.kasireddy@intel.com, linux-mm@kvack.org,
	linux-kernel@vger.kernel.org, Hongfu Li <lihongfu@kylinos.cn>
Subject: Re: [PATCH] mm/hugetlb: fix resv_huge_pages double decrement in memfd error path
Date: Fri, 28 Aug 2026 09:38:23 +0800	[thread overview]
Message-ID: <ee79e463-b1fe-4b5e-9bac-bcbaae334829@linux.dev> (raw)
In-Reply-To: <20260826194759.88487180eebf728c1df08f14@linux-foundation.org>


On 8/27/26 10:47 AM, Andrew Morton wrote:
> On Tue, 25 Aug 2026 10:10:13 +0800 Hongfu Li <hongfu.li@linux.dev> wrote:
>
>> From: Hongfu Li <lihongfu@kylinos.cn>
>>
>> alloc_hugetlb_folio_reserve() decrements h->resv_huge_pages when
>> dequeuing a folio, but unlike the use_global_reservation handling in
>> hugetlb_alloc_folio(), it does not set HPageRestoreReserve on the folio.
>>
>> Its sole caller memfd_alloc_folio() pre-allocates a reservation via
>> hugetlb_reserve_pages() before allocating. When hugetlb_add_to_page_cache()
>> fails, folio_put() drops the folio without HPageRestoreReserve set, so
>> free_huge_folio() does not restore the reservation. The subsequent
>> hugetlb_unreserve_pages() on the err_unresv path decrements the counter
>> a second time, leaving resv_huge_pages off by one for every failed
>> allocation.
>>
>> Set HPageRestoreReserve when consuming the reservation in
>> alloc_hugetlb_folio_reserve(). On the error path, free_huge_folio() then
>> restores the reservation before hugetlb_unreserve_pages() releases it.
>> The success path is unaffected, as hugetlb_add_to_page_cache() clears
>> the flag once the folio is added to the page cache.
>>
>> ...
>>
>> --- a/mm/hugetlb.c
>> +++ b/mm/hugetlb.c
>> @@ -2178,8 +2178,10 @@ struct folio *alloc_hugetlb_folio_reserve(struct hstate *h, int preferred_nid,
>>   
>>   	folio = dequeue_hugetlb_folio_nodemask(h, gfp_mask, preferred_nid,
>>   					       nmask);
>> -	if (folio)
>> +	if (folio) {
>> +		folio_set_hugetlb_restore_reserve(folio);
>>   		h->resv_huge_pages--;
>> +	}
>>   
>>   	spin_unlock_irq(&hugetlb_lock);
>>   	return folio;
> Thanks.  I pasted an AI-generated test case which might demonstrate
> this bug.  Requires fault-injection so I won't add cc:stable.

Thanks for sharing the test case.

I have run your provided test program and successfully triggered
the double‑decrement bug of resv_hugepages with fault‑injection.
Testing confirms that my patch fixes this bug.

> Also, Sashiko might have found an accounting issue in the nearby code
> (Sashiko doesn't like hugetlb.c):
>
> 	https://sashiko.dev/#/patchset/20260825021013.25672-1-hongfu.li@linux.dev

I have taken a close look at the nearby accounting issue reported
by Sashiko in hugetlb.c. It appears to be a real issue, and I intend
to submit a separate patch for review to fix this.

> #define _GNU_SOURCE
> #include <stdio.h>
> #include <stdlib.h>
> #include <unistd.h>
> #include <fcntl.h>
> #include <sys/mman.h>
> #include <linux/memfd.h>
>
> /* Read resv_hugepages from sysfs */
> static long get_resv_hugepages(void) {
>      FILE *f = fopen("/sys/kernel/mm/hugepages/hugepages-2048kB/resv_hugepages", "r");
>      if (!f) return -1;
>      long val = -1;
>      fscanf(f, "%ld", &val);
>      fclose(f);
>      return val;
> }
>
> int main(void) {
>      long orig_resv = get_resv_hugepages();
>      printf("[1] Initial resv_hugepages: %ld\n", orig_resv);
>
>      /* Enable fail_function for hugetlb_add_to_page_cache via debugfs */
>      system("echo hugetlb_add_to_page_cache > /sys/kernel/debug/fail_function/inject");
>      system("echo 100 > /sys/kernel/debug/fail_function/probability");
>
>      /* Create memfd and attempt write to allocate hugetlb folio */
>      int fd = memfd_create("test_memfd", MFD_HUGETLB);
>      if (fd >= 0) {
>          /* Force allocation path which triggers memfd_alloc_folio() */
>          ftruncate(fd, 2 * 1024 * 1024);
>          write(fd, "a", 1);
>          close(fd);
>      }
>
>      /* Disable error injection */
>      system("echo > /sys/kernel/debug/fail_function/inject");
>
>      long post_resv = get_resv_hugepages();
>      printf("[2] Post-failure resv_hugepages: %ld\n", post_resv);
>
>      if (post_resv < orig_resv) {
>          printf("[!] BUG DEMONSTRATED: resv_hugepages double-decremented by %ld!\n",
>                 orig_resv - post_resv);
>      }
>
>      return 0;
> }

-- 
Best regards,
Hongfu


      reply	other threads:[~2026-08-28  1:38 UTC|newest]

Thread overview: 4+ messages / expand[flat|nested]  mbox.gz  Atom feed  top
2026-08-25  2:10 Hongfu Li
2026-08-25  5:42 ` Muchun Song
2026-08-27  2:47 ` Andrew Morton
2026-08-28  1:38   ` Hongfu Li [this message]

Reply instructions:

You may reply publicly to this message via plain-text email
using any one of the following methods:

* Save the following mbox file, import it into your mail client,
  and reply-to-all from there: mbox

  Avoid top-posting and favor interleaved quoting:
  https://en.wikipedia.org/wiki/Posting_style#Interleaved_style

* Reply using the --to, --cc, and --in-reply-to
  switches of git-send-email(1):

  git send-email \
    --in-reply-to=ee79e463-b1fe-4b5e-9bac-bcbaae334829@linux.dev \
    --to=hongfu.li@linux.dev \
    --cc=akpm@linux-foundation.org \
    --cc=david@kernel.org \
    --cc=lihongfu@kylinos.cn \
    --cc=linux-kernel@vger.kernel.org \
    --cc=linux-mm@kvack.org \
    --cc=muchun.song@linux.dev \
    --cc=osalvador@suse.de \
    --cc=steven.sistare@oracle.com \
    --cc=vivek.kasireddy@intel.com \
    /path/to/YOUR_REPLY

  https://kernel.org/pub/software/scm/git/docs/git-send-email.html

* If your mail client supports setting the In-Reply-To header
  via mailto: links, try the mailto: link
Be sure your reply has a Subject: header at the top and a blank line before the message body.
This is a public inbox, see mirroring instructions
for how to clone and mirror all data and code used for this inbox

all inboxes | Powered by JetHome®