From: Binbin Wu <binbin.wu@linux.intel.com>
To: isaku.yamahata@intel.com, Xiaoyao Li <xiaoyao.li@intel.com>
Cc: kvm@vger.kernel.org, linux-kernel@vger.kernel.org,
isaku.yamahata@gmail.com, Paolo Bonzini <pbonzini@redhat.com>,
erdemaktas@google.com, Sean Christopherson <seanjc@google.com>,
Sagi Shahar <sagis@google.com>,
David Matlack <dmatlack@google.com>,
Kai Huang <kai.huang@intel.com>,
Zhi Wang <zhi.wang.linux@gmail.com>,
chen.bo@intel.com, hang.yuan@intel.com, tina.zhang@intel.com
Subject: Re: [RFC PATCH v4 05/16] KVM: TDX: Pass size to reclaim_page()
Date: Wed, 6 Sep 2023 09:48:35 +0800 [thread overview]
Message-ID: <383cf8d1-1d6f-f0d3-08de-fe4dc3ce1778@linux.intel.com> (raw)
In-Reply-To: <48b900ccfa2257ddbfeea475b9b43ee36fb52080.1690323516.git.isaku.yamahata@intel.com>
On 7/26/2023 6:23 AM, isaku.yamahata@intel.com wrote:
> From: Xiaoyao Li <xiaoyao.li@intel.com>
>
> A 2MB large page can be tdh_mem_page_aug()'ed to TD directly. In this case,
> it needs to reclaim and clear the page as 2MB size.
>
> Signed-off-by: Xiaoyao Li <xiaoyao.li@intel.com>
> Signed-off-by: Isaku Yamahata <isaku.yamahata@intel.com>
> ---
> arch/x86/kvm/vmx/tdx.c | 24 ++++++++++++++----------
> 1 file changed, 14 insertions(+), 10 deletions(-)
>
> diff --git a/arch/x86/kvm/vmx/tdx.c b/arch/x86/kvm/vmx/tdx.c
> index 3522ee232eda..86cfbf435671 100644
> --- a/arch/x86/kvm/vmx/tdx.c
> +++ b/arch/x86/kvm/vmx/tdx.c
> @@ -198,12 +198,13 @@ static void tdx_disassociate_vp_on_cpu(struct kvm_vcpu *vcpu)
> smp_call_function_single(cpu, tdx_disassociate_vp_arg, vcpu, 1);
> }
>
> -static void tdx_clear_page(unsigned long page_pa)
> +static void tdx_clear_page(unsigned long page_pa, int size)
> {
> const void *zero_page = (const void *) __va(page_to_phys(ZERO_PAGE(0)));
> void *page = __va(page_pa);
> unsigned long i;
>
> + WARN_ON_ONCE(size % PAGE_SIZE);
> /*
> * When re-assign one page from old keyid to a new keyid, MOVDIR64B is
> * required to clear/write the page with new keyid to prevent integrity
> @@ -212,7 +213,7 @@ static void tdx_clear_page(unsigned long page_pa)
> * clflush doesn't flush cache with HKID set. The cache line could be
> * poisoned (even without MKTME-i), clear the poison bit.
> */
> - for (i = 0; i < PAGE_SIZE; i += 64)
> + for (i = 0; i < size; i += 64)
> movdir64b(page + i, zero_page);
> /*
> * MOVDIR64B store uses WC buffer. Prevent following memory reads
> @@ -221,7 +222,8 @@ static void tdx_clear_page(unsigned long page_pa)
> __mb();
> }
>
> -static int tdx_reclaim_page(hpa_t pa, bool do_wb, u16 hkid)
> +static int tdx_reclaim_page(hpa_t pa, enum pg_level level,
> + bool do_wb, u16 hkid)
> {
> struct tdx_module_output out;
> u64 err;
> @@ -239,8 +241,10 @@ static int tdx_reclaim_page(hpa_t pa, bool do_wb, u16 hkid)
> pr_tdx_error(TDH_PHYMEM_PAGE_RECLAIM, err, &out);
> return -EIO;
> }
> + /* out.r8 == tdx sept page level */
> + WARN_ON_ONCE(out.r8 != pg_level_to_tdx_sept_level(level));
>
> - if (do_wb) {
> + if (do_wb && level == PG_LEVEL_4K) {
I was wondering if it is better to add a WARN_ON_ONCE() to ensure level is
PG_LEVEL_4K instead of skipping it silently. But later, I found the
warning of
comparing out.r8 and level has guaranteed that there will be a warning
if there
is a mismatch between do_wb and level.
> /*
> * Only TDR page gets into this path. No contention is expected
> * because of the last page of TD.
> @@ -252,7 +256,7 @@ static int tdx_reclaim_page(hpa_t pa, bool do_wb, u16 hkid)
> }
> }
>
> - tdx_clear_page(pa);
> + tdx_clear_page(pa, KVM_HPAGE_SIZE(level));
> return 0;
> }
>
> @@ -266,7 +270,7 @@ static void tdx_reclaim_td_page(unsigned long td_page_pa)
> * was already flushed by TDH.PHYMEM.CACHE.WB before here, So
> * cache doesn't need to be flushed again.
> */
> - if (tdx_reclaim_page(td_page_pa, false, 0))
> + if (tdx_reclaim_page(td_page_pa, PG_LEVEL_4K, false, 0))
> /*
> * Leak the page on failure:
> * tdx_reclaim_page() returns an error if and only if there's an
> @@ -474,7 +478,7 @@ void tdx_vm_free(struct kvm *kvm)
> * while operating on TD (Especially reclaiming TDCS). Cache flush with
> * TDX global HKID is needed.
> */
> - if (tdx_reclaim_page(kvm_tdx->tdr_pa, true, tdx_global_keyid))
> + if (tdx_reclaim_page(kvm_tdx->tdr_pa, PG_LEVEL_4K, true, tdx_global_keyid))
> return;
>
> free_page((unsigned long)__va(kvm_tdx->tdr_pa));
> @@ -1468,7 +1472,7 @@ static int tdx_sept_drop_private_spte(struct kvm *kvm, gfn_t gfn,
> * The HKID assigned to this TD was already freed and cache
> * was already flushed. We don't have to flush again.
> */
> - err = tdx_reclaim_page(hpa, false, 0);
> + err = tdx_reclaim_page(hpa, level, false, 0);
> if (KVM_BUG_ON(err, kvm))
> return -EIO;
> tdx_unpin(kvm, pfn);
> @@ -1501,7 +1505,7 @@ static int tdx_sept_drop_private_spte(struct kvm *kvm, gfn_t gfn,
> pr_tdx_error(TDH_PHYMEM_PAGE_WBINVD, err, NULL);
> return -EIO;
> }
> - tdx_clear_page(hpa);
> + tdx_clear_page(hpa, PAGE_SIZE);
> tdx_unpin(kvm, pfn);
> return 0;
> }
> @@ -1612,7 +1616,7 @@ static int tdx_sept_free_private_spt(struct kvm *kvm, gfn_t gfn,
> * already flushed. We don't have to flush again.
> */
> if (!is_hkid_assigned(kvm_tdx))
> - return tdx_reclaim_page(__pa(private_spt), false, 0);
> + return tdx_reclaim_page(__pa(private_spt), PG_LEVEL_4K, false, 0);
>
> /*
> * free_private_spt() is (obviously) called when a shadow page is being
next prev parent reply other threads:[~2023-09-06 1:48 UTC|newest]
Thread overview: 21+ messages / expand[flat|nested] mbox.gz Atom feed top
2023-07-25 22:23 [RFC PATCH v4 00/16] KVM TDX: TDP MMU: large page support isaku.yamahata
2023-07-25 22:23 ` [RFC PATCH v4 01/16] KVM: TDP_MMU: Go to next level if smaller private mapping exists isaku.yamahata
2023-09-05 8:10 ` Binbin Wu
2023-07-25 22:23 ` [RFC PATCH v4 02/16] KVM: TDX: Pass page level to cache flush before TDX SEAMCALL isaku.yamahata
2023-07-25 22:23 ` [RFC PATCH v4 03/16] KVM: TDX: Pass KVM page level to tdh_mem_page_add() and tdh_mem_page_aug() isaku.yamahata
2023-07-25 22:23 ` [RFC PATCH v4 04/16] KVM: TDX: Pass size to tdx_measure_page() isaku.yamahata
2023-07-25 22:23 ` [RFC PATCH v4 05/16] KVM: TDX: Pass size to reclaim_page() isaku.yamahata
2023-09-06 1:48 ` Binbin Wu [this message]
2023-07-25 22:23 ` [RFC PATCH v4 06/16] KVM: TDX: Update tdx_sept_{set,drop}_private_spte() to support large page isaku.yamahata
2023-07-25 22:23 ` [RFC PATCH v4 07/16] KVM: MMU: Introduce level info in PFERR code isaku.yamahata
2023-07-25 22:23 ` [RFC PATCH v4 08/16] KVM: TDX: Pin pages via get_page() right before ADD/AUG'ed to TDs isaku.yamahata
2023-09-07 5:26 ` Binbin Wu
2023-07-25 22:23 ` [RFC PATCH v4 09/16] KVM: TDX: Pass desired page level in err code for page fault handler isaku.yamahata
2023-07-25 22:23 ` [RFC PATCH v4 10/16] KVM: x86/tdp_mmu: Allocate private page table for large page split isaku.yamahata
2023-07-25 22:23 ` [RFC PATCH v4 11/16] KVM: x86/tdp_mmu: Split the large page when zap leaf isaku.yamahata
2023-07-25 22:23 ` [RFC PATCH v4 12/16] KVM: x86/tdp_mmu, TDX: Split a large page when 4KB page within it converted to shared isaku.yamahata
2023-07-25 22:23 ` [RFC PATCH v4 13/16] KVM: x86/tdp_mmu: Try to merge pages into a large page isaku.yamahata
2023-08-14 20:35 ` Isaku Yamahata
2023-07-25 22:24 ` [RFC PATCH v4 14/16] KVM: x86/tdp_mmu: TDX: Implement " isaku.yamahata
2023-07-25 22:24 ` [RFC PATCH v4 15/16] KVM: x86/mmu: Make kvm fault handler aware of large page of private memslot isaku.yamahata
2023-07-25 22:24 ` [RFC PATCH v4 16/16] KVM: TDX: Allow 2MB large page for TD GUEST isaku.yamahata
Reply instructions:
You may reply publicly to this message via plain-text email
using any one of the following methods:
* Save the following mbox file, import it into your mail client,
and reply-to-all from there: mbox
Avoid top-posting and favor interleaved quoting:
https://en.wikipedia.org/wiki/Posting_style#Interleaved_style
* Reply using the --to, --cc, and --in-reply-to
switches of git-send-email(1):
git send-email \
--in-reply-to=383cf8d1-1d6f-f0d3-08de-fe4dc3ce1778@linux.intel.com \
--to=binbin.wu@linux.intel.com \
--cc=chen.bo@intel.com \
--cc=dmatlack@google.com \
--cc=erdemaktas@google.com \
--cc=hang.yuan@intel.com \
--cc=isaku.yamahata@gmail.com \
--cc=isaku.yamahata@intel.com \
--cc=kai.huang@intel.com \
--cc=kvm@vger.kernel.org \
--cc=linux-kernel@vger.kernel.org \
--cc=pbonzini@redhat.com \
--cc=sagis@google.com \
--cc=seanjc@google.com \
--cc=tina.zhang@intel.com \
--cc=xiaoyao.li@intel.com \
--cc=zhi.wang.linux@gmail.com \
/path/to/YOUR_REPLY
https://kernel.org/pub/software/scm/git/docs/git-send-email.html
* If your mail client supports setting the In-Reply-To header
via mailto: links, try the mailto: link
Be sure your reply has a Subject: header at the top and a blank line
before the message body.
This is a public inbox, see mirroring instructions
for how to clone and mirror all data and code used for this inbox
Powered by JetHome