mirror of https://lore.kernel.org/lkml/
 help / color / mirror / Atom feed
From: Vincent Mailhol <mailhol@kernel.org>
To: Ard Biesheuvel <ardb+git@google.com>, linux-efi@vger.kernel.org
Cc: linux-kernel@vger.kernel.org, Ard Biesheuvel <ardb@kernel.org>,
	x86@kernel.org
Subject: Re: [PATCH v2 04/10] lib/ucs2_string: Split out ucs2_as_utf8_l() taking a separate limit
Date: Wed, 16 Sep 2026 12:26:22 +0200	[thread overview]
Message-ID: <1d5c8df2-2517-4e01-a7b2-6a03e31b4b1a@kernel.org> (raw)
In-Reply-To: <20260909115530.1924665-16-ardb+git@google.com>

On 09/09/2026 at 13:55, Ard Biesheuvel wrote:
> From: Ard Biesheuvel <ardb@kernel.org>
> 
> ucs2_as_utf() takes a maxlength argument, which specifies how many bytes
  ^^^^^^^^^^^^^
Typo: ucs2_as_utf8() (missing '8').

> the function is permitted to store into the destination buffer.
> 
> The same value is used as an upper bound for the ucs2_strnlen()
> invocation, which is reasonable in the general case, as each UCS-2
> character produces at least one byte of UTF-8 output, and so there is
> never a need to process more than 'maxlength' UCS-2 characters.
> 
> However, if the UCS-2 string is not NUL terminated, ucs2_strnlen() may
> read past the end of the buffer if 'maxlength' is set to a high value.
> 
> Current callers pass UCS-2 strings that are expected to be NUL
> terminated, but for processing the load options in the EFI stub, a
> version is needed that takes a separate limit argument. So split that
> off from the current implementation.
> 
> Signed-off-by: Ard Biesheuvel <ardb@kernel.org>
> ---
>  include/linux/ucs2_string.h | 11 ++++++++++-
>  lib/ucs2_string.c           |  6 +++---
>  2 files changed, 13 insertions(+), 4 deletions(-)
> 
> diff --git a/include/linux/ucs2_string.h b/include/linux/ucs2_string.h
> index c499ae809c7d..74f23ca5a967 100644
> --- a/include/linux/ucs2_string.h
> +++ b/include/linux/ucs2_string.h
> @@ -14,7 +14,16 @@ ssize_t ucs2_strscpy(ucs2_char_t *dst, const ucs2_char_t *src, size_t count);
>  int ucs2_strncmp(const ucs2_char_t *a, const ucs2_char_t *b, size_t len);
>  
>  unsigned long ucs2_utf8size(const ucs2_char_t *src);
> +unsigned long
> +ucs2_as_utf8_l(u8 *dest, const ucs2_char_t *src, unsigned long limit,
> +	       unsigned long maxlength);
> +
> +static inline
>  unsigned long ucs2_as_utf8(u8 *dest, const ucs2_char_t *src,
> -			   unsigned long maxlength);
> +			   unsigned long maxlength)
> +{
> +	return ucs2_as_utf8_l(dest, src, ucs2_strnlen(src, maxlength),
> +			      maxlength);
> +}
>  
>  #endif /* _LINUX_UCS2_STRING_H_ */
> diff --git a/lib/ucs2_string.c b/lib/ucs2_string.c
> index f75fb4f7961a..2df9bef79eea 100644
> --- a/lib/ucs2_string.c
> +++ b/lib/ucs2_string.c
> @@ -132,11 +132,11 @@ EXPORT_SYMBOL(ucs2_utf8size);
>   * final NUL character.
>   */

ucs2_as_utf8_l() still uses the old documentation of ucs2_as_utf8().
It currently reads:

  /*
   * copy at most maxlength bytes of whole utf8 characters to dest from the
   * ucs2 string src.
   *
   * The return value is the number of characters copied, not including the
   * final NUL character.
   */

That documentation should probably be moved to linux/ucs2_string.h so
that ucs2_as_utf8() remains documented after being turned into a
static inline wrapper.

As for ucs2_as_utf8_l(), it would be worth adding a new comment block
to highlight its specific behaviour: the new limit argument, the fact
that the string is only NUL-terminated if there is enough space and
that the return value is not the number of wide characters copied but
the number of bytes copied.

>  unsigned long
> -ucs2_as_utf8(u8 *dest, const ucs2_char_t *src, unsigned long maxlength)
> +ucs2_as_utf8_l(u8 *dest, const ucs2_char_t *src, unsigned long limit,
> +	       unsigned long maxlength)
>  {
>  	unsigned int i;
>  	unsigned long j = 0;
> -	unsigned long limit = ucs2_strnlen(src, maxlength);
>  
>  	for (i = 0; maxlength && i < limit; i++) {
>  		u16 c = src[i];
> @@ -163,7 +163,7 @@ ucs2_as_utf8(u8 *dest, const ucs2_char_t *src, unsigned long maxlength)
>  		dest[j] = '\0';
>  	return j;
>  }
> -EXPORT_SYMBOL(ucs2_as_utf8);
> +EXPORT_SYMBOL(ucs2_as_utf8_l);
>  
>  #ifndef __DISABLE_EXPORTS
>  MODULE_DESCRIPTION("UCS2 string handling");


Yours sincerely,
Vincent Mailhol

  reply	other threads:[~2026-09-16 10:26 UTC|newest]

Thread overview: 26+ messages / expand[flat|nested]  mbox.gz  Atom feed  top
2026-09-09 11:55 [PATCH v2 00/10] efi/libstub: Avoid UTF-16 conversion busywork Ard Biesheuvel
2026-09-09 11:55 ` [PATCH v2 01/10] x86/boot: Drop pointless re-implementation of panic() Ard Biesheuvel
2026-09-09 19:14   ` Borislav Petkov
2026-09-09 20:43     ` Ard Biesheuvel
2026-09-10 13:12       ` Kiryl Shutsemau
2026-09-11  7:32         ` Ard Biesheuvel
2026-09-12  6:10           ` Borislav Petkov
2026-09-12  8:41             ` Ard Biesheuvel
2026-09-12 17:54               ` Borislav Petkov
2026-09-13 16:35                 ` Ard Biesheuvel
2026-09-13 18:11                   ` Borislav Petkov
2026-09-09 11:55 ` [PATCH v2 02/10] lib/ucs2_string: Drop arbitrary input size limit and associated WARN() Ard Biesheuvel
2026-09-09 11:55 ` [PATCH v2 03/10] lib/ucs2_string: Suppress modinfo when __DISABLE_EXPORTS is set Ard Biesheuvel
2026-09-09 11:55 ` [PATCH v2 04/10] lib/ucs2_string: Split out ucs2_as_utf8_l() taking a separate limit Ard Biesheuvel
2026-09-16 10:26   ` Vincent Mailhol [this message]
2026-09-09 11:55 ` [PATCH v2 05/10] efi/libstub: Use ucs2_string library for UTF-16 to UTF-8 conversion Ard Biesheuvel
2026-09-15 17:00   ` Vincent Mailhol
2026-09-09 11:55 ` [PATCH v2 06/10] efi/libstub: Avoid efi_puts() for compile time constant strings Ard Biesheuvel
2026-09-15 17:01   ` Vincent Mailhol
2026-09-09 11:55 ` [PATCH v2 07/10] efi/libstub: Output UTF-16 directly from vsnprintf() Ard Biesheuvel
2026-09-09 11:55 ` [PATCH v2 08/10] efi/libstub: Add support for printing human readable GUIDs Ard Biesheuvel
2026-09-09 11:55 ` [PATCH v2 09/10] efi/libstub: Add efi_snprintf() to construct wide strings Ard Biesheuvel
2026-09-15 17:02   ` Vincent Mailhol
2026-09-09 11:55 ` [PATCH v2 10/10] efi/libstub: add initial Boot Loader Interface support Ard Biesheuvel
2026-09-15 17:02   ` Vincent Mailhol
2026-09-15 18:28 ` [PATCH v2 00/10] efi/libstub: Avoid UTF-16 conversion busywork Vincent Mailhol

Reply instructions:

You may reply publicly to this message via plain-text email
using any one of the following methods:

* Save the following mbox file, import it into your mail client,
  and reply-to-all from there: mbox

  Avoid top-posting and favor interleaved quoting:
  https://en.wikipedia.org/wiki/Posting_style#Interleaved_style

* Reply using the --to, --cc, and --in-reply-to
  switches of git-send-email(1):

  git send-email \
    --in-reply-to=1d5c8df2-2517-4e01-a7b2-6a03e31b4b1a@kernel.org \
    --to=mailhol@kernel.org \
    --cc=ardb+git@google.com \
    --cc=ardb@kernel.org \
    --cc=linux-efi@vger.kernel.org \
    --cc=linux-kernel@vger.kernel.org \
    --cc=x86@kernel.org \
    /path/to/YOUR_REPLY

  https://kernel.org/pub/software/scm/git/docs/git-send-email.html

* If your mail client supports setting the In-Reply-To header
  via mailto: links, try the mailto: link
Be sure your reply has a Subject: header at the top and a blank line before the message body.
This is a public inbox, see mirroring instructions
for how to clone and mirror all data and code used for this inbox

all inboxes | Powered by JetHome®