From: Vincent Mailhol <mailhol@kernel.org>
To: Ard Biesheuvel <ardb+git@google.com>, linux-efi@vger.kernel.org
Cc: linux-kernel@vger.kernel.org, Ard Biesheuvel <ardb@kernel.org>,
x86@kernel.org
Subject: Re: [PATCH v2 04/10] lib/ucs2_string: Split out ucs2_as_utf8_l() taking a separate limit
Date: Wed, 16 Sep 2026 12:26:22 +0200 [thread overview]
Message-ID: <1d5c8df2-2517-4e01-a7b2-6a03e31b4b1a@kernel.org> (raw)
In-Reply-To: <20260909115530.1924665-16-ardb+git@google.com>
On 09/09/2026 at 13:55, Ard Biesheuvel wrote:
> From: Ard Biesheuvel <ardb@kernel.org>
>
> ucs2_as_utf() takes a maxlength argument, which specifies how many bytes
^^^^^^^^^^^^^
Typo: ucs2_as_utf8() (missing '8').
> the function is permitted to store into the destination buffer.
>
> The same value is used as an upper bound for the ucs2_strnlen()
> invocation, which is reasonable in the general case, as each UCS-2
> character produces at least one byte of UTF-8 output, and so there is
> never a need to process more than 'maxlength' UCS-2 characters.
>
> However, if the UCS-2 string is not NUL terminated, ucs2_strnlen() may
> read past the end of the buffer if 'maxlength' is set to a high value.
>
> Current callers pass UCS-2 strings that are expected to be NUL
> terminated, but for processing the load options in the EFI stub, a
> version is needed that takes a separate limit argument. So split that
> off from the current implementation.
>
> Signed-off-by: Ard Biesheuvel <ardb@kernel.org>
> ---
> include/linux/ucs2_string.h | 11 ++++++++++-
> lib/ucs2_string.c | 6 +++---
> 2 files changed, 13 insertions(+), 4 deletions(-)
>
> diff --git a/include/linux/ucs2_string.h b/include/linux/ucs2_string.h
> index c499ae809c7d..74f23ca5a967 100644
> --- a/include/linux/ucs2_string.h
> +++ b/include/linux/ucs2_string.h
> @@ -14,7 +14,16 @@ ssize_t ucs2_strscpy(ucs2_char_t *dst, const ucs2_char_t *src, size_t count);
> int ucs2_strncmp(const ucs2_char_t *a, const ucs2_char_t *b, size_t len);
>
> unsigned long ucs2_utf8size(const ucs2_char_t *src);
> +unsigned long
> +ucs2_as_utf8_l(u8 *dest, const ucs2_char_t *src, unsigned long limit,
> + unsigned long maxlength);
> +
> +static inline
> unsigned long ucs2_as_utf8(u8 *dest, const ucs2_char_t *src,
> - unsigned long maxlength);
> + unsigned long maxlength)
> +{
> + return ucs2_as_utf8_l(dest, src, ucs2_strnlen(src, maxlength),
> + maxlength);
> +}
>
> #endif /* _LINUX_UCS2_STRING_H_ */
> diff --git a/lib/ucs2_string.c b/lib/ucs2_string.c
> index f75fb4f7961a..2df9bef79eea 100644
> --- a/lib/ucs2_string.c
> +++ b/lib/ucs2_string.c
> @@ -132,11 +132,11 @@ EXPORT_SYMBOL(ucs2_utf8size);
> * final NUL character.
> */
ucs2_as_utf8_l() still uses the old documentation of ucs2_as_utf8().
It currently reads:
/*
* copy at most maxlength bytes of whole utf8 characters to dest from the
* ucs2 string src.
*
* The return value is the number of characters copied, not including the
* final NUL character.
*/
That documentation should probably be moved to linux/ucs2_string.h so
that ucs2_as_utf8() remains documented after being turned into a
static inline wrapper.
As for ucs2_as_utf8_l(), it would be worth adding a new comment block
to highlight its specific behaviour: the new limit argument, the fact
that the string is only NUL-terminated if there is enough space and
that the return value is not the number of wide characters copied but
the number of bytes copied.
> unsigned long
> -ucs2_as_utf8(u8 *dest, const ucs2_char_t *src, unsigned long maxlength)
> +ucs2_as_utf8_l(u8 *dest, const ucs2_char_t *src, unsigned long limit,
> + unsigned long maxlength)
> {
> unsigned int i;
> unsigned long j = 0;
> - unsigned long limit = ucs2_strnlen(src, maxlength);
>
> for (i = 0; maxlength && i < limit; i++) {
> u16 c = src[i];
> @@ -163,7 +163,7 @@ ucs2_as_utf8(u8 *dest, const ucs2_char_t *src, unsigned long maxlength)
> dest[j] = '\0';
> return j;
> }
> -EXPORT_SYMBOL(ucs2_as_utf8);
> +EXPORT_SYMBOL(ucs2_as_utf8_l);
>
> #ifndef __DISABLE_EXPORTS
> MODULE_DESCRIPTION("UCS2 string handling");
Yours sincerely,
Vincent Mailhol
next prev parent reply other threads:[~2026-09-16 10:26 UTC|newest]
Thread overview: 26+ messages / expand[flat|nested] mbox.gz Atom feed top
2026-09-09 11:55 [PATCH v2 00/10] efi/libstub: Avoid UTF-16 conversion busywork Ard Biesheuvel
2026-09-09 11:55 ` [PATCH v2 01/10] x86/boot: Drop pointless re-implementation of panic() Ard Biesheuvel
2026-09-09 19:14 ` Borislav Petkov
2026-09-09 20:43 ` Ard Biesheuvel
2026-09-10 13:12 ` Kiryl Shutsemau
2026-09-11 7:32 ` Ard Biesheuvel
2026-09-12 6:10 ` Borislav Petkov
2026-09-12 8:41 ` Ard Biesheuvel
2026-09-12 17:54 ` Borislav Petkov
2026-09-13 16:35 ` Ard Biesheuvel
2026-09-13 18:11 ` Borislav Petkov
2026-09-09 11:55 ` [PATCH v2 02/10] lib/ucs2_string: Drop arbitrary input size limit and associated WARN() Ard Biesheuvel
2026-09-09 11:55 ` [PATCH v2 03/10] lib/ucs2_string: Suppress modinfo when __DISABLE_EXPORTS is set Ard Biesheuvel
2026-09-09 11:55 ` [PATCH v2 04/10] lib/ucs2_string: Split out ucs2_as_utf8_l() taking a separate limit Ard Biesheuvel
2026-09-16 10:26 ` Vincent Mailhol [this message]
2026-09-09 11:55 ` [PATCH v2 05/10] efi/libstub: Use ucs2_string library for UTF-16 to UTF-8 conversion Ard Biesheuvel
2026-09-15 17:00 ` Vincent Mailhol
2026-09-09 11:55 ` [PATCH v2 06/10] efi/libstub: Avoid efi_puts() for compile time constant strings Ard Biesheuvel
2026-09-15 17:01 ` Vincent Mailhol
2026-09-09 11:55 ` [PATCH v2 07/10] efi/libstub: Output UTF-16 directly from vsnprintf() Ard Biesheuvel
2026-09-09 11:55 ` [PATCH v2 08/10] efi/libstub: Add support for printing human readable GUIDs Ard Biesheuvel
2026-09-09 11:55 ` [PATCH v2 09/10] efi/libstub: Add efi_snprintf() to construct wide strings Ard Biesheuvel
2026-09-15 17:02 ` Vincent Mailhol
2026-09-09 11:55 ` [PATCH v2 10/10] efi/libstub: add initial Boot Loader Interface support Ard Biesheuvel
2026-09-15 17:02 ` Vincent Mailhol
2026-09-15 18:28 ` [PATCH v2 00/10] efi/libstub: Avoid UTF-16 conversion busywork Vincent Mailhol
Reply instructions:
You may reply publicly to this message via plain-text email
using any one of the following methods:
* Save the following mbox file, import it into your mail client,
and reply-to-all from there: mbox
Avoid top-posting and favor interleaved quoting:
https://en.wikipedia.org/wiki/Posting_style#Interleaved_style
* Reply using the --to, --cc, and --in-reply-to
switches of git-send-email(1):
git send-email \
--in-reply-to=1d5c8df2-2517-4e01-a7b2-6a03e31b4b1a@kernel.org \
--to=mailhol@kernel.org \
--cc=ardb+git@google.com \
--cc=ardb@kernel.org \
--cc=linux-efi@vger.kernel.org \
--cc=linux-kernel@vger.kernel.org \
--cc=x86@kernel.org \
/path/to/YOUR_REPLY
https://kernel.org/pub/software/scm/git/docs/git-send-email.html
* If your mail client supports setting the In-Reply-To header
via mailto: links, try the mailto: link
Be sure your reply has a Subject: header at the top and a blank line
before the message body.
This is a public inbox, see mirroring instructions
for how to clone and mirror all data and code used for this inbox
all inboxes | Powered by JetHome®