[PATCH v3 0/1] riscv: add vectorized memset, memcpy and memmove
Pincheng Wang <[email protected]> Thu, 5 Mar 2026 14:19:46 +0800
| Newsgroups | gmane.comp.lib.newlib |
|---|---|
| Message-ID | <[email protected]> |
Hi all, This v3 patch adds RISC-V Vector (RVV) optimized implementations for memset, memcpy and memmove. Changes since v2: - Changed conditional compilation order, so vector path will not be chosen when optimizing for size. Changes since v1: - Switch the conditional compilation macro from __riscv_v to __riscv_vector. - Replaced '.option arch,+v' with '.option arch,+zve32x'. - Removed an unnecessary unconditional jump instruction in memmove-asm.S. - In memcpy and memmove, when __riscv_misaligned_fast is not defined, the destination address is now aligned to SZREG to improve performance on systems with slow misaligned accesses. These implementations use the RVV extension with e8 element size and m8 LMUL, and are conditionally compiled only when __riscv_vector is defined, ensuring compatibility with non-vector RISC-V systems. Benchmark results on Spacemit X60 (Muse-pi) and Canaan K230 show significant improvements. memcpy: Up to 4.84x on Muse-pi and 4.66x on K230. memset: Up to 4.31x on Muse-pi and 3.14x on K230. memmove: Up to 2.87x on Muse-pi and 1.48x on K230. Comments and suggestions are greatly appreciated. Thank you for your time and review! Best regards, Pincheng Wang Pincheng Wang (1): riscv: add vectorized memset, memcpy and memmove newlib/libc/machine/riscv/memcpy-asm.S | 52 ++++++++++++++ newlib/libc/machine/riscv/memcpy.c | 2 +- newlib/libc/machine/riscv/memmove-asm.S | 93 +++++++++++++++++++++++++ newlib/libc/machine/riscv/memmove.c | 2 +- newlib/libc/machine/riscv/memset.S | 18 +++++ 5 files changed, 165 insertions(+), 2 deletions(-) -- 2.39.5