Re: [PATCH v3 0/1] riscv: add vectorized memset, memcpy and memmove
Corinna Vinschen <[email protected]> Tue, 17 Mar 2026 16:29:23 +0100
| Newsgroups | gmane.comp.lib.newlib |
|---|---|
| Message-ID | <[email protected]> |
Kito, ping? On Mar 5 14:19, Pincheng Wang wrote: > Hi all, > > This v3 patch adds RISC-V Vector (RVV) optimized implementations for > memset, memcpy and memmove. > > Changes since v2: > - Changed conditional compilation order, so vector path will not be > chosen when optimizing for size. > > Changes since v1: > - Switch the conditional compilation macro from __riscv_v to > __riscv_vector. > - Replaced '.option arch,+v' with '.option arch,+zve32x'. > - Removed an unnecessary unconditional jump instruction in > memmove-asm.S. > - In memcpy and memmove, when __riscv_misaligned_fast is not defined, > the destination address is now aligned to SZREG to improve performance > on systems with slow misaligned accesses. > > These implementations use the RVV extension with e8 element size and m8 > LMUL, and are conditionally compiled only when __riscv_vector is > defined, ensuring compatibility with non-vector RISC-V systems. > > Benchmark results on Spacemit X60 (Muse-pi) and Canaan K230 show > significant improvements. > > memcpy: Up to 4.84x on Muse-pi and 4.66x on K230. > memset: Up to 4.31x on Muse-pi and 3.14x on K230. > memmove: Up to 2.87x on Muse-pi and 1.48x on K230. > > Comments and suggestions are greatly appreciated. Thank you for your > time and review! > > Best regards, > Pincheng Wang > > Pincheng Wang (1): > riscv: add vectorized memset, memcpy and memmove > > newlib/libc/machine/riscv/memcpy-asm.S | 52 ++++++++++++++ > newlib/libc/machine/riscv/memcpy.c | 2 +- > newlib/libc/machine/riscv/memmove-asm.S | 93 +++++++++++++++++++++++++ > newlib/libc/machine/riscv/memmove.c | 2 +- > newlib/libc/machine/riscv/memset.S | 18 +++++ > 5 files changed, 165 insertions(+), 2 deletions(-) > > -- > 2.39.5