[PATCH v4 0/3] md/raid10: fix r10bio width mismatches across reshape
"Chen Cheng" <[email protected]>
| Newsgroups | gmane.linux.kernel,gmane.linux.raid |
|---|---|
| Message-ID | <[email protected]> |
From: Chen Cheng <[email protected]> Hi, This series fixes slab out-of-bounds accesses in raid10 when reshape changes the number of raid disks while regular I/O is still reusing r10bio objects allocated under the previous geometry. The bug is reproducible with a simple 4-disk to 5-disk reshape under write load, for example: mdadm -C /dev/md777 -l10 -n4 /dev/sda /dev/sdb /dev/sdc /dev/sdd mkfs.ext4 /dev/md777 mount /dev/md777 /mnt/test fsstress -d /mnt/test -n 24000 -p 8 -l 24 & mdadm /dev/md777 --add /dev/sde mdadm --grow /dev/md777 --raid-devices=5 \ --backup-file=/tmp/md-reshape-backup Without these changes, an r10bio allocated under the old geometry can later be reused, initialized, or freed after conf->geo.raid_disks has switched to the new geometry. This creates width mismatches between the object and the current devs[] walk/initialization width, which can trigger KASAN reports such as slab-out-of-bounds in __make_request(), put_all_bios(), or find_bio_disk(). This series addresses the problem in three steps: 1. ensure the sync_action=reshape caller suspends and locks before start_reshape 2. make the regular r10bio pool fixed-size across reshape transitions, and move the pool rebuild into the freeze window before the live geometry switch; 3. track the number of valid devs[] entries in each reused r10bio and use that recorded width when walking devs[] after reshape. Changes in v4: - The sync_action=reshape path, caller now invokes mddev_suspend_and_lock() before calling start_reshape() - The md-cluster and dm-raid paths are unchanged, that is reach start_reshape() with the mddev locked but without suspended. Changes in v3: - Replace freeze_array()/unfreeze_array() in raid10_start_reshape() with mddev_suspend_and_lock_nointr()/mddev_unlock_and_resume(). freeze_array() returns when nr_pending == nr_queued, which still allows retry-list items to hold pool objects; mddev_suspend() provides the correct upper-layer quiesce interface. (Suggested by Yu Kuai) Changes in v2: - add this cover letter - convert r10bio_pool to a fixed-size kmalloc mempool - rebuild r10bio_pool inside the freeze window before switching live reshape geometry - switch raid10_quiesce() to freeze_array()/unfreeze_array() Testing: - reproduced the original KASAN slab-out-of-bounds on 4-disk -> 5-disk raid10 reshape with fsstress - verified that this series fixes that reproducer - exercised the 5-disk -> 4-disk reshape direction as well Thanks, Chen Cheng Chen Cheng (3): md: suspend array before raid10 reshape via sync_action md/raid10: make r10bio_pool use fixed-size objects md/raid10: bound reused r10bio devs[] walks by used_nr_devs drivers/md/md.c | 22 ++++++++++++++---- drivers/md/raid10.c | 56 +++++++++++++++++++++++++++++++++------------ drivers/md/raid10.h | 4 +++- 3 files changed, 61 insertions(+), 21 deletions(-) -- 2.54.0