linux-next

mirror of https://git.kernel.org/pub/scm/linux/kernel/git/next/linux-next.git synced 2024-12-28 16:52:18 +00:00

History

Damien Le Moal 0b83c86b44 block: Prevent potential deadlock in blk_revalidate_disk_zones() The function blk_revalidate_disk_zones() calls the function disk_update_zone_resources() after freezing the device queue. In turn, disk_update_zone_resources() calls queue_limits_start_update() which takes a queue limits mutex lock, resulting in the ordering: q->q_usage_counter check -> q->limits_lock. However, the usual ordering is to always take a queue limit lock before freezing the queue to commit the limits updates, e.g., the code pattern: lim = queue_limits_start_update(q); ... blk_mq_freeze_queue(q); ret = queue_limits_commit_update(q, &lim); blk_mq_unfreeze_queue(q); Thus, blk_revalidate_disk_zones() introduces a potential circular locking dependency deadlock that lockdep sometimes catches with the splat: [ 51.934109] ====================================================== [ 51.935916] WARNING: possible circular locking dependency detected [ 51.937561] 6.12.0+ #2107 Not tainted [ 51.938648] ------------------------------------------------------ [ 51.940351] kworker/u16:4/157 is trying to acquire lock: [ 51.941805] ffff9fff0aa0bea8 (&q->limits_lock){+.+.}-{4:4}, at: disk_update_zone_resources+0x86/0x170 [ 51.944314] but task is already holding lock: [ 51.945688] ffff9fff0aa0b890 (&q->q_usage_counter(queue)#3){++++}-{0:0}, at: blk_revalidate_disk_zones+0x15f/0x340 [ 51.948527] which lock already depends on the new lock. [ 51.951296] the existing dependency chain (in reverse order) is: [ 51.953708] -> #1 (&q->q_usage_counter(queue)#3){++++}-{0:0}: [ 51.956131] blk_queue_enter+0x1c9/0x1e0 [ 51.957290] blk_mq_alloc_request+0x187/0x2a0 [ 51.958365] scsi_execute_cmd+0x78/0x490 [scsi_mod] [ 51.959514] read_capacity_16+0x111/0x410 [sd_mod] [ 51.960693] sd_revalidate_disk.isra.0+0x872/0x3240 [sd_mod] [ 51.962004] sd_probe+0x2d7/0x520 [sd_mod] [ 51.962993] really_probe+0xd5/0x330 [ 51.963898] __driver_probe_device+0x78/0x110 [ 51.964925] driver_probe_device+0x1f/0xa0 [ 51.965916] __driver_attach_async_helper+0x60/0xe0 [ 51.967017] async_run_entry_fn+0x2e/0x140 [ 51.968004] process_one_work+0x21f/0x5a0 [ 51.968987] worker_thread+0x1dc/0x3c0 [ 51.969868] kthread+0xe0/0x110 [ 51.970377] ret_from_fork+0x31/0x50 [ 51.970983] ret_from_fork_asm+0x11/0x20 [ 51.971587] -> #0 (&q->limits_lock){+.+.}-{4:4}: [ 51.972479] __lock_acquire+0x1337/0x2130 [ 51.973133] lock_acquire+0xc5/0x2d0 [ 51.973691] __mutex_lock+0xda/0xcf0 [ 51.974300] disk_update_zone_resources+0x86/0x170 [ 51.975032] blk_revalidate_disk_zones+0x16c/0x340 [ 51.975740] sd_zbc_revalidate_zones+0x73/0x160 [sd_mod] [ 51.976524] sd_revalidate_disk.isra.0+0x465/0x3240 [sd_mod] [ 51.977824] sd_probe+0x2d7/0x520 [sd_mod] [ 51.978917] really_probe+0xd5/0x330 [ 51.979915] __driver_probe_device+0x78/0x110 [ 51.981047] driver_probe_device+0x1f/0xa0 [ 51.982143] __driver_attach_async_helper+0x60/0xe0 [ 51.983282] async_run_entry_fn+0x2e/0x140 [ 51.984319] process_one_work+0x21f/0x5a0 [ 51.985873] worker_thread+0x1dc/0x3c0 [ 51.987289] kthread+0xe0/0x110 [ 51.988546] ret_from_fork+0x31/0x50 [ 51.989926] ret_from_fork_asm+0x11/0x20 [ 51.991376] other info that might help us debug this: [ 51.994127] Possible unsafe locking scenario: [ 51.995651] CPU0 CPU1 [ 51.996694] ---- ---- [ 51.997716] lock(&q->q_usage_counter(queue)#3); [ 51.998817] lock(&q->limits_lock); [ 52.000043] lock(&q->q_usage_counter(queue)#3); [ 52.001638] lock(&q->limits_lock); [ 52.002485] * DEADLOCK * Prevent this issue by moving the calls to blk_mq_freeze_queue() and blk_mq_unfreeze_queue() around the call to queue_limits_commit_update() in disk_update_zone_resources(). In case of revalidation failure, the call to disk_free_zone_resources() in blk_revalidate_disk_zones() is still done with the queue frozen as before. Fixes: `843283e96e` ("block: Fake max open zones limit when there is no limit") Cc: stable@vger.kernel.org Signed-off-by: Damien Le Moal <dlemoal@kernel.org> Reviewed-by: Christoph Hellwig <hch@lst.de> Link: https://lore.kernel.org/r/20241126104705.183996-1-dlemoal@kernel.org Signed-off-by: Jens Axboe <axboe@kernel.dk>		2024-11-26 07:56:43 -07:00
..
partitions	block: add partition uuid into uevent as "PARTUUID"	2024-10-22 08:15:17 -06:00
badblocks.c	badblocks: avoid checking invalid range in badblocks_check()	2023-12-23 18:38:08 -07:00
bdev.c	for-6.12/block-20240925	2024-09-25 14:56:40 -07:00
bfq-cgroup.c	Revert "block, bfq: merge bfq_release_process_ref() into bfq_put_cooperator()"	2024-11-19 19:05:32 -07:00
bfq-iosched.c	Revert "block, bfq: merge bfq_release_process_ref() into bfq_put_cooperator()"	2024-11-19 19:05:32 -07:00
bfq-iosched.h	block, bfq: remove bfq_log_bfqg()	2024-09-10 16:32:09 -06:00
bfq-wf2q.c	block, bfq: inject I/O to underutilized actuators	2023-01-29 15:18:33 -07:00
bio-integrity.c	blk-integrity: remove seed for user mapped buffers	2024-10-30 07:49:32 -06:00
bio.c	block: Error an attempt to split an atomic write in bio_split()	2024-11-11 08:35:46 -07:00
blk-cgroup-fc-appid.c	block: Replace all non-returning strlcpy with strscpy	2023-06-01 09:13:31 -06:00
blk-cgroup-rwstat.c	blk-cgroup: use group allocation/free of per-cpu counters API	2024-04-03 09:10:17 -06:00
blk-cgroup-rwstat.h	block: Use the new blk_opf_t type	2022-07-14 12:14:30 -06:00
blk-cgroup.c	blk-ioprio: remove per-disk structure	2024-07-28 16:47:51 -06:00
blk-cgroup.h	blk-cgroup: Remove unused declaration blkg_path()	2024-08-16 15:07:27 -06:00
blk-core.c	block: add a rq_list type	2024-11-13 12:04:58 -07:00
blk-crypto-fallback.c	block: Rework bio_split() return value	2024-11-11 08:35:46 -07:00
blk-crypto-internal.h	blk-crypto: remove blk_crypto_insert_cloned_request()	2023-03-16 09:35:09 -06:00
blk-crypto-profile.c	blk-crypto: use dynamic lock class for blk_crypto_profile::lock	2023-07-05 16:36:12 -06:00
blk-crypto-sysfs.c	block: make kobj_type structures constant	2023-02-09 09:38:16 -07:00
blk-crypto.c	blk-crypto: make blk_crypto_evict_key() more robust	2023-03-16 09:35:09 -06:00
blk-flush.c	for-6.11/block-20240710	2024-07-15 14:20:22 -07:00
blk-ia-ranges.c	block: make kobj_type structures constant	2023-02-09 09:38:16 -07:00
blk-integrity.c	blk-integrity: remove seed for user mapped buffers	2024-10-30 07:49:32 -06:00
blk-ioc.c	block: replace call_rcu by kfree_rcu for simple kmem_cache_free callback	2024-10-22 08:16:40 -06:00
blk-iocost.c	blk_iocost: remove some duplicate irq disable/enables	2024-10-02 07:15:43 -06:00
blk-iolatency.c	block: add blk_time_get_ns() and blk_time_get() helpers	2024-02-05 10:07:22 -07:00
blk-ioprio.c	blk-ioprio: remove per-disk structure	2024-07-28 16:47:51 -06:00
blk-ioprio.h	blk-ioprio: remove per-disk structure	2024-07-28 16:47:51 -06:00
blk-lib.c	block: fix detection of unsupported WRITE SAME in blkdev_issue_write_zeroes	2024-08-28 08:49:25 -06:00
blk-map.c	block: don't free the integrity payload in bio_integrity_unmap_free_user	2024-07-03 10:21:16 -06:00
blk-merge.c	block: req->bio is always set in the merge code	2024-11-19 19:06:57 -07:00
blk-mq-cpumap.c	blk-mq: include <linux/blk-mq.h> in block/blk-mq.h	2023-04-13 06:52:29 -06:00
blk-mq-debugfs.c	block: Catch possible entries missing from rqf_name[]	2024-07-19 09:32:49 -06:00
blk-mq-debugfs.h	block: Replace zone_wlock debugfs entry with zone_wplugs entry	2024-04-17 08:44:03 -06:00
blk-mq-pci.c	blk-mq: include <linux/blk-mq.h> in block/blk-mq.h	2023-04-13 06:52:29 -06:00
blk-mq-sched.c	blk-mq: Remove the hctx 'run' debugfs attribute	2024-01-17 14:16:34 -07:00
blk-mq-sched.h	blk-mq: make sure elevator callbacks aren't called for passthrough request	2023-05-18 19:42:54 -06:00
blk-mq-sysfs.c	blk-mq: include <linux/blk-mq.h> in block/blk-mq.h	2023-04-13 06:52:29 -06:00
blk-mq-tag.c	block: Fix lockdep warning in blk_mq_mark_tag_wait	2024-08-15 19:25:03 -06:00
blk-mq-virtio.c	blk-mq: include <linux/blk-mq.h> in block/blk-mq.h	2023-04-13 06:52:29 -06:00
blk-mq.c	block: Remove extra part pointer NULLify in blk_rq_init()	2024-11-25 08:42:14 -07:00
blk-mq.h	block: add a rq_list type	2024-11-13 12:04:58 -07:00
blk-pm.c	block: Remove blk_set_runtime_active()	2023-11-20 10:22:40 -07:00
blk-pm.h	block: Remove unused blk_pm_*() function definitions	2021-02-22 06:33:48 -07:00
blk-rq-qos.c	block: remove redundant explicit memory barrier from rq_qos waiter and waker	2024-10-22 14:05:09 -06:00
blk-rq-qos.h	block: skip QUEUE_FLAG_STATS and rq-qos for passthrough io	2023-12-01 18:29:18 -07:00
blk-settings.c	block: Support atomic writes limits for stacked devices	2024-11-19 10:30:02 -07:00
blk-stat.c	blk-throttle: remove CONFIG_BLK_DEV_THROTTLING_LOW	2024-05-09 09:44:55 -06:00
blk-stat.h	block: delete redundant function declaration	2024-05-27 13:58:06 -06:00
blk-sysfs.c	block: fix uaf for flush rq while iterating tags	2024-11-18 18:31:57 -07:00
blk-throttle.c	block: flush all throttled bios when deleting the cgroup	2024-10-22 08:16:43 -06:00
blk-throttle.h	blk-throttle: remove last_low_overflow_time	2024-09-10 16:31:41 -06:00
blk-timeout.c	block: blk-timeout: delete duplicated word	2020-07-31 16:29:47 -06:00
blk-wbt.c	blk-wbt: don't throttle swap writes in direct reclaim	2024-07-01 06:51:53 -06:00
blk-wbt.h	blk-wbt: remove the separate write cache tracking	2023-12-26 09:28:10 -07:00
blk-zoned.c	block: Prevent potential deadlock in blk_revalidate_disk_zones()	2024-11-26 07:56:43 -07:00
blk.h	block: lift bio_is_zone_append to bio.h	2024-11-11 09:06:45 -07:00
bounce.c	block: split integrity support out of bio.h	2024-07-03 10:21:15 -06:00
bsg-lib.c	scsi: bsg: Pass dev to blk_mq_alloc_queue()	2024-05-30 20:22:15 -04:00
bsg.c	SCSI misc on 20230629	2023-06-30 11:57:07 -07:00
disk-events.c	block: move bdev_mark_dead out of disk_check_media_change	2023-10-28 13:29:23 +02:00
early-lookup.c	wrapper for access to ->bd_partno	2024-05-02 17:48:09 -04:00
elevator.c	block: don't verify IO lock for freeze/unfreeze in elevator_init_mq()	2024-11-07 16:27:22 -07:00
elevator.h	block: return void from the queue_sysfs_entry load_module method	2024-10-22 08:16:22 -06:00
fops.c	fs/block: Check for IOCB_DIRECT in generic_atomic_write_valid()	2024-10-19 16:48:22 -06:00
genhd.c	block: fix uaf for flush rq while iterating tags	2024-11-18 18:31:57 -07:00
holder.c	block: fix deadlock between bd_link_disk_holder and partition scan	2024-02-23 07:44:19 -07:00
ioctl.c	block: implement async io_uring discard cmd	2024-09-11 10:45:28 -06:00
ioprio.c	block: move __get_task_ioprio() into header file	2024-01-08 12:27:39 -07:00
Kconfig	block: remove the blk_integrity_profile structure	2024-06-14 10:20:06 -06:00
Kconfig.iosched	block: Default to use cgroup support for BFQ	2023-01-30 09:42:42 -07:00
kyber-iosched.c	blk-mq: pass a flags argument to elevator_type->insert_requests	2023-04-13 06:52:30 -06:00
Makefile	block: remove the blk_integrity_profile structure	2024-06-14 10:20:06 -06:00
mq-deadline.c	block/mq-deadline: Fix the tag reservation code	2024-07-02 08:47:45 -06:00
opal_proto.h	block: sed-opal: handle empty atoms when parsing response	2024-02-16 15:52:45 -07:00
sed-opal.c	block: sed-opal: add ioctl IOC_OPAL_SET_SID_PW	2024-10-22 08:16:40 -06:00
t10-pi.c	move asm/unaligned.h to linux/unaligned.h	2024-10-02 17:23:23 -04:00