From: Eric Farman <farman@linux.ibm.com>
To: Matthew Rosato <mjrosato@linux.ibm.com>,
linux-s390@vger.kernel.org, kvm@vger.kernel.org,
linux-kernel@vger.kernel.org
Cc: Halil Pasic <pasic@linux.ibm.com>,
Christian Borntraeger <borntraeger@linux.ibm.com>,
stable@vger.kernel.org
Subject: Re: [PATCH v8 08/10] s390/vfio_ccw: move cp cleanup out of not operational
Date: Mon, 27 Jul 2026 23:23:12 -0400 [thread overview]
Message-ID: <3b83fd20-434c-4be9-96d0-4f162fe769b8@linux.ibm.com> (raw)
In-Reply-To: <3048391c-003a-4bde-84c9-e1eca0063766@linux.ibm.com>
On 7/27/26 10:08 PM, Matthew Rosato wrote:
> On 7/27/26 9:35 PM, Eric Farman wrote:
>> The fsm_notoper() routine is called when the device has been
>> lost, and is (by definition) no longer operational. Since this
>> can happen asynchronously from the normal behavior of the
>> driver, the cleanup may happen when holding other locks
>> in the calling sequence (notably, the cio subchannel lock).
>>
>> Push the cleanup of the private->cp resources to a workqueue,
>> where it can be done out from under that lock sequence and
>> safely under a common locking mechanism.
>
> Nit: you mention 'safely under a common locking mechanism' but that
> doesn't actually happen until next patch.
>
> Maybe add something like 'so a subsequent patch can...'
>
>>
>> Fixes: 204b394a23ad ("vfio/ccw: Move FSM open/close to MDEV open/close")
>> Cc: stable@vger.kernel.org
>> Signed-off-by: Eric Farman <farman@linux.ibm.com>
>> ---
>> drivers/s390/cio/vfio_ccw_drv.c | 9 +++++++++
>> drivers/s390/cio/vfio_ccw_fsm.c | 3 +--
>> drivers/s390/cio/vfio_ccw_ops.c | 6 +++++-
>> drivers/s390/cio/vfio_ccw_private.h | 3 +++
>> 4 files changed, 18 insertions(+), 3 deletions(-)
>>
>> diff --git a/drivers/s390/cio/vfio_ccw_drv.c b/drivers/s390/cio/vfio_ccw_drv.c
>> index 1a095085bc72..c197ad5ab580 100644
>> --- a/drivers/s390/cio/vfio_ccw_drv.c
>> +++ b/drivers/s390/cio/vfio_ccw_drv.c
>> @@ -125,6 +125,15 @@ void vfio_ccw_crw_todo(struct work_struct *work)
>> eventfd_signal(private->crw_trigger);
>> }
>>
>> +void vfio_ccw_notoper_todo(struct work_struct *work)
>> +{
>> + struct vfio_ccw_private *private;
>> +
>> + private = container_of(work, struct vfio_ccw_private, notoper_work);
>> +
>> + cp_free(&private->cp);
>> +}
>> +
>> /*
>> * Css driver callbacks
>> */
>> diff --git a/drivers/s390/cio/vfio_ccw_fsm.c b/drivers/s390/cio/vfio_ccw_fsm.c
>> index 4d7988ea47ef..4d47a3c7b9a0 100644
>> --- a/drivers/s390/cio/vfio_ccw_fsm.c
>> +++ b/drivers/s390/cio/vfio_ccw_fsm.c
>> @@ -170,8 +170,7 @@ static void fsm_notoper(struct vfio_ccw_private *private,
>> css_sched_sch_todo(sch, SCH_TODO_UNREG);
>> private->state = VFIO_CCW_STATE_NOT_OPER;
>>
>> - /* This is usually handled during CLOSE event */
>> - cp_free(&private->cp);
>> + queue_work(vfio_ccw_work_q, &private->notoper_work);
>> }
>>
>> /*
>> diff --git a/drivers/s390/cio/vfio_ccw_ops.c b/drivers/s390/cio/vfio_ccw_ops.c
>> index bc8eb485d03f..1cca3ecdae45 100644
>> --- a/drivers/s390/cio/vfio_ccw_ops.c
>> +++ b/drivers/s390/cio/vfio_ccw_ops.c
>> @@ -54,6 +54,7 @@ static int vfio_ccw_mdev_init_dev(struct vfio_device *vdev)
>> INIT_LIST_HEAD(&private->crw);
>> INIT_WORK(&private->io_work, vfio_ccw_sch_io_todo);
>> INIT_WORK(&private->crw_work, vfio_ccw_crw_todo);
>> + INIT_WORK(&private->notoper_work, vfio_ccw_notoper_todo);
>>
>> private->cp.guest_cp = kzalloc_objs(struct ccw1, CCWCHAIN_LEN_MAX);
>> if (!private->cp.guest_cp)
>> @@ -133,7 +134,9 @@ static void vfio_ccw_mdev_release_dev(struct vfio_device *vdev)
>>
>> /*
>> * Ensure these work items are fully drained, so none can
>> - * fire after being released.
>> + * fire after being released. The notoper_work struct is
>> + * only meaningful if the device had been opened, which
>> + * means it would have been cleaned in an earlier close.
>> */
>> cancel_work_sync(&private->io_work);
>> cancel_work_sync(&private->crw_work);
>> @@ -216,6 +219,7 @@ static void vfio_ccw_mdev_close_device(struct vfio_device *vdev)
>> */
>> cancel_work_sync(&private->io_work);
>> cancel_work_sync(&private->crw_work);
>> + flush_work(&private->notoper_work);
>
> Sashiko found a path where you might never open/close, so I guess we
> need to flush in both close and release after all?
Argh.
The distinction I overlooked is that while the device might not be
opened, meaning there's no cp stuff to free, the fsm_notoper call will
enqueue the workqueue itself anyway. It makes no distinction of whether
there will be anything for the cp_free logic to actually do.
>
> I still think flush vs cancel to ensure that we process the cp_free()
> work if it's pending else we risk leaking the cp resources.
I have been going back and forth, and agree flush is probably better.
Seems wrong to just blindly cancel it, even if there -shouldn't- be
anything to release.
>
> I think this is the last issue of note, so this series is very close as
> far as I'm concerned.
>
> Thanks,
> Matt
>
>>
>> vfio_ccw_unregister_dev_regions(private);
>> }
>> diff --git a/drivers/s390/cio/vfio_ccw_private.h b/drivers/s390/cio/vfio_ccw_private.h
>> index 0501d4bbcdbd..e2256402b089 100644
>> --- a/drivers/s390/cio/vfio_ccw_private.h
>> +++ b/drivers/s390/cio/vfio_ccw_private.h
>> @@ -102,6 +102,7 @@ struct vfio_ccw_parent {
>> * @req_trigger: eventfd ctx for signaling userspace to return device
>> * @io_work: work for deferral process of I/O handling
>> * @crw_work: work for deferral process of CRW handling
>> + * @notoper_work: work for deferred processing in not-operational state
>> */
>> struct vfio_ccw_private {
>> struct vfio_device vdev;
>> @@ -125,11 +126,13 @@ struct vfio_ccw_private {
>> struct eventfd_ctx *req_trigger;
>> struct work_struct io_work;
>> struct work_struct crw_work;
>> + struct work_struct notoper_work;
>> } __aligned(8);
>>
>> int vfio_ccw_sch_quiesce(struct subchannel *sch);
>> void vfio_ccw_sch_io_todo(struct work_struct *work);
>> void vfio_ccw_crw_todo(struct work_struct *work);
>> +void vfio_ccw_notoper_todo(struct work_struct *work);
>>
>> extern struct mdev_driver vfio_ccw_mdev_driver;
>>
>
next prev parent reply other threads:[~2026-07-28 3:23 UTC|newest]
Thread overview: 15+ messages / expand[flat|nested] mbox.gz Atom feed top
2026-07-28 1:34 [PATCH v8 00/10] s390/vfio_ccw fixes Eric Farman
2026-07-28 1:35 ` [PATCH v8 01/10] s390/vfio_ccw: free all memory if cp_init() fails Eric Farman
2026-07-28 1:35 ` [PATCH v8 02/10] s390/vfio_ccw: limit the number of channel program segments Eric Farman
2026-07-28 1:35 ` [PATCH v8 03/10] s390/vfio_ccw: fix out of bounds check on CCW array Eric Farman
2026-07-28 1:35 ` [PATCH v8 04/10] s390/vfio_ccw: ensure first IDAW remains constant Eric Farman
2026-07-28 1:35 ` [PATCH v8 05/10] s390/vfio_ccw: calculate idal length based on idaw type Eric Farman
2026-07-28 1:35 ` [PATCH v8 06/10] s390/vfio_ccw: ensure index for read/write regions are within range Eric Farman
2026-07-28 1:35 ` [PATCH v8 07/10] s390/vfio_ccw: cancel existing workqueues Eric Farman
2026-07-28 2:07 ` Matthew Rosato
2026-07-28 1:35 ` [PATCH v8 08/10] s390/vfio_ccw: move cp cleanup out of not operational Eric Farman
2026-07-28 2:08 ` Matthew Rosato
2026-07-28 3:23 ` Eric Farman [this message]
2026-07-28 1:35 ` [PATCH v8 09/10] s390/vfio_ccw: selectively expand io_mutex Eric Farman
2026-07-28 2:08 ` Matthew Rosato
2026-07-28 1:35 ` [PATCH v8 10/10] s390/vfio_ccw: implement a crw lock Eric Farman
Reply instructions:
You may reply publicly to this message via plain-text email
using any one of the following methods:
* Save the following mbox file, import it into your mail client,
and reply-to-all from there: mbox
Avoid top-posting and favor interleaved quoting:
https://en.wikipedia.org/wiki/Posting_style#Interleaved_style
* Reply using the --to, --cc, and --in-reply-to
switches of git-send-email(1):
git send-email \
--in-reply-to=3b83fd20-434c-4be9-96d0-4f162fe769b8@linux.ibm.com \
--to=farman@linux.ibm.com \
--cc=borntraeger@linux.ibm.com \
--cc=kvm@vger.kernel.org \
--cc=linux-kernel@vger.kernel.org \
--cc=linux-s390@vger.kernel.org \
--cc=mjrosato@linux.ibm.com \
--cc=pasic@linux.ibm.com \
--cc=stable@vger.kernel.org \
/path/to/YOUR_REPLY
https://kernel.org/pub/software/scm/git/docs/git-send-email.html
* If your mail client supports setting the In-Reply-To header
via mailto: links, try the mailto: link
Be sure your reply has a Subject: header at the top and a blank line
before the message body.
This is a public inbox, see mirroring instructions
for how to clone and mirror all data and code used for this inbox
all inboxes | Powered by JetHome®