Re: [RFC PATCH v2 1/7] block: Extand commit_rqs() to do batch processing

[RFC PATCH v2 0/7] Add MMC packed request support · Baolin Wang <baolin.wang7@gmail.com> · 2020-04-26
[RFC PATCH v2 1/7] block: Extand commit_rqs() to do batch processing · Baolin Wang <baolin.wang7@gmail.com> · 2020-04-26
Re: [RFC PATCH v2 1/7] block: Extand commit_rqs() to do batch processing · Christoph Hellwig <hch@infradead.org> · 2020-04-27
Re: [RFC PATCH v2 1/7] block: Extand commit_rqs() to do batch processing · Baolin Wang <baolin.wang7@gmail.com> · 2020-04-28
Re: [RFC PATCH v2 1/7] block: Extand commit_rqs() to do batch processing · Sagi Grimberg <sagi@grimberg.me> · 2020-05-08
Re: [RFC PATCH v2 1/7] block: Extand commit_rqs() to do batch processing · Ming Lei <hidden> · 2020-05-08
Re: [RFC PATCH v2 1/7] block: Extand commit_rqs() to do batch processing · Sagi Grimberg <sagi@grimberg.me> · 2020-05-08
Re: [RFC PATCH v2 1/7] block: Extand commit_rqs() to do batch processing · Ming Lei <hidden> · 2020-05-08
Re: [RFC PATCH v2 1/7] block: Extand commit_rqs() to do batch processing · Baolin Wang <baolin.wang7@gmail.com> · 2020-05-09
Re: [RFC PATCH v2 1/7] block: Extand commit_rqs() to do batch processing · Ming Lei <hidden> · 2020-05-09
Re: [RFC PATCH v2 1/7] block: Extand commit_rqs() to do batch processing · Sagi Grimberg <sagi@grimberg.me> · 2020-05-10
Re: [RFC PATCH v2 1/7] block: Extand commit_rqs() to do batch processing · Ming Lei <hidden> · 2020-05-11
Re: [RFC PATCH v2 1/7] block: Extand commit_rqs() to do batch processing · Sagi Grimberg <sagi@grimberg.me> · 2020-05-11
Re: [RFC PATCH v2 1/7] block: Extand commit_rqs() to do batch processing · Ming Lei <hidden> · 2020-05-11
Re: [RFC PATCH v2 1/7] block: Extand commit_rqs() to do batch processing · Sagi Grimberg <sagi@grimberg.me> · 2020-05-12
Re: [RFC PATCH v2 1/7] block: Extand commit_rqs() to do batch processing · Ming Lei <hidden> · 2020-05-12
[RFC PATCH v2 2/7] mmc: Add MMC packed request support for MMC software queue · Baolin Wang <baolin.wang7@gmail.com> · 2020-04-26
[RFC PATCH v2 3/7] mmc: host: sdhci: Introduce ADMA3 transfer mode · Baolin Wang <baolin.wang7@gmail.com> · 2020-04-26
[RFC PATCH v2 4/7] mmc: host: sdhci: Factor out the command configuration · Baolin Wang <baolin.wang7@gmail.com> · 2020-04-26
[RFC PATCH v2 5/7] mmc: host: sdhci: Remove redundant sg_count member of struct sdhci_host · Baolin Wang <baolin.wang7@gmail.com> · 2020-04-26
[RFC PATCH v2 6/7] mmc: host: sdhci: Add MMC packed request support · Baolin Wang <baolin.wang7@gmail.com> · 2020-04-26
[RFC PATCH v2 7/7] mmc: host: sdhci-sprd: Add MMC packed request support · Baolin Wang <baolin.wang7@gmail.com> · 2020-04-26

From: Ming Lei <hidden>
Date: 2020-05-11 01:30:05
Also in: linux-mmc, lkml

On Sun, May 10, 2020 at 12:44:53AM -0700, Sagi Grimberg wrote:

quoted

You're mostly correct. This is exactly why an I/O scheduler may be
applicable here IMO. Mostly because I/O schedulers tend to optimize for
something specific and always present tradeoffs. Users need to
understand what they are optimizing for.

Hence I'd say this functionality can definitely be available to an I/O
scheduler should one exist.

I guess it is just that there can be multiple requests available from
scheduler queue. Actually it can be so for other non-nvme drivers in
case of none, such as SCSI.

Another way is to use one per-task list(such as plug list) to hold the
requests for dispatch, then every drivers may see real .last flag, so they
may get chance for optimizing batch queuing. I will think about the
idea further and see if it is really doable.

How about my RFC v1 patch set[1], which allows dispatching more than
one request from the scheduler to support batch requests?

[1]
https://lore.kernel.org/patchwork/patch/1210034/
https://lore.kernel.org/patchwork/patch/1210035/

Basically, my idea is to dequeue request one by one, and for each
dequeued request:

- we try to get a budget and driver tag, if both succeed, add the
request to one per-task list which can be stored in stack variable,
then continue to dequeue more request

- if either budget or driver tag can't be allocated for this request,
marks the last request in the per-task list as .last, and send the
batching requests stored in the list to LLD

- when queueing batching requests to LLD, if one request isn't queued
to driver successfully, calling .commit_rqs() like before, meantime
adding the remained requests in the per-task list back to scheduler
queue or hctx->dispatch.

Sounds good to me.

quoted

One issue is that this way might degrade sequential IO performance if
the LLD just tells queue busy to blk-mq via return value of .queue_rq(),
so I guess we still may need one flag, such as BLK_MQ_F_BATCHING_SUBMISSION.

Why is that degrading sequential I/O performance? because the specific

Some devices may only return BLK_STS_RESOURCE from .queue_rq(), then more
requests are dequeued from scheduler queue if we always queue batching IOs
to LLD, and chance of IO merge is reduced, so sequential IO performance will
be effected.

Such as some scsi device which doesn't use sdev->queue_depth for
throttling IOs.

For virtio-scsi or virtio-blk, we may stop queue for avoiding the
potential affect.

device will do better without batching submissions? If so, the driver

It isn't related with batching submission, IMO.


Thanks,
Ming

`h`	back out one level
`j`	next message in thread
`k`	previous message in thread
`l`	drill in
`Esc`	close help / fold thread tree
`?`	toggle this help