提交 · 0ffbce80c263821161190f20e74a12f7aa8eab7b · openeuler / Kernel

25 6月, 2014 1 次提交

blk-mq: blk_mq_start_hw_queue() should use blk_mq_run_hw_queue() · 0ffbce80

由 Jens Axboe 提交于 6月 25, 2014

Currently it calls __blk_mq_run_hw_queue(), which depends on the
CPU placement being correct. This means it's not possible to call
blk_mq_start_hw_queues(q) from a context that is correct for all
queues, leading to triggering the

WARN_ON(!cpumask_test_cpu(raw_smp_processor_id(), hctx->cpumask));

in __blk_mq_run_hw_queue().
Reported-by: NMing Lei <tom.leiming@gmail.com>
Signed-off-by: NJens Axboe <axboe@fb.com>

0ffbce80

14 6月, 2014 2 次提交

C
blk-mq: merge blk_mq_drain_queue and __blk_mq_drain_queue · 95ed0681
由 Christoph Hellwig 提交于 6月 13, 2014
```
Signed-off-by: NChristoph Hellwig <hch@lst.de>
Signed-off-by: NJens Axboe <axboe@fb.com>
```
95ed0681

blk-mq: properly drain stopped queues · 8f5280f4

由 Christoph Hellwig 提交于 6月 13, 2014

If we need to drain a queue we need to run all queues, even if they
are marked stopped to make sure the driver has a chance to error out
on all queued requests.

This fixes surprise removal with scsi-mq.
Reported-by: NBart Van Assche <bvanassche@acm.org>
Tested-by: NBart Van Assche <bvanassche@acm.org>
Signed-off-by: NJens Axboe <axboe@fb.com>

8f5280f4

10 6月, 2014 1 次提交

blk-mq: add timer in blk_mq_start_request · 2b8393b4

由 Ming Lei 提交于 6月 10, 2014

This way will become consistent with non-mq case, also
avoid to update rq->deadline twice for mq.

The comment said: "We do this early, to ensure we are on
the right CPU.", but no percpu stuff is used in blk_add_timer(),
so it isn't necessary. Even when inserting from plug list, there
is no such guarantee at all.
Signed-off-by: NMing Lei <tom.leiming@gmail.com>
Signed-off-by: NJens Axboe <axboe@fb.com>

2b8393b4

09 6月, 2014 1 次提交

blk-mq: always initialize request->start_time · 3ee32372

由 Jens Axboe 提交于 6月 09, 2014

The blk-mq core only initializes this if io stats are enabled, since
blk-mq only reads the field in that case. But drivers could
potentially use it internally, so ensure that we always set it to
the current time when the request is allocated.
Reported-by: NMing Lei <tom.leiming@gmail.com>
Signed-off-by: NJens Axboe <axboe@fb.com>

3ee32372

07 6月, 2014 2 次提交

blk-mq: ->timeout should be cleared in blk_mq_rq_ctx_init() · f6be4fb4

由 Jens Axboe 提交于 6月 06, 2014

It'll be used in blk_mq_start_request() to set a potential timeout
for the request, so clear it to zero at alloc time to ensure that
we know if someone has set it or not.

Fixes random early timeouts on NVMe testing.
Signed-off-by: NJens Axboe <axboe@fb.com>

f6be4fb4

blk-mq: don't allow queue entering for a dying queue · 3b632cf0

由 Keith Busch 提交于 6月 06, 2014

If the queue is going away, don't let new allocs or queueing
happen on it. Go through the normal wait process, and exit with
ENODEV in that case.
Signed-off-by: NKeith Busch <keith.busch@intel.com>
Signed-off-by: NJens Axboe <axboe@fb.com>

3b632cf0

06 6月, 2014 1 次提交

blk-mq: bump max tag depth to 10K tags · a4391c64

由 Jens Axboe 提交于 6月 05, 2014

For some scsi-mq cases, the tag map can be huge. So increase the
max number of tags we support.

Additionally, don't fail with EINVAL if a user requests too many
tags. Warn that the tag depth has been adjusted down, and store
the new value inside the tag_set passed in.
Signed-off-by: NJens Axboe <axboe@fb.com>

a4391c64

05 6月, 2014 1 次提交

blk-mq: let blk_mq_tag_to_rq() take blk_mq_tags as the main parameter · 0e62f51f

由 Jens Axboe 提交于 6月 04, 2014

We currently pass in the hardware queue, and get the tags from there.
But from scsi-mq, with a shared tag space, it's a lot more convenient
to pass in the blk_mq_tags instead as the hardware queue isn't always
directly available. So instead of having to re-map to a given
hardware queue from rq->mq_ctx, just pass in the tags structure.
Signed-off-by: NJens Axboe <axboe@fb.com>

0e62f51f

04 6月, 2014 5 次提交

blk-mq: fix regression from commit · f899fed4

由 Jens Axboe 提交于 6月 04, 2014

When the code was collapsed to avoid duplication, the recent patch
for ensuring that a queue is idled before free was dropped, which was
added by commit 19c5d84f.

Add back the blk_mq_tag_idle(), to ensure we don't leak a reference
to an active queue when it is freed.
Signed-off-by: NJens Axboe <axboe@fb.com>

f899fed4

blk-mq: handle NULL req return from blk_map_request in single queue mode · ff87bcec

由 Jens Axboe 提交于 6月 03, 2014

blk_mq_map_request() can return NULL if we fail entering the queue
(dying, or removed), in which case it has already ended IO on the
bio. So nothing more to do, except just return.
Signed-off-by: NJens Axboe <axboe@fb.com>

ff87bcec

blk-mq: fix sparse warning on missed __percpu annotation · e6cdb092

由 Ming Lei 提交于 6月 03, 2014

'struct blk_mq_ctx' is  __percpu, so add the annotation
and fix the sparse warning reported from Fengguang:

	[block:for-linus 2/3] block/blk-mq.h:75:16: sparse: incorrect
	type in initializer (different address spaces)
Reported-by: Nkbuild test robot <fengguang.wu@intel.com>
Signed-off-by: NMing Lei <tom.leiming@gmail.com>
Signed-off-by: NJens Axboe <axboe@fb.com>

e6cdb092

blk-mq: fix schedule from atomic context · cb96a42c

由 Ming Lei 提交于 6月 01, 2014

blk_mq_put_ctx() has to be called before io_schedule() in
bt_get().

This patch fixes the problem by taking similar approach from
percpu_ida allocation for the situation.
Signed-off-by: NMing Lei <tom.leiming@gmail.com>
Signed-off-by: NJens Axboe <axboe@fb.com>

cb96a42c

blk-mq: move blk_mq_get_ctx/blk_mq_put_ctx to mq private header · 1aecfe48

由 Ming Lei 提交于 6月 01, 2014

The blk-mq tag code need these helpers.
Signed-off-by: NMing Lei <tom.leiming@gmail.com>
Signed-off-by: NJens Axboe <axboe@fb.com>

1aecfe48

31 5月, 2014 2 次提交

blk-mq: push IPI or local end_io decision to __blk_mq_complete_request() · ed851860

由 Jens Axboe 提交于 5月 30, 2014

We have callers outside of the blk-mq proper (like timeouts) that
want to call __blk_mq_complete_request(), so rename the function
and put the decision code for whether to use ->softirq_done_fn
or blk_mq_endio() into __blk_mq_complete_request().

This also makes the interface more logical again.
blk_mq_complete_request() attempts to atomically mark the request
completed, and calls __blk_mq_complete_request() if successful.
__blk_mq_complete_request() then just ends the request.
Signed-off-by: NJens Axboe <axboe@fb.com>

ed851860

blk-mq: remember to start timeout handler for direct queue · feff6894

由 Jens Axboe 提交于 5月 30, 2014

Commit 07068d5b added a direct-to-hw-queue mode, but this mode
needs to remember to add the request timeout handler as well.
Without it, we don't track timeouts for these requests.
Signed-off-by: NJens Axboe <axboe@fb.com>

feff6894

30 5月, 2014 3 次提交

blk-mq: make the sysfs mq/ layout reflect current mappings · 67aec14c

由 Jens Axboe 提交于 5月 30, 2014

Currently blk-mq registers all the hardware queues in sysfs,
regardless of whether it uses them (e.g. they have CPU mappings)
or not. The unused hardware queues lack the cpux/ directories,
and the other sysfs entries (like active, pending, etc) are all
zeroes.

Change this so that sysfs correctly reflects the current mappings
of the hardware queues.
Signed-off-by: NJens Axboe <axboe@fb.com>

67aec14c

blk-mq: blk_mq_tag_to_rq should handle flush request · 22302375

由 Shaohua Li 提交于 5月 30, 2014

flush request is special, which borrows the tag from the parent
request. Hence blk_mq_tag_to_rq needs special handling to return
the flush request from the tag.
Signed-off-by: NShaohua Li <shli@fusionio.com>
Signed-off-by: NJens Axboe <axboe@fb.com>

22302375

blk-mq: request initialization optimizations · 4b570521

由 Jens Axboe 提交于 5月 29, 2014

We currently clear a lot more than we need to, so make that a bit
more clever. Make some of the init dependent on features, like
only setting start_time if we are going to use it.
Signed-off-by: NJens Axboe <axboe@fb.com>

4b570521

29 5月, 2014 3 次提交

block: add queue flag for disabling SG merging · 05f1dd53

由 Jens Axboe 提交于 5月 29, 2014

If devices are not SG starved, we waste a lot of time potentially
collapsing SG segments. Enough that 1.5% of the CPU time goes
to this, at only 400K IOPS. Add a queue flag, QUEUE_FLAG_NO_SG_MERGE,
which just returns the number of vectors in a bio instead of looping
over all segments and checking for collapsible ones.

Add a BLK_MQ_F_SG_MERGE flag so that drivers can opt-in on the sg
merging, if they so desire.
Signed-off-by: NJens Axboe <axboe@fb.com>

05f1dd53

blk-mq: remove alloc_hctx and free_hctx methods · cdef54dd

由 Christoph Hellwig 提交于 5月 28, 2014

There is no need for drivers to control hardware context allocation
now that we do the context to node mapping in common code.
Signed-off-by: NChristoph Hellwig <hch@lst.de>
Signed-off-by: NJens Axboe <axboe@fb.com>

cdef54dd

blk-mq: add file comments and update copyright notices · 75bb4625

由 Jens Axboe 提交于 5月 28, 2014

None of the blk-mq files have an explanatory comment at the top
for what that particular file does. Add that and add appropriate
copyright notices as well.
Signed-off-by: NJens Axboe <axboe@fb.com>

75bb4625

28 5月, 2014 8 次提交

blk-mq: remove blk_mq_alloc_request_pinned · d852564f

由 Christoph Hellwig 提交于 5月 27, 2014

We now only have one caller left and can open code it there in a cleaner
way.
Signed-off-by: NChristoph Hellwig <hch@lst.de>
Signed-off-by: NJens Axboe <axboe@fb.com>

d852564f

blk-mq: do not use blk_mq_alloc_request_pinned in blk_mq_map_request · 793597a6

由 Christoph Hellwig 提交于 5月 27, 2014

We already do a non-blocking allocation in blk_mq_map_request, no need
to repeat it.  Just call __blk_mq_alloc_request to wait directly.
Signed-off-by: NChristoph Hellwig <hch@lst.de>
Signed-off-by: NJens Axboe <axboe@fb.com>

793597a6

blk-mq: remove blk_mq_wait_for_tags · a3bd7756

由 Christoph Hellwig 提交于 5月 27, 2014

The current logic for blocking tag allocation is rather confusing, as we
first allocated and then free again a tag in blk_mq_wait_for_tags, just
to attempt a non-blocking allocation and then repeat if someone else
managed to grab the tag before us.

Instead change blk_mq_alloc_request_pinned to simply do a blocking tag
allocation itself and use the request we get back from it.
Signed-off-by: NChristoph Hellwig <hch@lst.de>
Signed-off-by: NJens Axboe <axboe@fb.com>

a3bd7756

blk-mq: initialize request in __blk_mq_alloc_request · 5dee8577

由 Christoph Hellwig 提交于 5月 27, 2014

Both callers if __blk_mq_alloc_request want to initialize the request, so
lift it into the common path.
Signed-off-by: NChristoph Hellwig <hch@lst.de>
Signed-off-by: NJens Axboe <axboe@fb.com>

5dee8577

blk-mq: merge blk_mq_alloc_reserved_request into blk_mq_alloc_request · 4ce01dd1

由 Christoph Hellwig 提交于 5月 27, 2014

Instead of having two almost identical copies of the same code just let
the callers pass in the reserved flag directly.
Signed-off-by: NChristoph Hellwig <hch@lst.de>
Signed-off-by: NJens Axboe <axboe@fb.com>

4ce01dd1

blk-mq: add helper to insert requests from irq context · 6fca6a61

由 Christoph Hellwig 提交于 5月 28, 2014

Both the cache flush state machine and the SCSI midlayer want to submit
requests from irq context, and the current per-request requeue_work
unfortunately causes corruption due to sharing with the csd field for
flushes.  Replace them with a per-request_queue list of requests to
be requeued.

Based on an earlier test by Ming Lei.
Signed-off-by: NChristoph Hellwig <hch@lst.de>
Reported-by: NMing Lei <tom.leiming@gmail.com>
Tested-by: NMing Lei <tom.leiming@gmail.com>
Signed-off-by: NJens Axboe <axboe@fb.com>

6fca6a61

blk-mq: allow non-softirq completions · 95f09684

由 Jens Axboe 提交于 5月 27, 2014

Right now we export two ways of completing a request:

1) blk_mq_complete_request(). This uses an IPI (if needed) and
   completes through q->softirq_done_fn(). It also works with
   timeouts.

2) blk_mq_end_io(). This completes inline, and ignores any timeout
   state of the request.

Let blk_mq_complete_request() handle non-softirq_done_fn completions
as well, by just completing inline. If a driver has enough completion
ports to place completions correctly, it need not define a
mq_ops->complete() and we can avoid an indirect function call by
doing the completion inline.
Signed-off-by: NJens Axboe <axboe@fb.com>

95f09684

blk-mq: pass in suggested NUMA node to ->alloc_hctx() · f14bbe77

由 Jens Axboe 提交于 5月 27, 2014

Drivers currently have to figure this out on their own, and they
are missing information to do it properly. The ones that did
attempt to do it, do it wrong.

So just pass in the suggested node directly to the alloc
function.
Signed-off-by: NJens Axboe <axboe@fb.com>

f14bbe77

27 5月, 2014 4 次提交

block: only allocate/free mq_usage_counter in blk-mq · 3d2936f4

由 Ming Lei 提交于 5月 27, 2014

The percpu counter is only used for blk-mq, so move
its allocation and free inside blk-mq, and don't
allocate it for legacy queue device.
Signed-off-by: NMing Lei <tom.leiming@gmail.com>
Signed-off-by: NJens Axboe <axboe@fb.com>

3d2936f4

blk-mq: avoid code duplication · 624dbe47

由 Ming Lei 提交于 5月 27, 2014

blk_mq_exit_hw_queues() and blk_mq_free_hw_queues()
are introduced to avoid code duplication.
Signed-off-by: NMing Lei <tom.leiming@gmail.com>
Signed-off-by: NJens Axboe <axboe@fb.com>

624dbe47

blk-mq: fix leak of hctx->ctx_map · 1f9f07e9

由 Ming Lei 提交于 5月 27, 2014

hctx->ctx_map should have been freed inside blk_mq_free_queue().
Signed-off-by: NMing Lei <tom.leiming@gmail.com>
Signed-off-by: NJens Axboe <axboe@fb.com>

1f9f07e9

blk-mq: idle all hardware contexts before freeing a queue · 19c5d84f

由 Christoph Hellwig 提交于 5月 26, 2014

Without this we can leak the active_queues reference if a command is
freed while it is considered active.
Signed-off-by: NChristoph Hellwig <hch@lst.de>
Signed-off-by: NJens Axboe <axboe@fb.com>

19c5d84f

24 5月, 2014 1 次提交

blk-mq: allow setting of per-request timeouts · c22d9d8a

由 Jens Axboe 提交于 5月 23, 2014

Currently blk-mq uses the queue timeout for all requests. But
for some commands, drivers may want to set a specific timeout
for special requests. Allow this to be passed in through
request->timeout, and use it if set.
Signed-off-by: NJens Axboe <axboe@fb.com>

c22d9d8a

23 5月, 2014 1 次提交

blk-mq: split make request handler for multi and single queue · 07068d5b

由 Jens Axboe 提交于 5月 22, 2014

We want slightly different behavior from them:

- On single queue devices, we currently use the per-process plug
  for deferred IO and for merging.

- On multi queue devices, we don't use the per-process plug, but
  we want to go straight to hardware for SYNC IO.

Split blk_mq_make_request() into a blk_sq_make_request() for single
queue devices, and retain blk_mq_make_request() for multi queue
devices. Then we don't need multiple checks for q->nr_hw_queues
in the request mapping.
Signed-off-by: NJens Axboe <axboe@fb.com>

07068d5b

22 5月, 2014 2 次提交

blk-mq: save memory by freeing requests on unused hardware queues · 484b4061

由 Jens Axboe 提交于 5月 21, 2014

Depending on the topology of the machine and the number of queues
exposed by a device, we can end up in a situation where some of
the hardware queues are unused (as in, they don't map to any
software queues). For this case, free up the memory used by the
request map, as we will not use it. This can be a substantial
amount of memory, depending on the number of queues vs CPUs and
the queue depth of the device.
Signed-off-by: NJens Axboe <axboe@fb.com>

484b4061

blk-mq: allow the hctx cpu hotplug notifier to return errors · e814e71b

由 Jens Axboe 提交于 5月 21, 2014

Prepare this for the next patch which adds more smarts in the
plugging logic, so that we can save some memory.
Signed-off-by: NJens Axboe <axboe@fb.com>

e814e71b

21 5月, 2014 2 次提交

blk-mq: Micro-optimize blk_queue_nomerges() check · da41a589

由 Robert Elliott 提交于 5月 20, 2014

In blk_mq_make_request(), do the blk_queue_nomerges() check
outside the call to blk_attempt_plug_merge() to eliminate
function call overhead when nomerges=2 (disabled)
Signed-off-by: NRobert Elliott <elliott@hp.com>
Signed-off-by: NJens Axboe <axboe@fb.com>

da41a589

blk-mq: initialize q->nr_requests after calling blk_queue_make_request() · eba71768

由 Jens Axboe 提交于 5月 20, 2014

blk_queue_make_requests() overwrites our set value for q->nr_requests,
turning it into the default of 128. Set this appropriately after
initializing queue values in blk_queue_make_request().
Signed-off-by: NJens Axboe <axboe@fb.com>

eba71768

openeuler / Kernel 1 年多 前同步成功

openeuler / Kernel
1 年多前同步成功