提交 · 1c65df3d7b1b311a73f5636a4ad6611f067baf1e · openeuler / raspberrypi-kernel

18 6月, 2014 1 次提交

floppy: format block0 read error message properly · 1c65df3d

由 Jiri Kosina 提交于 6月 18, 2014

In case reading of block 0 fails, line without trailing newline
is printed causing dmesg to look horrible.
Signed-off-by: NJiri Kosina <jkosina@suse.cz>

1c65df3d

28 5月, 2014 1 次提交

floppy: do not corrupt bio.bi_flags when reading block 0 · 6314a108

由 Jiri Kosina 提交于 5月 28, 2014

Commit 41a55b4d ("floppy: silence warning during disk test") caused
bio.bi_flags being overwritten, and its initialization to BIO_UPTODATE
in bio_init() to be lost.

This was unnoticed until 7b7b68bb ("floppy: bail out in open() if
drive is not responding to block0 read"), because the error value wasn't
checked for in the bio completion callback.

Now we are actually looking at the error, and the loss of BIO_UPTODATE
causes EIO to be wrongly passed to the callback, which confuses the
FD_OPEN_SHOULD_FAIL_BIT logic.

Fix this by not destroying previous value of bi_flags when setting
BIO_QUIET.

Cc: Stephen Hemminger <shemminger@vyatta.com>
Reported-by: NTakashi Iwai <tiwai@suse.de>
Signed-off-by: NJiri Kosina <jkosina@suse.cz>

6314a108

21 5月, 2014 1 次提交

mtip32xx: move error handling to service thread · 9b204fbf

由 Asai Thambi S P 提交于 5月 20, 2014

Move error handling to service thread, and use mtip_set_timeout()
to set timeouts for HDIO_DRIVE_TASK and HDIO_DRIVE_CMD IOCTL commands.
Signed-off-by: NSelvan Mani <smani@micron.com>
Signed-off-by: NAsai Thambi S P <asamymuthupa@micron.com>
Signed-off-by: NJens Axboe <axboe@fb.com>

9b204fbf

16 5月, 2014 1 次提交

virtio_blk: fix race between start and stop queue · 0c29e93e

由 Ming Lei 提交于 5月 16, 2014

When there isn't enough vring descriptor for adding to vq,
blk-mq will be put as stopped state until some of pending
descriptors are completed & freed.

Unfortunately, the vq's interrupt may come just before
blk-mq's BLK_MQ_S_STOPPED flag is set, so the blk-mq will
still be kept as stopped even though lots of descriptors
are completed and freed in the interrupt handler. The worst
case is that all pending descriptors are freed in the
interrupt handler, and the queue is kept as stopped forever.

This patch fixes the problem by starting/stopping blk-mq
with holding vq_lock.

Cc: Jens Axboe <axboe@kernel.dk>
Cc: Rusty Russell <rusty@rustcorp.com.au>
Signed-off-by: NMing Lei <tom.leiming@gmail.com>
Signed-off-by: NJens Axboe <axboe@fb.com>

0c29e93e

14 5月, 2014 3 次提交

mtip32xx: stop block hardware queues before quiescing IO · 9acf03cf

由 Jens Axboe 提交于 5月 14, 2014

We need to stop the block layer queues to prevent new "normal"
IO from entering the driver, while we wait for existing commands
to finish.
Signed-off-by: NJens Axboe <axboe@fb.com>

9acf03cf

mtip32xx: blk_mq_init_queue() returns an ERR_PTR · a8a642cc

由 Dan Carpenter 提交于 5月 14, 2014

We changed this from blk_alloc_queue_node() to blk_mq_init_queue() so
the check needs to be updated as well.

Fixes: ffc771b3 ('mtip32xx: convert to use blk-mq')
Signed-off-by: NDan Carpenter <dan.carpenter@oracle.com>
Signed-off-by: NJens Axboe <axboe@fb.com>

a8a642cc

mtip32xx: convert to use blk-mq · ffc771b3

由 Jens Axboe 提交于 5月 09, 2014

This rips out timeout handling, requeueing, etc in converting
it to use blk-mq instead.
Acked-by: NAsai Thambi S P <asamymuthupa@micron.com>
Signed-off-by: NJens Axboe <axboe@fb.com>

ffc771b3

06 5月, 2014 12 次提交

cdrom: Remove unnecessary prototype for cdrom_get_disc_info · bd6f0bba

由 Joe Perches 提交于 5月 04, 2014

Move the function to the proper spot instead.
Signed-off-by: NJoe Perches <joe@perches.com>
Signed-off-by: NJens Axboe <axboe@fb.com>

bd6f0bba

cdrom: Remove unnecessary prototype for cdrom_mrw_exit · 40569c61

由 Joe Perches 提交于 5月 04, 2014

Move the function to appropriate locations instead.
Signed-off-by: NJoe Perches <joe@perches.com>
Signed-off-by: NJens Axboe <axboe@fb.com>

40569c61

cdrom: Remove cdrom_count_tracks prototype · a803393b

由 Joe Perches 提交于 5月 04, 2014

Move function to proper location instead.
Fix whitespace and embedded if too.
Signed-off-by: NJoe Perches <joe@perches.com>
Signed-off-by: NJens Axboe <axboe@fb.com>

a803393b

cdrom: Remove cdrom_get_next_writeable prototype · dac1c5cf

由 Joe Perches 提交于 5月 04, 2014

Move the function to the right spot instead.
Signed-off-by: NJoe Perches <joe@perches.com>
Signed-off-by: NJens Axboe <axboe@fb.com>

dac1c5cf

cdrom: Remove cdrom_get_last_written prototype · e1e60fda

由 Joe Perches 提交于 5月 04, 2014

Move the function instead.
Signed-off-by: NJoe Perches <joe@perches.com>
Signed-off-by: NJens Axboe <axboe@fb.com>

e1e60fda

J
cdrom: Move mmc_ioctls above cdrom_ioctl to remove unnecessary prototype · 2e9aa08a
由 Joe Perches 提交于 5月 04, 2014
```
Neaten the spacing too.
Signed-off-by: NJoe Perches <joe@perches.com>
Signed-off-by: NJens Axboe <axboe@fb.com>
```
2e9aa08a

cdrom: Remove unnecessary sanitize_format prototype · 1237d1ee

由 Joe Perches 提交于 5月 04, 2014

It's defined below without being called.
Signed-off-by: NJoe Perches <joe@perches.com>
Signed-off-by: NJens Axboe <axboe@fb.com>

1237d1ee

cdrom: Remove unnecessary check_for_audio_disc prototype · b62d6e82

由 Joe Perches 提交于 5月 04, 2014

The actual static is defined below it but not used until later.
Signed-off-by: NJoe Perches <joe@perches.com>
Signed-off-by: NJens Axboe <axboe@fb.com>

b62d6e82

cdrom: Remove prototype for open_for_data · 82b91540

由 Joe Perches 提交于 5月 04, 2014

Move static function to the appropriate place to remove
the now unnecessary prototype.
Signed-off-by: NJoe Perches <joe@perches.com>
Signed-off-by: NJens Axboe <axboe@fb.com>

82b91540

cdrom: Remove obfuscating IOCTL_IN and IOCTL_OUT macros · a09c391d

由 Joe Perches 提交于 5月 04, 2014

Macros with hidden control flow aren't nice.
Just use copy_to/from_user directly instead.
Signed-off-by: NJoe Perches <joe@perches.com>
Signed-off-by: NJens Axboe <axboe@fb.com>

a09c391d

cdrom: Remove unused CHECKAUDIO macro · e3b6b9ef

由 Joe Perches 提交于 5月 04, 2014

It's unused, make it disappear.
Signed-off-by: NJoe Perches <joe@perches.com>
Signed-off-by: NJens Axboe <axboe@fb.com>

e3b6b9ef

cdrom: convert cdinfo to cd_dbg · 5944b2ce

由 Joe Perches 提交于 5月 04, 2014

It's a debugging message, mark it so.
Signed-off-by: NJoe Perches <joe@perches.com>
Signed-off-by: NJens Axboe <axboe@fb.com>

5944b2ce

01 5月, 2014 21 次提交

block: null_blk: fix use after free · fc27691f

由 Ming Lei 提交于 5月 01, 2014

entry(cmd->ll_list) may belong to new request once end_cmd()
returns, so fix the bug with the patch.

Without the change, it is easy to observe oops when
doing null_blk(timer) test.
Signed-off-by: NMing Lei <tom.leiming@gmail.com>
Signed-off-by: NJens Axboe <axboe@fb.com>

fc27691f

drbd: use list_first_entry_or_null in first_peer_device/first_connection · ec4a3407

由 Lars Ellenberg 提交于 4月 28, 2014

If there are no peer_devices or connections, I'd rather have NULL
than some "arbitrary" address pretending to point to a struct.

Helps to avoid hard to debug symptoms, in case we ever try to use
and dereference a drbd_connection or drbd_peer_device
where we in fact don't have any connection at all.
Signed-off-by: NPhilipp Reisner <philipp.reisner@linbit.com>
Signed-off-by: NLars Ellenberg <lars.ellenberg@linbit.com>
Signed-off-by: NJens Axboe <axboe@fb.com>

ec4a3407

drbd: Allow attaching of a newly created device to any backing device · babea49e

由 Philipp Reisner 提交于 4月 28, 2014

A newly created device was never exposed before, i.e. has a
exposed_data_uuid of 0. Then it is valid to attach to any current_uuid
of a backing device (of course also to a newly created one (4))
Signed-off-by: NPhilipp Reisner <philipp.reisner@linbit.com>
Signed-off-by: NLars Ellenberg <lars.ellenberg@linbit.com>
Signed-off-by: NJens Axboe <axboe@fb.com>

babea49e

drbd: Test cstate while holding req_lock · 02df6fe1

由 Philipp Reisner 提交于 4月 28, 2014

In case a connection transitions into C_TIMEOUT within the timer
function (request_timer_fn()) we need to make sure that the receiver
thread (potentially running on a different CPU) sees the updated
cstate later on.
Signed-off-by: NPhilipp Reisner <philipp.reisner@linbit.com>
Signed-off-by: NLars Ellenberg <lars.ellenberg@linbit.com>
Signed-off-by: NJens Axboe <axboe@fb.com>

02df6fe1

drbd: use blk_set_stacking_limits() · c1b3156f

由 Philipp Reisner 提交于 4月 28, 2014

...instead directly assigning to q->limits.discard_zeroes_data
Signed-off-by: NPhilipp Reisner <philipp.reisner@linbit.com>
Signed-off-by: NLars Ellenberg <lars.ellenberg@linbit.com>
Signed-off-by: NJens Axboe <axboe@fb.com>

c1b3156f

drbd: evaluate disk and network timeout on different requests · 08535466

由 Lars Ellenberg 提交于 4月 28, 2014

Just because it is the oldest not yet completed request
does not make it the oldest request waiting for disk.
Or waiting for the peer.

And we completely missed already completed requests
that would still hold references to activity log extents,
waiting only for the barrier ack.

Find two oldest not yet completely processed requests,
one that is still waiting for local completion,
and one that is still waiting for some response from the peer.
These may or may not be the same request object.

Then separately apply the network and disk timeouts, respectively.
Signed-off-by: NPhilipp Reisner <philipp.reisner@linbit.com>
Signed-off-by: NLars Ellenberg <lars.ellenberg@linbit.com>
Signed-off-by: NJens Axboe <axboe@fb.com>

08535466

drbd: Fix a hole in the challange-response connection authentication · 67cca286

由 Philipp Reisner 提交于 4月 28, 2014

In the implementation as it was, the two peers sent each other
a challenge, and expects the challenge hashed with the shared
secret back.

A attacker could simply wait for the challenge of the peer, and
send the same challenge back. Then it waits for the response, and
sends the same response back.

Prevent this by not accepting a challenge from the peer that is
the same as the challenge sent to the peer.
Signed-off-by: NPhilipp Reisner <philipp.reisner@linbit.com>
Signed-off-by: NLars Ellenberg <lars.ellenberg@linbit.com>
Signed-off-by: NJens Axboe <axboe@fb.com>

67cca286

drbd: always implicitly close last epoch when idle · f9c78128

由 Lars Ellenberg 提交于 4月 28, 2014

Once our sender thread needs to wait_for_work(),
and actually needs to schedule(), just before we do that,
we already check if it is useful to implicitly close the last epoch.

The condition was too strict: only implicitly close the epoch,
if there have been no new (write) requests at all.

The assumption was that if there were new requests, they would
always be communicated one way or another, and would send necessary
epoch separating barriers explicitly.

This is not always true, e.g. when becoming diskless,
or while explicitly starting a full resync.

The last communicated epoch could stay open for a long time,
locking down corresponding activity log extents.

It is safe to always implicitly send that last barrier, as soon as we
determin that there cannot be more requests in the last communicated
epoch, even if there have been (uncommunicated) new requests in new
epochs meanwhile.
Signed-off-by: NPhilipp Reisner <philipp.reisner@linbit.com>
Signed-off-by: NLars Ellenberg <lars.ellenberg@linbit.com>
Signed-off-by: NJens Axboe <axboe@fb.com>

f9c78128

drbd: add back some fairness to AL transactions · e4d7d6f4

由 Lars Ellenberg 提交于 4月 28, 2014

When batching more updates to the activity log into single transactions,
we lost the ability for new requests to force themselves into the active
set: all preparation steps became non-blocking, and if all currently
hot extents keep busy, they could starve out new incoming requests
to cold extents for quite a while.

This can only happen if your IO backend accepts more IO operations per
average DRBD replication round trip time than you have al-extents
configured.

If we have incoming requests to cold extents,
at least do one blocking update per transaction.

In an artificial worst-case workload on SSD with an asynchronous 600 ms
replication link, with al-extents = 7 (the minimum we allow), and
concurrent full resynch, without this patch, some write requests have
been observed to be starved for 40 seconds.
With this patch, application observed a worst case latency of twice the
replication round trip time.
Signed-off-by: NPhilipp Reisner <philipp.reisner@linbit.com>
Signed-off-by: NLars Ellenberg <lars.ellenberg@linbit.com>
Signed-off-by: NJens Axboe <axboe@fb.com>

e4d7d6f4

drbd: keep max-bio size during detach/attach on disconnected primary · fa090e70

由 Lars Ellenberg 提交于 4月 28, 2014

We want to store in persistent meta data what the peer DRBD can handle,
which, due to spreading requests to multiple bios,
may be more than its backing device can handle.

Otherwise, if a disconnected Primary temporarily loses access to its local data
as well, we may accidentally shrink the max-bio setting, portentially causing
already assembled, but not yet processed, application bios to be spuriously
failed due to device limits.
Signed-off-by: NPhilipp Reisner <philipp.reisner@linbit.com>
Signed-off-by: NLars Ellenberg <lars.ellenberg@linbit.com>
Signed-off-by: NJens Axboe <axboe@fb.com>

fa090e70

drbd: fix a race between start_resync and send_and_submit · 074f4afe

由 Lars Ellenberg 提交于 4月 28, 2014

In the drbd make request function, specifically in
drbd_send_and_submit(), we decide whether we want to send the actual
write request, or only a "set this block out of sync" information.

We do so based on the current connection state, while holding the req_lock.
The connection state is not supposed to change while holding the req_lock.

But in drbd_start_resync, we did change that state anyways,
while only holding the global_state_lock, which is enough to change
sync-after dependencies (paused vs active resync), but
not good enough to change the connection state.

Fix: in drbd_start_resync, first grab the req_lock to serialize with
drbd_send_and_submit(), before grabbing the global_state_lock
to be able to evaluate the sync-after dependencies.
Signed-off-by: NPhilipp Reisner <philipp.reisner@linbit.com>
Signed-off-by: NLars Ellenberg <lars.ellenberg@linbit.com>
Signed-off-by: NJens Axboe <axboe@fb.com>

074f4afe

drbd: Enable QUEUE_FLAG_DISCARD only if the peer can recieve P_TRIM · 20c68fde

由 Lars Ellenberg 提交于 4月 28, 2014

Allow the user of REQ_DISCARD.
Signed-off-by: NPhilipp Reisner <philipp.reisner@linbit.com>
Signed-off-by: NLars Ellenberg <lars.ellenberg@linbit.com>
Signed-off-by: NJens Axboe <axboe@fb.com>

20c68fde

drbd: prepare sending side for REQ_DISCARD · 2f632aeb

由 Lars Ellenberg 提交于 4月 28, 2014

Note that I do NOT call __drbd_chk_io_error for failed REQ_DISCARD.
That may be wrong, though, or needs to differ between EOPNOTSUPP and
other errors...
Signed-off-by: NPhilipp Reisner <philipp.reisner@linbit.com>
Signed-off-by: NLars Ellenberg <lars.ellenberg@linbit.com>
Signed-off-by: NJens Axboe <axboe@fb.com>

2f632aeb

drbd: prepare receiving side for REQ_DISCARD · a0fb3c47

由 Lars Ellenberg 提交于 4月 28, 2014

If the receiver needs to serve a discard request on a queue that does
not announce to be discard cabable, it falls back to do synchronous
blkdev_issue_zeroout().

We expect only "reasonably" large (up to one activity log extent?)
discard requests.

We do this to not to not block the receiver for too long in this
fallback code path, and to not set/clear too many bits inside one
spinlock_irq_save() in drbd_set_in_sync/drbd_set_out_of_sync,
Signed-off-by: NPhilipp Reisner <philipp.reisner@linbit.com>
Signed-off-by: NLars Ellenberg <lars.ellenberg@linbit.com>
Signed-off-by: NJens Axboe <axboe@fb.com>

a0fb3c47

drbd: allow parallel promote/demote actions · 9e276872

由 Lars Ellenberg 提交于 4月 28, 2014

We plan to use genl_family->parallel_ops = true in the future,
but need to review all possible interactions first.

For now, only selectively drop genl_lock() in drbd_set_role(),
instead serializing on our own internal resource->conf_update mutex.

We now can be promoted/demoted on many resources in parallel,
which may significantly improve cluster failover times
when fencing is required.
Signed-off-by: NPhilipp Reisner <philipp.reisner@linbit.com>
Signed-off-by: NLars Ellenberg <lars.ellenberg@linbit.com>
Signed-off-by: NJens Axboe <axboe@fb.com>

9e276872

drbd: perpare for genetlink parallel_ops · a910b123

由 Lars Ellenberg 提交于 4月 28, 2014

Because all administrative requests via genetlink have been globally
serialized via genl_lock(), we used to have one static struct
drbd_config_context "admin context".

Move this on-stack to the respective callback functions.

This will allow us to selectively drop the genl_lock()
(or use genl_family->parallel_ops) in the future.
Signed-off-by: NPhilipp Reisner <philipp.reisner@linbit.com>
Signed-off-by: NLars Ellenberg <lars.ellenberg@linbit.com>
Signed-off-by: NJens Axboe <axboe@fb.com>

a910b123

drbd: Do not BUG() when connection breaks in a special way · 88ea685d

由 Philipp Reisner 提交于 4月 28, 2014

When a 'cluster wide' disconnect executes, the result comes back
from the peer, and immediately after that the connection breaks
then _conn_rq_cond() reported back SS_CW_SUCCESS.
Therefore _conn_request_state() calls conn_set_state(), which
has a BUG() in it.
The BUG() is hit because conn_is_valid_transition() does not like
the transaction. Which goes back to is_valid_soft_transition()
returning SS_OUTDATE_WO_CONN.

This fix is to consider an error reported by is_valid_soft_transition()
even when the peer agreed to the transaction.
Signed-off-by: NPhilipp Reisner <philipp.reisner@linbit.com>
Signed-off-by: NLars Ellenberg <lars.ellenberg@linbit.com>
Signed-off-by: NJens Axboe <axboe@fb.com>

88ea685d

drbd: don't let application IO pre-empt resync too often · e8299874

由 Lars Ellenberg 提交于 4月 28, 2014

Before, application IO could pre-empt resync activity
for up to hardcoded 20 seconds per resync request.
A very busy server could throttle the effective resync bandwidth
down to one request per 20 seconds.

Now, we only let application IO pre-empt resync traffic
while the current resync rate estimate is above c-min-rate.

If you disable the c-min-rate throttle feature (set c-min-rate = 0),
application IO will no longer pre-empt resync traffic at all.
Signed-off-by: NPhilipp Reisner <philipp.reisner@linbit.com>
Signed-off-by: NLars Ellenberg <lars.ellenberg@linbit.com>
Signed-off-by: NJens Axboe <axboe@fb.com>

e8299874

drbd: fix potential distributed deadlock during verify or resync · 0e49d7b0

由 Lars Ellenberg 提交于 4月 28, 2014

If max-buffers and socket buffer sizes are "too small" for the chosen
resync rate, this could lead potentially lead to a distributed deadlock,
which may or may not resolve itself via the "ko-count" and request
timeout mechanism, or could be resolved by forced disconnect.

One option to deal with this is proper configuration:
use larger max-buffer and socket buffers settings,
or reduce the resync rate.

But even with bad configuration we should not deadlock,
but "gracefully" recover.

The issue is avoided by using only up to max-buffers/2 for resync
requests, and by using max-buffers not as a hard limit for data buffer
allocations, but as a throttle threshold only.
Signed-off-by: NPhilipp Reisner <philipp.reisner@linbit.com>
Signed-off-by: NLars Ellenberg <lars.ellenberg@linbit.com>
Signed-off-by: NJens Axboe <axboe@fb.com>

0e49d7b0

drbd: resync: fix too large bursts for very slow rates · 6377b923

由 Lars Ellenberg 提交于 4月 28, 2014

While merging adjacent dirty blocks into resync requests,
the resync rate throttle was disregarded.
For very low resync rates, the effective rate may have exceeded
the intended rate by a larger margin.
Signed-off-by: NPhilipp Reisner <philipp.reisner@linbit.com>
Signed-off-by: NLars Ellenberg <lars.ellenberg@linbit.com>
Signed-off-by: NJens Axboe <axboe@fb.com>

6377b923

drbd: fix stalled resync detection in /proc/drbd · 9ae47260

由 Lars Ellenberg 提交于 4月 28, 2014

If we don't make resync or verify progress for "too long",
we want to flag it as "stalled".

Since 2010, "use rolling marks for resync speed calculation"
this "too long" was wrong by a factor of HZ.
With HZ 250, it would have been flagged as stalled
after 100 minutes.

Hardcode 3 minutes instead.
Signed-off-by: NPhilipp Reisner <philipp.reisner@linbit.com>
Signed-off-by: NLars Ellenberg <lars.ellenberg@linbit.com>
Signed-off-by: NJens Axboe <axboe@fb.com>

9ae47260