提交 · 3c3f2e32effd4c6acc3a9434bd7eecb0af653d89 · openeuler / raspberrypi-kernel

23 3月, 2010 1 次提交

ceph: fix connection fault con_work reentrancy problem · 3c3f2e32

由 Sage Weil 提交于 3月 18, 2010

The messenger fault was clearing the BUSY bit, for reasons unclear. This
made it possible for the con->ops->fault function to reopen the connection,
and requeue work in the workqueue--even though the current thread was
already in con_work.

This avoids a problem where the client busy loops with connection failures
on an unreachable OSD, but doesn't address the root cause of that problem.
Signed-off-by: NSage Weil <sage@newdream.net>

3c3f2e32

21 3月, 2010 1 次提交

ceph: fix authenticator timeout · 63733a0f

由 Sage Weil 提交于 3月 15, 2010

We were failing to reconnect to services due to an old authenticator, even
though we had the new ticket, because we weren't properly retrying the
connect handshake, because we were calling an old/incorrect helper that
left in_base_pos incorrect.  The result was a failure to reconnect to the
OSD or MDS (with an authentication error) if the MDS restarted after the
service had been up a few hours (long enough for the original authenticator
to be invalid).  This was only a problem if the AUTH_X authentication was
enabled.

Now that the 'negotiate' and 'connect' stages are fully separated, we
should use the prepare_read_connect() helper instead, and remove the
obsolete one.
Signed-off-by: NSage Weil <sage@newdream.net>

63733a0f

02 3月, 2010 2 次提交

ceph: reset front len on return to msgpool; BUG on mismatched front iov · 3ca02ef9

由 Sage Weil 提交于 3月 01, 2010

Reset msg front len when a message is returned to the pool: the caller
may have changed it.

BUG if we try to send a message with a hdr.front_len that doesn't match
the front iov.
Signed-off-by: NSage Weil <sage@newdream.net>

3ca02ef9

ceph: reset bits on connection close · 1679f876

由 Sage Weil 提交于 2月 26, 2010

Clear LOSSYTX bit, so that if/when we reconnect, said reconnect
will retry on failure.

Clear _PENDING bits too, to avoid polluting subsequent
connection state.

Drop unused REGISTERED bit.
Signed-off-by: NSage Weil <sage@newdream.net>

1679f876

26 2月, 2010 2 次提交

ceph: fix connection fault STANDBY check · e80a52d1

由 Sage Weil 提交于 2月 25, 2010

Move any out_sent messages to out_queue _before_ checking if
out_queue is empty and going to STANDBY, or else we may drop
something that was never acked.

And clean up the code a bit (less goto).
Signed-off-by: NSage Weil <sage@newdream.net>

e80a52d1

ceph: invalidate_authorizer without con->mutex held · 161fd65a

由 Sage Weil 提交于 2月 25, 2010

This fixes lock ABBA inversion, as the ->invalidate_authorizer()
op may need to take a lock (or even call back into the
messenger).
Signed-off-by: NSage Weil <sage@newdream.net>

161fd65a

24 2月, 2010 1 次提交

ceph: fix up unexpected message handling · 5b3a4db3

由 Sage Weil 提交于 2月 19, 2010

Fix skipping of unexpected message types from osd, mon.

Clean up pr_info and debug output.
Signed-off-by: NSage Weil <sage@newdream.net>

5b3a4db3

17 2月, 2010 2 次提交

ceph: cancel delayed work when closing connection · 91e45ce3

由 Sage Weil 提交于 2月 15, 2010

This ensures that if/when we reopen the connection, we can requeue work on
the connection immediately, without waiting for an old timer to expire.
Queue new delayed work inside con->mutex to avoid any race.

This fixes problems with clients failing to reconnect to the MDS due to
the client_reconnect message arriving too late (due to waiting for an old
delayed work timeout to expire).
Signed-off-by: NSage Weil <sage@newdream.net>

91e45ce3

ceph: allow connection to be reopened by fault callback · e2663ab6

由 Sage Weil 提交于 2月 16, 2010

Fix the messenger to allow a ceph_con_open() during the fault callback.
Previously the work wasn't getting queued on the connection because the
fault path avoids requeued work (normally spurious).  Loop on reopening by
checking for the OPENING state bit.

This fixes OSD reconnects when a TCP connection drops.
Signed-off-by: NSage Weil <sage@newdream.net>

e2663ab6

14 2月, 2010 1 次提交

ceph: fix msgr to keep sent messages until acked · 6c5d1a49

由 Sage Weil 提交于 2月 13, 2010

The test was backwards from commit b3d1dbbd: keep the message if the
connection _isn't_ lossy.  This allows the client to continue when the
TCP connection drops for some reason (network glitch) but both ends
survive.
Signed-off-by: NSage Weil <sage@newdream.net>

6c5d1a49

11 2月, 2010 1 次提交

ceph: allow renewal of auth credentials · 9bd2e6f8

由 Sage Weil 提交于 2月 02, 2010

Add infrastructure to allow the mon_client to periodically renew its auth
credentials.  Also add a messenger callback that will force such a renewal
if a peer rejects our authenticator.
Signed-off-by: NYehuda Sadeh <yehuda@hq.newdream.net>
Signed-off-by: NSage Weil <sage@newdream.net>

9bd2e6f8

30 1月, 2010 1 次提交

ceph: include type in ceph_entity_addr, filepath · ac8839d7

由 Sage Weil 提交于 1月 27, 2010

Include a type/version in ceph_entity_addr and filepath.  Include extra
byte in filepath encoding as necessary.
Signed-off-by: NSage Weil <sage@newdream.net>

ac8839d7

26 1月, 2010 4 次提交

ceph: keep reserved replies on the request structure · 0d59ab81

由 Yehuda Sadeh 提交于 1月 13, 2010

This includes treating all the data preallocation and revokation
at the same place, not having to have a special case for
the reserved pages.
Signed-off-by: NYehuda Sadeh <yehuda@hq.newdream.net>

0d59ab81

ceph: alloc message data pages and check if tid exists · 0547a9b3

由 Yehuda Sadeh 提交于 1月 11, 2010

Now doing it in the same callback that is also responsible for
allocating the 'front' part of the message. If we get a message
that we haven't got a corresponding tid for, mark it for skipping.

Moving the mutex unlock/lock from the osd alloc_msg callback
to the calling function in the messenger.
Signed-off-by: NYehuda Sadeh <yehuda@hq.newdream.net>

0547a9b3

Y
ceph: refactor messages data section allocation · 9d7f0f13
由 Yehuda Sadeh 提交于 1月 11, 2010
```
Signed-off-by: NYehuda Sadeh <yehuda@hq.newdream.net>
```
9d7f0f13

ceph: allocate middle of message before stating to read · 2450418c

由 Yehuda Sadeh 提交于 1月 08, 2010

Both front and middle parts of the message are now being
allocated at the ceph_alloc_msg().
Signed-off-by: NYehuda Sadeh <yehuda@hq.newdream.net>

2450418c

15 1月, 2010 1 次提交

ceph: remove unused erank field · 103e2d3a

由 Sage Weil 提交于 1月 07, 2010

The ceph_entity_addr erank field is obsolete; remove it.  Get rid of
trivial addr comparison helpers while we're at it.
Signed-off-by: NSage Weil <sage@newdream.net>

103e2d3a

24 12月, 2009 4 次提交

ceph: support ceph_pagelist for message payload · 58bb3b37

由 Sage Weil 提交于 12月 23, 2009

The ceph_pagelist is a simple list of whole pages, strung together via
their lru list_head.  It facilitates encoding to a "buffer" of unknown
size.  Allow its use in place of the ceph_msg page vector.

This will be used to fix the huge buffer preallocation woes of MDS
reconnection.
Signed-off-by: NSage Weil <sage@newdream.net>

58bb3b37

ceph: add feature bits to connection handshake (protocol change) · 04a419f9

由 Sage Weil 提交于 12月 23, 2009

Define supported and required feature set.  Fail connection if the server
requires features we do not support (TAG_FEATURES), or if the server does
not support features we require.
Signed-off-by: NSage Weil <sage@newdream.net>

04a419f9

ceph: control access to page vector for incoming data · 350b1c32

由 Sage Weil 提交于 12月 22, 2009

When we issue an OSD read, we specify a vector of pages that the data is to
be read into. The request may be sent multiple times, to multiple OSDs, if
the osdmap changes, which means we can get more than one reply.

Only read data into the page vector if the reply is coming from the
OSD we last sent the request to. Keep track of which connection is using
the vector by taking a reference. If another connection was already
using the vector before and a new reply comes in on the right connection,
revoke the pages from the other connection.
Signed-off-by: NSage Weil <sage@newdream.net>

350b1c32

ceph: use connection mutex to protect read and write stages · ec302645

由 Sage Weil 提交于 12月 22, 2009

Use a single mutex (previously out_mutex) to protect both read and write
activity from concurrent ceph_con_* calls.  Drop the mutex when doing
callbacks to avoid nested locking (the callback may need to call something
like ceph_con_close).
Signed-off-by: NSage Weil <sage@newdream.net>

ec302645

22 12月, 2009 7 次提交

Y
ceph: remove unaccessible code · 169e16ce
由 Yehuda Sadeh 提交于 12月 16, 2009
```
Signed-off-by: NYehuda Sadeh <yehuda@hq.newdream.net>
```
169e16ce

ceph: plug leak of incoming message during connection fault/close · cf3e5c40

由 Sage Weil 提交于 12月 11, 2009

If we explicitly close a connection, or there is a socket error, we need
to drop any partially received message.
Signed-off-by: NSage Weil <sage@newdream.net>

cf3e5c40

S
ceph: hex dump corrupt server data to KERN_DEBUG · 9ec7cab1
由 Sage Weil 提交于 12月 14, 2009
```
Also, print fsid using standard format, NOT hex dump.
Signed-off-by: NSage Weil <sage@newdream.net>
```
9ec7cab1

ceph: don't save sent messages on lossy connections · b3d1dbbd

由 Sage Weil 提交于 12月 14, 2009

For lossy connections we drop all state on socket errors, so there is no
reason to keep sent ceph_msg's around.
Signed-off-by: NSage Weil <sage@newdream.net>

b3d1dbbd

ceph: detect lossy state of connection · 92ac41d0

由 Sage Weil 提交于 12月 14, 2009

The server indicates whether a connection is lossy; set our LOSSYTX bit
appropriately.  Do not set lossy bit on outgoing connections.
Signed-off-by: NSage Weil <sage@newdream.net>

92ac41d0

S
ceph: plug msg leak in con_fault · 5e095e8b
由 Sage Weil 提交于 12月 14, 2009
```
Signed-off-by: NSage Weil <sage@newdream.net>
```
5e095e8b

ceph: carry explicit msg reference for currently sending message · c86a2930

由 Sage Weil 提交于 12月 14, 2009

Carry a ceph_msg reference for connection->out_msg.  This will allow us to
make out_sent optional.
Signed-off-by: NSage Weil <sage@newdream.net>

c86a2930

08 12月, 2009 2 次提交

S
ceph: use kref for ceph_msg · c2e552e7
由 Sage Weil 提交于 12月 07, 2009
```
Signed-off-by: NSage Weil <sage@newdream.net>
```
c2e552e7

ceph: simplify ceph_buffer interface · b6c1d5b8

由 Sage Weil 提交于 12月 07, 2009

We never allocate the ceph_buffer and buffer separtely, so use a single
constructor.

Disallow put on NULL buffer; make the caller check.
Signed-off-by: NSage Weil <sage@newdream.net>

b6c1d5b8

21 11月, 2009 1 次提交

ceph: reset msgr backoff during open, not after successful handshake · 03c677e1

由 Sage Weil 提交于 11月 20, 2009

Reset the backoff delay when we reopen the connection, so that the delays
for any initial connection problems are reasonable. We were resetting only
after a successful handshake, which was of limited utility.
Signed-off-by: NSage Weil <sage@newdream.net>

03c677e1

19 11月, 2009 2 次提交

ceph: negotiate authentication protocol; implement AUTH_NONE protocol · 4e7a5dcd

由 Sage Weil 提交于 11月 18, 2009

When we open a monitor session, we send an initial AUTH message listing
the auth protocols we support, our entity name, and (possibly) a previously
assigned global_id.  The monitor chooses a protocol and responds with an
initial message.

Initially implement AUTH_NONE, a dummy protocol that provides no security,
but works within the new framework.  It generates 'authorizers' that are
used when connecting to (mds, osd) services that simply state our entity
name and global_id.

This is a wire protocol change.
Signed-off-by: NSage Weil <sage@newdream.net>

4e7a5dcd

ceph: remove unnecessary ceph_con_shutdown · 71ececda

由 Sage Weil 提交于 11月 18, 2009

We require that ceph_con_close be called before we drop the connection,
so this is unneeded.  Just BUG if con->sock != NULL.
Signed-off-by: NSage Weil <sage@newdream.net>

71ececda

11 11月, 2009 1 次提交

ceph: separate banner and connect during handshake into distinct stages · eed0ef2c

由 Sage Weil 提交于 11月 10, 2009

We need to make sure we only swab the address during the banner once. So
break process_banner out of process_connect, and clean up the surrounding
code so that these are distinct phases of the handshake.
Signed-off-by: NSage Weil <sage@newdream.net>

eed0ef2c

05 11月, 2009 1 次提交

ceph: convert port endianness · f28bcfbe

由 Sage Weil 提交于 11月 04, 2009

The port is informational only, but we should make it correct.
Signed-off-by: NSage Weil <sage@newdream.net>

f28bcfbe

04 11月, 2009 1 次提交

ceph: use fixed endian encoding for ceph_entity_addr · 63f2d211

由 Sage Weil 提交于 11月 03, 2009

We exchange struct ceph_entity_addr over the wire and store it on disk.
The sockaddr_storage.ss_family field, however, is host endianness.  So,
fix ss_family endianness to big endian when sending/receiving over the
wire.
Signed-off-by: NSage Weil <sage@newdream.net>

63f2d211

10 10月, 2009 1 次提交

ceph: update to mon client protocol v15 · 13e38c8a

由 Sage Weil 提交于 10月 09, 2009

The mon request headers now include session_mon information that must
be properly initialized.
Signed-off-by: NSage Weil <sage@newdream.net>

13e38c8a

07 10月, 2009 1 次提交

ceph: messenger library · 31b8006e

由 Sage Weil 提交于 10月 06, 2009

A generic message passing library is used to communicate with all
other components in the Ceph file system.  The messenger library
provides ordered, reliable delivery of messages between two nodes in
the system.

This implementation is based on TCP.
Signed-off-by: NSage Weil <sage@newdream.net>

31b8006e