提交 · 77812a1ef139d84270d27faacc0630c887411013 · openeuler / raspberrypi-kernel

07 1月, 2011 3 次提交

fs: dcache remove dcache_lock · b5c84bf6

由 Nick Piggin 提交于 1月 07, 2011

dcache_lock no longer protects anything. remove it.
Signed-off-by: NNick Piggin <npiggin@kernel.dk>

b5c84bf6

fs: dcache scale subdirs · 2fd6b7f5

由 Nick Piggin 提交于 1月 07, 2011

Protect d_subdirs and d_child with d_lock, except in filesystems that aren't
using dcache_lock for these anyway (eg. using i_mutex).

Note: if we change the locking rule in future so that ->d_child protection is
provided only with ->d_parent->d_lock, it may allow us to reduce some locking.
But it would be an exception to an otherwise regular locking scheme, so we'd
have to see some good results. Probably not worthwhile.
Signed-off-by: NNick Piggin <npiggin@kernel.dk>

2fd6b7f5

fs: dcache scale dentry refcount · b7ab39f6

由 Nick Piggin 提交于 1月 07, 2011

Make d_count non-atomic and protect it with d_lock. This allows us to ensure a
0 refcount dentry remains 0 without dcache_lock. It is also fairly natural when
we start protecting many other dentry members with d_lock.
Signed-off-by: NNick Piggin <npiggin@kernel.dk>

b7ab39f6

18 11月, 2010 1 次提交

BKL: remove extraneous #include <smp_lock.h> · 451a3c24

由 Arnd Bergmann 提交于 11月 17, 2010

The big kernel lock has been removed from all these files at some point,
leaving only the #include.

Remove this too as a cleanup.
Signed-off-by: NArnd Bergmann <arnd@arndb.de>
Signed-off-by: NLinus Torvalds <torvalds@linux-foundation.org>

451a3c24

10 11月, 2010 1 次提交

ceph: make page alignment explicit in osd interface · b7495fc2

由 Sage Weil 提交于 11月 09, 2010

We used to infer alignment of IOs within a page based on the file offset,
which assumed they matched. This broke with direct IO that was not aligned
to pages (e.g., 512-byte aligned IO). We were also trusting the alignment
specified in the OSD reply, which could have been adjusted by the server.

Explicitly specify the page alignment when setting up OSD IO requests.
Signed-off-by: NSage Weil <sage@newdream.net>

b7495fc2

09 11月, 2010 2 次提交

ceph: fix update of ctime from MDS · d8672d64

由 Sage Weil 提交于 11月 08, 2010

The client can have a newer ctime than the MDS due to AUTH_EXCL and
XATTR_EXCL caps as well; update the check in ceph_fill_file_time
appropriately.

This fixes cases where ctime/mtime goes backward under the right sequence
of local updates (e.g. chmod) and mds replies (e.g. subsequent stat that
goes to the MDS).
Signed-off-by: NSage Weil <sage@newdream.net>

d8672d64

ceph: fix version check on racing inode updates · 8bd59e01

由 Sage Weil 提交于 11月 08, 2010

We may get updates on the same inode from multiple MDSs; generally we only
pay attention if the update is newer than what we already have.  The
exception is when an MDS sense unstable information, in which case we
always update.

The old > check got this wrong when our version was odd (e.g. 3) and the
reply version was even (e.g. 2): the older stale (v2) info would be
applied.  Fixed and clarified the comment.
Signed-off-by: NSage Weil <sage@newdream.net>

8bd59e01

08 11月, 2010 3 次提交

ceph: fix rdcache_gen usage and invalidate · cd045cb4

由 Sage Weil 提交于 11月 04, 2010

We used to use rdcache_gen to indicate whether we "might" have cached
pages. Now we just look at the mapping to determine that. However, some
old behavior remains from that transition.

First, rdcache_gen == 0 no longer means we have no pages. That can happen
at any time (presumably when we carry FILE_CACHE). We should not reset it
to zero, and we should not check that it is zero.

That means that the only purpose for rdcache_revoking is to resolve races
between new issues of FILE_CACHE and an async invalidate. If they are
equal, we should invalidate. On success, we decrement rdcache_revoking,
so that it is no longer equal to rdcache_gen. Similarly, if we success
in doing a sync invalidate, set revoking = gen - 1. (This is a small
optimization to avoid doing unnecessary invalidate work and does not
affect correctness.)
Signed-off-by: NSage Weil <sage@newdream.net>

cd045cb4

ceph: only let auth caps update max_size · 912a9b03

由 Sage Weil 提交于 11月 07, 2010

Only the auth MDS has a meaningful max_size value for us, so only update it
in fill_inode if we're being issued an auth cap. Otherwise, a random
stat result from a non-auth MDS can clobber a meaningful max_size, get
the client<->mds cap state out of sync, and make writes hang.

Specifically, even if the client re-requests a larger max_size (which it
will), the MDS won't respond because as far as it knows we already have a
sufficiently large value.
Signed-off-by: NSage Weil <sage@newdream.net>

912a9b03

ceph: fix bad pointer dereference in ceph_fill_trace · d8b16b3d

由 Sage Weil 提交于 11月 06, 2010

We dereference *in a few lines down, but only set it on rename.  It is
apparently pretty rare for this to trigger, but I have been hitting it
with a clustered MDSs.
Signed-off-by: NSage Weil <sage@newdream.net>

d8b16b3d

21 10月, 2010 1 次提交

ceph: factor out libceph from Ceph file system · 3d14c5d2

由 Yehuda Sadeh 提交于 4月 06, 2010

This factors out protocol and low-level storage parts of ceph into a
separate libceph module living in net/ceph and include/linux/ceph.  This
is mostly a matter of moving files around.  However, a few key pieces
of the interface change as well:

 - ceph_client becomes ceph_fs_client and ceph_client, where the latter
   captures the mon and osd clients, and the fs_client gets the mds client
   and file system specific pieces.
 - Mount option parsing and debugfs setup is correspondingly broken into
   two pieces.
 - The mon client gets a generic handler callback for otherwise unknown
   messages (mds map, in this case).
 - The basic supported/required feature bits can be expanded (and are by
   ceph_fs_client).

No functional change, aside from some subtle error handling cases that got
cleaned up in the refactoring process.
Signed-off-by: NSage Weil <sage@newdream.net>

3d14c5d2

14 9月, 2010 1 次提交

ceph: fix dn offset during readdir_prepopulate · 467c5251

由 Sage Weil 提交于 9月 13, 2010

When adding the readdir results to the cache, ceph_set_dentry_offset was
clobbered our just-set offset.  This can cause the readdir result offsets
to get out of sync with the server.  Add an argument to the helper so
that it does not.

This bug was introduced by 1cd3935b.
Signed-off-by: NSage Weil <sage@newdream.net>

467c5251

26 8月, 2010 1 次提交

ceph: ceph_get_inode() returns an ERR_PTR · ac1f12ef

由 Dan Carpenter 提交于 8月 25, 2010

ceph_get_inode() returns an ERR_PTR and it doesn't return a NULL.
Signed-off-by: NDan Carpenter <error27@gmail.com>
Signed-off-by: NSage Weil <sage@newdream.net>

ac1f12ef

23 8月, 2010 1 次提交

ceph: don't improperly set dir complete when holding EXCL cap · 12451491

由 Sage Weil 提交于 8月 22, 2010

If we hold the EXCL cap, we cannot trust the dir stats from the MDS (num
files, subdirs) and must not incorrectly conclude that the directory is
empty.  If we do, we get can bad results from lookup (bad ENOENT) and
bad readdir results.
Signed-off-by: NSage Weil <sage@newdream.net>

12451491

02 8月, 2010 1 次提交

ceph: perform lazy reads when file mode and caps permit · 2962507c

由 Sage Weil 提交于 5月 27, 2010

If the file mode is marked as "lazy," perform cached/buffered reads when
the caps permit it.  Adjust the rdcache_gen and invalidation logic
accordingly so that we manage our cache based on the FILE_CACHE -or-
FILE_LAZYIO cap bits.
Signed-off-by: NSage Weil <sage@newdream.net>

2962507c

28 7月, 2010 1 次提交

ceph: use complete_all and wake_up_all · 03066f23

由 Yehuda Sadeh 提交于 7月 27, 2010

This fixes an issue triggered by running concurrent syncs. One of the syncs
would go through while the other would just hang indefinitely. In any case, we
never actually want to wake a single waiter, so the *_all functions should
be used.
Signed-off-by: NYehuda Sadeh <yehuda@hq.newdream.net>
Signed-off-by: NSage Weil <sage@newdream.net>

03066f23

24 7月, 2010 1 次提交

ceph: fix leak of dentry in ceph_init_dentry() error path · 8c696737

由 Sage Weil 提交于 7月 22, 2010

If we fail to allocate a ceph_dentry_info, don't leak the dn reference.
Signed-off-by: NSage Weil <sage@newdream.net>

8c696737

22 6月, 2010 1 次提交

ceph: handle splice_dentry/d_materialize_unique error in readdir_prepopulate · d69ed05a

由 Sage Weil 提交于 6月 21, 2010

Handle a splice_dentry failure (due to a d_materialize_unique error)
without crashing.  (Also, report the error code.)
Signed-off-by: NSage Weil <sage@newdream.net>

d69ed05a

02 6月, 2010 1 次提交

ceph: fix d_subdirs ordering problem · 13a4214c

由 Henry C Chang 提交于 6月 01, 2010

We misused list_move_tail() to order the dentry in d_subdirs.
This will screw up the d_subdirs order.

This bug can be reliably reproduced by:
1. mount ceph fs.
2. on ceph fs, git clone git://ceph.newdream.net/git/ceph.git
3. Run autogen.sh in ceph directory.
(Note: Errors only occur at the first time you run autogen.sh.)
Signed-off-by: NHenry C Chang <henry_c_chang@tcloudcomputing.com>
Signed-off-by: NSage Weil <sage@newdream.net>

13a4214c

30 5月, 2010 1 次提交

fs/ceph: Use ERR_CAST · 7e34bc52

由 Julia Lawall 提交于 5月 22, 2010

Use ERR_CAST(x) rather than ERR_PTR(PTR_ERR(x)).  The former makes more
clear what is the purpose of the operation, which otherwise looks like a
no-op.

In the case of fs/ceph/inode.c, ERR_CAST is not needed, because the type of
the returned value is the same as the type of the enclosing function.

The semantic patch that makes this change is as follows:
(http://coccinelle.lip6.fr/)

// <smpl>
@@
type T;
T x;
identifier f;
@@

T f (...) { <+...
- ERR_PTR(PTR_ERR(x))
+ x
 ...+> }

@@
expression x;
@@

- ERR_PTR(PTR_ERR(x))
+ ERR_CAST(x)
// </smpl>
Signed-off-by: NJulia Lawall <julia@diku.dk>
Signed-off-by: NSage Weil <sage@newdream.net>

7e34bc52

18 5月, 2010 7 次提交

ceph: use common helper for aborted dir request invalidation · 167c9e35

由 Sage Weil 提交于 5月 14, 2010

We invalidate I_COMPLETE and dentry leases in two places: on aborted mds
request and on request replay. Use common helper to avoid duplicate code.
Signed-off-by: NSage Weil <sage@newdream.net>

167c9e35

ceph: set dn offset when spliced · 1cd3935b

由 Sage Weil 提交于 5月 03, 2010

We want to assign an offset when the dentry goes from null to linked, which
is always done by splice_dentry().  Notably, we should NOT assign an
offset when a dentry is first created and is still null.

BUG if we try to splice a non-null dentry (we shouldn't).
Signed-off-by: NSage Weil <sage@newdream.net>

1cd3935b

ceph: don't clobber i_max_offset on already complete dir · 1b7facc4

由 Sage Weil 提交于 4月 16, 2010

This can screw up offsets assigned to new dentries and break dcache
readdir results.
Signed-off-by: NSage Weil <sage@newdream.net>

1b7facc4

S
ceph: skip set_dentry_offset work if directory not I_COMPLETE · e8a74987
由 Sage Weil 提交于 4月 15, 2010
```
Signed-off-by: NSage Weil <sage@newdream.net>
```
e8a74987

ceph: fix xattr dangling pointer / double free · a6424e48

由 Sage Weil 提交于 4月 29, 2010

If we use the xattr_blob, clear the pointer so we don't release the memory
at the bottom of the fuction.
Reported-by: NHenry C Chang <henry_c_chang@tcloudcomputing.com>
Signed-off-by: NSage Weil <sage@newdream.net>

a6424e48

ceph: use ceph_sb_to_client instead of ceph_client · 640ef79d

由 Cheng Renquan 提交于 3月 26, 2010

ceph_sb_to_client and ceph_client are really identical, we need to dump
one; while function ceph_client is confusing with "struct ceph_client",
ceph_sb_to_client's definition is more clear; so we'd better switch all
call to ceph_sb_to_client.

  -static inline struct ceph_client *ceph_client(struct super_block *sb)
  -{
  -	return sb->s_fs_info;
  -}
Signed-off-by: NCheng Renquan <crquan@gmail.com>
Signed-off-by: NSage Weil <sage@newdream.net>

640ef79d

ceph: invalidate affected dentry leases on aborted requests · 81a6cf2d

由 Sage Weil 提交于 5月 14, 2010

If we abort a request, we return to caller, but the request may still
complete. And if we hold the dir FILE_EXCL bit, we may not release a
lease when sending a request. A simple un-tar, control-c, un-tar again
will reproduce the bug (manifested as a 'Cannot open: File exists').

Ensure we invalidate affected dentry leases (as well dir I_COMPLETE) so
we don't have valid (but incorrect) leases. Do the same, consistently, at
other sites where I_COMPLETE is similarly cleared.
Signed-off-by: NSage Weil <sage@newdream.net>

81a6cf2d

12 5月, 2010 1 次提交

ceph: fix open file counting on snapped inodes when mds returns no caps · 04d000eb

由 Sage Weil 提交于 5月 07, 2010

It's possible the MDS will not issue caps on a snapped inode, in which case
an open request may not __ceph_get_fmode(), botching the open file
counting.  (This is actually a server bug, but the client shouldn't BUG out
in this case.)
Signed-off-by: NSage Weil <sage@newdream.net>

04d000eb

04 5月, 2010 1 次提交

ceph: clear dir complete on d_move · c10f5e12

由 Sage Weil 提交于 4月 16, 2010

d_move() reorders the d_subdirs list, breaking the readdir result caching.
Unless/until d_move preserves that ordering, clear CEPH_I_COMPLETE on
rename.
Signed-off-by: NSage Weil <sage@newdream.net>

c10f5e12

31 3月, 2010 1 次提交

ceph: fix dentry rehashing on virtual .snap dir · 9358c6d4

由 Sage Weil 提交于 3月 30, 2010

If a lookup fails on the magic .snap directory, we bind it to a magic
snap directory inode in ceph_lookup_finish().  That code assumes the dentry
is unhashed, but a recent server-side change started returning NULL leases
on lookup failure, causing the .snap dentry to be hashed and NULL by
ceph_fill_trace().

This causes dentry hash chain corruption, or a dies when d_rehash()
includes
	BUG_ON(!d_unhashed(entry));

So, avoid processing the NULL dentry lease if it the dentry matches the
snapdir name in ceph_fill_trace().  That allows the lookup completion to
properly bind it to the snapdir inode.  BUG there if dentry is hashed to
be sure.
Signed-off-by: NSage Weil <sage@newdream.net>

9358c6d4

21 3月, 2010 1 次提交

ceph: fix inode removal from snap realm when racing with migration · 8b218b8a

由 Sage Weil 提交于 3月 09, 2010

When an inode was dropped while being migrated between two MDSs,
i_cap_exporting_issued was non-zero such that issue caps were non-zero and
__ceph_is_any_caps(ci) was true.  This prevented the inode from being
removed from the snap realm, even as it was dropped from the cache.

Fix this by dropping any residual i_snap_realm ref in destroy_inode.
Signed-off-by: NSage Weil <sage@newdream.net>

8b218b8a

20 2月, 2010 1 次提交

ceph: don't truncate dirty pages in invalidate work thread · c9af9fb6

由 Yehuda Sadeh 提交于 2月 19, 2010

Instead of truncating the whole range of pages, we skip those
pages that are dirty or in the middle of writeback. Those pages
will be cleared later when the writeback completes.
Signed-off-by: NYehuda Sadeh <yehuda@hq.newdream.net>
Signed-off-by: NSage Weil <sage@newdream.net>

c9af9fb6

18 2月, 2010 1 次提交
- S
  ceph: fix typo in ceph_queue_writeback debug output · 2c27c9a5
  由 Sage Weil 提交于 2月 17, 2010
```
Signed-off-by: NSage Weil <sage@newdream.net>
```
  2c27c9a5
12 2月, 2010 2 次提交

ceph: cleanup async writeback, truncation, invalidate helpers · 3c6f6b79

由 Sage Weil 提交于 2月 09, 2010

Grab inode ref in helper.  Make work functions static, with consistent
naming.
Signed-off-by: NSage Weil <sage@newdream.net>

3c6f6b79

ceph: fix truncation when not holding caps · 3d497d85

由 Yehuda Sadeh 提交于 2月 09, 2010

A truncation should occur when either we have the
specified caps for the file, or (in cases where we are
not the only ones referencing the file) when it is mapped
or when it is opened. The latter two cases were not
handled.
Signed-off-by: NYehuda Sadeh <yehuda@hq.newdream.net>
Signed-off-by: NSage Weil <sage@newdream.net>

3d497d85

30 1月, 2010 1 次提交

ceph: remove unreachable code · 0f26c4b2

由 Yehuda Sadeh 提交于 1月 29, 2010

We never truncate to a smaller size without contacting the MDS.
Signed-off-by: NYehuda Sadeh <yehuda@hq.newdream.net>
Signed-off-by: NSage Weil <sage@newdream.net>

0f26c4b2

26 1月, 2010 1 次提交

ceph: properly handle aborted mds requests · 5b1daecd

由 Sage Weil 提交于 1月 25, 2010

Previously, if the MDS request was interrupted, we would unregister the
request and ignore any reply. This could cause the caps or other cache
state to become out of sync. (For instance, aborting dbench and doing
rm -r on clients would complain about a non-empty directory because the
client didn't realize it's aborted file create request completed.)

Even we don't unregister, we still can't process the reply normally because
we are no longer holding the caller's locks (like the dir i_mutex).

So, mark aborted operations with r_aborted, and in the reply handler, be
sure to process all the caps. Do not process the namespace changes,
though, since we no longer will hold the dir i_mutex. The dentry lease
state can also be ignored as it's more forgiving.
Signed-off-by: NSage Weil <sage@newdream.net>

5b1daecd

15 1月, 2010 1 次提交

ceph: change dentry offset and position after splice_dentry · 4baa75ef

由 Yehuda Sadeh 提交于 1月 07, 2010

This fixes a bug, where we had the parent list have dentries with
offsets that are not monotonically increasing, which caused the ceph
dcache_readdir to skip entries.
Signed-off-by: NYehuda Sadeh <yehuda@hq.newdream.net>
Signed-off-by: NSage Weil <sage@newdream.net>

4baa75ef

22 12月, 2009 1 次提交

ceph: ensure rename target dentry fails revalidation · c4a29f26

由 Sage Weil 提交于 12月 21, 2009

This works around a bug in vfs_rename_dir() that rehashes the target
dentry.  Ensure such dentries always fail revalidation by timing out the
dentry lease and kicking it out of the current directory lease gen.

This can be reverted when the vfs bug is fixed.
Signed-off-by: NSage Weil <sage@newdream.net>

c4a29f26

08 12月, 2009 1 次提交

ceph: simplify ceph_buffer interface · b6c1d5b8

由 Sage Weil 提交于 12月 07, 2009

We never allocate the ceph_buffer and buffer separtely, so use a single
constructor.

Disallow put on NULL buffer; make the caller check.
Signed-off-by: NSage Weil <sage@newdream.net>

b6c1d5b8