提交 · 4850f524b2c4c8a4e9f8ef4dd9c7c4afde2f2b2c · openeuler / raspberrypi-kernel

01 3月, 2010 4 次提交

GFS2: print glock numbers in hex · 4818972e

由 Bob Peterson 提交于 2月 23, 2010

This patch changes glock numbers from printing in decimal to hex.
Since DLM prints corresponding resource IDs in hex, it makes debugging
easier.
Signed-off-by: NBob Peterson <rpeterso@redhat.com>
Signed-off-by: NSteven Whitehouse <swhiteho@redhat.com>

4818972e

GFS2: ordered writes are backwards · e5884636

由 Dave Chinner 提交于 2月 05, 2010

When we queue data buffers for ordered write, the buffers are added
to the head of the ordered write list. When the log needs to push
these buffers to disk, it also walks the list from the head. The
result is that the the ordered buffers are submitted to disk in
reverse order.

For large writes, this means that whenever the log flushes large
streams of reverse sequential order buffers are pushed down into the
block layers. The elevators don't handle this particularly well, so
IO rates tend to be significantly lower than if the IO was issued in
ascending block order.

Queue new ordered buffers to the tail of the ordered buffer list to
ensure that IO is dispatched in the order it was submitted. This
should significantly improve large sequential write speeds. On a
disk capable of 85MB/s, speeds increase from 50MB/s to 65MB/s for
noop and from 38MB/s to 50MB/s for cfq.
Signed-off-by: NDave Chinner <dchinner@redhat.com>
Signed-off-by: NSteven Whitehouse <swhiteho@redhat.com>

e5884636

GFS2: Remove loopy umount code · c1184f8a

由 Steven Whitehouse 提交于 1月 08, 2010

As a consequence of the previous patch, we can now remove the
loop which used to be required due to the circular dependency
between the inodes and glocks. Instead we can just invalidate
the inodes, and then clear up any glocks which are left.

Also we no longer need the rwsem since there is no longer any
danger of the inode invalidation calling back into the glock
code (and from there back into the inode code).
Signed-off-by: NSteven Whitehouse <swhiteho@redhat.com>

c1184f8a

GFS2: Metadata address space clean up · 009d8518

由 Steven Whitehouse 提交于 12月 08, 2009

Since the start of GFS2, an "extra" inode has been used to store
the metadata belonging to each inode. The only reason for using
this inode was to have an extra address space, the other fields
were unused. This means that the memory usage was rather inefficient.

The reason for keeping each inode's metadata in a separate address
space is that when glocks are requested on remote nodes, we need to
be able to efficiently locate the data and metadata which relating
to that glock (inode) in order to sync or sync and invalidate it
(depending on the remotely requested lock mode).

This patch adds a new type of glock, which has in addition to
its normal fields, has an address space. This applies to all
inode and rgrp glocks (but to no other glock types which remain
as before). As a result, we no longer need to have the second
inode.

This results in three major improvements:
 1. A saving of approx 25% of memory used in caching inodes
 2. A removal of the circular dependency between inodes and glocks
 3. No confusion between "normal" and "metadata" inodes in super.c

Although the first of these is the more immediately apparent, the
second is just as important as it now enables a number of clean
ups at umount time. Those will be the subject of future patches.
Signed-off-by: NSteven Whitehouse <swhiteho@redhat.com>

009d8518

12 2月, 2010 2 次提交

GFS2: Fix bmap allocation corner-case bug · 07ccb7bf

由 Steven Whitehouse 提交于 2月 12, 2010

This patch solves a corner case during allocation which occurs if both
metadata (indirect) and data blocks are required but there is an
obstacle in the filesystem (e.g. a resource group header or another
allocated block) such that when the allocation is requested only
enough blocks for the metadata are returned.

By changing the exit condition of this loop, we ensure that a
minimum of one data block will always be returned.
Signed-off-by: NSteven Whitehouse <swhiteho@redhat.com>

07ccb7bf

GFS2: Fix error code · 0e5a9fb0

由 Abhijith Das 提交于 2月 05, 2010

We need this one-liner to signal the mount helper of the 'insufficient journals' condition.
Signed-off-by: NAbhijith Das <adas@redhat.com>
Signed-off-by: NSteven Whitehouse <swhiteho@redhat.com>

0e5a9fb0

03 2月, 2010 2 次提交

GFS2: Extend umount wait coverage to full glock lifetime · 8f05228e

由 Steven Whitehouse 提交于 1月 29, 2010

Although all glocks are, by the time of the umount glock wait,
scheduled for demotion, some of them haven't made it far
enough through the process for the original set of waiting
code to wait for them.

This extends the ref count to the whole glock lifetime in order
to ensure that the waiting does catch all glocks. It does make
it a bit more invasive, but it seems the only sensible solution
at the moment.
Signed-off-by: NSteven Whitehouse <swhiteho@redhat.com>

8f05228e

GFS2: Wait for unlock completion on umount · e402746a

由 Steven Whitehouse 提交于 1月 25, 2010

This patch adds a wait on umount between the point at which we
dispose of all glocks and the point at which we unmount the
lock protocol. This ensures that we've received all the replies
to our unlock requests before we stop the locking.
Signed-off-by: NSteven Whitehouse <swhiteho@redhat.com>
Reported-by: NFabio M. Di Nitto <fdinitto@redhat.com>

e402746a

01 2月, 2010 3 次提交

GFS2: Use GFP_NOFS for alloc structure · ea8d62da

由 Steven Whitehouse 提交于 1月 29, 2010

This is called under a glock, so its a good plan to use GFP_NOFS
Signed-off-by: NSteven Whitehouse <swhiteho@redhat.com>

ea8d62da

GFS2: Fix previous patch · 7fe3ec6f

由 Steven Whitehouse 提交于 1月 29, 2010

The do_div() call needs to remain.
Signed-off-by: NSteven Whitehouse <swhiteho@redhat.com>

7fe3ec6f

GFS2: Don't withdraw on partial rindex entries · 55f0b4c5

由 Benjamin Marzinski 提交于 1月 25, 2010

ince gfs2 writes the rindex file a block at a time, and releases the
exclusive lock after each block, it is possible that another process
will grab the lock in the middle of the write. Since rindex entries are
not an even divisor of blocks, that other process may see partial
entries. On grows, this is fine. The process can simply ignore the the
partial entires. Previously, the code withdrew when it saw partial
entries. Now it simply ignores them.
Signed-off-by: NBenjamin Marzinski <bmarzins@redhat.com>
Signed-off-by: NSteven Whitehouse <swhiteho@redhat.com>

55f0b4c5

12 1月, 2010 1 次提交

GFS2: Fix refcnt leak on gfs2_follow_link() error path · 0f585f14

由 OGAWA Hirofumi 提交于 1月 12, 2010

If ->follow_link handler return the error, it should decrement
nd->path refcnt.

This patch fix it.
Signed-off-by: NOGAWA Hirofumi <hirofumi@mail.parknet.co.jp>
Signed-off-by: NSteven Whitehouse <swhiteho@redhat.com>

0f585f14

11 1月, 2010 1 次提交

GFS2: Use MAX_LFS_FILESIZE for meta inode size · ba198098

由 Steven Whitehouse 提交于 1月 08, 2010

Using ~0ULL was cauing sign issues in filemap_fdatawrite_range, so
use MAX_LFS_FILESIZE instead.
Signed-off-by: NSteven Whitehouse <swhiteho@redhat.com>

ba198098

08 1月, 2010 3 次提交

GFS2: Fix gfs2_xattr_acl_chmod() · e412bdb1

由 Steven Whitehouse 提交于 12月 21, 2009

The ref counting for the bh returned by gfs2_ea_find() was
wrong. This patch ensures that we always drop the ref count
to that bh correctly.
Signed-off-by: NSteven Whitehouse <swhiteho@redhat.com>

e412bdb1

GFS2: Fix locking bug in rename · 24b977b5

由 Steven Whitehouse 提交于 12月 09, 2009

The rename code was taking a resource group lock in cases where
it wasn't actually needed, this caused problems if the rename
was resulting in an inode being unlinked. The patch ensures that
we only take the rgrp lock early if it is really needed.
Signed-off-by: NSteven Whitehouse <swhiteho@redhat.com>

24b977b5

GFS2: Ensure uptodate inode size when using O_APPEND · 56aa616a

由 Steven Whitehouse 提交于 12月 08, 2009

The VFS reads the inode size during generic_file_aio_write() but
with no locking around it. In order to get the expected result
from O_APPEND opens, this patch updated the inode size before
calling generic_file_aio_write()

There is of course still a race here, in that there is nothing to
prevent another node coming in and extending the file in the
mean time. On the other hand, when used with file locking this
will ensure that the expected results are obtained.
Signed-off-by: NSteven Whitehouse <swhiteho@redhat.com>

56aa616a

18 12月, 2009 2 次提交

Revert "task_struct: make journal_info conditional" · b6e3224f

由 Linus Torvalds 提交于 12月 17, 2009

This reverts commit e4c570c4, as
requested by Alexey:

 "I think I gave a good enough arguments to not merge it.
  To iterate:
   * patch makes impossible to start using ext3 on EXT3_FS=n kernels
     without reboot.
   * this is done only for one pointer on task_struct"

  None of config options which define task_struct are tristate directly
  or effectively."
Requested-by: NAlexey Dobriyan <adobriyan@gmail.com>
Acked-by: NAndrew Morton <akpm@linux-foundation.org>
Signed-off-by: NLinus Torvalds <torvalds@linux-foundation.org>

b6e3224f

kill I_LOCK · eaff8079

由 Christoph Hellwig 提交于 12月 17, 2009

After I_SYNC was split from I_LOCK the leftover is always used together with
I_NEW and thus superflous.
Signed-off-by: NChristoph Hellwig <hch@lst.de>
Signed-off-by: NAl Viro <viro@zeniv.linux.org.uk>

eaff8079

17 12月, 2009 1 次提交

sanitize xattr handler prototypes · 431547b3

由 Christoph Hellwig 提交于 11月 13, 2009

Add a flags argument to struct xattr_handler and pass it to all xattr
handler methods.  This allows using the same methods for multiple
handlers, e.g. for the ACL methods which perform exactly the same action
for the access and default ACLs, just using a different underlying
attribute.  With a little more groundwork it'll also allow sharing the
methods for the regular user/trusted/secure handlers in extN, ocfs2 and
jffs2 like it's already done for xfs in this patch.

Also change the inode argument to the handlers to a dentry to allow
using the handlers mechnism for filesystems that require it later,
e.g. cifs.

[with GFS2 bits updated by Steven Whitehouse <swhiteho@redhat.com>]
Signed-off-by: NChristoph Hellwig <hch@lst.de>
Reviewed-by: NJames Morris <jmorris@namei.org>
Acked-by: NJoel Becker <joel.becker@oracle.com>
Signed-off-by: NAl Viro <viro@zeniv.linux.org.uk>

431547b3

16 12月, 2009 2 次提交

fs/gfs2/sys.c: use %pUB to print UUIDs · f0b34ae6

由 Joe Perches 提交于 12月 14, 2009

Signed-off-by: NJoe Perches <joe@perches.com>
Cc: Steven Whitehouse <swhiteho@redhat.com>
Signed-off-by: NAndrew Morton <akpm@linux-foundation.org>
Signed-off-by: NLinus Torvalds <torvalds@linux-foundation.org>

f0b34ae6

task_struct: make journal_info conditional · e4c570c4

由 Hiroshi Shimamoto 提交于 12月 14, 2009

journal_info in task_struct is used in journaling file system only.  So
introduce CONFIG_FS_JOURNAL_INFO and make it conditional.
Signed-off-by: NHiroshi Shimamoto <h-shimamoto@ct.jp.nec.com>
Cc: Chris Mason <chris.mason@oracle.com>
Cc: "Theodore Ts'o" <tytso@mit.edu>
Cc: Steven Whitehouse <swhiteho@redhat.com>
Cc: KONISHI Ryusuke <konishi.ryusuke@lab.ntt.co.jp>
Signed-off-by: NAndrew Morton <akpm@linux-foundation.org>
Signed-off-by: NLinus Torvalds <torvalds@linux-foundation.org>

e4c570c4

03 12月, 2009 19 次提交

GFS2: Fix glock refcount issues · 26bb7505

由 Steven Whitehouse 提交于 11月 27, 2009

This patch fixes some ref counting issues. Firstly by moving
the point at which we drop the ref count after a dlm lock
operation has completed we ensure that we never call
gfs2_glock_hold() on a lock with a zero ref count.

Secondly, by using atomic_dec_and_lock() in gfs2_glock_put()
we ensure that at no time will a glock with zero ref count
appear on the lru_list. That means that we can remove the
check for this in our shrinker (which was racy).
Signed-off-by: NSteven Whitehouse <swhiteho@redhat.com>

26bb7505

writeback: remove unused nonblocking and congestion checks (gfs2) · c29cd900

由 Wu Fengguang 提交于 11月 18, 2009

No one is calling wb_writeback and write_cache_pages with
wbc.nonblocking=1 any more. And lumpy pageout will want to do
nonblocking writeback without the congestion wait.
Signed-off-by: NWu Fengguang <fengguang.wu@intel.com>
Signed-off-by: NSteven Whitehouse <swhiteho@redhat.com>

c29cd900

GFS2: drop rindex glock to refresh rindex list · 9ae3c6de

由 Benjamin Marzinski 提交于 11月 10, 2009

When a gfs2 filesystem is grown, it needs to rebuild the rindex list to be able
to use the new space. gfs2 does this when the rindex is marked not uptodate,
which happens when the rindex glock is dropped. However, on a single node
setup, there is never any reason to drop the rindex glock, so gfs2 never
invalidates the the rindex. This patch makes gfs2 automatically drop the
rindex glock after filesystem grows, so it can refresh the rindex list.
Signed-off-by: NBenjamin Marzinski <bmarzins@redhat.com>
Signed-off-by: NSteven Whitehouse <swhiteho@redhat.com>

9ae3c6de

GFS2: Tag all metadata with jid · 0ab7d13f

由 Steven Whitehouse 提交于 11月 06, 2009

There are two spare field in the header common to all GFS2
metadata. One is just the right size to fit a journal id
in it, and this patch updates the journal code so that each
time a metadata block is modified, we tag it with the journal
id of the node which is performing the modification.

The reason for this is that it should make it much easier to
debug issues which arise if we can tell which node was the
last to modify a particular metadata block.

Since the field is updated before the block is written into
the journal, each journal should only contain metadata which
is tagged with its own journal id. The one exception to this
is the journal header block, which might have a different node's
id in it, if that journal was recovered by another node in the
cluster.

Thus each journal will contain a record of which nodes recovered
it, via the journal header.

The other field in the metadata header could potentially be
used to hold information about what kind of operation was
performed, but for the time being we just zero it on each
transaction so that if we use it for that in future, we'll
know that the information (where it exists) is reliable.

I did consider using the other field to hold the journal
sequence number, however since in GFS2's journaling we write
the modified data into the journal and not the original
data, this gives no information as to what action caused the
modification, so I think we can probably come up with a better
use for those 64 bits in the future.
Signed-off-by: NSteven Whitehouse <swhiteho@redhat.com>

0ab7d13f

GFS2: Locking order fix in gfs2_check_blk_state · 2c776349

由 Steven Whitehouse 提交于 11月 06, 2009

In some cases we already have the rindex lock when
we enter this function.
Signed-off-by: NSteven Whitehouse <swhiteho@redhat.com>

2c776349

GFS2: Remove dirent_first() function · 1579343a

由 Steven Whitehouse 提交于 11月 06, 2009

This function only had one caller left, and that caller only
called it for leaf blocks, hence one branch of the "if" was
never taken. In addition the call to get_left had already
verified the metadata type, so the function can be reduced
to a single line of code in its caller.
Signed-off-by: NSteven Whitehouse <swhiteho@redhat.com>

1579343a

GFS2: Display nobarrier option in /proc/mounts · cdcfde62

由 Steven Whitehouse 提交于 10月 30, 2009

Since the default is barriers on, this only displays the
nobarrier option when that is active.
Signed-off-by: NSteven Whitehouse <swhiteho@redhat.com>

cdcfde62

GFS2: add barrier/nobarrier mount options · f25934c5

由 Christoph Hellwig 提交于 10月 30, 2009

Currently gfs2 issues barrier unconditionally.  There are various reasons
to disable them, be that just for testing or for stupid devices flushing
large battert backed caches.  Add a nobarrier option that matches xfs and
btrfs for this.  Also add a symmetric barrier option to turn it back on
at remount time.
Signed-off-by: NChristoph Hellwig <hch@lst.de>
Signed-off-by: NSteven Whitehouse <swhiteho@redhat.com>

f25934c5

GFS2: remove division from new statfs code · c14f5735

由 Benjamin Marzinski 提交于 10月 26, 2009

It's not necessary to do any 64bit division for the statfs sync code, so
remove it.
Signed-off-by: NBenjamin Marzinski <bmarzins@redhat.com>
Signed-off-by: NSteven Whitehouse <swhiteho@redhat.com>

c14f5735

GFS2: Improve statfs and quota usability · 3d3c10f2

由 Benjamin Marzinski 提交于 10月 20, 2009

GFS2 now has three new mount options, statfs_quantum, quota_quantum and
statfs_percent. statfs_quantum and quota_quantum simply allow you to
set the tunables of the same name. Setting setting statfs_quantum to 0
will also turn on the statfs_slow tunable. statfs_percent accepts an
integer between 0 and 100. Numbers between 1 and 100 will cause GFS2 to
do any early sync when the local number of blocks free changes by at
least statfs_percent from the totoal number of blocks free. Setting
statfs_percent to 0 disables this.
Signed-off-by: NBenjamin Marzinski <bmarzins@redhat.com>
Signed-off-by: NSteven Whitehouse <swhiteho@redhat.com>

3d3c10f2

GFS2: Use dquot_send_warning() · 2ec46505

由 Steven Whitehouse 提交于 9月 28, 2009

This adds support to GFS2 to send quota warnings via netlink.
Also it removes a stray \r which was left over from when the
code used to print warnings on the console.
Signed-off-by: NSteven Whitehouse <swhiteho@redhat.com>

2ec46505

GFS2: Add set_xquota support · e285c100

由 Steven Whitehouse 提交于 9月 23, 2009

This patch adds the ability to set GFS2 quota limit and
warning levels via the XFS quota API.
Signed-off-by: NSteven Whitehouse <swhiteho@redhat.com>

e285c100

GFS2: Add get_xquota support · 113d6b3c

由 Steven Whitehouse 提交于 9月 28, 2009

This adds support for viewing the current GFS2 quota settings
via the XFS quota API. The setting of quotas will be addressed
in a later patch. Fields which are not supported here are left
set to zero.
Signed-off-by: NSteven Whitehouse <swhiteho@redhat.com>
Reviewed-by: NBob Peterson <rpeterso@redhat.com>

113d6b3c

GFS2: Clean up gfs2_adjust_quota() and do_glock() · 1e72c0f7

由 Steven Whitehouse 提交于 9月 15, 2009

Both of these functions contained confusing and in one case
duplicate code. This patch adds a new check in do_glock()
so that we report -ENOENT if we are asked to sync a quota
entry which doesn't exist. Due to the previous patch this is
now reported correctly to userspace.

Also there are a few new comments, and I hope that the code
is easier to understand now.
Signed-off-by: NSteven Whitehouse <swhiteho@redhat.com>

1e72c0f7

GFS2: Remove constant argument from qd_get() · 6a6ada81

由 Steven Whitehouse 提交于 9月 15, 2009

This function was only ever called with the "create"
argument set to true, so we can remove it.
Signed-off-by: NSteven Whitehouse <swhiteho@redhat.com>

6a6ada81

GFS2: Remove constant argument from qdsb_get() · 33a82529

由 Steven Whitehouse 提交于 9月 15, 2009

The "create" argument to qdsb_get() was only ever set to true,
so this patch removes that argument.
Signed-off-by: NSteven Whitehouse <swhiteho@redhat.com>

33a82529

GFS2: Add proper error reporting to quota sync via sysfs · ea762338

由 Steven Whitehouse 提交于 9月 15, 2009

For some reason, the errors were not making it to userspace.
Signed-off-by: NSteven Whitehouse <swhiteho@redhat.com>

ea762338

GFS2: Add get_xstate quota function · 1d371b5e

由 Steven Whitehouse 提交于 9月 11, 2009

This allows querying of the quota state via the XFS quota
API.
Signed-off-by: NSteven Whitehouse <swhiteho@redhat.com>

1d371b5e

GFS2: Remove obsolete code in quota.c · 91094d0f

由 Steven Whitehouse 提交于 9月 11, 2009

There is no point in testing for GLF_DEMOTE here, we might as
well always release the glock at that point.
Signed-off-by: NSteven Whitehouse <swhiteho@redhat.com>

91094d0f