提交 · 44d964d609c7c11b330a3d1caf30767fa13c7be3 · openeuler / raspberrypi-kernel

04 1月, 2012 24 次提交

A
vfs: spread struct mount mnt_set_mountpoint child argument · 44d964d6
由 Al Viro 提交于 11月 24, 2011
```
Signed-off-by: NAl Viro <viro@zeniv.linux.org.uk>
```
44d964d6
A
vfs: spread struct mount - clone_mnt/copy_tree argument · 87129cc0
由 Al Viro 提交于 11月 24, 2011
```
Signed-off-by: NAl Viro <viro@zeniv.linux.org.uk>
```
87129cc0
A
vfs: spread struct mount - shrink_submounts/select_submounts · 692afc31
由 Al Viro 提交于 11月 24, 2011
```
Signed-off-by: NAl Viro <viro@zeniv.linux.org.uk>
```
692afc31
A
vfs: spread struct mount - umount_tree argument · 761d5c38
由 Al Viro 提交于 11月 24, 2011
```
Signed-off-by: NAl Viro <viro@zeniv.linux.org.uk>
```
761d5c38

vfs: the first spoils - mnt_hash moved · 1b8e5564

由 Al Viro 提交于 11月 24, 2011

taken out of struct vfsmount into struct mount
Signed-off-by: NAl Viro <viro@zeniv.linux.org.uk>

1b8e5564

A
vfs: spread struct mount to remaining users of ->mnt_hash · d5e50f74
由 Al Viro 提交于 11月 24, 2011
```
Signed-off-by: NAl Viro <viro@zeniv.linux.org.uk>
```
d5e50f74
A
vfs: spread struct mount - clone_mnt/copy_tree result · cb338d06
由 Al Viro 提交于 11月 24, 2011
```
Signed-off-by: NAl Viro <viro@zeniv.linux.org.uk>
```
cb338d06
A
vfs: spread struct mount - change_mnt_propagation/set_mnt_shared · 0f0afb1d
由 Al Viro 提交于 11月 24, 2011
```
Signed-off-by: NAl Viro <viro@zeniv.linux.org.uk>
```
0f0afb1d
A
vfs: spread struct mount - alloc_vfsmnt/free_vfsmnt/mnt_alloc_id/mnt_free_id · b105e270
由 Al Viro 提交于 11月 24, 2011
```
Signed-off-by: NAl Viro <viro@zeniv.linux.org.uk>
```
b105e270
A
vfs: spread struct mount - tree_contains_unbindable · cbbe362c
由 Al Viro 提交于 11月 24, 2011
```
Signed-off-by: NAl Viro <viro@zeniv.linux.org.uk>
```
cbbe362c
A
vfs: spread struct mount - attach_recursive_mnt · 0fb54e50
由 Al Viro 提交于 11月 24, 2011
```
Signed-off-by: NAl Viro <viro@zeniv.linux.org.uk>
```
0fb54e50
A
vfs: spread struct mount - mount group id handling · 4b8b21f4
由 Al Viro 提交于 11月 24, 2011
```
Signed-off-by: NAl Viro <viro@zeniv.linux.org.uk>
```
4b8b21f4
A
vfs: spread struct mount - commit_tree · 4b2619a5
由 Al Viro 提交于 11月 24, 2011
```
Signed-off-by: NAl Viro <viro@zeniv.linux.org.uk>
```
4b2619a5
A
vfs: spread struct mount - attach_mnt/detach_mnt · 419148da
由 Al Viro 提交于 11月 24, 2011
```
Signed-off-by: NAl Viro <viro@zeniv.linux.org.uk>
```
419148da

vfs: spread struct mount - namespace.c internal iterators · 315fc83e

由 Al Viro 提交于 11月 24, 2011

next_mnt() return value, first argument
skip_mnt_tree() return value and argument
Signed-off-by: NAl Viro <viro@zeniv.linux.org.uk>

315fc83e

vfs: spread struct mount - __lookup_mnt() result · c7105365

由 Al Viro 提交于 11月 24, 2011

switch __lookup_mnt() to returning struct mount *; callers adjusted.
Signed-off-by: NAl Viro <viro@zeniv.linux.org.uk>

c7105365

vfs: start hiding vfsmount guts series · 7d6fec45

由 Al Viro 提交于 11月 23, 2011

Almost all fields of struct vfsmount are used only by core VFS (and
a fairly small part of it, at that).  The plan: embed struct vfsmount
into struct mount, making the latter visible only to core parts of VFS.
Then move fields from vfsmount to mount, eventually leaving only
mnt_root/mnt_sb/mnt_flags in struct vfsmount.  Filesystem code still
gets pointers to struct vfsmount and remains unchanged; all such
pointers go to struct vfsmount embedded into the instances of struct
mount allocated by fs/namespace.c.  When fs/namespace.c et.al. get
a pointer to vfsmount, they turn it into pointer to mount (using
container_of) and work with that.

This is the first part of series; struct mount is introduced,
allocation switched to using it.
Signed-off-by: NAl Viro <viro@zeniv.linux.org.uk>

7d6fec45

vfs: mnt_drop_write_file() · 2a79f17e

由 Al Viro 提交于 12月 09, 2011

new helper (wrapper around mnt_drop_write()) to be used in pair with
mnt_want_write_file().
Signed-off-by: NAl Viro <viro@zeniv.linux.org.uk>

2a79f17e

vfs: make do_kern_mount() static · 79e801a9

由 Al Viro 提交于 12月 12, 2011

the only user outside of fs/namespace.c has died
Signed-off-by: NAl Viro <viro@zeniv.linux.org.uk>

79e801a9

A
vfs: dentry_reset_mounted() doesn't use vfsmount argument · aa0a4cf0
由 Al Viro 提交于 11月 24, 2011
```
lose it
Signed-off-by: NAl Viro <viro@zeniv.linux.org.uk>
```
aa0a4cf0
A
unexport put_mnt_ns(), make create_mnt_ns() static outright · 6c449c8d
由 Al Viro 提交于 11月 16, 2011
```
Signed-off-by: NAl Viro <viro@zeniv.linux.org.uk>
```
6c449c8d

vfs: more mnt_parent cleanups · afac7cba

由 Al Viro 提交于 11月 23, 2011

a) mount --move is checking that ->mnt_parent is non-NULL before
looking if that parent happens to be shared; ->mnt_parent is never
NULL and it's not even an misspelled !mnt_has_parent()

b) pivot_root open-codes is_path_reachable(), poorly.

c) so does path_is_under(), while we are at it.
Signed-off-by: NAl Viro <viro@zeniv.linux.org.uk>

afac7cba

vfs: new internal helper: mnt_has_parent(mnt) · b2dba1af

由 Al Viro 提交于 11月 23, 2011

vfsmounts have ->mnt_parent pointing either to a different vfsmount
or to itself; it's never NULL and termination condition in loops
traversing the tree towards root is mnt == mnt->mnt_parent.  At least
one place (see the next patch) is confused about what's going on;
let's add an explicit helper checking it right way and use it in
all places where we need it.  Not that there had been too many,
but...
Signed-off-by: NAl Viro <viro@zeniv.linux.org.uk>

b2dba1af

vfs: kill pointless helpers in namespace.c · aa9c0e07

由 Al Viro 提交于 11月 23, 2011

mnt_{inc,dec}_count() is not cleaner than doing the corresponding
mnt_add_count() directly and mnt_set_count() is not used at all.
Signed-off-by: NAl Viro <viro@zeniv.linux.org.uk>

aa9c0e07

07 12月, 2011 1 次提交

fix apparmor dereferencing potentially freed dentry, sanitize __d_path() API · 02125a82

由 Al Viro 提交于 12月 05, 2011

__d_path() API is asking for trouble and in case of apparmor d_namespace_path()
getting just that.  The root cause is that when __d_path() misses the root
it had been told to look for, it stores the location of the most remote ancestor
in *root.  Without grabbing references.  Sure, at the moment of call it had
been pinned down by what we have in *path.  And if we raced with umount -l, we
could have very well stopped at vfsmount/dentry that got freed as soon as
prepend_path() dropped vfsmount_lock.

It is safe to compare these pointers with pre-existing (and known to be still
alive) vfsmount and dentry, as long as all we are asking is "is it the same
address?".  Dereferencing is not safe and apparmor ended up stepping into
that.  d_namespace_path() really wants to examine the place where we stopped,
even if it's not connected to our namespace.  As the result, it looked
at ->d_sb->s_magic of a dentry that might've been already freed by that point.
All other callers had been careful enough to avoid that, but it's really
a bad interface - it invites that kind of trouble.

The fix is fairly straightforward, even though it's bigger than I'd like:
	* prepend_path() root argument becomes const.
	* __d_path() is never called with NULL/NULL root.  It was a kludge
to start with.  Instead, we have an explicit function - d_absolute_root().
Same as __d_path(), except that it doesn't get root passed and stops where
it stops.  apparmor and tomoyo are using it.
	* __d_path() returns NULL on path outside of root.  The main
caller is show_mountinfo() and that's precisely what we pass root for - to
skip those outside chroot jail.  Those who don't want that can (and do)
use d_path().
	* __d_path() root argument becomes const.  Everyone agrees, I hope.
	* apparmor does *NOT* try to use __d_path() or any of its variants
when it sees that path->mnt is an internal vfsmount.  In that case it's
definitely not mounted anywhere and dentry_path() is exactly what we want
there.  Handling of sysctl()-triggered weirdness is moved to that place.
	* if apparmor is asked to do pathname relative to chroot jail
and __d_path() tells it we it's not in that jail, the sucker just calls
d_absolute_path() instead.  That's the other remaining caller of __d_path(),
BTW.
        * seq_path_root() does _NOT_ return -ENAMETOOLONG (it's stupid anyway -
the normal seq_file logics will take care of growing the buffer and redoing
the call of ->show() just fine).  However, if it gets path not reachable
from root, it returns SEQ_SKIP.  The only caller adjusted (i.e. stopped
ignoring the return value as it used to do).
Reviewed-by: NJohn Johansen <john.johansen@canonical.com>
ACKed-by: NJohn Johansen <john.johansen@canonical.com>
Signed-off-by: NAl Viro <viro@zeniv.linux.org.uk>
Cc: stable@vger.kernel.org

02125a82

23 11月, 2011 1 次提交

mount_subtree() pointless use-after-free · d31da0f0

由 Al Viro 提交于 11月 22, 2011

d'oh... we'd carefully pinned mnt->mnt_sb down, dropped mnt and attempt
to grab s_umount on mnt->mnt_sb.  The trouble is, *mnt might've been
overwritten by now...
Signed-off-by: NAl Viro <viro@zeniv.linux.org.uk>

d31da0f0

17 11月, 2011 2 次提交

new helper: mount_subtree() · ea441d11

由 Al Viro 提交于 11月 16, 2011

takes vfsmount and relative path, does lookup within that vfsmount
(possibly triggering automounts) and returns the result as root
of subtree suitable for return by ->mount() (i.e. a reference to
dentry and an active reference to its superblock grabbed, superblock
locked exclusive).

btrfs and nfs switched to it instead of open-coding the sucker.
Signed-off-by: NAl Viro <viro@zeniv.linux.org.uk>

ea441d11

switch create_mnt_ns() to saner calling conventions, fix double mntput() in nfs · c1334495

由 Al Viro 提交于 11月 16, 2011

Life is much saner if create_mnt_ns(mnt) drops mnt in case of error...
Switch it to such calling conventions, switch callers, fix double mntput() in
fs/nfs/super.c one.
Signed-off-by: NAl Viro <viro@zeniv.linux.org.uk>

c1334495

28 10月, 2011 1 次提交

vfs: add "device" tag to /proc/self/mountstats · a877ee03

由 Bryan Schumaker 提交于 10月 07, 2011

nfsiostat was failing to find mounted filesystems on kernels after
2.6.38 because of changes to show_vfsstat() by commit
c7f404b4.  This patch adds back the
"device" tag before the nfs server entry so scripts can parse the
mountstats file correctly.
Signed-off-by: NBryan Schumaker <bjschuma@netapp.com>
CC: stable@kernel.org [>=2.6.39]
Signed-off-by: NChristoph Hellwig <hch@lst.de>

a877ee03

27 9月, 2011 1 次提交

VFS: Fix the remaining automounter semantics regressions · 815d405c

由 Trond Myklebust 提交于 9月 26, 2011

The concensus seems to be that system calls such as stat() etc should
not trigger an automount.  Neither should the l* versions.

This patch therefore adds a LOOKUP_AUTOMOUNT flag to tag those lookups
that _should_ trigger an automount on the last path element.
Signed-off-by: NTrond Myklebust <Trond.Myklebust@netapp.com>
[ Edited to leave out the cases that are already covered by LOOKUP_OPEN,
  LOOKUP_DIRECTORY and LOOKUP_CREATE - all of which also fundamentally
  force automounting for their own reasons   - Linus ]
Signed-off-by: NLinus Torvalds <torvalds@linux-foundation.org>

815d405c

24 7月, 2011 1 次提交

VFS : mount lock scalability for internal mounts · 423e0ab0

由 Tim Chen 提交于 7月 19, 2011

For a number of file systems that don't have a mount point (e.g. sockfs
and pipefs), they are not marked as long term. Therefore in
mntput_no_expire, all locks in vfs_mount lock are taken instead of just
local cpu's lock to aggregate reference counts when we release
reference to file objects.  In fact, only local lock need to have been
taken to update ref counts as these file systems are in no danger of
going away until we are ready to unregister them.

The attached patch marks file systems using kern_mount without
mount point as long term.  The contentions of vfs_mount lock
is now eliminated.  Before un-registering such file system,
kern_unmount should be called to remove the long term flag and
make the mount point ready to be freed.
Signed-off-by: NTim Chen <tim.c.chen@linux.intel.com>
Signed-off-by: NAl Viro <viro@zeniv.linux.org.uk>

423e0ab0

21 7月, 2011 1 次提交

fs: seq_file - add event counter to simplify poll() support · f1514638

由 Kay Sievers 提交于 7月 12, 2011

Moving the event counter into the dynamically allocated 'struc seq_file'
allows poll() support without the need to allocate its own tracking
structure.

All current users are switched over to use the new counter.

Requested-by: Andrew Morton akpm@linux-foundation.org
Acked-by: NNeilBrown <neilb@suse.de>
Tested-by: Lucas De Marchi lucas.demarchi@profusion.mobi
Signed-off-by: NKay Sievers <kay.sievers@vrfy.org>
Signed-off-by: NAl Viro <viro@zeniv.linux.org.uk>

f1514638

26 5月, 2011 1 次提交

fs/namespace.c: bound mount propagation fix · 7c6e984d

由 Roman Borisov 提交于 5月 25, 2011

This issue was discovered by users of busybox.  And the bug is actual for
busybox users, I don't know how it affects others.  Apparently, mount is
called with and without MS_SILENT, and this affects mount() behaviour.
But MS_SILENT is only supposed to affect kernel logging verbosity.

The following script was run in an empty test directory:

mkdir -p mount.dir mount.shared1 mount.shared2
touch mount.dir/a mount.dir/b
mount -vv --bind         mount.shared1 mount.shared1
mount -vv --make-rshared mount.shared1
mount -vv --bind         mount.shared2 mount.shared2
mount -vv --make-rshared mount.shared2
mount -vv --bind mount.shared2 mount.shared1
mount -vv --bind mount.dir     mount.shared2
ls -R mount.dir mount.shared1 mount.shared2
umount mount.dir mount.shared1 mount.shared2 2>/dev/null
umount mount.dir mount.shared1 mount.shared2 2>/dev/null
umount mount.dir mount.shared1 mount.shared2 2>/dev/null
rm -f mount.dir/a mount.dir/b mount.dir/c
rmdir mount.dir mount.shared1 mount.shared2

mount -vv was used to show the mount() call arguments and result.
Output shows that flag argument has 0x00008000 = MS_SILENT bit:

mount: mount('mount.shared1','mount.shared1','(null)',0x00009000,'(null)'):0
mount: mount('','mount.shared1','',0x0010c000,''):0
mount: mount('mount.shared2','mount.shared2','(null)',0x00009000,'(null)'):0
mount: mount('','mount.shared2','',0x0010c000,''):0
mount: mount('mount.shared2','mount.shared1','(null)',0x00009000,'(null)'):0
mount: mount('mount.dir','mount.shared2','(null)',0x00009000,'(null)'):0
mount.dir:
a
b

mount.shared1:

mount.shared2:
a
b

After adding --loud option to remove MS_SILENT bit from just one mount cmd:

mkdir -p mount.dir mount.shared1 mount.shared2
touch mount.dir/a mount.dir/b
mount -vv --bind         mount.shared1 mount.shared1 2>&1
mount -vv --make-rshared mount.shared1               2>&1
mount -vv --bind         mount.shared2 mount.shared2 2>&1
mount -vv --loud --make-rshared mount.shared2               2>&1  # <-HERE
mount -vv --bind mount.shared2 mount.shared1         2>&1
mount -vv --bind mount.dir     mount.shared2         2>&1
ls -R mount.dir mount.shared1 mount.shared2      2>&1
umount mount.dir mount.shared1 mount.shared2 2>/dev/null
umount mount.dir mount.shared1 mount.shared2 2>/dev/null
umount mount.dir mount.shared1 mount.shared2 2>/dev/null
rm -f mount.dir/a mount.dir/b mount.dir/c
rmdir mount.dir mount.shared1 mount.shared2

The result is different now - look closely at mount.shared1 directory listing.
Now it does show files 'a' and 'b':

mount: mount('mount.shared1','mount.shared1','(null)',0x00009000,'(null)'):0
mount: mount('','mount.shared1','',0x0010c000,''):0
mount: mount('mount.shared2','mount.shared2','(null)',0x00009000,'(null)'):0
mount: mount('','mount.shared2','',0x00104000,''):0
mount: mount('mount.shared2','mount.shared1','(null)',0x00009000,'(null)'):0
mount: mount('mount.dir','mount.shared2','(null)',0x00009000,'(null)'):0

mount.dir:
a
b

mount.shared1:
a
b

mount.shared2:
a
b

The analysis shows that MS_SILENT flag which is ON by default in any
busybox-> mount operations cames to flags_to_propagation_type function and
causes the error return while is_power_of_2 checking because the function
expects only one bit set.  This doesn't allow to do busybox->mount with
any --make-[r]shared, --make-[r]private etc options.

Moreover, the recently added flags_to_propagation_type() function doesn't
allow us to do such operations as --make-[r]private --make-[r]shared etc.
when MS_SILENT is on.  The idea or clearing the MS_SILENT flag came from
to Denys Vlasenko.
Signed-off-by: NRoman Borisov <ext-roman.borisov@nokia.com>
Reported-by: NDenys Vlasenko <vda.linux@googlemail.com>
Cc: Chuck Ebbert <cebbert@redhat.com>
Cc: Alexander Shishkin <virtuoso@slind.org>
Cc: Al Viro <viro@zeniv.linux.org.uk>
Cc: Christoph Hellwig <hch@lst.de>
Signed-off-by: NAndrew Morton <akpm@linux-foundation.org>
Signed-off-by: NAl Viro <viro@zeniv.linux.org.uk>

7c6e984d

13 4月, 2011 1 次提交

Revert "vfs: Export file system uuid via /proc/<pid>/mountinfo" · be85bcca

由 Linus Torvalds 提交于 4月 12, 2011

This reverts commit 93f1c20b.

It turns out that libmount misparses it because it adds a '-' character
in the uuid string, which libmount then incorrectly confuses with the
separator string (" - ") at the end of all the optional arguments.

Upstream libmount (in the util-linux tree) has been fixed, but until
that fix actually percolates up to users, we'd better not expose this
change in the kernel.

Let's revisit this later (possibly by exposing the UUID without any '-'
characters in it, avoiding the user-space bug).
Reported-by: NDave Jones <davej@redhat.com>
Cc: Aneesh Kumar K.V <aneesh.kumar@linux.vnet.ibm.com>
Cc: Al Viro <viro@zeniv.linux.org.uk>
Cc: Karel Zak <kzak@redhat.com>
Cc: Ram Pai <linuxram@us.ibm.com>
Cc: Miklos Szeredi <mszeredi@suse.cz>
Cc: Eric Sandeen <sandeen@redhat.com>
Signed-off-by: NLinus Torvalds <torvalds@linux-foundation.org>

be85bcca

23 3月, 2011 1 次提交

fs: use appropriate printk priority levels · 80cdc6da

由 Mandeep Singh Baines 提交于 3月 22, 2011

printk()s without a priority level default to KERN_WARNING.  To reduce
noise at KERN_WARNING, this patch set the priority level appriopriately
for unleveled printks()s.  This should be useful to folks that look at
dmesg warnings closely.
Signed-off-by: NMandeep Singh Baines <msb@chromium.org>
Cc: Jens Axboe <axboe@kernel.dk>
Cc: Al Viro <viro@zeniv.linux.org.uk>
Signed-off-by: NAndrew Morton <akpm@linux-foundation.org>
Signed-off-by: NLinus Torvalds <torvalds@linux-foundation.org>

80cdc6da

18 3月, 2011 4 次提交

change the locking order for namespace_sem · b12cea91

由 Al Viro 提交于 3月 18, 2011

Have it nested inside ->i_mutex.  Instead of using follow_down()
under namespace_sem, followed by grabbing i_mutex and checking that
mountpoint to be is not dead, do the following:
	grab i_mutex
	check that it's not dead
	grab namespace_sem
	see if anything is mounted there
	if not, we've won
	otherwise
		drop locks
		put_path on what we had
		replace with what's mounted
		retry everything with new mountpoint to be

New helper (lock_mount()) does that.  do_add_mount(), do_move_mount(),
do_loopback() and pivot_root() switched to it; in case of the last
two that eliminates a race we used to have - original code didn't
do follow_down().
Signed-off-by: NAl Viro <viro@zeniv.linux.org.uk>

b12cea91

fix deadlock in pivot_root() · 27cb1572

由 Al Viro 提交于 3月 18, 2011

Don't hold vfsmount_lock over the loop traversing ->mnt_parent;
do check_mnt(new.mnt) under namespace_sem instead; combined with
namespace_sem held over all that code it'll guarantee the stability
of ->mnt_parent chain all the way to the root.

Doing check_mnt() outside of namespace_sem in case of pivot_root()
is wrong anyway.
Signed-off-by: NAl Viro <viro@zeniv.linux.org.uk>

27cb1572

vfs: split off vfsmount-related parts of vfs_kern_mount() · 9d412a43

由 Al Viro 提交于 3月 17, 2011

new function: mount_fs().  Does all work done by vfs_kern_mount()
except the allocation and filling of vfsmount; returns root dentry
or ERR_PTR().

vfs_kern_mount() switched to using it and taken to fs/namespace.c,
along with its wrappers.

alloc_vfsmnt()/free_vfsmnt() made static.

functions in namespace.c slightly reordered.
Signed-off-by: NAl Viro <viro@zeniv.linux.org.uk>

9d412a43

kill simple_set_mnt() · 474a00ee

由 Al Viro 提交于 3月 17, 2011

not needed anymore, since all users (->get_sb() instances) are gone.
Signed-off-by: NAl Viro <viro@zeniv.linux.org.uk>

474a00ee

17 3月, 2011 1 次提交

vfs: new superblock methods to override /proc/*/mount{s,info} · c7f404b4

由 Al Viro 提交于 3月 16, 2011

a) ->show_devname(m, mnt) - what to put into devname columns in mounts,
mountinfo and mountstats
b) ->show_path(m, mnt) - what to put into relative path column in mountinfo

Leaving those NULL gives old behaviour.  NFS switched to using those.
Signed-off-by: NAl Viro <viro@zeniv.linux.org.uk>

c7f404b4