Eric Lee / smarc-fsl-linux-kernel

12 Jun, 2009

40 commits

964f53696 fs/qnx4: sanitize includes ... Browse Code »

fs-internal parts of qnx4_fs.h taken to fs/qnx4/qnx4.h, includes adjusted,
qnx4_fs.h doesn't need unifdef anymore.

Signed-off-by: Al Viro

Al Viro
2009-06-12 09:36:12 +0800
79d257675 Sanitize qnx4 fsync handling ... Browse Code »

* have directory operations use mark_buffer_dirty_inode(),
so that sync_mapping_buffers() would get those.
* make qnx4_write_inode() honour its last argument.
* get rid of insane copies of very ancient "walk the indirect blocks"
in qnx4/fsync - they never matched the actual fs layout and, fortunately,
never'd been called. Again, all this junk is not needed; ->fsync()
should just do sync_mapping_buffers + sync_inode (and if we implement
block allocation for qnx4, we'll need to use mark_buffer_dirty_inode()
for extent blocks)

Signed-off-by: Al Viro

Al Viro
2009-06-12 09:36:11 +0800
d5aacad54 New helper - simple_fsync() ... Browse Code »

writes associated buffers, then does sync_inode() to write
the inode itself (and to make it clean). Depends on
->write_inode() honouring the second argument.

Signed-off-by: Al Viro

Al Viro
2009-06-12 09:36:11 +0800
8688b8635 linux/magic.h: move cramfs magic out of cramfs_fs.h ... Browse Code »

Signed-off-by: Mike Frysinger
CC: Alexander Viro
Signed-off-by: Al Viro

Mike Frysinger
2009-06-12 09:36:10 +0800
28ad0c118 fs: Rearrange inode structure elements to avoid waste due to padding ... Browse Code »

Signed-off-by: "Theodore Ts'o"
Cc: linux-fsdevel@vger.kernel.org
Signed-off-by: Al Viro

Theodore Ts'o
2009-06-12 09:36:09 +0800
9fd5746fd fs: Remove i_cindex from struct inode ... Browse Code »

The only user of the i_cindex element in the inode structure is used
is by the firewire drivers. As part of an attempt to slim down the
inode structure to save memory --- since a typical Linux system will
have hundreds of thousands if not millions of inodes cached, a
reduction in the size inode has high leverage.

The firewire driver does not need i_cindex in any fast path, so it's
simple enough to calculate when it is needed, instead of wasting space
in the inode structure.

Signed-off-by: "Theodore Ts'o"
Cc: krh@redhat.com
Cc: stefanr@s5r6.in-berlin.de
Cc: linux-fsdevel@vger.kernel.org
Signed-off-by: Al Viro

Theodore Ts'o
2009-06-12 09:36:09 +0800
62c6943b4 Trim a bit of crap from fs.h ... Browse Code »

do_remount_sb() is fs/internal.h fodder, fsync_no_super() is long gone.

Signed-off-by: Al Viro

Al Viro
2009-06-12 09:36:07 +0800
f3da392e9 dcache: extrace and use d_unlinked() ... Browse Code »

d_unlinked() will be used in middle-term to ban checkpointing when opened
but unlinked file is detected, and in long term, to detect such situation
and special case on it.

Signed-off-by: Alexey Dobriyan
Signed-off-by: Al Viro

Alexey Dobriyan
2009-06-12 09:36:06 +0800
c3f8a40c1 quota: Introduce writeout_quota_sb() (version 4) ... Browse Code »

Introduce this function which just writes all the quota structures but
avoids all the syncing and cache pruning work to expose quota structures
to userspace. Use this function from __sync_filesystem when wait == 0.

Signed-off-by: Jan Kara
Signed-off-by: Al Viro

Jan Kara
2009-06-12 09:36:04 +0800
850b201b0 quota: cleanup dquota sync functions (version 4) ... Browse Code »

Currently the VFS calls vfs_dq_sync to sync out disk quotas for a given
superblock. This is a small wrapper around sync_dquots which for the
case of a non-NULL superblock is a small wrapper around quota_sync_sb.

Just make quota_sync_sb global (rename it to sync_quota_sb) and call it
directly. Also call it directly for those cases in quota.c that have a
superblock and leave sync_dquots purely an iterator over sync_quota_sb and
remove it's superblock argument.

To make this nicer move the check for the lack of a quota_sync method
from the callers into sync_quota_sb.

[folded build fix from Alexander Beregalov ]

Signed-off-by: Christoph Hellwig
Signed-off-by: Jan Kara
Signed-off-by: Al Viro

Christoph Hellwig
2009-06-12 09:36:04 +0800
60b0680fa vfs: Rename fsync_super() to sync_filesystem() (version 4) ... Browse Code »

Rename the function so that it better describe what it really does. Also
remove the unnecessary include of buffer_head.h.

Signed-off-by: Jan Kara
Signed-off-by: Al Viro

Jan Kara
2009-06-12 09:36:04 +0800
c15c54f5f vfs: Move syncing code from super.c to sync.c (version 4) ... Browse Code »

Move sync_filesystems(), __fsync_super(), fsync_super() from
super.c to sync.c where it fits better.

[build fixes folded]

Signed-off-by: Jan Kara
Signed-off-by: Al Viro

Jan Kara
2009-06-12 09:36:04 +0800
5cee5815d vfs: Make sys_sync() use fsync_super() (version 4) ... Browse Code »

It is unnecessarily fragile to have two places (fsync_super() and do_sync())
doing data integrity sync of the filesystem. Alter __fsync_super() to
accommodate needs of both callers and use it. So after this patch
__fsync_super() is the only place where we gather all the calls needed to
properly send all data on a filesystem to disk.

Nice bonus is that we get a complete livelock avoidance and write_supers()
is now only used for periodic writeback of superblocks.

sync_blockdevs() introduced a couple of patches ago is gone now.

[build fixes folded]

Signed-off-by: Jan Kara
Signed-off-by: Al Viro

Jan Kara
2009-06-12 09:36:03 +0800
429479f03 vfs: Make __fsync_super() a static function (version 4) ... Browse Code »

__fsync_super() does the same thing as fsync_super(). So change the only
caller to use fsync_super() and make __fsync_super() static. This removes
unnecessarily duplicated call to sync_blockdev() and prepares ground
for the changes to __fsync_super() in the following patches.

Signed-off-by: Jan Kara
Signed-off-by: Al Viro

Jan Kara
2009-06-12 09:36:03 +0800
876a9f76a remove s_async_list ... Browse Code »

Remove the unused s_async_list in the superblock, a leftover of the
broken async inode deletion code that leaked into mainline. Having this
in the middle of the sync/unmount path is not helpful for the following
cleanups.

Signed-off-by: Christoph Hellwig
Signed-off-by: Al Viro

Christoph Hellwig
2009-06-12 09:36:02 +0800
96029c4e0 fs: introduce mnt_clone_write ... Browse Code »

This patch speeds up lmbench lat_mmap test by about another 2% after the
first patch.

Before:
avg = 462.286
std = 5.46106

After:
avg = 453.12
std = 9.58257

(50 runs of each, stddev gives a reasonable confidence)

It does this by introducing mnt_clone_write, which avoids some heavyweight
operations of mnt_want_write if called on a vfsmount which we know already
has a write count; and mnt_want_write_file, which can call mnt_clone_write
if the file is open for write.

After these two patches, mnt_want_write and mnt_drop_write go from 7% on
the profile down to 1.3% (including mnt_clone_write).

[AV: mnt_want_write_file() should take file alone and derive mnt from it;
not only all callers have that form, but that's the only mnt about which
we know that it's already held for write if file is opened for write]

Cc: Dave Hansen
Signed-off-by: Nick Piggin
Signed-off-by: Al Viro

npiggin@suse.de
2009-06-12 09:36:02 +0800
d3ef3d735 fs: mnt_want_write speedup ... Browse Code »

This patch speeds up lmbench lat_mmap test by about 8%. lat_mmap is set up
basically to mmap a 64MB file on tmpfs, fault in its pages, then unmap it.
A microbenchmark yes, but it exercises some important paths in the mm.

Before:
avg = 501.9
std = 14.7773

After:
avg = 462.286
std = 5.46106

(50 runs of each, stddev gives a reasonable confidence, but there is quite
a bit of variation there still)

It does this by removing the complex per-cpu locking and counter-cache and
replaces it with a percpu counter in struct vfsmount. This makes the code
much simpler, and avoids spinlocks (although the msync is still pretty
costly, unfortunately). It results in about 900 bytes smaller code too. It
does increase the size of a vfsmount, however.

It should also give a speedup on large systems if CPUs are frequently operating
on different mounts (because the existing scheme has to operate on an atomic in
the struct vfsmount when switching between mounts). But I'm most interested in
the single threaded path performance for the moment.

[AV: minor cleanup]

Cc: Dave Hansen
Signed-off-by: Nick Piggin
Signed-off-by: Al Viro

npiggin@suse.de
2009-06-12 09:36:02 +0800
3174c21b7 Move junk from proc_fs.h to fs/proc/internal.h ... Browse Code »

Signed-off-by: Al Viro

Al Viro
2009-06-12 09:36:01 +0800
1c755af4d switch lookup_mnt() ... Browse Code »

Signed-off-by: Al Viro

Al Viro
2009-06-12 09:36:01 +0800
9393bd07c switch follow_down() ... Browse Code »

Signed-off-by: Al Viro

Al Viro
2009-06-12 09:36:01 +0800
589ff870e Switch collect_mounts() to struct path ... Browse Code »

Signed-off-by: Al Viro

Al Viro
2009-06-12 09:36:01 +0800
bab77ebf5 switch follow_up() to struct path ... Browse Code »

Signed-off-by: Al Viro

Al Viro
2009-06-12 09:36:00 +0800
e64c390ca switch rqst_exp_parent() ... Browse Code »

Signed-off-by: Al Viro

Al Viro
2009-06-12 09:36:00 +0800
91c9fa8f7 switch rqst_exp_get_by_name() ... Browse Code »

Signed-off-by: Al Viro

Al Viro
2009-06-12 09:36:00 +0800
2a7378711 Cache root in nameidata ... Browse Code »

New field: nd->root. When pathname resolution wants to know the root,
check if nd->root.mnt is non-NULL; use nd->root if it is, otherwise
copy current->fs->root there. After path_walk() is finished, we check
if we'd got a cached value in nd->root and drop it. Before calling
path_walk() we should either set nd->root.mnt to NULL *or* copy (and
pin down) some path to nd->root. In the latter case we won't be
looking at current->fs->root at all.

Signed-off-by: Al Viro

Al Viro
2009-06-12 09:35:59 +0800
73422811d reiserfs: allow exposing privroot w/ xattrs enabled ... Browse Code »

This patch adds an -oexpose_privroot option to allow access to the privroot.

Signed-off-by: Jeff Mahoney
Signed-off-by: Al Viro

Jeff Mahoney
2009-06-12 09:35:58 +0800
3bb66d7f8 Merge branch 'for-linus' of git://git.infradead.org/users/eparis/notify ... Browse Code »

* 'for-linus' of git://git.infradead.org/users/eparis/notify:
fsnotify: allow groups to set freeing_mark to null
inotify/dnotify: should_send_event shouldn't match on FS_EVENT_ON_CHILD
dnotify: do not bother to lock entry->lock when reading mask
dnotify: do not use ?true:false when assigning to a bool
fsnotify: move events should indicate the event was on a child
inotify: reimplement inotify using fsnotify
fsnotify: handle filesystem unmounts with fsnotify marks
fsnotify: fsnotify marks on inodes pin them in core
fsnotify: allow groups to add private data to events
fsnotify: add correlations between events
fsnotify: include pathnames with entries when possible
fsnotify: generic notification queue and waitq
dnotify: reimplement dnotify using fsnotify
fsnotify: parent event notification
fsnotify: add marks to inodes so groups can interpret how to handle those inodes
fsnotify: unified filesystem notification backend

Linus Torvalds
2009-06-12 05:22:55 +0800
512626a04 Merge branch 'for-linus' of git://linux-arm.org/linux-2.6 ... Browse Code »

* 'for-linus' of git://linux-arm.org/linux-2.6:
kmemleak: Add the corresponding MAINTAINERS entry
kmemleak: Simple testing module for kmemleak
kmemleak: Enable the building of the memory leak detector
kmemleak: Remove some of the kmemleak false positives
kmemleak: Add modules support
kmemleak: Add kmemleak_alloc callback from alloc_large_system_hash
kmemleak: Add the vmalloc memory allocation/freeing hooks
kmemleak: Add the slub memory allocation/freeing hooks
kmemleak: Add the slob memory allocation/freeing hooks
kmemleak: Add the slab memory allocation/freeing hooks
kmemleak: Add documentation on the memory leak detector
kmemleak: Add the base support

Manual conflict resolution (with the slab/earlyboot changes) in:
drivers/char/vt.c
init/main.c
mm/slab.c

Linus Torvalds
2009-06-12 05:15:57 +0800
8a1ca8ced Merge branch 'perfcounters-for-linus' of git://git.kernel.org/pub/scm/linux/kern… ... Browse Code »

…el/git/tip/linux-2.6-tip

* 'perfcounters-for-linus' of git://git.kernel.org/pub/scm/linux/kernel/git/tip/linux-2.6-tip: (574 commits)
perf_counter: Turn off by default
perf_counter: Add counter->id to the throttle event
perf_counter: Better align code
perf_counter: Rename L2 to LL cache
perf_counter: Standardize event names
perf_counter: Rename enums
perf_counter tools: Clean up u64 usage
perf_counter: Rename perf_counter_limit sysctl
perf_counter: More paranoia settings
perf_counter: powerpc: Implement generalized cache events for POWER processors
perf_counters: powerpc: Add support for POWER7 processors
perf_counter: Accurate period data
perf_counter: Introduce struct for sample data
perf_counter tools: Normalize data using per sample period data
perf_counter: Annotate exit ctx recursion
perf_counter tools: Propagate signals properly
perf_counter tools: Small frequency related fixes
perf_counter: More aggressive frequency adjustment
perf_counter/x86: Fix the model number of Intel Core2 processors
perf_counter, x86: Correct some event and umask values for Intel processors
...

Linus Torvalds
2009-06-12 05:01:07 +0800
b640f042f Merge branch 'topic/slab/earlyboot' of git://git.kernel.org/pub/scm/linux/kernel… ... Browse Code »

…/git/penberg/slab-2.6

* 'topic/slab/earlyboot' of git://git.kernel.org/pub/scm/linux/kernel/git/penberg/slab-2.6:
vgacon: use slab allocator instead of the bootmem allocator
irq: use kcalloc() instead of the bootmem allocator
sched: use slab in cpupri_init()
sched: use alloc_cpumask_var() instead of alloc_bootmem_cpumask_var()
memcg: don't use bootmem allocator in setup code
irq/cpumask: make memoryless node zero happy
x86: remove some alloc_bootmem_cpumask_var calling
vt: use kzalloc() instead of the bootmem allocator
sched: use kzalloc() instead of the bootmem allocator
init: introduce mm_init()
vmalloc: use kzalloc() instead of alloc_bootmem()
slab: setup allocators earlier in the boot sequence
bootmem: fix slab fallback on numa
bootmem: use slab if bootmem is no longer available

Linus Torvalds
2009-06-12 03:25:06 +0800
ff52cc215 fsnotify: move events should indicate the event was on a child ... Browse Code »

fsnotify tells its listeners explicitly when an event happened on the given
inode verses on the child of the given inode. (see __fsnotify_parent)
However, the semantics of fsnotify_move() are such that we deliver events
directly to the two parent directories in question (old_dir and new_dir)
directly without using the __fsnotify_parent() call. fsnotify should be
adding FS_EVENT_ON_CHILD for the notifications to these parents.

Signed-off-by: Eric Paris

Eric Paris
2009-06-12 02:57:54 +0800
63c882a05 inotify: reimplement inotify using fsnotify ... Browse Code »

Reimplement inotify_user using fsnotify. This should be feature for feature
exactly the same as the original inotify_user. This does not make any changes
to the in kernel inotify feature used by audit. Those patches (and the eventual
removal of in kernel inotify) will come after the new inotify_user proves to be
working correctly.

Signed-off-by: Eric Paris
Acked-by: Al Viro
Cc: Christoph Hellwig

Eric Paris
2009-06-12 02:57:54 +0800
164bc6195 fsnotify: handle filesystem unmounts with fsnotify marks ... Browse Code »

When an fs is unmounted with an fsnotify mark entry attached to one of its
inodes we need to destroy that mark entry and we also (like inotify) send
an unmount event.

Signed-off-by: Eric Paris
Acked-by: Al Viro
Cc: Christoph Hellwig

Eric Paris
2009-06-12 02:57:54 +0800
e4aff1173 fsnotify: allow groups to add private data to events ... Browse Code »

inotify needs per group information attached to events. This patch allows
groups to attach private information and implements a callback so that
information can be freed when an event is being destroyed.

Signed-off-by: Eric Paris
Acked-by: Al Viro
Cc: Christoph Hellwig

Eric Paris
2009-06-12 02:57:54 +0800
47882c6f5 fsnotify: add correlations between events ... Browse Code »

As part of the standard inotify events it includes a correlation cookie
between two dentry move operations. This patch includes the same behaviour
in fsnotify events. It is needed so that inotify userspace can be
implemented on top of fsnotify.

Signed-off-by: Eric Paris
Acked-by: Al Viro
Cc: Christoph Hellwig

Eric Paris
2009-06-12 02:57:54 +0800
62ffe5dfb fsnotify: include pathnames with entries when possible ... Browse Code »

When inotify wants to send events to a directory about a child it includes
the name of the original file. This patch collects that filename and makes
it available for notification.

Signed-off-by: Eric Paris
Acked-by: Al Viro
Cc: Christoph Hellwig

Eric Paris
2009-06-12 02:57:53 +0800
a2d8bc6cb fsnotify: generic notification queue and waitq ... Browse Code »

inotify needs to do asyc notification in which event information is stored
on a queue until the listener is ready to receive it. This patch
implements a generic notification queue for inotify (and later fanotify) to
store events to be sent at a later time.

Signed-off-by: Eric Paris
Acked-by: Al Viro
Cc: Christoph Hellwig

Eric Paris
2009-06-12 02:57:53 +0800
3c5119c05 dnotify: reimplement dnotify using fsnotify ... Browse Code »

Reimplement dnotify using fsnotify.

Signed-off-by: Eric Paris
Acked-by: Al Viro
Cc: Christoph Hellwig

Eric Paris
2009-06-12 02:57:53 +0800
c28f7e56e fsnotify: parent event notification ... Browse Code »

inotify and dnotify both use a similar parent notification mechanism. We
add a generic parent notification mechanism to fsnotify for both of these
to use. This new machanism also adds the dentry flag optimization which
exists for inotify to dnotify.

Signed-off-by: Eric Paris
Acked-by: Al Viro
Cc: Christoph Hellwig

Eric Paris
2009-06-12 02:57:53 +0800
3be25f49b fsnotify: add marks to inodes so groups can interpret how to handle those inodes ... Browse Code »

This patch creates a way for fsnotify groups to attach marks to inodes.
These marks have little meaning to the generic fsnotify infrastructure
and thus their meaning should be interpreted by the group that attached
them to the inode's list.

dnotify and inotify will make use of these markings to indicate which
inodes are of interest to their respective groups. But this implementation
has the useful property that in the future other listeners could actually
use the marks for the exact opposite reason, aka to indicate which inodes
it had NO interest in.

Signed-off-by: Eric Paris
Acked-by: Al Viro
Cc: Christoph Hellwig

Eric Paris
2009-06-12 02:57:53 +0800