Eric Lee / smarc-fsl-linux-kernel

12 Aug, 2008

1 commit

23a0ee908 Merge branch 'core/locking' into core/urgent Browse Code »

Ingo Molnar
2008-08-12 06:11:49 +0800

11 Aug, 2008

2 commits

3295f0ef9 lockdep: rename map_[acquire|release]() => lock_map_[acquire|release]() ... Browse Code »

the names were too generic:

drivers/uio/uio.c:87: error: expected identifier or '(' before 'do'
drivers/uio/uio.c:87: error: expected identifier or '(' before 'while'
drivers/uio/uio.c:113: error: 'map_release' undeclared here (not in a function)

Signed-off-by: Ingo Molnar

Ingo Molnar
2008-08-11 16:30:30 +0800
4f3e7524b lockdep: map_acquire ... Browse Code »

Most the free-standing lock_acquire() usages look remarkably similar, sweep
them into a new helper.

Signed-off-by: Peter Zijlstra
Signed-off-by: Ingo Molnar

Peter Zijlstra
2008-08-11 15:30:23 +0800

05 Aug, 2008

1 commit

529ae9aaa mm: rename page trylock ... Browse Code »

Converting page lock to new locking bitops requires a change of page flag
operation naming, so we might as well convert it to something nicer
(!TestSetPageLocked_Lock => trylock_page, SetPageLocked => set_page_locked).

This also facilitates lockdeping of page lock.

Signed-off-by: Nick Piggin
Acked-by: KOSAKI Motohiro
Acked-by: Peter Zijlstra
Acked-by: Andrew Morton
Acked-by: Benjamin Herrenschmidt
Signed-off-by: Linus Torvalds

Nick Piggin
2008-08-05 12:31:34 +0800

01 Aug, 2008

1 commit

e9e34f4e8 jbd2: don't abort if flushing file data failed ... Browse Code »

In ordered mode, the current jbd2 aborts the journal if a file data buffer
has an error. But this behavior is unintended, and we found that it has
been adopted accidentally.

This patch undoes it and just calls printk() instead of aborting the
journal. Unlike a similar patch for ext3/jbd, file data buffers are
written via generic_writepages(). But we also need to set AS_EIO
into their mappings because wait_on_page_writeback_range() clears
AS_EIO before a user process sees it.

Signed-off-by: Hidehiro Kawai
Signed-off-by: "Theodore Ts'o"

Hidehiro Kawai
2008-08-01 10:26:04 +0800

27 Jul, 2008

1 commit

00b32b7fb ext4: unexport jbd2_journal_update_superblock ... Browse Code »

Remove the unused EXPORT_SYMBOL(jbd2_journal_update_superblock).

Signed-off-by: Adrian Bunk
Signed-off-by: "Theodore Ts'o"

Theodore Ts'o
2008-07-27 05:33:53 +0800

14 Jul, 2008

1 commit

530576bbf jbd2: fix race between jbd2_journal_try_to_free_buffers() and jbd2 commit transaction ... Browse Code »

journal_try_to_free_buffers() could race with jbd commit transaction
when the later is holding the buffer reference while waiting for the
data buffer to flush to disk. If the caller of
journal_try_to_free_buffers() request tries hard to release the buffers,
it will treat the failure as error and return back to the caller. We
have seen the directo IO failed due to this race. Some of the caller of
releasepage() also expecting the buffer to be dropped when passed with
GFP_KERNEL mask to the releasepage()->journal_try_to_free_buffers().

With this patch, if the caller is passing the GFP_KERNEL to indicating
this call could wait, in case of try_to_free_buffers() failed, let's
waiting for journal_commit_transaction() to finish commit the current
committing transaction , then try to free those buffers again with
journal locked.

Signed-off-by: Mingming Cao
Reviewed-by: Badari Pulavarty
Signed-off-by: "Theodore Ts'o"

Mingming Cao
2008-07-14 09:06:39 +0800

12 Jul, 2008

4 commits

cd1aac329 ext4: Add ordered mode support for delalloc ... Browse Code »

This provides a new ordered mode implementation which gets rid of using
buffer heads to enforce the ordering between metadata change with the
related data chage. Instead, in the new ordering mode, it keeps track
of all of the inodes touched by each transaction on a list, and when
that transaction is committed, it flushes all of the dirty pages for
those inodes. In addition, the new ordered mode reverses the lock
ordering of the page lock and transaction lock, which provides easier
support for delayed allocation.

Signed-off-by: Aneesh Kumar K.V
Signed-off-by: Mingming Cao
Signed-off-by: "Theodore Ts'o"

Aneesh Kumar K.V
2008-07-12 07:27:31 +0800
87c89c232 jbd2: Remove data=ordered mode support using jbd buffer heads ... Browse Code »

Signed-off-by: Jan Kara

Jan Kara
2008-07-12 07:27:31 +0800
c851ed540 jbd2: Implement data=ordered mode handling via inodes ... Browse Code »

This patch adds necessary framework into JBD2 to be able to track inodes
with each transaction and write-out their dirty data during transaction
commit time.

This new ordered mode brings all sorts of advantages such as possibility
to get rid of journal heads and buffer heads for data buffers in ordered
mode, better ordering of writes on transaction commit, simplification of
some JBD code, no more anonymous pages when truncate of data being
committed happens. Also with this new ordered mode, delayed allocation
on ordered mode is much simpler.

Signed-off-by: Jan Kara

Jan Kara
2008-07-12 07:27:31 +0800
736603ab2 jbd2: Add commit time into the commit block ... Browse Code »

Carlo Wood has demonstrated that it's possible to recover deleted
files from the journal. Something that will make this easier is if we
can put the time of the commit into commit block.

Signed-off-by: "Theodore Ts'o"

Theodore Ts'o
2008-07-12 07:27:31 +0800

07 Jun, 2008

1 commit

624080ede jbd2: If a journal checksum error is detected, propagate the error to ext4 ... Browse Code »

If a journal checksum error is detected, the ext4 filesystem will call
ext4_error(), and the mount will either continue, become a read-only
mount, or cause a kernel panic based on the superblock flags
indicating the user's preference of what to do in case of filesystem
corruption being detected.

Signed-off-by: "Theodore Ts'o"

Theodore Ts'o
2008-06-07 05:50:40 +0800

04 Jun, 2008

1 commit

034772b06 jbd2: Fix barrier fallback code to re-lock the buffer head ... Browse Code »

If the device doesn't support write barriers, the write is retried
without ordered mode. But the buffer head needs to be re-locked or
submit_bh will fail with on BUG(!buffer_locked(bh)).

Signed-off-by: "Theodore Ts'o"

Theodore Ts'o
2008-06-04 10:31:11 +0800

26 May, 2008

1 commit

8ea76900b jbd2: Fix memory leak when verifying checksums in the journal ... Browse Code »

Cc: Andreas Dilger
Cc: Girish Shilamkar
Signed-off-by: "Theodore Ts'o"

Theodore Ts'o
2008-05-26 22:28:09 +0800

16 May, 2008

1 commit

02c471cb1 jbd2: update transaction t_state to T_COMMIT fix ... Browse Code »

Updating the current transaction's t_state is protected by j_state_lock. We
need to do the same when updating the t_state to T_COMMIT.

Acked-by: Jan Kara
Signed-off-by: Mingming Cao
Signed-off-by: "Theodore Ts'o"
Signed-off-by: Andrew Morton

Mingming Cao
2008-05-16 02:46:17 +0800

13 May, 2008

1 commit

f36f21ecc Fix misuses of bdevname() ... Browse Code »

bdevname() fills the buffer that it is given as a parameter, so calling
strcpy() or snprintf() on the returned value is redundant (and probably not
guaranteed to work - I don't think strcpy and snprintf support overlapping
buffers.)

Signed-off-by: Jean Delvare
Cc: Stephen Tweedie
Cc: Jens Axboe
Signed-off-by: Andrew Morton
Signed-off-by: Linus Torvalds

Jean Delvare
2008-05-13 23:02:26 +0800

30 Apr, 2008

1 commit

620de4e19 jbd2: only create debugfs and stats entries if init is successful ... Browse Code »

jbd2 debugfs and stats entries should only be created if cache initialisation
is successful. At the moment they are being created unconditionally which
will leave them dangling if cache (and hence module) initialisation fails.

Signed-off-by: Duane Griffin
Cc:
Signed-off-by: Andrew Morton
Signed-off-by: Mingming Cao
Signed-off-by: "Theodore Ts'o"

Duane Griffin
2008-04-30 10:02:47 +0800

29 Apr, 2008

1 commit

79da3664f jbd2: use non-racy method for proc entries creation ... Browse Code »

Use proc_create()/proc_create_data() to make sure that ->proc_fops and ->data
be setup before gluing PDE to main tree.

Signed-off-by: Denis V. Lunev
Cc:
Cc: Alexey Dobriyan
Cc: "Eric W. Biederman"
Signed-off-by: Andrew Morton
Signed-off-by: Linus Torvalds

Denis V. Lunev
2008-04-29 23:06:20 +0800

28 Apr, 2008

1 commit

9fa27c85d jbd2: tidy up revoke cache initialisation and destruction ... Browse Code »

Make revocation cache destruction safe to call if initialisation fails
partially or entirely. This allows it to be used to cleanup in the case of
initialisation failure, simplifying the code slightly.

Signed-off-by: Duane Griffin
Signed-off-by: Mingming Cao
Signed-off-by: "Theodore Ts'o"
Signed-off-by: Aneesh Kumar K.V

Duane Griffin
2008-04-28 21:40:00 +0800

17 Apr, 2008

6 commits

329d291f5 jdb2: replace remaining __FUNCTION__ occurrences ... Browse Code »

__FUNCTION__ is gcc-specific, use __func__

Signed-off-by: Harvey Harrison
Cc:
Signed-off-by: Andrew Morton
Signed-off-by: "Theodore Ts'o"

Harvey Harrison
2008-04-17 22:38:59 +0800
5648ba5b2 jbd2: fix kernel-doc notation ... Browse Code »

Fix kernel-doc notation in jbd2.

Signed-off-by: Randy Dunlap
Signed-off-by: Mingming Cao
Signed-off-by: Andrew Morton
Signed-off-by: "Theodore Ts'o"

Randy Dunlap
2008-04-17 22:38:59 +0800
8a9362eb4 jbd2: replace potentially false assertion with if block ... Browse Code »

If an error occurs during jbd2 cache initialisation it is possible for the
journal_head_cache to be NULL when jbd2_journal_destroy_journal_head_cache is
called. Replace the J_ASSERT with an if block to handle the situation
correctly.

Note that even with this fix things will break badly if jbd2 is statically
compiled in and cache initialisation fails.

Signed-off-by: Duane Griffin
Cc:
Signed-off-by: Andrew Morton
Signed-off-by: Mingming Cao
Signed-off-by: "Theodore Ts'o"

Duane Griffin
2008-04-17 22:38:59 +0800
83c49523c jbd2: eliminate duplicated code in revocation table init/destroy functions ... Browse Code »

The revocation table initialisation/destruction code is repeated for each of
the two revocation tables stored in the journal. Refactoring the duplicated
code into functions is tidier, simplifies the logic in initialisation in
particular, and slightly reduces the code size.

There should not be any functional change.

Signed-off-by: Duane Griffin
Signed-off-by: Mingming Cao
Signed-off-by: "Theodore Ts'o"

Duane Griffin
2008-04-17 22:38:59 +0800
1dfc3220d jbd2: fix possible journal overflow issues ... Browse Code »

There are several cases where the running transaction can get buffers
added to its BJ_Metadata list which it never dirtied, which makes its
t_nr_buffers counter end up larger than its t_outstanding_credits
counter.

This will cause issues when starting new transactions as while we are
logging buffers we decrement t_outstanding_buffers, so when
t_outstanding_buffers goes negative, we will report that we need less
space in the journal than we actually need, so transactions will be
started even though there may not be enough room for them. In the worst
case scenario (which admittedly is almost impossible to reproduce) this
will result in the journal running out of space.

The fix is to only refile buffers from the committing transaction to the
running transactions BJ_Modified list when b_modified is set on that
journal, which is the only way to be sure if the running transaction has
modified that buffer.

This patch also fixes an accounting error in journal_forget, it is
possible that we can call journal_forget on a buffer without having
modified it, only gotten write access to it, so instead of freeing a
credit, we only do so if the buffer was modified. The assert will help
catch if this problem occurs. Without these two patches I could hit
this assert within minutes of running postmark, with them this issue no
longer arises.

Cc:
Cc: Jan Kara
Signed-off-by: Josef Bacik
Signed-off-by: Andrew Morton
Signed-off-by: "Theodore Ts'o"

Josef Bacik
2008-04-17 22:38:59 +0800
9fc7c63a1 jbd2: fix the way the b_modified flag is cleared ... Browse Code »

Currently at the start of a journal commit we loop through all of the buffers
on the committing transaction and clear the b_modified flag (the flag that is
set when a transaction modifies the buffer) under the j_list_lock.

The problem is that everywhere else this flag is modified only under the jbd2
lock buffer flag, so it will race with a running transaction who could
potentially set it, and have it unset by the committing transaction.

This is also a big waste, you can have several thousands of buffers that you
are clearing the modified flag on when you may not need to. This patch
removes this code and instead clears the b_modified flag upon entering
do_get_write_access/journal_get_create_access, so if that transaction does
indeed use the buffer then it will be accounted for properly, and if it does
not then we know we didn't use it.

That will be important for the next patch in this series. Tested thoroughly
by myself using postmark/iozone/bonnie++.

Cc:
Cc: Jan Kara
Signed-off-by: Josef Bacik
Signed-off-by: Andrew Morton
Signed-off-by: "Theodore Ts'o"

Josef Bacik
2008-04-17 22:38:59 +0800

31 Mar, 2008

1 commit

1076d17ac jbd/jbd2 NULL noise ... Browse Code »

Signed-off-by: Al Viro
Signed-off-by: Linus Torvalds

Al Viro
2008-03-31 05:18:41 +0800

20 Mar, 2008

1 commit

d00256766 jbd2: correctly unescape journal data blocks ... Browse Code »

Fix a long-standing typo (predating git) that will cause data corruption if a
journal data block needs unescaping. At the moment the wrong buffer head's
data is being unescaped.

To test this case mount a filesystem with data=journal, start creating and
deleting a bunch of files containing only JBD2_MAGIC_NUMBER (0xc03b3998), then
pull the plug on the device. Without this patch the files will contain zeros
instead of the correct data after recovery.

Signed-off-by: Duane Griffin
Acked-by: Jan Kara
Cc:
Cc:
Signed-off-by: Andrew Morton
Signed-off-by: Linus Torvalds

Duane Griffin
2008-03-20 09:53:36 +0800

10 Feb, 2008

1 commit

c4e35e07a JBD2: Clear buffer_ordered flag for barried IO request on success ... Browse Code »

In JBD2 jbd2_journal_write_commit_record(), clear the buffer_ordered
flag for the bh after barried IO has succeed. This prevents later, if
the same buffer head were submitted to the underlying device, which has
been reconfigured to not support barrier request, the JBD2 commit code
could treat it as a normal IO (without barrier).

This is a port from JBD/ext3 fix from Neil Brown.

More details from Neil:

Some devices - notably dm and md - can change their behaviour in
response to BIO_RW_BARRIER requests. They might start out accepting
such requests but on reconfiguration, they find out that they cannot
any more. JBD2 deal with this by always testing if BIO_RW_BARRIER
requests fail with EOPNOTSUPP, and retrying the write
requests without the barrier (probably after waiting for any pending
writes to complete).

However there is a bug in the handling this in JBD2 for ext4 .

When ext4/JBD2 to submit a BIO_RW_BARRIER request,
it sets the buffer_ordered flag on the buffer head.
If the request completes successfully, the flag STAYS SET.

Other code might then write the same buffer_head after the device has
been reconfigured to not accept barriers. This write will then fail,
but the "other code" is not ready to handle EOPNOTSUPP errors and the
error will be treated as fatal.

Cc: Neil Brown
Signed-off-by: Dave Kleikamp
Signed-off-by: Mingming Cao
Signed-off-by: "Theodore Ts'o"

Dave Kleikamp
2008-02-10 14:09:32 +0800

07 Feb, 2008

1 commit

e86e14385 BKL-removal: remove incorrect comment refering to lock_kernel() from jbd/jbd2 ... Browse Code »

None of the callers of this function does actually take the BKL as far as I
can see. So remove the comment refering to the BKL.

Signed-off-by: Andi Kleen
Cc:
Cc: Theodore Ts'o
Signed-off-by: Andrew Morton
Signed-off-by: Linus Torvalds

Andi Kleen
2008-02-07 02:41:20 +0800

05 Feb, 2008

3 commits

4d6051797 JBD2: Use the incompat macro for testing the incompat feature. ... Browse Code »

JBD2_FEATURE_INCOMPAT_ASYNC_COMMIT needs to be checked with
JBD2_HAS_INCOMPAT_FEATURE

Signed-off-by: Aneesh Kumar K.V
Signed-off-by: Mingming Cao
Signed-off-by: "Theodore Ts'o"

Aneesh Kumar K.V
2008-02-05 23:56:15 +0800
c4b8e635f jbd2: Fix reference counting on the journal commit block's buffer head ... Browse Code »

With journal checksum patch we added asynchronous commits of journal
commit headers, and accidentally dropped taking a reference on the
buffer head.

(Before the change, sync_dirty_buffer did the get_bh(). The associative
put_bh is done by journal_wait_on_commit_record().)

Signed-off-by: Aneesh Kumar K.V
Signed-off-by: Mingming Cao
Signed-off-by: "Theodore Ts'o"

Aneesh Kumar K.V
2008-02-05 23:55:26 +0800
b048d8462 jbd2: Add error check to journal_wait_on_commit_record to avoid oops ... Browse Code »

The buffer head pointer passed to journal_wait_on_commit_record() could
be NULL if the previous journal_submit_commit_record() failed or journal
has already aborted.

Looking at the jbd2 debug messages, before the oops happened, the jbd2
is aborted due to trying to access the next log block beyond the end
of device. This might be caused by using a corrupted image.

We need to check the error returns from journal_submit_commit_record()
and avoid calling journal_wait_on_commit_record() in the failure case.

This addresses Kernel Bugzilla #9849

Signed-off-by: Mingming Cao
Signed-off-by: "Theodore Ts'o"

Mingming Cao
2008-02-05 21:52:45 +0800

30 Jan, 2008

1 commit

95c354fe9 spinlock: lockbreak cleanup ... Browse Code »

The break_lock data structure and code for spinlocks is quite nasty.
Not only does it double the size of a spinlock but it changes locking to
a potentially less optimal trylock.

Put all of that under CONFIG_GENERIC_LOCKBREAK, and introduce a
__raw_spin_is_contended that uses the lock data itself to determine whether
there are waiters on the lock, to be used if CONFIG_GENERIC_LOCKBREAK is
not set.

Rename need_lockbreak to spin_needbreak, make it use spin_is_contended to
decouple it from the spinlock implementation, and make it typesafe (rwlocks
do not have any need_lockbreak sites -- why do they even get bloated up
with that break_lock then?).

Signed-off-by: Nick Piggin
Signed-off-by: Ingo Molnar
Signed-off-by: Thomas Gleixner

Nick Piggin
2008-01-30 20:31:20 +0800

29 Jan, 2008

7 commits

4019191be jbd2: sparse pointer use of zero as null ... Browse Code »

Get rid of sparse related warnings from places that use integer as NULL
pointer. (Ported from upstream ext3/jbd changes.)

Signed-off-by: Mingming Cao
Signed-off-by: "Theodore Ts'o"

Mingming Cao
2008-01-29 12:58:27 +0800
db857da33 jbd2: Use round-jiffies() function for the "5 second" ext4/jbd2 wakeup ... Browse Code »

While "every 5 seconds" doesn't sound as a problem, there can be many
of these (and these timers do add up over all the kernel). The "5
second" wakeup isn't really timing sensitive; in addition even with
rounding it'll still happen every 5 seconds (with the exception of the
very first time, which is likely to be rounded up to somewhere closer
to 6 seconds)

(Ported from similar JBD patch made by Arjan van de Ven to
fs/jbd/transaction.c)

Cc: Arjan van de Ven
Cc: Andrew Morton
Signed-off-by: Mingming Cao
Signed-off-by: "Theodore Ts'o"

Mingming Cao
2008-01-29 12:58:27 +0800
77160957e jbd2: Mark jbd2 slabs as SLAB_TEMPORARY ... Browse Code »

This patch marks slab allocations by jbd2 as short-lived in support of
Mel Gorman's "Group short-lived and reclaimable kernel allocations"
patch. (Ported from similar changes made to fs/jbd/journal.c and
fs/jbd/revoke.c in Mel's patch.)

Cc: Mel Gorman
Cc: Andrew Morton
Signed-off-by: Mingming Cao
Signed-off-by: "Theodore Ts'o"

Mingming Cao
2008-01-29 12:58:27 +0800
7b7510662 jbd2: add lockdep support ... Browse Code »

Ported from similar patch for the jbd layer.

Signed-off-by: Mingming Cao
Signed-off-by: "Theodore Ts'o"

Mingming Cao
2008-01-29 12:58:27 +0800
818d276ce ext4: Add the journal checksum feature ... Browse Code »

The journal checksum feature adds two new flags i.e
JBD2_FEATURE_INCOMPAT_ASYNC_COMMIT and JBD2_FEATURE_COMPAT_CHECKSUM.

JBD2_FEATURE_CHECKSUM flag indicates that the commit block contains the
checksum for the blocks described by the descriptor blocks.
Due to checksums, writing of the commit record no longer needs to be
synchronous. Now commit record can be sent to disk without waiting for
descriptor blocks to be written to disk. This behavior is controlled
using JBD2_FEATURE_ASYNC_COMMIT flag. Older kernels/e2fsck should not be
able to recover the journal with _ASYNC_COMMIT hence it is made
incompat.
The commit header has been extended to hold the checksum along with the
type of the checksum.

For recovery in pass scan checksums are verified to ensure the sanity
and completeness(in case of _ASYNC_COMMIT) of every transaction.

Signed-off-by: Andreas Dilger
Signed-off-by: Girish Shilamkar
Signed-off-by: Dave Kleikamp
Signed-off-by: Mingming Cao

Girish Shilamkar
2008-01-29 12:58:27 +0800
8e85fb3f3 jbd2: jbd2 stats through procfs ... Browse Code »

The patch below updates the jbd stats patch to 2.6.20/jbd2.
The initial patch was posted by Alex Tomas in December 2005
(http://marc.info/?l=linux-ext4&m=113538565128617&w=2).
It provides statistics via procfs such as transaction lifetime and size.

Sometimes, investigating performance problems, i find useful to have
stats from jbd about transaction's lifetime, size, etc. here is a
patch for review and inclusion probably.

for example, stats after creation of 3M files in htree directory:

[root@bob ~]# cat /proc/fs/jbd/sda/history
R/C tid wait run lock flush log hndls block inlog ctime write drop close
R 261 8260 2720 0 0 750 9892 8170 8187
C 259 750 0 4885 1
R 262 20 2200 10 0 770 9836 8170 8187
R 263 30 2200 10 0 3070 9812 8170 8187
R 264 0 5000 10 0 1340 0 0 0
C 261 8240 3212 4957 0
R 265 8260 1470 0 0 4640 9854 8170 8187
R 266 0 5000 10 0 1460 0 0 0
C 262 8210 2989 4868 0
R 267 8230 1490 10 0 4440 9875 8171 8188
R 268 0 5000 10 0 1260 0 0 0
C 263 7710 2937 4908 0
R 269 7730 1470 10 0 3330 9841 8170 8187
R 270 0 5000 10 0 830 0 0 0
C 265 8140 3234 4898 0
C 267 720 0 4849 1
R 271 8630 2740 20 0 740 9819 8170 8187
C 269 800 0 4214 1
R 272 40 2170 10 0 830 9716 8170 8187
R 273 40 2280 0 0 3530 9799 8170 8187
R 274 0 5000 10 0 990 0 0 0

where,

R - line for transaction's life from T_RUNNING to T_FINISHED
C - line for transaction's checkpointing
tid - transaction's id
wait - for how long we were waiting for new transaction to start
(the longest period journal_start() took in this transaction)
run - real transaction's lifetime (from T_RUNNING to T_LOCKED
lock - how long we were waiting for all handles to close
(time the transaction was in T_LOCKED)
flush - how long it took to flush all data (data=ordered)
log - how long it took to write the transaction to the log
hndls - how many handles got to the transaction
block - how many blocks got to the transaction
inlog - how many blocks are written to the log (block + descriptors)
ctime - how long it took to checkpoint the transaction
write - how many blocks have been written during checkpointing
drop - how many blocks have been dropped during checkpointing
close - how many running transactions have been closed to checkpoint this one

all times are in msec.

[root@bob ~]# cat /proc/fs/jbd/sda/info
280 transaction, each upto 8192 blocks
average:
1633ms waiting for transaction
3616ms running transaction
5ms transaction was being locked
1ms flushing data (in ordered mode)
1799ms logging transaction
11781 handles per transaction
5629 blocks per transaction
5641 logged blocks per transaction

Signed-off-by: Johann Lombardi
Signed-off-by: Mariusz Kozlowski
Signed-off-by: Mingming Cao
Signed-off-by: Eric Sandeen

Johann Lombardi
2008-01-29 12:58:27 +0800
f5a7a6b0d jbd2: Fix assertion failure in fs/jbd2/checkpoint.c ... Browse Code »

Before we start committing a transaction, we call
__journal_clean_checkpoint_list() to cleanup transaction's written-back
buffers.

If this call happens to remove all of them (and there were already some
buffers), __journal_remove_checkpoint() will decide to free the transaction
because it isn't (yet) a committing transaction and soon we fail some
assertion - the transaction really isn't ready to be freed :).

We change the check in __journal_remove_checkpoint() to free only a
transaction in T_FINISHED state. The locking there is subtle though (as
everywhere in JBD ;(). We use j_list_lock to protect the check and a
subsequent call to __journal_drop_transaction() and do the same in the end
of journal_commit_transaction() which is the only place where a transaction
can get to T_FINISHED state.

Probably I'm too paranoid here and such locking is not really necessary -
checkpoint lists are processed only from log_do_checkpoint() where a
transaction must be already committed to be processed or from
__journal_clean_checkpoint_list() where kjournald itself calls it and thus
transaction cannot change state either. Better be safe if something
changes in future...

Signed-off-by: Jan Kara
Cc:
Signed-off-by: Andrew Morton

Jan Kara
2008-01-29 12:58:27 +0800