From: Linus Torvalds <torvalds@linux-foundation.org>
To: Ingo Molnar <mingo@elte.hu>
Cc: Frederic Weisbecker <fweisbec@gmail.com>,
LKML <linux-kernel@vger.kernel.org>,
Jeff Mahoney <jeffm@suse.com>,
ReiserFS Development List <reiserfs-devel@vger.kernel.org>,
Chris Mason <chris.mason@oracle.com>,
Alexander Beregalov <a.beregalov@gmail.com>,
Alessio Igor Bogani <abogani@texware.it>,
Jonathan Corbet <corbet@lwn.net>,
Alexander Viro <viro@zeniv.linux.org.uk>
Subject: Re: [PATCH 0/6] kill-the-BKL/reiserfs3: performance improvements, faster than Bkl based scheme
Date: Fri, 1 May 2009 14:11:31 -0700 (PDT) [thread overview]
Message-ID: <alpine.LFD.2.00.0905011350500.5379@localhost.localdomain> (raw)
In-Reply-To: <20090501203639.GA9396@elte.hu>
On Fri, 1 May 2009, Ingo Molnar wrote:
> * Ingo Molnar <mingo@elte.hu> wrote:
>
> > The BKL is clearly removed at a faster reate with such debugging
> > measures in place. With such measures the BKL _really_ hurts, and
> > very visibly so - and that results in active removal.
>
> Btw., if you can think of any way to create a cleaner debug tool
> here i'd be glad to start a new tree that adds only _that_ tool -
> and that could perhaps be dealt with independently and possibly
> pulled upstream too, if clean enough.
Quite frankly, for something as clearly defined as as a filesystem, I
would literally suggest starting out with a few trivial stages:
Stage 1:
- raw search-and-replace of [un]lock_kernel()
sed 's/unlock_kernel()/reiserfs_unlock(sb)/g'
sed 's/lock_kernel()/reiserfs_lock(sb)/g'
and then obviously fix it up so that 'sb' always exists (ie you would
need to add a a few
struct super_block *sb = inode->i_sb;
lines manually to make the end result work.
- add that 'reiserfs_[un]lock()' pair around every single VFS entry-point
that currently gets called with lock_kernel(). They'd _still_ get
called with lock-kernel too (you're not fixing that), but now they'd
_also_ take the new lock.
- make reiserfs_unlock(sb) just do [un]lock_kernel(), and use that as a
known good starting point. Nothing has really changed (and you in fact
_increased_ the nesting on the BKL, since now the BKL is gotten twice
in those areas that were called from the VFS with the BKL held), but
you now have a mindless and pretty trivial conversion where you still
have a working system, and nothing really changed - except you now have
the beginnings of a nice abstraction.
IOW, "stage 1" is all totally mindless, and never breaks anything. It's
very much designed to be "obviously correct".
Stage 2:
- switch over reiserfs_[un]lock() to a per-superblock mutex, and enable
lockdep, and start looking for nesting problems.
IOW, "stage 2" is when you make a _minimal_ change to now enable all the
good lockdep infrastructure. But it's also designed to just do _one_
thing: look at nesting. Nothing else.
Stage 3:
- once you've fixed all nesting problems (well, most of them - enable a
swapfile on that filesystem and I bet you'll find a few more), you
might want to look at performance issues. In particular, you want to
drop the lock over at least the most _common_ blocking operations, if
at all possible.
And stage 3 is where it would make absolutely _tons_ of sense to just add
some very simple infrastructure for having a per-thread "IO warning
counter", and then simply increment/decrement that counter in
reiserfs_[un]lock().
But exactly because we do NOT want to get warnings about BKL use in all
the _other_ subsystems, we don't want to mix this up with the BKL counter
that we already have. You could probably even use the tracing
infrastructure to do this: enable tracing in reiserfs_lock(), disable it
in reiserfs_unlock(), and look at what you catch in between.
I dunno. That's how I would go about it if I just wanted to clean up a
specific subsystem, and didn't want to tie it together with getting
everything else right too.
Linus
next prev parent reply other threads:[~2009-05-01 21:18 UTC|newest]
Thread overview: 39+ messages / expand[flat|nested] mbox.gz Atom feed top
2009-05-01 2:44 Frederic Weisbecker
2009-05-01 2:44 ` [PATCH 1/6] kill-the-BKL/reiserfs: release write lock on fs_changed() Frederic Weisbecker
2009-05-01 6:31 ` Andi Kleen
2009-05-01 13:28 ` Frederic Weisbecker
2009-05-01 13:44 ` Chris Mason
2009-05-01 14:01 ` Frederic Weisbecker
2009-05-01 14:14 ` Chris Mason
2009-05-02 1:19 ` Frederic Weisbecker
2009-05-01 2:44 ` [PATCH 2/6] kill-the-BKL/reiserfs: release the write lock before rescheduling on do_journal_end() Frederic Weisbecker
2009-05-01 7:09 ` Ingo Molnar
2009-05-01 13:31 ` Frederic Weisbecker
2009-05-01 22:20 ` Frederic Weisbecker
2009-05-01 2:44 ` [PATCH 3/6] kill-the-BKL/reiserfs: release write lock while rescheduling on prepare_for_delete_or_cut() Frederic Weisbecker
2009-05-01 2:44 ` [PATCH 4/6] kill-the-BKL/reiserfs: release the write lock inside get_neighbors() Frederic Weisbecker
2009-05-01 5:51 ` Ingo Molnar
2009-05-01 13:25 ` Frederic Weisbecker
2009-05-01 13:29 ` Chris Mason
2009-05-01 13:31 ` Ingo Molnar
2009-05-01 2:44 ` [PATCH 5/6] kill-the-BKL/reiserfs: release the write lock inside reiserfs_read_bitmap_block() Frederic Weisbecker
2009-05-01 5:47 ` Ingo Molnar
2009-05-01 13:19 ` Frederic Weisbecker
2009-05-01 13:30 ` Chris Mason
2009-05-01 13:51 ` Frederic Weisbecker
2009-05-01 2:44 ` [PATCH 6/6] kill-the-BKL/reiserfs: release the write lock on flush_commit_list() Frederic Weisbecker
2009-05-01 5:42 ` Ingo Molnar
2009-05-01 13:13 ` Frederic Weisbecker
2009-05-01 13:23 ` Ingo Molnar
2009-05-01 13:26 ` Chris Mason
2009-05-01 13:29 ` Ingo Molnar
2009-05-01 13:54 ` Chris Mason
2009-05-01 5:35 ` [PATCH 0/6] kill-the-BKL/reiserfs3: performance improvements, faster than Bkl based scheme Ingo Molnar
2009-05-01 12:18 ` Thomas Meyer
2009-05-01 14:12 ` Frederic Weisbecker
2009-05-01 19:59 ` Linus Torvalds
2009-05-01 20:33 ` Ingo Molnar
2009-05-01 20:36 ` Ingo Molnar
2009-05-01 21:11 ` Linus Torvalds [this message]
2009-05-01 21:32 ` Ingo Molnar
2009-05-02 1:39 ` Frederic Weisbecker
Reply instructions:
You may reply publicly to this message via plain-text email
using any one of the following methods:
* Save the following mbox file, import it into your mail client,
and reply-to-all from there: mbox
Avoid top-posting and favor interleaved quoting:
https://en.wikipedia.org/wiki/Posting_style#Interleaved_style
* Reply using the --to, --cc, and --in-reply-to
switches of git-send-email(1):
git send-email \
--in-reply-to=alpine.LFD.2.00.0905011350500.5379@localhost.localdomain \
--to=torvalds@linux-foundation.org \
--cc=a.beregalov@gmail.com \
--cc=abogani@texware.it \
--cc=chris.mason@oracle.com \
--cc=corbet@lwn.net \
--cc=fweisbec@gmail.com \
--cc=jeffm@suse.com \
--cc=linux-kernel@vger.kernel.org \
--cc=mingo@elte.hu \
--cc=reiserfs-devel@vger.kernel.org \
--cc=viro@zeniv.linux.org.uk \
/path/to/YOUR_REPLY
https://kernel.org/pub/software/scm/git/docs/git-send-email.html
* If your mail client supports setting the In-Reply-To header
via mailto: links, try the mailto: link
Be sure your reply has a Subject: header at the top and a blank line
before the message body.
This is a public inbox, see mirroring instructions
for how to clone and mirror all data and code used for this inbox
all inboxes | Powered by JetHome®