Commits · 2600d2dd5085ab6fb09540226138a60055abf335 · nexedi / linux

23 Feb, 2010 4 commits

ceph: drop messages on unregistered mds sessions; cleanup · 2600d2dd

Sage Weil authored Feb 22, 2010

Verify the mds session is currently registered before handling
incoming messages.  Clean up message handlers to pull mds out
of session->s_mds instead of less trustworthy src field.

Clean up con_{get,put} debug output.
Signed-off-by: Sage Weil <sage@newdream.net>

2600d2dd

ceph: fix comments, locking in destroy_inode · a6369741

Sage Weil authored Feb 22, 2010

The destroy_inode path needs no inode locks since there are no
inode references.  Update __ceph_remove_cap comment to reflect
that it is called without cap->session->s_mutex in this case.
Signed-off-by: Sage Weil <sage@newdream.net>

a6369741

ceph: move dereference after NULL test · 4ce1e9ad

Alexander Beregalov authored Feb 22, 2010

Signed-off-by: Alexander Beregalov <a.beregalov@gmail.com>
Signed-off-by: Sage Weil <sage@newdream.net>

4ce1e9ad

ceph: fix up unexpected message handling · 5b3a4db3

Sage Weil authored Feb 19, 2010

Fix skipping of unexpected message types from osd, mon.

Clean up pr_info and debug output.
Signed-off-by: Sage Weil <sage@newdream.net>

5b3a4db3

19 Feb, 2010 4 commits

ceph: cleanup redundant code in handle_cap_grant · bcd2cbd1

Yehuda Sadeh authored Feb 19, 2010

There is no state in local vars that requires us to loop after temporarily
dropping i_lock.
Signed-off-by: Yehuda Sadeh <yehuda@hq.newdream.net>
Signed-off-by: Sage Weil <sage@newdream.net>

bcd2cbd1

ceph: don't truncate dirty pages in invalidate work thread · c9af9fb6

Yehuda Sadeh authored Feb 19, 2010

Instead of truncating the whole range of pages, we skip those
pages that are dirty or in the middle of writeback. Those pages
will be cleared later when the writeback completes.
Signed-off-by: Yehuda Sadeh <yehuda@hq.newdream.net>
Signed-off-by: Sage Weil <sage@newdream.net>

c9af9fb6

ceph: remove page upon writeback completion if lost cache cap · e63dc5c7

Yehuda Sadeh authored Feb 19, 2010

This page should have been removed earlier when the cache cap was
revoked, but a writeback was in flight, so it was skipped. We truncate
it here just as the writeback finishes, while it's still locked.
Signed-off-by: Yehuda Sadeh <yehuda@hq.newdream.net>
Signed-off-by: Sage Weil <sage@newdream.net>

e63dc5c7

ceph: fix check for invalidate_mapping_pages success · 5ecad6fd

Sage Weil authored Feb 17, 2010

We need to know whether there was any page left behind, and not the
return value (the total number of pages invalidated).  Look at the mapping
to see if we were successful or not.

Move it all into a helper to simplify the two callers.
Signed-off-by: Yehuda Sadeh <yehuda@hq.newdream.net>
Signed-off-by: Sage Weil <sage@newdream.net>

5ecad6fd

17 Feb, 2010 12 commits

ceph: fix typo in ceph_queue_writeback debug output · 2c27c9a5
Sage Weil authored Feb 17, 2010
```
Signed-off-by: Sage Weil <sage@newdream.net>
```
2c27c9a5
ceph: v0.19 release · a17d6473
Sage Weil authored Feb 17, 2010
```
Signed-off-by: Sage Weil <sage@newdream.net>
```
a17d6473

ceph: use rbtree for pg pools; decode new osdmap format · 4fc51be8

Sage Weil authored Feb 16, 2010

Since we can now create and destroy pg pools, the pool ids will be sparse,
and an array no longer makes sense for looking up by pool id.  Use an
rbtree instead.

The OSDMap encoding also no longer has a max pool count (previously used to
allocate the array).  There is a new pool_max, that is the largest pool id
we've ever used, although we don't actually need it in the client.
Signed-off-by: Sage Weil <sage@newdream.net>

4fc51be8

ceph: fix memory leak when destroying osdmap with pg_temp mappings · 9794b146
Sage Weil authored Feb 16, 2010
```
Also move _lookup_pg_mapping into a helper.
Signed-off-by: Sage Weil <sage@newdream.net>
```
9794b146

ceph: fix iterate_caps removal race · 7c1332b8

Sage Weil authored Feb 16, 2010

We need to be able to iterate over all caps on a session with a
possibly slow callback on each cap.  To allow this, we used to
prevent cap reordering while we were iterating.  However, we were
not safe from races with removal: removing the 'next' cap would
make the next pointer from list_for_each_entry_safe be invalid,
and cause a lock up or similar badness.

Instead, we keep an iterator pointer in the session pointing to
the current cap.  As before, we avoid reordering.  For removal,
if the cap isn't the current cap we are iterating over, we are
fine.  If it is, we clear cap->ci (to mark the cap as pending
removal) but leave it in the session list.  In iterate_caps, we
can safely finish removal and get the next cap pointer.

While we're at it, clean up put_cap to not take a cap reservation
context, as it was never used.
Signed-off-by: Sage Weil <sage@newdream.net>

7c1332b8

ceph: clean up readdir caps reservation · 85ccce43

Sage Weil authored Feb 17, 2010

Use a global counter for the minimum number of allocated caps instead of
hard coding a check against readdir_max.  This takes into account multiple
client instances, and avoids examining the superblock mount options when a
cap is dropped.
Signed-off-by: Sage Weil <sage@newdream.net>

85ccce43

ceph: fix authentication races, auth_none oops · 5ce6e9db

Sage Weil authored Feb 15, 2010

Call __validate_auth() under monc->mutex, and use helper for
initial hello so that the pending_auth flag is set.  This fixes
possible races in which we have an authentication request (hello
or otherwise) pending and send another one.  In particular, with
auth_none, we _never_ want to call ceph_build_auth() from
__validate_auth(), since the ->build_request() method is NULL.
Signed-off-by: Sage Weil <sage@newdream.net>

5ce6e9db

ceph: use rbtree for mon statfs requests · 85ff03f6

Sage Weil authored Feb 15, 2010

An rbtree is lighter weight, particularly given we will generally have
very few in-flight statfs requests.
Signed-off-by: Sage Weil <sage@newdream.net>

85ff03f6

ceph: use rbtree for snap_realms · a105f00c

Sage Weil authored Feb 15, 2010

Switch from radix tree to rbtree for snap realms.  This is much more
appropriate given that realm keys are few and far between.
Signed-off-by: Sage Weil <sage@newdream.net>

a105f00c

ceph: use rbtree for mds requests · 44ca18f2

Sage Weil authored Feb 15, 2010

The rbtree is a more appropriate data structure than a radix_tree.  It
avoids extra memory usage and simplifies the code.

It also fixes a bug where the debugfs 'mdsc' file wasn't including the
most recent mds request.
Signed-off-by: Sage Weil <sage@newdream.net>

44ca18f2

ceph: cancel delayed work when closing connection · 91e45ce3

Sage Weil authored Feb 15, 2010

This ensures that if/when we reopen the connection, we can requeue work on
the connection immediately, without waiting for an old timer to expire.
Queue new delayed work inside con->mutex to avoid any race.

This fixes problems with clients failing to reconnect to the MDS due to
the client_reconnect message arriving too late (due to waiting for an old
delayed work timeout to expire).
Signed-off-by: Sage Weil <sage@newdream.net>

91e45ce3

ceph: allow connection to be reopened by fault callback · e2663ab6

Sage Weil authored Feb 16, 2010

Fix the messenger to allow a ceph_con_open() during the fault callback.
Previously the work wasn't getting queued on the connection because the
fault path avoids requeued work (normally spurious).  Loop on reopening by
checking for the OPENING state bit.

This fixes OSD reconnects when a TCP connection drops.
Signed-off-by: Sage Weil <sage@newdream.net>

e2663ab6

15 Feb, 2010 1 commit

ceph: reset osd connections after fault · 153a008b

Sage Weil authored Feb 15, 2010

A single osd connection fault (e.g. tcp disconnect) wasn't
reopening the connection, which causes all current and future
requests for that osd to hang.
Signed-off-by: Sage Weil <sage@newdream.net>

153a008b

14 Feb, 2010 1 commit

ceph: fix msgr to keep sent messages until acked · 6c5d1a49

Sage Weil authored Feb 13, 2010

The test was backwards from commit b3d1dbbd: keep the message if the
connection _isn't_ lossy.  This allows the client to continue when the
TCP connection drops for some reason (network glitch) but both ends
survive.
Signed-off-by: Sage Weil <sage@newdream.net>

6c5d1a49

11 Feb, 2010 14 commits

ceph: remove bogus invalidate_mapping_pages · 80310491