git.apps.os.sepia.ceph.com Git - ceph.git/log

]> git.apps.os.sepia.ceph.com Git - ceph.git/log

projects / ceph.git / log

summary | shortlog | log | commit | commitdiff | tree
first ⋅ prev ⋅ next

commit | commitdiff | tree

Samuel Just [Wed, 28 Nov 2012 23:10:43 +0000 (15:10 -0800)]

PG: scrubber.end should be exactly a boundary

Let scrubber.end be (foo, HEAD, 10) where the oid is foo , HEAD is the
snap, and 10 is the hash and scrubber.begin similarly be (bar, 5, 1).

After choosing to scan [(bar, 5, 1), (foo, HEAD, 10)), we block writes
on that interval.

1) A write might then come in for foo (which isn't blocked) which
creates a new snap (foo, 400, 10) which happens to fall in the interval.
This will result in a crash in _scrub() when it attempts to compare
clones since it will get (foo, 400, 10) but not the head object
(foo, HEAD, 10).

2) Alternately, the write from 1) has already happened.  When we scan
the log, we find 34'10 and 34'11 are the clone operation creating
(foo, 400, 10) and the modify on (foo, HEAD, 10) respectively.  Both
primary and replica will wait for last_update_applied to be 34'10
before scanning, but last_update_applied will in fact skip to 34'11
since 34'10 and 34'11 happened in the same transaction.  This can
result in IO hanging on the scrubber interval.

Instead, we ensure that scrubber.end is exactly a hash boundary
(min hobject_t a with the specified hash).  No such object can
exist since we don't create objects with empty oids, so no writes
can occur on that object.

Signed-off-by: Samuel Just <sam.just@inktank.com>

commit | commitdiff | tree

Samuel Just [Thu, 15 Nov 2012 21:35:47 +0000 (13:35 -0800)]

ReplicatedPG: remove from snap_collections even without objects to trim

Also, make sure to write_info after updating snap_collections.

Signed-off-by: Samuel Just <sam.just@inktank.com>

commit | commitdiff | tree

Samuel Just [Thu, 29 Nov 2012 19:28:25 +0000 (11:28 -0800)]

OSD: get_or_create_pg return null if pool is gone

Signed-off-by: Samuel Just <sam.just@inktank.com>

commit | commitdiff | tree

Samuel Just [Wed, 28 Nov 2012 00:00:03 +0000 (16:00 -0800)]

OSD: history.last_epoch_started should start at 0

history.last_epoch_started marks a lower bound on the last epoch at
which the pg went active. As with info.last_epoch_started, it should be
0 prior to the first activation.

Signed-off-by: Samuel Just <sam.just@inktank.com>

commit | commitdiff | tree

Samuel Just [Wed, 21 Nov 2012 21:59:22 +0000 (13:59 -0800)]

PG: maintain osd local last_epoch_started for find_best_info

In order to proceed with peering, we need an osd with a log including
the last commit sent to a client.  This translates to the oldest
last_update from the infos of the most recent acting set to go active.
history.last_epoch_started gives us a lower bound on the last time the
entire acting set persisted authoratative logs/infos.  However, it
doesn't indicate anything about the info/log on the osd which sent it.
Thus, we will maintain an osd local info.last_epoch_started to determine
which osds were actually active (and thus have the required log
entries).  The max info.last_epoch_started in the prior set gives us an
upper bound on the last interval during which writes occurred.  The min
last_update among the infos with that last_epoch_started must therefore
be an upper bound on the oldest operation which clients consider
committed.  Any osd with an info.last_updated past that version must be
sufficient.

The observed bug was there was an empty pg info with a
last_epoch_started at the most recent interval which pushed
min_last_update_acceptable to eversion_t().  There were two down osds,
but peering proceeded since the backfill peer did survive.  However,
its info was later disregarded due to incomplete.  An empty osd was
then chosen as the best_info since it's last_update was equal to
min_last_update_acceptable.  This caused the contents of the pg to be
lost.

Signed-off-by: Samuel Just <sam.just@inktank.com>

commit | commitdiff | tree

Samuel Just [Thu, 29 Nov 2012 21:51:41 +0000 (13:51 -0800)]

hobject_t: make max private

Signed-off-by: Samuel Just <sam.just@inktank.com>

commit | commitdiff | tree

Sage Weil [Mon, 19 Nov 2012 05:20:36 +0000 (21:20 -0800)]

Merge remote-tracking branch 'gh/wip-mon-parsing' into next

commit | commitdiff | tree

Sage Weil [Sun, 18 Nov 2012 22:37:22 +0000 (14:37 -0800)]

Merge branch 'wip-mon-leaks-fix' into next

commit | commitdiff | tree

Sage Weil [Sun, 18 Nov 2012 16:34:35 +0000 (08:34 -0800)]

mon: shutdown async signal handler sooner

Before the mon, and lockdep, in particular.

#0  __pthread_mutex_lock (mutex=0x30) at pthread_mutex_lock.c:50
#1  0x0000000000816092 in ceph::log::Log::submit_entry (this=0x0, e=0x2f4a270) at log/Log.cc:138
#2  0x00000000007ee0f8 in handle_fatal_signal (signum=11) at global/signal_handler.cc:100
#3  <signal handler called>
#4  0x00000000008e1300 in lockdep_will_lock (name=0x959aa7 "SignalHandler::lock", id=17) at common/lockdep.cc:163
#5  0x00000000008867fc in Mutex::_will_lock (this=0x2f20428) at ./common/Mutex.h:56
#6  0x0000000000886605 in Mutex::Lock (this=0x2f20428, no_lockdep=false) at common/Mutex.cc:81
#7  0x00000000007eeb95 in SignalHandler::entry (this=0x2f20300) at global/signal_handler.cc:198
#8  0x00000000008b0bd1 in Thread::_entry_func (arg=0x2f20300) at common/Thread.cc:43
#9  0x00007f36fefd6b50 in start_thread (arg=<optimized out>) at pthread_create.c:304
#10 0x00007f36fd80b6dd in clone () at ../sysdeps/unix/sysv/linux/x86_64/clone.S:112
#11 0x0000000000000000 in ?? ()

#0  0x00007f36fefd7e75 in pthread_join (threadid=139874129766144, thread_return=0x0) at pthread_join.c:89
#1  0x00000000008b11ec in Thread::join (this=0x2f20300, prval=0x0) at common/Thread.cc:130
#2  0x00000000007eeae7 in SignalHandler::shutdown (this=0x2f20300) at global/signal_handler.cc:186
#3  0x00000000007ee9cf in SignalHandler::~SignalHandler (this=0x2f20300, __in_chrg=<optimized out>) at global/signal_handler.cc:175
#4  0x00000000007eea58 in SignalHandler::~SignalHandler (this=0x2f20300, __in_chrg=<optimized out>) at global/signal_handler.cc:176
#5  0x00000000007ee643 in shutdown_async_signal_handler () at global/signal_handler.cc:324
#6  0x00000000006de9d2 in main (argc=7, argv=0x7fffbfb8a1e8) at ceph_mon.cc:439

Signed-off-by: Sage Weil <sage@inktank.com>

commit | commitdiff | tree

Sage Weil [Sun, 18 Nov 2012 16:00:16 +0000 (08:00 -0800)]

mon/AuthMonitor: refactor assign_global_id

Move the failure logic into the caller, where we easier to do something
about it and return the right value to the caller.

Signed-off-by: Sage Weil <sage@inktank.com>

commit | commitdiff | tree

Sage Weil [Sun, 18 Nov 2012 15:52:49 +0000 (07:52 -0800)]

mon/AuthMonitor: reorder session->put()

Signed-off-by: Sage Weil <sage@inktank.com>

commit | commitdiff | tree

Sage Weil [Sun, 18 Nov 2012 04:57:50 +0000 (20:57 -0800)]

msg/Pipe: remove useless reader_joining

We set it but do not read it.

Signed-off-by: Sage Weil <sage@inktank.com>

commit | commitdiff | tree

Sage Weil [Sun, 18 Nov 2012 04:56:50 +0000 (20:56 -0800)]

msg/Pipe: join previous reader threads

We may stop and then restart the reader thread. Join previous threads
before we create new ones.

Signed-off-by: Sage Weil <sage@inktank.com>

commit | commitdiff | tree

Sage Weil [Sun, 18 Nov 2012 00:36:44 +0000 (16:36 -0800)]

msg/DispatchQueue: fix message leak from discard_queue()

We need to drop the Message ref() here; the msgr owns one ref
independent of those from the intrusive_ptr's in the queue itself.

Signed-off-by: Sage Weil <sage@inktank.com>

commit | commitdiff | tree

Sage Weil [Sun, 18 Nov 2012 00:01:13 +0000 (16:01 -0800)]

msg/SimpleMessenger: use put() on local_connection

This aids leak debugging; not much else.

Signed-off-by: Sage Weil <sage@inktank.com>

commit | commitdiff | tree

Sage Weil [Sat, 17 Nov 2012 22:21:07 +0000 (14:21 -0800)]

mon: clean up Subsription xlists

Signed-off-by: Sage Weil <sage@inktank.com>

commit | commitdiff | tree

Sage Weil [Sun, 18 Nov 2012 16:19:41 +0000 (08:19 -0800)]

mon: drop con->session reference in remove_session()

This captures all callers.

Signed-off-by: Sage Weil <sage@inktank.com>

commit | commitdiff | tree

Sage Weil [Sat, 17 Nov 2012 16:52:42 +0000 (08:52 -0800)]

mon: sessions get cleaned up before dtor

Signed-off-by: Sage Weil <sage@inktank.com>

commit | commitdiff | tree

Sage Weil [Fri, 16 Nov 2012 20:21:14 +0000 (12:21 -0800)]

msg/Pipe: don't leak session_security

Make sure we free old instances of sesseion_security before we reset the
pointer.

Signed-off-by: Sage Weil <sage@inktank.com>

commit | commitdiff | tree

Joao Eduardo Luis [Thu, 15 Nov 2012 02:16:53 +0000 (02:16 +0000)]

mon: Monitor: make MSG_MON_PAXOS case a bit more consistent

Signed-off-by: Joao Eduardo Luis <joao.luis@inktank.com>

commit | commitdiff | tree

Joao Eduardo Luis [Thu, 15 Nov 2012 02:16:17 +0000 (02:16 +0000)]

mon: Paxos{,Service}: finish contexts and put messages on shutdown

Signed-off-by: Joao Eduardo Luis <joao.luis@inktank.com>

commit | commitdiff | tree

Joao Eduardo Luis [Wed, 14 Nov 2012 15:54:17 +0000 (15:54 +0000)]

mon: Monitor: finish contexts on shutdown

Signed-off-by: Joao Eduardo Luis <joao.luis@inktank.com>

commit | commitdiff | tree

Joao Eduardo Luis [Tue, 13 Nov 2012 16:57:34 +0000 (16:57 +0000)]

mon: Monitor: drop election messages if entity doesn't have enough caps

Signed-off-by: Joao Eduardo Luis <joao.luis@inktank.com>

commit | commitdiff | tree

Sage Weil [Tue, 13 Nov 2012 01:43:51 +0000 (17:43 -0800)]

mon: remove all sessions on shutdown

Signed-off-by: Sage Weil <sage@inktank.com>

commit | commitdiff | tree

Joao Eduardo Luis [Mon, 12 Nov 2012 23:14:55 +0000 (23:14 +0000)]

ceph_mon: cleanup on shutdown

Properly cleanup the throttlers, 'g_ceph_context' and the
async_singnal_handler.

Signed-off-by: Joao Eduardo Luis <joao.luis@inktank.com>

commit | commitdiff | tree

Chen Baozi [Sun, 18 Nov 2012 06:34:21 +0000 (14:34 +0800)]

rgw: add -lresolv flags to Makefile.am

radosgw depends on libresolv since since the commit 951c6be. So we need to
add -lresolve flags, or it cannot link right library.

Signed-off-by: Chen Baozi <baozich@gmail.com>

commit | commitdiff | tree

Sage Weil [Sun, 4 Nov 2012 16:21:50 +0000 (08:21 -0800)]

mon/MonClient: use thread-safe RNG for picking monitors

Avoid using shared-state rand() when picking monitors. This way we don't
screw with library users like test_librbd_fsx that rely on srand() and
rand() being deterministic.

Signed-off-by: Sage Weil <sage@inktank.com>

commit | commitdiff | tree

Sage Weil [Sat, 17 Nov 2012 05:26:30 +0000 (21:26 -0800)]

Merge remote-tracking branch 'gh/wip-3431' into next

commit | commitdiff | tree

Josh Durgin [Sat, 17 Nov 2012 01:13:50 +0000 (17:13 -0800)]

Makefile.am: fix LDADD for test_objectcacher_stress

Signed-off-by: Josh Durgin <josh.durgin@inktank.com>

commit | commitdiff | tree

Sage Weil [Sat, 17 Nov 2012 01:36:34 +0000 (17:36 -0800)]

Merge branch 'wip-coverity' into next

Reviewed-by: Josh Durgin <josh.durgin@inktank.com>

commit | commitdiff | tree

Sage Weil [Sat, 17 Nov 2012 01:36:16 +0000 (17:36 -0800)]

client: fix lock leak in lazio_*() failure paths

CID 743400 (#1 of 1): Missing unlock (LOCK)
At (5): Returning without unlocking "this->client_lock._m".

CID 743399 (#1 of 1): Missing unlock (LOCK)
At (5): Returning without unlocking "this->client_lock._m".

Signed-off-by: Sage Weil <sage@inktank.com>

commit | commitdiff | tree

Josh Durgin [Sat, 17 Nov 2012 00:43:00 +0000 (16:43 -0800)]

Merge branch 'wip-oc-hang' into next

Reviewed-by: Sage Weil <sage.weil@inktank.com>

commit | commitdiff | tree

Sage Weil [Sat, 17 Nov 2012 00:19:00 +0000 (16:19 -0800)]

upstart: set high open file limits

The default 1024 limit is easily hit on larger clusters.

Signed-off-by: Sage Weil <sage@inktank.com>

commit | commitdiff | tree

Sage Weil [Sat, 17 Nov 2012 00:10:30 +0000 (16:10 -0800)]

msg/Accepter: only close socket if >= 0

It is possible for rebind() to fail, in which case the OSD will go through
it's shutdown procedure and call stop(). This is simpler than trying to
avoid calling stop() when rebind() fails.

Fixes: #3504
Signed-off-by: Sage Weil <sage@inktank.com>

commit | commitdiff | tree

Sage Weil [Sat, 17 Nov 2012 00:04:13 +0000 (16:04 -0800)]

osd: default journal size to 5GB

Signed-off-by: Sage Weil <sage@inktank.com>

commit | commitdiff | tree

Josh Durgin [Fri, 16 Nov 2012 23:27:52 +0000 (15:27 -0800)]

librbd: take cache lock when discarding data from cache

Signed-off-by: Josh Durgin <josh.durgin@inktank.com>

commit | commitdiff | tree

Josh Durgin [Fri, 16 Nov 2012 00:20:33 +0000 (16:20 -0800)]

ObjectCacher: fix off-by-one error in split

This error left a completion that should have been attached
to the right BufferHead on the left BufferHead, which would
result in the completion never being called unless the buffers
were merged before it's original read completed. This would cause
a hang in any higher level waiting for a read to complete.

The existing loop went backwards (using a forward iterator),
but stopped when the iterator reached the beginning of the map,
or when a waiter belonged to the left BufferHead.

If the first list of waiters should have been moved to the right
BufferHead, it was skipped because at that point the iterator
was at the beginning of the map, which was the main condition
of the loop.

Restructure the waiters-moving loop to go forward in the map instead,
so it's harder to make an off-by-one error.

Possibly-fixes: #3286
Signed-off-by: Josh Durgin <josh.durgin@inktank.com>

commit | commitdiff | tree

Josh Durgin [Thu, 15 Nov 2012 18:41:32 +0000 (10:41 -0800)]

ObjectCacher: begin at the right place when iterating over BufferHeads

If the desired offset overlaps a BH, data.lower_bound() will return
the element after it, since it's indexed by the start of a range.

The confusingly similarly named data_lower_bound() method fixes this,
and returns the correct starting element.

Signed-off-by: Josh Durgin <josh.durgin@inktank.com>

commit | commitdiff | tree

Josh Durgin [Fri, 16 Nov 2012 01:32:08 +0000 (17:32 -0800)]

ObjectCacher: add debug function to check BufferHead consistency

This isn't called because it's potentially expensive, but calling it
in various places can help future debugging.

Signed-off-by: Josh Durgin <josh.durgin@inktank.com>

commit | commitdiff | tree

Josh Durgin [Fri, 16 Nov 2012 00:53:37 +0000 (16:53 -0800)]

ObjectCacher: more debugging for read completions

Signed-off-by: Josh Durgin <josh.durgin@inktank.com>

commit | commitdiff | tree

Josh Durgin [Thu, 15 Nov 2012 18:35:57 +0000 (10:35 -0800)]

ObjectCacher: assert lock is held everywhere

Signed-off-by: Josh Durgin <josh.durgin@inktank.com>

commit | commitdiff | tree

Josh Durgin [Fri, 16 Nov 2012 19:56:46 +0000 (11:56 -0800)]

ObjectCacher: debug read waiters

Now we can tell which ones will be called.

Signed-off-by: Josh Durgin <josh.durgin@inktank.com>

commit | commitdiff | tree

Josh Durgin [Fri, 16 Nov 2012 23:16:24 +0000 (15:16 -0800)]

ObjectCacher: don't needlessly increment iterator

This iterator is now reset on each run through the loop,
so there's no point in incrementing it here.

Signed-off-by: Josh Durgin <josh.durgin@inktank.com>

commit | commitdiff | tree

Josh Durgin [Fri, 16 Nov 2012 20:26:16 +0000 (12:26 -0800)]

ObjectCacher: retry reads when they are incomplete

Skipping these callbacks when there's a racing write or
a gap in the results causes the original reads they represent
to never be completed. If the read falls within the range
of a BufferHead, retry all waiters no matter what.

Signed-off-by: Josh Durgin <josh.durgin@inktank.com>

commit | commitdiff | tree

Sage Weil [Fri, 16 Nov 2012 22:19:25 +0000 (14:19 -0800)]

common/ceph_argparse: fix malloc failure check

CID 743418 (#1 of 1): Dereference before null check (REVERSE_INULL)
Null-checking "argv" suggests that it may be null, but it has already been dereferenced on all paths leading to the check.

Signed-off-by: Sage Weil <sage@inktank.com>

commit | commitdiff | tree

Sage Weil [Fri, 16 Nov 2012 22:18:21 +0000 (14:18 -0800)]

mon/MonClient: initialize ptr in ctor

CID 743433 (#1 of 1): Uninitialized pointer field (UNINIT_CTOR)
At (2): Non-static class member "authorize_handler_registry" is not initialized in this constructor nor in any functions that it calls.

Signed-off-by: Sage Weil <sage@inktank.com>

commit | commitdiff | tree

Sage Weil [Fri, 16 Nov 2012 22:15:23 +0000 (14:15 -0800)]

os/FileStore: fix fd leak in _rmattr

CID 743405 (#2 of 2): Resource leak (RESOURCE_LEAK)
At (16): Handle variable "fd" going out of scope leaks the handle.

Signed-off-by: Sage Weil <sage@inktank.com>

commit | commitdiff | tree

Sage Weil [Fri, 16 Nov 2012 22:14:17 +0000 (14:14 -0800)]

os/FileStore: fix fd leaks in _setattrs

CID 743406 (#3 of 3): Resource leak (RESOURCE_LEAK)
At (26): Handle variable "fd" going out of scope leaks the handle.

Signed-off-by: Sage Weil <sage@inktank.com>

commit | commitdiff | tree

Sage Weil [Fri, 16 Nov 2012 22:11:05 +0000 (14:11 -0800)]

osdc/ObjectCacher: faux use-after-free

CID 743435 (#1 of 1): Use after free (USE_AFTER_FREE)
At (68): Passing freed pointer "rd" as an argument to function "std::basic_ostream<char, std::char_traits<char> >::operator <<(void const *)".

Signed-off-by: Sage Weil <sage@inktank.com>

commit | commitdiff | tree

Josh Durgin [Tue, 13 Nov 2012 18:28:32 +0000 (10:28 -0800)]

test: add ObjectCacher stress test that does not use a cluster

Use a fake writeback handler and respond to all requests with -ENOENT.
This tests that all operations will complete, and the cache doesn't
lose waiters or callbacks.

Signed-off-by: Josh Durgin <josh.durgin@inktank.com>

commit | commitdiff | tree

Josh Durgin [Tue, 13 Nov 2012 18:01:30 +0000 (10:01 -0800)]

ObjectCacher: more debugging for BufferHeads

This is useful for checking for lost waiters.

Signed-off-by: Josh Durgin <josh.durgin@inktank.com>

commit | commitdiff | tree

Gary Lowell [Fri, 16 Nov 2012 08:46:41 +0000 (00:46 -0800)]

build: update for boost_thread library.

There is a difference in naming conventions between debian and
rpm based distributions for this library. In configure.ac we
check first for boost_thread-mt, then if it's not found check
for boost_thread. A side effect of the AC_CEHCK_LIB macro is
to add the library to the $LIBS, so the explicit -llibboost_thread
in the Makefile has been removed.
(cherry picked from commit f0c7bb363000037bbf7d58ac6e2d39d0f10200fe)

commit | commitdiff | tree

Joao Eduardo Luis [Fri, 16 Nov 2012 15:30:53 +0000 (15:30 +0000)]

mon: OSDMonitor: clarify some command replies

Signed-off-by: Joao Eduardo Luis <joao.luis@inktank.com>

commit | commitdiff | tree

Joao Eduardo Luis [Fri, 16 Nov 2012 15:30:24 +0000 (15:30 +0000)]

mon: OSDMonitor: fix spacing when outputting items on command reply

Signed-off-by: Joao Eduardo Luis <joao.luis@inktank.com>

commit | commitdiff | tree

Sage Weil [Fri, 16 Nov 2012 00:50:39 +0000 (16:50 -0800)]

os/FileStore: only try BTRFS_IOC_SUBVOL_CREATE on btrfs

Only try to create a btrfs subvolume if the fs is btrfs. Otherwise, just
create a directory. Then we can error out on *any* ioctl error, and not
rely on the ioctl error code to determine if we failed because we are on
a non-btrfs or a real error.

Fixes: #3052
Signed-off-by: Sage Weil <sage@inktank.com>
Reviewed-by: Dan Mick <dan.mick@inktank.com>

commit | commitdiff | tree

Sage Weil [Fri, 16 Nov 2012 00:35:53 +0000 (16:35 -0800)]

mon: clean up 'ceph osd ...' list output

No more 'osd.0 is already inosd.1 is already in' crap.

Signed-off-by: Sage Weil <sage@inktank.com>

commit | commitdiff | tree

Sage Weil [Fri, 16 Nov 2012 00:24:38 +0000 (16:24 -0800)]

mon: correctly identify crush names

get_item_id() returns 0 if the name already exists; that's not what we
want here. Verify the name exists before checking its id.

Signed-off-by: Sage Weil <sage@inktank.com>

commit | commitdiff | tree

Sage Weil [Fri, 16 Nov 2012 00:23:48 +0000 (16:23 -0800)]

mon: use parse_osd_id() throughout

Signed-off-by: Sage Weil <sage@inktank.com>

commit | commitdiff | tree

Samuel Just [Fri, 16 Nov 2012 00:01:18 +0000 (16:01 -0800)]

PrioritizedQueue: remove internal lock, not used

Signed-off-by: Samuel Just <sam.just@inktank.com>
Reviewed-by: Sage Weil <sage@inktank.com>

commit | commitdiff | tree

Samuel Just [Thu, 15 Nov 2012 23:56:46 +0000 (15:56 -0800)]

DispatchQueue: lock DispatchQueue when for get_queue_len()

Signed-off-by: Samuel Just <sam.just@inktank.com>
Reviewed-by: Sage Weil <sage@inktank.com>

commit | commitdiff | tree

Alex Elder [Thu, 15 Nov 2012 23:51:34 +0000 (17:51 -0600)]

run_xfstests.sh: activate more tests that now work

I've gone through the set of xfstests that were previously found to
not work.  Some of those now do work, and with the addition of an
option to pass to "mkfs.xfs" a large number of other tests now
produce expected output as well.

This patch updates the default list of tests to run to reflect
the result of this exercise.  The following 50 additional tests
are now run by default:

    029 074 078 084-087 100 105 117 121 124 126 129-134
    164 165 167 174 181 184 186 187 192 214-216 227 236
    237 241 243 245-249 257-259 261 277 278 280 285 286

Test 127 completed without error, but it took from 1-3 hours so I
kept that out of the list.

Signed-off-by: Alex Elder <elder@inktank.com>

commit | commitdiff | tree

Sage Weil [Thu, 15 Nov 2012 18:06:07 +0000 (10:06 -0800)]

msg/Pipe: fix leak of Authorizer

Reported-by: Joao Luis <joao.luis@inktank.com>
Signed-off-by: Sage Weil <sage@inktank.com>

commit | commitdiff | tree

Sage Weil [Thu, 15 Nov 2012 17:48:25 +0000 (09:48 -0800)]

Merge remote-tracking branch 'gh/wip-3477' into next

Reviewed-by: Greg Farnum <greg@inktank.com>

commit | commitdiff | tree

Samuel Just [Thu, 15 Nov 2012 00:30:51 +0000 (16:30 -0800)]

msg/DispatchQueue: release throttle on messages when dropping an id

Signed-off-by: Samuel Just <sam.just@inktank.com>

commit | commitdiff | tree

Samuel Just [Thu, 15 Nov 2012 00:30:05 +0000 (16:30 -0800)]

PrioritizedQueue: allow remove_by_class to return removed items

Signed-off-by: Samuel Just <sam.just@inktank.com>

commit | commitdiff | tree

Sage Weil [Thu, 15 Nov 2012 00:59:06 +0000 (16:59 -0800)]

librbd: use delete[] properly

==4986== Mismatched free() / delete / delete []
==4986==    at 0x4C2658C: operator delete(void*) (in /usr/lib/valgrind/vgpreload_memcheck-amd64-linux.so)
==4986==    by 0x4ED8EA9: librbd::ImageCtx::~ImageCtx() (ImageCtx.cc:100)
==4986==    by 0x4EF3827: librbd::close_image(librbd::ImageCtx*) (internal.cc:1869)
==4986==    by 0x4EE8FB8: librbd::clone(librados::IoCtx&, char const*, char const*, librados::IoCtx&, char const*, unsigned long, int*, unsigned long, int) (internal.cc:900)
==4986==    by 0x4EC363C: rbd_clone2 (librbd.cc:553)
==4986==    by 0x404C85: do_clone (fsx.c:836)
==4986==    by 0x405639: test (fsx.c:1048)
==4986==    by 0x406369: main (fsx.c:1523)
==4986==  Address 0xd498b30 is 0 bytes inside a block of size 37 alloc'd
==4986==    at 0x4C26CF7: operator new[](unsigned long) (in /usr/lib/valgrind/vgpreload_memcheck-amd64-linux.so)
==4986==    by 0x4ED9B4D: librbd::ImageCtx::init_layout() (ImageCtx.cc:164)
==4986==    by 0x4ED9845: librbd::ImageCtx::init() (ImageCtx.cc:142)
==4986==    by 0x4EF3449: librbd::open_image(librbd::ImageCtx*, bool) (internal.cc:1828)
==4986==    by 0x4EE89E0: librbd::clone(librados::IoCtx&, char const*, char const*, librados::IoCtx&, char const*, unsigned long, int*, unsigned long, int) (internal.cc:871)
==4986==    by 0x4EC363C: rbd_clone2 (librbd.cc:553)
==4986==    by 0x404C85: do_clone (fsx.c:836)
==4986==    by 0x405639: test (fsx.c:1048)
==4986==    by 0x406369: main (fsx.c:1523)

Signed-off-by: Sage Weil <sage@inktank.com>

commit | commitdiff | tree

Sage Weil [Thu, 15 Nov 2012 00:54:17 +0000 (16:54 -0800)]

objecter: fix leak of out_handlers

The error paths don't use the handlers. Make sure they get cleaned up.

Fixes: #3446
Signed-off-by: Sage Weil <sage@inktank.com>

commit | commitdiff | tree

Sage Weil [Thu, 15 Nov 2012 01:00:57 +0000 (17:00 -0800)]

mon: calculate failed_since relative to message receive time

Instead of looking at the current time we process the message, look at the
receive time. This gives us a more real failure time given that messages
may be requeued.

It doesn't solve the problem when messages are forwarded between monitors
due to an election, but that's ok; this is still a net improvement.

Signed-off-by: Sage Weil <sage@inktank.com>

commit | commitdiff | tree

Yehuda Sadeh [Thu, 15 Nov 2012 00:42:11 +0000 (16:42 -0800)]

rgw: update post policy parser

json parser semantics changed a little bit, so
needed to update the policy parser.

Signed-off-by: Yehuda Sadeh <yehuda@inktank.com>

commit | commitdiff | tree

Sage Weil [Thu, 15 Nov 2012 00:26:58 +0000 (16:26 -0800)]

mon: set default port when binding to random local ip

Fixes #3135
Signed-off-by: Sage Weil <sage@inktank.com>

commit | commitdiff | tree

Sage Weil [Thu, 15 Nov 2012 00:22:27 +0000 (16:22 -0800)]

Merge remote-tracking branch 'gh/wip-asok' into next

Reviewed-by: Josh Durgin <josh.durgin@inktank.com>

commit | commitdiff | tree

Yehuda Sadeh [Wed, 14 Nov 2012 19:30:34 +0000 (11:30 -0800)]

rgw: relax date format check

Don't try to parse beyond the GMT or UTC. Some clients use
special date formatting. If we end up misparsing the date
it'll fail in the authorization, so don't need to be too
restrictive.

Signed-off-by: Yehuda Sadeh <yehuda@inktank.com>

commit | commitdiff | tree

Sage Weil [Wed, 14 Nov 2012 02:18:24 +0000 (18:18 -0800)]

client: register admin socket commands without lock held

Avoid a lock cycle.

existing dependency Client::client_lock (11) -> AdminSocket::m_lock (16) at:
ceph version 0.54-578-g7926ef5 (7926ef53935313501d4a7fe0e587f3e3b00b313c)
1: (Mutex::Lock(bool)+0x41) [0x831337]
2: (AdminSocket::register_command(std::string, AdminSocketHook*, std::string)+0x40) [0x873a32]
3: (Client::init()+0x454) [0x6f4c24]
4: (main()+0x637) [0x6ea399]
5: (__libc_start_main()+0xed) [0x7fd97bbca76d]
6: ./ceph-fuse() [0x6e9c59]

-4> 2012-11-13 18:14:48.619714 7fd97b1a3700 0 new dependency AdminSocket::m_lock (16) -> Client::client_lock (11) creates a cycle at
ceph version 0.54-578-g7926ef5 (7926ef53935313501d4a7fe0e587f3e3b00b313c)
1: (Mutex::Lock(bool)+0x41) [0x831337]
2: (Objecter::RequestStateHook::call(std::string, std::string, ceph::buffer::list&)+0x7a) [0x90627e]
3: (AdminSocket::do_accept()+0xb1b) [0x87318f]
4: (AdminSocket::entry()+0x2fa) [0x8725fe]
5: (Thread::_entry_func(void*)+0x23) [0x86b335]
6: (()+0x7e9a) [0x7fd97d279e9a]
7: (clone()+0x6d) [0x7fd97bc9ccbd]

Signed-off-by: Sage Weil <sage@inktank.com>

commit | commitdiff | tree

Sage Weil [Wed, 14 Nov 2012 02:17:55 +0000 (18:17 -0800)]

objecter: separate locked and unlocked init/shutdown

We don't want to hold the lock while we register the admin socket commands
or else we create a lock cycle when we try to process them later.

Signed-off-by: Sage Weil <sage@inktank.com>

commit | commitdiff | tree

Gary Lowell [Wed, 14 Nov 2012 01:29:47 +0000 (17:29 -0800)]

Merge branch 'next'

Conflicts:
configure.ac
src/rgw/rgw_common.cc

commit | commitdiff | tree

Samuel Just [Wed, 14 Nov 2012 00:45:49 +0000 (16:45 -0800)]

osd/: add config helper for min_size and update build_simple*

min_size should never be set to 0 on a pool. config.h
now has a helper to determine the correct default value.

Signed-off-by: Samuel Just <sam.just@inktank.com>
Reviewed-by: Sage Weil <sage@inktank.com>

commit | commitdiff | tree

Sage Weil [Wed, 14 Nov 2012 01:11:34 +0000 (17:11 -0800)]

doc/release-notes: fix heading

Signed-off-by: Sage Weil <sage@inktank.com>

commit | commitdiff | tree

Sage Weil [Wed, 14 Nov 2012 00:24:23 +0000 (16:24 -0800)]

doc: release-notes for v0.54

Signed-off-by: Sage Weil <sage@inktank.com>

commit | commitdiff | tree

Sage Weil [Tue, 13 Nov 2012 22:34:53 +0000 (14:34 -0800)]

doc: update crush weight ramping process

Signed-off-by: Sage Weil <sage@inktank.com>

commit | commitdiff | tree

Yehuda Sadeh [Tue, 13 Nov 2012 23:42:52 +0000 (15:42 -0800)]

rgw: fix warning

Signed-off-by: Yehuda Sadeh <yehuda@inktank.com>

commit | commitdiff | tree

Sage Weil [Tue, 13 Nov 2012 23:39:42 +0000 (15:39 -0800)]

Merge branch 'wip-min-size'

Reviewed-by: Sam Just <sam.just@inktank.com>

commit | commitdiff | tree

Sage Weil [Tue, 13 Nov 2012 23:16:56 +0000 (15:16 -0800)]

osd: default pool min_size to 0 (which gives us size-size/2)

Signed-off-by: Sage Weil <sage@inktank.com>

commit | commitdiff | tree

Sage Weil [Tue, 13 Nov 2012 23:11:42 +0000 (15:11 -0800)]

mon: default min_size to size-size/2 if min_size default is 0

Signed-off-by: Sage Weil <sage@inktank.com>

commit | commitdiff | tree

Sage Weil [Tue, 13 Nov 2012 21:27:45 +0000 (13:27 -0800)]

osd: default min_size to size - size/2

size -> min_size:
5 -> 3
4 -> 2
3 -> 2
2 -> 1

Basically, default to tolerating minority down.

Signed-off-by: Sage Weil <sage@inktank.com>

commit | commitdiff | tree

Sage Weil [Tue, 13 Nov 2012 21:25:50 +0000 (13:25 -0800)]

mon: helpful warning in 'health detail' output about incomplete pgs

Signed-off-by: Sage Weil <sage@inktank.com>

commit | commitdiff | tree

Sage Weil [Tue, 13 Nov 2012 21:25:31 +0000 (13:25 -0800)]

osd: start_boot() after init()

The previous trigger for start_boot() was racy, depending on whether we
got our rotating keys quickly.

Signed-off-by: Sage Weil <sage@inktank.com>

commit | commitdiff | tree

Dan Mick [Tue, 13 Nov 2012 21:24:15 +0000 (13:24 -0800)]

vstart.sh: support -X by adding 'auth required = none' entries

Signed-off-by: Dan Mick <dan.mick@inktank.com>

commit | commitdiff | tree

Sage Weil [Tue, 13 Nov 2012 22:50:42 +0000 (14:50 -0800)]

Merge remote-tracking branch 'gh/wip-rgw-integration'

Conflicts:
src/common/config_opts.h

commit | commitdiff | tree

Gary Lowell [Tue, 13 Nov 2012 21:18:07 +0000 (13:18 -0800)]

v0.54

commit | commitdiff | tree

Yehuda Sadeh [Tue, 13 Nov 2012 21:06:22 +0000 (13:06 -0800)]

rgw: compile with -Woverloaded-virtual

This will trigger a warning if RGWRados api changes while
RGWCache doesn't.

Signed-off-by: Yehuda Sadeh <yehuda@inktank.com>

commit | commitdiff | tree

Yehuda Sadeh [Tue, 13 Nov 2012 20:09:05 +0000 (12:09 -0800)]

rgw: fix RGWCache api

RGWCache api diverted form RGWRados, crippling the cache.

Signed-off-by: Yehuda Sadeh <yehuda@inktank.com>

commit | commitdiff | tree

Yehuda Sadeh [Tue, 13 Nov 2012 20:09:05 +0000 (12:09 -0800)]

rgw: fix RGWCache api

RGWCache api diverted form RGWRados, crippling the cache.

Signed-off-by: Yehuda Sadeh <yehuda@inktank.com>

commit | commitdiff | tree

Sage Weil [Mon, 12 Nov 2012 15:06:42 +0000 (07:06 -0800)]

osd: remove dead rotating key code from init

Ancient, dead.

Signed-off-by: Sage Weil <sage@inktank.com>

commit | commitdiff | tree

Sage Weil [Mon, 12 Nov 2012 15:06:25 +0000 (07:06 -0800)]

osd: defer boot until we have rotating keys

Make sure we have our rotating keys before we start booting. This
ensures we can open connections with peers *before* we add ourselves to
the osdmap. This behaviors marks instances of #3292, although it is
not clear whether it is responsible for the actual crash.

Signed-off-by: Sage Weil <sage@inktank.com>
Reviewed-by: Sam Just <sam.just@inktank.com>

commit | commitdiff | tree

Samuel Just [Tue, 13 Nov 2012 18:56:22 +0000 (10:56 -0800)]

Merge branches 'wip_persist_missing' and 'wip_recovery_qos'

Reviewed-by: Sage Weil <sage@inktank.com>

commit | commitdiff | tree

Samuel Just [Mon, 5 Nov 2012 23:40:43 +0000 (15:40 -0800)]

PG: persist divergent_priors in ondisklog

Consider the following logs:

a) 10'10(5'7) foo
   12'11(4'3) bar

b) 10'10(5'7) foo
   13'11(4'4) baz

When the osd with a merges primary log b, bar is deleted and
added to the missing set with need=4'3 and have=0'0.  If
the osd then dies after deleting bar, but before recovering
bar, PG::read_state() on start up will fail to re-add bar
to the missing set, and bar will be incorrect on that osd.

Now, (4'3, bar) will be added to the divergent_priors mapping
to be scanned during read_state along with the log.

Signed-off-by: Samuel Just <sam.just@inktank.com>

commit | commitdiff | tree

Samuel Just [Mon, 5 Nov 2012 19:33:13 +0000 (11:33 -0800)]

PG::merge_old_entry: fix case for divergent prior_version

Previously, we asserted that a log entry with a divergent
prior_version must be a clone.  Consider the following
case:

6'11(6'2)  m foo
7'12(6'3) m bar
7'13(7'12) m bar

If this is merged with:

6'11(6'2)  m foo
8'12(6'4) m baz

we will hit the assert.

Merging a divergent entry with prior_version after current
tail, but not in the log implies that prior_version was a
divergent entry which we have already merged.  The missing
set and filestore collection must therefore have already
been adjusted.

Signed-off-by: Samuel Just <sam.just@inktank.com>

commit | commitdiff | tree

Sage Weil [Sat, 10 Nov 2012 11:57:23 +0000 (03:57 -0800)]

PrioritizedQueue: use iterator to streamlink SubQueue::remove_by_class()

Signed-off-by: Sage Weil <sage@inktank.com>

commit | commitdiff | tree

Sage Weil [Sat, 10 Nov 2012 11:51:25 +0000 (03:51 -0800)]

PrioritizedQueue: avoid double-lookup on create_queue()

Signed-off-by: Sage Weil <sage@inktank.com>

commit | commitdiff | tree

Samuel Just [Sat, 29 Sep 2012 00:27:39 +0000 (17:27 -0700)]

osd/: de-prioritize recovery ops relative to client ops

Signed-off-by: Samuel Just <sam.just@inktank.com>

Unnamed repository; edit this file 'description' to name the repository.