[PATCH] ocfs2/dlm: fixes * fix a hang which can occur during shutdown migration * do not allow nodes to join during recovery * when restarting lock mastery, do not ignore nodes which come up * more than one node could become recovery master, fix this * sleep to allow some time for heartbeat state to catch up to network * extra debug info for bad recovery state problems * make DLM_RECO_NODE_DATA_DONE a valid state for non-master recovery nodes * prune all locks from dead nodes on $RECOVERY lock resources * do NOT automatically add new nodes to mle nodemaps until they have properly joined the domain * make sure dlm_pick_recovery_master only exits when all nodes have synced * properly handle dlmunlock errors in dlm_pick_recovery_master * do not propagate network errors in dlm_send_begin_reco_message * dead nodes were not being put in the recovery map sometimes, fix this * dlmunlock was failing to clear the unlock actions on DLM_DENIED Signed-off-by: Kurt Hackel <kurt.hackel@oracle.com> Signed-off-by: Mark Fasheh <mark.fasheh@oracle.com>

commit: e2faea4ce340f199c1957986c4c3dc2de76f5746 [log] [tgz]
author: Kurt Hackel <kurt.hackel@oracle.com> Thu Jan 12 14:24:55 2006 -0800
committer: Mark Fasheh <mark.fasheh@oracle.com> Fri Feb 03 13:47:20 2006 -0800
tree: 2336b06cf270b3cff2ff39ba75fc67639dc63df9
parent: 0d419a6a95ee158675aa184c6c3e476b22d02145 [diff] [blame]
diff --git a/fs/ocfs2/dlm/dlmunlock.c b/fs/ocfs2/dlm/dlmunlock.c
index cec2ce1..c95f08d 100644
--- a/fs/ocfs2/dlm/dlmunlock.c
+++ b/fs/ocfs2/dlm/dlmunlock.c

@@ -188,6 +188,19 @@
 			actions &= ~(DLM_UNLOCK_REMOVE_LOCK|
 				     DLM_UNLOCK_REGRANT_LOCK|
 				     DLM_UNLOCK_CLEAR_CONVERT_TYPE);
+		} else if (status == DLM_RECOVERING || 
+			   status == DLM_MIGRATING || 
+			   status == DLM_FORWARD) {
+			/* must clear the actions because this unlock
+			 * is about to be retried.  cannot free or do
+			 * any list manipulation. */
+			mlog(0, "%s:%.*s: clearing actions, %s\n",
+			     dlm->name, res->lockname.len,
+			     res->lockname.name,
+			     status==DLM_RECOVERING?"recovering":
+			     (status==DLM_MIGRATING?"migrating":
+			      "forward"));
+			actions = 0;
 		}
 		if (flags & LKM_CANCEL)
 			lock->cancel_pending = 0;
commit	e2faea4ce340f199c1957986c4c3dc2de76f5746	[log] [tgz]
author	Kurt Hackel <kurt.hackel@oracle.com>	Thu Jan 12 14:24:55 2006 -0800
committer	Mark Fasheh <mark.fasheh@oracle.com>	Fri Feb 03 13:47:20 2006 -0800
tree	2336b06cf270b3cff2ff39ba75fc67639dc63df9
parent	0d419a6a95ee158675aa184c6c3e476b22d02145 [diff] [blame]