Optimization fixes

Two primary fixes.  First, the save/restore mechanism for FP callee saves
was broken if there were any holes in the save mask (the Arm ld/store
multiple instructions for floating point use a start + count mechanism,
rather than the bit-mask mechanism used for core registers).

The second fix corrects a problem introduced by the recent enhancements
to loading floating point literals.  The load->copy optimization mechanism
for literal loads used the value of the loaded literal to identify
redundant loads.  However, it used only the first 32 bits of the
literal - which worked fine previously because 64-bit literal loads
were treated as a pair of 32-bit loads.  The fix was to use the
label of the literal rather than the value in the aliasInfo - which
works for all sizes.

Change-Id: Ic4779adf73b2c7d80059a988b0ecdef39921a81f
diff --git a/src/compiler/CompilerIR.h b/src/compiler/CompilerIR.h
index 934139b..2d4f83e 100644
--- a/src/compiler/CompilerIR.h
+++ b/src/compiler/CompilerIR.h
@@ -266,7 +266,7 @@
     int numIns;
     int numOuts;
     int numRegs;            // Unlike struct Method, does not include ins
-    int numSpills;          // NOTE: includes numFPSpills
+    int numCoreSpills;
     int numFPSpills;
     int numPadding;         // # of 4-byte padding cells
     int regsOffset;         // sp-relative offset to beginning of Dalvik regs