Commits · 47ff14b091ee66a4c5f159e7fb8714ca9a66d2f9 · Roger Ferrer / llvm-epi-0.8

Jan 21, 2011

Convert -enable-sched-cycles and -enable-sched-hazard to -disable · 47ff14b0

Andrew Trick authored Jan 21, 2011

flags. They are still not enable in this revision.

Added TargetInstrInfo::isZeroCost() to fix a fundamental problem with
the scheduler's model of operand latency in the selection DAG.

Generalized unit tests to work with sched-cycles.

llvm-svn: 123969

47ff14b0

fix PR9013, an infinite loop in instcombine. · b5e15d19
Chris Lattner authored Jan 21, 2011
```
llvm-svn: 123968
```
b5e15d19
update obsolete comment. · f4ca47bd
Chris Lattner authored Jan 21, 2011
```
llvm-svn: 123965
```
f4ca47bd

Don't try to pull vector bitcasts that change the number of elements through · 6a083cf8

Nick Lewycky authored Jan 21, 2011

a select. A vector select is pairwise on each element so we'd need a new
condition with the right number of elements to select on. Fixes PR8994.

llvm-svn: 123963

6a083cf8

Object: Fix type punned pointer issues by making DataRefImpl a union and using intptr_t. · 0324b672
Michael J. Spencer authored Jan 21, 2011
```
llvm-svn: 123962
```
0324b672

Add a constant folding of casts from zero to zero. Fixes PR9011! · 39b12c05

Nick Lewycky authored Jan 21, 2011

While here, I'd like to complain about how vector is not an aggregate type
according to llvm::Type::isAggregateType(), but they're listed under aggregate
types in the LangRef and zero vectors are stored as ConstantAggregateZero.

llvm-svn: 123956

39b12c05

Don't be overly aggressive with CSE of "ldr constantpool". If it's a pc-relative · 028ccbfc

Evan Cheng authored Jan 20, 2011

value, the "add pc" must be CSE'ed at the same time. We could follow the same
approach as T2 by adding pseudo instructions that combine the ldr + "add pc".
But the better approach is to use movw + movt (which I will enable soon), so
I'll leave this as a TODO.

llvm-svn: 123949

028ccbfc

Jan 20, 2011

Implement requiredTransitive · f07426b4

Tobias Grosser authored Jan 20, 2011

The PassManager did not implement the transitivity of requiredTransitive. This
was unnoticed since 2006.

llvm-svn: 123942

f07426b4

Fix the encoding and parsing of clrex instruction · e965f06f
Bruno Cardoso Lopes authored Jan 20, 2011
```
llvm-svn: 123936
```
e965f06f
Change instruction names for consistency · ef8cab90
Bruno Cardoso Lopes authored Jan 20, 2011
```
llvm-svn: 123930
```
ef8cab90
Add cdp/cdp2 instructions for thumb/thumb2 · d8f9b37f
Bruno Cardoso Lopes authored Jan 20, 2011
```
llvm-svn: 123929
```
d8f9b37f

- Use a more appropriate name for Owen's ARM Parser isMCR hack since the same... · 33461ecc

Bruno Cardoso Lopes authored Jan 20, 2011

- Use a more appropriate name for Owen's ARM Parser isMCR hack since the same operands can be present
 in cdp/cdp2 instructions. Also increase the hack with cdp/cdp2 instructions.
- Fix the encoding of cdp/cdp2 instructions for ARM (no thumb and thumb2 yet) and add testcases for t
hem.

llvm-svn: 123927

33461ecc

SplitKit requires that all defs are in place before calling useIntv(). · 8a46e26b

Jakob Stoklund Olesen authored Jan 20, 2011

The value mapping gets confused about which original values have multiple new
definitions so they may need phi insertions.

This could probably be simplified by letting enterIntvBefore() take a live range
to be added following the instruction. As long as the range stays inside the
same basic block, value mapping shouldn't be a problem.

llvm-svn: 123926

8a46e26b

Add LiveIntervalMap::dumpCache() to print out the cache used by the ssa update algorithm. · 04e6b3bd
Jakob Stoklund Olesen authored Jan 20, 2011
```
llvm-svn: 123925
```
04e6b3bd
Add mcr*2 and mr*c2 support to thumb2 targets · 4d4b490f
Bruno Cardoso Lopes authored Jan 20, 2011
```
llvm-svn: 123919
```
4d4b490f
Add mcr* and mr*c support to thumb targets · cf99dc7e
Bruno Cardoso Lopes authored Jan 20, 2011
```
llvm-svn: 123917
```
cf99dc7e
Allow sign-extending of i8 and i16 to i128 on SPU. · 6e5a54b3
Kalle Raiskila authored Jan 20, 2011
```
llvm-svn: 123912
```
6e5a54b3

At -O123 the early-cse pass is run before instcombine has run. According to my · 8fb2c382

Duncan Sands authored Jan 20, 2011

auto-simplier the transform most missed by early-cse is (zext X) != 0 -> X != 0.
This patch adds this transform and some related logic to InstructionSimplify
and removes some of the logic from instcombine (unfortunately not all because
there are several situations in which instcombine can improve things by making
new instructions, whereas instsimplify is not allowed to do this). At -O2 this
often results in more than 15% more simplifications by early-cse, and results in
hundreds of lines of bitcode being eliminated from the testsuite. I did see some
small negative effects in the testsuite, for example a few additional instructions
in three programs. One program, 483.xalancbmk, got an additional 35 instructions,
which seems to be due to a function getting an additional instruction and then
being inlined all over the place.

llvm-svn: 123911

8fb2c382

Refactor mcr* and mr*c instructions into classes with the same encoding. No functionality change. · 32f9b756
Bruno Cardoso Lopes authored Jan 20, 2011
```
llvm-svn: 123910
```
32f9b756
My editor's indent went crazy. Fix. · 37c4a8be
Eric Christopher authored Jan 20, 2011
```
llvm-svn: 123909
```
37c4a8be

Expand invalid return values for umulo and smulo. Handle these similarly · 785db078

Eric Christopher authored Jan 20, 2011

to add/sub by doing the normal operation and then checking for overflow
afterwards. This generally relies on the DAG handling the later invalid
operations as well.

Fixes the 64-bit part of rdar://8622122 and rdar://8774702.

llvm-svn: 123908

785db078

Correct itinerary entry for t2MOV_pic_ga_add_pc. · 7af85533
Evan Cheng authored Jan 20, 2011
```
llvm-svn: 123907
```
7af85533

Sorry, several patches in one. · b8b0ad80

Evan Cheng authored Jan 20, 2011

TargetInstrInfo:
Change produceSameValue() to take MachineRegisterInfo as an optional argument.
When in SSA form, targets can use it to make more aggressive equality analysis.

Machine LICM:
1. Eliminate isLoadFromConstantMemory, use MI.isInvariantLoad instead.
2. Fix a bug which prevent CSE of instructions which are not re-materializable.
3. Use improved form of produceSameValue.

ARM:
1. Teach ARM produceSameValue to look pass some PIC labels.
2. Look for operands from different loads of different constant pool entries
   which have same values.
3. Re-implement PIC GA materialization using movw + movt. Combine the pair with
   a "add pc" or "ldr [pc]" to form pseudo instructions. This makes it possible
   to re-materialize the instruction, allow machine LICM to hoist the set of
   instructions out of the loop and make it possible to CSE them. It's a bit
   hacky, but it significantly improve code quality.
4. Some minor bug fixes as well.

With the fixes, using movw + movt to materialize GAs significantly outperform the
load from constantpool method. 186.crafty and 255.vortex improved > 20%, 254.gap
and 176.gcc ~10%.

llvm-svn: 123905

b8b0ad80

Object: Add ELF support. · b60a18de
Michael J. Spencer authored Jan 20, 2011
```
llvm-svn: 123896
```
b60a18de
Object: Add COFF Support. · 8e90adaf
Michael J. Spencer authored Jan 20, 2011
```
llvm-svn: 123895
```
8e90adaf

Selection DAG scheduler register pressure heuristic fixes. · 2cd1f0be

Andrew Trick authored Jan 20, 2011

Added a check for already live regs before claiming HighRegPressure.
Fixed a few cases of checking the wrong number of successors.
Added some tracing until these heuristics are better understood.

llvm-svn: 123892

2cd1f0be

Check that a live range exists before shortening it. This fixes PR8989. · 4060abb4
Jakob Stoklund Olesen authored Jan 20, 2011
```
The live range may have been deleted earlier because of rematerialization.

llvm-svn: 123891
```
4060abb4
Add hidden -verify-coalescing to run the machine code verifier before and after · 145755f1
Jakob Stoklund Olesen authored Jan 20, 2011
```
register coalescing.

llvm-svn: 123890
```
145755f1
Sparc backend: Implements a delay slot filler that attempt to fill delay slots · 058e1247
Venkatraman Govindaraju authored Jan 20, 2011
```
with useful instructions.

llvm-svn: 123884
```
058e1247
Update a comment. · 050eec1d
Cameron Zwarich authored Jan 20, 2011
```
llvm-svn: 123879
```
050eec1d
Fix bug found by new clang warning. · 5acd4a64
Jakob Stoklund Olesen authored Jan 20, 2011
```
llvm-svn: 123872
```
5acd4a64
Use only one API at a time. · b2139f65
Eric Christopher authored Jan 20, 2011
```
llvm-svn: 123866
```
b2139f65

If we can, lower the multiply part of a umulo/smulo call to a libcall · bb14f656

Eric Christopher authored Jan 20, 2011

with an invalid type then split the result and perform the overflow check
normally.

Fixes the 32-bit parts of rdar://8622122 and rdar://8774702.

llvm-svn: 123864

bb14f656

Fix debug info for merged global. · 2d9e532a
Devang Patel authored Jan 20, 2011
```
llvm-svn: 123862
```
2d9e532a
Divert Hopfield network debug output. It is very noisy. · 79be8aec
Jakob Stoklund Olesen authored Jan 19, 2011
```
llvm-svn: 123859
```
79be8aec

Don't accidentally leave small gaps in the live ranges when leaving the active · 509089f5

Jakob Stoklund Olesen authored Jan 19, 2011

interval after an instruction. The leaveIntvAfter() method only adds liveness
from the instruction's boundary index to the inserted copy.

Ideally, SplitKit should be smarter about this, perhaps by combining useIntv()
and leaveIntvAfter() into one method that guarantees continuity.

llvm-svn: 123858

509089f5

Make sure to propogate the error code when we fail to parse a modifier. · 493c0fbd
Jim Grosbach authored Jan 19, 2011
```
llvm-svn: 123857
```
493c0fbd
Fix register address expression. Patch by Ken Dyck. · 8698f09d
Devang Patel authored Jan 19, 2011
```
llvm-svn: 123856
```
8698f09d

Jan 19, 2011

Implement RAGreedy::splitAroundRegion and remove loop splitting. · 9fb04015

Jakob Stoklund Olesen authored Jan 19, 2011

Region splitting includes loop splitting as a subset, and it is more generic.
The splitting heuristics for variables that are live in more than one block are
now:

1. Try to create a region that covers multiple basic blocks.
2. Try to create a new live range for each block with multiple uses.
3. Spill.

Steps 2 and 3 are similar to what the standard spiller is doing.

llvm-svn: 123853

9fb04015

Similarly, analyze truncate through multiply. · 5c901f34
Nick Lewycky authored Jan 19, 2011
```
llvm-svn: 123842
```
5c901f34