Commits · d6d9ccca38f47594a61a5636b58dc99c6d2ec0ea · Roger Ferrer / llvm-epi-0.8

Oct 18, 2012
- Temporarily revert the TargetTransform changes. · d6d9ccca
  Bob Wilson authored Oct 18, 2012
  
  The TargetTransform changes are breaking LTO bootstraps of clang. I am working with Nadav to figure out the problem, but I am reverting it for now to get our buildbots working. This reverts svn commits: 165665 165669 165670 165786 165787 165997 and I have also reverted clang svn 165741 llvm-svn: 166168
  d6d9ccca
- Remove the use of dominators and AA. · 642efbcd
  Nadav Rotem authored Oct 18, 2012
  
  llvm-svn: 166167
  642efbcd
- Vectorizer: Add support for loops with an unknown count. For example: · b52f7174
  Nadav Rotem authored Oct 18, 2012
  
  for (i=0; i<n; i++){ a[i] = b[i+1] + c[i+3]; } llvm-svn: 166165
  b52f7174
- Revert r166157 because some tests fail... · 4be9013e
  Bill Wendling authored Oct 17, 2012
  
  llvm-svn: 166159
  4be9013e
- Check that the operand of the GEP is not the GEP itself. This occurred during an LTO build of LLVM. · 7c234246
  Bill Wendling authored Oct 17, 2012
  
  llvm-svn: 166157
  7c234246
- Revert part of r166049 back and enable test case in r166125. · 3ac8201e
  Michael Liao authored Oct 17, 2012
  
  - Folding (trunc (concat ... X )) to (concat ... (trunc X) ...) is valid when '...' are all 'undef's. - r166125 relies on this transformation. llvm-svn: 166155
  3ac8201e
- LoopVectorize.cpp: Fix a warning. [-Wunused-variable] · 78574157
  NAKAMURA Takumi authored Oct 17, 2012
  
  llvm-svn: 166153
  78574157
- Remove redundant SetInsertPoint call. · 68e5dfdd
  Jakub Staszak authored Oct 17, 2012
  
  llvm-svn: 166138
  68e5dfdd
- Revert r166049 · c87d98db
  Michael Liao authored Oct 17, 2012
  
  - In general, it's unsafe for this transformation. llvm-svn: 166135
  c87d98db
- Add conditional branch instructions and their patterns. · 6743924a
  Reed Kotler authored Oct 17, 2012
  
  llvm-svn: 166134
  6743924a
Oct 17, 2012

Fix some typos and wrong indenting. · 4955ec31
Roman Divacky authored Oct 17, 2012
```
llvm-svn: 166128
```
4955ec31

Teach DAG combine to fold (extract_subvec (concat v1, ..) i) to v_i · 7a442c80

Michael Liao authored Oct 17, 2012

- If the extracted vector has the same type of all vectored being concatenated
  together, it should be simplified directly into v_i, where i is the index of
  the element being extracted.

llvm-svn: 166125

7a442c80

Switch MRI::UsedPhysRegs to a register unit bit vector. · 7a9f0c09

Jakob Stoklund Olesen authored Oct 17, 2012

This is a more compact, less redundant representation, and it avoids
scanning long lists of aliases for ARM D-registers, for example.

llvm-svn: 166124

7a9f0c09

Add a really faster pre-RA scheduler (-pre-RA-sched=linearize). It doesn't use · 839fb650

Evan Cheng authored Oct 17, 2012

any scheduling heuristics nor does it build up any scheduling data structure
that other heuristics use. It essentially linearize by doing a DFA walk but
it does handle glues correctly.

IMPORTANT: it probably can't handle all the physical register dependencies so
it's not suitable for x86. It also doesn't deal with dbg_value nodes right now
so it's definitely is still WIP.

rdar://12474515

llvm-svn: 166122

839fb650

Merge MRI::isPhysRegOrOverlapUsed() into isPhysRegUsed(). · 07364426

Jakob Stoklund Olesen authored Oct 17, 2012

All callers of these functions really want the isPhysRegOrOverlapUsed()
functionality which also checks aliases. For historical reasons, targets
without register aliases were calling isPhysRegUsed() instead.

Change isPhysRegUsed() to also check aliases, and switch all
isPhysRegOrOverlapUsed() callers to isPhysRegUsed().

llvm-svn: 166117

07364426

Add a loop vectorizer. · 6b94c2a0
Nadav Rotem authored Oct 17, 2012
```
llvm-svn: 166112
```
6b94c2a0

Check for empty YMM use-def lists in X86VZeroUpper. · a10c0980

Jakob Stoklund Olesen authored Oct 17, 2012

The previous MRI.isPhysRegUsed(YMM0) would also return true when the
function contains a call to a function that may clobber YMM0. That's
most of them.

Checking the use-def chains allows us to skip functions that don't
explicitly mention YMM registers.

llvm-svn: 166110

a10c0980

Fix fallout from RegInfo => FrameLowering refactoring on MSP430. · 0a69176c
Anton Korobeynikov authored Oct 17, 2012
```
Patch by Job Noorman!

llvm-svn: 166108
```
0a69176c
misched: Better handling of invalid latencies in the machine model · 0b1d8d04
Andrew Trick authored Oct 17, 2012
```
llvm-svn: 166107
```
0b1d8d04

Support: Don't remove special files on signals. · 511479dd

Daniel Dunbar authored Oct 17, 2012

 - Similar to Path::eraseFromDisk(), we don't want LLVM to remove things like
   /dev/null, even if it has the permission.

llvm-svn: 166105

511479dd

[asan] better debug diagnostics in asan compiler module · 20343351
Kostya Serebryany authored Oct 17, 2012
```
llvm-svn: 166102
```
20343351

This just in, it is a *bad idea* to use 'udiv' on an offset of · 6fab42aa

Chandler Carruth authored Oct 17, 2012

a pointer. A very bad idea. Let's not do that. Fixes PR14105.

Note that this wasn't *that* glaring of an oversight. Originally, these
routines were only called on offsets within an alloca, which are
intrinsically positive. But over the evolution of the pass, they ended
up being called for arbitrary offsets, and things went downhill...

llvm-svn: 166095

6fab42aa

Fix a really annoying "bug" introduced in r165941. The change from that · 40617f59

Chandler Carruth authored Oct 17, 2012

revision makes no sense. We cannot use the address space of the *post
indexed* type to conclude anything about a *pre indexed* pointer type's
size. More importantly, this index can never be over a pointer. We are
indexing over arrays and vectors here.

Of course, I have no test case here. Neither did the original patch. =/

llvm-svn: 166091

40617f59

Check SSSE3 instead of SSE4.1 · cef9541d

Michael Liao authored Oct 17, 2012

- All shuffle insns required, especially PSHUB, are added in SSSE3.

llvm-svn: 166086

cef9541d

Fix setjmp on models with non-Small code model nor non-Static relocation model · 6f720613

Michael Liao authored Oct 17, 2012

- MBB address is only valid as an immediate value in Small & Static
  code/relocation models. On other models, LEA is needed to load IP address of
  the restore MBB.
- A minor fix of MBB in MC lowering is added as well to enable target
  relocation flag being propagated into MC.

llvm-svn: 166084

6f720613

Use a SparseSet instead of a BitVector for UsedInInstr in RAFast. · a2136be1

Jakob Stoklund Olesen authored Oct 17, 2012

This is just as fast, and it makes it possible to avoid leaking the
UsedPhysRegs BitVector implementation through
MachineRegisterInfo::addPhysRegsUsed().

llvm-svn: 166083

a2136be1

Use a typedef to reduce some typing and reformat code accordingly. · 494109b0
Eric Christopher authored Oct 16, 2012
```
llvm-svn: 166077
```
494109b0
Variable name cleanup. · 02509481
Eric Christopher authored Oct 16, 2012
```
llvm-svn: 166076
```
02509481

Avoid rematerializing a redef immediately after the old def. · 4df59a9f

Jakob Stoklund Olesen authored Oct 16, 2012

PR14098 contains an example where we would rematerialize a MOV8ri
immediately after the original instruction:

  %vreg7:sub_8bit<def> = MOV8ri 9; GR32_ABCD:%vreg7
  %vreg22:sub_8bit<def> = MOV8ri 9; GR32_ABCD:%vreg7

Besides being pointless, it is also wrong since the original instruction
only redefines part of the register, and the value read by the new
instruction is wrong.

The problem was the LiveRangeEdit::allUsesAvailableAt() didn't
special-case OrigIdx == UseIdx and found the wrong SSA value.

llvm-svn: 166068

4df59a9f

Revert r166046 "Switch back to the old coalescer for now to fix the 32 bit bit" · 2043329e
Jakob Stoklund Olesen authored Oct 16, 2012
```
A fix for PR14098, including the test case is in the next commit.

llvm-svn: 166067
```
2043329e

Oct 16, 2012

[InstCombine] Teach InstCombine how to handle an obfuscated splat. · 02a1141e

Michael Gottesman authored Oct 16, 2012

An obfuscated splat is where the frontend poorly generates code for a splat
using several different shuffles to create the splat, i.e.,

  %A = load <4 x float>* %in_ptr, align 16
  %B = shufflevector <4 x float> %A, <4 x float> undef, <4 x i32> <i32 0, i32 0, i32 undef, i32 undef>
  %C = shufflevector <4 x float> %B, <4 x float> %A, <4 x i32> <i32 0, i32 1, i32 4, i32 undef>
  %D = shufflevector <4 x float> %C, <4 x float> %A, <4 x i32> <i32 0, i32 1, i32 2, i32 4>

llvm-svn: 166061

02a1141e

[ms-inline asm] Add the helper function, isParseringInlineAsm(). To be used in a future commit. · e4ad2a0b
Chad Rosier authored Oct 16, 2012
```
llvm-svn: 166054
```
e4ad2a0b
Simplify code. No functionality change. · 8f46e914
Jakub Staszak authored Oct 16, 2012
```
llvm-svn: 166053
```
8f46e914
Check .rela instead of ELF64 for the compensation vaue resetting · d6f3168a
Michael Liao authored Oct 16, 2012
```
llvm-svn: 166051
```
d6f3168a
80-col fixup. · 25dcab1e
Jakub Staszak authored Oct 16, 2012
```
llvm-svn: 166050
```
25dcab1e
Teach DAG combine to fold (trunc (fptoXi x)) to (fptoXi x) · 19006206
Michael Liao authored Oct 16, 2012
```
llvm-svn: 166049
```
19006206
Switch back to the old coalescer for now to fix the 32 bit bit · b58be2c5
Rafael Espindola authored Oct 16, 2012
```
llvm+clang+compiler-rt bootstrap.

llvm-svn: 166046
```
b58be2c5
Simplify potentially quadratic behavior while erasing elements from std::vector. · ba34fdb0
Jakub Staszak authored Oct 16, 2012
```
llvm-svn: 166045
```
ba34fdb0

Support v8f32 to v8i8/vi816 conversion through custom lowering · 02ca3454

Michael Liao authored Oct 16, 2012

- Add custom FP_TO_SINT on v8i16 (and v8i8 which is legalized as v8i16 due to
  vector element-wise widening) to reduce DAG combiner and its overhead added
  in X86 backend.

llvm-svn: 166036

02ca3454

This patch addresses PR13949. · 48081cad

Bill Schmidt authored Oct 16, 2012

For the PowerPC 64-bit ELF Linux ABI, aggregates of size less than 8
bytes are to be passed in the low-order bits ("right-adjusted") of the
doubleword register or memory slot assigned to them. A previous patch
addressed this for aggregates passed in registers. However, small
aggregates passed in the overflow portion of the parameter save area are
still being passed left-adjusted.

The fix is made in PPCTargetLowering::LowerCall_Darwin_Or_64SVR4 on the
caller side, and in PPCTargetLowering::LowerFormalArguments_64SVR4 on
the callee side. The main fix on the callee side simply extends
existing logic for 1- and 2-byte objects to 1- through 7-byte objects,
and correcting a constant left over from 32-bit code. There is also a
fix to a bogus calculation of the offset to the following argument in
the parameter save area.

On the caller side, again a constant left over from 32-bit code is
fixed. Additionally, some code for 1, 2, and 4-byte objects is
duplicated to handle the 3, 5, 6, and 7-byte objects for SVR4 only. The
LowerCall_Darwin_Or_64SVR4 logic is getting fairly convoluted trying to
handle both ABIs, and I propose to separate this into two functions in a
future patch, at which time the duplication can be removed.

The patch adds a new test (structsinmem.ll) to demonstrate correct
passing of structures of all seven sizes. Eight dummy parameters are
used to force these structures to be in the overflow portion of the
parameter save area.

As a side effect, this corrects the case when aggregates passed in
registers are saved into the first eight doublewords of the parameter
save area: Previously they were stored left-justified, and now are
properly stored right-justified. This requires changing the expected
output of existing test case structsinregs.ll.

llvm-svn: 166022

48081cad