Commits · 7bf08432d2c618de728bc7671b8c5360966050bf · Roger Ferrer / llvm-epi-0.8

Feb 09, 2010
- Move Intrinsic::objectsize lowering back to InstCombineCalls and · 7b7028fd
  Eric Christopher authored Feb 09, 2010
```
enable constant 0 offset lowering.

llvm-svn: 95691
```
  7b7028fd
- Pull these back out, they're a little too aggressive and time · ad1aa862
  Eric Christopher authored Feb 09, 2010
```
consuming for a simple optimization.

llvm-svn: 95671
```
  ad1aa862
- simplify this code, duh. · f4c8d3ce
  Chris Lattner authored Feb 09, 2010
```
llvm-svn: 95643
```
  f4c8d3ce
- fix PR6193, only considering sign extensions *from i1* for this · 9b6a1789
  Chris Lattner authored Feb 09, 2010
```
xform.

llvm-svn: 95642
```
  9b6a1789
- Add file in here too. · be2f0b2b
  Eric Christopher authored Feb 09, 2010
```
llvm-svn: 95641
```
  be2f0b2b
- Add a new pass to do llvm.objsize lowering using SCEV. · 9f85e7eb
  Eric Christopher authored Feb 09, 2010
```
Initial skeleton and SCEVUnknown lowering implemented,
the rest should come relatively quickly.  Move testcase
to new directory.

Move pass to right before SimplifyLibCalls - which is
moved down a bit so we can take advantage of a few opts.

llvm-svn: 95628
```
  9f85e7eb
- fix some problems handling large vectors reported in PR6230 · b22423c8
  Chris Lattner authored Feb 08, 2010
```
llvm-svn: 95616
```
  b22423c8
Feb 06, 2010

Reintroduce the InlineHint function attribute. · 74bb06c0

Jakob Stoklund Olesen authored Feb 06, 2010

This time it's for real! I am going to hook this up in the frontends as well.

The inliner has some experimental heuristics for dealing with the inline hint.
When given a -respect-inlinehint option, functions marked with the inline
keyword are given a threshold just above the default for -O3.

We need some experiments to determine if that is the right thing to do.

llvm-svn: 95466

74bb06c0

Don't unroll loops containing function calls. · 5f9ead27
Jakob Stoklund Olesen authored Feb 05, 2010
```
llvm-svn: 95454
```
5f9ead27

Feb 05, 2010

Teach SimplifyCFG about magic pointer constants. · 916f48a0

Jakob Stoklund Olesen authored Feb 05, 2010

Weird code sometimes uses pointer constants other than null. This patch
teaches SimplifyCFG to build switch instructions in those cases.

Code like this:

void f(const char *x) {
  if (!x)
    puts("null");
  else if ((uintptr_t)x == 1)
    puts("one");
  else if (x == (char*)2 || x == (char*)3)
    puts("two");
  else if ((intptr_t)x == 4)
    puts("four");
  else
    puts(x);
}

Now becomes a switch:

define void @f(i8* %x) nounwind ssp {
entry:
  %magicptr23 = ptrtoint i8* %x to i64            ; <i64> [#uses=1]
  switch i64 %magicptr23, label %if.else16 [
    i64 0, label %if.then
    i64 1, label %if.then2
    i64 2, label %if.then9
    i64 3, label %if.then9
    i64 4, label %if.then14
  ]

Note that LLVM's own DenseMap uses magic pointers.

llvm-svn: 95439

916f48a0

fix logical-select to invoke filecheck right, and fix hte instcombine · 64ffd11d

Chris Lattner authored Feb 05, 2010

xform it is checking to actually pass.  There is no need to match
m_SelectCst<0, -1> since instcombine canonicalizes that into not(sext).

Add matches for sext(not(x)) in addition to not(sext(x)).

llvm-svn: 95420

64ffd11d

Implement releaseMemory in CodeGenPrepare and free the BackEdges · 4739e41c

Dan Gohman authored Feb 05, 2010

container data. This prevents it from holding onto dangling
pointers and potentially behaving unpredictably.

llvm-svn: 95409

4739e41c

Use a SmallSetVector instead of a SetVector; this code showed up as a · 8abb67df
Dan Gohman authored Feb 05, 2010
```
malloc caller in a profile.

llvm-svn: 95407
```
8abb67df
Remove this code for now. I have a better idea and will rewrite with · 04371b4f
Eric Christopher authored Feb 05, 2010
```
that in mind.

llvm-svn: 95402
```
04371b4f

Do not reassociate expressions with i1 type. SimplifyCFG converts some · 27dfb1e1

Bob Wilson authored Feb 04, 2010

short-circuited conditions to AND/OR expressions, and those expressions
are often converted back to a short-circuited form in code gen.  The
original source order may have been optimized to take advantage of the
expected values, and if we reassociate them, we change the order and
subvert that optimization.  Radar 7497329.

llvm-svn: 95333

27dfb1e1

Feb 04, 2010

Increase inliner thresholds by 25. · 113fb54b

Jakob Stoklund Olesen authored Feb 04, 2010

This makes the inliner about as agressive as it was before my changes to the
inliner cost calculations. These levels give the same performance and slightly
smaller code than before.

llvm-svn: 95320

113fb54b

Temporarily revert this since it appears to have caused a build · 107a1fbf
Eric Christopher authored Feb 04, 2010
```
failure.

llvm-svn: 95294
```
107a1fbf

Rework constant expr and array handling for objectsize instcombining. · 42fa84a8

Eric Christopher authored Feb 04, 2010

Fix bugs where we would compute out of bounds as in bounds, and where
we couldn't know that the linker could override the size of an array.

Add a few new testcases, change existing testcase to use a private
global array instead of extern.

llvm-svn: 95283

42fa84a8

If we're dealing with a zero-length array, don't lower to any · f12e18db
Eric Christopher authored Feb 03, 2010
```
particular size, we just don't know what the length is yet.

llvm-svn: 95266
```
f12e18db

Feb 03, 2010

Adjust the heuristics used to decide when SROA is likely to be profitable. · 04365c5f

Bob Wilson authored Feb 03, 2010

The SRThreshold value makes perfect sense for checking if an entire aggregate
should be promoted to a scalar integer, but it is not so good for splitting
an aggregate into its separate elements. A struct may contain a large embedded
array along with some scalar fields that would benefit from being split apart
by SROA. Even if the total aggregate size is large, it may still be good to
perform SROA. Thus, the most important piece of this patch is simply moving
the aggregate size comparison vs. SRThreshold so that it guards only the
aggregate promotion.

We have also been checking the number of elements to decide if an aggregate
should be split up. The limit of "SRThreshold/4" seemed rather arbitrary,
and I don't think it's very useful to derive this limit from SRThreshold
anyway. I've collected some data showing that the current default limit of
32 (since SRThreshold defaults to 128) is a reasonable cutoff for struct
types. One thing suggested by the data is that distinguishing between structs
and arrays might be useful. There are (obviously) a lot more large arrays
than large structs (as measured by the number of elements and not the total
size -- a large array inside a struct still counts as a single element given
the way we do SROA right now). Out of 8377 arrays where we successfully
performed SROA while compiling a large set of benchmarks, only 16 of them had
more than 8 elements. And, for those 16 arrays, it's not at all clear that
SROA was actually beneficial. So, to offset the compile time cost of
investigating more large structs for SROA, the patch lowers the limit on array
elements to 8.

This fixes Apple Radar 7563690.

llvm-svn: 95224

04365c5f

Revert 94937 and move the noreturn check to codegen. · 27a41d54
Evan Cheng authored Feb 03, 2010
```
llvm-svn: 95198
```
27a41d54
Fix some comment typos. · 76e8c595
Bob Wilson authored Feb 03, 2010
```
llvm-svn: 95170
```
76e8c595
Recommit this, looks like it wasn't the cause. · d86233c1
Eric Christopher authored Feb 03, 2010
```
llvm-svn: 95165
```
d86233c1
Hopefully temporarily revert this. · e67d01a9
Eric Christopher authored Feb 02, 2010
```
llvm-svn: 95154
```
e67d01a9

Feb 02, 2010
- Reformat my last patch slightly. · f9553572
  Eric Christopher authored Feb 02, 2010
```
llvm-svn: 95147
```
  f9553572
- Re-add strcmp and known size object size checking optimization. · 4264e7e4
  Eric Christopher authored Feb 02, 2010
```
Passed bootstrap and nightly test run here.

llvm-svn: 95145
```
  4264e7e4
- don't turn (A & (C0?-1:0)) | (B & ~(C0?-1:0)) -> C0 ? A : B · 8e2c4716
  Chris Lattner authored Feb 02, 2010
```
for vectors.  Codegen is generating awful code or segfaulting
in various cases (e.g. PR6204).

llvm-svn: 95058
```
  8e2c4716
- fix a crash in loop unswitch on a loop invariant vector condition. · 302240d7
  Chris Lattner authored Feb 02, 2010
```
llvm-svn: 95055
```
  302240d7
- LangRef.html says that inttoptr and ptrtoint always use zero-extension · 949458d0
  Dan Gohman authored Feb 02, 2010
```
when the cast is extending.

llvm-svn: 95046
```
  949458d0
- Don't need to check the last argument since it'll always be bool. We also · 14dfc3f6
  Eric Christopher authored Feb 02, 2010
```
don't use TargetData here.

llvm-svn: 95040
```
  14dfc3f6
- More indentation/tabification fixes. · 9afa9732
  Eric Christopher authored Feb 02, 2010
```
llvm-svn: 95036
```
  9afa9732
- Untabify previous commit. · 14082347
  Eric Christopher authored Feb 02, 2010
```
llvm-svn: 95035
```
  14082347
- Formatting. · 56e4182c
  Eric Christopher authored Feb 01, 2010
```
llvm-svn: 95027
```
  56e4182c
Feb 01, 2010

Add an option to GVN to remove all partially redundant loads. This is currently · d517b520

Bob Wilson authored Feb 01, 2010

disabled by default.  This divides the existing load PRE code into 2 phases:
first it checks that it is safe to move the load to each of the predecessors
where it is unavailable, and then if it is safe, the code is changed to move
the load.  Radar 7571861.

llvm-svn: 95007

d517b520

cleanups. · 9306ffa0
Chris Lattner authored Feb 01, 2010
```
llvm-svn: 94995
```
9306ffa0

fix rdar://7590304 , a miscompilation of objc apps on arm. The caller · 846a52e2

Chris Lattner authored Feb 01, 2010

of objc message send was getting marked arm_apcscc, but the prototype
isn't.  This is fine at runtime because objcmsgsend is implemented in
assembly.  Only turn a mismatched caller and callee into 'unreachable'
if the callee is a definition.

llvm-svn: 94986

846a52e2

fix rdar://7590304 , an infinite loop in instcombine. In the invoke · 2cecedf0

Chris Lattner authored Feb 01, 2010

case, instcombine can't zap the invoke for fear of changing the CFG.
However, we have to do something to prevent the next iteration of
instcombine from inserting another store -> undef before the invoke
thereby getting into infinite iteration between dead store elim and
store insertion.

Just zap the callee to null, which will prevent the next iteration
from doing anything.

llvm-svn: 94985

2cecedf0

Fix pr6198 by moving the isSized() check to an outer conditional. · f65ba356

Bob Wilson authored Feb 01, 2010

The testcase from pr6198 does not crash for me -- I don't know what's up with
that -- so I'm not adding it to the tests.

llvm-svn: 94984

f65ba356

Jan 31, 2010
- Simplify/generalize the xor+add->sign-extend instcombine. · a2cc2875
  Eli Friedman authored Jan 31, 2010
```
llvm-svn: 94943
```
  a2cc2875
- Add a small transform: transform -(X<<Y) to (-X<<Y) when the shift has a single · 37a8197b
  Eli Friedman authored Jan 31, 2010
```
use and X is free to negate.

llvm-svn: 94941
```
  37a8197b