- Apr 26, 2021
-
-
Alexey Bader authored
-
- Apr 27, 2021
-
-
Wenlei He authored
CSSPGO only supports text and extended binary profile now. Raw binary does not have the metadata section CSSPGO needs, and Compact binary profile needs special handling for GUID based context names, which is not yet implemented. Disasslow these two format for CSSPGO profile writing to avoid silently generating invalid profiles. Differential Revision: https://reviews.llvm.org/D101300
-
- Apr 26, 2021
-
-
Michael Kruse authored
When replacing a conditional branch by an unconditional one because the condition is a constant, transfer the metadata to the new branch instruction. Part of fix for llvm.org/PR50060 Reviewed By: lebedev.ri Differential Revision: https://reviews.llvm.org/D101141
-
Frederik Gossen authored
Ensure to preserve the correct type during when folding and canonicalization. `shape.broadcast` of of a single operand can only be folded away if the argument type is correct. Differential Revision: https://reviews.llvm.org/D101158
-
Mark de Wever authored
Seems our check-generated-output doesn't work as intended. Will investigate it further.
-
Sebastian Neubauer authored
Use MapVector instead of SmallDenseMap because it has a deterministic iteration order. Differential Revision: https://reviews.llvm.org/D101299
-
Alexey Bader authored
Initial version of the document covers address space handling Differential Revision: https://reviews.llvm.org/D99488
-
Simon Pilgrim authored
Noticed while triaging the rG2149aa73 regressions - the LXADD ops are load+store RMW instructions, not just loads.
-
Greg McGary authored
Add option to limit (or remove limits) on the number of errors printed before exiting. This option exists in the other lld ports: COFF & ELF. Differential Revision: https://reviews.llvm.org/D101274
-
Dávid Bolvanský authored
This reverts commit b125c034. Some bots say " missing binary operator before token "(" "
-
Nigel Perks authored
The headers shipped with the XMOS XCore compiler expect __xcore__ to be defined. The __XS1B__ macro, already defined, is for the default subtarget. No other targets affected.
-
Paul C. Anagnostopoulos authored
-
Dávid Bolvanský authored
This is needed after https://reviews.llvm.org/rG9658d045926545e62cc3f963fe611d7c5d0c9d98 which introduced code that at least didn't compile clean with clang 8. Reviewed By: xbolva00 Differential Revision: https://reviews.llvm.org/D101288
-
Paul C. Anagnostopoulos authored
Differential Revision: https://reviews.llvm.org/D100854
-
Alexey Bader authored
Default address space (applies when no explicit address space was specified) maps to generic (4) address space. Added SYCL named address spaces `sycl_global`, `sycl_local` and `sycl_private` defined as sub-sets of the default address space. Static variables without address space now reside in global address space when compile for SPIR target, unless they have an explicit address space qualifier in source code. Differential Revision: https://reviews.llvm.org/D89909
-
Dávid Bolvanský authored
Proof: https://alive2.llvm.org/ce/z/mncA9K Solves https://bugs.llvm.org/show_bug.cgi?id=50104 Reviewed By: spatel Differential Revision: https://reviews.llvm.org/D101257
-
Tim Northover authored
-
Emre Kultursay authored
This change is similar in spirit to the change at: https://reviews.llvm.org/rG34c697c85e9d0af11a72ac4df5578aac94a627b3 It fixes the problem where the layout of a type was being accessed while its base classes were not populated yet; which caused an incorrect layout to be produced and cached. This fixes PR50054 Reviewed By: teemperor Differential Revision: https://reviews.llvm.org/D100977
-
Tim Renouf authored
Stops llvm-objdump tests failing when AMDGPU target is not supported. Change-Id: Ic4ae443958c41c303ff6bee0966e5f21ab7a1851
-
Tim Renouf authored
1. Add an accessor function to MCSymbolizer to retrieve addresses referenced by a symbolizable operand, but not resolved to a symbol. That way, the caller can synthesize labels at those addresses and then retry disassembling the section. 2. Implement that in AMDGPU -- a failed symbol lookup results in the address being added to a vector returned by the new function. 3. Use that in llvm-objdump when using MCSymbolizer (which only happens on AMDGPU) and SymbolizeOperands is on. Differential Revision: https://reviews.llvm.org/D101145 Change-Id: I19087c3bbfece64bad5a56ee88bcc9110d83989e
-
Sanjay Patel authored
The script warns that this could be buggy when updating.
-
Yuanbo Li authored
When transforming a loop terminating condition into a "max" comparison, the DebugLoc from the old condition should be set on the newly created comparison. They are the same operation, just optimized. Fixes PR48067. Differential Revision: https://reviews.llvm.org/D98218
-
Bradley Smith authored
Differential Revision: https://reviews.llvm.org/D101058
-
Dávid Bolvanský authored
-
David Green authored
This expands the VMOVRRD(extract(..(build_vector(a, b, c, d)))) pattern, to also handle insert_vectors. Providing we can find the correct insert, this helps further simplify patterns by removing the redundant VMOVRRD. Differential Revision: https://reviews.llvm.org/D100245
-
David Green authored
-
David Sherwood authored
We can already vectorize loops that involve int<>int, fp<>fp, int<>fp and fp<>int conversions, however we didn't previously have any tests for them. This patch adds some tests for each conversion type. Differential Revision: https://reviews.llvm.org/D99951
-
David Sherwood authored
When vectorising for AArch64 targets if you specify the SVE attribute we automatically then treat masked loads and stores as legal. Also, since we have no cost model for masked memory ops we believe it's cheap to use the masked load/store intrinsics even for fixed width vectors. This can lead to poor code quality as the intrinsics will currently be scalarised in the backend. This patch adds a basic cost model that marks fixed-width masked memory ops as significantly more expensive than for scalable vectors. Tests for the cost model are added here: Transforms/LoopVectorize/AArch64/masked-op-cost.ll Differential Revision: https://reviews.llvm.org/D100745
-
Florian Hahn authored
When iterating over const blocks, the base type in the lambdas needs to use const VPBlockBase *, otherwise it cannot be used with input iterators over const VPBlockBase. Also adjust the type of the input iterator range to const &, as it does not take ownership of the input range.
-
Andrzej Warzynski authored
When generating output for `-fdebug-dump-symbols`, make sure that BuildRuntimeDerivedTypeTables is also run. This change is needed in order to make the implementation of `-fdebug-dump-symbols` in `flang-new` consistent with `f18`. It also allows us to port more tests to use the new driver whenever it is enabled. Differential Revision: https://reviews.llvm.org/D100649
-
Jeremy Morse authored
The test added in D97533 (and modified by this patch) has some overly strict printed metadata ordering requirements, specifically the interleaving of DILocalVariable nodes and DILocation nodes. Slight changes in metadata emission can easily break this unfortunately. This patch stops after clang codegen rather than allowing the coro splitter to run, and reduces the need for ordering: it picks out the DILocalVariable nodes being sought, in any order (CHECK-DAG), and doesn't examine any DILocations. The implicit CHECK-NOT is what's important: the test seeks to ensure a duplicate set of DILocalVariables aren't emitted in the same scope. Differential Revision: https://reviews.llvm.org/D100298
-
Krasimir Georgiev authored
Clang-format was indenting the lines following the `?` in the added test case by +5 instead of +4. This only happens in a very specific situation, where the `?` is followed by a multiline block comment, as in the example. This fix addresses this without regressing any of the existing tests. Differential Revision: https://reviews.llvm.org/D101033
-
David Green authored
CGP can move instructions like a ptrtoint into a loop, but the MVETailPredication when converting them will currently assume invariant trip counts. This tries to ensure the operands are loop invariant, and bails if not. Differential Revision: https://reviews.llvm.org/D100550
-
Hsiangkai Wang authored
-
Djordje Todorovic authored
Initial (D96045) patch didn't handle split dwarf cases, so this fixes that bug. In addition, before applying this patch, we had a slowdown that happened after the D96045. With this patch, the slowdown will be fixed as well. Differential Revision: https://reviews.llvm.org/D100951
-
Jan Svoboda authored
-
Jan Svoboda authored
Add option to `clang-scan-deps` to enable/disable generation of command-line arguments with absolute paths. This is essentially a revert of D100533, but with improved naming and added test. Reviewed By: dexonsmith Differential Revision: https://reviews.llvm.org/D101051
-
Sven van Haastregt authored
Mirror the remaining C++ for OpenCL specific builtins from opencl-c.h to the TableGen builtin functions. Fixes PR50041 (part 2).
-
Hsiangkai Wang authored
Differential Revision: https://reviews.llvm.org/D100824
-