- Jun 29, 2023
-
-
Philip Reames authored
-
Snehasish Kumar authored
Add an overload for InstrProfWriter::write so that users can emit the buffer to a string. Also use this new overload for existing unit test usecases. Reviewed By: tejohnson Differential Revision: https://reviews.llvm.org/D153904
-
David Green authored
See D153507. The existing test is over-simplified, as written it should have been simpified prior to SLP vectorization. I have left it as-is to ensure the crash it was protecting against doesn't arise again. A new test with valid inputs is also added to show the incorrect costs of alt cmp vectorization.
-
Fraser Cormack authored
D143505 fixed/simplified folding of operations with SNaN operands. In doing so it introduced a crash when handling scalable vector types, wherein the scalable-vector ConstantVector was cast to a ConstantFP. Since we know by that point in the code that if we've found a NaN, we're dealing with a scalable-vector splat (as there are no other kinds of scalable-vector constant for which that holds), we can grab the splatted value and re-use the existing code, which will automatically splat the new NaN back to a scalable vector for us. Reviewed By: arsenm Differential Revision: https://reviews.llvm.org/D153566
-
Valentin Clement authored
Some symbols were not resolved in the device, host and self clause resulting in an `Internal: no symbol found` error. This patch adds symbol resolution for these clauses. Reviewed By: razvanlupusoru Differential Revision: https://reviews.llvm.org/D153919
-
Valentin Clement authored
Some compiler treat `acc routine` without a parallelism clause as if seq is present. Relax the parser rule to allow acc routine without clause. The default clause will be handled in lowering. Reviewed By: razvanlupusoru Differential Revision: https://reviews.llvm.org/D153896
-
- Jun 28, 2023
-
-
Paul Robinson authored
-
LLVM GN Syncbot authored
-
Yusra Syeda authored
Revert "[SystemZ][z/OS] This patch adds support for the ADA (associated data area), doing the following:" This reverts commit 9df0f66a.
-
Jeffrey Byrnes authored
[AMDGPU] NFC: Add schedule-relaxed-occupancy to relax occupancy targets for wave-limited/membound kernels Default scheduling behavior for these types of kernels is to chase high occupancy goals with scheduling heuristics, but allow occupancy drops if we are unable to reach the target. This (experimental, off-by-default) feature relaxes occupancy target from the beginning, which enables scheduler to produce better ILP schedules. Differential Revision: https://reviews.llvm.org/D153925 Change-Id: I112833214e2db869704591f4df3c4574d0fcbb1b
-
Shilei Tian authored
the color of finished task
-
Craig Topper authored
The functions are identical except for the opcode of the node. We can have a single function and use N->getOpcode(). Reviewed By: luke, paulwalker-arm Differential Revision: https://reviews.llvm.org/D153929
-
Paul Robinson authored
Differential Revision: https://reviews.llvm.org/D153884
-
LLVM GN Syncbot authored
-
Yusra Syeda authored
- Creates the ADA table to handle displacements - Emits the ADA section in the SystemZAsmPrinter - Lowers the ADA_ENTRY node into the appropriate load instruction Differential Revision: https://reviews.llvm.org/D153788
-
LLVM GN Syncbot authored
-
Felipe de Azevedo Piovezan authored
This concludes the migration of accelerator tables from LLDB code to LLVM code. Differential Revision: https://reviews.llvm.org/D153868
-
David Green authored
-
Guillaume Chatelet authored
Reviewed By: courbet Differential Revision: https://reviews.llvm.org/D153958
-
Joseph Huber authored
Currently, AMDGPU more or less only supports linking with LTO. If the user does not either pass `-flto` or `-Wl,-plugin-opt=mcpu=` manually linking will fail because the architecture's aren't compatible. THis patch simply passes `-mcpu` by default if it was specified. Should be a no-op if it's not actually used. Reviewed By: yaxunl Differential Revision: https://reviews.llvm.org/D153909
-
Jie Fu authored
[flang] Build broken due to no member named 'getNumScalableDims' in 'mlir::VectorType' after D153412 (NFC) /data/llvm-project/flang/lib/Optimizer/Dialect/FIROps.cpp:971:46: error: no member named 'getNumScalableDims' in 'mlir::VectorType' if (mlir::dyn_cast<mlir::VectorType>(ty).getNumScalableDims() == 0) ~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~ ^ 1 error generated. -
Youngsuk Kim authored
Partial progress towards removing in-tree uses of `Type::getPointerTo`, before we can deprecate the API. If the API is used solely to support an unnecessary bitcast, get rid of the bitcast as well. Reviewed By: nikic Differential Revision: https://reviews.llvm.org/D153933
-
Felipe de Azevedo Piovezan authored
This commit is replacing really old LLDB code, and we've found some odd behavior while doing this replacement. While the changes here are largely NFC, there are some subtle changes that fix such odd behavior. The most curious example of this is the method `FindCompleteObjCClassName`, which has a flag `must_be_implementation`. This flag was _only_ being respected for accelerator tables containing the atom `type_flags`, which seems counter-intuitive. The implementation for DWARF 5 tables does not do that and neither does the code introduced in this patch. There were other weird cases, for example, we found boolean logic that was always true in a code path: look for a `if !has_qualified_name...` deleted line; that condition was true by simple if/else analysis. Differential Revision: https://reviews.llvm.org/D153867
-
Serge Pavlov authored
Previously function template instantiations occurred with FP options that were in effect at the end of translation unit. It was a problem for late template parsing as these FP options were used as attributes of AST nodes and may result in crash. To fix it FP options are set to the state of the point of template definition. Differential Revision: https://reviews.llvm.org/D143241
-
Alexey Bataev authored
Added some extra checks for comapreCMP function if IsCompatibility is false to make it meat the strict weak ordering requirements to be correctly used in sort functions.
-
Kevin P. Neal authored
Add missed change requested in D153371.
-
Andrzej Warzynski authored
This is a follow-up of https://reviews.llvm.org/D153372 in which `numScalableDims` (single integer) was effectively replaced with `isScalableDim` bitmask. This change is a part of a larger effort to enable scalable vectorisation in Linalg. See this RFC for more context: * https://discourse.llvm.org/t/rfc-scalable-vectorisation-in-linalg/ Differential Revision: https://reviews.llvm.org/D153412
-
Nikita Popov authored
-
David Truby authored
On platforms which support COMDAT sections we should use them when linkonce or linkonce_odr linkage is requested. This is required on Windows (PE/COFF) and provides better behaviour than weak symbols on ELF-based platforms. This patch also reverts string literals to use linkonce instead of internal linkage now that comdats are supported. Differential Revision: https://reviews.llvm.org/D153768
-
Jingu Kang authored
Differential Revision: https://reviews.llvm.org/D153847
-
Nikita Popov authored
-
Felipe de Azevedo Piovezan authored
All the new code should match the behavior of the old exactly. Of note, the custom queries used to be implemented inside `HashedNameToDIE.cpp` (which is the LLDB implementation of the tables). However, when porting to LLVM, we believe they don't belong inside the LLVM table implementation: 1. They don't require any knowledge about the table itself 2. They are not relevant for other users of these classes. 3. They use LLDB data structures. As such, we implement these custom queries inside AppleDWARFIndex.cpp. Types and Objective-C tables are done separately, as they have slightly different functionality that require rewriting more code. Differential Revision: https://reviews.llvm.org/D153866
-
John Brawn authored
When we only have a 16-bit pc-relative branch instruction we generate a table of address for a jump table. Currently this is placed inline, but this won't work with execute-only memory. In this case generate the jump table out-of-line. Differential Revision: https://reviews.llvm.org/D153774
-
Kevin P. Neal authored
The sort of the elements in the GET_SUBTARGETINFO_MACRO block is done on the "Name" field of each record. This field is not guaranteed to be unique, is not guaranteed to even have a value at all, and is not used in the output anyway. Change to sort on the "FieldName" field which should be unique. Problem spotted when lib/Target/PowerPC/PPCGenSubtargetInfo.inc changed unexpectedly. Differential Revision: https://reviews.llvm.org/D153371
-
Nikita Popov authored
-
Florian Hahn authored
This patch extends the existing logic to handle cases where we have branch conditions of the form (AND icmp, icmp) where the first icmp implies the second. This can improve results in some cases, e.g. if SimplifyCFG folded conditions from multiple branches to an AND. The implementation handles this by adding a new type of check (AndImpliedCheck), which are queued before conditional facts for the same block. When encountering AndImpliedChecks during solving, the first condition is optimistically added to the constraint system, then we check if the second icmp can be simplified, and finally the newly added entries are removed. The reason for doing things this way is to avoid clashes with signed <-> unsigned condition transfer, which require us to re-order facts to increase effectiveness. Reviewed By: nikic, antoniofrighetto Differential Revision: https://reviews.llvm.org/D151799
-
Tue Ly authored
-
Haojian Wu authored
-
Florian Hahn authored
This allows easier re-use of the checking logic. Split off from D151799.
-
Tue Ly authored
Clean up exhaustive tests. Let check functions return number of failures instead of passed/failed. Reviewed By: sivachandra Differential Revision: https://reviews.llvm.org/D153682
-