- Sep 27, 2022
-
-
WANG Xuerui authored
Differential Revision: https://reviews.llvm.org/D134582
-
Guray Ozen authored
This revision refactors gpu block id generator lambda that is used in the transform dialect. It removes the lambda and instead uses a static function that's name generateGpuBlockIds. It also simplifies arguments that the function takes. Reviewed By: nicolasvasilache Differential Revision: https://reviews.llvm.org/D134724
-
Alvin Wong authored
When running `target module lookup` command, show the name of absolute symbols. Also fix indentation issue after printing an absolute symbol. Reviewed By: clayborg, DavidSpickett Differential Revision: https://reviews.llvm.org/D134516
-
Alvin Wong authored
This adds a line break between each result address in the output of the lldb command `target modules lookup`. Before this change, a new address result will be printed on the same line as the summary of the last result, making the output difficult to view. Also adds a test for this command. Reviewed By: labath Differential Revision: https://reviews.llvm.org/D134111
-
Simon Pilgrim authored
Simplifies the diff for D134697
-
Florian Hahn authored
The expanded values when rewriting exit values need to preserve LCSSA. Ask SCEVExpander to preserve LCSSA to ensure that. Fixes #58007.
-
Nikita Popov authored
-
Daniel Kiss authored
A given function is compatible with all previous arch versions. To avoid compering values of the attribute this logic adds all predecessor architecture values. Reviewed By: dmgreen, DavidSpickett Differential Revision: https://reviews.llvm.org/D134353
-
Nikita Popov authored
This updates checkFunctionMemoryAccess() to infer a precise FunctionModRefBehavior, rather than an approximation split into read/write and argmemonly. Afterwards, we still map this back to imprecise function attributes. This still allows us to infer some cases that we previously did not handle, namely inaccessiblememonly and inaccessiblemem_or_argmemonly. In practice, this means we get better memory attributes in the presence of intrinsics like @llvm.assume. Differential Revision: https://reviews.llvm.org/D134527
-
Nikita Popov authored
Currently if both LIBCXX_STATICALLY_LINK_ABI_IN_STATIC_LIBRARY (only applies to static library) and LIBCXX_ENABLE_ABI_LINKER_SCRIPT (only applies to shared library) are enabled, the former will be silently ignored. The shared library will use a linker script, while the static library fails to link libc++abi.a entirely. This is caused by what appears to be an implementation bug: The LIBCXX_STATICALLY_LINK_ABI_IN_*_LIBRARY options are declared as dependent options of LIBCXX_ENABLE_STATIC_ABI_LIBRARY, rather than simply using it as the default value. Of course, the combination of LIBCXX_STATICALLY_LINK_ABI_IN_SHARED_LIBRARY and LIBCXX_ENABLE_ABI_LINKER_SCRIPT still results in a cmake error, because these would be conflicting requests for the shared library. Differential Revision: https://reviews.llvm.org/D134644
-
David Sherwood authored
Adds support for the Neoverse V2 CPU to the AArch64 backend. Differential Revision: https://reviews.llvm.org/D134352
-
Florian Hahn authored
After unrolling a loop, the block and loop dispositions need to be cleared. As we don't know which SCEVs in the loop/blocks may be impacted, completely clear the cache. This should also fix some cases where deleted loops remained in the LoopDispositions cache. This fixes a verification failure surfaced by D134531. I am planning on reviewing/updating the existing uses of forgetLoopDispositions to check if they should be replaced by forgetBlockAndLoopDispositions. Reviewed By: nikic Differential Revision: https://reviews.llvm.org/D134612
-
LLVM GN Syncbot authored
-
Carlos Alberto Enciso authored
It allows finding all intervals that overlap with any given point. At this time, it does not support any deletion or rebalancing operations. The IntervalTree is designed to be set up once, and then queried without any further additions. Reviewed By: psamolysov, probinson Differential Revision: https://reviews.llvm.org/D125776
-
Paulo Matos authored
Initial table.get/set implementation would match and lower combinations of GEP+load/store to table.get/set instructions. However, this is error prone due to potential combinations of GEP+load/store we don't implement, and load/store optimizations. By changing the code to using intrinsics, we avoid both issues and simplify the code. New builtins implemented: * @llvm.wasm.table.get.externref * @llvm.wasm.table.get.funcref * @llvm.wasm.table.set.externref * @llvm.wasm.table.set.funcref Reviewed By: asb, tlively Differential Revision: https://reviews.llvm.org/D134436
-
Fraser Cormack authored
-
Vitaly Buka authored
Test sometimes fail with: https://lab.llvm.org/buildbot/#/builders/168/builds/9110
-
Yeting Kuo authored
Add vp.maxnum and vp.minnum which are vector predicted intrinsics of llvm.maxnum and llvm.minnum. Reviewed By: craig.topper Differential Revision: https://reviews.llvm.org/D134639
-
Vitaly Buka authored
Break msan bots. Details in D134666. This reverts commit 0ce96e06.
-
Katherine Rasmussen authored
Write a semantics test for the atomic intrinsic subroutine, atomic_define. Reviewed By: rouson Differential Revision: https://reviews.llvm.org/D134288
-
jacquesguan authored
Reviewed By: craig.topper Differential Revision: https://reviews.llvm.org/D134631
-
Paul Scoropan authored
This feature implements support for making entries in the exception section on XCOFF on the direct assembly path using the ".except" pseudo-op. It also provides functionality to lower entries (comprised of language and reason codes) into the exception section through the use of annotation metadata attached to llvm.ppc.trap/trapd/tw/tdw intrinsics. Integrated assembler support will be provided in another review. https://reviews.llvm.org/D133030 needs to merge first for LIT tests Reviewed By: shchenz, RKSimon Differential Revision: https://reviews.llvm.org/D132146
-
Yaxun (Sam) Liu authored
HIP is able to unbundle archive of bundled bitcode. However currently there are two bugs: 1. archives passed by -l: are not unbundled. 2. archives passed as input files are not unbundled The actual file name of an archive passed by -l: should not be prefixed with lib and appended with '.a', but the file path is prefixed with paths in '-L' options. The actual file name of an archive passed as an input file stays the same, not affected by the '-L' options.
-
Sebastian Peryt authored
Second patch in the series to remove legacy PM and associated -enable-new-pm=0 flag targets pass that has not been ported to new PM - PruneEH. Discussion about this can be found in D44415. Reviewed By: aeubanks Differential Revision: https://reviews.llvm.org/D134686
-
Tue Ly authored
Simplify `tanf` implementation and improve its performance. Completely reuse the implementation of `sinf`, `cosf`, `sincosf` and use the definition `tan(x) = sin(x)/cos(x)`. Performance benchmark using perf tool from the CORE-MATH project on Ryzen 1700: ``` $ CORE_MATH_PERF_MODE="rdtsc" ./perf.sh tanf GNU libc version: 2.35 GNU libc release: stable CORE-MATH reciprocal throughput : 18.558 System LIBC reciprocal throughput : 49.919 BEFORE: LIBC reciprocal throughput : 36.480 LIBC reciprocal throughput : 27.217 (with `-msse4.2` flag) LIBC reciprocal throughput : 20.205 (with `-mfma` flag) AFTER: LIBC reciprocal throughput : 30.337 LIBC reciprocal throughput : 21.072 (with `-msse4.2` flag) LIBC reciprocal throughput : 15.804 (with `-mfma` flag) $ CORE_MATH_PERF_MODE="rdtsc" ./perf.sh tanf --latency GNU libc version: 2.35 GNU libc release: stable CORE-MATH latency : 56.702 Syst...
-
Yuanfang Chen authored
And correct a few typos.
-
Xiang Li authored
-
Vitaly Buka authored
-
Vitaly Buka authored
Supports either default of -f[no-]sanitize-memory-param-retval. Reviewed By: aeubanks, MaskRay Differential Revision: https://reviews.llvm.org/D134683
-
Adrian Prantl authored
This reverts commit c1ce19021da0cf1c88722024e6ff9cee7aabc7b6.
-
Fangrui Song authored
This is the AArch64 counterpart of D134127. Daniel Kiss will change more `BUILTIN` to `TARGET_BUILTIN`. Fix #57802
-
Alex Yermolovich authored
There is an implicit circular dependency in a debug information coming from GCC. This results in a coredump. It tries to resolve scope for DIE TAGs that are scopeless. Reviewed By: dblaikie Differential Revision: https://reviews.llvm.org/D132900
-
Tony Tye authored
Correct naming of DW_OP_LLVM_aspace_bregx in AMDGPUDwarfExtensionsForHeterogeneousDebugging.rst . Reviewed By: dblaikie Differential Revision: https://reviews.llvm.org/D134467
-
Aart Bik authored
Removed some whitespace, improved some of the ops doc. Reviewed By: wrengr Differential Revision: https://reviews.llvm.org/D134672
-
Mircea Trofin authored
Differential Revision: https://reviews.llvm.org/D134664
-
Min-Yih Hsu authored
Add support for translating AtomicRMWInst and AtomicCmpXchgInst from LLVM IR. Differential Revision: https://reviews.llvm.org/D134450
-
Craig Topper authored
I've left the getAggregateElement as a fast path for non-ConstantExprs to avoid a call to getSplatValue in release builds. Fixes PR57989. Reviewed By: spatel Differential Revision: https://reviews.llvm.org/D134670
-
Akira Hatanaka authored
checkArrayElementAlignment in Sema::BuildCXXNew This commit fixes a bug that was introduced by adaf62ce and reported here: https://reviews.llvm.org/D133711#3814717
-
Fangrui Song authored
For RVC, GNU assembler and LLVM integrated assembler add c.nop followed by a sequence of 4-byte nops. Even if remove % 4 == 0, we have to split one 4-byte nop and therefore need to write the code sequence, otherwise we create an incorrect c.unimp.
-
Florian Hahn authored
-