- Sep 14, 2023
-
-
Stefan Gränitz authored
We don't have to set or clear the Thumb bit in relocation fixup values. It's not part of the branch range and the respective encoding functions like encodeImmBT4BlT1BlxT2() shift out the least significant bit anyway. This was a leftover from the initial patch before we switched to store Thumb state in target-flags with D146641.
-
Reid Kleckner authored
This reverts commit 2ca4d136. Also revert the followup, "[InlineAsm] fix botched merge conflict resolution" This reverts commit 8b9bf3a9. There were SystemZ and Mips build errors, too many to fix forward.
-
Luke Lau authored
-
Louis Dionne authored
This partially reverts commit e30a148b, which removed the base template for std::char_traits. That base template had been marked as deprecated since LLVM 16 and we were planning to remove it in LLVM 18. However, as explained in the post-commit comments in https://reviews.llvm.org/D157058, the deprecation mechanism didn't work as expected. Basically, the deprecation warnings were never shown to users since libc++ headers are system headers and Clang doesn't show warnings in system headers. As a result, this removal came with basically no lead time as far as users are concerned, which is a poor experience. For this reason, I am re-introducing the deprecated char_traits specialization until we have a proper way of phasing it out in a way that is not a surprise for users.
-
Philip Reames authored
A splat index means the operation is reading from (writing to) the same memory location. Generally, zero is the cheapest value to splat. As such, we'd prefer to add the splatted value to the base, and use a constant zero as the index operand.
-
Simon Pilgrim authored
If we know the condition mask sign bit of the demanded elements then we can directly return the LHS/RHS selection
-
frgossen authored
-
Nick Desaulniers authored
Link: #66003
-
Jonas Devlieghere authored
This file is heavily trafficked and the various formatting and whitespace issues make it tedious to work on.
-
Krzysztof Drewniak authored
Currently, the VectorToLLVM patterns are built into a library along with the corresponding pass, which also pulls in all the platform-specific vector dialects (like AMXDialect) to apply all the vector to LLVM conversions. This causes dependency bloat when writing libraries - for example the GPU to LLVM passes, which use the vector to LLVM patterns, don't need the X86Vector dialect to be present at all. This commit partitions the library into VectorToLLVM and VectorToLLVMPass, where the latter pulls in all the other vector transformations. Reviewed By: nicolasvasilache, mehdi_amini Differential Revision: https://reviews.llvm.org/D158287
-
Vitaly Buka authored
Not sure if this is real leak of lsan and clang-repl incompatibility. ca8d2533 causes leak report in ClangReplInterpreterExceptionTests.
-
Krzysztof Drewniak authored
Reviewed By: fmorac Differential Revision: https://reviews.llvm.org/D152299
-
- Sep 13, 2023
-
-
Joshua Cranmer authored
Reviewed By: arsenm, aeubanks Differential Revision: https://reviews.llvm.org/D127221
-
David Spickett authored
I missed this before landing.
-
Mircea Trofin authored
The tests check against an expected output file - trivial change. Reference: either D158068 or e6b85c30
-
Nick Desaulniers authored
Similar to commit 2fad6e69 ("[InlineAsm] wrap Kind in enum class NFC") Fix the TODOs added in commit 93bd4287 ("[InlineAsm] refactor InlineAsm class NFC (#65649)")
-
Yinying Li authored
lvlTypes = [ "compressed" ] to map = (d0) -> (d0 : compressed) lvlTypes = [ "dense" ] to map = (d0) -> (d0 : dense)
-
Siva Chandra authored
This is part of a libc wide CMake cleanup which aims to eliminate certain explicitly duplicated logic which is available in CMake-3.20. This change in particular makes the entrypoint aliases real library targets so that they can be treated as normal library targets by other libc build rules.
-
Eymen Ünay authored
Added support and test for ELF::R_ARM_JUMP24 Reviewed By: sgraenitz Differential Revision: https://reviews.llvm.org/D157541
-
Eymen Ünay authored
- Added WritableArmRelocation and ArmRelocation Structs - Encode/Decode funcs for B/BL A1 and BLX A2 encodings - Add ARM helper functions, consistent with the existing Thumb helper functions - Add Test for ELF::R_ARM_CALL Reviewed By: sgraenitz Differential Revision: https://reviews.llvm.org/D157533
-
Timm Bäder authored
-
sstwcw authored
Dictionary literal keys and strings in TypeScript type declarations can not be broken. The problem was pointed out by @alexfh and @e-kud here: https://reviews.llvm.org/D154093#4644512
-
lorenzo chelini authored
Currently, the compiler crashes if the number of tiles provided exceeds the number of loops.
-
Zequan Wu authored
This test doesn't explicitly require lld but clang's flag -fuse-ld=lld uses lld. This fixes the error that clang cannot invoke lld.
-
CarolineConcatto authored
The new ACLE PR#225[1] now combines the slice parameters for some builtins. Slice specifies the ZA slice number directly and needs to be explicity implemented by the "user" with the base register plus the immediate offset [1]https://github.com/ARM-software/acle/pull/225/files
-
Martin Erhart authored
Revert "[mlir][bufferization] Remove allow-return-allocs and create-deallocs pass options, remove bufferization.escape attribute" This reverts commit 6a91dfed. This caused problems in downstream projects. We are reverting to give them more time for integration.
-
Martin Erhart authored
Revert "[mlir][bufferization] Update linalg integration tests to lower ops created by bufferization-to-memref pass" This reverts commit 6f35401f. This caused problems in downstream projects. We are reverting to give them more time for integration.
-
Martin Erhart authored
This reverts commit 1bebb60a. This caused problems in downstream projects. We are reverting to give them more time for integration.
-
Martin Erhart authored
This reverts commit 29d86175. This caused problems in downstream projects. We are reverting to give them more time for integration.
-
Martin Erhart authored
This reverts commit 89117f18. This caused problems in downstream projects. We are reverting to give them more time for integration.
-
Martin Erhart authored
This reverts commit 1356e853. This caused problems in downstream projects. We are reverting to give them more time for integration.
-
Martin Erhart authored
This reverts commit f0c46639. This caused problems in downstream projects. We are reverting to give them more time for integration.
-
Martin Erhart authored
This reverts commit cb5fe6ce. This caused problems in downstream projects. We are reverting to give them more time for integration.
-
Mikhail R. Gadelha authored
This patch changes the size of time_t to be an int64_t. This still follows the POSIX standard which only requires time_t to be an integer. Making time_t a 64-bit integer also fixes two cases in 32 bits platforms that use SYS_clock_nanosleep_time64 and SYS_clock_gettime64, as the name of these calls implies, they require a 64-bit time_t. For instance, in rv32, the 32-bit version of these syscalls is not available. We also follow glibc here, where time_t is still a 32-bit integer in arm32. Reviewed By: sivachandra Differential Revision: https://reviews.llvm.org/D159125
-
Joseph Huber authored
Summary: Currently, there is an assertion that prevents us from emitting an AMDGPU global with a non-target specific address space (i.e. numerical attribute). I'm unsure what the original intentions of this assertion were, but we should be able to use OpenCL address spaces when compiling directly to AMDGPU from C++. This is permitted on NVPTX so I'm unsure what this assertion is guarding. The patch simply removes the assertion and adds a test to ensure that these emit the expected address spaces. Fixes https://github.com/llvm/llvm-project/issues/65069
-
Guray Ozen authored
clang was used for local testing. The PR changes it to `mlir-cpu-runner`
-
Joseph Huber authored
Summary: We use the `llvm.amgcn.abi.version` varaible to control code generation. This is emitted in every module now to indicate what should be used when compiling. Previously, the logic caused us to emit an external reference to this variable when creating the code for the `none` type. This would then cause us not to emit the actual definition. This patch refines the logic to create the external reference, and then update it if it is found unset by the time we emit the global. I had to remove the reference to `GetOrCreateLLVmGlobal` because it did not accept the proper address space.
-
Guray Ozen authored
An integration test for the 128b Swizzling TMA. TMA with 128B Swizzle loads data as follows (each numbered cell is 16 bytes). The program tests this pattern for `128x64xf16` type. ``` |-------------------------------| | 0 | 1 | 2 | 3 | 4 | 5 | 6 | 7 | | 1 | 0 | 3 | 2 | 5 | 4 | 7 | 6 | | 2 | 3 | 0 | 1 | 6 | 7 | 4 | 5 | | 3 | 2 | 1 | 0 | 7 | 6 | 5 | 4 | | 4 | 5 | 6 | 7 | 0 | 1 | 2 | 3 | | 5 | 4 | 7 | 6 | 1 | 0 | 3 | 2 | | 6 | 7 | 4 | 5 | 2 | 3 | 0 | 1 | |-------------------------------| | ... pattern repeats ... | |-------------------------------| ```
-
martinboehme authored
Otherwise, the test doesn't actually do anything.
-
Guray Ozen authored
The register number of predicate is calculated incorrectly. This PR fixes that.
-