- Sep 14, 2023
-
-
Louis Dionne authored
Thanks @Gadal for noticing. Fixes #63703
-
Maryam Moghadas authored
This patch reverts commit 7614ba0a to optimize VPERM when one of its vector operands is XXSWAPD, similar to XXPERM. It also reorganizes the little-endian swap code on LE, swapping the vector operand after adjusting the mask operand. This ensures that the vector operand is swapped at the correct point in the code, resulting in a valid constant pool for the mask operand. Reviewed By: stefanp Differential Revision: https://reviews.llvm.org/D149083
-
Jakub Kuderski authored
This is a cleanup in preparation for adding a second conversion path using the KHR cooperative matrix extension. Make the existing lowering explicit about emitting ops from the NV coop matrix extension. Clean up surrounding code.
-
Joseph Huber authored
Summary: This comment is outddated and can be removed, also mention an option for limiting parallelism during tests in the documentation.
-
Jordan Rupprecht authored
These were added in 2011 in 54da06ca and 0d901121. It was then incompletely removed in 2015 in c268835e: the video itself was removed, but the VideoJS css & script is still loaded.
-
Paul T Robinson authored
-
Terry Wilmarth authored
Fix /tmp approach, and add environment variable method as third fallback during library registration The /tmp fallback for /dev/shm did not write to a fixed filename, so multiple instances of the runtime would not be able to detect each other. Now, we create the /tmp file in much the same way as the /dev/shm file was created, since mkstemp approach would not work to create a file that other instances of the runtime would detect. Also, add the environment variable method as a third fallback to /dev/shm and /tmp for library registration, as some systems do not have either. Also, add ability to fallback to a subsequent method should a failure occur during any part of the registration process. When unregistering, it is assumed that the method chosen during registration should work, so errors at that point are ignored. This also avoids a problem with multiple threads trying to unregister the library.
-
Peiming Liu authored
Many frontends canonicalize scalar into 0-ranked tensor, it change will hopefully make the operation easier to use for those cases.
-
Jeffrey Byrnes authored
Adds test to show impact on cross block CopyToReg & CopyFromReg handling for n x i8, and shows NFC on CC Differential Revision: https://reviews.llvm.org/D159303 Change-Id: Ib6d9802dbebe8e3245e4ccfd4a6f23357de8c480
-
Justin Bogner authored
This adds more validation that a dxil triple is actually useable when compiling HLSL. The OS field of the triple needs to be a versioned shader model. Later, we should set a default if this is empty and check that the version is a shader model we can actually handle. The Environment field of the triple needs to be specified and be a valid shader stage. I'd like to allow this to be empty and treat it like library, but allowing that currently crashes in DXIL metadata handling. Differential Revision: https://reviews.llvm.org/D159103
-
Vitaly Buka authored
-
michaelrj-google authored
This patch adds the long double table option for printf into the new configuration scheme. This allows it to be set for most targets but unset for baremetal.
-
chbessonova authored
Use symbol's flags saved in `NMSymbol::SymFlags` inside `NMSymbol::isDefined()` since `BasicSymbolRef::getFlags()` requires the symbol's containing entity object to exist (which doesn't, causing llvm-nm to crash).
-
Philip Reames authored
No need to special case add 0, N. SelectionDAG::getNode contains the canonicalization and simplification for this case, so no need to duplicate it here.
-
Vitaly Buka authored
LLVMSymbolizer::DemangleName returns the same input if it can't demangle. We can't tell if this is already demangled or format is unsupported. Internally DemangleName uses nonMicrosoftDemangle which can report a failure.
-
Yitzhak Mandelbaum authored
[clang][dataflow] Change `diagnoseFunction` to use `llvm::SmallVector` instead of `std::vector`. (#66014) The template is agnostic as to the type used by the list, as long as it is compatible with `llvm::move` and `std::back_inserter`. In practice, we've encountered analyses which use different types (`llvm::SmallVector` vs `std::vector`), so it seems preferable to leave this open to the caller.
-
Joseph Huber authored
Summary: This patch simply adds the necessary config to enable qsort and bsearch on the GPU. It is *highly* unlikely that anyone will use these, as they are single threaded, but we may as well support all entrypoints that we can.
-
Stefan Gränitz authored
-
Stefan Gränitz authored
-
Stefan Gränitz authored
We don't have to set or clear the Thumb bit in relocation fixup values. It's not part of the branch range and the respective encoding functions like encodeImmBT4BlT1BlxT2() shift out the least significant bit anyway. This was a leftover from the initial patch before we switched to store Thumb state in target-flags with D146641.
-
Reid Kleckner authored
This reverts commit 2ca4d136. Also revert the followup, "[InlineAsm] fix botched merge conflict resolution" This reverts commit 8b9bf3a9. There were SystemZ and Mips build errors, too many to fix forward.
-
Luke Lau authored
-
Louis Dionne authored
This partially reverts commit e30a148b, which removed the base template for std::char_traits. That base template had been marked as deprecated since LLVM 16 and we were planning to remove it in LLVM 18. However, as explained in the post-commit comments in https://reviews.llvm.org/D157058, the deprecation mechanism didn't work as expected. Basically, the deprecation warnings were never shown to users since libc++ headers are system headers and Clang doesn't show warnings in system headers. As a result, this removal came with basically no lead time as far as users are concerned, which is a poor experience. For this reason, I am re-introducing the deprecated char_traits specialization until we have a proper way of phasing it out in a way that is not a surprise for users.
-
Philip Reames authored
A splat index means the operation is reading from (writing to) the same memory location. Generally, zero is the cheapest value to splat. As such, we'd prefer to add the splatted value to the base, and use a constant zero as the index operand.
-
Simon Pilgrim authored
If we know the condition mask sign bit of the demanded elements then we can directly return the LHS/RHS selection
-
frgossen authored
-
Nick Desaulniers authored
Link: #66003
-
Jonas Devlieghere authored
This file is heavily trafficked and the various formatting and whitespace issues make it tedious to work on.
-
Krzysztof Drewniak authored
Currently, the VectorToLLVM patterns are built into a library along with the corresponding pass, which also pulls in all the platform-specific vector dialects (like AMXDialect) to apply all the vector to LLVM conversions. This causes dependency bloat when writing libraries - for example the GPU to LLVM passes, which use the vector to LLVM patterns, don't need the X86Vector dialect to be present at all. This commit partitions the library into VectorToLLVM and VectorToLLVMPass, where the latter pulls in all the other vector transformations. Reviewed By: nicolasvasilache, mehdi_amini Differential Revision: https://reviews.llvm.org/D158287
-
Vitaly Buka authored
Not sure if this is real leak of lsan and clang-repl incompatibility. ca8d2533 causes leak report in ClangReplInterpreterExceptionTests.
-
Krzysztof Drewniak authored
Reviewed By: fmorac Differential Revision: https://reviews.llvm.org/D152299
-
- Sep 13, 2023
-
-
Joshua Cranmer authored
Reviewed By: arsenm, aeubanks Differential Revision: https://reviews.llvm.org/D127221
-
David Spickett authored
I missed this before landing.
-
Mircea Trofin authored
The tests check against an expected output file - trivial change. Reference: either D158068 or e6b85c30
-
Nick Desaulniers authored
Similar to commit 2fad6e69 ("[InlineAsm] wrap Kind in enum class NFC") Fix the TODOs added in commit 93bd4287 ("[InlineAsm] refactor InlineAsm class NFC (#65649)")
-
Yinying Li authored
lvlTypes = [ "compressed" ] to map = (d0) -> (d0 : compressed) lvlTypes = [ "dense" ] to map = (d0) -> (d0 : dense)
-
Siva Chandra authored
This is part of a libc wide CMake cleanup which aims to eliminate certain explicitly duplicated logic which is available in CMake-3.20. This change in particular makes the entrypoint aliases real library targets so that they can be treated as normal library targets by other libc build rules.
-
Eymen Ünay authored
Added support and test for ELF::R_ARM_JUMP24 Reviewed By: sgraenitz Differential Revision: https://reviews.llvm.org/D157541
-
Eymen Ünay authored
- Added WritableArmRelocation and ArmRelocation Structs - Encode/Decode funcs for B/BL A1 and BLX A2 encodings - Add ARM helper functions, consistent with the existing Thumb helper functions - Add Test for ELF::R_ARM_CALL Reviewed By: sgraenitz Differential Revision: https://reviews.llvm.org/D157533
-
Timm Bäder authored
-