- Dec 08, 2021
-
-
Mehdi Amini authored
-
Vitaly Buka authored
kIgnored didn't exist when the code was added, but it should be equivalent to kReachable. The goal is to refactor MarkInvalidPCCb to avoid StackDepotGet in StopTheWorld.
-
Haohai Wen authored
i64 mul cost is 1cy for all cpu that support avx512. Currently all X86 cpu uses i64 mul cost in X64 cost table which is not true for cpu that support avx512 (skx, icx). Reviewed By: pengfei, RKSimon Differential Revision: https://reviews.llvm.org/D115016
-
Lang Hames authored
MaterializationUnit::Interface holds the values that make up the interface (for ORC's purposes) of a materialization unit: the symbol flags map and initializer symbol. Having a type for this will make functions that build materializer interfaces more readable and maintainable.
-
Chen Zheng authored
-
Alex Lorenz authored
This patch extends LLVM IR to add metadata that can be used to emit macho files with two build version load commands. It utilizes "darwin.target_variant.triple" and "darwin.target_variant.SDK Version" metadata names for that, which will be set by a future patch in clang. MachO uses two build version load commands to represent an object file / binary that is targeting both the macOS target, and the Mac Catalyst target. At runtime, a dynamic library that supports both targets can be loaded from either a native macOS or a Mac Catalyst app on a macOS system. We want to add support to this to upstream to LLVM to be able to build compiler-rt for both targets, to finish the complete support for the Mac Catalyst platform, which is right now targetable by upstream clang, but the compiler-rt bits aren't supported because of the lack of this multiple build version support. Differential Revision: https://reviews.llvm.org/D112189
-
Vitaly Buka authored
It was removed in D114934 and failed clang-thumbv7-full-2stage.
-
Mehdi Amini authored
The new form of printing attribute in the declarative assembly is eliding the `#dialect.mnemonic` prefix to only keep the `<....>` part. Differential Revision: https://reviews.llvm.org/D113873
-
Chen Zheng authored
Reviewed By: nemanjai Differential Revision: https://reviews.llvm.org/D114062
-
Nico Weber authored
This reverts commit bbccf499. 1ee32055 relanded in 5dda2efd.
-
Phoebe Wang authored
A follow up of D114162. Reviewed By: asavonic Differential Revision: https://reviews.llvm.org/D114782
-
Mircea Trofin authored
This reverts commit b2fbd45d. D114922 fixed the reason of the 2nd revert. This patch also re-applies 39e9f5d3. Differential Revision: https://reviews.llvm.org/D112012
-
Kirill Stoimenov authored
Reviewed By: kstoimenov Differential Revision: https://reviews.llvm.org/D115307
-
Kirill Stoimenov authored
Follow up from D115271. Reviewed By: vitalybuka Differential Revision: https://reviews.llvm.org/D115293
-
Jez Ng authored
Reviewed By: MaskRay Differential Revision: https://reviews.llvm.org/D115281
-
Michael Jones authored
In https://reviews.llvm.org/D114302 I renamed ForLoopCopy to for_loop_copy, but these got missed. Differential Revision: https://reviews.llvm.org/D115303
-
Petr Hosek authored
Some of the compiler-rt runtimes use custom instrumented libc++ build. Use the runtimes build for building this custom libc++. Differential Revision: https://reviews.llvm.org/D114922
-
Michael Jones authored
adds the .yaml files clang-tidy generates as byproducts, which means that they will be updated properly and cleaned by `ninja -t clean` Reviewed By: lntue Differential Revision: https://reviews.llvm.org/D115290
-
Aart Bik authored
also fixes two typos in IR doc Reviewed By: bixia Differential Revision: https://reviews.llvm.org/D115288
-
James Farrell authored
Revert "Revert "Use VersionTuple for parsing versions in Triple, fixing issues that caused the original change to be reverted. This makes it possible to distinguish between "16" and "16.0" after parsing, which previously was not possible."" This reverts commit 63a6348c. Differential Revision: https://reviews.llvm.org/D115254
-
Michael Berg authored
This reverts commit 8487981a.
-
Yaxun (Sam) Liu authored
nvcc defines __CUDACC_RDC__ for both host and device compilation when -rdc=true is specified (https://docs.nvidia.com/cuda/cuda-compiler-driver-nvcc/index.html#nvcc-identification-macro) This patch defines __CLANG_RDC__ when -fgpu-rdc is specified for CUDA/HIP. Reviewed by: Artem Belevich Differential Revision: https://reviews.llvm.org/D114812
-
Jim Ingham authored
Differential Revision: https://reviews.llvm.org/D115017
-
Michael Berg authored
Both these preference helper functions have initial support with this change. The loop unrolling preferences are set with initial settings to control thresholds, size and attributes of loops to unroll with some tuning done. The peeling preferences may need some tuning as well as the initial support looks much like what other architectures utilize. Reviewed By: craig.topper Differential Revision: https://reviews.llvm.org/D113798
-
Carlo Bertolli authored
Prepare amdgpu plugin for asynchronous implementation. This patch switches to using HSA API for asynchronous memory copy. Moving away from hsa_memory_copy means that plugin is responsible for locking/unlocking host memory pointers. Reviewed By: JonChesterfield Differential Revision: https://reviews.llvm.org/D115279
-
AlexisPerry authored
This patch is part of the upstreaming effort from the fir-dev branch. Address review comments - move CHECK blocks to after the mlir code in the test file - fix style with respect to anonymous namespaces: only include class definitions in the namespace and make functions static and outside the namespace - fix a few nits - remove TODO in favor of notifyMatchFailure - removed unnecessary CHECK line from convert-to-llvm.fir - rebase on main - add TODO back in - get successfull test of TODO in AllocMemOp converion of derived type with LEN params - clearer comments and reduced use of auto - move defintion of computeDerivedTypeSize to fix build error Co-authored-by:
Eric Schweitz <eschweitz@nvidia.com> Co-authored-by:
Jean Perier <jperier@nvidia.com> Reviewed By: awarzynski, clementval, kiranchandramohan, schweitz Differential Revision: https://reviews.llvm.org/D114104
-
Kirill Stoimenov authored
This is additional cleanup as follow-up of D104279. Reviewed By: vitalybuka Differential Revision: https://reviews.llvm.org/D115271
-
Martin Storsjö authored
This does mostly the same as D112126, but for the runtimes cmake files. Most of that is straightforward, but the interdependency between libcxx and libunwind is tricky: Libunwind is built at the same time as libcxx, but libunwind is not installed yet. LIBCXXABI_USE_LLVM_UNWINDER makes libcxx link directly against the just-built libunwind, but the compiler implicit -lunwind isn't found. This patch avoids that by adding --unwindlib=none if supported, if we are going to link explicitly against a newly built unwinder anyway. Reapplying this after db32c4f4, which should fix the issues that were reported last time this was applied. Differential Revision: https://reviews.llvm.org/D113253
-
Kiran Chandramohan authored
The XRebox Op is formed by the codegen rewrite which makes it easier to convert the operation to LLVM. The XRebox op includes the information from the rebox op and the associated slice, shift, and shape ops. During the conversion process a new descriptor is created for reboxing. Co-authored-by:
Jean Perier <jperier@nvidia.com> Co-authored-by:
Eric Schweitz <eschweitz@nvidia.com> Co-authored-by:
Val Donaldson <vdonaldson@nvidia.com> Reviewed By: clementval Differential Revision: https://reviews.llvm.org/D114709
-
Kirill Stoimenov authored
This reverts commit e4800fc0. Reviewed By: kstoimenov Differential Revision: https://reviews.llvm.org/D115286
-
Kirill Stoimenov authored
This is additional cleanup as follow-up of D104279. Reviewed By: vitalybuka Differential Revision: https://reviews.llvm.org/D115271
-
Nikolas Klauser authored
clang has `= default` as an extension in c++03, so just use it. Reviewed By: ldionne, Quuxplusone, #libc Spies: libcxx-commits Differential Revision: https://reviews.llvm.org/D115275
-
LLVM GN Syncbot authored
-
Jonas Devlieghere authored
This reverts commit 0fe61ecc because it breaks the modules build. https://green.lab.llvm.org/green/job/clang-stage2-rthinlto/4858/ https://green.lab.llvm.org/green/view/LLDB/job/lldb-cmake/39112/
-
Carlo Bertolli authored
At present, amdgpu plugin merges both asynchronous and synchronous kernel launch implementations into a single synchronous version. This patch prepares the plugin for asynchronous implementation by: - Privatizing actual kernel launch code (valid in both cases) into an anonymous namespace base function Actual separation of kernel launch code (async vs sync) is a following patch. Reviewed By: JonChesterfield Differential Revision: https://reviews.llvm.org/D115267
-
Sanjay Patel authored
We avoid this fold in the more general cases where we use `FoldOpIntoSelect`. That's because -- unlike most binary opcodes -- 'rem' can't usually be speculated with a variable divisor since it can have immediate UB. But in the case where both arms of the select are constants, we can safely evaluate both sides and eliminate 'rem' completely. This should fix: https://llvm.org/PR52102 The same optimization for 'div' is planned as a follow-up patch. Differential Revision: https://reviews.llvm.org/D115173
-
Sanjay Patel authored
More coverage for D115173 and its div sibling. These are existing transforms handled by simplifyDivRemOfSelectWithZeroOp().
-
Sanjay Patel authored
Similar patterns as tests for: D115173 (and we might want to make a single patch for all of these).
-
Martin Storsjö authored
In the "runtimes" setup, the runtime (e.g. OpenMP) can be built for a target entirely different from the current host build (where LLVM and Clang are built). If profiling is enabled, libomptarget links against LLVMSupport (which only has been built for the host). Thus, don't enable profiling by default in this setup. This should allow relanding D113253. Differential Revision: https://reviews.llvm.org/D114083
-
Valentin Clement authored
This patch adds the runtime function to allocate and deallocate ragged arrays. This patch is part of the upstreaming effort from fir-dev branch. Reviewed By: klausler Differential Revision: https://reviews.llvm.org/D114534 Co-authored-by:
Eric Schweitz <eschweitz@nvidia.com>
-