- Feb 12, 2022
-
-
Min-Yih Hsu authored
This patch refactors all the existing M68k arithmetic instructions to use the new VarLenCodeEmitterGen infrastructure. This patch is tested by the existing MC test cases. Note that one of the codegen tests needed to be updated because the ordering of two equivalent instructions were switched. Differential Revision: https://reviews.llvm.org/D115234
-
Min-Yih Hsu authored
Full write up: https://gist.github.com/mshockwave/66e98d099256deefc062633909bb7b5b The existing CodeEmitterGen infrastructure is unable to generate encoder function for ISAs with variable-length instructions. This patch introduces a new infrastructure to support variable-length instruction encoding, including a new TableGen syntax for writing instruction encoding directives and a new TableGen backend component, VarLenCodeEmitterGen, built on top of CodeEmitterGen. Differential Revision: https://reviews.llvm.org/D115128
-
Jay Foad authored
Remove ARTIFICIAL_VGPR which only existed to make VReg_1 not empty. Differential Revision: https://reviews.llvm.org/D119552
-
Joe Loser authored
`ranges_swap_ranges.h` includes `<type_traits>` but does not use anything from it. So, remove the include. Differential Revision: https://reviews.llvm.org/D119491
-
Sebastian Neubauer authored
Use a subtarget feature instead of a command line argument to reduce global state. We want to enable flat scratch for graphics in some cases and this doesn't work well with command line options. Differential Revision: https://reviews.llvm.org/D119425
-
Sameer Sahasrabuddhe authored
The module flag to indicate use of hostcall is insufficient to catch all cases where hostcall might be in use by a kernel. This is now replaced by a function attribute that gets propagated to top-level kernel functions via their respective call-graph. If the attribute "amdgpu-no-hostcall-ptr" is absent on a kernel, the default behaviour is to emit kernel metadata indicating that the kernel uses the hostcall buffer pointer passed as an implicit argument. The attribute may be placed explicitly by the user, or inferred by the AMDGPU attributor by examining the call-graph. The attribute is inferred only if the function is not being sanitized, and the implictarg_ptr does not result in a load of any byte in the hostcall pointer argument. Reviewed By: jdoerfert, arsenm, kpyzhov Differential Revision: https://reviews.llvm.org/D119216
-
Julien Pages authored
Add a new llvm.fptrunc.round intrinsic to precisely control the rounding mode when converting from f32 to f16. Differential Revision: https://reviews.llvm.org/D110579
-
Florian Hahn authored
Test from https://github.com/llvm/llvm-project/issues/48253.
-
Hongtao Yu authored
When generating nested CS profile with all calling contexts of a function duplicated into a base profile under `--generate-merged-base-profiles`, do not recount callee samples when computing profile summary. This fixes the profile summary mismatch between flat cs profile and nested cs profile, for both extbinary and text format. Reviewed By: wenlei Differential Revision: https://reviews.llvm.org/D119494
-
Geoffrey Martin-Noble authored
-
Kai Nacke authored
The XPLINK return `b 2(7)` has size 4 bytes, while the Linux return `br 7` only has size 2 bytes. Thus a new alias is required to have correct instruction byte count. It also fixes the conditional return code. Reviewed By: uweigand Differential Revision: https://reviews.llvm.org/D119437
-
Simon Pilgrim authored
Avoid the need for a forward declaration. Cleanup prep for Issue #53760
-
Mark de Wever authored
Reviewed By: #libc, ldionne Differential Revision: https://reviews.llvm.org/D119350
-
Mark de Wever authored
Reviewed By: #libc, ldionne Differential Revision: https://reviews.llvm.org/D119349
-
Simon Pilgrim authored
Replace the *_EXTEND node with the raw operands, this will make it easier to use combineToExtendBoolVectorInReg for any boolvec extension combine. Cleanup prep for Issue #53760
-
Mark de Wever authored
This avoids using an libc++ internal macro in our tests. This version doesn't depend on the internal macro but redefines it. Reviewed By: #libc, ldionne Differential Revision: https://reviews.llvm.org/D119460
-
LLVM GN Syncbot authored
-
Nikolas Klauser authored
Implement ranges::min_element Reviewed By: Quuxplusone, Mordante, #libc Spies: miscco, libcxx-commits, mgorny Differential Revision: https://reviews.llvm.org/D117025
-
LLVM GN Syncbot authored
-
Nikita Popov authored
-
Nikolas Klauser authored
Add `ranges::in_fun_result` Reviewed By: Quuxplusone, #libc, var-const Spies: CaseyCarter, var-const, libcxx-commits, mgorny Differential Revision: https://reviews.llvm.org/D116974
-
- Feb 11, 2022
-
-
OCHyams authored
Dexter saves various files to a new results directory each time it is run (including when it's run by lit tests) and there isn't a way to opt-out. This patch reconfigures the behaviour to be opt-in by removing the default `--results-directory` location. Now results are only saved if `--results-directory` is specified. Reviewed By: jmorse Differential Revision: https://reviews.llvm.org/D119545
-
Matt Arsenault authored
If we had some source value we could infer an address space from that went through a ptrtoint/inttoptr pair, this would fail since bitcast can't change the address space. Fixes issue 53665.
-
Nikita Popov authored
It's not the same GEP if the source element type is different.
-
Simon Pilgrim authored
All paths have already dereferenced the CodeCompleter pointer in the ResultBuilder constructor
-
Simon Pilgrim authored
All paths have already dereferenced the Prev pointer
-
Simon Pilgrim authored
All paths have already dereferenced the block pointer
-
Jez Ng authored
... to use hyphens instead of underscores, making it consistent with our other substitutions like %no-arg-lld and %lld-watchos. Reviewed By: keith Differential Revision: https://reviews.llvm.org/D119513
-
Dávid Bolvanský authored
Discussed here: https://reviews.llvm.org/D119061#3310822 Reviewed By: aaron.ballman Differential Revision: https://reviews.llvm.org/D119451
-
Anton Zabaznov authored
OpenCL C 3.0 __opencl_c_subgroups feature is slightly different then other equivalent features and extensions (fp64 and 3d image writes): OpenCL C 3.0 device can support the extension but not the feature. cl_khr_subgroups requires subgroup independent forward progress. This patch adjusts the check which is used when translating language builtins to check either the extension or feature is supported. Reviewed By: Anastasia Differential Revision: https://reviews.llvm.org/D118999
-
Haojian Wu authored
-
AndreyChurbanov authored
Fixed mistaken iterations distribution between different target regions. Differential Revision: https://reviews.llvm.org/D118393
-
Anton Afanasyev authored
-
Marek Kurdej authored
-
Marek Kurdej authored
-
Louis Dionne authored
-
Jay Foad authored
-
Amy Kwan authored
This patch adds an option (no-clang-tools) to disable building clang-tools-extra when performing release testing. Prior to this patch, clang-tools-extra was built by default, but on some platforms (such as AIX), clang-tools-extra is not supported, and so we do not normally build it. Furthermore, this change should not change the invocation for targets that build clang-tools-extra normally. Differential Revision: https://reviews.llvm.org/D119520
-
Nikita Popov authored
Rather than checking that the type is the same (which is always the case, given how these are part of the same phi) check that the source element type is the same. With opaque pointers, this is no longer implied.
-
Zakk Chen authored
Masked reduction intrinsics are specical cases which don't need to have policy operand. The mask only affects which elements are read. It doesn't effect the destination register. The reduction intrinsics have a dedicated destination operand. If it is undef, we use tail agnostic. If it not undef we use tail undisturbed. Co-Authored-by:
Craig Topper <craig.topper@sifive.com> Differential Revision: https://reviews.llvm.org/D117681
-