- Sep 09, 2023
-
-
Zahira Ammarguellat authored
https://lab.llvm.org/buildbot/#/builders/38/builds/14427. Differential Revision: https://reviews.llvm.org/D159486
-
Dmitri Gribenko authored
-
bzEq authored
-
Valentin Clement (バレンタイン クレメン) authored
OpenACC 3.3 - 2.10 The cache directive may appear at the top of (inside of) a loop. This patch adds a semantic check to ensure the cache directive is inside a loop.
-
Jon Phillips authored
Avoid unnecessarily aggressive line-breaking when using "LambdaBodyIndentation: OuterScope" with argument bin-packing. Differential Revision: https://reviews.llvm.org/D148131
-
Fabian Mora authored
This patch employs the updated promise mechanism to enforce Target Attribute IR constraints. Due to this patch, TargetAttributes implementations no longer have to be registered before executing translation to LLVM IR in cases where they are not needed, like when translating `gpu.binary` operations.
-
Bjorn Pettersson authored
Make sure the TestCount const definition is guarded the same way as the use of the constant. This is an attempt to fix buildbot failures related to -Wunused-const-variable.
-
Thomas authored
recommit https://github.com/llvm/llvm-project/pull/65432 with minor bug fix for bitcasts
-
Alexey Bataev authored
If the scalar does not need to be scheduled and it was vectorized already in one of the vector nodes, we still can try to vectorize it in another node. Just does not need account its cost in the scalar total cost, as it will be handled in the main vectorized node. Differential Revision: https://reviews.llvm.org/D159205
-
Philip Reames authored
This covers the bug identified in review of pr 65777.
-
Jakub Kuderski authored
`concept` is a c++20 keyword and makes some buildbots fail: https://lab.llvm.org/buildbot/#/builders/160/builds/24345
-
Siva Chandra authored
-
jwanggit86 authored
For the AMD GFX90A GPU, the SCC instruction modifier is allowed for certain classes of instructions. However, the current assembler generates an error message, "scc is not supported on this GPU", regardless of the instruciton. This fix modifies the message as well as the logic for generating the message. Related tests are moved from gfx90a_err.s to gfx90a_asm_features.s. Co-authored-by:Jun Wang <jun.wang7@amd.com>
-
Shilei Tian authored
This reverts commit c5525a6e. AMD BB is not happy again.
-
Mehdi Amini authored
Aligning the section about pull-request with the `gh` tools with the section using the web UI (#65795) This is fairly minor, but start addressing the post-review comments
-
Shilei Tian authored
This reverts commit e592c2dc that reverts e91e3cf0.
-
Anton Korobeynikov authored
-
Shilei Tian authored
`AAAddressSpace` currently only works for `LoadInst` and `StoreInst` currently. For `StoreInst`, the corresponding use can be the pointer operand, or value operand, or both. When it is used as value operand, it can prevent `AMDGPUPromoteAlloca` from optimization in certain cases. This patch changes the manifest method such that only pointer operand will be rewritten.
-
Joseph Huber authored
Summary: The GPU uses separate implementations to perform file IO. This is all done through the RPC interface and we kept it minimal such that we could treat a `stdin`, `stdout`, or `stderr` handle from the CPU correctly on the GPU. The RPC implementation uses different opcodes for whether or not we are using one of the standard streams. This is so we do not need to initialize anything to access the CPU's standard stream, because the server knows that it should print to `stdout` if it gets the `STDOUT` variant of the opcode. It also saves us an RPC call, which are expensive relatively speaking. This patch simply cleans up this interface to make them all use a common function. This is done in preparation to implement some more file IO functions like getc or putc.
-
Johannes Doerfert authored
Through the new `Attributor::checkForAllCallees` we can look through indirect calls and visit all potential callees if they are known. Most AAs will do that implicitly now via `AACalleeToCallSite`, thus, most AAs are able to deal with missing callees for call site IR positions. Differential Revision: https://reviews.llvm.org/D112290
-
Zero Omega authored
There are missing include and using in TextStubTests and AsmPrinterDwarfTest and they causes build failures when using vanilla GoogleTest v1.14.0. This patch fixes this issue.
-
Tom Stellard authored
Secrets are not available for workflows triggered by PRs, so we need to split the pr-subscriber action into two separate actions. The first will listen for new labels on PRs and the second will add a comment with the team mention. See https://securitylab.github.com/research/github-actions-preventing-pwn-requests/
-
Siva Chandra authored
-
Konstantin Varlamov authored
This is to bring `main` in sync with `release/17.x` after https://reviews.llvm.org/D159171. Differential Revision: https://reviews.llvm.org/D159454
-
Michael Liao authored
-
Joseph Huber authored
Summary: AMDGPU binaries use a "code object" as the ABI indicator. We are currently trying to move over to a newer code object. We want these library functions to use the "generic" or default ABI such that it is specified when linked into the user application. Currently this will default to v4 as the startup code will use whatever the current default is.
-
Ethan Luis McDonough authored
Executable allocate directives require that list items show up in the corresponding allocate statement. This patch is dependent on revision D150428 and applies the semantic check introduced there to allocate directives associated with allocate statements. Reviewed By: kiranchandramohan Differential Revision: https://reviews.llvm.org/D150483
-
Jakub Kuderski authored
- Fix operand/attribute order - Use ODS for parsing/printing - Allow for stride to be any integer type
-
Nathan Chancellor authored
When building the LoongArch Linux kernel without `CONFIG_DYNAMIC_FTRACE`, the build fails to link because the mcount symbol is `mcount`, not `_mcount` like GCC generates and the kernel expects: ``` ld.lld: error: undefined symbol: mcount >>> referenced by version.c >>> init/version.o:(early_hostname) in archive vmlinux.a >>> referenced by do_mounts.c >>> init/do_mounts.o:(rootfs_init_fs_context) in archive vmlinux.a >>> referenced by main.c >>> init/main.o:(__traceiter_initcall_level) in archive vmlinux.a >>> referenced 97011 more times >>> did you mean: _mcount >>> defined in: vmlinux.a(arch/loongarch/kernel/mcount.o) ``` Set `MCountName` in `LoongArchTargetInfo` to `_mcount`, which resolves the build failure.
-
Matthias Springer authored
This commit provides a default implementation for all ops that implement the `DestinationStyleOpInterface`. Result values of such ops are tied to operand, and those have the same type.
-
Tom Stellard authored
-
Jake Egan authored
The test is hardcoded to fail after passing `test_ill_formed_utf16()`. It passes on 32-bit AIX if we remove this. Reviewed By: Mordante, #libc, ldionne Differential Revision: https://reviews.llvm.org/D150273
-
Jakub Kuderski authored
- Fix order of operands/attributes - Allow for stride to be any integer type - Use ODS for parsing/printing - Update examples and tests - Fix a typo in SPIR-V tblgen code
-
Dmitri Gribenko authored
This reverts commit db5d845c. As per PR discussion "Looks like we've missed lowering of bitcasts between v2f16 and v2i16 and it breaks XLA."
-
Tom Stellard authored
The labeler doesn't seem to support these.
-
David Green authored
It appears that these were the wrong way around, with the wrong type sizes taking extra cycles. The smaller i8 sizes are now the ones marked as taking longer.
-
Tom Stellard authored
-
Shraiysh authored
-
Hongtao Yu authored
With `-fpseudo-probe-for-profiling`, the dwarf discriminator for a callsite will be overwritten to pseudo probe related information for that callsite. The probe information is encoded in a special format (i.e., with all lowest three digits be one) in order to be distinguished from regular dwarf discriminator. The special encoding format will be decoded to zero by the regular discriminator logic. This means all callsites would have a zero discriminator in both the sample profile and the compiler, for classic AutoFDO. This is inconvenient in that no decent classic AutoFDO can be generated from a pseudo probe build. I'm mitigating the issue by allowing callsite probe id to be used as the base dwarf discriminator for classic AutoFDO, since probe id is also unique and can be used to differentiate callsites on the same source line.
-
Tom Stellard authored
-