- Feb 05, 2020
-
-
Jonas Devlieghere authored
Re-land "[lldb] [testsuite] generalize `DWARFASTParserClangTests` based on `DWARFExpressionTest`'s YAML"
-
Jon Chesterfield authored
Summary: [libomptarget] Implement wavefront functions for amdgcn Reviewers: jdoerfert, ABataev, grokos, arsenm Reviewed By: arsenm Subscribers: saiislam, wdng, arsenm, jvesely, openmp-commits Tags: #openmp Differential Revision: https://reviews.llvm.org/D73077
-
Dimitry Andric authored
This was broken by 28c91219, which added `struct_stack_t_sz` only to sanitizer_platform_limits_posix.{cpp,h}.
-
Jan Kratochvil authored
Revert: [lldb] [testsuite] generalize `DWARFASTParserClangTests` based on `DWARFExpressionTest`'s YAML It is causing a failure on OSX, to be investigated more. Differential Revision: https://reviews.llvm.org/D73279
-
Simon Pilgrim authored
We only accept tok::TokenKind::comment enum values so no need to add other cases. Seems to be a cut+paste typo.
-
Matt Arsenault authored
We were handling the case where the result was a struct with an extracted SGPR component, but not for the simple case.
-
Jonas Devlieghere authored
Unless trivially copyable, SB classes that are passed by value should be treated as const references by the reproducer infrastructure and their address should be serialized. This is already the case on the serialization side, but not on the deserialization side. This fixes that by removing the NotImplementedTag.
-
Jonas Devlieghere authored
Currently SBFile isn't really instrumented, which was causing trouble when capturing and replaying the Python test suite. The class is particularly tricky because one of its constructors takes a FileSP which isn't instrumented. Until we have proper shadowing in place, we'll simply always record a nullptr. Differential revision: https://reviews.llvm.org/D73992
-
Simon Pilgrim authored
Be more specific that getOperandConstraint should return -1 or a uint8_t value
-
Matt Arsenault authored
-
Matt Arsenault authored
We don't need to expand to full shifts for the > 32-bit case. This just switches to a sext_inreg of the high half.
-
Matt Arsenault authored
Split the VALU 64-bit case in RegBankSelect.
-
Austin Kerbow authored
https://reviews.llvm.org/D72312 introduced an infinite loop which involves DAGCombiner::visitFMA and AMDGPUTargetLowering::performFNegCombine. fma( a, fneg(b), fneg(c) ) => fneg( fma (a, b, c) ) => fma( a, fneg(b), fneg(c) ) ... This only breaks with types where 'isFNegFree' returns flase, e.g. v4f32. Reproducing the issue also needs the attribute 'no-signed-zeros-fp-math', and no source mods allowed on one of the users of the Op. This fix makes changes to indicate that it is not free to negate a fma if it has users with source mods. Differential Revision: https://reviews.llvm.org/D73939
-
Evgenii Stepanov authored
Subscribers: #sanitizers, llvm-commits Tags: #sanitizers, #llvm Differential Revision: https://reviews.llvm.org/D73987
-
Matt Arsenault authored
The legalization has improved since this was added, and the tests relying on this no longer need it.
-
Michael Liao authored
Summary: - The device compilation needs to have a consistent source code compared to the corresponding host compilation. If macros based on the host-specific target processor is not properly populated, the device compilation may fail due to the inconsistent source after the preprocessor. So far, only the host triple is used to build the macros. If a detailed host CPU target or certain features are specified, macros derived from them won't be populated properly, e.g. `__SSE3__` won't be added unless `+sse3` feature is present. On Windows compilation compatible with MSVC, that missing macros result in that intrinsics are not included and cause device compilation failure on the host-side source. - This patch addresses this issue by introducing two `cc1` options, i.e., `-aux-target-cpu` and `-aux-target-feature`. If a specific host CPU target or certain features are specified, the compiler driver will append them during the construction of the offline compilation actions. Then, the toolchain in `cc1` phase will populate macros accordingly. - An internal option `--gpu-use-aux-triple-only` is added to fall back the original behavior to help diagnosing potential issues from the new behavior. Reviewers: tra, yaxunl Subscribers: cfe-commits Tags: #clang Differential Revision: https://reviews.llvm.org/D73942
-
Richard Smith authored
are equally constrained.
-
Richard Smith authored
-
Tyker authored
-
Craig Topper authored
This time with correct types for the data result from the SUB. Original commit message: Our normal lowering for ISD::SETCC uses X86ISD::SUB to enable CSE unless the RHS is 0. optimizeCompareInstr called by the peephole pass can turn subs with unused results into cmps to clean this up. This commit makes other places that create X86ISD::CMP have the same behavior.
-
River Riddle authored
This allows for reusing the internal state of the printer, which is more efficient and also allows for using type aliases
-
Teresa Johnson authored
Summary: It can be useful to tune the default inline threshold without overriding other inlining thresholds (e.g. in code compiled for size). The existing `-inline-threshold` flag overrides other thresholds, so it is insufficient in codebases where there is a mix of code compiled for size and speed. Patch by Michael Holman <michael.holman@microsoft.com> Reviewers: eraman, tejohnson Reviewed By: tejohnson Subscribers: tejohnson, mtrofin, davidxl, hiraditya, haicheng, llvm-commits Tags: #llvm Differential Revision: https://reviews.llvm.org/D73217
-
Matt Arsenault authored
-
Matt Arsenault authored
-
Paula Toth authored
Reviewers: sivachandra, abrachet Reviewed By: sivachandra, abrachet Subscribers: libc-commits, MaskRay, tschuett Tags: #libc-project, #llvm Differential Revision: https://reviews.llvm.org/D72248
-
Tyker authored
Summary: Changes: - Calls to consteval function are now evaluated in constant context but IR is still generated for them. - Add diagnostic for taking address of a consteval function in non-constexpr context. - Add diagnostic for address of consteval function accessible at runtime. - Add tests Reviewers: rsmith, aaron.ballman Reviewed By: rsmith Subscribers: mgrang, riccibruno, cfe-commits Tags: #clang Differential Revision: https://reviews.llvm.org/D63960
-
Aaron Watry authored
Fixes a wimpy-mode CTS failure for asin(float). Passes non-wimpy for both float/double on RX580. Signed-off-by:
Aaron Watry <awatry@gmail.com> Tested-by:
Jan Vesely <jan.vesely@rutgers.edu> Reviewed-by:
Jan Vesely <jan.vesely@rutgers.edu>
-
Matt Arsenault authored
shouldOptimizeForSize is showing up in a profile, spending around 10% of the pass time in one function. This should probably not be so slow, but the much cheaper attribute check should be done first anyway.
-
aartbik authored
Summary: A few details were missing in the description. These changes makes the documented code "compile". Reviewers: nicolasvasilache, andydavis1 Reviewed By: nicolasvasilache, andydavis1 Subscribers: merge_guards_bot, mehdi_amini, rriddle, jpienaar, burmako, shauheen, antiagainst, nicolasvasilache, arpith-jacob, mgester, lucyrfox, liufengdb, Joonsoo, llvm-commits Tags: #llvm Differential Revision: https://reviews.llvm.org/D73923
-
Jonathan Coe authored
Summary: Add comments to the list of tokens that can follow the ']' at the end of a C# attribute specifier to prevent comments after attribute specifiers from being formatted as continuations. Reviewers: MyDeveloperDay, krasimir Reviewed By: MyDeveloperDay Tags: #clang-format Differential Revision: https://reviews.llvm.org/D73977
-
Matt Arsenault authored
Prepare to accurately track the future denormal-fp-math attribute changes. The way to actually set these separately is not wired in yet. This is just a mechanical change, and mostly still assumes the input and output mode match. This should be refined for some cases. For example, fcanonicalize lowering should use the flushing variant if either input or output flushing is enabled
-
Vedant Kumar authored
-
Vedant Kumar authored
In order to synthesize tail call frames, the stack frame list must not be empty (otherwise, there is no "previous" frame to infer a tail call from). This case is hard to hit. To trigger it, we must first fail to push `unwind_frame_sp` because we either fail to get its SymbolContext, or given its SymbolContext the GetParentOfInlineScope call fails. This causes m_concrete_frames_fetched to be incremented while m_frames remains empty. Then, the next frame in the stack may fail within SynthesizeTailCallFrames. This crash arose during a kernel debugging session. rdar://59147051
-
Jacques Pienaar authored
-
Fangrui Song authored
Make usage more consistent, and make it possible to enable LongOptionsUseDoubleDash.
-
Matt Arsenault authored
The usage of the Imm out argument from SelectSMRDOffset is pretty confusing. Stop trying to reject CI immediates in the case where the offset field can be used. It's not an illegal way to encode the immediate, so just prefer the better encoding pattern with AddedComplexity. We probably don't even really need the different opcodes for the different offset types anymore, but that will be more work to cleanup. The SMRD non-buffer load patterns could also use a cleanup to be done separately.
-
Matt Arsenault authored
-
Simon Pilgrim authored
We weren't account for load latencies in the SSE42/AES/CLMUL schedule classes
-
Matt Arsenault authored
-
Yitzhak Mandelbaum authored
Summary: Spells out some `auto`s explicitly and adds another test for the matcher `isExpandedFromMacro`. Reviewers: aaron.ballman Subscribers: gribozavr, cfe-commits Tags: #clang Differential Revision: https://reviews.llvm.org/D73975
-