mikros/rust - rust - Gitea.pterpstra.com

Author	SHA1	Message	Date
Josh Stone	1b79bb937f	Add inline comments why we're forcing the target cpu	2024-05-01 16:54:20 -07:00
Josh Stone	706f06c39a	Use an explicit x86-64 cpu in tests that are sensitive to it There are a few tests that depend on some target features not being enabled by default, and usually they are correct with the default x86-64 target CPU. However, in downstream builds we have modified the default to fit our distros -- `x86-64-v2` in RHEL 9 and `x86-64-v3` in RHEL 10 -- and the latter especially trips tests that expect not to have AVX. These cases are few enough that we can just set them back explicitly.	2024-05-01 15:25:26 -07:00
Matthias Krüger	d81e444c8e	Rollup merge of #124543 - maurer:llvm-range, r=nikic codegen tests: Tolerate `range()` qualifications in enum tests Current LLVM can infer range bounds on the i8s involved with these tests, and annotates it. Accept these bounds if present. `@rustbot` label: +llvm-main cc `@durin42`	2024-04-30 06:43:43 +02:00
Matthew Maurer	8101884b37	codegen tests: Tolerate `range()` qualifications in enum tests Current LLVM can infer range bounds on the i8s involved with these tests, and annotates it. Accept these bounds if present.	2024-04-30 00:02:49 +00:00
Krasimir Georgiev	52ea73a540	adapt a codegen test for llvm 19 No functional changes intended. Found by our experimental rust + LLVM @ HEAD bot: https://buildkite.com/llvm-project/rust-llvm-integrate-prototype/builds/27747#018f2570-018c-4b12-9c5a-38cf81453683/957-965	2024-04-29 13:03:45 +00:00
bors	284f94f9c0	Auto merge of #121298 - nikic:writable, r=cuviper Set writable and dead_on_unwind attributes for sret arguments Set the `writable` and `dead_on_unwind` attributes for `sret` arguments. This allows call slot optimization to remove more memcpy's. See https://llvm.org/docs/LangRef.html#parameter-attributes for the specification of these attributes. In short, the statement we're making here is that: * The return slot is writable. * The return slot will not be read if the function unwinds. Fixes https://github.com/rust-lang/rust/issues/90595.	2024-04-25 04:31:56 +00:00
Nikita Popov	976267b514	Add needs-unwind to codegen test When compiled with -C panic=abort we'd generate an extra panic_cannot_unwind shim in the variant calling C-unwind.	2024-04-25 11:44:32 +09:00
Nikita Popov	137775dd63	Fix incorrect CHECK-LABEL	2024-04-25 11:43:47 +09:00
Nikita Popov	3695af697e	Set writable and dead_on_unwind attributes for sret arguments	2024-04-25 11:43:47 +09:00
Gary Guo	cfee72aa24	Fix tests and bless	2024-04-24 13:12:33 +01:00
Oli Scherer	aef0f4024a	Error on using `yield` without also using `#[coroutine]` on the closure And suggest adding the `#[coroutine]` to the closure	2024-04-24 08:05:29 +00:00
bors	29a56a3b1c	Auto merge of #122053 - erikdesjardins:alloca, r=nikic Stop using LLVM struct types for alloca The alloca type has no semantic meaning, only the size (and alignment, but we specify it explicitly) matter. Using `[N x i8]` is a more direct way to specify that we want `N` bytes, and avoids relying on LLVM's struct layout. It is likely that a future LLVM version will change to an untyped alloca representation. Split out from #121577. r? `@ghost`	2024-04-24 03:00:44 +00:00
Matthias Krüger	918304b190	Rollup merge of #124003 - WaffleLapkin:dellvmization, r=scottmcm,RalfJung,antoyo Dellvmize some intrinsics (use `u32` instead of `Self` in some integer intrinsics) This implements https://github.com/rust-lang/compiler-team/issues/693 minus what was implemented in #123226. Note: I decided to _not_ change `shl`/... builder methods, as it just doesn't seem worth it. r? ``@scottmcm``	2024-04-23 20:17:51 +02:00
Markus Reiter	33e68aadc9	Stabilize generic `NonZero`.	2024-04-22 18:48:47 +02:00
Mark Rousskov	f1ae5314be	Avoid reloading Vec::len across grow_one in push This saves an extra load from memory.	2024-04-20 21:07:00 -04:00
Scott McMurray	986d9f104b	Make `checked` ops emit unchecked LLVM operations where feasible For things with easily pre-checked overflow conditions -- shifts and unsigned subtraction -- write then checked methods in such a way that we stop emitting wrapping versions of them. For example, today <https://rust.godbolt.org/z/qM9YK8Txb> neither ```rust a.checked_sub(b).unwrap() ``` nor ```rust a.checked_sub(b).unwrap_unchecked() ``` actually optimizes to `sub nuw`. After this PR they do.	2024-04-18 18:11:21 -07:00
Scott McMurray	d05545c05d	At debuginfo=0, don't inline debuginfo when inlining	2024-04-18 09:35:35 -07:00
Maybe Waffle	c2046c4b09	Add codegen tests for changed intrinsics	2024-04-16 12:35:22 +00:00
bors	5dcb678ad8	Auto merge of #122917 - saethlin:atomicptr-to-int, r=nikic Add the missing inttoptr when we ptrtoint in ptr atomics Ralf noticed this here: https://github.com/rust-lang/rust/pull/122220#discussion_r1535172094 Our previous codegen forgot to add the cast back to integer type. The code compiles anyway, because of course all locals are in-memory to start with, so previous codegen would do the integer atomic, store the integer to a local, then load a pointer from that local. Which is definitely _not_ what we wanted: That's an integer-to-pointer transmute, so all pointers returned by these `AtomicPtr` methods didn't have provenance. Yikes. Here's the IR for `AtomicPtr::fetch_byte_add` on 1.76: https://godbolt.org/z/8qTEjeraY ```llvm define noundef ptr `@atomicptr_fetch_byte_add(ptr` noundef nonnull align 8 %a, i64 noundef %v) unnamed_addr #0 !dbg !7 { start: %0 = alloca ptr, align 8, !dbg !12 %val = inttoptr i64 %v to ptr, !dbg !12 call void `@llvm.lifetime.start.p0(i64` 8, ptr %0), !dbg !28 %1 = ptrtoint ptr %val to i64, !dbg !28 %2 = atomicrmw add ptr %a, i64 %1 monotonic, align 8, !dbg !28 store i64 %2, ptr %0, align 8, !dbg !28 %self = load ptr, ptr %0, align 8, !dbg !28 call void `@llvm.lifetime.end.p0(i64` 8, ptr %0), !dbg !28 ret ptr %self, !dbg !33 } ``` r? `@RalfJung` cc `@nikic`	2024-04-15 08:07:47 +00:00
Matthias Krüger	4a0e9e0deb	Rollup merge of #123249 - goolmoos:naked_variadics, r=pnkfelix do not add prolog for variadic naked functions fixes #99858	2024-04-12 17:41:33 +02:00
Erik Desjardins	daaaacdcb3	remove alloca type from issue-105386-ub-in-debuginfo It's irrelevant for the purposes of this test (there is only one alloca) and its size changes depending on the target, so it can't be matched easily.	2024-04-12 08:36:22 -04:00
Guy Shefy	9139d7252d	do not add prolog for variadic naked functions fixes #99858	2024-04-12 15:29:39 +03:00
Erik Desjardins	f4426c189f	use [N x i8] for alloca types	2024-04-11 21:42:35 -04:00
Matthew Maurer	e70cf014b8	codegen tests: Tolerate `nuw` `nsw` on `trunc` llvm/llvm-project#87910 infers `nuw` and `nsw` on some `trunc` instructions we're doing `FileCheck` on. Tolerate but don't require them to support both release and head LLVM.	2024-04-11 17:20:08 +00:00
León Orell Valerian Liehr	aac3f24054	Rollup merge of #122470 - tgross35:f16-f128-step4-libs-min, r=Amanieu `f16` and `f128` step 4: basic library support This is the next step after https://github.com/rust-lang/rust/pull/121926, another portion of https://github.com/rust-lang/rust/pull/114607 Tracking issue: https://github.com/rust-lang/rust/issues/116909 This PR adds the most basic operations to `f16` and `f128` that get lowered as LLVM intrinsics. This is a very small step but it seemed reasonable enough to add unopinionated basic operations before the larger modules that are built on top of them. r? ```@Amanieu``` since you were pretty involved in the RFC cc ```@compiler-errors``` ```@rustbot``` label +T-libs-api +S-blocked +F-f16_and_f128	2024-04-11 01:56:23 +02:00
Trevor Gross	454de78ea3	Add basic library support for `f16` and `f128` Implement basic operation traits that get lowered to intrinsics. This includes codegen tests for implemented operations.	2024-04-10 13:50:27 -04:00
bors	c2239bca5b	Auto merge of #123185 - scottmcm:more-typed-copy, r=compiler-errors Remove my `scalar_copy_backend_type` optimization attempt I added this back in https://github.com/rust-lang/rust/pull/111999 , but I no longer think it's a good idea - It had to get scaled back to only power-of-two things to not break a bunch of targets - LLVM seems to be getting better at memcpy removal anyway - Introducing vector instructions has seemed to sometimes (https://github.com/rust-lang/rust/pull/115515#issuecomment-1750069529) make autovectorization worse So this removes it from the codegen crates entirely, and instead just tries to use <https://doc.rust-lang.org/nightly/nightly-rustc/rustc_codegen_ssa/traits/builder/trait.BuilderMethods.html#method.typed_place_copy> instead of direct `memcpy` so things will still use load/store when a type isn't `OperandValue::Ref`.	2024-04-10 16:32:41 +00:00
Scott McMurray	593e900ad2	Update 122805 test for PR 123185	2024-04-10 08:28:43 -07:00
Matthias Krüger	2ddf984594	Rollup merge of #123612 - kxxt:riscv-target-abi, r=jieyouxu,nikic,DianQK Set target-abi module flag for RISC-V targets Fixes cross-language LTO on RISC-V targets (Fixes #121924)	2024-04-10 04:27:40 +02:00
Scott McMurray	b5376ba601	Remove my `scalar_copy_backend_type` optimization attempt I added this back in 111999, but I no longer think it's a good idea - It had to get scaled back to only power-of-two things to not break a bunch of targets - LLVM seems to be getting better at memcpy removal anyway - Introducing vector instructions has seemed to sometimes (115515) make autovectorization worse So this removes it from the codegen crates entirely, and instead just tries to use <https://doc.rust-lang.org/nightly/nightly-rustc/rustc_codegen_ssa/traits/builder/trait.BuilderMethods.html#method.typed_place_copy> instead of direct `memcpy` so things will still use load/store for immediates.	2024-04-09 08:51:32 -07:00
kxxt	f19c48e7a8	Set target-abi module flag for RISC-V targets Fixes cross-language LTO on RISC-V targets (Fixes #121924)	2024-04-09 05:25:51 +02:00
bors	59c808fcd9	Auto merge of #122387 - DianQK:re-enable-early-otherwise-branch, r=cjgillot Re-enable the early otherwise branch optimization Closes #95162. Fixes #119014. This is the first part of #121397. An invalid enum discriminant can come from anywhere. We have to check to see if all successors contain the discriminant statement. This should have a pass to hoist instructions. r? cjgillot	2024-04-09 01:02:29 +00:00
bors	ab5bda1aa7	Auto merge of #123645 - matthiaskrgr:rollup-yd8d7f1, r=matthiaskrgr Rollup of 9 pull requests Successful merges: - #122781 (Fix argument ABI for overaligned structs on ppc64le) - #123367 (Safe Transmute: Compute transmutability from `rustc_target::abi::Layout`) - #123518 (Fix `ByMove` coroutine-closure shim (for 2021 precise closure capturing behavior)) - #123547 (bootstrap: remove unused pub fns) - #123564 (Don't emit divide-by-zero panic paths in `StepBy::len`) - #123578 (Restore `pred_known_to_hold_modulo_regions`) - #123591 (Remove unnecessary cast from `LLVMRustGetInstrProfIncrementIntrinsic`) - #123632 (parser: reduce visibility of unnecessary public `UnmatchedDelim`) - #123635 (CFI: Fix ICE in KCFI non-associated function pointers) r? `@ghost` `@rustbot` modify labels: rollup	2024-04-08 20:31:08 +00:00
Matthias Krüger	9570ac4d28	Rollup merge of #123564 - scottmcm:step-by-div-zero, r=joboet Don't emit divide-by-zero panic paths in `StepBy::len` I happened to notice today that there's actually two such calls emitted in the assembly: <https://rust.godbolt.org/z/1Wbbd3Ts6> Since they're impossible, hopefully telling LLVM that will also help optimizations elsewhere.	2024-04-08 22:06:22 +02:00
Matthias Krüger	ecfc3384f1	Rollup merge of #122781 - nikic:ppc-abi-fix, r=cuviper Fix argument ABI for overaligned structs on ppc64le When passing a 16 (or higher) aligned struct by value on ppc64le, it needs to be passed as an array of `i128` rather than an array of `i64`. This will force the use of an even starting doubleword. For the case of a 16 byte struct with alignment 16 it is important that `[1 x i128]` is used instead of `i128` -- apparently, the latter will get treated similarly to `[2 x i64]`, not exhibiting the correct ABI. Add a `force_array` flag to `Uniform` to support this. The relevant clang code can be found here: `fe2119a7b0/clang/lib/CodeGen/Targets/PPC.cpp (L878-L884)` `fe2119a7b0/clang/lib/CodeGen/Targets/PPC.cpp (L780-L784)` I think the corresponding psABI wording is this: > Fixed size aggregates and unions passed by value are mapped to as > many doublewords of the parameter save area as the value uses in > memory. Aggregrates and unions are aligned according to their > alignment requirements. This may result in doublewords being > skipped for alignment. In particular the last sentence. Though I didn't find any wording for Clang's behavior of clamping the alignment to 16. Fixes https://github.com/rust-lang/rust/issues/122767. r? `@cuviper`	2024-04-08 22:06:20 +02:00
bors	211518e5fb	Auto merge of #120614 - DianQK:simplify-switch-int, r=cjgillot Transforms match into an assignment statement Fixes #106459. We should be able to do some similar transformations, like `enum` to `enum`. r? mir-opt	2024-04-08 18:28:50 +00:00
bors	537aab7a2e	Auto merge of #120131 - oli-obk:pattern_types_syntax, r=compiler-errors Implement minimal, internal-only pattern types in the type system rebase of https://github.com/rust-lang/rust/pull/107606 You can create pattern types with `std::pat::pattern_type!(ty is pat)`. The feature is incomplete and will panic on you if you use any pattern other than integral range patterns. The only way to create or deconstruct a pattern type is via `transmute`. This PR's implementation differs from the MCP's text. Specifically > This means you could implement different traits for different pattern types with the same base type. Thus, we just forbid implementing any traits for pattern types. is violated in this PR. The reason is that we do need impls after all in order to make them usable as fields. constants of type `std::time::Nanoseconds` struct are used in patterns, so the type must be structural-eq, which it only can be if you derive several traits on it. It doesn't need to be structural-eq recursively, so we can just manually implement the relevant traits on the pattern type and use the pattern type as a private field. Waiting on: * [x] move all unrelated commits into their own PRs. * [x] fix niche computation (see 2db07f94f44f078daffe5823680d07d4fded883f) * [x] add lots more tests * [x] T-types MCP https://github.com/rust-lang/types-team/issues/126 to finish * [x] some commit cleanup * [x] full self-review * [x] remove 61bd325da19a918cc3e02bbbdce97281a389c648, it's not necessary anymore I think. * [ ] ~~make sure we never accidentally leak pattern types to user code (add stability checks or feature gate checks and appopriate tests)~~ we don't even do this for the new float primitives * [x] get approval that [the scope expansion to trait impls](https://rust-lang.zulipchat.com/#narrow/stream/326866-t-types.2Fnominated/topic/Pattern.20types.20types-team.23126/near/427670099) is ok r? `@BoxyUwU`	2024-04-08 16:25:23 +00:00
Oli Scherer	84acfe86de	Actually create ranged int types in the type system.	2024-04-08 12:02:19 +00:00
DianQK	928c57dc9a	Add test case for #119014	2024-04-08 19:20:04 +08:00
DianQK	1f061f47e2	Transforms match into an assignment statement	2024-04-08 19:00:53 +08:00
Philippe-Cholet	7a2678de7d	Add invariant to VecDeque::pop_* that len < cap if pop successful Similar to #114370 for VecDeque instead of Vec. It now uses `core::hint::assert_unchecked`.	2024-04-08 12:12:13 +02:00
Kai Luo	d8d1e6ce21	Limited to little endian target	2024-04-08 11:11:11 +08:00
Nikita Popov	009280c5e3	Fix argument ABI for overaligned structs on ppc64le When passing a 16 (or higher) aligned struct by value on ppc64le, it needs to be passed as an array of `i128` rather than an array of `i64`. This will force the use of an even starting register. For the case of a 16 byte struct with alignment 16 it is important that `[1 x i128]` is used instead of `i128` -- apparently, the latter will get treated similarly to `[2 x i64]`, not exhibiting the correct ABI. Add a `force_array` flag to `Uniform` to support this. The relevant clang code can be found here: `fe2119a7b0/clang/lib/CodeGen/Targets/PPC.cpp (L878-L884)` `fe2119a7b0/clang/lib/CodeGen/Targets/PPC.cpp (L780-L784)` I think the corresponding psABI wording is this: > Fixed size aggregates and unions passed by value are mapped to as > many doublewords of the parameter save area as the value uses in > memory. Aggregrates and unions are aligned according to their > alignment requirements. This may result in doublewords being > skipped for alignment. In particular the last sentence. Fixes https://github.com/rust-lang/rust/issues/122767.	2024-04-08 11:15:36 +09:00
bors	4e431fad67	Auto merge of #123561 - saethlin:str-unchecked-sub-index, r=scottmcm Use unchecked_sub in str indexing https://github.com/rust-lang/rust/pull/108763 applied this logic to indexing for slices, but of course `str` has its own separate impl. Found this by skimming over the codegen for https://github.com/oxidecomputer/hubris/; their dist builds enable overflow checks so the lack of `unchecked_sub` was producing an impossible-to-hit overflow check and also inhibiting some inlining. r? scottmcm	2024-04-07 12:49:15 +00:00
bors	0e3235f85b	Auto merge of #123555 - DianQK:update-llvm-18, r=cuviper Update to LLVM 18.1.3 Fixes #122805. This should work on all targets: https://rust.godbolt.org/z/svW8ha31z. r? `@cuviper`	2024-04-07 06:33:58 +00:00
DianQK	5acfe772fa	Add the test case for #122805	2024-04-07 13:01:54 +08:00
Scott McMurray	00bd24766f	Don't emit divide-by-zero panic paths in `StepBy::len` I happened to notice today that there's actually two such calls emitted in the assembly: <https://rust.godbolt.org/z/1Wbbd3Ts6> Since they're impossible, hopefully telling LLVM that will also help optimizations elsewhere.	2024-04-06 11:37:57 -07:00
Ben Kimock	712aab72df	Use unchecked_sub in str indexing	2024-04-06 14:09:03 -04:00
Ben Kimock	a7912cb421	Put checks that detect UB under their own flag below debug_assertions	2024-04-06 11:21:47 -04:00
Matthias Krüger	ad3df4919d	Rollup merge of #123525 - maurer:no-id-dyn2, r=compiler-errors CFI: Don't rewrite ty::Dynamic directly Now that we're using a type folder, the arguments in predicates are processed automatically - we don't need to descend manually. We also want to keep projection clauses around, and this does so. r? `@compiler-errors`	2024-04-06 08:56:35 +02:00

1 2 3 4 5 ...

585 Commits