SPIRV-Tools

mirror of https://github.com/KhronosGroup/SPIRV-Tools synced 2024-11-29 22:41:03 +00:00

Author	SHA1	Message	Date
greg-lunarg	d11725b1d4	Add --relax-float-ops and --convert-relaxed-to-half (#2808 ) The first pass applies the RelaxedPrecision decoration to all executable instructions with float32 based type results. The second pass converts all executable instructions with RelaxedPrecision result to the equivalent float16 type, inserting converts where necessary.	2019-09-03 13:22:13 -04:00
Jamie Madill	1c9ca422dd	GN: Make SPIRV-Tools target use public_deps. (#2828 ) Should prevent invalid header usage warnings in dependent targets. See http://anglebug.com/3876 for context.	2019-09-03 12:57:32 -04:00
Steven Perron	b54d950298	Fold Fmix should accept vector operands. (#2826 ) Fixes #2819	2019-09-03 09:17:18 -04:00
Alastair Donaldson	2c5ed16ba9	Fix end comments in header files (#2829 ) The end comments for the #ifndef ... #endif macros in various header files containd a stray #define.	2019-09-02 17:31:27 -04:00
Ben Clayton	65e362b7ae	AggressiveDCEPass: Set modified to true when appending to to_kill_ (#2825 ) Also add an assertion that these `modified` is true if to_kill_ has a non-zero size to catch this sort of issue in the pass. Fixes: #2824	2019-08-30 16:27:22 -04:00
Steven Perron	d67130caca	Replace SwizzleInvocationsAMD extended instruction. (#2823 ) Part of #2814	2019-08-30 14:07:24 -04:00
Steven Perron	ad71c057c7	Replace SwizzleInvocationsMaskedAMD extended instruction. (#2822 ) Part of #2814	2019-08-30 10:48:42 -04:00
Corentin Wallez	4ae9b71651	Fix gn check (#2821 ) spriv-opt was missing a dependency on the AMD ballot extension that was needed because it uses a header in the AMD ext to KHR ext pass.	2019-08-30 09:09:34 -04:00
Steven Perron	35d98be3bc	Amd ext to khr (#2811 ) Add the first steps to removing the AMD extension VK_AMD_shader_ballot. Splitting up to make the PRs smaller. Adding utilities to add capabilities and change the version of the module. Replaces the instructions: OpGroupIAddNonUniformAMD = 5000 OpGroupFAddNonUniformAMD = 5001 OpGroupFMinNonUniformAMD = 5002 OpGroupUMinNonUniformAMD = 5003 OpGroupSMinNonUniformAMD = 5004 OpGroupFMaxNonUniformAMD = 5005 OpGroupUMaxNonUniformAMD = 5006 OpGroupSMaxNonUniformAMD = 5007 and extentend instructions WriteInvocationAMD = 3 MbcntAMD = 4 Part of #2814	2019-08-29 12:48:17 -04:00
Ben Clayton	5a581e738c	spvtools::Optimizer - don't assume original_binary and optimized_binary are aliased (#2799 ) If they are not aliased, the function will always print the message: "Binary unexpectedly changed despite optimizer saying there was no change" Which is (usually) totally bogus. Fixes #2798	2019-08-29 10:04:55 -04:00
Steven Perron	73422a0a5e	Check feature mgr in context consistency check (#2818 ) We add a check that the feature manager is correcter after each pass. This resulted in a couple failing tests cases. Those are fixed. Part of #2814	2019-08-28 11:49:16 -04:00
Steven Perron	15fc19d091	Refactor instruction folders (#2815 ) * Refactor instruction folders We want to refactor the instruction folder to allow different sets of rules to be added to the instruction folder. We might want different sets of rules in different circumstances. We also need a way to add rules for extended instructions. Changes are made to the FoldingRules class and ConstFoldingRules class to enable that. We added tests to check that we can fold extended instructions using the new framework. At the same time, I noticed that there were two tests that did not tests what they were suppose to. They could not be easily salvaged. #2813 was opened to track adding the new tests.	2019-08-26 18:54:11 -04:00
jonahryandavis	1eb89172a8	Add missing files to BUILD.gn (#2809 ) New files missing from BUILD.gn caused build failures in Chromium and ANGLE: remove_relaxed_precision_decoration_opportunity_finder.cpp remove_relaxed_precision_decoration_opportunity_finder.h	2019-08-23 13:14:34 -04:00
Alastair Donaldson	8336d1925f	Extend reducer to remove relaxed precision decorations (#2797 ) Adds a reduction pass that removes OpDecorate and OpMemberDecorate instructions that annotate instructions and members with RelaxedPrecision. As well as being useful in its own right, removing such references allows other passes to remove further instructions.	2019-08-22 23:33:09 +01:00
Steven Perron	b00ef0d26e	Handle Id overflow in private-to-local (#2807 ) We need to handle id overflow in the private to local pass. Fixes https://crbug.com/962295	2019-08-22 09:14:48 -04:00
Steven Perron	aef8f92b2b	Even more id overflow in sroa (#2806 ) Now we need to handle id overflow when we overflow while replacing uses of the variable. While looking at this code, I noticed an error in the way we handle access chains that cannot be replaced because of overflow. Name it will make some change, and then give up by returning SuccessWithoutChange. But it was changed. This is fixed up by returning Failure if we notice the error at the time of rewriting the users. This is for both id overflow or out-of-bounds accesses. Code is added to "CheckUses" to remove variables that have out-of-bounds accesses from the candidate list, so we don't even try to rewrite its uses. Fixes https://crbug.com/995032	2019-08-21 13:12:42 -04:00
Steven Perron	c5d1dab99e	Add name for variables in desc sroa (#2805 ) Fixes #2802.	2019-08-21 10:55:02 -04:00
David Neto	0cbdc7a2c3	Remove unimplemented method declaration (#2804 )	2019-08-20 08:53:27 -04:00
Steven Perron	bc62722b80	Handle overflow in wrap-opkill (#2801 ) Fixes https://crbug/994203	2019-08-18 19:00:18 -04:00
Steven Perron	9cd07272a6	More handle overflow in sroa (#2800 ) If we run out of ids when creating a new variable, sroa does not recognize the error, and continues doing work. This leads to segmentation faults. Fixes https://crbug/969655	2019-08-16 13:15:17 -04:00
greg-lunarg	06407250a1	Instrument: Add support for Buffer Device Address extension (#2792 )	2019-08-16 09:18:34 -04:00
Toomas Remmelg	7b4e5bd5ec	Update remquo validation to match the OpenCL Extended Instruction Set Specification (#2791 )	2019-08-15 09:38:37 -04:00
Jaebaek Seo	dac9210dcb	Use ascii code based characters (#2796 ) My local python reports that CHANGES uses a non-ascii code character and fail in running utils/update_build_version.py.	2019-08-15 08:46:59 -04:00
Jaebaek Seo	ff872dc6bf	Change the way to include header (#2795 ) `#include <source/util/string_utils.h>` works only when we specify `include_directories(${CMAKE_CURRENT_SOURCE_DIR}/)` in cmake. It is hard to set the source directory as a include path in some build systems e.g., bazel. Using the relative path easily solves this issue. This commit uses `#include "source/util/string_utils.h"` instead of `#include <source/util/string_utils.h>`.	2019-08-14 18:09:20 -04:00
alan-baker	bbd80462f5	Fix validation of constant matrices (#2794 ) Fixes #2793 * Don't special case matrix validation compared to other composites * just check the constituents are constants or undefs * later checking validates the column type * new test	2019-08-14 11:26:41 -04:00
Steven Perron	60043edfa1	Replace OpKill With function call. (#2790 ) We are no able to inline OpKill instructions into a continue construct. See #2433. However, we have to be able to inline to correctly do legalization. This commit creates a pass that will wrap OpKill instructions into a function of its own. That way we are able to inline the rest of the code. The follow up to this will be to not inline any function that contains an OpKill. Fixes #2726	2019-08-14 09:27:12 -04:00
Steven Perron	f701237f2d	Remove useless semi-colons (#2789 ) Later versions of clang seem to pick up more useless semi-colons. I've removed them.	2019-08-12 08:52:39 -04:00
greg-lunarg	95386f9e45	Instrument: Fix version 2 output record write for tess eval shaders. (#2782 ) Fix output record write for tess eval shaders. Also change command line for bindless instrumentation to use use output record version 2.	2019-08-09 08:22:41 -04:00
Steven Perron	22ce39c8e1	Start SPIRV-Tools v2019.5	2019-08-08 10:57:18 -04:00
Steven Perron	d65513e92c	Finalize SPIRV-Tools v2019.4	2019-08-08 10:55:48 -04:00
Steven Perron	4b64beb1ae	Add descriptor array scalar replacement (#2742 ) Creates a pass that will replace a descriptor array with individual variables. See #2740 for details. Fixes #2740.	2019-08-08 10:53:19 -04:00
Steven Perron	c26c2615f3	Update CHANGES	2019-08-08 10:52:01 -04:00
greg-lunarg	29af42df12	Add SPV_EXT_physical_storage_buffer to opt whitelists (#2779 ) This also fixes ADCE to not remove possibly needed OpTypeForwardPointer. The bug, its fix and the corresponding test have a circular dependency with the extension, so they are packaged together.	2019-08-08 09:45:59 -04:00
Steven Perron	b029d3697e	Handle RelaxedPrecision in SROA (#2788 ) If a member of a struct has a relaxed precision, sroa will not split the struct. This means we do not get all cases. This commit handles these cases. The other part is that the decoration needs to be passed on to the new variables. Fixes #2786	2019-08-07 12:17:26 -04:00
Ryan Harrison	370375d235	Add -fextra-semi to Clang builds (#2787 ) This will catch instances on the bots where PRs introduce unneeded semi-colons, which are going to cause downstream users problems. Fixes #2781	2019-08-07 11:09:55 -04:00
Alastair Donaldson	698b56a8f0	Add 'copy object' transformation (#2766 ) This transformation can introduce an instruction that uses OpCopyObject to make a copy of some other result id. This change introduces the transformation, but does not yet introduce a fuzzer pass to actually apply it.	2019-08-05 18:00:13 +01:00
Paul Thomson	4f14b4c8cc	fuzz: change output extension and fix usage string (#2778 )	2019-08-02 10:09:41 +01:00
Geoff Lang	0b70972a29	Remove extra ';' after member function definition. (#2780 ) This fixes a clang compiler warning about extra semicolons.	2019-08-01 19:33:55 -04:00
Ryan Harrison	5ada98d0bb	Update WebGPU validation rules of OpAtomic*s (#2777 ) Fixes #2723	2019-07-31 17:15:47 -04:00
alan-baker	3726b500b1	Treat access chain indexes as signed in SROA (#2776 ) Fixes #2768 * In scalar replacement, interpret access chain indexes as signed counts * Use Constant::GetSignExtendedValue and Constant::GetZeroExtendedValue where appropriate * new tests	2019-07-31 15:39:33 -04:00
David Neto	31590104ec	Add pass to inject code for robust-buffer-access semantics (#2771 ) spirv-opt: Add --graphics-robust-access Clamps access chain indices so they are always in bounds. Assumes: - Logical addressing mode - No runtime-array-descriptor-indexing - No variable pointers Adds stub code for clamping coordinate and samples for OpImageTexelPointer. Adds SinglePassRunAndFail optimizer test fixture. Android.mk: add source/opt/graphics_robust_access_pass.cpp Adds Constant::GetSignExtendedValue, Constant::GetZeroExtendedValue	2019-07-30 19:52:46 -04:00
Ryan Harrison	4a28259cc8	Update OpMemoryBarriers rules for WebGPU (#2775 ) Part of #2724	2019-07-30 14:50:55 -04:00
David Neto	7621034aae	Add opt test fixture method SinglePassRunAndFail (#2770 ) Checks for failure status code and matches against the expected error message.	2019-07-30 10:38:46 -04:00
David Neto	ac3d131054	Element type is const for analysis::Vector,Matrix,RuntimeArray (#2765 ) This makes it symmetric with the result type of ...->element_type which returns a const Type. So now we can write code like this: analysis::Vector v = ... analysis::Vector(v->element_type(), 2);	2019-07-29 22:55:18 -04:00
Diego Novillo	49797609b7	Protect against out-of-bounds references when folding OpCompositeExtract (#2774 ) This fixes #2608. The original test case had an out-of-bounds reference that ended up folding into OpCompositeExtract that was indexing right outside the constant composite. The returned constant would then cause a segfault during constant propagation.	2019-07-29 13:27:40 -07:00
alan-baker	7fd2365b06	Don't move debug or decorations when folding (#2772 ) Fixes #2764 * Don't replace all uses when simplifying instructions, instead only update non-debug, non-decoration uses * added a test * Add a new version of RAUW that takes a predicate to decide whether to replace the use or not * used in simplification pass	2019-07-29 16:20:43 -04:00
Ryan Harrison	7bafeda284	Update OpControlBarriers rules for WebGPU (#2769 ) * Update OpControlBarriers rules for WebGPU Part of #2724	2019-07-29 12:53:27 -04:00
Diego Novillo	9559cdbdf0	Fix #2609 - Handle out-of-bounds scalar replacements. (#2767 ) * Fix #2609 - Handle out-of-bounds scalar replacements. When SROA tries to do a replacement for an OpAccessChain that is exactly one element out of bounds, the code was trying to access its internal array of replacements and segfaulting. This protects the code from doing this, and it additionally fixes the way SROA works by not returning failure when it refuses to do a replacement. Instead of failing the optimization pass, SROA will now simply refuse to do the replacement and keep going. Additionally, this patch fixes the SROA logic to now return a proper status so we can correctly state that the pass made no changes to the IR if it only found invalid references.	2019-07-26 12:33:40 -04:00
Alastair Donaldson	f54b8653dd	Limit fuzzer tests so that they take less time to run (#2763 ) The recently added fuzzer_replayer and fuzzer_shrinker tests were rather heavyweight and were leading to CI timeouts. This change reduces the runtime of those tests by having them do fewer iterations.	2019-07-25 13:09:49 -04:00
Steven Perron	bb0e2f65bb	Fix check for unreachable blocks in merge-return (#2762 ) Merge return expects unreachable merge block to look a certain way, and unreachable continue blocks to look a certain way. What if an unreachable block is both a merge and a continue? The continue is suppose to take precedent, but merge-return implements it with the merge taking precedent. This change flips that around. Fixes #2746	2019-07-25 09:34:18 -04:00

1 2 3 4 5 ...

2223 Commits