SPIRV-Tools

mirror of https://github.com/KhronosGroup/SPIRV-Tools synced 2024-12-26 09:41:03 +00:00

Author	SHA1	Message	Date
Steven Perron	8df947d2d6	Handle instructions not in blocks in code sinking. (#2308 ) When looking at the uses of the result of an instruction, code sinking assumes that all uses are in a basic block. However, this is not true if there is a decoration or name for the result of that insturction. This commit checks for this. Fixes https://crbug.com/923243.	2019-01-21 12:09:56 -05:00
Steven Perron	d6c067630d	Handle extract with no index in VDCE. (#2305 ) It is legal, but not generated by any SPIR-V producer: an OpCompositeExtract with no indexes. This is essentially just a copy of the object, so we treat them that way. We simply propagate the live variables of the result to the operand. Fixes https://crbug.com/919181.	2019-01-18 15:43:36 -05:00
Steven Perron	81fb2649bf	Handle access chain with no index in SROA. (#2304 ) It is legal, but not generated by any SPIR-V producer: an OpAccessChain with no indexes. This is essentially just a copy of the pointer. I have decided to treat it like an OpCopyObject. In CheckUses, we return that it is not okay. When looking at this I realized that we had code in GetUsedComponents that cannot be reached. If there is a use in an OpCopyObject the it will not call GetUsedComponents. I removed that dead code. Fixes https://crbug.com/918311.	2019-01-18 14:19:43 -05:00
Steven Perron	213e15e100	Fix overflow when negating INT_MIN. (#2293 ) When doing (-INT_MIN) is considered overflow, so we cannot fold it by actually performing the negation. Fixes https://crbug.com/917991	2019-01-17 17:01:55 -05:00
Steven Perron	99c2c21cf4	Fix memory leak in unrolling. (#2301 ) During unrolling a new loop is created, but its ownership is not clear as it gets passed through the code. Changed something to unique_ptr to make that clearer. Fixes #2299. Fixing other memory leaks at the same time. Fixes #2296 Fixes #2297	2019-01-17 16:02:43 -05:00
Steven Perron	dd4157dcee	Sink (#2284 ) Add code sinking pass. It will move OpLoad and OpAccessChain instructions as close as possible to their uses. Part of #1611.	2019-01-17 15:56:36 -05:00
greg-lunarg	8d2d66f30c	Fix vertex instrumentation to use VertexIndex and InstanceIndex (#2294 ) ...instead of VertexId and InstanceId	2019-01-16 18:02:07 -05:00
Steven Perron	49b5b0abc6	Fix up bit shifts by 32. (#2292 ) In C++, a bit shift of the same size as the type is undefined, but it is defined in spir-v. When folding those cases, we have to be careful. We cannot simply do the shift in C++. Fixes https://crbug.com/917697.	2019-01-16 15:52:23 -05:00
greg-lunarg	83bfdc976a	Instrumentation: Add ArrayStride decoration to debug output buffer array (#2290 )	2019-01-16 10:01:40 -05:00
alan-baker	06c9dc07bd	Upgrade modf and frexp (#2266 ) Fixes #2138 * Modf and frexp are upgraded to use the struct version of the instruction and generate an explicit store whose flags can be upgraded separately * Fixed major bug where availability and visibility were reversed for non-copy memory instructions * Fixed bug where availability and visibility scope operands were reversed for copy memory * Upgraded all opt tests to use SPV_ENV_UNIVERSAL_1_3 * Upgrade tests moved into unified tests and removed standalone test	2019-01-07 12:36:38 -05:00
Steven Perron	241644a5a3	Have replace load size handle extact with no index. (#2261 ) Fixes https://crbug.com/917774	2019-01-03 13:02:10 -05:00
Steven Perron	9f36c8bb72	Handle CompositeInsert with no indices in VDCE (#2258 ) * Handle CompositeInsert with no indices in VDCE In the spec, there it nothing that forces an OpCompositeInsert to have an index, but VDCE assumes there is at least 1 in a couple places. This commit updates VDCE to handle these cases.	2019-01-02 14:00:04 -05:00
Steven Perron	bdc2ab9356	In LICM don't place code between merge instruction and branch. (#2252 ) Fixes #2210.	2018-12-20 18:33:52 -05:00
kholtnv	e49bd96f2c	Added additional changes for the new AccelerationStructureNV type. (#2218 ) * Added additional changes for the new AccelerationStructureNV type. * Added additional changes for the new AccelerationStructureNV type. Change tabs to space... * Added additional changes for the new accelerationStructureNV type -- add proper type name. Fix TypeManager.TypeStrings test: [----------] 29 tests from TypeManager [ RUN ] TypeManager.TypeStrings [ OK ] TypeManager.TypeStrings (7 ms)	2018-12-19 21:42:39 +00:00
Steven Perron	68b69e16aa	Update the continue target in merge return. (#2249 ) When we are predicating the continue target for a loop, it can no longer be the continue target because it will have a branch that exits the loop and is not the bach edge. The continue target will have to be the target of that branch that is still in the loop. Fixes #2211.	2018-12-19 21:24:49 +00:00
Steven Perron	ac7feace90	Fix missing OpPhi after merge return. (#2248 ) The function `UpdatePhiNodes` was being called inconsistently. In one case, the cfg had already been updated to include the new edge, and in another place the cfg was not updated. This caused the function to miss flagging a block as needing new phi nodes. I picked that the cfg should not be updated before making the call. I documented it, and change the call sites to match. Fixes #2207.	2018-12-19 18:17:42 +00:00
Steven Perron	9d04f82bef	Ensure SROA gets the correct pointer type. (#2247 ) We initially assumed that if the type manager returned the correct id for the pointee type, that we would get the correct pointer type back, but that is not true. See the unit test added with this commit. We need to fall back to the linear search any time we are looking for a pointer to a type that may not be unique. At the same time, SROA considered an OpName on a variable to be a use of the entire variable. That has been fixed. Fixes #2209.	2018-12-19 17:07:29 +00:00
Steven Perron	9e81c337f9	Place load after OpPhi instructions in block. (#2246 ) We currently place the load instructions at the start of the basic block that dominates all of the loads. If that basic block contains OpPhi instructions, then this will generate invalid code. We just need to search for a location that comes after all of the OpPhi instructions. Fixes #2204.	2018-12-19 15:18:22 +00:00
Steven Perron	5ec2d1a8cd	Don't fold specialized branches in loop unswitch (#2245 ) * Don't fold specialized branchs in loop unswitch Folding branches can have a lot of special cases, and can be a little error prone. So I only want it in one place. That will be in dead branch elimination. I will change loop unswitching to set the branches that were being folded to have a constant condition. Then subsequent pass of dead branch elimination will be able to remove the code. At the same time, I added a check that loop unswitching will not unswitch a branch with a constant condition. It is not useful to do it because dead branch elimination will simple fold the branch anyway. Also it avoid an infinite loop that would other wise be introduced by my first change. Fixes #2203.	2018-12-19 04:40:30 +00:00
Steven Perron	1254335d13	Don't unswitch the latch block. (#2205 ) Loop unswitching is unswitching the conditional branch that creates the back-edge. In the version of the loop, where the bachedge is not taken, there is no back-edge. This is what causes the validator to complain. The solution I will go with will be to now unswitch a condition with a back-edge. At this time we do not now if loop unswitching is used. We do not include it in the optimization sets provided, nor is it used in glslang's set. When there are opportunities and no breaks from the loop, the loop with either be a single iteration loop, or an infinite loop. There is no performance advantage to performing loop unswitching in either of those cases. If there is a break, maintaining structured control flow will be tricky. Unless we see a clear advantage to handling these case, I would go with the safer simpler solution. Fixes #2201.	2018-12-18 18:15:00 +00:00
Steven Perron	ff07c6df83	SSA-rewriter: make sure phi entries are unique. (#2206 ) If there are multiple edges to a basic block, then the ssa rewriter will create OpPhi instructions with duplicate entries. This is invalid, and it is fixed in this commit. Fixes #2202.	2018-12-18 18:14:27 +00:00
Jeff Bolz	24328a0554	Recognize OpTypeAccelerationStructureNV as a type instruction (#2190 )	2018-12-11 19:03:55 -05:00
Steven Perron	e07dabc25f	Invalidate the decoration manager at the start of ADCE. (#2189 ) * Invalidate the decoration manager at the start of ADCE. If the decoration manager is kept live the the contex will try to keep it up to date. ADCE deals with group decorations by changing the operands in \|OpGroupDecorate\| instructions directly without informing the decoration manager. This puts it in an invalid state, which will cause an error when the context tries to update it. To Avoid this problem, we will invalidate the decoration manager upfront. At the same time, the decoration manager is now considered when checking the consistency of the decoration manager.	2018-12-10 13:24:33 -05:00
Jeff Bolz	8fc8dfe4a5	Add PCH_FILE for upgrade_memory_model target. MSVC doesn't like building pass_utils.cpp twice in the same folder with different PCH settings. (#2186 )	2018-12-10 10:54:19 -05:00
Steven Perron	0bc66a8ba9	Fix invalid OpPhi generated by merge-return. (#2172 ) * Fix invalid OpPhi generated by merge-return. When we create a new phi node for a value say %10, we have to replace all of the uses of %10 that are no longer dominated by the def of %10 by the result id of the new phi. However, if the use is in a phi node, it is possible that the bb contains the use is not dominated by either. In this case, needs to be handled differently. * Split loop headers before add a new branch to them. In merge return, Phi node in loop header that are also merges for loop do not get updated correctly. Those cases do not fit in with our current analysis. Doing this will simplify the code by reducing the number of cases that have to be handled.	2018-12-07 14:10:30 -05:00
David Neto	6df6194db8	Validate Uniform decoration (#2181 )	2018-12-07 09:32:57 -05:00
Steven Perron	2e4563d94f	Document in the context what happens with id overflow. (#2159 ) Added documentation to the ir context to indicates that TakeNextId() returns 0 when the max id is reached. TODOs were added to each call sight so that we know where we have to start to handle this case. Handle id overflow in \|SplitLoopHeader\|. Handle id overflow in \|GetOrCreatePreHeaderBlock\|. Handle failure to create preheader in LICM. Part of https://github.com/KhronosGroup/SPIRV-Tools/issues/1841.	2018-12-06 09:07:00 -05:00
Steven Perron	17cba4695c	Remove undefined behaviour when folding shifts. (#2157 ) We currently simulate all shift operations when the two operand are constants. The problem is that if the shift amount is larger than 32, the result is undefined. I'm changing the folder to return 0 if the shift value is too high. That way, we will have defined behaviour. https://crbug.com/910937.	2018-12-04 10:04:02 -05:00
alan-baker	e510b1bac5	Update memory model (#1904 ) Upgrade to VulkanKHR memory model * Converts Logical GLSL450 memory model to Logical VulkanKHR * Adds extension and capability * Removes deprecated decorations and replaces them with appropriate flags on downstream instructions * Support for Workgroup upgrades * Support for copy memory * Adding support for image functions * Adding barrier upgrades and tests * Use QueueFamilyKHR scope instead of device	2018-11-30 14:15:51 -05:00
Steven Perron	2d2a512691	Don't inline recursive functions. (#2130 ) * Move ProcessFunction* function from pass to the context. There are a few functions that are used to traverse the call tree. They currently live in the Pass class, but they have nothing to do with a pass, and may be needed outside of a pass. They would be better in the ir context, or in a specific call tree class if we ever have a need for it. * Don't inline recursive functions. Inlining does not check if a function is recursive or not. This has been fine as long as the shader was a Vulkan shader, which forbid recursive functions. However, not all shaders are vulkan, so either we limit inlining to Vulkan shaders or we teach it to look for recursive functions. I prefer to keep the passes as general as is reasonable. The change does not require much new code in inlining and gives a reason to refactor some other code. The changes are to add a member function to the Function class that checks if that function is recursive or not. Then this is used in inlining to not inlining a function call if it calls a recursive function. * Add id to function analysis There are a few places that build a map from ids to Function whose result is that id. I decided to add an analysis to the context for this to reduce that code, and simplify some of the functions. * Add missing file.	2018-11-29 14:24:58 -05:00
alan-baker	3d56cddb75	Validate pointer variables (#2111 ) Fixes #2104 * Checks the rules for logical addressing and variable pointers * Has an out for relaxed logical pointers * Updated PassFixture to expose validator options * enabled relaxed logical pointers for some tests * New validator tests	2018-11-27 16:47:10 -05:00
Ryan Harrison	d7cd1203a4	Ensure for OpVariable that result type and storage class operand agree (#2052 ) From SPIR-V spec, section 3.32.8 on OpVariable: Its Storage Class operand must be the same as the Storage Class operand of the result type. Fixes #941	2018-11-16 11:22:11 -05:00
greg-lunarg	c37388f1ad	Add passes to propagate and eliminate redundant line instructions (#2027 ). (#2039 ) These are bookend passes designed to help preserve line information across passes which delete, move and clone instructions. The propagation pass attaches a debug line instruction to every instruction based on SPIR-V line propagation rules. It should be performed before optimization. The redundant line elimination pass eliminates all line instructions which match the previous line instruction. This pass should be performed at the end of optimization to reduce physical SPIR-V file size. Fixes #2027.	2018-11-15 14:06:17 -05:00
Steven Perron	dc9d155d62	Fix folding of volatile store. (#2048 ) When looking for the Volatile mask on a store, the instruction folder accesses an out-of-bounds element. We fix that up. Fixes crbug.com/903530.	2018-11-14 13:52:18 -05:00
Steven Perron	ec5574a9c6	Instruction::GetBaseAddress to handle OpPtrAccessChain (#2050 ) That function currently only handled OpPtrAccessChain if it was in the middle of the chain, but not at the start. Fixing that up. Fixes crbug.com/905271.	2018-11-14 12:42:25 -05:00
greg-lunarg	1e9fc1aac1	Add base and core bindless validation instrumentation classes (#2014 ) * Add base and core bindless validation instrumentation classes * Fix formatting. * Few more formatting fixes * Fix build failure * More build fixes * Need to call non-const functions in order. Specifically, these are functions which call TakeNextId(). These need to be called in a specific order to guarantee that tests which do exact compares will work across all platforms. c++ pretty much does not guarantee order of evaluation of operands, so any such functions need to be called separately in individual statements to guarantee order. * More ordering. * And more ordering. * And more formatting. * Attempt to fix NDK build * Another attempt to address NDK build problem. * One more attempt at NDK build failure * Add instrument.hpp to BUILD.gn * Some name improvement in instrument.hpp * Change all types in instrument.hpp to int. * Improve documentation in instrument.hpp * Format fixes * Comment clean up in instrument.hpp * imageInst -> image_inst * Fix GetLabel() issue.	2018-11-08 13:54:54 -05:00
greg-lunarg	6721478ef1	Don't assume one return means function can be inlined. (#2018 ) (#2025 ) If there is only 1 return and it is in a loop, then the function cannot be inlined. Fix condition when inlined code needs one-trip loop wrapper. The dummy loop is needed when there is a return inside a selection construct. Even if there is only 1 return.	2018-11-08 09:11:20 -05:00
Jeff Bolz	c06a35b902	Rename PCH macro to spvtools_pch to avoid conflicts with other projects. Also add pch to test/opt. (#2034 )	2018-11-07 09:15:04 -05:00
Jeff Bolz	60fac96c6b	Enable precompiled headers for spirv-tools(-shared) and some unit tests (#2026 )	2018-11-06 09:26:23 -05:00
Steven Perron	f2cc71e5cb	Handle OpMemberDecorateStringGOOGLE in ACDE (#2029 ) Add missing case to the switch statement for the annotation instructions. See https://github.com/KhronosGroup/glslang/issues/1561.	2018-11-02 13:42:45 -04:00
dan sinclair	9e6f5134d1	Reduce number of test targets (#2024 ) This CL takes the various opt unit tests and makes a single executable instead of one per test. This reduces the number of build targets by ~125 when building with ninja.	2018-11-01 10:19:37 -04:00
Steven Perron	6647884a13	Remove MemberDecorateStringGOOGLE during stript-refect. (#2021 ) The strip-reflect pass is not removing the reflection decorations that are decorating members. With this commit, they will now be removed. Fixes #2019.	2018-10-30 16:17:35 -04:00
Steven Perron	18fe6d59e5	Fix dead branch elim infinite loop. (#2009 ) When looking for a break from a selection construct, we do not realize that a jump to the continue target of a loop containing the selection is a break. This causes and infinit loop, or possibly other failures. Fixes #2004.	2018-10-24 09:10:30 -04:00
Steven Perron	0ba35798c3	Fix dead branch elim infinite loop. (#1997 ) When looking for a break from a selection construct, we do not need to look inside nested constructs. However, if a loop header has an unconditional branch, then we enter the loop. Entering the loop causes an infinite loop because we keep going through the loop. The solution is to look for a merge block, if one exsits, even for block terminated by an OpBranch. Fixes #1979.	2018-10-22 13:59:20 -04:00
alan-baker	6e85d1a6fc	Fix restrictions in if conversion (#1998 ) Fixes #1991 * Improved identification of potential conditional branches * Pass changed to only work for shaders * added a test to catch the bug	2018-10-19 15:16:46 -04:00
Steven Perron	715afb0cea	Add a nullptr check to array copy propagation. (#1987 ) We are missing a check for a nullptr that is causing things to fail. Added an extra test case, and fixed up others. This is the fix for https://github.com/Microsoft/DirectXShaderCompiler/issues/1598.	2018-10-19 12:53:40 -04:00
greg-lunarg	c4687889b7	Fix ADCE to treat OpUnreachable correctly during liveness analysis (#1984 ) ADCE liveness algorithm should treat OpUnreachable at least like other branch instructions. It was being treated as always live which was preventing useless structured constructs from being eliminated. OpUnreachable is generated by dead branch elimination which is now being required by merge return, so this fix should accompany that change.	2018-10-19 10:16:35 -04:00
Steven Perron	0e68bb3632	Only run merge-returnon reachable functions. (#1983 ) We currently run merge-return on all functions, but dead-branch-elimination only runs on function reachable from an entry point or exported function. Since dead-branch-elimination is needed for merge-return, they have to match. Fixes #1976.	2018-10-18 08:48:27 -04:00
greg-lunarg	ab45d69154	Fix ADCE liveness to include all enclosing control structures. (#1975 ) Was removing control structures which didn't have data dependency with enclosed live loop and otherwise did not contain live code. An example is a counting loop around a live loop. Fixes #1967.	2018-10-16 08:00:07 -04:00
Nuno Subtil	5bc30788fd	Fix gtest.h include in test/opt/pass_utils.h Fixes builds where googletest is outside the SPIRV-Tools tree.	2018-10-12 10:22:25 -04:00

1 2 3 4 5 ...

400 Commits