SPIRV-Tools

mirror of https://github.com/KhronosGroup/SPIRV-Tools synced 2024-11-28 14:11:04 +00:00

Author	SHA1	Message	Date
Steven Perron	bdc2ab9356	In LICM don't place code between merge instruction and branch. (#2252 ) Fixes #2210.	2018-12-20 18:33:52 -05:00
Steven Perron	5ec2d1a8cd	Don't fold specialized branches in loop unswitch (#2245 ) * Don't fold specialized branchs in loop unswitch Folding branches can have a lot of special cases, and can be a little error prone. So I only want it in one place. That will be in dead branch elimination. I will change loop unswitching to set the branches that were being folded to have a constant condition. Then subsequent pass of dead branch elimination will be able to remove the code. At the same time, I added a check that loop unswitching will not unswitch a branch with a constant condition. It is not useful to do it because dead branch elimination will simple fold the branch anyway. Also it avoid an infinite loop that would other wise be introduced by my first change. Fixes #2203.	2018-12-19 04:40:30 +00:00
Steven Perron	1254335d13	Don't unswitch the latch block. (#2205 ) Loop unswitching is unswitching the conditional branch that creates the back-edge. In the version of the loop, where the bachedge is not taken, there is no back-edge. This is what causes the validator to complain. The solution I will go with will be to now unswitch a condition with a back-edge. At this time we do not now if loop unswitching is used. We do not include it in the optimization sets provided, nor is it used in glslang's set. When there are opportunities and no breaks from the loop, the loop with either be a single iteration loop, or an infinite loop. There is no performance advantage to performing loop unswitching in either of those cases. If there is a break, maintaining structured control flow will be tricky. Unless we see a clear advantage to handling these case, I would go with the safer simpler solution. Fixes #2201.	2018-12-18 18:15:00 +00:00
Steven Perron	2e4563d94f	Document in the context what happens with id overflow. (#2159 ) Added documentation to the ir context to indicates that TakeNextId() returns 0 when the max id is reached. TODOs were added to each call sight so that we know where we have to start to handle this case. Handle id overflow in \|SplitLoopHeader\|. Handle id overflow in \|GetOrCreatePreHeaderBlock\|. Handle failure to create preheader in LICM. Part of https://github.com/KhronosGroup/SPIRV-Tools/issues/1841.	2018-12-06 09:07:00 -05:00
greg-lunarg	1e9fc1aac1	Add base and core bindless validation instrumentation classes (#2014 ) * Add base and core bindless validation instrumentation classes * Fix formatting. * Few more formatting fixes * Fix build failure * More build fixes * Need to call non-const functions in order. Specifically, these are functions which call TakeNextId(). These need to be called in a specific order to guarantee that tests which do exact compares will work across all platforms. c++ pretty much does not guarantee order of evaluation of operands, so any such functions need to be called separately in individual statements to guarantee order. * More ordering. * And more ordering. * And more formatting. * Attempt to fix NDK build * Another attempt to address NDK build problem. * One more attempt at NDK build failure * Add instrument.hpp to BUILD.gn * Some name improvement in instrument.hpp * Change all types in instrument.hpp to int. * Improve documentation in instrument.hpp * Format fixes * Comment clean up in instrument.hpp * imageInst -> image_inst * Fix GetLabel() issue.	2018-11-08 13:54:54 -05:00
Jeff Bolz	60fac96c6b	Enable precompiled headers for spirv-tools(-shared) and some unit tests (#2026 )	2018-11-06 09:26:23 -05:00
alan-baker	fae1e61ab8	Fix bug in construct block calculation (#1964 ) Fixes #1960 * Only allows blocks that are dominated by the header * Fixed a bad loop fusion test * Added a test derived from the reported bug	2018-10-10 11:14:01 -04:00
Steven Perron	c4c68712c4	Make EFFCEE required (#1943 ) Fixes #1912. Remove the non-effcee build as EFFCEE is now required.	2018-10-04 10:00:11 -04:00
Lei Zhang	9bfe0eb25e	Wrap tests needing effcee inside SPIRV_EFFCEE	2018-09-20 11:51:40 -04:00
Steven Perron	9fbcce4ca1	Add unrolling to the legalization passes (#1903 ) Adds unrolling to the legalization passes. After enabling unrolling I found a bug when there is a self-referencing phi node. That has been fixed. The test that checks for that the order of optimizations is correct also needed to be updated.	2018-09-19 16:40:09 -04:00
dan sinclair	eda2cfbe12	Cleanup includes. (#1795 ) This Cl cleans up the include paths to be relative to the top level directory. Various include-what-you-use fixes have been added.	2018-08-03 15:06:09 -04:00
dan sinclair	7861df9bb3	Remove using namespace commands. (#1794 ) This CL removes the last two 'using namespace' commands.	2018-08-03 08:05:52 -04:00
dan sinclair	e477e7573e	Remove the module from opt::Function. (#1717 ) The function class provides a {Set\|Get}Parent call in order to provide the context to the LoopDescriptor methods. This CL removes the module from Function and provides the needed context directly to LoopDescriptor on creation.	2018-07-12 14:42:05 -04:00
dan sinclair	f96b7f1cb9	use Pass::Run to set the context on each pass. (#1708 ) Currently the IRContext is passed into the Pass::Process method. It is then up to the individual pass to store the context into the context_ variable. This CL changes the Run method to store the context before calling Process which no-longer receives the context as a parameter.	2018-07-12 09:08:45 -04:00
dan sinclair	2cce2c5b97	Move tests into namespaces (#1689 ) This CL moves the test into namespaces based on their directories.	2018-07-11 09:24:49 -04:00
Sean Purcell	9532aede29	Fix unused param errors when Effcee not present	2018-07-10 11:57:42 -04:00
dan sinclair	a3e3869540	Convert validation to use libspriv::Instruction where possible. (#1663 ) For the instructions which execute after the IdPass check we can provide the Instruction instead of the spv_parsed_instruction_t. This Instruction class provides a bit more context (like the source line) that is not available from spv_parsed_instruction_t.	2018-07-10 10:57:52 -04:00
dan sinclair	e6b953361d	Move the ir namespace to opt. (#1680 ) This CL moves the files in opt/ to consistenly be under the opt:: namespace. This frees up the ir:: namespace so it can be used to make a shared ir represenation.	2018-07-09 11:32:29 -04:00
Corentin Wallez	ba602c9059	Add a WIP WebGPU environment. It disallows OpUndef Add SPV_ENV_WEBGPU_0 for work-in-progress WebGPU. val: Disallow OpUndef in WebGPU env Silence unused variable warnings when !defined(SPIRV_EFFCE) Limit visibility of validate_instruction.cpp's symbols Only InstructionPass needs to be visible so all other functions are put in an anonymous namespace inside the libspirv namespace.	2018-06-21 15:53:15 -04:00
David Neto	700ebd3442	Make fewer test executables Try to reduce the amount of disk space used by especially by debug builds, which may be contributing to AppVeyor failures. Collapses tests in categories: - validator - loop optimizations - dominator analysis - linker Contributes to #1615	2018-06-12 09:48:42 -04:00
alan-baker	18ad1be7f9	Fixing MacOS compiler error	2018-05-15 12:23:27 -04:00
Stephen McGroarty	1c2cbaf569	Add GetContinueBlock to loop class. Previously, the loop class used the terms latch and continue block interchangeably. This patch splits the two and corrects and tests some uses of the old uses of GetLatchBlock.	2018-05-03 14:30:41 -04:00
Toomas Remmelg	1dc2458060	Add a loop fusion pass. This pass will look for adjacent loops that are compatible and legal to be fused. Loops are compatible if: - they both have one induction variable - they have the same upper and lower bounds - same initial value - same condition - they have the same update step - they are adjacent - there are no break/continue in either of them Fusion is legal if: - fused loops do not have any dependencies with dependence distance greater than 0 that did not exist in the original loops. - there are no function calls in the loops (could have side-effects) - there are no barriers in the loops It will fuse all such loops as long as the number of registers used for the fused loop stays under the threshold defined by max_registers_per_loop.	2018-05-01 15:40:37 -04:00
Stephen McGroarty	9a5dd6fe88	Support loop fission. Adds support for spliting loops whose register pressure exceeds a user provided level. This pass will split a loop into two or more loops given that the loop is a top level loop and that spliting the loop is legal. Control flow is left intact for dead code elimination to remove. This pass is enabled with the --loop-fission flag to spirv-opt.	2018-05-01 15:15:10 -04:00
David Neto	7a59283587	Another fix for old XCode: std::set explicit ctor in test code	2018-04-20 15:58:01 -04:00
GregF	1c89da46ff	Test/DependencyAnalysis: Fix uninitialized variables	2018-04-19 15:34:15 -04:00
Toomas Remmelg	0f335cf87e	Add support for MIV and Delta test dependence analysis. GCD MIV test as described in Chapter 3 of "Optimizing Compilers for Modern Architectures: A Dependence-Based Approach" by Randy Allen, and Ken Kennedy. Delta test as described in Figure 3 of "Practical Dependence Testing" by Gina Goff, Ken Kennedy, and Chau-Wen Tseng from PLDI '91.	2018-04-17 13:57:02 -04:00
Victor Lomuller	10e5d7cf13	Add a loop peeling pass. For each loop in a function, the pass walks the loops from inner to outer most loop and tries to peel loop for which a certain amount of iteration can be done before or after the loop. To limit code growth, peeling will not happen if the growth in code size goes above a configurable threshold.	2018-04-11 15:41:29 +01:00
Alexander Johnston	61b50b3bfa	ZIV and SIV loop dependence analysis. Provides functionality to perform ZIV and SIV dependency analysis tests between a load and store within the same loop. Dependency tests rely on scalar analysis to prove and disprove dependencies with regard to the loop being analysed. Based on the 1990 paper Practical Dependence Testing by Goff, Kennedy, Tseng Adds support for marking loops in the loop nest as IRRELEVANT. Loops are marked IRRELEVANT if the analysed instructions contain no induction variables for the loops, i.e. the loops induction variable is not relevent to the dependence of the store and load.	2018-04-11 09:32:42 -04:00
Victor Lomuller	bdf421cf40	Add loop peeling utility The loop peeler util takes a loop as input and create a new one before. The iterator of the duplicated loop then set to accommodate the number of iteration required for the peeling. The loop peeling pass that decided to do the peeling and profitability analysis is left for a follow-up PR.	2018-03-20 10:21:10 -04:00
Steven Perron	b3daa93b46	Change merge return pass to handle structured cfg. We are seeing shaders that have multiple returns in a functions. These functions must get inlined for legalization purposes; however, the inliner does not know how to inline functions that have multiple returns. The solution we will go with it to improve the merge return pass to handle structured control flow. Note that the merge return pass will assume the cfg has been cleanedup by dead branch elimination. Fixes #857.	2018-03-19 13:49:04 -04:00
Victor Lomuller	3497a94460	Add loop unswitch pass. It moves all conditional branching and switch whose conditions are loop invariant and uniform. Before performing the loop unswitch we check that the loop does not contain any instruction that would prevent it (barriers, group instructions etc.).	2018-02-27 08:52:46 -05:00
Stephen McGroarty	e354984b09	Unroller support for multiple induction variables Support for multiple induction variables within a loop and support for loop condition operands <= and >=.	2018-02-27 11:50:08 +00:00
Stephen McGroarty	dd8400e150	Initial support for loop unrolling. This patch adds initial support for loop unrolling in the form of a series of utility classes which perform the unrolling. The pass can be run with the command spirv-opt --loop-unroll. This will unroll loops within the module which have the unroll hint set. The unroller imposes a number of requirements on the loops it can unroll. These are documented in the comments for the LoopUtils::CanPerformUnroll method in loop_utils.h. Some of the restrictions will be lifted in future patches.	2018-02-14 15:44:38 -05:00
Alexander Johnston	84ccd0b9ae	Loop invariant code motion initial implementation	2018-02-08 22:55:47 -05:00
Victor Lomuller	50e85c865c	Add LoopUtils class to gather some loop transformation support. This patch adds LoopUtils class to handle some loop related transformations. For now it has 2 transformations that simplifies other transformations such as loop unroll or unswitch: - Dedicate exit blocks: this ensure that all exit basic block (out-of-loop basic blocks that have a predecessor in the loop) have all their predecessors in the loop; - Loop Closed SSA (LCSSA): this ensure that all definitions in a loop are used inside the loop or in a phi instruction in an exit basic block. It also adds the following capabilities: - Loop::IsLCSSA to test if the loop is in a LCSSA form - Loop::GetOrCreatePreHeaderBlock that can build a loop preheader if required; - New methods to allow on the fly updates of the loop descriptors. - New methods to allow on the fly updates of the CFG analysis. - Instruction::SetOperand to allow expression of the index relative to Instruction::NumOperands (to be compatible with the index returned by DefUseManager::ForEachUse)	2018-02-01 15:35:09 -05:00
Victor Lomuller	6018de81de	Add LoopDescriptor as an IRContext analysis. Move some function definitions from header to source to avoid circular definition.	2018-01-25 16:12:32 -05:00
Victor Lomuller	e8ad02f3dd	Add loop descriptors and some required dominator tree extensions. Add post-order tree iterator. Add DominatorTreeNode extensions: - Add begin/end methods to do pre-order and post-order tree traversal from a given DominatorTreeNode Add DominatorTree extensions: - Add begin/end methods to do pre-order and post-order tree traversal - Tree traversal ignore by default the pseudo entry block - Retrieve a DominatorTreeNode from a basic block Add loop descriptor: - Add a LoopDescriptor class to register all loops in a given function. - Add a Loop class to describe a loop: - Loop parent - Nested loops - Loop depth - Loop header, merge, continue and preheader - Basic blocks that belong to the loop Correct a bug that forced dominator tree to be constantly rebuilt.	2018-01-08 09:31:13 -05:00

38 Commits