SPIRV-Tools

mirror of https://github.com/KhronosGroup/SPIRV-Tools synced 2024-11-23 04:00:05 +00:00

Author	SHA1	Message	Date
David Neto	886859159e	Fix generation of Vim syntax file	2018-02-09 17:47:51 -05:00
Józef Kucia	4e4a254bc8	Do not hardcode libdir and includedir in pkg config files	2018-02-09 12:43:03 +01:00
Steven Perron	1a849ffb60	Add header files missing from CMakeLists.txt	2018-02-08 23:02:22 -05:00
Alexander Johnston	84ccd0b9ae	Loop invariant code motion initial implementation	2018-02-08 22:55:47 -05:00
GregF	ca4457b4b6	SROA: Do replacement on structs with no partial references.	2018-02-08 15:20:02 -05:00
Steven Perron	06cdb96984	Make use of the instruction folder. Implementation of the simplification pass. - Create pass that calls the instruction folder on each instruction and propagate instructions that fold to a copy. This will do copy propagation as well. - Did not use the propagator engine because I want to modify the instruction as we go along. - Change folding to not allocate new instructions, but make changes in place. This change had a big impact on compile time. - Add simplification pass to the legalization passes in place of insert-extract elimination. - Added test cases for new folding rules. - Added tests for the simplification pass - Added a method to the CFG to apply a function to the basic blocks in reverse post order. Contributes to #1164.	2018-02-07 23:01:47 -05:00
Andrey Tuganov	a61e4c1356	Disable check which fails Vulkan CTS	2018-02-07 13:31:35 -05:00
Andrey Tuganov	2f0c3aaa11	Add Vulkan-specific validation rules for atomics Added atomic instructions validation rules from https://www.khronos.org/registry/vulkan/specs/1.0/html/vkspec.html#spirvenv-module-validation	2018-02-07 13:31:35 -05:00
Józef Kucia	3013897556	Build SPIRV-Tools as shared library Add pkg-config file for shared libraries Properly build SPIRV-Tools DLL Test C interface with shared library Set PATH to shared library file for c_interface_shared test Otherwise, the test won't find SPIRV-Tools-shared.dll. Do not use private functions when testing with shared library Make all symbols hidden by default for shared library target	2018-02-07 10:43:32 -05:00
David Neto	b1c9c4e8c0	Enable Visual Studio 2013 again Disable use of Effcee and RE2 with MSVC compilers older than Visual Studio 2015 since RE2 doesn't support them.	2018-02-06 14:40:28 -05:00
David Neto	e7fafdaa68	Fix test inclusion when Effcee is absent	2018-02-06 12:10:50 -05:00
David Neto	c452bfd054	Update CHANGES	2018-02-06 11:47:44 -05:00
Alan Baker	871022772e	Registering a type now rebuilds it out of memory owned by the manager. * Added TypeManager::RebuildType * rebuilds the type and its constituent types in terms of memory owned by the manager. * Used by TypeManager::RegisterType to properly allocate memory * Adding an unit test to expose the issue * Added some tests to provide coverage of RebuildType * Added an accessor to the target pointer for a forward pointer	2018-02-06 10:17:56 -05:00
GregF	860b2ee5fc	ADCE: Fix combinator initialization The combinator initialization was only looking at the capabilities in the shader and not the inferred capabilities. Geometry and tessellation shaders were not setting the Shader capability which is inferred. So the combinator set was not initialized correctly causing problems for ADCE.	2018-02-05 16:54:03 -05:00
David Neto	9e19fc0f31	VS2013: LoopDescriptor LoopContainerType can't contain unique_ptr The loop descriptor must explicitly manage the storage for contained Loop objects. Fixes #1262	2018-02-05 14:19:21 -05:00
Andrey Tuganov	12e6860d07	Add barrier instructions validation pass	2018-02-05 13:14:55 -05:00
David Neto	3ef4bb600f	Avoid vector copies in range-for loops in opt/types.cpp Also be more explicit about iterated types in other range-for loops.	2018-02-05 13:08:39 -05:00
David Neto	87f9cfaba3	Disambiguate between const and nonconst ForEachSuccessorLabel This helps VisualStudio 2013 compile the code. Contributes to #1262	2018-02-02 17:54:40 -05:00
Steven Perron	bc1ec9418b	Add general folding infrastructure. Create the folding engine that will 1) attempt to fold an instruction. 2) iterates on the folding so small folding rules can be easily combined. 3) insert new instructions when needed. I've added the minimum number of rules needed to test the features above.	2018-02-02 12:24:11 -05:00
David Neto	1c0056c339	Start v2018.1-dev	2018-02-02 11:27:07 -05:00
David Neto	c430a41ae3	Finalize v2018.0	2018-02-02 11:25:58 -05:00
David Neto	bcd4f23930	Update CHANGES	2018-02-02 11:23:33 -05:00
Alan Baker	abe113219e	Reordering performance passes ordering to produce better opts * Moved initial insert/extract passes later to cover more opportunities * Added an extra set of passes to clean up opportunities exposed later in the pipeline	2018-02-01 18:01:10 -05:00
Victor Lomuller	50e85c865c	Add LoopUtils class to gather some loop transformation support. This patch adds LoopUtils class to handle some loop related transformations. For now it has 2 transformations that simplifies other transformations such as loop unroll or unswitch: - Dedicate exit blocks: this ensure that all exit basic block (out-of-loop basic blocks that have a predecessor in the loop) have all their predecessors in the loop; - Loop Closed SSA (LCSSA): this ensure that all definitions in a loop are used inside the loop or in a phi instruction in an exit basic block. It also adds the following capabilities: - Loop::IsLCSSA to test if the loop is in a LCSSA form - Loop::GetOrCreatePreHeaderBlock that can build a loop preheader if required; - New methods to allow on the fly updates of the loop descriptors. - New methods to allow on the fly updates of the CFG analysis. - Instruction::SetOperand to allow expression of the index relative to Instruction::NumOperands (to be compatible with the index returned by DefUseManager::ForEachUse)	2018-02-01 15:35:09 -05:00
Steven Perron	61d8c0384b	Add pass to reaplce invalid opcodes Creates a pass that will remove instructions that are invalid for the current shader stage. For the instruction to be considered for replacement 1) The opcode must be valid for a shader modules. 2) The opcode must be invalid for the current shader stage. 3) All entry points to the module must be for the same shader stage. 4) The function containing the instruction must be reachable from an entry point. Fixes #1247.	2018-02-01 15:25:09 -05:00
Andrey Tuganov	d37869c842	Added OpenCL ExtInst validation rules	2018-02-01 14:14:13 -05:00
Jeremy Hayes	cd68f2b176	Add adjacency validation pass Validate OpPhi predecessors. Validate OpLoopMerge successors. Validate OpSelectionMerge successors. Fix collateral damage to existing tests. Remove ValidateIdWithMessage.OpSampledImageUsedInOpPhiBad.	2018-02-01 14:10:55 -05:00
Andrey Tuganov	905536c519	Fixed harmless uninit var warning	2018-01-31 17:49:01 -05:00
David Neto	ac537c71a8	Use SPIR-V headers from "unified1" directory	2018-01-31 15:36:50 -05:00
Alan Baker	2735e0851e	Remove constexpr from Analysis operators * Had to remove templating from InstructionBuilder as a result * now preserved analyses are specified as a constructor argument * updated tests and uses * changed static_assert to a runtime assert * this should probably get further changes in the future	2018-01-31 14:44:43 -05:00
GregF	0aa0ac52f7	Opt: Add ScalarReplacement to RegisterSizePasses	2018-01-31 10:19:17 -05:00
Andrey Tuganov	44d88c8d9c	Add memory semantics checks to validate atomics	2018-01-30 18:00:01 -05:00
David Neto	38f297c194	Update CHANGES	2018-01-30 17:47:00 -05:00
Alan Baker	16949236fe	Prevent unnecessary changes to the IR in dead branch elim * When handling unreachable merges and continues, do not optimize to the same IR * pass did not check whether the unreachable blocks were in the optimized form before transforming them * added a test to catch this issue	2018-01-30 16:51:58 -05:00
Andrey Tuganov	c86cb76a22	Improved error message in val capabilities	2018-01-30 16:22:10 -05:00
Alan Baker	e661da7941	Enhancements to block merging * Should handle all possibilities * Stricter checks for what is disallowed: * header and header * merge and merge * Allow header and merge blocks to be merged * Erases the structured control declaration if merging header and merge blocks together.	2018-01-30 16:05:51 -05:00
Alan Baker	6704233d39	Fix dereference of possibly nullptr * If the dead branch elim is performed on a module without structured control flow, the OpSelectionMerge may not be present * Add a check for pointer validity before dereferencing * Added a test to catch the bug	2018-01-30 10:15:43 -05:00
GregF	f28b106173	InsertExtractElim: Split out DeadInsertElim as separate pass	2018-01-30 08:52:14 -05:00
Alan Baker	1b46f7ecad	Fixes in CCP for #1228 * Forces traversal of phis if the def has changed to varying * Mark a phi as varying if all incoming values are varying * added a test to catch the bug	2018-01-29 15:12:05 -05:00
Victor Lomuller	6018de81de	Add LoopDescriptor as an IRContext analysis. Move some function definitions from header to source to avoid circular definition.	2018-01-25 16:12:32 -05:00
Greg Fischer	684997eb72	DeadInsertElim: Detect and DCE dead Inserts This adds Dead Insert Elimination to the end of the --eliminate-insert-extract pass. See the new tests for examples of code that will benefit. Essentially, this removes OpCompositeInsert instructions which are not used, either because there is no instruction which uses the value at the index it is inserted, or because a subsequent insert intercepts any such use. This code has been seen to remove significant amounts of dead code from real-life HLSL shaders being ported to Vulkan. In fact, it is needed to remove dead texture samples which cause Vulkan validation layer errors (unbound textures and samplers) if not removed . Such DCE is thus required for fxc equivalence and legalization. This analysis operates across "chains" of Inserts which can also contain Phi instructions.	2018-01-25 16:07:21 -05:00
Alan Baker	2e93e806e4	Initial implementation of if conversion * Handles simple cases only * Identifies phis in blocks with two predecessors and attempts to convert the phi to an select * does not perform code motion currently so the converted values must dominate the join point (e.g. can't be defined in the branches) * limited for now to two predecessors, but can be extended to handle more cases * Adding if conversion to -O and -Os	2018-01-25 09:42:00 -08:00
Andrey Tuganov	b2eb840468	Validator: restricted some atomic ops for shaders Ban floating point case for OpAtomicLoad, OpAtomicExchange, OpAtomicCompareExchange. In graphics (Shader) environments, these instructions only operate on scalar integers. Ban the floating point case. OpenCL supports atomic_float.	2018-01-24 14:06:06 -08:00
Andrey Tuganov	bdc78377bc	Added Vulkan-specifc checks to image validation Implemented Vulkan-specific rules: - OpTypeImage must declare a scalar 32-bit float or 32-bit integer type for the “Sampled Type”. - OpSampledImage must only consume an “Image” operand whose type has its “Sampled” operand set to 1.	2018-01-24 14:05:42 -08:00
Steven Perron	c4835e1bd8	Use id_map in Fold*ToConstant The folding routines are suppose to use the id_map provided to map the ids in the instruction. The ones I just added are missing it.	2018-01-22 16:27:31 -05:00
Steven Perron	6c409e30a2	Add generic folding function and use in CCP The current folding routines have a very cumbersome interface, make them harder to use, and not a obvious how to extend. This change is to create a new interface for the folding routines, and show how it can be used by calling it from CCP. This does not make a significant change to the behaviour of CCP. In general it should produce the same code as before; however it is possible that an instruction that takes 32-bit integers as inputs and the result is not a 32-bit integer or bool will not be folded as before. It seems like andriod has a problem with INT32_MAX and the like. I'll explicitly define those if the are not already defined.	2018-01-22 14:26:49 -05:00
Alan Baker	3b780db7f8	Fixes infinite loop in ADCE * Addresses how breaks are indentified to prevent infinite loops when back to back loop share a merge and header * Added test to catch the bug	2018-01-19 11:08:46 -05:00
Victor Lomuller	cf3b2a58c4	Introduce an instruction builder helper class. The class factorize the instruction building process. Def-use manager analysis can be updated on the fly to maintain coherency. To be updated to take into account more analysis.	2018-01-19 10:17:45 -05:00
Alan Baker	73940aba1b	Simplifying code for adding instructions to worklist * AddToWorklist can now be called unconditionally * It will only add instructions that have not already been marked as live * Fixes a case where a merge was not added to the worklist because the branch was already marked as live * Added two similar tests that fail without the fix	2018-01-18 20:36:46 -05:00
Steven Perron	34d4294c2c	Create a pass to work around a driver bug related to OpUnreachable. We have come across a driver bug where and OpUnreachable inside a loop is causing the shader to go into an infinite loop. This commit will try to avoid this bug by turning OpUnreachable instructions that are contained in a loop into branches to the loop merge block. This is not added to "-O" and "-Os" because it should only be used if the driver being targeted has this problem. Fixes #1209.	2018-01-18 20:31:46 -05:00

1 2 3 4 5 ...

1278 Commits