SPIRV-Cross

Author	SHA1	Message	Date
Bill Hollings	284ccf5d2d	Fixes from code review of adding writable images to iOS Tier2 argument buffers.	2023-01-08 21:22:23 -05:00
Bill Hollings	643b7be196	MSL: Add support for writable images in iOS Tier2 argument buffers. - Add CompilerMSL::Options::argument_buffers_tier as an enumeration to allow calling app to specify platform argument buffer tier capabilities. - Support iOS writable images in Tier2 argument buffers when specified. Tier capabilities based on recommendations from Apple engineering.	2022-12-28 12:40:37 -05:00
Chip Davis	aa5a8c482e	MSL: Prevent stores to storage resources in discarded fragments. Some Metal devices have a bug where storage resources can still be written to even if the fragment is discarded. This is obviously a bug in Metal, but bothering Apple to fix it will only fix it for newer versions; therefore, a workaround is needed for older versions. I have made this an option so that, in case the bug is ever fixed, the workaround can be disabled. This workaround is simple: if a fragment shader may discard its fragment and writes to a storage resource, a variable representing the `HelperInvocation` built-in is created and passed to all functions. The flag is checked on all resource writes; writes do not occur when `HelperInvocation` is `true`. This relies on the earlier workaround to update `HelperInvocation` when the fragment is discarded. Fixes at least 3 failures in the CTS.	2022-11-20 01:29:41 -08:00
Chip Davis	c7ce92a95b	MSL: Manually update `BuiltInHelperInvocation` when a fragment is discarded. Some Metal devices have a bug where `simd_is_helper_thread()` won't return true after a fragment has been discarded. We can work around this by manually setting `gl_HelperInvocation` upon discarding a fragment. This is fairly unintrusive, so it is enabled by default. I've made it an option so that, when the bug is fixed, we can disable it.	2022-11-19 23:48:26 -08:00
Hans-Kristian Arntzen	47c7fc16eb	HLSL: Add option to bind vertex input smemantics by name.	2022-10-26 12:41:23 +02:00
Chip Davis	0b679334e4	MSL: Don't flatten arrayed per-patch output blocks in tessellation shaders. Flattening doesn't play well with dynamic indices. In this case, it's better to leave it as an array of structs. (I wanted to do this for named blocks generally. Trouble is, the builtin `gl_out` block is also a named block...) Fixes six more CTS tests, under `dEQP-VK.tessellation.user_defined_io.per_patch_block_array.*`.	2022-10-18 15:04:42 -07:00
Chip Davis	a171087180	MSL: Support "raw" buffer input in tessellation evaluation shaders. Using vertex-style stage input is complex, and it doesn't support nesting of structures or arrays. By using raw buffer input instead, we get this support "for free," and everything becomes much simpler. Arguably, this is the way I should've done this in the first place. Eventually, I'd like to make this the default, and then remove the option altogether. (And I still need to do that with `multi_patch_workgroup`...) Should help fix 66 tests in the Vulkan CTS, under the following trees: - `dEQP-VK.pipeline..interface_matching.` - `dEQP-VK.tessellation.user_defined_io.` - `dEQP-VK.clipping.user_defined.`	2022-10-18 14:58:59 -07:00
Hans-Kristian Arntzen	f09ba27777	Merge pull request #2035 from KhronosGroup/fix-2032 HLSL: Improve support for VertexInfo aux struct.	2022-10-03 14:54:07 +02:00
Hans-Kristian Arntzen	b5386e3ea9	HLSL: Improve support for VertexInfo aux struct. Add concept of explicit bindings for aux structs and allows query if these aux structs are required.	2022-10-03 13:31:27 +02:00
Hans-Kristian Arntzen	f3b1375b13	Add reflection support for shader record buffers. Reflect naming scheme in a context sensitive way that matches the frontend. GLSL -> use block name HLSL (DXC) -> use instance name.	2022-10-03 12:20:08 +02:00
Chip Davis	064eaebe72	MSL: Add a mechanism to fix up shader outputs. This is analogous to the existing support for fixing up shader inputs. It is intended to be used with tessellation to add implicit builtins that are read from a later stage, despite not being written in an earlier stage. (Believe it or not, this is in fact legal in Vulkan.) Helps fix 8 CTS tests under `dEQP-VK.pipeline.*.no_position`. (Eight other tests work solely by accident without this change.)	2022-09-09 17:06:34 -07:00
Hans-Kristian Arntzen	4c345166dc	GLSL: Implement task shaders. Due to bugged glslang / spirv-tools w.r.t. terminator instructions, add a hack to ignore invalid SPIR-V for the time being.	2022-09-05 12:31:22 +02:00
Hans-Kristian Arntzen	06ca9accd7	HLSL: Add option to emit entry point name 1:1 instead of main(). MSL backend supports emitting custom name, and there's no reason for HLSL to not support that as well, but we have to make it an option to not break existing users.	2022-07-22 12:04:33 +02:00
Sergii Penner	1bba4d5137	Fix typo Add a missing coma	2022-06-20 09:26:34 -06:00
Hans-Kristian Arntzen	0b303aab16	Add --stage handling for ray tracing.	2022-05-10 17:14:54 +02:00
Stefan Lienhard	05c9a14422	cli: display missing memory qualifiers for reflect and dump-resources	2022-04-25 22:05:34 +02:00
Hans-Kristian Arntzen	31be74a853	Add relax_nan_checks options. Makes codegen from typical D3D emulation SPIR-V more readable. Also makes cross compilation with NotEqual more sensible. It's very rare to actually need the strict NaN-checks in practice. Also, glslang now emits UnordNotEqual by default it seems, so give up trying to assume OrdNotEqual. Harmonize for UnordNotEqual as the sane default.	2022-03-03 14:50:56 +01:00
Daniel Thornburgh	44c3333a1c	Qualify std::move. Clang added -Wunqualified-std-cast-call in https://reviews.llvm.org/D119670, which warns on unqualified std::move and std::forward calls. This change qualifies these calls to allow the project to build on HEAD Clang -Werror.	2022-03-02 23:17:58 +00:00
Hans-Kristian Arntzen	188dc8b13c	Merge pull request #1862 from flokart-world/feature/flatten-ubo-for-hlsl HLSL: Make --flatten-ubo work correctly	2022-02-16 16:39:45 +01:00
Shintaro Sakahara	ed4ded040e	HLSL: Make --flatten-ubo work correctly	2022-02-16 21:53:24 +09:00
Hans-Kristian Arntzen	c716a9a5dd	Add debug option to modify maximum number of compile iterations. Should be seen as a hack, but it's pragmatic in some scenarios.	2022-02-16 12:12:27 +01:00
Hans-Kristian Arntzen	bb04156d3c	CLI/HLSL: Don't set explicit binding for synthesized NumWorkgroups CBV.	2021-09-30 14:30:49 +02:00
Jon Leech	f2a65545b8	Finish adding SPDX tags and setup a reuse checked in Github Actions CI	2021-06-29 11:03:52 +02:00
Hans-Kristian Arntzen	d75666b170	GLSL: Emit num_views for OVR_multiview2.	2021-06-28 12:56:27 +02:00
Hans-Kristian Arntzen	585fc6f3cb	MSL: Always enable support for base vertex/index on iOS. No good reason to not just enable it in CLI.	2021-06-03 11:27:49 +02:00
Hans-Kristian Arntzen	c87cb54499	MSL: Add CLI option for sampler suffix.	2021-05-21 16:47:41 +02:00
Hans-Kristian Arntzen	26a4986009	GLSL: Implement noncoherent framebuffer fetch.	2021-05-21 14:22:57 +02:00
Hans-Kristian Arntzen	b4a380a04c	Support reflecting builtins. They were ignored in input/output variables.	2021-04-19 12:10:49 +02:00
Hans-Kristian Arntzen	ee85bb345e	Fix print_help comment.	2021-04-19 12:10:49 +02:00
Hans-Kristian Arntzen	9c1cadd440	Add --mask-stage-output-* CLI options.	2021-04-19 12:10:49 +02:00
Hans-Kristian Arntzen	4704482bbc	meta: Update copyright headers to 2021.	2021-01-14 16:07:49 +01:00
Hans-Kristian Arntzen	ce18d1b8a5	CLI: Fix silly regression with handling of -V.	2021-01-08 10:51:49 +01:00
Hans-Kristian Arntzen	02b7f9cbe9	CLI: Add stdin support.	2021-01-06 11:06:41 +01:00
Hans-Kristian Arntzen	cf1e9e0643	Add MIT dual license for the SPIRV-Cross API.	2020-12-01 16:47:08 +01:00
Chip Davis	fd738e3387	MSL: Adjust FragCoord for sample-rate shading. In Metal, the `[[position]]` input to a fragment shader remains at fragment center, even at sample rate, like OpenGL and Direct3D. In Vulkan, however, when the fragment shader runs at sample rate, the `FragCoord` builtin moves to the sample position in the framebuffer, instead of the fragment center. To account for this difference, adjust the `FragCoord`, if present, by the sample position. The -0.5 offset is because the fragment center is at (0.5, 0.5). Also, add an option to force sample-rate shading in a fragment shader. Since Metal has no explicit control for this, this is done by adding a dummy `[[sample_id]]` which is otherwise unused, if none is already present. This is intended to be used from e.g. MoltenVK when a pipeline's `minSampleShading` value is nonzero. Instead of checking if any `Input` variables have `Sample` interpolation, I've elected to check that the `SampleRateShading` capability is present. Since `SampleId`, `SamplePosition`, and the `Sample` interpolation decoration require this cap, this should be equivalent for any valid SPIR-V module. If this isn't acceptable, let me know.	2020-11-23 10:30:24 -06:00
Chip Davis	68908355a9	MSL: Expand subgroup support. Add support for declaring a fixed subgroup size. Metal, like Vulkan with `VK_EXT_subgroup_size_control`, allows the thread execution width to vary depending on factors such as register usage. Unfortunately, this breaks several tests that depend on the subgroup size being what the device says it is. So we'll fix the subgroup size at the size the device declares. The extra invocations in the subgroup will appear to be inactive. Because of this, the ballot mask builtins are now ANDed with the active subgroup mask. Add support for emulating a subgroup of size 1. This is intended to be used by Vulkan Portability implementations (e.g. MoltenVK) when the hardware/software combo provides insufficient support for subgroups. Luckily for us, Vulkan 1.1 only requires that the subgroup size be at least 1. Add support for quadgroup and SIMD-group functions which were added to iOS in Metal 2.2 and 2.3. This will allow clients to take advantage of expanded quadgroup and SIMD-group support in recent Metal versions and on recent Apple GPUs (families 6 and 7). Gut emulation of subgroup builtins in fragment shaders. It turns out codegen for the SIMD-group functions in fragment wasn't implemented for AMD on Mojave; it's a safe bet that it wasn't implemented for the other drivers either. Subgroup support in fragment shaders now requires Metal 2.2.	2020-11-20 15:55:49 -06:00
Hans-Kristian Arntzen	6fc2a0581a	Run format_all.sh.	2020-11-08 13:59:52 +01:00
Hans-Kristian Arntzen	b3344174f7	HLSL: Add option to flatten matrix vertex input semantics. Helps translation layers where we expect inputs to be multiple float vectors rather than an indexed matrix.	2020-11-03 11:18:32 +01:00
Chip Davis	c20d5945a2	MSL: Allow framebuffer fetch on Mac in MSL 2.3. Another Apple GPU feature that will now be supported on Apple Silicon Macs.	2020-10-29 10:50:59 -05:00
Chip Davis	5845e009ea	MSL: Handle Offset and Grad operands for 1D-as-2D textures.	2020-10-15 12:51:00 -05:00
Chip Davis	21d38f74ce	MSL: Fix calculation of atomic image buffer address. Fix reversed coordinates: `y` should be used to calculate the row address. Align row address to the row stride. I've made the row alignment a function constant; this makes it possible to override it at pipeline compile time. Honestly, I don't know how this worked at all for Epic. It definitely didn't work in the CTS prior to this.	2020-10-13 20:51:56 -05:00
Chip Davis	4cf840ee7b	MSL: Support layered input attachments. These need to use arrayed texture types, or Metal will complain when binding the resource. The target layer is addressed relative to the Layer output by the vertex pipeline, or to the ViewIndex if in a multiview pipeline. Unlike with the s/t coordinates, Vulkan does not forbid non-zero layer coordinates here, though this cannot be expressed in Vulkan GLSL. Supporting 3D textures will require additional work. Part of the problem is that Metal does not allow texture views to subset a 3D texture, so we need some way to pass the base depth to the shader.	2020-09-02 09:18:25 -05:00
Chip Davis	cab7335e64	MSL: Don't set the layer for multiview if the device doesn't support it. Some older iOS devices don't support layered rendering. In that case, don't set `[[render_target_array_index]]`, because the compiler will reject the shader in that case. The client will then have to unroll the render pass manually.	2020-09-01 19:30:28 -05:00
Hans-Kristian Arntzen	57c93d44ac	GLSL: Add option to force flattening IO blocks. It is not always desirable to use actual blocks. A prime example in the case where EXT_shader_io_blocks is not supported on the target implementation.	2020-07-28 15:16:06 +02:00
Tomek Ponitka	18f23c47d9	Enabling setting a fixed sampleMask in Metal fragment shaders. In Metal render pipelines don't have an option to set a sampleMask parameter, the only way to get that functionality is to set the sample_mask output of the fragment shader to this value directly. We also need to take care to combine the fixed sample mask with the one that the shader might possibly output.	2020-07-24 11:19:46 +02:00
Chip Davis	688c5fcbda	MSL: Add support for processing more than one patch per workgroup. This should hopefully reduce underutilization of the GPU, especially on GPUs where the thread execution width is greater than the number of control points. This also simplifies initialization by reading the buffer directly instead of using Metal's vertex-attribute-in-compute support. It turns out the only way in which shader stages are allowed to differ in their interfaces is in the number of components per vector; the base type must be the same. Since we are using the raw buffer instead of attributes, we can now also emit arrays and matrices directly into the buffer, instead of flattening them and then unpacking them. Structs are still flattened, however; this is due to the need to handle vectors with fewer components than were output, and I think handling this while also directly emitting structs could get ugly. Another advantage of this scheme is that the extra invocations needed to read the attributes when there were more input than output points are now no more. The number of threads per workgroup is now lcm(SIMD-size, output control points). This should ensure we always process a whole number of patches per workgroup. To avoid complexity handling indices in the tessellation control shader, I've also changed the way vertex shaders for tessellation are handled. They are now compute kernels using Metal's support for vertex-style stage input. This lets us always emit vertices into the buffer in order of vertex shader execution. Now we no longer have to deal with indexing in the tessellation control shader. This also fixes a long-standing issue where if an index were greater than the number of vertices to draw, the vertex shader would wind up writing outside the buffer, and the vertex would be lost. This is a breaking change, and I know SPIRV-Cross has other clients, so I've hidden this behind an option for now. In the future, I want to remove this option and make it the default.	2020-07-23 17:59:54 -05:00
Hans-Kristian Arntzen	d573a95a9c	Run format_all.sh.	2020-07-01 11:42:58 +02:00
Chip Davis	5281d9997e	MSL: Fix up input variables' vector lengths in all stages. Metal is picky about interface matching. If the types don't match exactly, down to the number of vector components, Metal fails pipline compilation. To support pipelines where the number of components consumed by the fragment shader is less than that produced by the vertex shader, we have to fix up the fragment shader to accept all the components produced.	2020-06-16 14:50:30 -05:00
Hans-Kristian Arntzen	2d5200650a	HLSL: Add native support for 16-bit types. Adds support for templated load/store in SM 6.2 to deal with small types.	2020-06-04 12:33:56 +02:00
Hans-Kristian Arntzen	165392a2b0	Document all CLI options.	2020-05-28 13:23:33 +02:00

1 2 3 4

193 Commits