fdo-mirrors/mesa

mirror of https://gitlab.freedesktop.org/mesa/mesa.git synced 2026-05-21 21:58:10 +02:00

Author	SHA1	Message	Date
Samuel Pitoiset	7e433e25c8	radv: add nir_intrinsic_load_sample_positions_amd in the ABI Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Rhys Perry <pendingchaos02@gmail.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18615>	2022-09-20 09:52:37 +00:00
Vinson Lee	97d307406b	radv: Use count_tes_user_sgprs return value. Fix defect reported by Coverity Scan. Useless call (USELESS_CALL) side_effect_free: Calling count_tes_user_sgprs(key) is only useful for its return value, which is ignored. Fixes: `8253ec3855` ("radv: add shader arguments for dynamic patch control points") Signed-off-by: Vinson Lee <vlee@freedesktop.org> Reviewed-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18659>	2022-09-20 06:14:47 +00:00
Bas Nieuwenhuizen	b2972cf410	radv: Add scratch stack to reduce LDS stack in RT traversal. The current stack size is a significant limiter for occupancy, and hence we need smaller stacks in LDS. Rhys earlier had a patch that just put the N entries closest to the root in LDS and the rest in scratch. However, this is not ideal for performance as most of the activity is happening away from the root, near the leaves. Of course we can't just switch it around, as the leaf activity likely isn't happening all the way at the end of the stack. So what we do is make the LDS stack kinda a ringbuffer by always accessing it using the stack index modulo the buffer size (always a power of two so we can efficiently mask). If we then do not have free space in this buffer we evict the entries closest to the root to scratch and if we hit the "bottom" of the LDS space we load from scratch. Some rough perf numbers for indication with Q2RTX: \| evicting \| LDS entries \| perf \| \|----------\|-------------\|------\| \| no \| 76 \| 55% \| \| no \| 32 \| 100% \| \| no \| 24 \| 105% \| \| yes \| 32 \| 95% \| \| yes \| 16 \| 100% \| \| yes \| 8 \| 90% \| \| yes \| 4 \| 75% \| (For the case with 4 entries we need to do some extra accounting as a full batch may not be available to evict) So an obvious choice is to use a stack of 16 entries. One might wonder if Q2RTX perf is mainly good due to BVHs with very little geometry and hence low depth, so I also did some profiling with control. This is done with RGP instruction timing, so this is instructions executed not weighted for enabled masks, i.e. divergence effects included. \| game \| LDS entries \| scratch action \| fraction of iterations \| \|---------\|-------------\|----------------\|------------------------\| \| Control \| 8 \| store \| 10.3% \| \| Control \| 8 \| load \| 34.8% \| \| Control \| 16 \| store \| 0.58% \| \| Control \| 16 \| load \| 2.62% \| \| Q2RTX \| 16 \| store \| 1.00% \| \| Q2RTX \| 16 \| load \| 3.07% \| So Q2RTX doesn't seem like an unreasonably good case for this algorithm. On the implementation side, we can always place the scratch stack at address 0 by just reserving the scratch space, and in the case of fixed callstack size moving that up. In the dynamic case the dynamic stack base already takes any reserved scratch space into account. Reviewed-by: Konstantin Seurer <konstantin.seurer@gmail.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18541>	2022-09-20 01:39:20 +00:00
Rhys Perry	7d26fafacf	radv: fix dynamic RT stack size with VGPR spilling VGPR spilling might cause VGPRs to be spilled at scratch offset 0, so we can't use that. fossil-db (Sienna Cichlid, Q2RTX and Control): Totals from 4 (0.26% of 1524) affected shaders: Instrs: 8734 -> 8737 (+0.03%) CodeSize: 48492 -> 48504 (+0.02%) Latency: 384375 -> 384369 (-0.00%) InvThroughput: 256250 -> 256246 (-0.00%) Copies: 1312 -> 1313 (+0.08%) Branches: 256 -> 258 (+0.78%) Signed-off-by: Rhys Perry <pendingchaos02@gmail.com> Reviewed-by: Konstantin Seurer <konstantin.seurer@gmail.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18541>	2022-09-20 01:39:20 +00:00
Bas Nieuwenhuizen	ca04f968d9	radv: Use nested ifs for pushing child nodes in traversal loop. Avoids a bunch of overhead costs if the previous child was empty already. Reviewed-by: Konstantin Seurer <konstantin.seurer@gmail.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18538>	2022-09-20 01:29:05 +02:00
Bas Nieuwenhuizen	91a4cd26b3	radv: Use constant for ray traversal exit condition. Make the stack base ssa def dead in the loop, can save a register. Reviewed-by: Konstantin Seurer <konstantin.seurer@gmail.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18538>	2022-09-20 01:29:04 +02:00
Bas Nieuwenhuizen	40a235c9a8	Revert "radv/rt: use derefs for the traversal stack" This reverts commit `3750663c72`. Doing things with derefs adds extra instructions for multiplying the index with the element size, e.g. BBF0_13: s_waitcnt vmcnt(0) v_mov_b32_e32 v27, v55 s_mov_b32 s23, exec_lo v_cmpx_ne_i32_e32 -1, v27 s_cbranch_execz _L14 BBF0_14: v_lshlrev_b32_e32 v48, 2, v46 <-- ds_write_b32 v48, v27 v_add_nc_u32_e32 v46, 32, v46 _L14: s_mov_b32 exec_lo, s23 v_mov_b32_e32 v27, v54 s_mov_b32 s23, exec_lo v_cmpx_ne_i32_e32 -1, v27 s_cbranch_execz _L15 BBF0_15: v_lshlrev_b32_e32 v48, 2, v46 <-- ds_write_b32 v48, v27 v_add_nc_u32_e32 v46, 32, v46 On Q2RTC indirect lighting this saves about 2.3 VALU instructions per loop iteration, which is ~4% of VALU instructions (we're at 58 per iteration now according to RGP). Reviewed-by: Konstantin Seurer <konstantin.seurer@gmail.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18538>	2022-09-20 01:29:04 +02:00
Bas Nieuwenhuizen	85ca0b12a2	radv: Store top of stack in a register. Saves a bunch of processing and a lot of LDS traffic. Improves perf of the indirect lighting RT pass in Q2RTX by ~3%. This is mostly due to the -5% VALU instructions and -25% LDS instructions. Reviewed-by: Konstantin Seurer <konstantin.seurer@gmail.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18538>	2022-09-20 01:29:04 +02:00
Bas Nieuwenhuizen	f7f48251b0	radv: Don't flatten bottom AS exit if statement. The flattening by ACO is more efficient than the nir condmask. Reviewed-by: Konstantin Seurer <konstantin.seurer@gmail.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18538>	2022-09-20 01:29:04 +02:00
James Park	b7d4897df9	meson,amd: Remove Windows libelf wrap Functionality isn't worth the maintenance cost. Reviewed-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18478>	2022-09-19 12:51:12 +00:00
Rhys Perry	e122d95d73	radv: remove unnecessary .align_mul=4 The builders can pick a default using the component size. Signed-off-by: Rhys Perry <pendingchaos02@gmail.com> Reviewed-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18465>	2022-09-16 13:19:55 +00:00
Rhys Perry	ee1a75bd74	radv: use nir_ubfe_imm Signed-off-by: Rhys Perry <pendingchaos02@gmail.com> Reviewed-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18465>	2022-09-16 13:19:55 +00:00
Rhys Perry	272d37fa72	radv: shrink zero-initialization in vkCmdSetVertexInputEXT Signed-off-by: Rhys Perry <pendingchaos02@gmail.com> Reviewed-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18465>	2022-09-16 13:19:55 +00:00
Rhys Perry	891cb799aa	radv: disable EXT_vertex_input_dynamic_state when using DGC This simplifies the DGC path and removes some untested code. The only user of the partial DGC implementation (vkd3d-proton) doesn't use EXT_vertex_input_dynamic_state. Signed-off-by: Rhys Perry <pendingchaos02@gmail.com> Reviewed-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18465>	2022-09-16 13:19:55 +00:00
Samuel Pitoiset	299d294304	Revert "radv: upload the PS epilog in the existing pipeline BO" This is completely broken because the PS epilog has refcount and radv_upload_shaders() updates its VA. This reverts commit `7c34b31db2`. Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Bas Nieuwenhuizen <bas@basnieuwenhuizen.nl> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18632>	2022-09-16 11:38:28 +00:00
Qiang Yu	074f3216f2	ac/nir/ngg: support gs streamout Port from radeonsi. Reviewed-by: Timur Kristóf <timur.kristof@gmail.com> Signed-off-by: Qiang Yu <yuq825@gmail.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/17654>	2022-09-16 08:51:28 +00:00
Qiang Yu	5ec79f9899	ac/nir/ngg: nogs support streamout Port from radeonsi. Works on both GFX11 and GFX10. Although GFX10 can do atomic GDS add on all threads, now we just disable the NGG streamout for GFX10, so it's OK. There's a difference for the GFX11 implementation with radeonsi that we do all 4 buffer/stream info calc on a single thread. It's just because this is simple, we need to update GDS on a single thread anyway, and streamout is not that performance critical to loss a small amount of instruction. We may change to a better implementation when using register based streamout. When streamout enabled, ES threads need to save all vertex attributes to LDS besides position. This is because we don't know where in the streamout buffer to export the attributes to and wheter there are space in the streamout buffer. Streamout is done in primitives, so we need to check if there is space and where the current primitive should be written to by GDS atomic add, then in GS threads do the streamout. Reviewed-by: Timur Kristóf <timur.kristof@gmail.com> Signed-off-by: Qiang Yu <yuq825@gmail.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/17654>	2022-09-16 08:51:28 +00:00
Samuel Pitoiset	f5ba4e855e	radv: do not remove PSIZ for VS when the topology is unknown When compiling only the pre-rast stages in a library, the input assembly state might not be present and the topology would be 0. Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Timur Kristóf <timur.kristof@gmail.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18519>	2022-09-16 08:22:16 +00:00
Samuel Pitoiset	7f91555d4c	radv: enable the VS prologs cache if graphicsPipelineLibrary is enabled GPL will re-use most of the VS prologs code. Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Timur Kristóf <timur.kristof@gmail.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18519>	2022-09-16 08:22:16 +00:00
Samuel Pitoiset	c199a5160a	radv: bind the VS input state for prologs created with GPL If we have a VS that needs a prolog without using the dynamic state, that means that it comes from a library, so we can overwrite the cmdbuf VS input state. Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Timur Kristóf <timur.kristof@gmail.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18519>	2022-09-16 08:22:16 +00:00
Samuel Pitoiset	0feab7b9cf	radv: prepare the VS input state for prologs created with GPL This state will be bound at pipeline bind time. Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Timur Kristóf <timur.kristof@gmail.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18519>	2022-09-16 08:22:16 +00:00
Samuel Pitoiset	fdfa59d7bf	radv: rename radv_pipeline_key:🆚:dynamic_vs_input to has_prolog With GPL it's possible to create VS prologs without this dynamic state, so it seems better to rename. Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Timur Kristóf <timur.kristof@gmail.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18519>	2022-09-16 08:22:16 +00:00
Samuel Pitoiset	57b3bff41a	radv: disable VK_EXT_graphics_pipeline_library with LLVM Epilogs/prologs aren't supported at all. Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Timur Kristóf <timur.kristof@gmail.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18609>	2022-09-16 07:12:40 +00:00
Konstantin Seurer	b12cc5c4fe	radv: Cleanup radv_GetInstanceProcAddr Reviewed-by: Bas Nieuwenhuizen <bas@basnieuwenhuizen.nl> Reviewed-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18600>	2022-09-15 20:50:17 +00:00
Hans-Kristian Arntzen	f5b46a643f	radv: Implement VK_EXT_mutable_descriptor_type. Trivial promotion from VALVE, just rename enums and types. Reviewed-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18610>	2022-09-15 16:55:25 +00:00
Samuel Pitoiset	731116da1a	radv: stop checking for NULL pipelines in radv_CmdBindPipeline() This should never happen now. Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-By: Mike Blumenkrantz <michael.blumenkrantz@gmail.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18567>	2022-09-14 19:13:43 +00:00
Samuel Pitoiset	949d76d174	radv: stop dirtying the graphics pipeline when restoring it radv_CmdBindPipeline() does it already. Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-By: Mike Blumenkrantz <michael.blumenkrantz@gmail.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18567>	2022-09-14 19:13:42 +00:00
Samuel Pitoiset	73f1155193	radv: reset the compute pipeline when the saved one was NULL Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-By: Mike Blumenkrantz <michael.blumenkrantz@gmail.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18567>	2022-09-14 19:13:42 +00:00
Samuel Pitoiset	3cceaaf5cd	radv: do not bind NULL graphics pipeline when restoring the meta state It's invalid to bind NULL pipelines, but make sure to reset it to its previous NULL state. Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-By: Mike Blumenkrantz <michael.blumenkrantz@gmail.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18567>	2022-09-14 19:13:42 +00:00
Samuel Pitoiset	9ebaa62a34	radv: stop setting redundant viewport/scissor for internal operations Only emit them when it's needed. Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-By: Mike Blumenkrantz <michael.blumenkrantz@gmail.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18567>	2022-09-14 19:13:42 +00:00
Mike Blumenkrantz	5505811f64	radv: avoid bottlenecking on sequential sparse buffer binds it's more costly to submit individual sparse buffer binds than to merge them and submit bigger binds, so try to pre-compare and flatten out the bind array as much as possible to reduce ioctl counts Reviewed-by: Bas Nieuwenhuizen <bas@basnieuwenhuizen.nl> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18507>	2022-09-14 14:40:02 +00:00
Konstantin Seurer	6a19950b61	radv: Explicitly store the VA of accel structs Gets rid of a bit of code and fixes the RRA accel_struct_vas table if the BO is freed before vkDestroyAccelerationStructureKHR is called. Signed-off-by: Konstantin Seurer <konstantin.seurer@gmail.com> Reviewed-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18530>	2022-09-14 09:05:25 +00:00
Konstantin Seurer	7da66f8f25	radv/rra: Replace aliasing assert with a warning Signed-off-by: Konstantin Seurer <konstantin.seurer@gmail.com> Reviewed-by: Friedrich Vock <friedrich.vock@gmx.de> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18530>	2022-09-14 09:05:25 +00:00
Konstantin Seurer	916621e5a5	radv: Make the radv_buffer_get_va parameter const Signed-off-by: Konstantin Seurer <konstantin.seurer@gmail.com> Reviewed-by: Friedrich Vock <friedrich.vock@gmx.de> Reviewed-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18530>	2022-09-14 09:05:25 +00:00
Konstantin Seurer	d94cb8b595	radv/rra: Remove redundant bounds validation Signed-off-by: Konstantin Seurer <konstantin.seurer@gmail.com> Reviewed-by: Friedrich Vock <friedrich.vock@gmx.de> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18530>	2022-09-14 09:05:25 +00:00
Konstantin Seurer	97d8bb9bc6	radv/rra: Map accel struct VAs to handles When validating a BVH, rra_validate_node uses _mesa_hash_table_u64_search to lookup, whether a BLAS pointer is valid. Since _mesa_hash_table_u64_search returns the data field of the found entry, we need to populate it. Otherwise, the NULL-check won't work. Fixes: `5749806` ("radv: Add Radeon Raytracing Analyzer trace dumping utilities") Signed-off-by: Konstantin Seurer <konstantin.seurer@gmail.com> Reviewed-by: Friedrich Vock <friedrich.vock@gmx.de> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18530>	2022-09-14 09:05:25 +00:00
Georg Lehmann	1e7a930e10	radv: Enable VK_EXT_load_store_op_none. VK_ATTACHMENT_STORE_OP_NONE_EXT is already supported through VK_KHR_dynamic_rendering. It doesn't seem like we need to do anything special for VK_ATTACHMENT_LOAD_OP_NONE_EXT. Closes: https://gitlab.freedesktop.org/mesa/mesa/-/issues/7246 Signed-off-by: Georg Lehmann <dadschoorse@gmail.com> Reviewed-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18545>	2022-09-13 09:56:11 +00:00
Samuel Pitoiset	075d8aeb67	radv: advertise extendedDynamicState2PatchControlPoints For less stuttering with Zink, also required by Zink for full GPL. Closes: https://gitlab.freedesktop.org/mesa/mesa/-/issues/6584 Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Timur Kristóf <timur.kristof@gmail.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18344>	2022-09-13 08:24:14 +00:00
Samuel Pitoiset	eef1511437	radv: implement dynamic patch control points Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Timur Kristóf <timur.kristof@gmail.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18344>	2022-09-13 08:24:14 +00:00
Samuel Pitoiset	76960e2d93	radv: move emitting GE_CNTL for non-NGG pipelines from the cmdbuf GE_CNTL is the equivalent of IA_MULTI_VGT_PARAM on GFX9 and older. Calling this function for every draw shouldn't really hurt in practice because only non-NGG pipelines need this. Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Timur Kristóf <timur.kristof@gmail.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18344>	2022-09-13 08:24:14 +00:00
Samuel Pitoiset	0bf822144f	radv: move emitting PRIMGROUP_SIZE for <= GFX9 from the cmdbuf The number of tessellation patches that is computed from the number of patch control points might change dynamically too. Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Timur Kristóf <timur.kristof@gmail.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18344>	2022-09-13 08:24:14 +00:00
Samuel Pitoiset	556b297977	radv: pass the number of patch control points to si_get_ia_multi_vgt_param() To prepare for dynamic patch control points. Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Timur Kristóf <timur.kristof@gmail.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18344>	2022-09-13 08:24:14 +00:00
Samuel Pitoiset	5bfac03c8a	radv: add ABI lowering support for dynamic patch control points The number of patch control points (TCS) and the number of patches (TCS/TES) is read from user SGPRs. Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Timur Kristóf <timur.kristof@gmail.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18344>	2022-09-13 08:24:14 +00:00
Samuel Pitoiset	8253ec3855	radv: add shader arguments for dynamic patch control points This introduces two new user SGPRS: - tcs_offchip_layout: input patch size and number of patches in TCS - tes_num_patches: number of patches in TES Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Timur Kristóf <timur.kristof@gmail.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18344>	2022-09-13 08:24:14 +00:00
Samuel Pitoiset	21d9390b0e	radv: set workgroup_size to 256 when patch control points is dynamic It's the maximum possible value. This is to ensure that compilers don't optimize away barriers, like in ACO when workgroup_size is less than or equal to wave_size, s_barrier is considered a no-op. Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Timur Kristóf <timur.kristof@gmail.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18344>	2022-09-13 08:24:14 +00:00
Samuel Pitoiset	9373dbdfcc	radv: skip computing some tess info when patch control points is dynamic We don't know the value. Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Timur Kristóf <timur.kristof@gmail.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18344>	2022-09-13 08:24:14 +00:00
Samuel Pitoiset	0cea8276bc	radv: add radv_pipeline_key::dynamic_patch_control_points This will be used to compile different tessellation shaders when patch control points is dynamic. Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Timur Kristóf <timur.kristof@gmail.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18344>	2022-09-13 08:24:14 +00:00
Samuel Pitoiset	54bd5851ba	radv: emit the LDS size for TCS directly from the pipeline on GFX9+ To be consistent with the LDS shader config for LS, and this will be emitted from the cmdbuf for dynamic patch control points. Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Timur Kristóf <timur.kristof@gmail.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18344>	2022-09-13 08:24:14 +00:00
Samuel Pitoiset	9a67edaa56	radv: reword a comment about dynamic states when rasterization is disabled Make it more generic instead of listing all states. Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Timur Kristóf <timur.kristof@gmail.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18344>	2022-09-13 08:24:14 +00:00
Samuel Pitoiset	7c34b31db2	radv: upload the PS epilog in the existing pipeline BO This reduces the number of BOs needed. Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Bas Nieuwenhuizen <bas@basnieuwenhuizen.nl> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18363>	2022-09-13 07:03:14 +00:00

1 2 3 4 5 ...

5903 commits