fdo-mirrors/mesa

mirror of https://gitlab.freedesktop.org/mesa/mesa.git synced 2026-05-20 02:38:07 +02:00

Author	SHA1	Message	Date
Boris Brezillon	5b810c7de3	panfrost/midgard: Add missing lowering passes for type/size conversion ops Replace the manual type/size conversion lowering description by one that's automatically generated and covers all type/size conversions. Signed-off-by: Boris Brezillon <boris.brezillon@collabora.com> Reviewed-by: Alyssa Rosenzweig <alyssa.rosenzweig@collabora.com> Tested-by: Marge Bot <https://gitlab.freedesktop.org/mesa/mesa/merge_requests/3478> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/merge_requests/3478>	2020-01-22 15:31:28 +00:00
Boris Brezillon	fcceeaffae	panfrost/midgard: Add 64 bits float <-> int converters The 64 bit converter cases were missing, add them now. Signed-off-by: Boris Brezillon <boris.brezillon@collabora.com> Reviewed-by: Alyssa Rosenzweig <alyssa.rosenzweig@collabora.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/merge_requests/3478>	2020-01-22 15:31:28 +00:00
Boris Brezillon	fe5fbadd46	panfrost/midgard: Fix mir_print_instruction() for branch instructions Branch instructions should not be treated as regular ALUs. Signed-off-by: Boris Brezillon <boris.brezillon@collabora.com> Reviewed-by: Alyssa Rosenzweig <alyssa.rosenzweig@collabora.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/merge_requests/3478>	2020-01-22 15:31:28 +00:00
Boris Brezillon	e1f9e8d60b	panfrost/midgard: Add f2f64 support So we can convert floats into doubles. Signed-off-by: Boris Brezillon <boris.brezillon@collabora.com> Reviewed-by: Alyssa Rosenzweig <alyssa.rosenzweig@collabora.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/merge_requests/3478>	2020-01-22 15:31:28 +00:00
Boris Brezillon	f53a0799c7	panfrost/midgard: Factorize f2f and u2u handling Those size conversion operations work the same way apart from f2f using an fmov op code and u2u using an imov. Let's handle them in the same case block to avoid code duplication. Signed-off-by: Boris Brezillon <boris.brezillon@collabora.com> Reviewed-by: Alyssa Rosenzweig <alyssa.rosenzweig@collabora.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/merge_requests/3478>	2020-01-22 15:31:28 +00:00
Boris Brezillon	6548d01b3d	panfrost/midgard: Make sure promote_fmov() only promotes 32-bit imovs mir_constant_float() assumes we're dealing with 32-bit integers/floats, which is only the case if reg_mode is equal to midgard_reg_mode_32. Signed-off-by: Boris Brezillon <boris.brezillon@collabora.com> Reviewed-by: Alyssa Rosenzweig <alyssa.rosenzweig@collabora.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/merge_requests/3478>	2020-01-22 15:31:28 +00:00
Boris Brezillon	9566f26ed4	panfrost/midgard: Rework mir_adjust_constants() to make it type/size agnostic Right now, constant combining is not supported in 16 bit mode, and 64 bit mode is simply ignored. Let's rework the function to make it type/bit-size agnostic. Signed-off-by: Boris Brezillon <boris.brezillon@collabora.com> Reviewed-by: Alyssa Rosenzweig <alyssa.rosenzweig@collabora.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/merge_requests/3478>	2020-01-22 15:31:28 +00:00
Boris Brezillon	15c92d158c	panfrost/midgard: Use a union to manipulate embedded constants Each instruction bundle can contain up to 16 constant bytes. The meaning of those byte is instruction dependent: it depends on the instruction native type (int, uint or float) and the instruction reg_mode (8, 16, 32 or 64 bit). Those different layouts can be exposed as a union to facilitate constants manipulation. Signed-off-by: Boris Brezillon <boris.brezillon@collabora.com> Reviewed-by: Alyssa Rosenzweig <alyssa.rosenzweig@collabora.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/merge_requests/3478>	2020-01-22 15:31:28 +00:00
Boris Brezillon	9134f22df2	panfrost/midgard: Print the actual source register for store operations Store operation use r26/r27 but have a word->reg set to 0 or 1 (base is r26). Let's take this base offset into account in print_load_store_instr(). Signed-off-by: Boris Brezillon <boris.brezillon@collabora.com> Reviewed-by: Alyssa Rosenzweig <alyssa.rosenzweig@collabora.com> Tested-by: Marge Bot <https://gitlab.freedesktop.org/mesa/mesa/merge_requests/3482> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/merge_requests/3482>	2020-01-21 14:57:12 +00:00
Alyssa Rosenzweig	14b37ebd44	panfrost: Add pandecode entries for ASTC/ETC formats Signed-off-by: Alyssa Rosenzweig <alyssa.rosenzweig@collabora.com> Tested-by: Marge Bot <https://gitlab.freedesktop.org/mesa/mesa/merge_requests/3414> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/merge_requests/3414>	2020-01-21 08:35:23 -05:00
Icecream95	31bd3b5279	panfrost: Add ASTC texture formats Acked-by: Daniel Stone <daniels@collabora.com> Reviewed-by: Alyssa Rosenzweig <alyssa.rosenzweig@collabora.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/merge_requests/3414>	2020-01-21 08:35:23 -05:00
Icecream95	960fe9daea	panfrost: Add ETC1/ETC2 texture formats Acked-by: Daniel Stone <daniels@collabora.com> Reviewed-by: Alyssa Rosenzweig <alyssa.rosenzweig@collabora.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/merge_requests/3414>	2020-01-21 08:35:23 -05:00
Alyssa Rosenzweig	2091d311c9	panfrost: Rework linear<--->tiled conversions There's a lot going on here (it's a ton of commits squashed together since otherwise this would be impossible to review...) 1. We have a fast path for linear->tiled for whole (aligned) tiles, but we have to use a slow path for unaligned accesses. We can get a pretty major win for partial updates by using this slow path simply on the borders of the update region, and then hit the fast path for the tile-aligned interior. This does require some shuffling. 2. Mark the LUTs constant, which allows the compiler to inline them, which pairs well with loop unrolling (eliminating the memory accesses and just becoming some immediates.. which are not as immediate on aarch64 as I'd like..) 3. Add fast path for bpp1/2/8/16. These use the same algorithm and we have native types for them, so may as well get the fast path. 4. Drop generic path for bpp != 1/2/8/16, since these formats are generally awful and there's no way to tile them efficienctly and honestly there's not a good reason too either. Lima doesn't support any of these formats; Panfrost can make the opinionated choice to make them linear. 5. Specialize the unaligned routines. They don't have to be fully generic, they just can't assume alignment. So now they should be nearly as fast as the aligned versions (which get some extra tricks to be even faster but the difference might be neglible on some workloads). 6. Specialize also for the size of the tile, to allow 4x4 tiling as well as 16x16 tiling. This allows compressed textures to be efficiently tiled with the same routines (so we add support for tiling ASTC/ETC textures while we're at it) Signed-off-by: Alyssa Rosenzweig <alyssa.rosenzweig@collabora.com> Reviewed-by: Vasily Khoruzhick <anarsoul@gmail.com> Tested-by: Vasily Khoruzhick <anarsoul@gmail.com> #lima on Mali400 Part-of: <https://gitlab.freedesktop.org/mesa/mesa/merge_requests/3414>	2020-01-21 08:35:19 -05:00
Alyssa Rosenzweig	f2d876b2b2	panfrost,lima: De-Galliumize tiling routines There's an implicit dependence on Gallium here that will add more complexity than needed when testing/optimizing out of driver as well as potentially Vulkanizing. We don't need a full pipe_box, just the x/y/w/h properties directly. Signed-off-by: Alyssa Rosenzweig <alyssa.rosenzweig@collabora.com> Reviewed-by: Vasily Khoruzhick <anarsoul@gmail.com> Tested-by: Vasily Khoruzhick <anarsoul@gmail.com> #lima on Mali400 Part-of: <https://gitlab.freedesktop.org/mesa/mesa/merge_requests/3414>	2020-01-21 08:35:16 -05:00
Alyssa Rosenzweig	0ca7ab1c97	panfrost: Compile tiling routines with -O3 These are major hot spots for panfrost and lima; better let the compiler do its thing even on debug builds. Signed-off-by: Alyssa Rosenzweig <alyssa.rosenzweig@collabora.com> Reviewed-by: Vasily Khoruzhick <anarsoul@gmail.com> Tested-by: Vasily Khoruzhick <anarsoul@gmail.com> #lima on Mali400 Part-of: <https://gitlab.freedesktop.org/mesa/mesa/merge_requests/3414>	2020-01-21 08:35:01 -05:00
Alyssa Rosenzweig	4af8d5b064	pan/midgard: Fix recursive csel scheduling Corner case causing invalid scheduling on shaders with nested csels, i.e. GLSL code resembling: (foo ? bool1 : bool2) ? x : y By explicitly disallowing csels this is fixed. Fixes INSTR_INVALID_ENC on a glamor shader (noticeable with slowdown and visual corruption when scrolling "too far" on GTK apps). Signed-off-by: Alyssa Rosenzweig <alyssa.rosenzweig@collabora.com> Tested-by: Marge Bot <https://gitlab.freedesktop.org/mesa/mesa/merge_requests/3463> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/merge_requests/3463>	2020-01-18 14:40:05 +00:00
Alyssa Rosenzweig	564a782ff7	panfrost: Identify un/pack colour opcodes We still need to identify formats in the disassembler, but this will at least get the opcode name clear. Signed-off-by: Alyssa Rosenzweig <alyssa.rosenzweig@collabora.com> Tested-by: Marge Bot <https://gitlab.freedesktop.org/mesa/mesa/merge_requests/3462> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/merge_requests/3462>	2020-01-18 14:18:48 +00:00
Alyssa Rosenzweig	13c32e5fed	pan/midgard: Bytemasks should round up, not round down Otherwise we'll lost components in DCE. Signed-off-by: Alyssa Rosenzweig <alyssa.rosenzweig@collabora.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/merge_requests/3462>	2020-01-18 14:18:48 +00:00
Boris Brezillon	6af63c939b	panfrost/midgard: Fix swizzle for store instructions The current logic considers that the nir_intrinsic_component(store_intr) encodes the source components start, but it actually encodes the destination one. Source component offset adjustment is taken care of in install_registers_instr(), when offset_swizzle() is called. This fixes dEQP-GLES2.functional.shaders.random.all_features.fragment.45 when PAN_MESA_DEBUG=deqp (looks like exposing GLES3 features has an impact on the varyings layout). Signed-off-by: Boris Brezillon <boris.brezillon@collabora.com> Reviewed-by: Alyssa Rosenzweig <alyssa.rosenzweig@collabora.com> Tested-by: Marge Bot <https://gitlab.freedesktop.org/mesa/mesa/merge_requests/3429> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/merge_requests/3429>	2020-01-17 12:54:31 +00:00
Robert Foss	62adb6522b	panfrost: Prefix schedule_program to prevent collision Currently the schedule_program implementation being used is picked at compile time, which on the Android platform means that the bifrost compiler & scheduler is used for all targets, including midgard based hardware. This commit disambiguates between the two schedule_program functions. Signed-off-by: Robert Foss <robert.foss@collabora.com> Reviewed-by: Alyssa Rosenzweig <alyssa.rosenzweig@collabora.com>	2020-01-15 22:30:17 +00:00
Alyssa Rosenzweig	6bd9c4dc57	panfrost: Fix linear depth textures As pointed out by Boris, what we were calling PAN_LINEAR depth textures was in fact u-interleaved tiled (!), but we never noticed since we flipped the flag used for sampling, leading to all sorts of fun bugs when attempting to directly acess depth textures from the CPU. Which begs the question -- if what we called LINEAR was tiled, how do we actually render linear depth textures? It turns out the flags for AFBC form a mali_block_format 2-bit code just like their render-target counterparts, so we can render to any of the above. Signed-off-by: Alyssa Rosenzweig <alyssa.rosenzweig@collabora.com> Reported-by: Boris Brezillon <boris.brezillon@collabora.com> Tested-by: Marge Bot <https://gitlab.freedesktop.org/mesa/mesa/merge_requests/3393> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/merge_requests/3393>	2020-01-14 19:42:20 +00:00
Afonso Bordado	22217f24ec	pan/midgard: Fix midgard_compile.h includes We now use enum mali_format which is defined in panfrost-job.h Tested-by: Marge Bot <https://gitlab.freedesktop.org/mesa/mesa/merge_requests/3243> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/merge_requests/3243>	2020-01-14 17:16:11 +00:00
Boris Brezillon	440b0d6eec	panfrost: Remove unneeded phi nodes Add a pass to remove unneeded phi nodes as done in other drivers. Signed-off-by: Boris Brezillon <boris.brezillon@collabora.com> Reviewed-by: Alyssa Rosenzweig <alyssa.rosenzweig@collabora.com> Tested-by: Marge Bot <https://gitlab.freedesktop.org/mesa/mesa/merge_requests/3294> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/merge_requests/3294>	2020-01-13 14:09:47 +00:00
Alyssa Rosenzweig	59d30fd4bc	pan/midgard: Support indirect UBO offsets ...in case we have arrays in a UBO block that we'd like to access indirectly. Signed-off-by: Alyssa Rosenzweig <alyssa.rosenzweig@collabora.com> Tested-by: Marge Bot <https://gitlab.freedesktop.org/mesa/mesa/merge_requests/3352> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/merge_requests/3352>	2020-01-10 17:48:42 -05:00
Icecream95	f2f1277624	panfrost: Add negative lod bias support Reviewed-by: Alyssa Rosenzweig <alyssa.rosenzweig@collabora.com>	2020-01-10 06:51:42 +00:00
Alyssa Rosenzweig	4cd3dc94ad	panfrost: Don't double-flip Z/W for 2D arrays We need to mindful that we don't clobber the shadow comparator. Fixes dEQP-GLES3.functional.shaders.texture_functions.texture.sampler2darrayshadow_* Signed-off-by: Alyssa Rosenzweig <alyssa.rosenzweig@collabora.com> Reviewed-by: Tomeu Vizoso <tomeu.vizoso@collabora.com>	2020-01-07 08:54:52 +01:00
Alyssa Rosenzweig	bc4c853b49	pan/midgard: Account for z/w flip in texelFetch Required for proper txf of 2D arrays. Fixes dEQP-GLES3.functional.shaders.texture_functions.texelfetch.2darray Signed-off-by: Alyssa Rosenzweig <alyssa.rosenzweig@collabora.com> Reviewed-by: Tomeu Vizoso <tomeu.vizoso@collabora.com>	2020-01-07 08:54:47 +01:00
Roman Stratiienko	ed0fa78b46	panfrost: Fix Android build Include missing `encoder/pan_props.c` into the build. Signed-off-by: Roman Stratiienko <roman.stratiienko@globallogic.com> Reviewed-by: Alyssa Rosenzweig <alyssa.rosenzweig@collabora.com>	2020-01-04 16:54:38 +00:00
Alyssa Rosenzweig	3759b84926	pan/midgard: Use upper ALU tags for MFBD writeout It's not clear yet what the distinction is. Signed-off-by: Alyssa Rosenzweig <alyssa.rosenzweig@collabora.com>	2020-01-02 17:27:23 -05:00
Alyssa Rosenzweig	2d1e18ee83	pan/midgard: Identity ld_color_buffer as 32-bit I'm not sure why I mistakenly identified it as an 8-bit op before. Signed-off-by: Alyssa Rosenzweig <alyssa.rosenzweig@collabora.com>	2020-01-02 15:20:55 -05:00
Alyssa Rosenzweig	5063ab6a9c	pan/midgard: Remove old comment Signed-off-by: Alyssa Rosenzweig <alyssa.rosenzweig@collabora.com>	2020-01-02 15:20:55 -05:00
Alyssa Rosenzweig	5bc62af2a0	pan/midgard: Generate MRT writeout loops They need a very particular form; the naive way we did before is not sufficient in practice, it doesn't look like. So let's follow the rough structure of the blob's writeout since this is fixed code anyway. Signed-off-by: Alyssa Rosenzweig <alyssa.rosenzweig@collabora.com>	2020-01-02 15:20:55 -05:00
Alyssa Rosenzweig	db879b034a	pan/midgard: Generalize IS_ALU and quadword_size There are more ALU tags, let's do some cleanup while we're at it. Signed-off-by: Alyssa Rosenzweig <alyssa.rosenzweig@collabora.com>	2020-01-02 15:20:55 -05:00
Alyssa Rosenzweig	66f98ffab0	pan/midgard: Use better heuristic for shader termination This still may not be perfect (in the sense that legal shaders might still get cut off) but this fits how writeout is done with both Panfrost and the blob, so it's good enough for what we need and allows MRT shaders to be sanely disassembled. Signed-off-by: Alyssa Rosenzweig <alyssa.rosenzweig@collabora.com>	2020-01-02 15:20:55 -05:00
Alyssa Rosenzweig	c298f25c4e	pan/midgard: Fix memory corruption in constant combining It's a long story... but we'd try to insert constants that weren't there and end up clobbering fields in the bundle following the constant array... Signed-off-by: Alyssa Rosenzweig <alyssa.rosenzweig@collabora.com>	2020-01-02 15:20:55 -05:00
Tomeu Vizoso	ed3eede296	panfrost: Dynamically allocate array of texture pointers With 3D textures we can have lots of layers, so better allocate it dynamically at runtime. Signed-off-by: Tomeu Vizoso <tomeu.vizoso@collabora.com> Reviewed-by: Alyssa Rosenzweig <alyssa.rosenzweig@collabora.com>	2020-01-02 12:41:02 -05:00
Afonso Bordado	525cbe85ef	pan/midgard: Optimize branches with inverted arguments Remove the invert on arguments to branches, and invert the branch condition instead. This saves one instruction per inverted argument. Closes #2088 Signed-off-by: Afonso Bordado <afonsobordado@az8.co> Reviewed-by: Alyssa Rosenzweig <alyssa.rosenzweig@collabora.com>	2019-12-31 20:01:16 +00:00
Afonso Bordado	0e83688f47	pan/midgard: Move midgard_is_branch_unit to helpers Signed-off-by: Afonso Bordado <afonsobordado@az8.co> Reviewed-by: Alyssa Rosenzweig <alyssa.rosenzweig@collabora.com>	2019-12-31 20:01:12 +00:00
Alyssa Rosenzweig	a0d65d860d	pan/midgard: Remove prepacked_branch It's an ugly hack that's no longer used. Signed-off-by: Alyssa Rosenzweig <alyssa.rosenzweig@collabora.com>	2019-12-31 03:26:24 +00:00
Alyssa Rosenzweig	02f503ef00	pan/midgard: Convert fragment writeout to proper branches This eliminates the only use of prepacked_branch, which is a such a hack anyway. Signed-off-by: Alyssa Rosenzweig <alyssa.rosenzweig@collabora.com>	2019-12-31 03:26:24 +00:00
Alyssa Rosenzweig	8f4b15636b	panfrost: Remove MRT indirection in blend shaders Since we have a separate blend shader for each render target, let's simplify this structure and reduce the options memory footprint by 88% or something goofy like that. Should also enable separate blending per render target. Signed-off-by: Alyssa Rosenzweig <alyssa.rosenzweig@collabora.com>	2019-12-30 17:11:08 -05:00
Alyssa Rosenzweig	67fe2afa51	panfrost: Implement integer varyings We need to actually work out the varying format on demand, rather than assuming rgba32f. Fixes dEQP-GLES3.functional.fragment_out.basic.int.* Signed-off-by: Alyssa Rosenzweig <alyssa.rosenzweig@collabora.com>	2019-12-30 17:11:08 -05:00
Alyssa Rosenzweig	71df7c69bc	panfrost: Identify glProvokingVertex flag Signed-off-by: Alyssa Rosenzweig <alyssa.rosenzweig@collabora.com>	2019-12-30 17:11:08 -05:00
Alyssa Rosenzweig	c17a441666	pan/midgard: Implement flat shading We need to shuffle around some lowerings but it's just a flag. Signed-off-by: Alyssa Rosenzweig <alyssa.rosenzweig@collabora.com>	2019-12-30 17:11:08 -05:00
Alyssa Rosenzweig	66c2696fda	pan/midgard: Use type-appropriate st_vary We would like to store (u)ints as well. Signed-off-by: Alyssa Rosenzweig <alyssa.rosenzweig@collabora.com>	2019-12-30 17:11:08 -05:00
Caio Marcelo de Oliveira Filho	b0203b561c	panfrost: Fix Makefile.sources Add missing `\`. Fixes Android build. Reviewed-by: Eric Engestrom <eric.engestrom@intel.com> Fixes: `de077c2078` ("panfrost: Remove mali_alt_func")	2019-12-28 12:31:41 -08:00
Alyssa Rosenzweig	65e5c1942a	panfrost: Remove 32-bit next_job path It has been unused for a while; let's just remove the abstraction. Technically the hardware does support 32-bit job descriptors, but we don't and we can't keep them from breaking so let's not pretend they work. Signed-off-by: Alyssa Rosenzweig <alyssa.rosenzweig@collabora.com> Suggested-by: Boris Brezillon <boris.brezillon@collabora.com>	2019-12-27 13:03:22 -05:00
Alyssa Rosenzweig	95ba661b49	panfrost; Update comment about work/uniform_count Signed-off-by: Alyssa Rosenzweig <alyssa.rosenzweig@collabora.com>	2019-12-27 13:01:17 -05:00
Alyssa Rosenzweig	de077c2078	panfrost: Remove mali_alt_func There's only one way to encode comparison functions in the command stream, not two. It's just that the semantics for texture comparisons are flipped from the semantics of stencil comparison. We can factor out that flip to common Panfrost code, rather than tying it to a second Gallium routine. Signed-off-by: Alyssa Rosenzweig <alyssa.rosenzweig@collabora.com>	2019-12-27 12:58:00 -05:00
Alyssa Rosenzweig	bc1fc29e21	panfrost: Add missing #include in common header Fixes way back when... Signed-off-by: Alyssa Rosenzweig <alyssa.rosenzweig@collabora.com>	2019-12-27 12:58:00 -05:00

1 2 3 4 5 ...

610 commits