fdo-mirrors/mesa

mirror of https://gitlab.freedesktop.org/mesa/mesa.git synced 2026-05-20 17:48:15 +02:00

Author	SHA1	Message	Date
Alyssa Rosenzweig	c3eb81fd16	asahi: Identify XML for anisotropic filtering Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/20123>	2022-12-10 21:51:04 -05:00
Alyssa Rosenzweig	8dcf7648f1	agx: Lower VBOs in NIR Now we support all the vertex formats! This means we don't hit u_vbuf for format translation, which helps performance in lots of applications. By doing the lowering in NIR, the vertex fetch code itself can be optimized by NIR (e.g. nir_opt_algebraic) which can improve generated code quality. In my first implementation of this, I had a big switch statement mapping format enums to interchange formats and post-processing code. This ends up being really unwieldly, the combinatorics of bit packing + conversion + swizzles is enormous and for performance we want to support everything (no u_vbuf fallbacks). To keep the combinatorics in check, we rely on parsing the util_format_description to separate out the issues of bit packing, conversion, and swizzling, allowing us to handle bizarro formats like B10G10R10A2_SNORM with no special casing. In an effort to support everything in one shot, this handles all the formats needed for the extensions EXT_vertex_array_bgra, ARB_vertex_type_2_10_10_10_rev, and ARB_vertex_type_10f_11f_11f_rev. Passes dEQP-GLES3.functional.vertex_arrays.* Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19996>	2022-12-02 06:25:20 +00:00
Asahi Lina	112830f1a0	asahi: Pass through layer alignment flag to the hardware Signed-off-by: Asahi Lina <lina@asahilina.net> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/20031>	2022-11-28 19:50:18 +00:00
Alyssa Rosenzweig	597e303b5b	agx: Add merge helpers to GenXML From panfrost. Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/20013>	2022-11-28 16:48:38 +00:00
Alyssa Rosenzweig	debee344a2	agx: Make empty texture pack to all-zeroes So we can do partial textures. Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/20013>	2022-11-28 16:48:38 +00:00
Asahi Lina	f5a26cc646	asahi: Fix remaining build issues on macOS Signed-off-by: Asahi Lina <lina@asahilina.net> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/20030>	2022-11-28 16:10:19 +00:00
Alyssa Rosenzweig	20cdc35fdb	asahi: Add missing #include Noticed when shuffling headers. Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19999>	2022-11-25 18:56:48 +00:00
Asahi Lina	6f15873d44	asahi: Introduce compressed resource support Signed-off-by: Asahi Lina <lina@asahilina.net> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19999>	2022-11-25 18:56:48 +00:00
Asahi Lina	78948c03f0	asahi: Identify compression-related XML Signed-off-by: Asahi Lina <lina@asahilina.net> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19999>	2022-11-25 18:56:48 +00:00
Alyssa Rosenzweig	b102f045ab	asahi: Set GPR count accurately for background/EOT Better occupancy, which is especially important when the background shader does memory access (for reloads). On my 4K monitor, glmark2 -bdesktop fullscreen from 95fps to 133fps. At default settings, glmark2 -bterrain from 63fps to 71fps. Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19997>	2022-11-25 18:02:42 +00:00
Alyssa Rosenzweig	04360a270e	asahi: Copy panfrost's bo cache Massive performance gains, some fps before/after numbers from glmark2: [shading] 1486 -> 2391 [refract] 87 -> 127 [terrain] 32 -> 56 ...and it's basically for free with enough copy/paste, so thank you to Boris Brezillon for an excellent Asahi patch, the LRU cache seems to work great on M1 :-p There are a few minor changes I made from panfrost, notably adjusting the constants to account for 16KiB pages and switching from pthread_mutex to simple_mtx to be less weird in Mesa. For context on the design, the following commits evolved it in Panfrost and their commit messages may be useful... The logic in this module is the product of years of mistakes and correcting course :-) `f06809cdca` ("panfrost: Evict the BO cache when allocation fails") `77d0498913` ("panfrost: Fix major flaw in BO cache") `ee82f9f07e` ("panfrost: Try to evict unused BOs from the cache") `2225383af8` ("panfrost: Make sure the BO is 'ready' when picked from the cache") `9af4aeaaf7` ("panfrost: Don't return imported/exported BOs to the cache") Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19971>	2022-11-24 23:37:48 +00:00
Alyssa Rosenzweig	7c8e3963bd	asahi: Stop aligning pool allocations to 4KiB This defeats the point of specifying alignments and of packing allocations together with the BO cache. We're a real driver now, let's allocate memory like one. Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19971>	2022-11-24 23:37:48 +00:00
Alyssa Rosenzweig	860f5d77c6	asahi: Label BOs internally This will help debugging memory usage problems. Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19971>	2022-11-24 23:37:48 +00:00
Alyssa Rosenzweig	70f40ea4d3	asahi: Wire up all BCn formats We have these native. Passes the relevant piglits. Large reduction in memory usage on Xonotic on higher settings (8x less memory per texture), which allows Xonotic to run at high settings without OOMing. Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Tested-by: Asahi Lina <lina@asahilina.net> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19903>	2022-11-21 22:33:43 +00:00
Alyssa Rosenzweig	74e92274af	asahi,agx: Use new tilebuffer infrastructure Flag day change to replace the previous hardcoded background/end-of-tile shaders and the API-style load/store_output in fragment shaders with the generated shaders and lowered *_agx intrinsics. This gets us working non-UNORM8 render targets and working MRT. It's also a step in the direction of working MSAA but that needs a lot more work, since the multisampling programming model on AGX is quite different from any of the APIs (including Metal). Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19871>	2022-11-19 20:25:41 +00:00
Alyssa Rosenzweig	c5c0ea39f6	asahi: Add new clear/reload/store infrastructure With multiple render targets, it's not practical to generate all variants of the background and end-of-tile programs at start up. Rather than trying, add a hash table of meta program keys to background programs, and compile variants as they're needed. With the new infrastructure, it's sensible to handle clears with the same code path as reloads. In addition to getting us closer to multiple render target support, this gets us support for non-RGBA8 render targets, as the u8norm tilebuffer format was baked into the hardcoded clear shader and store shaders used before. Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19871>	2022-11-19 20:25:41 +00:00
Alyssa Rosenzweig	b1f5004ee7	asahi: Add agx_usc_shared_none helper Convenience for vertex USC programs. Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19871>	2022-11-19 20:25:41 +00:00
Alyssa Rosenzweig	c713197c25	asahi: Add R16 SNORM formats For completeness, since we do have hardware for this. Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19871>	2022-11-19 20:25:41 +00:00
Alyssa Rosenzweig	d637189d36	asahi: Add more XML via PowerVR These bits are the same as RGX. Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19871>	2022-11-19 20:25:41 +00:00
Alyssa Rosenzweig	a3907e92da	asahi: Add note to XML about 16-bit varyings Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19871>	2022-11-19 20:25:41 +00:00
Alyssa Rosenzweig	94a8fe51d5	asahi: Identify more depth-related fields in XML Needed for gl_FragDepth writes. Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19871>	2022-11-19 20:25:41 +00:00
Alyssa Rosenzweig	6ce615d852	asahi: Add XML for layered rendering We don't need to support this for a while but it's good to know the mechanism. Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19871>	2022-11-19 20:25:41 +00:00
Alyssa Rosenzweig	74de571402	asahi: Add NIR pass to lower tilebuffer access The compiler can't handle load/store_output directly for nontrivial tilebuffer layouts. Add a NIR pass to lower these intrinsics, applying a given layout. Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19871>	2022-11-19 20:25:41 +00:00
Alyssa Rosenzweig	66a680a043	asahi: Add tilebuffer layout helpers Laying out the tilebuffer is nontrivial and a task shared between GL and VK, so add unit-tested helpers. Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19871>	2022-11-19 20:25:41 +00:00
Alyssa Rosenzweig	5d3243ea2d	asahi: Add some notes about unknowns to the XML Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19871>	2022-11-19 20:25:41 +00:00
Alyssa Rosenzweig	363ffa779d	asahi: Identify multisampling fields of shared layout Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19871>	2022-11-19 20:25:41 +00:00
Alyssa Rosenzweig	5a20c90508	asahi: Add _with_bo pool uploads Will be useful for managing our meta shaders. Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19871>	2022-11-19 20:25:41 +00:00
Alyssa Rosenzweig	8781aef6b4	asahi: Make libasahi_lib depend on libasahi_decode The track_alloc and track_free symbols are used, we need to link them in. Depending on build flags / environment / etc, fixes the potential build error hit by a CI job: mold: error: undefined symbol: agxdecode_track_alloc >>> referenced by agx_device.c >>> src/asahi/lib/libasahi_lib.a(src/asahi/lib/libasahi_lib.a.p/agx_device.c.o):(agx_shmem_alloc)>>> referenced by agx_device.c >>> src/asahi/lib/libasahi_lib.a(src/asahi/lib/libasahi_lib.a.p/agx_device.c.o):(agx_bo_create) mold: error: undefined symbol: agxdecode_track_free >>> referenced by agx_device.c >>> src/asahi/lib/libasahi_lib.a(src/asahi/lib/libasahi_lib.a.p/agx_device.c.o):(agx_bo_unreference) ...when trying to link with libasahi_lib without libasahi_decode for unit tests. Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19871>	2022-11-19 20:25:41 +00:00
Alyssa Rosenzweig	6ee6cfec41	asahi: Use PIPE_FORMATs for driver-compiler ABI This avoids exposing the ISA-internal agx_format to the driver, instead hiding it behind a real PIPE_FORMAT. This lets us use real pipe formats in formatted intrinsics in NIR, which is convenient; it will allow us to simplify the compiler/driver ABI; and it lets us use common format helpers (e.g. util_format_get_blocksize) for the internal formats in driver lowering. Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19871>	2022-11-19 20:25:41 +00:00
Alyssa Rosenzweig	f328207475	asahi: Split out agx_usc.h into a common file So the tilebuffer helpers can build the "shared" USC word. Also because Ella will probably want to use these O:) Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19811>	2022-11-19 04:27:11 +00:00
Alyssa Rosenzweig	8be506039d	asahi: Note some magic bits used with memoryless RTs Obviously there can't actually be memoryless render targets, because how would partial renders work? The control stream with memoryless looks like everything would if it went to memory (e.g. full 2D MSAA attachments for the partial loads/stores even if only a resolved 2D image for the final store). Except the memoryless attachments all load from the same address 0xeeee0000. Clearly that's not actually what happens, so what gives? Unclear... but I see the magic bits mentioned here set, and I assume there are some firmware (or kernel) shenanigans used to JIT allocate the backing storage for partial renders. Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19811>	2022-11-19 04:27:11 +00:00
Alyssa Rosenzweig	3fa87e47d5	asahi: Identify "Sample mask after depth/stencil" bit Corresponds to Metal [[sample_mask,post_depth_coverage]]. Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19811>	2022-11-19 04:27:11 +00:00
Alyssa Rosenzweig	ff616099ce	asahi: Identify the pass type enum Via PowerVR. Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19811>	2022-11-19 04:27:10 +00:00
Alyssa Rosenzweig	2e6369f5f6	asahi: Identify PBE sample count Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19811>	2022-11-19 04:27:10 +00:00
Alyssa Rosenzweig	1f0edc0158	asahi: Identify Dimension for Render Target Metal uses when rendering to multisampled 2D. Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19811>	2022-11-19 04:27:10 +00:00
Alyssa Rosenzweig	9c52001a1d	asahi: Implement stencil texturing Stencil texturing is easy: S8_UINT is textured like R8_UINT (with a little swizzle fixup), and stencil is always S8_UINT thanks to u_transfer_helper. So we just need to do some fixups to make u_transfer_helper's seperate_stencil work and everything will work out. Passes dEQP-GLES31.functional.stencil_texturing.* Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19811>	2022-11-19 04:27:10 +00:00
Alyssa Rosenzweig	1ffbd53aa2	asahi: Add internal formats for RGB10A2 We need to use I16 as the interchange format here. Fixes: dEQP-GLES3.functional.fragment_out.basic.uint.rgb10_a2ui* Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19811>	2022-11-19 04:27:10 +00:00
Alyssa Rosenzweig	efb5aef935	asahi: Implement perf_debug Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19811>	2022-11-19 04:27:10 +00:00
Alyssa Rosenzweig	a57b4577a1	asahi: Fix indexed draw decode Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19606>	2022-11-17 02:47:10 +00:00
Asahi Lina	b774ed7c18	asahi: Stub import/export code It will be used on Linux, and it is convenient to be able to compile the same code on macOS in the mean time. Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19606>	2022-11-17 02:47:10 +00:00
Asahi Lina	7147313d0a	asahi: Support XRGB formats Just treat them like ARGB. Not sure if this is sane, but it works for now... Signed-off-by: Asahi Lina <lina@asahilina.net> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19606>	2022-11-17 02:47:10 +00:00
Asahi Lina	7c59e75481	asahi: Add renderonly to device Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19606>	2022-11-17 02:47:10 +00:00
Alyssa Rosenzweig	eac8cbb049	asahi: Identify counts for compute kernels In the same place as for vertex/fragment. Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19265>	2022-10-29 19:23:51 +00:00
Alyssa Rosenzweig	9061e960b2	asahi: Add group tests Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18813>	2022-10-22 14:59:25 -04:00
Alyssa Rosenzweig	8b464f4c59	asahi: Don't use unnecessary test fixture Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18813>	2022-10-22 14:59:23 -04:00
Alyssa Rosenzweig	537a77ea6b	asahi: Rename LOD clamps tests to fit other packing We'll use for testing the "groups" encoding. Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18813>	2022-10-22 14:59:21 -04:00
Alyssa Rosenzweig	721c4f2186	asahi: Remove "padding" field Trivial. Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18813>	2022-10-22 14:58:48 -04:00
Alyssa Rosenzweig	06cb242a54	asahi: Identify more shader-related fields The big discovery is the "number of uniform registers" field. I learned about this one accidentally when my preamble shaders weren't working right, because we had inadvertently hardcoded "at most 32 registers" :-) In the course of identifying that field, I found that the pipeline address is used as a tagged pointer, with some unknown field in the bottom bits and alignment demanded. The XML is updated to account for this. I later found that there's also a "number of general purpose registers used by the preamble shader" field. I missed this one first, because the encoding is slightly different from the usual "number of general purpose registers in the main shader" field. The specification is slightly coarser. I don't know why the hardware needs that information anyway -- occupancy of the preamble shader should be irrelevant -- but it's not a big deal. Finally I found that the "more than 4 textures?" bit is... not that. I do not yet know what it is, but it is... not that. These all use the new groups() modifier for GenXML Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18813>	2022-10-22 14:58:37 -04:00
Alyssa Rosenzweig	24bfa5af88	asahi: Identify "Uniform high" USC word The start field in the Uniform USC word is only 8-bits, whereas 9-bits are required to address the entire uniform register file. This other word gets used for the high half, with start indexed from u128l in the natural way. Apparently spending the evening stuffing too many uniforms into Metal is paying off. Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18813>	2022-10-22 14:54:07 -04:00
Alyssa Rosenzweig	e126338394	asahi: Precompile for shader-db This gets shader-db's runner working, in conjunction with a shader-db ./run modified to set ASAHI_MESA_DEBUG=precompile. This flag triggers precompiles of all shaders witha default key so we can exercise the compiler. Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18813>	2022-10-22 14:54:07 -04:00

1 2 3 4 5

231 commits