fdo-mirrors/mesa

mirror of https://gitlab.freedesktop.org/mesa/mesa.git synced 2026-06-19 19:18:21 +02:00

Author	SHA1	Message	Date
Nicolai Hähnle	bdf1bf1cb5	radeonsi: generate an explicit switch instruction over vertex streams SimplifyCFG generates a switch instruction anyway when all four streams are present, but is simultaneously not smart enough to eliminate some redundant jumps that it generates. The generated assembly is still a bit silly, probably because the control flow annotation doesn't know how to handle a switch with uniform condition. Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2016-12-12 09:04:49 +01:00
Nicolai Hähnle	bae929f96e	radeonsi: fetch only outputs of current vertex stream from the GSVS ring Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2016-12-12 09:04:46 +01:00
Nicolai Hähnle	dfb69cac33	radeonsi: only export from GS copy shader for vertex stream 0 When running the copy shader for vertex streams != 0, the SX does not need any data from us (there is no rasterization for the higher vertex streams, only streamout). Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2016-12-12 09:04:43 +01:00
Nicolai Hähnle	21f2bb22a3	radeonsi: do not export VS outputs from vertex streams != 0 This affects for GS copy shaders. When an output is meant for vertex stream != 0, then we don't have to make it available to the pixel shader. There is a minor inefficiency here because the GLSL varying packing pass does not group varyings of the same vertex stream together, but it shouldn't be important in practice. Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2016-12-12 09:04:36 +01:00
Nicolai Hähnle	fc0e009aa7	radeonsi: pull iteration over vertex streams into GS copy shader logic The iteration is not needed for normal vertex shaders. Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2016-12-12 09:04:33 +01:00
Nicolai Hähnle	180ae18ec5	radeonsi: group streamout writes by vertex stream Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2016-12-12 09:04:30 +01:00
Nicolai Hähnle	d89592836a	radeonsi: load the streamout buf descriptors closer to their use LLVM can still decide to hoist the loads since they're marked invariant. Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2016-12-12 09:04:27 +01:00
Nicolai Hähnle	564f17f0d7	radeonsi: extract writing of a single streamout output Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2016-12-12 09:04:24 +01:00
Nicolai Hähnle	b41dd00235	radeonsi: separate the call to si_llvm_emit_streamout from exports Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2016-12-12 09:04:22 +01:00
Nicolai Hähnle	5ad6e56ca3	radeonsi: plumb the output vertex_stream through to si_shader_output_values Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2016-12-12 09:04:19 +01:00
Nicolai Hähnle	2985708fa0	radeonsi: rename members of si_shader_output_values Be a bit more verbose and avoid confusion in future patches. Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2016-12-12 09:04:16 +01:00
Nicolai Hähnle	88509518b0	radeonsi: fix an off-by-one error in the bounds check for max_vertices The spec actually says that calling EmitStreamVertex is undefined when you exceed max_vertices. But we do need to avoid trampling over memory outside the GSVS ring. Cc: mesa-stable@lists.freedesktop.org Reviewed-by: Edward O'Callaghan <funfunctor@folklore1984.net> Reviewed-by: Michel Dänzer <michel.daenzer@amd.com> Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2016-12-12 09:04:13 +01:00
Nicolai Hähnle	7655bccce8	radeonsi: do not kill GS with memory writes Vertex emits beyond the specified maximum number of vertices are supposed to have no effect, which is why we used to always kill GS that reached the limit. However, if the GS also writes to memory (SSBO, atomics, shader images), then we must keep going and only skip the vertex emit itself. Cc: mesa-stable@lists.freedesktop.org Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2016-12-12 09:04:10 +01:00
Nicolai Hähnle	7b5b3d63c5	radeonsi: update all GSVS ring descriptors for new buffer allocations Fixes GL45-CTS.gtf40.GL3Tests.transform_feedback3.transform_feedback3_geometry_instanced. Cc: mesa-stable@lists.freedesktop.org Reviewed-by: Edward O'Callaghan <funfunctor@folklore1984.net> Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2016-12-12 09:04:06 +01:00
Nicolai Hähnle	2eaacba7f2	st/glsl_to_tgsi: plumb the GS output stream qualifier through to TGSI Allow drivers to emit GS outputs in a smarter way. Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2016-12-12 09:04:03 +01:00
Nicolai Hähnle	cc34a6f0bd	tgsi/scan: collect information about output usagemasks Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2016-12-12 09:04:01 +01:00
Nicolai Hähnle	cf8e9778fc	tgsi/scan: collect information about output vertex streams Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2016-12-12 09:03:57 +01:00
Nicolai Hähnle	81d0dc5e55	gallium: extract individual streamout output structure So that we can pass pointers to individual array entries around. Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2016-12-12 09:03:54 +01:00
Nicolai Hähnle	04811354c8	tgsi: add Stream{X,Y,Z,W} fields to tgsi_declaration_semantic This is for geometry shader outputs. Without it, drivers have no way of knowing which stream each output is intended for, and have to conservatively write all outputs to all streams. Separate stream numbers for each component are required due to output packing. Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2016-12-12 09:03:51 +01:00
Nicolai Hähnle	173d80b401	glsl: remember per-component vertex streams for packed varyings Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2016-12-12 09:03:47 +01:00
Grazvydas Ignotas	6092169b96	i965/blorp: fix release build unused variable warning Signed-off-by: Grazvydas Ignotas <notasas@gmail.com> Reviewed-by: Eduardo Lima Mitev <elima@igalia.com>	2016-12-12 07:09:33 +01:00
Edward O'Callaghan	5e6b2b05a5	virgl: Fix a strict-aliasing violation in the encoder As per the C spec, it is illegal to alias pointers to different types. This results in undefined behaviour after optimization passes, resulting in very subtle bugs that happen only on a full moon.. Use a memcpy() as a well defined coercion between the double to uint64_t interpretations of the memory. V.2: Use static_assert() instead of assert(). V.3: Use C99 compat STATIC_ASSERT() over C11 static_assert(). Signed-off-by: Edward O'Callaghan <funfunctor@folklore1984.net> Acked-by: Dave Airlie <airlied@redhat.com>	2016-12-12 16:50:15 +11:00
Kenneth Graunke	35c5a9a64d	i965: Print out cycle estimates at the start of block annotations. We now print START B15 <-B14 (42774 cycles) indicating that we estimate B15 will take 42,774 cycles. Printing this should make it easier where time is spent in the program. Signed-off-by: Kenneth Graunke <kenneth@whitecape.org> Reviewed-by: Matt Turner <mattst88@gmail.com>	2016-12-11 16:33:05 -08:00
Kenneth Graunke	713cd23d8e	mesa: Return LINEAR encoding for winsys FBO depth/stencil. GetFramebufferAttachmentParameteriv should return GL_LINEAR for the window system default framebuffer's GL_DEPTH or GL_STENCIL attachments when there are zero depth or stencil bits. The GL 4.5 spec's GetFramebufferAttachmentParameteriv section says: "If the value of FRAMEBUFFER_ATTACHMENT_OBJECT_TYPE is not NONE, these queries apply to all other framebuffer types: [...] If attachment is not a color attachment, or no data storage or texture image has been specified for the attachment, then params will contain the value LINEAR." Note that we already return LINEAR for the case where there is an actual depth or stencil renderbuffer attached. In the case modified by this patch, FRAMEBUFFER_ATTACHMENT_OBJECT_TYPE returns FRAMEBUFFER_DEFAULT rather than NONE. Fixes a CTS test when run in a visual without depth / stencil buffers: GL45-CTS.gtf30.GL3Tests.framebuffer_srgb.framebuffer_srgb_default_encoding Signed-off-by: Kenneth Graunke <kenneth@whitecape.org> Reviewed-by: Jordan Justen <jordan.l.justen@intel.com>	2016-12-11 16:33:05 -08:00
Grazvydas Ignotas	b58d1eecc6	intel/aubinator: fix 32bit shift overflow warning Doesn't look like this can work on 32bit, just rids of annoying warning. Signed-off-by: Grazvydas Ignotas <notasas@gmail.com> Reviewed-by: Eduardo Lima Mitev <elima@igalia.com>	2016-12-11 20:04:15 +01:00
Grazvydas Ignotas	3a1b15c392	anv: fix release build unused variable warnings Signed-off-by: Grazvydas Ignotas <notasas@gmail.com> Reviewed-by: Eduardo Lima Mitev <elima@igalia.com>	2016-12-11 20:03:14 +01:00
Grazvydas Ignotas	90c29784c6	radv/ac: some fix maybe-uninitialized warnings Mark some paths unreachable so that compiler knows variables are initialized in all valid paths. Signed-off-by: Grazvydas Ignotas <notasas@gmail.com> Reviewed-by: Bas Nieuwenhuizen <bas@basnieuwenhuizen.nl>	2016-12-10 21:46:56 +01:00
Grazvydas Ignotas	ec08666a28	radv/meta: use VK_NULL_HANDLE for handles Otherwise we get 32bit warnings because handle is plain uint64_t there and NULL is not suited to initialize that. Signed-off-by: Grazvydas Ignotas <notasas@gmail.com> Reviewed-by: Bas Nieuwenhuizen <bas@basnieuwenhuizen.nl>	2016-12-10 21:46:56 +01:00
Grazvydas Ignotas	9bff2c9884	radv: fix release build unused variable warnings Just mark with MAYBE_UNUSED. Signed-off-by: Grazvydas Ignotas <notasas@gmail.com> Reviewed-by: Bas Nieuwenhuizen <bas@basnieuwenhuizen.nl>	2016-12-10 21:46:56 +01:00
Grazvydas Ignotas	15e12ab8fc	softpipe: fix release build unused variable warning Signed-off-by: Grazvydas Ignotas <notasas@gmail.com> Signed-off-by: Marek Olšák <marek.olsak@amd.com>	2016-12-10 21:25:45 +01:00
Grazvydas Ignotas	c81a89f662	radeonsi: fix release build unused variable warnings Signed-off-by: Grazvydas Ignotas <notasas@gmail.com> Signed-off-by: Marek Olšák <marek.olsak@amd.com>	2016-12-10 21:19:59 +01:00
Chad Versace	42011be1e2	i965/mt: Disable HiZ when sharing depth buffer externally (v2) intel_miptree_make_shareable() discarded and disabled CCS. Fix it so that it discards and disables HiZ too. Fixes dEQP-EGL.functional.image.render_multiple_contexts.gles2_renderbuffer_depth16_depth_buffer on Skylake. v2: Actually do what the commit message says. Discard the HiZ buffer. Fixes: https://bugs.freedesktop.org/show_bug.cgi?id=98329 Reviewed-by: Topi Pohjolainen <topi.pohjolainen@intel.com> Reviewed-by: Kenneth Graunke <kenneth@whitecape.org> Cc: Nanley Chery <nanley.g.chery@intel.com Cc: Haixia Shi <hshi@chromium.org> Cc: mesa-stable@lists.freedesktop.org	2016-12-10 08:05:11 -08:00
Chad Versace	1c8be049be	i965/mt: Disable aux surfaces after making miptree shareable The entire goal of intel_miptree_make_shareable() is to permanently disable the miptree's aux surfaces. So set intel_mipmap_tree:disable_aux_buffers after the function's done with discarding down the aux surfaces. References: https://bugs.freedesktop.org/show_bug.cgi?id=98329 Reviewed-by: Topi Pohjolainen <topi.pohjolainen@intel.com> Reviewed-by: Kenneth Graunke <kenneth@whitecape.org> Cc: Nanley Chery <nanley.g.chery@intel.com Cc: Haixia Shi <hshi@chromium.org> Cc: mesa-stable@lists.freedesktop.org	2016-12-10 08:05:11 -08:00
Jason Ekstrand	da1c49171d	spirv: Use a simpler and more correct implementaiton of tanh() The new implementation is more correct because it clamps the incoming value to 10 to avoid floating-point overflow. It also uses a much reduced version of the formula which only requires 1 exp() rather than 2. This fixes all of the dEQP-VK.glsl.builtin.precision.tanh.* tests. Reviewed-by: Kenneth Graunke <kenneth@whitecape.org> Cc: "13.0" <mesa-dev@lists.freedesktop.org>	2016-12-09 18:38:21 -08:00
Jason Ekstrand	9807f502eb	glsl: Use a simpler formula for tanh The formula we have used in the past is a trivial reduction from the definition by simply multiplying both the numerator and denominator of the formula by 2. However, multiplying by e^x, you can further reduce it. This allows us to get rid of one side of the clamp and two of exponential functions which should make it faster. The new formula still passes the dEQP precision tests for tanh so it should be fine. Reviewed-by: Roland Scheidegger <sroland@vmware.com> Reviewed-by: Kenneth Graunke <kenneth@whitecape.org>	2016-12-09 18:38:21 -08:00
Edward O'Callaghan	efe9d1cde3	anv: Clean up some unused variables Following on from the spirit of commit `011e5570f`. Signed-off-by: Edward O'Callaghan <funfunctor@folklore1984.net> Reviewed-by: Jason Ekstrand <jason@jlekstrand.net>	2016-12-10 11:59:59 +11:00
Tim Rowley	2a127b780b	swr: [rasterizer common/core/jitter] fetch support for GL_FIXED v2: use fmul(1/65536) instead of fdiv(65535) Reviewed-by: Bruce Cherniak <bruce.cherniak@intel.com>	2016-12-09 16:20:13 -06:00
Emil Velikov	d0d21532f9	configure: cleanup GLX_USE_TLS handling Mesa requires ax_pthread_ok = yes, thus we can fold/rewrite the conditional to follow the more common "if test" pattern. No functional change intended. Signed-off-by: Emil Velikov <emil.velikov@collabora.com> Reviewed-by: Eric Anholt <eric@anholt.net>	2016-12-09 19:22:03 +00:00
Emil Velikov	b83153e77b	configure: enable glx-tls by default In the (not too) distant future we'd want to remove this option and effectively drop the other codepath(s) we have in our dispatch. Linux distributions have been using --enable-glx-tls for a number of years. Some/most BSD platforms still don't support this, yet this should serve as an encouragement to move things forwards. Note: we had many bug reports were opened due to the wrong default option. See the list below for details. v2: - Correct default option in help string (Andreas) - Add bugzilla references. Bugzilla: https://bugs.freedesktop.org/show_bug.cgi?id=70623 Bugzilla: https://bugs.freedesktop.org/show_bug.cgi?id=72902 Bugzilla: https://bugs.freedesktop.org/show_bug.cgi?id=73778 Bugzilla: https://bugs.freedesktop.org/show_bug.cgi?id=89043 Cc: Jean-Sébastien Pédron <dumbbell@FreeBSD.org> Cc: Jonathan Gray <jsg@jsg.id.au> Cc: mesa-maintainers@lists.freedesktop.org Signed-off-by: Emil Velikov <emil.velikov@collabora.com> Reviewed-by: Eric Anholt <eric@anholt.net> Reviewed-by: Andreas Boll <andreas.boll.dev@gmail.com>	2016-12-09 19:21:41 +00:00
Emil Velikov	0715ba4be6	docs: document how to (self-) reject stable patches Document what has been the unofficial way to self-reject stable patches. Namely: drop the mesa-stable tag and push the commit. Signed-off-by: Emil Velikov <emil.velikov@collabora.com> Reviewed-by: Nanley Chery <nanley.g.chery@intel.com>	2016-12-09 17:37:30 +00:00
Emil Velikov	26541a1fcc	egl: add and enable EGL_KHR_config_attribs Extension is already implemented in the main code. Signed-off-by: Emil Velikov <emil.velikov@collabora.com> Reviewed-by: Eric Engestrom <eric.engestrom@imgtec.com>	2016-12-09 17:36:28 +00:00
Emil Velikov	bf384a2d85	egl/surfaceless: remove duplicate KHR_image_base enablement Already set by the core code - dri2_create_screen/dri2_setup_screen Cc: Chad Versace <chadversary@chromium.org> Signed-off-by: Emil Velikov <emil.velikov@collabora.com> Reviewed-by: Eric Engestrom <eric.engestrom@imgtec.com>	2016-12-09 17:36:26 +00:00
Eric Engestrom	9e1d35ca75	egl: unexport _eglConvertIntsToAttribs Nobody else makes use of this function. We can always re-export it if someone ever needs it. Signed-off-by: Eric Engestrom <eric@engestrom.ch> Reviewed-by: Emil Velikov <emil.velikov@collabora.com>	2016-12-09 17:33:43 +00:00
Eric Engestrom	4729e1b511	egl: rename static functions to match convention Signed-off-by: Eric Engestrom <eric@engestrom.ch> Reviewed-by: Emil Velikov <emil.velikov@collabora.com>	2016-12-09 17:33:36 +00:00
Haixia Shi	d4983390a8	compiler/glsl: fix precision problem of tanh Clamp input scalar value to range [-10, +10] to avoid precision problems when the absolute value of input is too large. Fixes dEQP-GLES3.functional.shaders.builtin_functions.precision.tanh.* test failures. v2: added more explanation in the comment. v3: fixed a typo in the comment. Signed-off-by: Haixia Shi <hshi@chromium.org> Reviewed-by: Jason Ekstrand <jason@jlekstrand.net> Reviewed-by: Kenneth Graunke <kenneth@whitecape.org> Cc: "13.0" <mesa-dev@lists.freedesktop.org>	2016-12-09 09:14:20 -08:00
Tim Rowley	7aea08667c	swr: [rasterizer core/memory] Finish R24_UNORM_X8_TYPELESS for AVX512 This one-off specialization was missed. Reviewed-by: Bruce Cherniak <bruce.cherniak@intel.com>	2016-12-09 10:41:31 -06:00
Bas Nieuwenhuizen	53e1c970ef	radv: Use enum for memory types. Inspired by patches from Eric Engestrom. Signed-off-by: Bas Nieuwenhuizen <basni@google.com> Cc: Eric Engestrom <eric@engestrom.ch> Reviewed-by: Edward O'Callaghan <funfunctor@folklore1984.net> Reviewed-by: Dave Airlie <airlied@redhat.com>	2016-12-09 08:53:05 +01:00
Bas Nieuwenhuizen	4ae84efbc5	radv: Use enum for memory heaps. Inspired by patches from Eric Engestrom. Signed-off-by: Bas Nieuwenhuizen <basni@google.com> Cc: Eric Engestrom <eric@engestrom.ch> Reviewed-by: Edward O'Callaghan <funfunctor@folklore1984.net> Reviewed-by: Dave Airlie <airlied@redhat.com>	2016-12-09 08:53:05 +01:00
Bas Nieuwenhuizen	011e5570f8	radv: Clean up some unused variables. Leftovers from anv? Signed-off-by: Bas Nieuwenhuizen <basni@google.com> Reviewed-by: Edward O'Callaghan <funfunctor@folklore1984.net> Reviewed-by: Dave Airlie <airlied@redhat.com>	2016-12-09 08:53:05 +01:00
Timothy Arceri	8977cd4fdd	i965: delay adding built-in uniforms to Parameters list This is a step towards using NIR optimisations over GLSL IR optimisations. Delaying adding built-in uniforms until after we convert to NIR gives it a chance to optimise them away. V2: move the new code back to brw_link_shader() Reviewed-by: Kenneth Graunke <kenneth@whitecape.org>	2016-12-09 16:29:10 +11:00

... 97 98 99 100 101 ...

92185 commits