fdo-mirrors/mesa

mirror of https://gitlab.freedesktop.org/mesa/mesa.git synced 2026-05-05 00:58:05 +02:00

Author	SHA1	Message	Date
Roland Scheidegger	aceae09ef0	glsl: fix compile errors with mingw due to missing PRIx64 definitions define __STDC_FORMAT_MACROS and include <inttypes.h> (same as ir_builder_print_visitor.cpp already does). Otherwise, some mingw build errors out (since `8e7e1ae036` and `bbce1c538d` presumably) with: src/compiler/glsl/ir_print_visitor.cpp:479:40: error: expected ‘)’ before ‘PRIu64’ case GLSL_TYPE_UINT64:fprintf(f, "%" PRIu64, ir->value.u64[i]); break; (Note even with that fix I get other format specifier warnings: src/compiler/glsl/ir_print_visitor.cpp:473:47: warning: unknown conversion type character ‘a’ in format [-Wformat=] fprintf(f, "%a", ir->value.f[i]); ^ src/compiler/glsl/ir_print_visitor.cpp:473:47: warning: too many arguments for format [-Wformat-extra-args] but it still compiles at least) Reviewed-by: Jose Fonseca <jfonseca@vmware.com>	2017-01-24 19:12:46 +01:00
Roland Scheidegger	f4df21ed95	gallivm: don't try to use fast rcp for fdiv The use of fast rcp instruction is disabled, and will always fall back to use a division instead (1 / x). Hence, if we get a division opcode, it doesn't make much sense trying to split that into rcp/mul. Reviewed-by: Jose Fonseca <jfonseca@vmware.com>	2017-01-24 19:12:46 +01:00
Roland Scheidegger	25208949d7	gallivm: (trivial) fix ddiv cpu implementation we can't use the cpu implementation of fdiv, as this one uses different lp_build_context, which causes assertion failure. Just use default fdiv action (there is no fast rcp for doubles which we could potentially use anyway). Cc: 17.0 <mesa-stable@lists.freedesktop.org> Reviewed-by: Nicolai Hähnle <nicolai.haehnle@amd.com> Reviewed-by: Jose Fonseca <jfonseca@vmware.com>	2017-01-24 19:12:46 +01:00
Roland Scheidegger	3b575a955c	tgsi: implement ddiv opcode softpipe (along with llvmpipe) claims to support arb_gpu_shader_fp64, so we really need to support that opcode. Cc: 17.0 <mesa-stable@lists.freedesktop.org> Reviewed-by: Nicolai Hähnle <nicolai.haehnle@amd.com> Reviewed-by: Jose Fonseca <jfonseca@vmware.com>	2017-01-24 19:12:46 +01:00
Jason Ekstrand	4c180f9633	i965/blorp: Use the correct ISL format for combined depth/stencil In brw_blorp_copyteximage, we use the format from the render buffer. This could be a combined depth/stencil format. In this case, we handle stencil properly but we give blorp the wrong ISL format. Specifically, we would give blorp ISL_FORMAT_R32G32B32A32_FLOAT which is the wrong size was causing GPU hangs. Fixes: GL45-CTS.gtf30.GL3Tests.packed_depth_stencil.packed_depth_stencil_copyteximage Reviewed-by: Chad Versace <chadversary@chromium.org> Reviewed-by: Topi Pohjolainen <topi.pohjolainen@intel.com> Cc: "13.0 17.0" <mesa-stable@lists.freedesktop.org>	2017-01-24 10:06:07 -08:00
Samuel Pitoiset	0054dded03	st/glsl_to_tgsi: fix compilation warnings since int64 types state_tracker/st_glsl_to_tgsi.cpp:302:28: warning: ‘glsl_to_tgsi_instruction::tex_type’ is too small to hold all values of ‘enum glsl_base_type’ glsl_base_type tex_type:4; Fixes: `8ce53d4a2f` ("glsl: Add basic ARB_gpu_shader_int64 types") Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Nicolai Hähnle <nicolai.haehnle@amd.com>	2017-01-24 12:45:39 +01:00
Samuel Pitoiset	d90d37db73	gallium/radeon: undef the very specific UPDATE_COUNTER macro Also, wrap this into a do { ... } while (0). Suggested by Nicolai. Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Nicolai Hähnle <nicolai.haehnle@amd.com>	2017-01-24 11:17:25 +01:00
Topi Pohjolainen	ba6399df94	i965/blorp: Add also depth and stencil buffers to render cache v2 (Jason, Curro): Add stencil also even though it is not enabled yet. Cc: 17.0 <mesa-stable@lists.freedesktop.org> Reviewed-by: Kenneth Graunke <kenneth@whitecape.org> Reviewed-by: Jason Ekstrand <jason@jlekstrand.net> Signed-off-by: Topi Pohjolainen <topi.pohjolainen@intel.com>	2017-01-24 10:41:58 +02:00
Ben Widawsky	e63ab36d0e	gbm: Fix width height getters return type (trivial) v2: Other way round... to make consistent, make both return type have the fixed width - uint32_t. Cc: Daniel Stone <daniel@fooishbar.org> Signed-off-by: Ben Widawsky <ben@bwidawsk.net> Reviewed-by: Eric Engestrom <eric.engestrom@imgtec.com> Acked-by: Daniel Stone <daniels@collabora.com>	2017-01-23 21:43:38 -08:00
Ben Widawsky	bb9ff98b4c	gbm: Move getters to match order in header file (trivial) Other things are out of order, but I need to add a getter so I'm just fixing those. This helps people adding to GBM know where the right place to put things is. Signed-off-by: Ben Widawsky <ben@bwidawsk.net> Reviewed-by: Eric Engestrom <eric.engestrom@imgtec.com> Acked-by: Daniel Stone <daniels@collabora.com>	2017-01-23 21:43:34 -08:00
Emil Velikov	530cd248f5	docs: add news item and link release notes for 12.0.6 Signed-off-by: Emil Velikov <emil.velikov@collabora.com>	2017-01-24 02:15:30 +00:00
Emil Velikov	9b16bd8b6c	docs: use correct year for the 12.0.6 release notes Signed-off-by: Emil Velikov <emil.velikov@collabora.com> (cherry picked from commit `13953f012d`)	2017-01-24 02:15:30 +00:00
Emil Velikov	c16e7e0a60	docs: add sha256 checksums for 12.0.6 Signed-off-by: Emil Velikov <emil.velikov@collabora.com> (cherry picked from commit `36e3f2542d`)	2017-01-24 02:15:30 +00:00
Emil Velikov	b1137cb9de	docs: add release notes for 12.0.6 Signed-off-by: Emil Velikov <emil.velikov@collabora.com> (cherry picked from commit `555885a0bf`)	2017-01-24 02:15:30 +00:00
Emil Velikov	9924cdecd9	docs/releasing: remove stray "cd" Signed-off-by: Emil Velikov <emil.velikov@collabora.com>	2017-01-24 02:15:29 +00:00
Ilia Mirkin	b755f2f233	nv50: add support for MUL_ZERO_WINS property This is simply keyed off the vertex shader, as that's guaranteed to be present in any pipeline. Signed-off-by: Ilia Mirkin <imirkin@alum.mit.edu>	2017-01-23 20:37:14 -05:00
Ilia Mirkin	8c764a2321	nvc0: add support for MUL_ZERO_WINS property This sets the dnz flag on all the relevant multiplication operations. At emission time, this will only be supported by nvc0+, so nv50 will need a different solution. Signed-off-by: Ilia Mirkin <imirkin@alum.mit.edu>	2017-01-23 20:37:14 -05:00
Ilia Mirkin	e1346f25bf	st/nine: set the MUL_ZERO_WINS flag when supported Signed-off-by: Ilia Mirkin <imirkin@alum.mit.edu> Reviewed-by: Axel Davy <axel.davy@ens.fr>	2017-01-23 20:37:10 -05:00
Ilia Mirkin	6e40938fbc	gallium: add PIPE_CAP_TGSI_MUL_ZERO_WINS Signed-off-by: Ilia Mirkin <imirkin@alum.mit.edu> Reviewed-by: Nicolai Hähnle <nicolai.haehnle@amd.com> Reviewed-by: Axel Davy <axel.davy@ens.fr>	2017-01-23 20:36:47 -05:00
Ilia Mirkin	a2b2cd81d1	gallium: add TGSI_PROPERTY_MUL_ZERO_WINS This will be useful for proper D3D9 emulation, where this behavior is expected by some shaders. Signed-off-by: Ilia Mirkin <imirkin@alum.mit.edu> Reviewed-by: Nicolai Hähnle <nicolai.haehnle@amd.com> Reviewed-by: Axel Davy <axel.davy@ens.fr>	2017-01-23 20:35:55 -05:00
Marek Olšák	573bf0940a	radeonsi: always set the TCL1_ACTION_ENA when invalidating L2 Some CIK-VI docs say this is the default behavior on SI. That doesn't answer whether it's also the default behavior on CIK-VI. Cc: 17.0 13.0 <mesa-stable@lists.freedesktop.org> Reviewed-by: Nicolai Hähnle <nicolai.haehnle@amd.com>	2017-01-23 23:43:38 +01:00
Marek Olšák	5d3dd70cab	radeonsi: don't declare LDS in TES not used since we started using the offchip tess ring Reviewed-by: Nicolai Hähnle <nicolai.haehnle@amd.com>	2017-01-23 23:43:38 +01:00
Marek Olšák	59c5da40ed	radeonsi: preload PS inputs only if KILL is used so that most shaders can get lower VGPR usage thanks to lazy input loading. I think this is a more accurate constraint that prevents the black transitions in Witcher 2. Affected shaders (7758): Max Waves: 57437 -> 58231 (1.38 %) Reviewed-by: Nicolai Hähnle <nicolai.haehnle@amd.com>	2017-01-23 23:43:38 +01:00
Marek Olšák	7b32ae4df5	gallium/radeon: adjust the rule for using the LINEAR_ALIGNED layout Reviewed-by: Nicolai Hähnle <nicolai.haehnle@amd.com>	2017-01-23 23:43:38 +01:00
Marek Olšák	e248390e93	winsys/amdgpu: drop all IBs if at least one was rejected within the context The corruption is inevitable and hangs are possible too. Reviewed-by: Nicolai Hähnle <nicolai.haehnle@amd.com>	2017-01-23 23:43:38 +01:00
Marek Olšák	1840800860	winsys/amdgpu: report a rejected IB as a lost context Reviewed-by: Nicolai Hähnle <nicolai.haehnle@amd.com>	2017-01-23 23:43:38 +01:00
Dave Airlie	dcfcb3047c	vulkan: import latest registry for 1.0.39 extensions. Acked-by: Jason Ekstrand <jason@jlekstrand.net> Signed-off-by: Dave Airlie <airlied@redhat.com>	2017-01-24 08:13:37 +10:00
Dave Airlie	e38bee34bf	vulkan: bump vulkan.h to 1.0.39 version This introduces a bunch of new extension defines. Acked-by: Jason Ekstrand <jason@jlekstrand.net> Signed-off-by: Dave Airlie <airlied@redhat.com>	2017-01-24 08:13:23 +10:00
Grazvydas Ignotas	f65b3641c3	radv: don't resubmit the same cs over and over while tracing Fixes: `97dfff54` ("radv: Dump command buffer on hang.") Signed-off-by: Grazvydas Ignotas <notasas@gmail.com> Reviewed-by: Bas Nieuwenhuizen <bas@basnieuwenhuizen.nl> CC: <mesa-stable@lists.freedesktop.org>	2017-01-23 22:27:05 +01:00
Samuel Pitoiset	aa2ace8e49	gallium/radeon: add HUD queries for monitoring some hw blocks It's also possible to monitor them via performance counters but the hardware can only use two counters simultaneously. It seems easier to re-use the existing code which reads from MMIO instead of writing a multi-pass approach. v2: - add new lines after ':' Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2017-01-23 21:19:49 +01:00
Samuel Pitoiset	a704f19247	gallium/radeon: refactor the GRBM counters path This will allow to expose more queries in order to know which blocks are busy/idle. v2: - add new lines after ':' Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2017-01-23 21:19:49 +01:00
George Kyriazis	00847e4f14	swr: Align query results allocation Some query results struct contents are declared as cache line aligned. Use aligned malloc, and align the whole struct, to be safe. Fixes crash when compiling with clang. CC: <mesa-stable@lists.freedesktop.org> Reviewed-by: Bruce Cherniak <bruce.cherniak@intel.com>	2017-01-23 14:15:54 -06:00
Bruce Cherniak	b829206b07	swr: Prune empty nodes in CalculateProcessorTopology. CalculateProcessorTopology tries to figure out system topology by parsing /proc/cpuinfo to determine the number of threads, cores, and NUMA nodes. There are some architectures where the "physical id" begins with 1 rather than 0, which was creating and empty "0" node and causing a crash in CreateThreadPool. Bugzilla: https://bugs.freedesktop.org/show_bug.cgi?id=97102 Reviewed-By: George Kyriazis <george.kyriazis@intel.com> CC: <mesa-stable@lists.freedesktop.org>	2017-01-23 13:52:26 -06:00
Matt Turner	d349449a16	i965: Use UNUSED to silence unused variable (used in assert).	2017-01-23 10:50:20 -08:00
Rainer Hochecker	09b140abb5	dri: allow 16bit R/GR images to be exported via drm buffers This allows eglCreateImageKHR to access P010 surfaces created by vaapi Signed-off-by: Rainer Hochecker <fernetmenta@online.de> Acked-by: Ben Widawky <ben@bwidawsk.net>	2017-01-23 08:47:15 -08:00
Christian König	1338d912f5	st/va: make sure that we call begin_frame() only once v2 This fixes "st/va: delay calling begin_frame until we have all parameters". v2: call begin frame after decoder (re)creation as well. Signed-off-by: Christian König <christian.koenig@amd.com> Reviewed-by: Nayan Deshmukh <nayan26deshmukh@gmail.com> Tested-by: Andy Furniss <adf.lists@gmail.com>	2017-01-23 17:00:04 +01:00
Eric Engestrom	50141e131a	drirc: remove spurious tabs Signed-off-by: Eric Engestrom <eric@engestrom.ch> Reviewed-by: Edward O'Callaghan <funfunctor@folklore1984.net> Reviewed-by: Nicolai Hähnle <nicolai.haehnle@amd.com>	2017-01-23 16:34:58 +01:00
Nicolai Hähnle	cfabbbcfd7	st/glsl_to_tgsi: use DDIV instead of DRCP + DMUL Fixes GL45-CTS.gpu_shader_fp64.built_in_functions. v2: use DDIV unconditionally (Roland) Reviewed-by: Roland Scheidegger <sroland@vmware.com> (v1) Reviewed-by: Marek Olšák <marek.olsak@amd.com> (v1) Tested-by: Glenn Kennard <glenn.kennard@gmail.com> Tested-by: James Harvey <lothmordor@gmail.com> Cc: 17.0 <mesa-stable@lists.freedesktop.org>	2017-01-23 16:17:26 +01:00
Nicolai Hähnle	b71c415c3d	glsl: split DIV_TO_MUL_RCP into single- and double-precision flags Reviewed-by: Marek Olšák <marek.olsak@amd.com> Reviewed-by: Iago Toral Quiroga <itoral@igalia.com> Tested-by: Glenn Kennard <glenn.kennard@gmail.com> Tested-by: James Harvey <lothmordor@gmail.com> Cc: 17.0 <mesa-stable@lists.freedesktop.org>	2017-01-23 16:17:19 +01:00
Nicolai Hähnle	e4f8f9a638	r600: implement DDIV Tested-by: Glenn Kennard <glenn.kennard@gmail.com> Tested-by: James Harvey <lothmordor@gmail.com> Cc: 17.0 <mesa-stable@lists.freedesktop.org>	2017-01-23 16:17:15 +01:00
Nicolai Hähnle	488560cfe6	r600: factor out cayman_emit_unary_double_raw We will use it for DDIV. Tested-by: Glenn Kennard <glenn.kennard@gmail.com> Tested-by: James Harvey <lothmordor@gmail.com> Cc: 17.0 <mesa-stable@lists.freedesktop.org>	2017-01-23 16:17:12 +01:00
Nicolai Hähnle	76b02d2fe1	r600: double multiply can handle only one multiply at a time It seems clear that trying to multiply two pairs of doubles would result in the temporary register getting overwritten by the second pair. So make the code more explicit. Tested-by: Glenn Kennard <glenn.kennard@gmail.com> Tested-by: James Harvey <lothmordor@gmail.com> Cc: 17.0 <mesa-stable@lists.freedesktop.org>	2017-01-23 16:15:45 +01:00
Timothy Arceri	f3f9207786	glsl: fix tes linking regression Fixes regression caused by `cbeba6bd48`. I accidentally pushed the wrong version of the patch.	2017-01-23 19:07:22 +11:00
Timothy Arceri	38a67f020d	mesa: remove unused gl_shader_info field from gl_linked_shader Reviewed-by: Nicolai Hähnle <nicolai.haehnle@amd.com>	2017-01-23 14:48:04 +11:00
Timothy Arceri	79f07e87c9	mesa/glsl: set and get cs layouts to and from shader_info Reviewed-by: Nicolai Hähnle <nicolai.haehnle@amd.com>	2017-01-23 14:48:04 +11:00
Timothy Arceri	b96bddae67	mesa/glsl: set and get gs layouts directly to and from shader_info Reviewed-by: Nicolai Hähnle <nicolai.haehnle@amd.com>	2017-01-23 14:48:04 +11:00
Timothy Arceri	cbeba6bd48	mesa/glsl/i965: set and get tes layouts directly to and from shader_info Reviewed-by: Nicolai Hähnle <nicolai.haehnle@amd.com>	2017-01-23 14:48:04 +11:00
Timothy Arceri	64e201ab8f	glsl: use last_vert_prog to get last {clip,cull}_distance_array_size Reviewed-by: Nicolai Hähnle <nicolai.haehnle@amd.com>	2017-01-23 14:48:04 +11:00
Timothy Arceri	fc707f570f	mesa/glsl: set {clip,cull}_distance_array_size directly in gl_program There are some line wrapping violations here but those lines will get deleted in the following patch. Reviewed-by: Nicolai Hähnle <nicolai.haehnle@amd.com>	2017-01-23 14:48:04 +11:00
Timothy Arceri	f86d15ed94	st/mesa/glsl: change xfb_program field to last_vert_prog Now that the i965 backend doesn't depend on this field we can make it more generic and short circuit a bunch of code paths. The new field will be used in a following patch for another clean-up. Reviewed-by: Nicolai Hähnle <nicolai.haehnle@amd.com>	2017-01-23 14:48:04 +11:00

1 2 3 4 5 ...

88443 commits