fdo-mirrors/mesa

mirror of https://gitlab.freedesktop.org/mesa/mesa.git synced 2026-05-26 21:08:12 +02:00

Author	SHA1	Message	Date
Andy Furniss	a599302227	st/va Avoid VBR bitrate calculation overflow v2 VBR bitrate calc needs 64 bits at high rates. v2: use float. Signed-off-by: Andy Furniss <adf.lists@gmail.com> Reviewed-by: Christian König <christian.koenig@amd.com> Cc: mesa-stable@lists.freedesktop.org	2016-09-27 14:21:45 +02:00
Mark Thompson	a543f231d7	st/va: Fix vaSyncSurface with no outstanding operation Fixes crash if the application doesn't do what the state tracker expects. Reviewed-by: Christian König <christian.koenig@amd.com>	2016-09-27 14:21:44 +02:00
Samuel Pitoiset	f24b517858	nv50/ir: fix comments about instructions info Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Acked-by: Ilia Mirkin <imirkin@alum.mit.edu>	2016-09-26 21:59:37 +02:00
Rob Clark	ecd6fce261	mesa/st: support lowering multi-planar YUV Support multi-planar YUV for external EGLImage's (currently just in the dma-buf import path) by lowering to multiple texture fetch's for each plane and CSC in shader. There was some discussion of alternative approaches for tracking the additional UV or U/V planes: https://lists.freedesktop.org/archives/mesa-dev/2016-September/127832.html They all seemed worse than pipe_resource::next Signed-off-by: Rob Clark <robdclark@gmail.com>	2016-09-26 15:29:17 -04:00
Samuel Pitoiset	ac859d68f4	nvc0: allow to force compiling programs in debug build This adds a new envvar called NV50_PROG_CHIPSET which allows to compile shaders with a different target, especially useful for shader-db. Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Ilia Mirkin <imirkin@alum.mit.edu>	2016-09-26 19:39:04 +02:00
Samuel Pitoiset	e05042b367	nv50/ir: drop unused NVISA_XXX_CHIPSET constants Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Ilia Mirkin <imirkin@alum.mit.edu>	2016-09-26 19:39:04 +02:00
Samuel Pitoiset	be0535b8c7	gallium/util: make use of strtol() in debug_get_num_option() This allows to use hexadecimal numbers which are automatically detected by strtol() when the base is 0. Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Nicolai Hähnle <nicolai.haehnle@amd.com> Tested-by: Brian Paul <brianp@vmware.com>	2016-09-26 19:39:04 +02:00
Glenn Kennard	5da24242b3	r600g: Add support for PK2H/UP2H Based off of Ilia's original patch, but with output values replicated so that it matches the TGSI semantics. Signed-off-by: Glenn Kennard <glenn.kennard@gmail.com> Signed-off-by: Marek Olšák <marek.olsak@amd.com>	2016-09-26 17:08:49 +02:00
Brian Paul	c0d7b6073d	svga: set PIPE_BIND_DEPTH_STENCIL flag for new resources when possible When we create a depth/stencil texture, also check if we can render to it and set the PIPE_BIND_DEPTH_STENCIL flag. We were previously doing this for color textures (PIPE_BIND_RENDER_TARGET). Reviewed-by: Charmaine Lee <charmainel@vmware.com>	2016-09-23 19:54:42 -06:00
Brian Paul	f942a70340	svga: don't special case caps for SVGA3D_R32_FLOAT This may have been needed years ago during development, but not now. Prevents some regressions after introducing the next patch. Reviewed-by: Charmaine Lee <charmainel@vmware.com>	2016-09-23 19:54:42 -06:00
Brian Paul	14639cdf8f	svga: use new adjust_z_layer() helper in svga_pipe_blit.c To handle z/layer fix-ups for blitting and copying. Note that we weren't doing this properly in svga_blit() before. Also, remove redundant stex, dtex assignments. Reviewed-by: Charmaine Lee <charmainel@vmware.com>	2016-09-23 19:54:42 -06:00
Brian Paul	c42000545d	svga: simplify/improve the format compatibility check for region copies The util_is_format_compatible() function didn't quite do what we wanted for vgpu10. This check is more flexible and allows copies between formats such as R32G32B32A32_FLOAT and R32G32B32A32_INT. Reviewed-by: Charmaine Lee <charmainel@vmware.com>	2016-09-23 19:54:42 -06:00
Brian Paul	2ad4ba0727	svga: add const qualifier on svga_translate_format() Reviewed-by: Charmaine Lee <charmainel@vmware.com>	2016-09-23 19:54:42 -06:00
Brian Paul	4d04696524	svga: eliminate unneeded gotos in svga_validate_surface_view() Reviewed-by: Charmaine Lee <charmainel@vmware.com>	2016-09-23 19:54:42 -06:00
Neha Bhende	47f16f5e7f	svga: disable srgb format related code from svga_blit() With latest mesa and latest piglit tests srgb<->linear conversion is not required as per GL4.4 rules See commit `b662c70aea`. Reviewed-by: Charmaine Lee <charmainel@vmware.com>	2016-09-23 19:53:51 -06:00
Timothy Arceri	e60928f4c4	gallium: remove unused PIPE_CC_GCC_VERSION Acked-by: Edward O'Callaghan <funfunctor@folklore1984.net>	2016-09-23 16:18:21 +10:00
Eric Anholt	36f0f03182	nir: Allow opt_peephole_sel to be more aggressive in flattening IFs. VC4 was running into a major performance regression from enabling control flow in the glmark2 conditionals test, because of short if statements containing an ffract. This pass seems like it was was trying to ensure that we only flattened IFs that should be entirely a win by guaranteeing that there would be fewer bcsels than there were MOVs otherwise. However, if the number of ALU ops is small, we can avoid the overhead of branching (which itself costs cycles) and still get a win, even if it means moving real instructions out of the THEN/ELSE blocks. For now, just turn on aggressive flattening on vc4. i965 will need some tuning to avoid regressions. It does looks like this may be useful to replace freedreno code. Improves glmark2 -b conditionals:fragment-steps=5:vertex-steps=0 from 47 fps to 95 fps on vc4. vc4 shader-db: total instructions in shared programs: 101282 -> 99543 (-1.72%) instructions in affected programs: 17365 -> 15626 (-10.01%) total uniforms in shared programs: 31295 -> 31172 (-0.39%) uniforms in affected programs: 3580 -> 3457 (-3.44%) total estimated cycles in shared programs: 225182 -> 223746 (-0.64%) estimated cycles in affected programs: 26085 -> 24649 (-5.51%) v2: Update shader-db output. Reviewed-by: Ian Romanick <ian.d.romanick@intel.com> (v1)	2016-09-22 11:10:21 +03:00
Brian Paul	b35684543e	gallium/util: add comment on util_is_format_compatible() From reading the code, it's not obvious what is src/dest compatible. The list of a->b copy-compatible formats comes from Jose's original check-in message, with some format name updates. Reviewed-by: Nicolai Hähnle <nicolai.haehnle@amd.com> Reviewed-by: Roland Scheidegger <sroland@vmware.com>	2016-09-21 12:26:17 -06:00
Brian Paul	99d9f764b2	svga: minor simplification in svga_validate_surface_view() Get rid of unneeded local var. Reviewed-by: Charmaine Lee <charmainel@vmware.com>	2016-09-21 12:23:45 -06:00
Brian Paul	1cc7a76d73	svga: remove disable_shader debug variable Never used, AFAIK. Reviewed-by: Charmaine Lee <charmainel@vmware.com>	2016-09-21 12:23:45 -06:00
Nicolai Hähnle	1f291369e4	gallivm: support negation on 64-bit integers This should be analogous to 32-bit integers. Reviewed-by: Edward O'Callaghan <funfunctor@folklore1984.net> Signed-off-by: Nicolai Hähnle <nicolai.haehnle@amd.com>	2016-09-21 10:24:50 +02:00
Dave Airlie	4207612f9c	radeonsi: prepare 64-bit integer support. (v2) v2: - no PIPE_CAP_INT64 yet - emit DIV/MOD without the divide-by-zero workaround Reviewed-by: Marek Olšák <marek.olsak@amd.com> (v1) Reviewed-by: Edward O'Callaghan <funfunctor@folklore1984.net> Signed-off-by: Dave Airlie <airlied@redhat.com> Signed-off-by: Nicolai Hähnle <nicolai.haehnle@amd.com>	2016-09-21 10:24:38 +02:00
Dave Airlie	5561a37710	gallivm/llvmpipe: prepare support for ARB_gpu_shader_int64. This enables 64-bit integer support in gallivm and llvmpipe. v2: add conversion opcodes. v3: - PIPE_CAP_INT64 is not there yet - restrict DIV/MOD defaults to the CPU, as for 32 bits - TGSI_OPCODE_I2U64 becomes TGSI_OPCODE_U2I64 Reviewed-by: Roland Scheidegger <sroland@vmware.com> Reviewed-by: Edward O'Callaghan <funfunctor@folklore1984.net> Signed-off-by: Dave Airlie <airlied@redhat.com> Signed-off-by: Nicolai Hähnle <nicolai.haehnle@amd.com>	2016-09-21 10:24:30 +02:00
Dave Airlie	6b26039da3	tgsi/softpipe: prepare ARB_gpu_shader_int64 support. (v3) This adds all the opcodes to tgsi_exec for softpipe to use. v2: add conversion opcodes. v3: - no PIPE_CAP_INT64 yet - change TGSI_OPCODE_I2U64 to TGSI_OPCODE_U2I64 Reviewed-by: Roland Scheidegger <sroland@vmware.com> Reviewed-by: Edward O'Callaghan <funfunctor@folklore1984.net> Signed-off-by: Dave Airlie <airlied@redhat.com> Signed-off-by: Nicolai Hähnle <nicolai.haehnle@amd.com>	2016-09-21 10:24:11 +02:00
Dave Airlie	3985e6c044	gallium/tgsi: add support for 64-bit integer immediates. This adds support to TGSI for 64-bit integer immediates. Reviewed-by: Marek Olšák <marek.olsak@amd.com> Reviewed-by: Nicolai Hähnle <nicolai.haehnle@amd.com> Reviewed-by: Roland Scheidegger <sroland@vmware.com> Reviewed-by: Edward O'Callaghan <funfunctor@folklore1984.net> Signed-off-by: Dave Airlie <airlied@redhat.com>	2016-09-21 10:23:55 +02:00
Dave Airlie	6e1a34d545	gallium: add opcode and types for 64-bit integers. (v3) This just adds the basic support for 64-bit opcodes, and the new types. v2: add conversion opcodes. add documentation. v3: - make docs more consistent - change TGSI_OPCODE_I2U64 to TGSI_OPCODE_U2I64 Reviewed-by: Marek Olšák <marek.olsak@amd.com> (v2) Reviewed-by: Roland Scheidegger <sroland@vmware.com> Reviewed-by: Edward O'Callaghan <funfunctor@folklore1984.net> Signed-off-by: Dave Airlie <airlied@redhat.com> Signed-off-by: Nicolai Hähnle <nicolai.haehnle@amd.com>	2016-09-21 10:23:05 +02:00
Leo Liu	956f3e3bcd	radeon/vce: add firmware support for version 52.8.3 Signed-off-by: Leo Liu <leo.liu@amd.com> Reviewed-by: Christian König <christian.koenig@amd.com>	2016-09-20 15:58:56 -04:00
Indrajit Das	f9311265bf	st/omx/dec/h265: Correct the timestamping (derived from commit `3b6bda665a`) v2: fix the tabs(Leo) Reviewed-by: Christian König <christian.koenig@amd.com> Reviewed-by: Nishanth Peethambaran <nishanth.peethambaran@amd.com> Signed-off-by: Indrajit Das <indrajit-kumar.das@amd.com> Signed-off-by: Leo Liu <leo.liu@amd.com>	2016-09-20 15:58:56 -04:00
Nayan Deshmukh	0301858a31	st/va: flush the context before calling flush_frontbuffer(v2) so that the texture is rendered to back buffer before calling flush_frontbuffer and can be copied to a different buffer in the function v2: change comment style Signed-off-by: Nayan Deshmukh <nayan26deshmukh@gmail.com> Reviewed-by: Michel Dänzer <michel.daenzer@amd.com> Acked-by: Christian König <christian.koenig@amd.com>	2016-09-20 11:18:29 +02:00
Nayan Deshmukh	e4cc2276c1	st/vdpau: flush the context before calling flush_frontbuffer so that the texture is rendered to back buffer before calling flush_frontbuffer and can be copied to a different buffer in the function Signed-off-by: Nayan Deshmukh <nayan26deshmukh@gmail.com> Reviewed-by: Michel Dänzer <michel.daenzer@amd.com> Acked-by: Christian König <christian.koenig@amd.com>	2016-09-20 11:18:07 +02:00
Nayan Deshmukh	853e80f5a0	vl/dri3: handle the case of different GPU(v4.2) In case of prime when rendering is done on GPU other then the server GPU, use a seprate linear buffer for each back buffer which will be displayed using present extension. v2: Use a seprate linear buffer for each back buffer (Michel) v3: Change variable names and fix coding style (Leo and Emil) v4: Use PIPE_BIND_SAMPLER_VIEW for back buffer in case when a seprate linear buffer is used (Michel) v4.1: remove empty line v4.2: destroy the context and handle the case when create_context fails (Emil) Signed-off-by: Nayan Deshmukh <nayan26deshmukh@gmail.com> Reviewed-by: Leo Liu <leo.liu@amd.com> Acked-by: Michel Dänzer <michel.daenzer@amd.com> Acked-by: Christian König <christian.koenig@amd.com>	2016-09-20 11:17:02 +02:00
Ilia Mirkin	40d787ab05	st/vdpau: fix argument type to vlVdpOutputSurfaceDMABuf Signed-off-by: Ilia Mirkin <imirkin@alum.mit.edu> Reviewed-by: Christian König <christian.koenig@amd.com>	2016-09-20 11:13:05 +02:00
Tim Rowley	92ec820244	swr: [rasterizer core] Better thread destruction Signed-off-by: Tim Rowley <timothy.o.rowley@intel.com>	2016-09-19 20:10:19 -05:00
Tim Rowley	fdf2890423	swr: [rasterizer jitter] Fix missing end-of-file newline Signed-off-by: Tim Rowley <timothy.o.rowley@intel.com>	2016-09-19 20:10:19 -05:00
Tim Rowley	2f86a9577a	swr: [rasterizer core] Add macros for mapping ArchRast to buckets Switch all RDTSC_START/STOP macros to use AR_BEGIN/END macros. Signed-off-by: Tim Rowley <timothy.o.rowley@intel.com>	2016-09-19 20:10:19 -05:00
Samuel Pitoiset	6ed05fa4cb	nvc0: get rid of nvc0_stage_sampler_states_bind_range() Same thing as nvc0_stage_set_sampler_views_range(). Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Ilia Mirkin <imirkin@alum.mit.edu>	2016-09-19 20:03:24 +02:00
Samuel Pitoiset	407948df1b	nvc0: get rid of nvc0_stage_set_sampler_views_range() This function was quite similar to nvc0_stage_set_sampler_views() and I don't see any reasons to not remove it. Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Ilia Mirkin <imirkin@alum.mit.edu>	2016-09-19 20:03:20 +02:00
Samuel Pitoiset	557a29b51f	nv50/ir: optimize SUB(a, b) to MOV(a - b) This helps shaders in UE4 demos, especially with Elemental (+1% perf). This optimization reduces spilling usage in one shader which explains the little gain. GF100/GK104: total instructions in shared programs :2838551 -> 2838045 (-0.02%) total gprs used in shared programs :396706 -> 396684 (-0.01%) total local used in shared programs :34432 -> 34416 (-0.05%) local gpr inst bytes helped 1 19 112 112 hurt 0 0 0 0 Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Ilia Mirkin <imirkin@alum.mit.edu>	2016-09-18 16:42:39 +02:00
Samuel Pitoiset	d8b4f5fcca	gk110/ir: fix wrong emission of OP_NOT This should emit src0 instead of src1. Found by inspection. Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Ilia Mirkin <imirkin@alum.mit.edu> Cc: mesa-stable@lists.freedesktop.org	2016-09-18 16:42:33 +02:00
Martina Kollarova	15804c4b90	r600g/sb: fix struct/class declaration conflicts A couple of forward-declarations were causing warnings in clang: 'value' defined as a class here but previously declared as a struct [-Wmismatched-tags] Signed-off-by: Martina Kollarova <martina.kollarova@intel.com> Reviewed-by: Bas Nieuwenhuizen <bas@basnieuwenhuizen.nl>	2016-09-18 09:23:42 +02:00
Lars Hamre	ddd6116e32	tgsi: Enable returns from within loops Fixes the following piglit test (for softpipe): /spec/glsl-1.10/execution/fs-loop-return Signed-off-by: Lars Hamre <chemecse@gmail.com> Reviewed-by: Brian Paul <brianp@vmware.com>	2016-09-17 10:24:13 -06:00
Charmaine Lee	8a6391477e	svga: relax restriction of compressed formats for texture upload This patch relaxes the restriction of compressed formats for texture upload buffer. For now, 3D texture with compressed format is still not supported in the texture upload buffer path. As Brian noted, ETQW does many texture updates with glCompressedTexSubImage. This patch greatly improves the performance of the ETQW trace. Tested with ETQW, MTT piglit, glretrace, conform, viewperf v2: Per Brian's suggestion, removed the subregion boundary check. Reviewed-by: Brian Paul <brianp@vmware.com>	2016-09-17 10:24:13 -06:00
Brian Paul	15dee0fc1d	svga: skip query flush if we already have the query result This reduces the number of times we flush in some situations (the arbocclude demo is one trivial example). Tested with Piglit, ETQW, Sauerbraten, arbocclude. Reviewed-by: Charmaine Lee <charmainel@vmware.com>	2016-09-17 10:24:13 -06:00
Brian Paul	c71e82b8e9	svga: remove unneeded svga_context_flush() in svga_end_query() Since commit `99d8fe20ab` we don't have to flush the command buffer when we end a query. Tested with Piglit, Sauerbraten, arbocclude, ETQW (noticably faster now). Reviewed-by: José Fonseca <jfonseca@vmware.com>	2016-09-17 10:24:13 -06:00
Charmaine Lee	f1b3374d28	svga: use upload buffer for upload texture. With this patch, when running with vgpu10, instead of mapping directly to the guest backed memory for texture update, we'll use the texture upload buffer and use the transfer from buffer command to update the host side texture memory. This optimization yields about 20% performance improvement with Lightsmark2008 and about 40% with Tropics. Tested with Lightsmark2008, Tropics, Heaven, MTT piglit, glretrace, conform. Reviewed-by: Brian Paul <brianp@vmware.com>	2016-09-17 10:24:13 -06:00
Charmaine Lee	a9c4a861d5	svga: refactor svga_texture_transfer_map/unmap functions Split the functions into separate functions for dma and direct map to make the code more readable. Tested with MTT piglit, glretrace, viewperf, conform, various OpenGL apps Reviewed-by: Brian Paul <brianp@vmware.com>	2016-09-17 10:24:12 -06:00
Charmaine Lee	c8ef82d65a	svga: add SVGA3d_vgpu10_TransferFromBuffer() Also add the corresponding dump function to dump the TransferFromBuffer command. Reviewed-by: Brian Paul <brianp@vmware.com>	2016-09-17 10:24:12 -06:00
Charmaine Lee	2a4b019239	svga: single sample surface can be created as non-multisamples surface With this patch, single sample surface will be created as non-multisamples surface. Tested with piglit, glretrace. Reviewed-by: Brian Paul <brianp@vmware.com>	2016-09-17 10:24:12 -06:00
Charmaine Lee	5947d90830	svga: fix memory leak with sampler state This patch fixes a memory leak with sampler state when piglit is run with HW version 11. Sampler state clean up was incorrectly skipped in svga_cleanup_sampler_state() for vgpu9. Tested with piglit. Reviewed-by: Neha Bhende <bhenden@vmware.com> Reviewed-by: Brian Paul <brianp@vmware.com>	2016-09-17 10:24:12 -06:00
Brian Paul	12689efbbe	svga: fix prim type check/assignment in translate_indices() Left over test code spotted by Sinclair. Tested with piglit, Google Earth, Lightsmark, Heaven4, glretraces, etc. Reviewed-by: Sinclair Yeh <syeh@vmware.com>	2016-09-17 10:09:00 -06:00

... 8 9 10 11 12 ...

29173 commits