fdo-mirrors/mesa

mirror of https://gitlab.freedesktop.org/mesa/mesa.git synced 2026-05-27 14:28:22 +02:00

Author	SHA1	Message	Date
Vadim Girlin	17bb96b03d	r600g/sb: use MULADD workaround on R7xx for MULADD_IEEE Looks like the same issue that was seen with MULADD in trans slot on R7xx also affects MULADD_IEEE (maybe all OP3 instructions and MULADD is just a most frequently used?). So the workaround is to not allow affected instructions to be placed into the trans slot. Fixes https://bugs.freedesktop.org/show_bug.cgi?id=67927 Signed-off-by: Vadim Girlin <vadimgirlin@gmail.com> Cc: "9.2" <mesa-stable@lists.freedesktop.org>	2013-08-14 01:03:18 +04:00
Roland Scheidegger	6991f86945	gallivm: implement new float comparison instructions returning integer masks FSEQ/FSGE/FSLT/FSNE work just the same as SEQ/SGE/SLT/SNE except skip the select. And just for consistency use the same appropriate ordered/unordered comparisons for the old opcodes as well. Reviewed-by: Zack Rusin <zackr@vmware.com>	2013-08-13 19:09:17 +02:00
Roland Scheidegger	0930082ffd	tgsi: implement new float comparison instructions returning integer masks Also while here add a bunch of other forgotten (integer) instructions to tgsi_util_get_inst_usage_mask() (which isn't used for much except optimizing away unused input components), though it may still be incomplete. Reviewed-by: Zack Rusin <zackr@vmware.com>	2013-08-13 19:09:17 +02:00
Roland Scheidegger	e7a5bf7a34	gallium: add new float comparison instructions returning integer masks Newer graphic languages don't want messy float mask results but instead true "boolean" mask results for float comparisons. Otherwise just need to convert the floats back to integers. Need to keep the old opcodes however due to both legacy (gl and d3d9) needing them and because older hw can't really deal with integers. These new FSEQ/FSGE/FSLT/FSNE opcodes are part of integer API and hence must be supported if a driver claims to support glsl 1.30 (or PIPE_SHADER_CAP_INTEGERS). Reviewed-by: Zack Rusin <zackr@vmware.com>	2013-08-13 19:09:17 +02:00
Chia-I Wu	3b6cee1634	ilo: enable dumping of WM PCB It was disabled because it wasn't supported.	2013-08-13 16:28:24 +08:00
Chia-I Wu	0f8a86682f	ilo: no binding table change when constants are pushed When constants can be pushed, and nothing else requires new SURFACE_STATEs, there is no need to emit BINDING_TABLE_STATE.	2013-08-13 16:26:03 +08:00
Chia-I Wu	c6e1e0157b	ilo: support push constant model in shaders Source constants from URB constant data when the constant data can fit in the PCB.	2013-08-13 16:04:35 +08:00
Chia-I Wu	5e30ffbda6	ilo: support copying constant buffer 0 to PCB Add ILO_KERNEL_PCB_CBUF0_SIZE so that a kernel can specify how many bytes of constant buffer 0 need to be copied to PCB.	2013-08-13 15:52:41 +08:00
Chia-I Wu	5df62dce34	ilo: make constant buffer 0 upload optional Add ILO_KERNEL_SKIP_CBUF0_UPLOAD so that we can skip constant buffer 0 upload when the kernel does not need it.	2013-08-13 15:52:37 +08:00
Chia-I Wu	8b5b5fe394	Revert "ilo: initialize constant buffer SURFACE_STATE early" This reverts commit `a9b800aa81`. With push constant support, the constructed SURFACE_STATE is unused and wasted. The change only slows things down.	2013-08-13 15:24:58 +08:00
Roland Scheidegger	cd2f26090a	gallivm: fix exec_mask interaction with geometry shader after end of main Because we must maintain an exec_mask even if there's currently nothing on the mask stack, we can still have an exec_mask at the end of the program. Effectively, this mask should be set back to default when returning from main. Without relying on END/RET opcode (I think it's valid to have neither) it is actually difficult to do this, as there doesn't seem any reasonable place to do it, so instead let's just say the exec_mask is invalid outside main (which it really is effectively). The problem is that geometry shader called end_primitive outside the shader (in the epilogue), and as a result used a bogus mask, leading to bugs if we had to set the (somewhat misnamed) ret_in_main bit anywhere. So just avoid the mask combining function when called from outside the shader. Reviewed-by: Zack Rusin <zackr@vmware.com>	2013-08-12 23:33:00 +02:00
Roland Scheidegger	dfa7b72563	draw: simplify prim mask construction The code was quite weird, the second comparison was in fact a complete no-op and we can also do the comparison with the vector directly instead of scalar, which should not also be faster but it is way more obvious how that mask is actually going to look like. Reviewed-by: Zack Rusin <zackr@vmware.com>	2013-08-12 23:33:00 +02:00
Roland Scheidegger	7147094ff2	gallivm: simplify geometry shader mask handling a bit Instead of reducing masks to 0/1 simply use the mask directly as -1. Also use some signed comparison instead of unsigned (as far as I understand these values have to be (very) small and signed means llvm doesn't have to apply additional logic to do the unsigned comparisons the cpu can't do). Saves a couple of instructions in some test geometry shader here. v2: that was a bit to much optimization, don't skip combining the masks... Reviewed-by: Zack Rusin <zackr@vmware.com>	2013-08-12 23:33:00 +02:00
Roland Scheidegger	84fce45321	draw: (trivial) dump tgsi for geometry shaders with GALLIVM_DEBUG_TGSI And dump the variant key too (same as vs does). Just so I can stop wondering why I see the tgsi dump for fs and vs but not gs...	2013-08-12 23:33:00 +02:00
Roland Scheidegger	8c5283dc17	gallivm: (trivial) fix typo in argument declaration of lp_build_size_query_soa Was meant to match the name used elsewhere, spotted by Anthony.	2013-08-12 23:33:00 +02:00
Chia-I Wu	a9b800aa81	ilo: initialize constant buffer SURFACE_STATE early Fix ilo_gpe_init_view_surface_for_buffer to allow buffer to be NULL, and add ilo_gpe_set_view_surface_bo to set it later. This allows us to set up SURFACE_STATE early for constant buffers backed by user buffers.	2013-08-12 11:49:51 +08:00
Chia-I Wu	b2f79a3823	ilo: 3DSTATE_INDEX_BUFFER may be wrongly skipped In finalize_index_buffer(), when the current index buffer was destroyed due to u_upload_data(), it may happen that the new index buffer is at the same address as the old one. Comparing the pointers to the two buffers could fail to work, and 3DSTATE_INDEX_BUFFER would be incorrectly skipped. Holding a reference to the current index buffer before calling u_upload_data() should fix the problem.	2013-08-10 13:01:41 +08:00
Roland Scheidegger	894d4903e7	gallivm: set non-existing values really to zero in size queries for d3d10 My previous attempt at doing so double-failed miserably (minification of zero still gives one, and even if it would not the value was never written anyway). While here also rename the confusingly named int_vec bld as we have int vecs of different sizes, and rename need_nr_mips (as this also changes out-of-bounds behavior) to is_sviewinfo too. Reviewed-by: Zack Rusin <zackr@vmware.com>	2013-08-09 20:49:19 +02:00
Roland Scheidegger	b0f74250e1	gallivm: use texture target from shader instead of static state for size query d3d10 has no notion of distinct array resources neither at the resource nor sampler view level. However, shader dcl of resources certainly has, and d3d10 expects resinfo to return the values according to that - in particular a resource might have been a 1d texture with some array layers, then the sampler view might have only used 1 layer so it can be accessed both as 1d or 1d array texture (I think - the former definitely works). resinfo of a resource decleared as array needs to return number of array layers but non-array resource needs to return 0 (and not 1). Hence fix this by passing the target from the shader decl to emit_size_query and use that (in case of OpenGL the target will come from the instruction itself). Could probably do the same for actual sampling, though it may not matter there (as the bogus components will essentially get clamped away), possibly could wreak havoc though if it REALLY doesn't match (which is of course an error but still). Reviewed-by: Zack Rusin <zackr@vmware.com>	2013-08-09 20:49:18 +02:00
Roland Scheidegger	38ad404f76	gallivm: honor d3d10's wishes of out-of-bounds behavior for texture size query Specifically, must return 0 for non-existent mip levels (and non-existent textures which is an unsolved problem) for everything but total mip count. Reviewed-by: Zack Rusin <zackr@vmware.com>	2013-08-09 20:49:18 +02:00
Roland Scheidegger	836098f6b2	util: (trivial) fix asm input/output list for fxsave Otherwise gcc might do very unsafe optimizations, spotted by Uros Bizjak. Hopefully this time it's finally right?	2013-08-09 17:30:13 +02:00
Alex Deucher	c88783047e	r600g: disable GPUVM by default Cayman and trinity systems still seem to suffer from stability problems with GPUVM. This also fixes compute on these asics. It can still be enabled for testing by setting env var RADEON_VA=true. Fixes: https://bugs.freedesktop.org/show_bug.cgi?id=65958 Signed-off-by: Alex Deucher <alexander.deucher@amd.com> CC: "9.2" <mesa-stable@lists.freedesktop.org> CC: "9.1" <mesa-stable@lists.freedesktop.org> Reviewed-by: Christian König <christian.koenig@amd.com>	2013-08-09 10:51:25 -04:00
Zack Rusin	e8d8974f80	softpipe: fix the regressions softpipe has a really weird handling of the draw attrs, lets just not inject outputs in its data. Trivial.	2013-08-08 20:54:50 -04:00
Zack Rusin	662a4d4a12	draw: rewrite primitive assembler We can't be injecting the primitive id's in the pipeline because by that time the primitives have already been decomposed. To properly number the primitives we need to handle the adjacency primitives by hand. This patch moves the prim id injection into the original primitive assembler and completely removes the useless pipeline stage. Signed-off-by: Zack Rusin <zackr@vmware.com> Reviewed-by: Roland Scheidegger <sroland@vmware.com>	2013-08-08 20:54:25 -04:00
Zack Rusin	1d425c4c6d	draw: reset the vertex id when injecting new primitive id Without reseting the vertex id, with primitives where the same vertex is used with different primitives (e.g. tri/lines strips) our vbuf module won't re-emit those vertices with the changed primitive id. So lets reset the vertex id whenever injecting new primitive id to make sure that the vertex data is correctly emitted. Signed-off-by: Zack Rusin <zackr@vmware.com> Reviewed-by: Roland Scheidegger <sroland@vmware.com>	2013-08-08 20:54:03 -04:00
Zack Rusin	57cd326778	draw: cleanup the extra attribs Before inserting new front face and prim id outputs cleanup the old extra outputs, otherwise our cache will use previous output slots which will break as soon as outputs of the current shader don't match the last. Signed-off-by: Zack Rusin <zackr@vmware.com> Reviewed-by: Roland Scheidegger <sroland@vmware.com>	2013-08-08 20:53:40 -04:00
Dieter Nützel	8f40fa0e7f	util: (trivial) fix more compile errors in u_cpu_detect (gcc/x86 this time). Oops. Should fix https://bugs.freedesktop.org/show_bug.cgi?id=67921	2013-08-09 01:25:54 +02:00
Roland Scheidegger	43076a55c2	util: (trivial) fix compile error with MSVC on x86	2013-08-08 19:08:57 +02:00
Roland Scheidegger	6ce54a81b2	gallivm: honor d3d10 floating point rules for shadow comparisons d3d10 specifies ordered comparisons for everything but not_equal which is unordered (http://msdn.microsoft.com/en-us/library/windows/desktop/cc308050.aspx). OpenGL probably doesn't care. Reviewed-by: Zack Rusin <zackr@vmware.com>	2013-08-08 18:55:58 +02:00
Roland Scheidegger	aa84f1ad55	softpipe: don't clamp reference value for shadow comparison for float formats Clamping is only done for fixed-point formats as part of conversion to texture format. Reviewed-by: Zack Rusin <zackr@vmware.com>	2013-08-08 18:55:57 +02:00
Roland Scheidegger	e1590b9690	gallivm: don't clamp reference value for shadow comparison for float formats This is wrong both for OpenGL and d3d. (In fact clamping is a side effect of converting to depth format, so this should really do quantization too at least in d3d10 for the comparisons to be truly correct.) Reviewed-by: Zack Rusin <zackr@vmware.com>	2013-08-08 18:55:57 +02:00
Roland Scheidegger	eac57bc223	gallivm: propagate scalar_lod to emit_size_query too Clearly the returned values need to be per-element if the lod is per element. Does not actually change behavior yet. Reviewed-by: Zack Rusin <zackr@vmware.com>	2013-08-08 18:55:57 +02:00
Roland Scheidegger	c8572a9457	gallium: clarify SVIEWINFO opcode This opcode is quite problematic in tgsi, while it tries to mirror d3d10 resinfo it can't really do what's stated there due to missing the crazy return type modifiers. Hence specify this is ignored along with the swizzle. (Other options would be to have multiple opcodes or specify the ret type modifier maybe in dst_reg as there's padding bits left there but it is the only instruction allowing this.) Reviewed-by: Zack Rusin <zackr@vmware.com>	2013-08-08 18:55:57 +02:00
Roland Scheidegger	ce0e66af0a	gallivm: fix out-of-bounds behavior for fetch/ld For d3d10 and ARB_robust_buffer_access_behavior, we are required to return 0 for out-of-bounds coordinates (for which we can just enable the code already there was just disabled). Additionally, also need to return 0 for out-of-bounds mip level and out-of-bounds layer. This changes the logic so instead of clamping the level/layer, an out-of-bound mask is computed instead in this case (actual clamping then can be omitted just like with coordinates, since we set the fetch offset to zero if that happens anyway). Reviewed-by: Zack Rusin <zackr@vmware.com>	2013-08-08 18:55:57 +02:00
Roland Scheidegger	883987503f	util: try much harder to set DAZ flag While so far this only causes some harmless test failures, there's lots more cpus with DAZ. All 64bit capable ones can do it (particularly relevant for AMD cpus as they supported sse3 very very late) but if really necessary we can check support for that for real with some more magic. (In fact just about ANY cpu with sse2 can support DAZ, I believe the only exception are first gen P4 (Willamette) and from those only early steppings which can't do it it's almost like intel forgot to add it... - a real pity though docs say you can't just try to set it as they will throw a GPF.) While this was meant to address https://bugs.freedesktop.org/show_bug.cgi?id=67672 it does not fix it. Most likely the tests need fixing as I don't think there's any guarantee about denorm handling in the reference math library functions if the flags aren't set to standard values. Nevertheless enabling DAZ on all cpus which can do it should be the right thing to do. Reviewed-by: Jose Fonseca <jfonseca@vmware.com>	2013-08-08 18:55:57 +02:00
Roland Scheidegger	e3b5e2db1b	util: implement table-based + linear interpolation linear-to-srgb conversion Should be much faster, seems to work in softpipe. While here (also it's now disabled) fix up the pow factor - the former value is what is in GL core it is however not actually accurate to fp32 standard (as it is 1.0/2.4), and if someone would do all the accurate math there's no reason to waste 8 mantissa bits or so... v2: use real table generating function instead of just printing the values (might take a bit longer as it does calculations on some 3+ million floats but much more descriptive obviously). Also fix up another inaccurate pow factor (this time in the python code) - wondering where the couple one bit errors came from :-(. Reviewed-by: Jose Fonseca <jfonseca@vmware.com> Reviewed-by: Zack Rusin <zackr@vmware.com>	2013-08-08 18:55:57 +02:00
Roland Scheidegger	2d9fea95e8	gallivm: fix comment wrt srgb accuracy. I think it's actually not good enough now...	2013-08-08 18:55:57 +02:00
Chia-I Wu	f9a4288bd2	ilo: get rid of GPE tables completely Move the estimate functions out of the tables and kill the tables.	2013-08-08 13:46:01 +08:00
Chia-I Wu	19204081ce	ilo: clean up GPE header inclusions This reduces the number of source files need to be recompiled when GPE functions are changed other than regular clean ups.	2013-08-08 13:41:10 +08:00
Chia-I Wu	e292b9362a	ilo: initialize alpha test state in ilo_gpe_init_dsa This could speed up BLEND_STATE and COLOR_CALC_STATE emission a bit.	2013-08-08 13:30:34 +08:00
Chia-I Wu	02496cd2b6	ilo: fold gen6_translate_index_size into the caller There is only one caller so fold it.	2013-08-08 13:10:36 +08:00
Chia-I Wu	1c19d0bb81	ilo: fold gen6_translate_depth_format into the caller There is only one caller so fold it.	2013-08-08 13:02:17 +08:00
Courtney Goeltzenleuchter	c2c5366ff2	ilo: Call GPE emit functions directly. Eliminate pipeline and GPE function vectors and have the pipeline functions call the GPE emit functions directly.	2013-08-08 11:39:21 +08:00
Courtney Goeltzenleuchter	4bc9daf923	ilo: move emit functions so that they can be inlined.	2013-08-08 11:39:21 +08:00
Tom Stellard	d0c13fba17	r300g/compiler/tests: Pass the required LDFLAGS when building the test program CC: "9.2 <mesa-stable@lists.freedesktop.org>"	2013-08-07 17:28:19 -07:00
Tom Stellard	d691ba4d94	r300g/compiler/tests: Fix segfault CC: "9.2" <mesa-stable@lists.freedesktop.org>	2013-08-07 17:27:23 -07:00
Kristian Høgsberg	5575fdaccf	gallium-egl: Commit the rest of the native_wayland_drm_bufmgr_helper v2 patch I missed Anders v2 on the list which fixed non-wayland compilation: http://lists.freedesktop.org/archives/mesa-dev/2013-July/042062.html Signed-off-by: Kristian Høgsberg <krh@bitplanet.net>	2013-08-07 11:23:47 -07:00
Ander Conselvan de Oliveira	8d29b5271a	egl: Update to Wayland 1.2 server API Since Wayland 1.2, struct wl_buffer and a few functions are deprecated. References to wl_buffer are replaced with wl_resource and some getter functions and calls to deprecated functions are replaced with the proper new API. The latter changes are related to resource versioning. Signed-off-by: Ander Conselvan de Oliveira <ander.conselvan.de.oliveira@intel.com>	2013-08-07 10:37:58 -07:00
Ander Conselvan de Oliveira	602351dd58	gallium-egl: Don't add a listener for wl_drm twice in wayland platform A listener is added just after the interface is bound, in registry_handle_global(). Signed-off-by: Ander Conselvan de Oliveira <ander.conselvan.de.oliveira@intel.com>	2013-08-07 10:37:58 -07:00
Ander Conselvan de Oliveira	331a8fa41d	gallium-egl: Simplify native_wayland_drm_bufmgr_helper interface The helper provides a series of functions to easy the implementation of the WL_bind_wayland_display extension on different platforms. But even with the helpers there was still a bit of duplicated code between platforms, with the drm authentication being the only part that differs. This patch changes the bufmgr interface to provide a self contained object with a create function that takes a drm authentication callback as an argument. That way all the helper functions are made static and the "_helper" suffix was removed from the sources file name. This change also removes the mix of Wayland client and server code in the wayland drm platform source file. All the uses of libwayland-server are now contained in native_wayland_drm_bufmgr.c. Changes to the drm platform are only compile tested. Signed-off-by: Ander Conselvan de Oliveira <ander.conselvan.de.oliveira@intel.com>	2013-08-07 10:37:58 -07:00

1 2 3 4 5 ...

19000 commits