fdo-mirrors/mesa

mirror of https://gitlab.freedesktop.org/mesa/mesa.git synced 2026-05-20 06:58:16 +02:00

Author	SHA1	Message	Date
Brian Paul	7e8cf34546	svga: add new flush-time HUD query To measure the time spent flushing the command buffer. Reviewed-by: Charmaine Lee <charmainel@vmware.com>	2016-03-07 09:33:15 -07:00
Brian Paul	903afc370f	svga: also dump SVGA3D_BUFFER surfaces in svga_screen_cache_dump() Reviewed-by: Charmaine Lee <charmainel@vmware.com>	2016-03-07 09:33:15 -07:00
Kristian Høgsberg Kristensen	32aa01663f	anv: Quiet pTessellationState warning Some application pass a dummy for pTessellationState which results in a lot of noise. Only warn if we're actually given tessellation shadear stages.	2016-03-06 22:06:24 -08:00
Ilia Mirkin	0941ef3dd5	mesa: flip current tf object back to default if current is being deleted In the rather unusual case of Bind + Delete, we need to make sure that we unbind the current tf object. Fixes dEQP-GLES3.functional.lifetime.delete_bound.transform_feedback Signed-off-by: Ilia Mirkin <imirkin@alum.mit.edu> Reviewed-by: Matt Turner <mattst88@gmail.com>	2016-03-07 00:36:08 -05:00
Ilia Mirkin	f6827e20d1	glsl: avoid stack smashing when there are too many attributes This fixes a crash in dEQP-GLES3.functional.transform_feedback.array_element.separate.points.lowp_mat3x2 and likely others. The vertex shader has > 16 input variables (without explicit locations), which causes us to index outside of the to_assign array. Signed-off-by: Ilia Mirkin <imirkin@alum.mit.edu> Reviewed-by: Timothy Arceri <timothy.arceri@collabora.com> Cc: "11.1 11.2" <mesa-stable@lists.freedesktop.org>	2016-03-07 00:36:08 -05:00
Jason Ekstrand	23de78768b	anv: Create fences from the batch BO pool Applications may create a lot of fences, perhaps as much as one per vkQueueSubmit. Really, they're supposed to use ResetFence, but it's easy enough for us to make them crazy-cheap so we might as well.	2016-03-06 14:26:52 -08:00
Francisco Jerez	3dd0441f6c	i965/vec4: Propagate swizzles correctly during copy propagation. This simplifies the code that iterates over the per-component values found in the matching copy_entry struct and checks whether the register regions that were copied to each component are similar enough to be treated as a single (reswizzled) value which can be propagated into the current instruction. Aside from being scattered between opt_copy_propagation(), try_copy_propagate(), and try_constant_propagate(), what I found terribly confusing about the preexisting logic was that opt_copy_propagation() tried to reorder the array of values according to the swizzle of the instruction source, which meant one would have had to invert the reordering applied at the top level in order to find out which component to take from each value (we were just taking the i-th component from the i-th value, which is not correct in general). The saturate mask was also being swizzled incorrectly. This consolidates the logic for matching multiple components of a copy_entry into a single function which returns the result as a regular src_reg on success, as if the copy had been performed with a single MOV instruction copying all components of the src_reg into the destination. Fixes several ARB_vertex_program MOV test-cases from: https://cgit.freedesktop.org/~kwg/piglit/log/?h=arb_program Acked-by: Matt Turner <mattst88@gmail.com>	2016-03-06 12:22:40 -08:00
Francisco Jerez	c70b7c80e3	i965: Don't try copy propagation if constant propagation succeeded. It cannot get any better. Reviewed-by: Iago Toral Quiroga <itoral@igalia.com> Reviewed-by: Matt Turner <mattst88@gmail.com>	2016-03-06 12:22:40 -08:00
Francisco Jerez	dcf5e19e65	i965/vec4: Use swizzle() to swizzle immediates during constant propagation. Reviewed-by: Iago Toral Quiroga <itoral@igalia.com>	2016-03-06 12:22:40 -08:00
Francisco Jerez	ff7a2b489e	i965: Add support for swizzling arbitrary immediates to (brw_)swizzle(). Scalar immediates used to be handled correctly by swizzle() (as the identity) but since commit `58fa9d47b5` it will corrupt the contents of the immediate. Vector immediates were never handled correctly, but we had ad-hoc code to swizzle VF immediates in the vec4 copy propagation pass. This takes care of swizzling V and UV in addition. v2: Don't implement swizzling of V/UV immediates (Matt). If you need to swizzle an integer vector immediate in the future apply the following diff to go back to v1: --- a/src/mesa/drivers/dri/i965/brw_eu.c +++ b/src/mesa/drivers/dri/i965/brw_eu.c @@ -119,11 +119,10 @@ brw_swap_cmod(uint32_t cmod) static unsigned imm_shift(enum brw_reg_type type, unsigned i) { - assert(type != BRW_REGISTER_TYPE_UV && type != BRW_REGISTER_TYPE_V && - "Not implemented."); - if (type == BRW_REGISTER_TYPE_VF) return 8 * (i & 3); + else if (type == BRW_REGISTER_TYPE_UV \|\| type == BRW_REGISTER_TYPE_V) + return 4 * (i & 7); else return 0; } Reviewed-by: Iago Toral Quiroga <itoral@igalia.com>	2016-03-06 12:22:40 -08:00
Francisco Jerez	537d3df974	i965: Pass symbolic swizzle to brw_swizzle() as a single argument. And replace brw_swizzle1() with brw_swizzle(). Seems slightly cleaner and will allow reusing brw_swizzle() in the vec4 back-end more easily. Reviewed-by: Iago Toral Quiroga <itoral@igalia.com> Reviewed-by: Matt Turner <mattst88@gmail.com>	2016-03-06 12:22:39 -08:00
Ilia Mirkin	ff085d014e	nvc0: reset TFB bufctx when we no longer hold a reference to the buffers This fixes some use-after-free situations in dEQP when an xfb state is removed, and then a clear is triggered, which only does a partial validation. It would attempt to read the no-longer-valid buffers, resulting in crashes. Signed-off-by: Ilia Mirkin <imirkin@alum.mit.edu> Cc: "11.1 11.2" <mesa-stable@lists.freedesktop.org>	2016-03-06 10:14:52 -05:00
Jason Ekstrand	21ee5fd326	anv: Emit null render targets v2 (Francisco Jerez): Add the state_offset to the surface state offset	2016-03-05 20:47:10 -08:00
Ilia Mirkin	fa43c4bd99	nv50/ir: using sampleid/pos shouldn't force per-sample interpolation See https://www.khronos.org/bugzilla/show_bug.cgi?id=1462 Signed-off-by: Ilia Mirkin <imirkin@alum.mit.edu>	2016-03-05 23:26:03 -05:00
Ilia Mirkin	313205cb8f	st/mesa: don't force per-sample interp if only sampleid/pos are used The OES extensions clarify this behaviour to differentiate between per-sample invocation and per-sample interpolation. Using sampleid/pos will force per-sample invocation but not per-sample interpolation. See https://www.khronos.org/bugzilla/show_bug.cgi?id=1462 Signed-off-by: Ilia Mirkin <imirkin@alum.mit.edu> Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2016-03-05 23:26:03 -05:00
Ilia Mirkin	dcbf8377be	swrast: fix GL_ANY_SAMPLES_PASSED values in Result Since commit `922be4eab`, the expectation is that the query result contains the correct value. Unfortunately swrast does not distinguish between GL_SAMPLES_PASSED and GL_ANY_SAMPLES_PASSED. As a result, we must fix up the query result in a post-draw fixup. Bugzilla: https://bugs.freedesktop.org/show_bug.cgi?id=94274 Signed-off-by: Ilia Mirkin <imirkin@alum.mit.edu> Tested-by: Vinson Lee <vlee@freedesktop.org> Reviewed-by: Brian Paul <brianp@vmware.com> Cc: "11.2" <mesa-stable@lists.freedesktop.org>	2016-03-05 23:25:52 -05:00
Jason Ekstrand	8502794c12	anv/pipeline: Handle null wm_prog_data in 3DSTATE_CLIP	2016-03-05 14:42:16 -08:00
Kristian Høgsberg Kristensen	7b348ab8a0	anv: Fix rebase error	2016-03-05 14:33:50 -08:00
Kristian Høgsberg Kristensen	34326f46df	anv: Turn pipeline cache on by default Move the environment variable check to cache creation time so we block both lookups and uploads if it's turned off.	2016-03-05 13:54:24 -08:00
Kristian Høgsberg Kristensen	f2b37132cb	anv: Check if shader if present before uploading to cache Between the initial check the returns NO_KERNEL and compiling the shader, other threads may have added the shader to the cache. Before uploading the kernel, check again (under the mutex) that the compiled shader still isn't present.	2016-03-05 13:54:24 -08:00
Kristian Høgsberg Kristensen	30bbe28b7e	anv: Always use point size from the shader There is no API for setting the point size and the shader is always required to set it. Section 24.4: "If the value written to PointSize is less than or equal to zero, or if no value was written to PointSize, results are undefined." As such, we can just always program PointWidthSource to Vertex. This simplifies anv_pipeline a bit and avoids trouble when we enable the pipeline cache and don't have writes_point_size in the prog_data.	2016-03-05 13:54:24 -08:00
Kristian Høgsberg Kristensen	6139fe9a77	anv: Also cache the struct anv_pipeline_binding maps This is state the we generate when compiling the shaders and we need it for mapping resources from descriptor sets to binding table indices.	2016-03-05 13:50:07 -08:00
Kristian Høgsberg Kristensen	584f39c65e	anv: Don't re-upload shaders when merging Using anv_pipeline_cache_upload_kernel() will re-upload the kernel and prog_data when we merge caches. Since the kernel and prog_data is already in the program_stream, use anv_pipeline_cache_add_entry() instead to only add the entry to the hash table.	2016-03-05 13:50:07 -08:00
Kristian Høgsberg Kristensen	626559ed37	anv: Add anv_pipeline_cache_add_entry() This function will grow the cache to make room and then add the entry.	2016-03-05 13:50:07 -08:00
Kristian Høgsberg Kristensen	07441c344c	anv: Rename anv_pipeline_cache_add_entry() to 'set' This function is a helper that unconditionally sets a hash table entry and expects the cache to have enough room. Calling it 'add_entry' suggests it will grow the cache as needed.	2016-03-05 13:50:07 -08:00
Kristian Høgsberg Kristensen	87967a2c85	anv: Simplify pipeline cache control flow a bit No functional change, but the control flow around searching the cache and falling back to compiling is a bit simpler.	2016-03-05 13:50:07 -08:00
Kristian Høgsberg Kristensen	2b29342fae	anv: Store prog data in pipeline cache stream We have to keep it there for the cache to work, so let's not have an extra copy in struct anv_pipeline too.	2016-03-05 13:50:07 -08:00
Kristian Høgsberg Kristensen	37c5e70253	anv: Rename 'table' to 'hash_table' in anv_pipeline_cache A little less ambiguous.	2016-03-05 13:50:07 -08:00
Kristian Høgsberg Kristensen	c028ffea70	anv: Serialize as much pipeline cache as we can We can serialize as much as the application asks for and just stop once we run out of memory. This lets applications use a fixed amount of space for caching and still get some benefit.	2016-03-05 13:50:07 -08:00
Kristian Høgsberg Kristensen	cd812f086e	anv: Use 1.0 pipeline cache header The final version of the pipeline cache header adds a few more fields.	2016-03-05 13:50:07 -08:00
Kristian Høgsberg Kristensen	26ed943eb9	anv: Fix shader key hashing This was copied from inline code to a helper and wasn't updated to hash a pointer instead.	2016-03-05 13:50:07 -08:00
Kristian Høgsberg Kristensen	3baf8af947	anv: Remove excess whitespace	2016-03-05 13:50:07 -08:00
Kristian Høgsberg Kristensen	ab36eae5e7	anv: Remove left-over bits of sparse-descriptor code	2016-03-05 13:50:07 -08:00
Jason Ekstrand	1afdfc3e6e	anv/pipeline: Implement the depth compare EQUAL workaround on gen8+	2016-03-05 09:59:28 -08:00
Jason Ekstrand	7c1660aa14	anv: Don't allow D16_UNORM to be combined with stencil Among other things, this can cause the depth or stencil test to spurriously fail when the fragment shader uses discard.	2016-03-05 09:59:28 -08:00
Jason Ekstrand	9a90176d48	anv/pipeline: Calculate the correct max_source_attr for 3DSTATE_SBE	2016-03-05 09:59:28 -08:00
Brian Paul	a4678311be	st/mesa: 78-column wrapping in st_extensions.c Reviewed-by: Eduardo Lima Mitev <elima@igalia.com>	2016-03-05 09:21:05 -07:00
Brian Paul	9e6a6bd575	gallium/util: add new comments, assertions in u_debug_refcnt.c Reviewed-by: Eduardo Lima Mitev <elima@igalia.com>	2016-03-05 09:20:34 -07:00
Brian Paul	b6a607b221	gallium/util: update comments and URL in u_debug_refcnt.c Reviewed-by: Eduardo Lima Mitev <elima@igalia.com>	2016-03-05 09:20:28 -07:00
Brian Paul	cbca6964e2	gallium/util: make stream variable static in u_debug_refcnt.c Reviewed-by: Eduardo Lima Mitev <elima@igalia.com>	2016-03-05 09:20:23 -07:00
Brian Paul	fb0abedce7	gallium/util: re-indent u_debug_refcnt.[ch] Wrap comments to 78 columns, etc. Reviewed-by: Eduardo Lima Mitev <elima@igalia.com>	2016-03-05 09:20:14 -07:00
Brian Paul	a7ba29f6d8	gallium/tests: silence warning in compute.c compute.c: In function ‘launch_grid’: compute.c:435:20: warning: assignment discards ‘const’ qualifier from pointer target type [enabled by default] info.input = input; ^ Maybe the pipe_grid_info::input field should be const void *? Reviewed-by: Samuel Pitoiset <samuel.pitoiset@gmail.com>	2016-03-05 09:15:44 -07:00
Timothy Arceri	31943e6ba5	glsl: replace remaining tabs in link_varyings.cpp Reviewed-by: Thomas Helland <thomashelland90@gmail.com>	2016-03-05 20:50:10 +11:00
Timothy Arceri	e2415e8467	glsl: replace remaining tabs in link_uniforms.cpp Reviewed-by: Thomas Helland <thomashelland90@gmail.com>	2016-03-05 20:50:05 +11:00
Jordan Justen	81f30e2f50	anv/hsw: Move query code to genX file for Haswell This fixes many CTS cases, but will require an update to the kernel command parser register whitelist. (The CS GPRs and TIMESTAMP registers need to be whitelisted.) Signed-off-by: Jordan Justen <jordan.l.justen@intel.com>	2016-03-05 01:08:07 -08:00
Timothy Arceri	3322cb7b8d	docs: mark align layout qualifier as DONE Reviewed-by: Ian Romanick <ian.d.romanick@intel.com> Reviewed-by: Samuel Iglesias Gonsálvez <siglesias@igalia.com>	2016-03-05 19:39:13 +11:00
Timothy Arceri	037f68d81e	glsl: apply align layout qualifier rules to block offsets From Section 4.4.5 (Uniform and Shader Storage Block Layout Qualifiers) of the OpenGL 4.50 spec: "The align qualifier makes the start of each block member have a minimum byte alignment. It does not affect the internal layout within each member, which will still follow the std140 or std430 rules. The specified alignment must be a power of 2, or a compile-time error results. The actual alignment of a member will be the greater of the specified align alignment and the standard (e.g., std140) base alignment for the member's type. The actual offset of a member is computed as follows: If offset was declared, start with that offset, otherwise start with the next available offset. If the resulting offset is not a multiple of the actual alignment, increase it to the first offset that is a multiple of the actual alignment. This results in the actual offset the member will have. When align is applied to an array, it affects only the start of the array, not the array's internal stride. Both an offset and an align qualifier can be specified on a declaration. The align qualifier, when used on a block, has the same effect as qualifying each member with the same align value as declared on the block, and gets the same compile-time results and errors as if this had been done. As described in general earlier, an individual member can specify its own align, which overrides the block-level align, but just for that member. Reviewed-by: Ian Romanick <ian.d.romanick@intel.com> Reviewed-by: Samuel Iglesias Gonsálvez <siglesias@igalia.com>	2016-03-05 19:39:07 +11:00
Timothy Arceri	5a27fefffe	glsl: parse align layout qualifier Reviewed-by: Ian Romanick <ian.d.romanick@intel.com> Reviewed-by: Samuel Iglesias Gonsálvez <siglesias@igalia.com>	2016-03-05 19:39:01 +11:00
Timothy Arceri	22b0082b9d	docs: mark explicit byte offsets as DONE Reviewed-by: Edward O'Callaghan <eocallaghan@alterapraxis.com>	2016-03-05 19:38:55 +11:00
Timothy Arceri	802262c0af	glsl: use explicit offset when lowering buffer access Reviewed-by: Edward O'Callaghan <eocallaghan@alterapraxis.com>	2016-03-05 19:38:49 +11:00

... 66 67 68 69 70 ...

82384 commits