fdo-mirrors/mesa

mirror of https://gitlab.freedesktop.org/mesa/mesa.git synced 2026-01-06 08:50:09 +01:00

Author	SHA1	Message	Date
Eric Anholt	97bdb4c039	i965: Don't allocate a 1-level texture when GL_GENERATE_MIPMAP is set. Given that a teximage that calls us with this flag set will immediately proceed to allocate the other levels, we can probably just go ahead and allocate those levels now. Reduces miptree copies in piglit by about .05%. Reviewed-by: Chad Versace <chad.versace@linux.intel.com>	2013-09-30 14:35:42 -07:00
Eric Anholt	6ca9b532d8	i965: Stop allocating miptrees with first_level != 0. If the caller shows up with GL_BASE_LEVEL != 0, it doesn't mean that the texture will over the course of its lifetime have that nonzero baselevel, it means that the caller is filling the texture from the bottom up for some reason (one could imagine demand-loading detailed texture layers at runtime, for example). If we allocate from just the current baselevel, it means when they come along with the next level up, we'll have to allocate a new miptree and copy all of our bits out of the first miptree. Reviewed-by: Chad Versace <chad.versace@linux.intel.com>	2013-09-30 14:35:42 -07:00
Eric Anholt	3b9a2dc938	i965: Drop a special case for guessing small miptree levels. Let's say you started allocating your 2D texture with level 2 of a tree as a 1x1 image. The driver doesn't know if this means that level 0 is 4x4 or 4x1 or 1x4, so we would just allocate a single 1x1 and let it get copied in to the real location at texture validate time later. Since this is just a temporary allocation that will get copied, the extra space allocation of just taking the normal path which will happen to producing a 4x1 level 0, 2x1 level 1, and 1x1 level 2 is the right way to go, to reduce complexity in the normal case. No change in miptree copies over the course of a piglit run. Reviewed-by: Chad Versace <chad.versace@linux.intel.com>	2013-09-30 14:35:42 -07:00
Eric Anholt	7de88ac380	i965: Totally switch around how we handle nonzero baselevel-first_level. This has no effect currently, because intel_finalize_mipmap_tree() always makes mt->first_level == tObj->BaseLevel. The change I made before to handle it (`b1080cfbdb`) got very close to working, but after fixing some unrelated bugs in the series, it still left tex-miplevel-selection producing errors when testing textureLod(). The problem is that for explicit LODs, the sampler's LOD clamping is ignored, and only the surface's MIP clamping is respected. So we need to use surface mip clamping, which applies on top of the sampler's mip clamping, so the sampler change gets backed out. Now actually tested with a non-regressing series producing a non-zero computed baselevel. Reviewed-by: Chad Versace <chad.versace@linux.intel.com>	2013-09-30 14:35:42 -07:00
Eric Anholt	9c116d5eac	i965: Always look up from the object's mt when setting up texturing state. We know that the object's mt is equal to the firstimage's mt because it's gone through intel_finalize_mipmap_tree(). Saves a lookup of firstimage on pre-gen7. v2: Merge in the warning fix that appeared later in the series (noted by Chad) Reviewed-by: Chad Versace <chad.versace@linux.intel.com>	2013-09-30 14:35:42 -07:00
Vinson Lee	114ae47475	r600g/sb: Move variable dereference after null check. Fixes "Deference before null check" defect reported by Coverity. Signed-off-by: Vinson Lee <vlee@freedesktop.org> Reviewed-by: Vadim Girlin <vadimgirlin@gmail.com>	2013-09-30 10:27:52 -07:00
Brian Paul	0d441aac3d	st/mesa: fix comment typo	2013-09-30 09:06:52 -06:00
Marek Olšák	7b25f52a95	r600g,radeonsi: workaround for late shared screen initialization Accidentally broken by the consolidation.	2013-09-30 13:01:13 +02:00
Laurent Carlier	868791f0ba	r600g: Fix build failure introduced with r600_texture.c consolidation It seems that case with opencl enabled was forgotten Signed-off-by: Marek Olšák <marek.olsak@amd.com>	2013-09-29 22:01:04 +02:00
Marek Olšák	4e9aa6711f	radeon: make texture logging more useful This has been very useful for tracking down bugs in libdrm. The *_PRINT_TEXDEPTH environment variables were probably never used, so I removed them.	2013-09-29 15:18:10 +02:00
Marek Olšák	e64633e8c3	r600g,radeonsi: share r600_texture.c The function r600_choose_tiling is new and needs a review. The only change in functionality is that it enables 2D tiling for compressed textures on SI. It was probably accidentally turned off. v2: don't make scanout buffers linear	2013-09-29 15:18:10 +02:00
Marek Olšák	4069d39465	r600g: remove compute_global_transfer_* calls from texture_transfer_map/unmap Textures can never have target==PIPE_BUFFER.	2013-09-29 15:18:10 +02:00
Marek Olšák	ef6680d3ee	r600g: move the low-level buffer functions for multiple rings to drivers/radeon Also slightly optimize r600_buffer_map_sync_with_rings.	2013-09-29 15:18:09 +02:00
Marek Olšák	1bb77f81db	r600g,radeonsi: consolidate tiling_info initialization and the util_format_s3tc_init calls too.	2013-09-29 15:18:09 +02:00
Marek Olšák	09fc5d6e26	radeonsi: implement clear_buffer using CP DMA, initialize CMASK with it More work needs to be done for this to be entirely shared with r600g. I'm just trying to share r600_texture.c now. The reason I put the implementation to si_descriptors.c is that the emit function had already been there.	2013-09-29 15:18:09 +02:00
Marek Olšák	68f6dec32e	r600g: move aux_context and r600_screen_clear_buffer to drivers/radeon This will be used in the next commit.	2013-09-29 15:18:09 +02:00
Marek Olšák	0cb9de1dd0	radeonsi: move debug options to R600_DEBUG	2013-09-29 15:18:09 +02:00
Marek Olšák	ba650ccf91	r600g: move some debug options to drivers/radeon	2013-09-29 15:18:09 +02:00
Marek Olšák	2814202ef4	r600g,radeonsi: share the async dma interface r600_texture.c is one step closer to r600g.	2013-09-29 15:18:09 +02:00
Marek Olšák	e916267285	radeonsi: move radeonsi-specific functions out of r600_texture.c	2013-09-29 15:18:08 +02:00
Marek Olšák	31169400a0	r600g,radeonsi: remove unused code	2013-09-29 15:18:08 +02:00
Marek Olšák	6f21009cb3	r600g: move r600g-specific functions out of r600_texture.c	2013-09-29 15:18:08 +02:00
Marek Olšák	bfea9c498d	r600g,radeonsi: consolidate r600_texture structures	2013-09-29 15:18:08 +02:00
Marek Olšák	4ea2e5a4e7	r600g: get rid of r600_texture::is_rat It's always 0.	2013-09-29 15:18:08 +02:00
Marek Olšák	ba29324dba	r600g: get rid of r600_texture::array_mode	2013-09-29 15:18:08 +02:00
Marek Olšák	39801d4ba7	r600g,radeonsi: consolidate transfer, cmask, and fmask structures	2013-09-29 15:18:08 +02:00
Marek Olšák	a62cd6949c	radeon drivers: handle PIPE_CAP_MAX_VIEWPORTS	2013-09-29 15:18:07 +02:00
Marek Olšák	900b1863c8	radeon/llvm: fix TGSI_OPCODE_UCMP This doesn't fix any known issue (I haven't run piglit with this yet), but the code was obviously completely wrong. It looks like copy-pasted from CMP. Reviewed-by: Tom Stellard <thomas.stellard@amd.com>	2013-09-29 14:49:23 +02:00
Marek Olšák	2bda5f3298	st/mesa: fix GLSL mix(.., .., bvecN) v2: use CMP on drivers without native integer support	2013-09-29 14:42:42 +02:00
Tom Stellard	a64d3dd135	configure.ac: Add a more informative warning when libclc.pc is not found v2 v2: - Don't display an error message when the user doesn't ask for libclc. Reviewed-by: Matt Turner <mattst88@gmail.com>	2013-09-27 20:20:35 -07:00
Vinson Lee	b2d5757831	mesa: Include stdint.h in mtypes.h for uint32_t symbol. This patch fixes the MSVC build error introduced with commit `b2e327e08f`. api_arrayelt.c src\mesa\main/mtypes.h(1809) : error C2061: syntax error : identifier 'uint32_t' src\mesa\main/mtypes.h(1810) : error C2059: syntax error : '}' src\mesa\main/mtypes.h(1825) : error C2079: 'Minimum' uses undefined union 'gl_perf_monitor_counter_value' src\mesa\main/mtypes.h(1828) : error C2079: 'Maximum' uses undefined union 'gl_perf_monitor_counter_value' Signed-off-by: Vinson Lee <vlee@freedesktop.org>	2013-09-26 20:48:47 -07:00
Kenneth Graunke	aac75f877d	i965/fs: Don't double-accept operands of logical and/or/xor operations. If the argument to emit_bool_to_cond_code() is an ir_expression, we loop over the operands, calling accept() on each of them, which generates assembly code to compute that subexpression. We then emit one or two final instruction that perform the top-level operation on those operands. If it's not an expression (say, a boolean-valued variable), we simply call accept() on the whole value. In commit `80ecb8f1` (i965/fs: Avoid generating extra AND instructions on bool logic ops), Eric made logic operations jump out of the expression path to the non-expression path. Unfortunately, this meant that we would first accept() the two operands, skip generating any code that used them, then accept() the whole expression, generating code for the operands a second time. Dead code elimination would always remove the first set of redundant operand assembly, since nothing actually used them. But we shouldn't generate it in the first place. Signed-off-by: Kenneth Graunke <kenneth@whitecape.org> Reviewed-by: Matt Turner <mattst88@gmail.com>	2013-09-26 16:55:18 -07:00
Kenneth Graunke	e5c49bc25b	i965: Add #define for MI_REPORT_PERF_COUNT on Gen6+. This appears in Volume 1 Part 1 of the Sandybridge PRM on page 48. Signed-off-by: Kenneth Graunke <kenneth@whitecape.org>	2013-09-26 16:55:18 -07:00
Kenneth Graunke	0f2da77307	i965: Add support for GL_AMD_performance_monitor on Ironlake. Ironlake's counters are always enabled; userspace can simply send a MI_REPORT_PERF_COUNT packet to take a snapshot of them. This makes it easy to implement. The counters are documented in the source code for the intel-gpu-tools intel_perf_counters utility. v2: Adjust for core data structure changes. Add a table mapping buffer object offsets to exposed counters (which changes each generation). Finally, add report ID assertions to sanity check the BO layout (thanks to Carl Worth). v3: Update for core BeginPerfMonitor hook changes (requested by Brian). Signed-off-by: Kenneth Graunke <kenneth@whitecape.org> Reviewed-by: Ian Romanick <ian.d.romanick@intel.com>	2013-09-26 16:55:18 -07:00
Kenneth Graunke	b2e327e08f	mesa: Add core support for the GL_AMD_performance_monitor extension. This provides an interface for applications (and OpenGL-based tools) to access GPU performance counters. Since the exact performance counters available vary between vendors and hardware generations, the extension provides an API the application can use to get the names, types, and minimum/maximum values of all available counters. Counters are also organized into groups. Applications create "performance monitor" objects, select the counters they want to track, and Begin/End monitoring, much like OpenGL's query API. Multiple monitors can be in flight simultaneously. v2: Pass ctx to all driver hooks (suggested by Christoph), and attempt to fix overallocation of bitsets (caught by Christoph). Incomplete. v3: Significantly rework core data structures. Store counters in groups rather than in a global list. Use their array index in the group's counter list as the ID rather than trying to store a globally unique counter ID. Use bitsets for active counters within a group, and also track which groups are active so that's easy to query. v4: Remove _mesa_ prefix on static functions; detect out of memory conditions in new_performance_monitor(); make BeginPerfMonitor hook return a boolean rather than setting m->Active or raising an error. Switch to GLuint/unsigned for NumGroups, NumCounters, and MaxActiveCounters (which also means switching a bunch of temporary variable types). All suggested by Brian Paul. Also, remove commented out code at the bottom of the block. Finally, fix the dispatch sanity test (noticed by Ian Romanick). Signed-off-by: Kenneth Graunke <kenneth@whitecape.org> Reviewed-by: Brian Paul <brianp@vmware.com> [v3] Reviewed-by: Ian Romanick <ian.d.romanick@intel.com>	2013-09-26 16:55:18 -07:00
Kenneth Graunke	f91475d4ab	glsl: Create and use a has_uniform_buffer_objects() helper. This is better than overriding the extension enable based on the language version; it's robust against shaders that do: #version 140 #extension GL_ARB_uniform_buffer_object : disable Signed-off-by: Kenneth Graunke <kenneth@whitecape.org> Reviewed-by: Ian Romanick <ian.d.romanick@intel.com>	2013-09-26 16:55:18 -07:00
Kenneth Graunke	e4af55c78f	glsl: Create and use a has_explicit_attrib_location() helper. Explicit attribute locations are supported with GLSL 3.30, GLSL ES 3.00, or "#extension GL_ARB_explicit_attrib_location: enable". Using a helper function makes it easy to check for this. This enables support in GLSL 3.30, which was previously missing. Previously, we overrode the extension enable flag for ES 3.00. This is not robust against a shader such as: #version 330 #extension GL_ARB_explicit_attrib_location : disable Disabling extensions should not remove core language functionality. Signed-off-by: Kenneth Graunke <kenneth@whitecape.org> Reviewed-by: Ian Romanick <ian.d.romanick@intel.com>	2013-09-26 16:55:18 -07:00
Kenneth Graunke	e9b410b54d	mesa: Remove 'invalidate_state' parameter to _mesa_dirty_texobj(). Every caller passed true. Signed-off-by: Kenneth Graunke <kenneth@whitecape.org> Reviewed-by: Eric Anholt <eric@anholt.net>	2013-09-26 16:55:18 -07:00
Eric Anholt	1c904466aa	mesa: Remove some remaining FEATURE_* detritus. Reviewed-by: Matt Turner <mattst88@gmail.com> Reviewed-by: Kenneth Graunke <kenneth@whitecape.org>	2013-09-26 16:29:39 -07:00
Chris Forbes	fe2528c0b6	i965: Fix cube array coordinate normalization Hardware requires the magnitude of the largest component to not exceed 1; brw_cubemap_normalize ensures that this is the case. Unfortunately, we would previously multiply the array index for cube arrays by the normalization factor. The incorrect array index would then cause the sampler to attempt to access either the wrong cube, or memory outside the cube surface entirely, resulting in garbage rendering or in the worst case, hangs. Alter the normalization pass to only multiply the .xyz components. Fixes broken rendering in the arb_texture_cube_map_array-cubemap piglit, which was recently adjusted to provoke this behavior. V2: Fix indent. Signed-off-by: Chris Forbes <chrisf@ijw.co.nz> Cc: "9.2" mesa-stable@lists.freedesktop.org Reviewed-by: Eric Anholt <eric@anholt.net>	2013-09-26 18:24:22 +12:00
Zack Rusin	d83ef680e2	draw/clip: don't emit so many empty triangles Compress empty triangles (don't emit more than one in a row) and never emit empty triangles if we already generated a triangle covering a non-null area. We can't skip all null-triangles because c_primitives expects ones that were generated from vertices exactly at the clipping-plane, to be emitted. Signed-off-by: Zack Rusin <zackr@vmware.com> Reviewed-by: José Fonseca <jfonseca@vmware.com> Reviewed-by: Roland Scheidegger <sroland@vmware.com>	2013-09-25 19:42:22 -04:00
Zack Rusin	60c448faea	llvmpipe: count c_primitives before discarding null prims We need to count the clipper primitives before the rasterizer discards one it considers to be null. Signed-off-by: Zack Rusin <zackr@vmware.com> Reviewed-by: José Fonseca <jfonseca@vmware.com> Reviewed-by: Roland Scheidegger <sroland@vmware.com>	2013-09-25 19:41:02 -04:00
Zack Rusin	1291e833e7	llvmpipe: we need to subdivide if fb is bigger in either direction We need to subdivide triangles if either of the dimensions is larger than the max edge length, not when both of them are larger. Signed-off-by: Zack Rusin <zackr@vmware.com> Reviewed-by: José Fonseca <jfonseca@vmware.com> Reviewed-by: Roland Scheidegger <sroland@vmware.com>	2013-09-25 19:38:21 -04:00
Marek Olšák	028b26e2ef	radeon/llvm: fix shadow cube texturing for GL3.0 The fix is at the end (TGSI_TEXTURE_SHADOWCUBE handling), but I also restructured the code for it to be more readable. Fixes spec/!OpenGL 3.0/sampler-cube-shadow. Reviewed-by: Michel Dänzer <michel.daenzer@amd.com>	2013-09-25 20:45:23 +02:00
Marek Olšák	57f38e9f92	radeonsi: fix blitting the last 2 mipmap levels of compressed textures This fixes compressedteximage piglit tests. +10 piglits Evergreen and Cayman have the same issue. R600 and R700 don't. Cc: "9.2" <mesa-stable@lists.freedesktop.org> Reviewed-by: Michel Dänzer <michel.daenzer@amd.com>	2013-09-25 20:45:22 +02:00
Marek Olšák	296adb6de9	radeonsi: add missing colorbuffer formats (rework format translation) This fixes some piglits, e.g: spec/!OpenGL 3.0/required-renderbuffer-attachment-formats. This can be ported to r600g. Reviewed-by: Michel Dänzer <michel.daenzer@amd.com>	2013-09-25 20:45:22 +02:00
Marek Olšák	f9ea435ebc	radeonsi: bypass alpha-test for integer colorbuffers Fixes spec/EXT_texture_integer/fbo-blending. Reviewed-by: Michel Dänzer <michel.daenzer@amd.com>	2013-09-25 20:45:22 +02:00
Marek Olšák	f7d004b9ad	r600g: fix texture buffer object cache flushing Cc: "9.2" <mesa-stable@lists.freedesktop.org>	2013-09-25 20:45:22 +02:00
Marek Olšák	6317a3fb31	r600g: fix constant buffer cache flushing Cc: "9.2" <mesa-stable@lists.freedesktop.org>	2013-09-25 20:45:22 +02:00
Christian König	4871128e58	radeon/winsys: keep screen pointer in winsys v2 Only create one screen for each winsys instance. This helps with buffer sharing and interop handling. v2: rebased and some minor cleanup Signed-off-by: Christian König <christian.koenig@amd.com> Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2013-09-25 19:41:31 +02:00

1 2 3 4 5 ...

58690 commits