mesa.git - Unnamed repository; edit this file 'description' to name the repository.

	Commit message (Collapse)	Author	Age	Files	Lines
*	st/nine: Partial software vertex processing support	Axel Davy	2016-10-10	8	-60/+322
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	Software Vertex Processing allows: . Less limitations for shaders (more loops, etc) . Less limitations for ff (more enabled lights, 255 matrices for VertexBlend) In particular shaders can get more constants. This patch implements support for this (not using software rendering, but hardware rendering, as llvmpipe and dx10+ hw have the same limits...) This is considered a second class path. Even apps asking for "Mixed Vertex processing" (ie the ability to switch to swvp on demand) do not use the feature much. Some just initialize more constants than the normal limit at the start of the application, but never use more than the normal limit. When the apps do not need the software vertex processing features, they do not seem to turn it on. This means it is ok if that path is slow. Thus no care has been made to make the path optimized. Signed-off-by: Axel Davy <[email protected]>
*	st/nine: Rework vs int and bool constants buffer	Axel Davy	2016-10-10	4	-24/+37
\| \| \| \| \| \|	This will help to support swvp constants. Signed-off-by: Axel Davy <[email protected]>
*	st/nine: Change dirty tracking for vs int and bool constants	Axel Davy	2016-10-10	4	-22/+56
\| \| \| \| \| \| \|	This change makes easier to introduce tracking for swvp constants. Signed-off-by: Axel Davy <[email protected]>
*	st/nine: Drop unused constant upload path	Axel Davy	2016-10-10	3	-210/+4
\| \| \| \| \| \| \| \| \|	This path has been disabled for some time because of some bugs with it. It hasn't been updated to the new features, and is not faster. Signed-off-by: Axel Davy <[email protected]> Reviewed-by: Patrick Rudolph <[email protected]>
*	st/nine: Add support for swvp constants in shaders	Axel Davy	2016-10-10	3	-37/+125
\| \| \| \| \| \| \|	swvp has relaxed limits (more nested loops, etc). In particular it enables more constants. Signed-off-by: Axel Davy <[email protected]>
*	st/nine: Initial mixed vertex processing support	Axel Davy	2016-10-10	2	-5/+16
\| \| \| \| \| \| \| \| \| \|	In mixed vertex processing, the user can enable or disable software vertex processing. It is on hardware by default. This feature is not a state, and thus the setting doesn't need to be recorded by stateblocks. Signed-off-by: Axel Davy <[email protected]>
*	st/nine: Implement SetNPatchMode	Axel Davy	2016-10-10	1	-1/+1
\| \| \| \|	Signed-off-by: Axel Davy <[email protected]>
*	st/nine: Implement D3DUSAGE_SOFTWAREPROCESSING	Axel Davy	2016-10-10	1	-2/+3
\| \| \| \| \| \| \| \|	Buffers with this flag must be usable with both software and hardware vertex processing. Use Staging for fast cpu access. Signed-off-by: Axel Davy <[email protected]> Reviewed-by: Patrick Rudolph <[email protected]>
*	st/nine: Allocate more space for ATI1	Patrick Rudolph	2016-10-10	2	-6/+17
\| \| \| \| \| \| \| \| \| \| \| \| \| \|	ATIx are "unknown" formats that do not follow block format conventions. Tests showed that pitch*height bytes are allocated. apitrace used to depend on this behaviour. It used to copy more bytes than it has to for the ATI1 block format, but it didn't crash on Windows. Increase buffersize for ATI1 to fix this crash. The same issue was present in WINE but a patch has been sent by me. Signed-off-by: Patrick Rudolph <[email protected]> Reviewed-by: Axel Davy <[email protected]>
*	st/nine: Add missing break	Patrick Rudolph	2016-10-10	1	-0/+1
\| \| \| \| \| \| \|	Add missing break instruction. Signed-off-by: Patrick Rudolph <[email protected]> Reviewed-by: Axel Davy <[email protected]>
*	st/nine: Implement relative addressing for ps inputs	Axel Davy	2016-10-10	1	-3/+27
\| \| \| \| \| \| \| \| \|	To implement the feature we copy the ps inputs to a temp array. This is not optimal for performance, but it is the simplest solution. This is a feature that is very very rarely used. Signed-off-by: Axel Davy <[email protected]>
*	st/nine: Wait for pending tasks to execute in swapchain	Axel Davy	2016-10-10	1	-0/+4
\| \| \| \| \| \|	Fixes crash after Reset() when using thread_submit=true Signed-off-by: Axel Davy <[email protected]>
*	st/nine: Use fixed size arrays for swapchain buffers	Axel Davy	2016-10-10	2	-40/+11
\| \| \| \|	Signed-off-by: Axel Davy <[email protected]>
*	st/nine: Fix buffer count check for Ex devices	Patrick Rudolph	2016-10-10	1	-5/+4
\| \| \| \| \|	Signed-off-by: Patrick Rudolph <[email protected]> Reviewed-by: Axel Davy <[email protected]>
*	st/nine: Disable seamless cubemap for d3d	Axel Davy	2016-10-10	2	-2/+2
\| \| \| \| \| \|	d3d9 doesn't have seamless cubemap. Signed-off-by: Axel Davy <[email protected]>
*	st/nine: Fix some check flags	Axel Davy	2016-10-10	1	-11/+12
\| \| \| \| \| \| \|	Uses the new defines introduced in previous commit. See comment in the commit for more explanation. Signed-off-by: Axel Davy <[email protected]>
*	st/nine: Unify some check flags	Axel Davy	2016-10-10	2	-4/+12
\| \| \| \| \| \|	The new defines will be reused in a later patch. Signed-off-by: Axel Davy <[email protected]>
*	gallium/util: Really allow aliasing of dst for u_box_union_*	Axel Davy	2016-10-10	1	-11/+20
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	Gallium nine relies on aliasing to work with this function. Without this patch, dirty region tracking was incorrect, which could lead to incorrect textures or vertex buffers. Fixes several game bugs with nine. Fixes https://github.com/iXit/Mesa-3D/issues/234 Signed-off-by: Axel Davy <[email protected]> Reviewed-by: Patrick Rudolph <[email protected]> Reviewed-by: Nicolai Hähnle <[email protected]> Reviewed-by: Roland Scheidegger <[email protected]> Reviewed-by: Edward O'Callaghan <[email protected]> Cc: "12.0" <[email protected]>
*	softpipe: Cap to 2 GB on 32 bits	Axel Davy	2016-10-10	1	-0/+6
\| \| \| \| \| \| \| \| \| \| \| \| \| \|	On 32 bits system, application memory is quite limited. softpipe uses application memory. To help prevent memory exhaustion, limit reported memory availability to 2GB. Some gallium nine apps do check reported memory by allocating resources until memory is full. Gallium nine refuses allocations when 80% of the reported memory limit is used. This change helps some apps to start. Signed-off-by: Axel Davy <[email protected]> Reviewed-by: Roland Scheidegger <[email protected]>
*	llvmpipe: Cap to 2 GB on 32 bits	Axel Davy	2016-10-10	1	-0/+6
\| \| \| \| \| \| \| \| \| \| \| \| \| \|	On 32 bits system, application memory is quite limited. llvmpipe uses application memory. To help prevent memory exhaustion, limit reported memory availability to 2GB. Some gallium nine apps do check reported memory by allocating resources until memory is full. Gallium nine refuses allocations when 80% of the reported memory limit is used. This change helps some apps to start. Signed-off-by: Axel Davy <[email protected]> Reviewed-by: Roland Scheidegger <[email protected]>
*	gallium/os: Fix overflow on 32 bits	Axel Davy	2016-10-10	1	-4/+10
\| \| \| \| \| \| \| \| \| \| \| \| \|	On systems with more than 4GB of ram, os_get_total_physical_memory was triggering an integer overflow for the linux and haiku path, when on 32 bits. Bugzilla: https://bugs.freedesktop.org/show_bug.cgi?id=94561 Signed-off-by: Axel Davy <[email protected]> Reviewed-by: Roland Scheidegger <[email protected]> Reviewed-by: Marek Olšák <[email protected]>
*	st/nine: Memset pipe_resource templates	Axel Davy	2016-10-10	2	-0/+8
\| \| \| \| \| \| \| \| \| \| \| \|	Fixes regression introduced by ecd6fce2611e88ff8468a354cff8eda39f260a31 and is more future proof than just clearing the next field. Other nine usages did already zero out the templates. Signed-off-by: Axel Davy <[email protected]> Acked-by: Edward O'Callaghan <[email protected]>
*	nvc0: fix valid range for shader buffers	Samuel Pitoiset	2016-10-10	3	-0/+3
\| \| \| \| \| \| \| \|	When offset != 0, the valid range was wrong because the second argument of util_range_add() is end, not size. Signed-off-by: Samuel Pitoiset <[email protected]> Reviewed-by: Ilia Mirkin <[email protected]>
*	nvc0/ir: fix overwriting of value backing non-constant gather offset	Ilia Mirkin	2016-10-10	1	-2/+2
\| \| \| \| \| \| \| \| \| \| \| \| \|	Normally the value is an immediate, which is moved to some temporary, so there's no problem. In the case of a non-constant offset (as allowed by ARB_gpu_shader5), we have to take care to copy it first before using it to build up the bits. This fixes a compilation error observed in F1 2015. Signed-off-by: Ilia Mirkin <[email protected]> Reviewed-by: Samuel Pitoiset <[email protected]> Cc: [email protected]
*	glsl: Add missing cache_destroy stub function.	Vinson Lee	2016-10-10	1	-0/+5
\| \| \| \| \| \| \| \| \| \| \| \|	CC glsl/tests/cache_test.o glsl/tests/cache_test.c: In function ‘test_cache_create’: glsl/tests/cache_test.c:160:4: error: implicit declaration of function ‘cache_destroy’ [-Werror=implicit-function-declaration] cache_destroy(cache); ^ Fixes: 87ab26b2ab35 ("glsl: Add initial functions to implement an on-disk cache") Signed-off-by: Vinson Lee <[email protected]> Reviewed-by: Timothy Arceri <[email protected]>
*	egl: Unify the EGLint/EGLAttrib paths in eglCreateSync* (v3)	Chad Versace	2016-10-10	5	-52/+34
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	Pre-patch, there were two code paths for parsing EGLSync attribute lists: one path for old-style EGLint lists, used by eglCreateSyncKHR, and another for new-style EGLAttrib lists, used by eglCreateSync (1.5) and eglCreateSync64 (EGL_KHR_cl_event2). There were two attrib_list parsing functions, _eglParseSyncAttribList(_EGLSync sync, const EGLint attrib_list) _eglParseSyncAttribList64(_EGLSync sync, const EGLattrib attrib_list) This patch unifies the two attrib_list parsing functions into one, _eglParseSyncAttribList(_EGLSync sync, const EGLattrib attrib_list) Many internal EGLSync function signatures had two attrib_list parameters to accomodate both code paths: one parameter was an EGLint list and other an EGLAttrib list. At most one of the parameters was allowed to be non-null. This patch removes the `EGLint attrib_list` parameter, leaving only the `EGLAttrib attrib_list` parameter, for all internal EGLSync functions. v2: - Consistently use condition (sizeof(int_list[0]) == sizeof(attrib_list[0])). [for emil] v3: - Don't double-unlock the display in eglCreateSyncKHR. Reviewed-by: Emil Velikov <[email protected]> (v2)
*	intel: Fix bash-specific redirection.	Eric Anholt	2016-10-10	1	-1/+1
\| \| \| \|	Reviewed-by: Lionel Landwerlin <[email protected]>
*	nv50/ir: only stick one preret per function	Ilia Mirkin	2016-10-10	1	-4/+7
\| \| \| \| \| \| \| \| \| \| \|	A function with multiple returns would have had multiple preret settings at the top of the function. While this is unlikely to have caused issues since we don't use functions in earnest, it could have in some cases overflowed the call stack, in case a function had a lot of early returns. Signed-off-by: Ilia Mirkin <[email protected]> Reviewed-by: Samuel Pitoiset <[email protected]>
*	radeonsi: make more use of si_have_tgsi_compute	Nicolai Hähnle	2016-10-10	1	-3/+1
\| \| \| \| \|	Reviewed-by: Edward O'Callaghan <[email protected]> Reviewed-by: Marek Olšák <[email protected]>
*	gallium/radeon: assign a name to LLVM output variables in debug builds	Nicolai Hähnle	2016-10-10	1	-1/+6
\| \| \| \| \| \| \|	This can be helpful with R600_DEBUG=preoptir. Reviewed-by: Edward O'Callaghan <[email protected]> Reviewed-by: Marek Olšák <[email protected]>
*	gallium/radeon: avoid redundant work with overlapping in/out arrays	Nicolai Hähnle	2016-10-10	1	-1/+4
\| \| \| \| \|	Reviewed-by: Edward O'Callaghan <[email protected]> Reviewed-by: Marek Olšák <[email protected]>
*	radeonsi: support ARB_compute_variable_group_size	Nicolai Hähnle	2016-10-10	5	-17/+53
\| \| \| \| \| \| \| \|	Not sure if it's possible to avoid programming the block size twice (once for the userdata and once for the dispatch). Reviewed-by: Edward O'Callaghan <[email protected]> Reviewed-by: Marek Olšák <[email protected]>
*	anv: turn on samplerAnisotropy in VkPhysicalDeviceFeatures	Lionel Landwerlin	2016-10-10	2	-2/+1
\| \| \| \| \| \| \| \| \| \| \| \| \| \|	According to the Vulkan spec 5.63.4 : samplerAnisotropy indicates whether anisotropic filtering is supported. If this feature is not enabled, the maxAnisotropy member of the VkSamplerCreateInfo structure must be 1.0. Since we already set maxAnisotropy to 16 and program the hardware according to the VkSamplerCreateInfo.maxAnisotropy, it seems we can turn this on. Signed-off-by: Lionel Landwerlin <[email protected]> Reviewed-by: Jason Ekstrand <[email protected]>
*	radv: Use proper header guards over 'pragma once' directives	Edward O'Callaghan	2016-10-10	13	-14/+60
\| \| \| \| \|	Signed-off-by: Edward O'Callaghan <[email protected]> Reviewed-by: Eric Engestrom <[email protected]>
*	mesa: throw error if bufSize negative in GetSynciv on OpenGL ES	Tapani Pälli	2016-10-10	1	-0/+8
\| \| \| \| \| \| \| \| \| \| \| \| \| \|	Fixes following dEQP tests: dEQP-GLES31.functional.debug.negative_coverage.callbacks.state.get_synciv dEQP-GLES31.functional.debug.negative_coverage.get_error.state.get_synciv dEQP-GLES31.functional.debug.negative_coverage.log.state.get_synciv v2: drop _mesa_is_gles check (Kenneth) Signed-off-by: Tapani Pälli <[email protected]> Bugzilla: https://bugs.freedesktop.org/show_bug.cgi?id=98133 Reviewed-by: Kenneth Graunke <[email protected]>
*	glsl: prohibit lowp, mediump precision on atomic_uint	Tapani Pälli	2016-10-10	1	-0/+14
\| \| \| \| \| \| \| \| \| \| \| \|	Fixes following dEQP tests: dEQP-GLES31.functional.debug.negative_coverage.callbacks.atomic_counter.atomic_precision dEQP-GLES31.functional.debug.negative_coverage.get_error.atomic_counter.atomic_precision dEQP-GLES31.functional.debug.negative_coverage.log.atomic_counter.atomic_precision Signed-off-by: Tapani Pälli <[email protected]> Bugzilla: https://bugs.freedesktop.org/show_bug.cgi?id=98131 Reviewed-by: Kenneth Graunke <[email protected]>
*	glsl: optimize copy_propagation_elements pass	Tapani Pälli	2016-10-10	1	-50/+148
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	Changes make copy_propagation_elements pass faster, reducing link time spent in test case of bug 94477. Does not fix the actual issue but brings down the total time. No regressions seen in CI. v2 (idr): Formatting / whitespace fixes. Embed the acp_ref in the acp_entry. v3 (idr): Delete unused copy constructor. Use while(pop_head) instead of foreach() { remove }. Signed-off-by: Tapani Pälli <[email protected]> Signed-off-by: Ian Romanick <[email protected]> Reviewed-by: Kenneth Graunke <[email protected]>
*	intel: aubinator: enable loading dumps from standard input	Lionel Landwerlin	2016-10-08	1	-36/+129
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	In conjuction with an intel_aubdump change, you can now look at your application's output like this : $ intel_aubdump -c '/path/to/aubinator --gen=hsw' my_gl_app v2: Add print_help() comment about standard input handling (Eero) Remove shrinked gtt space debug workaround (Eero) v3: Use realloc rather than memcpy/free (Ben) Signed-off-by: Lionel Landwerlin <[email protected]> Reviewed-by: Sirisha Gandikota <[email protected]>
*	intel: aubinator: enable loading xml files from a given directory	Lionel Landwerlin	2016-10-08	3	-3/+81
\| \| \| \| \| \| \|	This might be useful for people who debug with out of tree descriptions. Signed-off-by: Lionel Landwerlin <[email protected]> Reviewed-by: Sirisha Gandikota <[email protected]>
*	intel: aubinator: generate a standalone binary	Lionel Landwerlin	2016-10-08	6	-53/+102
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	Embed the xml files into the binary, so aubinator can be used from any location. v2: Split generation packing into another patch (Jason) Check for xxd (Jason) v3: Fix out of tree builds (Jason) Generate custom variable name rather than names generated by xxd (Lionel) v4: Move generated _xml.h files to genxml/ (Sirisha) v5: Remove newline from makefile (Jason) v6: Add comment on gen*_xml.h creation (Jason) Signed-off-by: Lionel Landwerlin <[email protected]> Reviewed-by: Jason Ekstrand <[email protected]>
*	anv/TODO: Update the HiZ task	Nanley Chery	2016-10-07	1	-1/+1
\| \| \| \| \| \|	Signed-off-by: Nanley Chery <[email protected]> Reviewed-by: Jason Ekstrand <[email protected]> Reviewed-by: Chad Versace <[email protected]>
*	anv: Enable fast depth clears	Nanley Chery	2016-10-07	2	-2/+35
\| \| \| \| \| \| \| \| \|	Provides an FPS increase of ~30% on the Sascha triangle and multisampling demos. Signed-off-by: Nanley Chery <[email protected]> Reviewed-by: Jason Ekstrand <[email protected]> Reviewed-by: Chad Versace <[email protected]>
*	anv/cmd_buffer: Enable rendering to HiZ	Chad Versace	2016-10-07	2	-4/+40
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	Nanley Chery: (rebase) - Resolve conflicts with new anv_batch_emit macro (amend) - Handle a QPitch TODO - Emit 3DSTATE_HIER_DEPTH_BUFFER on pre-BDW systems - Only use HiZ for single-subpass renderpasses - Emit the HiZ instruction before the stencil instruction to follow the optimized clear sequence specified in the PRMs - Don't modify clear params - Enable resolves when a HiZ buffer is used to ensure depth buffer validity Provides an FPS increase of ~15% on the Sascha triangle and multisampling demos. Signed-off-by: Nanley Chery <[email protected]> Reviewed-by: Chad Versace <[email protected]> Reviewed-by: Jason Ekstrand <[email protected]>
*	anv/cmd_buffer: Add code for performing HZ operations	Nanley Chery	2016-10-07	3	-0/+197
\| \| \| \| \| \| \| \| \|	Create a function that performs one of three HiZ operations - depth/stencil clears, HiZ resolve, and depth resolves. Signed-off-by: Nanley Chery <[email protected]> Reviewed-by: Jason Ekstrand <[email protected]> Reviewed-by: Chad Versace <[email protected]>
*	anv/image: Memset hiz surfaces to 0 when binding memory	Jason Ekstrand	2016-10-07	1	-1/+30
\| \| \| \| \| \| \| \| \|	Nanley Chery (amend): - Change memset value from 0xff to 0 (a defined value for HiZ). Signed-off-by: Nanley Chery <[email protected]> Reviewed-by: Chad Versace <[email protected]> Reviewed-by: Jason Ekstrand <[email protected]>
*	anv: Move BindImageMemory to anv_image.c	Jason Ekstrand	2016-10-07	2	-20/+20
\| \| \| \| \| \|	Signed-off-by: Nanley Chery <[email protected]> Reviewed-by: Chad Versace <[email protected]> Reviewed-by: Jason Ekstrand <[email protected]>
*	anv: Allocate hiz surface	Chad Versace	2016-10-07	1	-3/+34
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	Nanley Chery: (rebase) - Use isl_surf_get_hiz_surf() (amend) - Only add a HiZ surface onto a depth/stencil attachment - Add comment above HiZ surface addition - Hide HiZ behind INTEL_VK_HIZ prior to BDW - Disable HiZ for untested cases - Remove DISABLE_AUX_BIT instead of preventing it from being added Signed-off-by: Nanley Chery <[email protected]> Reviewed-by: Jason Ekstrand <[email protected]> Reviewed-by: Chad Versace <[email protected]>
*	anv: Add func anv_image_has_hiz()	Chad Versace	2016-10-07	1	-0/+10
\| \| \| \| \|	Signed-off-by: Nanley Chery <[email protected]> Reviewed-by: Jason Ekstrand <[email protected]>
*	anv: Add anv_image::hiz_surface	Chad Versace	2016-10-07	1	-0/+2
\| \| \| \| \| \| \|	Unused. Signed-off-by: Nanley Chery <[email protected]> Reviewed-by: Jason Ekstrand <[email protected]>
*	isl: Correct a comment in the isl_format enum	Nanley Chery	2016-10-07	1	-1/+1
\| \| \| \| \| \| \| \|	HiZ is not a color surface, but an auxiliary depth surface. Signed-off-by: Nanley Chery <[email protected]> Reviewed-by: Chad Versace <[email protected]> Reviewed-by: Jason Ekstrand <[email protected]>