mesa.git - Unnamed repository; edit this file 'description' to name the repository.

	Commit message (Collapse)	Author	Age	Files	Lines
*	radeonsi: add CP DMA flags for greater control over synchronization	Marek Olšák	2017-01-06	3	-16/+31
\| \| \| \| \| \| \|	for L2 prefetch Reviewed-by: Edward O'Callaghan <[email protected]> Reviewed-by: Nicolai Hähnle <[email protected]>
*	radeonsi: cleanly communicate which CP DMA packet is first	Marek Olšák	2017-01-06	1	-11/+21
\| \| \| \| \|	Reviewed-by: Edward O'Callaghan <[email protected]> Reviewed-by: Nicolai Hähnle <[email protected]>
*	gallium/radeon: add new HUD query num-SDMA-IBs	Marek Olšák	2017-01-06	9	-2/+21
\| \| \| \| \|	Reviewed-by: Bas Nieuwenhuizen <[email protected]> Reviewed-by: Nicolai Hähnle <[email protected]>
*	gallium/radeon: rename the num-ctx-flushes query to num-GFX-IBs	Marek Olšák	2017-01-06	9	-14/+14
\| \| \| \| \|	Reviewed-by: Bas Nieuwenhuizen <[email protected]> Reviewed-by: Nicolai Hähnle <[email protected]>
*	radeonsi: add HUD queries for cache flush stats	Marek Olšák	2017-01-06	4	-0/+32
\| \| \| \| \|	Reviewed-by: Bas Nieuwenhuizen <[email protected]> Reviewed-by: Nicolai Hähnle <[email protected]>
*	radeonsi: don't count fast clears and prefetches into CP DMA stats	Marek Olšák	2017-01-06	1	-2/+6
\| \| \| \| \|	Reviewed-by: Bas Nieuwenhuizen <[email protected]> Reviewed-by: Nicolai Hähnle <[email protected]>
*	radeonsi: don't wait for compute shaders in texture_barrier	Marek Olšák	2017-01-06	1	-2/+1
\| \| \| \| \| \| \|	it doesn't interact with compute shaders in any way Reviewed-by: Bas Nieuwenhuizen <[email protected]> Reviewed-by: Nicolai Hähnle <[email protected]>
*	radeonsi: assume that a TES without POSITION precedes GS	Marek Olšák	2017-01-06	1	-1/+2
\| \| \| \| \|	Reviewed-by: Bas Nieuwenhuizen <[email protected]> Reviewed-by: Nicolai Hähnle <[email protected]>
*	radeonsi: unduplicate VS color export code	Marek Olšák	2017-01-06	1	-9/+2
\| \| \| \| \| \| \|	it's exactly the same as the other ones Reviewed-by: Bas Nieuwenhuizen <[email protected]> Reviewed-by: Nicolai Hähnle <[email protected]>
*	radeonsi: clean up more HAVE_LLVM #ifdefs	Marek Olšák	2017-01-06	2	-13/+20
\| \| \| \| \|	Reviewed-by: Bas Nieuwenhuizen <[email protected]> Reviewed-by: Nicolai Hähnle <[email protected]>
*	gallium/radeon: clean up HAVE_LLVM #ifdefs in r600_get_llvm_processor_name	Marek Olšák	2017-01-06	1	-17/+11
\| \| \| \|	Reviewed-by: Nicolai Hähnle <[email protected]>
*	i965: Properly flush in hsw_pause_transform_feedback().	Kenneth Graunke	2017-01-06	1	-0/+3
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	Fixes a number of transform feedback tests when run with Linux 4.8, which allows us to use the MI_LOAD_REGISTER_REG command, at which point we started using this new broken path. ES3-CTS.functional.transform_feedback.array_element.interleaved.lines.* and Piglit's arb_transform_feedback2/draw-auto are both fixed by this patch, for example. Thanks to Chris Wilson for catching this mistake! Bugzilla: https://bugs.freedesktop.org/show_bug.cgi?id=99030 Signed-off-by: Kenneth Graunke <[email protected]> Reviewed-by: Anuj Phogat <[email protected]>
*	i965: Fix texturing in the vec4 TCS and GS backends.	Kenneth Graunke	2017-01-06	1	-2/+12
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	We were failing to zero m0.2 of the sampler message header for TCS and GS messages in the simple case. fs_generator has done this for about a year now, but we missed it in vec4_generator. Fixes ES31-CTS.core.texture_cube_map_array.sampling, GL45-CTS.texture_cube_map_array.sampling, and many dEQP-GLES31.functional.shaders.opaque_type_indexing.sampler subtests: - dynamically_uniform.tessellation_control.isampler3d - dynamically_uniform.tessellation_control.isamplercube - dynamically_uniform.tessellation_control.sampler2d - dynamically_uniform.tessellation_control.usamplercube - dynamically_uniform.tessellation_control.sampler2darray - dynamically_uniform.tessellation_control.isampler2darray - dynamically_uniform.tessellation_control.usampler3d - dynamically_uniform.tessellation_control.usampler2darray - dynamically_uniform.tessellation_control.usampler2d - dynamically_uniform.tessellation_control.sampler3d - dynamically_uniform.tessellation_control.samplercube - dynamically_uniform.tessellation_control.isampler2d - uniform.tessellation_control.isampler3d - uniform.tessellation_control.isamplercube - uniform.tessellation_control.usampler2d - uniform.tessellation_control.usampler3d - uniform.tessellation_control.sampler2darray - uniform.tessellation_control.isampler2darray - uniform.tessellation_control.usampler2darray - uniform.tessellation_control.sampler2d - uniform.tessellation_control.usamplercube - uniform.tessellation_control.sampler3d - uniform.tessellation_control.samplercube - uniform.tessellation_control.isampler2d Cc: [email protected] Signed-off-by: Kenneth Graunke <[email protected]> Reviewed-by: Jason Ekstrand <[email protected]>
*	swr: [rasterizer core] rename OutputMerger functions	Tim Rowley	2017-01-06	2	-9/+9
\| \| \| \|	Reviewed-by: Bruce Cherniak <[email protected]>
*	swr: [rasterizer core] fix SIMD16 Transpose_16_16	Tim Rowley	2017-01-06	1	-2/+2
\| \| \| \| \| \| \|	Fix incorrect swizzling in SIMD16 Transpose_16_16 breaking the two-channel 16-bpc formats like R16G16_FLOAT. Reviewed-by: Bruce Cherniak <[email protected]>
*	swr: [rasterizer core] fix SIMD16 output merger	Tim Rowley	2017-01-06	2	-16/+22
\| \| \| \| \| \|	Honor the colorHottileEnable mask when accessing colorBuffer pointers. Reviewed-by: Bruce Cherniak <[email protected]>
*	swr: [rasterizer core] fix SIMD16 PackTraits pack() and unpack()	Tim Rowley	2017-01-06	3	-48/+82
\| \| \| \| \| \|	Fix routines for 8-bit and 16-bit formats used by optimized tile store. Reviewed-by: Bruce Cherniak <[email protected]>
*	swr: [rasterizer core] fix SIMD16 transpose functions	Tim Rowley	2017-01-06	3	-113/+225
\| \| \| \| \| \| \| \| \| \| \| \| \|	Fixed Transpose_16 methods of following formats: Transpose8_8_8_8 Transpose8_8 Transpose32_32 Transpose16_16_16_16 Transpose16_16_16 Transpose16_16 Reviewed-by: Bruce Cherniak <[email protected]>
*	swr: [rasterizer core] whitespace adjustments	Tim Rowley	2017-01-06	1	-2/+1
\| \| \| \|	Reviewed-by: Bruce Cherniak <[email protected]>
*	i965: Don't set EmitNoMainReturn.	Kenneth Graunke	2017-01-05	1	-1/+0
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	A while ago, we stopped using Luca's GLSL IR lower_jumps pass in favor of nir_lower_returns(). Marek's commit d3cb79e043338b0e55a3fba8df652f3 put it in do_common_optimization, which resulted in us calling it again. Dropping the EmitNoMainReturn setting makes us skip that pass again. Apparently that pass doesn't work properly, because this fixes Piglit's tests/spec/glsl-1.10/execution/vs-nested-return-sibling-loop. Bugzilla: https://bugs.freedesktop.org/show_bug.cgi?id=99287 Signed-off-by: Kenneth Graunke <[email protected]> Reviewed-by: Timothy Arceri <[email protected]>
*	vc4: Rewrite T image handling based on calling the LT handler.	Eric Anholt	2017-01-05	1	-34/+75
\| \| \| \| \| \| \| \| \| \| \| \| \|	The T images are composed of effectively swizzled-around blocks of LT (4x4 utile) images, so we can reduce the t_utile_address() calls by 16x by calling into the simpler LT loop. This also adds support for calling down with non-utile-aligned coordinates, which will be part of lifting the utile alignment requirement on our callers and avoiding the RMW on non-utile-aligned stores. Improves 1024x1024 TexSubImage by 2.55014% +/- 1.18584% (n=46) Improves 1024x1024 GetTexImage by 2.242% +/- 0.880954% (n=32)
*	vc4: Move the utile_width/height functions to header inlines.	Eric Anholt	2017-01-05	2	-37/+36
\| \| \| \| \| \|	I want these inlined in the callers, particularly with the tiling changes coming up, but we're not building with lto so some caller would suffer.
*	vc4: Make the load/store utile functions static.	Eric Anholt	2017-01-05	2	-4/+2
\| \| \| \| \|	They don't have any other callers outside of this file, and I'm hoping they get inlined soon.
*	vc4: Simplify the load/store utile functions.	Eric Anholt	2017-01-05	1	-10/+22
\| \| \| \| \| \| \| \| \|	They now have less of a dependency on the cpp, and don't have to do a divide. Hacking up mesa-demos teximage to do only one subtest and not draw points, I saw 1024x1024 glTexSubImage2D() improve by 4.86939% +/- 1.40408% (n=30) and glGetTexImage() by 2.18978% +/- 0.140268% (n=5).
*	vc4: Reuse a list function to simplify bufmgr code.	Eric Anholt	2017-01-05	1	-11/+2
\|
*	vc4: Flush the job early if we're referencing too many BOs.	Eric Anholt	2017-01-05	3	-0/+16
\| \| \| \| \| \| \| \| \| \| \|	If we get up toward 256MB (or whatever the CMA area size is), VC4_GEM_CREATE will start throwing errors. Even if we don't trigger that, when we flush the kernel's BO allocation for the CLs or bin memory may end up throwing an error, at which point our job won't get rendered at all. Just flush early (half of maximum CMA size) so that hopefully we never get to that point.
*	st/mesa/glsl: move SamplerTargets to gl_program	Timothy Arceri	2017-01-06	6	-13/+16
\| \| \| \| \| \| \| \|	This will help allow us to simplify the handling of samplers by storing them in a single location rather than duplicating them in both gl_linked_shader and gl_program. Reviewed-by: Eric Anholt <[email protected]>
*	st/mesa/glsl: set SamplersUsed directly in gl_program	Timothy Arceri	2017-01-06	5	-6/+3
\| \| \| \|	Reviewed-by: Eric Anholt <[email protected]>
*	mesa/glsl: set sampler units directly in gl_program	Timothy Arceri	2017-01-06	4	-17/+7
\| \| \| \| \| \| \|	Now that we create gl_program earlier there is no need to mess about copying things to gl_linked_shader then to gl_program. Reviewed-by: Eric Anholt <[email protected]>
*	mesa: simplify sampler setting code	Timothy Arceri	2017-01-06	1	-22/+11
\| \| \| \| \| \| \| \|	There is no need to loop over active samplers the code above this would have already exited if the sampler was inactive, or errored if the count was larger than the uniforms array size. Reviewed-by: Eric Anholt <[email protected]>
*	mesa/glsl: set num_textures per stage directly in shader_info	Timothy Arceri	2017-01-06	6	-6/+5
\| \| \| \|	Reviewed-by: Eric Anholt <[email protected]>
*	mesa: make _CurrentFragmentProgram a gl_program struct pointer	Timothy Arceri	2017-01-06	6	-30/+22
\| \| \| \| \| \| \| \|	Making this point to a gl_program struct rather than a gl_shader_program struct will allow use to later also make the CurrentProgram array hold gl_program structs which in turn will allow for code simpilifcation. Reviewed-by: Eric Anholt <[email protected]>
*	i965: stop passing gl_shader_program to the precompile and codegen functions	Timothy Arceri	2017-01-06	12	-87/+31
\| \| \| \| \| \| \| \|	We no longer need it. While we are at it we mark the vs, gs, and wm codegen functions as static. Reviewed-by: Eric Anholt <[email protected]>
*	mesa/glsl: remove hack to reset sampler units to zero	Timothy Arceri	2017-01-06	3	-20/+16
\| \| \| \| \| \| \| \| \| \|	Now that we have the is_arb_asm flag we can just skip the initialisation. V2: remove hack from standalone compiler where it was never needed since it only compiles glsl shaders. Reviewed-by: Eric Anholt <[email protected]>
*	i965: make use of new is_arb_asm flag	Timothy Arceri	2017-01-06	2	-13/+11
\| \| \| \|	Reviewed-by: Eric Anholt <[email protected]>
*	st/mesa/glsl: add new is_arb_asm flag in gl_program	Timothy Arceri	2017-01-06	13	-34/+45
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	Set the flag via the _mesa_init_gl_program() and NewProgram() helpers. In i965 we currently check for the existance of gl_shader_program to decide if this is an ARB assembly style program or not. Adding a flag makes the code clearer and will help removes a dependency on gl_shader_program in the i965 codegen functions. Also this will allow use to skip initialising sampler units for linked shaders, we currently memset it to zero again during linking. Reviewed-by: Eric Anholt <[email protected]>
*	i965: pass gl_program directly to brw_compile_tes()	Timothy Arceri	2017-01-06	3	-6/+4
\| \| \| \| \| \|	This is the only thing we use from gl_shader_program so pass it directly. Reviewed-by: Lionel Landwerlin <[email protected]>
*	i965: stop passing gl_shader_program to brw_nir_setup_glsl_uniforms()	Timothy Arceri	2017-01-06	8	-18/+13
\| \| \| \| \| \| \|	We can now just get the data needed from the gl_shader_program_data pointer in gl_program. Reviewed-by: Lionel Landwerlin <[email protected]>
*	i965: pass gl_program to brw_upload_ubo_surfaces()	Timothy Arceri	2017-01-06	6	-22/+20
\| \| \| \| \| \|	There is no need to pass gl_linked_shader anymore. Reviewed-by: Lionel Landwerlin <[email protected]>
*	i965: stop passing gl_shader_program to ↵	Timothy Arceri	2017-01-06	8	-32/+13
\| \| \| \| \| \| \| \| \|	brw_assign_common_binding_table_offsets() We now get everything we need directly from gl_program so there is no need for this. Reviewed-by: Lionel Landwerlin <[email protected]>
*	st/mesa/glsl/i965: move ShaderStorageBlocks to gl_program	Timothy Arceri	2017-01-06	6	-10/+11
\| \| \| \| \| \| \| \| \| \| \| \|	Having it here rather than in gl_linked_shader allows us to simplify the code. Also it is error prone to depend on the gl_linked_shader for programs in current use because a failed linking attempt will free infomation about the current program. In i965 we could be trying to recompile a shader variant but may have lost some required fields. Reviewed-by: Lionel Landwerlin <[email protected]>
*	st/mesa/glsl/i965: set num_ssbos directly in shader_info	Timothy Arceri	2017-01-06	9	-23/+24
\| \| \| \| \| \| \|	Here we also remove the duplicate field in gl_linked_shader and always get the value from shader_info instead. Reviewed-by: Lionel Landwerlin <[email protected]>
*	st/mesa/glsl/i965: move per stage UniformBlocks to gl_program	Timothy Arceri	2017-01-06	7	-21/+19
\| \| \| \| \| \| \|	This will help allow us to store pointers to gl_program structs in the CurrentProgram array resulting in a bunch of code simplifications. Reviewed-by: Lionel Landwerlin <[email protected]>
*	st/mesa/glsl/i965: set num_ubos directly in shader_info	Timothy Arceri	2017-01-06	9	-21/+17
\| \| \| \| \| \| \|	This also removes the duplicate field in gl_linked_shader, and gets num_ubos from shader_info instead. Reviewed-by: Lionel Landwerlin <[email protected]>
*	st/mesa/glsl/i965: move ImageUnits and ImageAccess fields to gl_program	Timothy Arceri	2017-01-06	12	-67/+50
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	Having it here rather than in gl_linked_shader allows us to simplify the code. Also it is error prone to depend on the gl_linked_shader for programs in current use because a failed linking attempt will free infomation about the current program. In i965 we could be trying to recompile a shader variant but may have lost some required fields. We drop the memset on ImageUnits because gl_program is already created using rzalloc(). Reviewed-by: Lionel Landwerlin <[email protected]>
*	i965: get InfoLog and LinkStatus via the pointer in gl_program	Timothy Arceri	2017-01-06	1	-4/+4
\| \| \| \|	Reviewed-by: Lionel Landwerlin <[email protected]>
*	i965: get shared_size from shader_info rather than gl_shader_program	Timothy Arceri	2017-01-06	1	-2/+2
\| \| \| \|	Reviewed-by: Lionel Landwerlin <[email protected]>
*	i965: stop depending on gl_shader_program for brw_compute_vue_map() params	Timothy Arceri	2017-01-06	1	-1/+1
\| \| \| \| \| \| \| \|	This removes another dependency on gl_shader_program from the codegen functions, this will help allow us to use gl_program for the CurrentProgram array rather than gl_shader_program. Reviewed-by: Lionel Landwerlin <[email protected]>
*	i965: pass gl_program to the brw_*_debug_recompile() functions	Timothy Arceri	2017-01-06	7	-138/+125
\| \| \| \| \| \| \| \| \| \| \| \| \| \|	Rather then passing gl_shader_program. The only field use was Name which is the same as the Id field in gl_program. For wm and vs we also make the functions static and move them before the codegen functions. This change reduces the codegen functions dependency on gl_shader_program. Reviewed-by: Lionel Landwerlin <[email protected]>
*	gallivm: (trivial) fix typo bug with small AoS format unpacking	Roland Scheidegger	2017-01-06	1	-1/+1
\| \| \| \| \| \| \|	Fix typo using wrong (uninitialized) build context introduced by 4634cb5921b985f04f2daf00cda2d28036143bd3. (This only affects very rare small packed formats which have a PIPE_SWIZZLE_0 channel, such as r4a4, which is never used by mesa/st. Nevertheless it broke lp_test_format.)