mesa.git - Unnamed repository; edit this file 'description' to name the repository.

	Commit message (Collapse)	Author	Age	Files	Lines
*	radeonsi: add TCS epilog	Marek Olšák	2016-02-21	4	-13/+155
\| \| \| \|	Reviewed-by: Nicolai Hähnle <[email protected]>
*	radeonsi: add VS epilog	Marek Olšák	2016-02-21	4	-11/+171
\| \| \| \| \| \| \| \| \|	It only exports the primitive ID. Also used by TES when it's compiled as VS. The VS input location of the primitive ID input is v2. Reviewed-by: Nicolai Hähnle <[email protected]>
*	radeonsi: add VS prolog	Marek Olšák	2016-02-21	4	-1/+267
\| \| \| \| \| \|	This is disabled with use_monolithic_shaders = true. Reviewed-by: Nicolai Hähnle <[email protected]>
*	radeonsi: first bits for non-monolithic shaders	Marek Olšák	2016-02-21	4	-14/+45
\| \| \| \|	Reviewed-by: Nicolai Hähnle <[email protected]>
*	radeonsi: add code for dumping all shader parts together (v2)	Marek Olšák	2016-02-21	1	-12/+34
\| \| \| \| \| \|	v2: unify some code into si_get_shader_binary_size Reviewed-by: Michel Dänzer <[email protected]>
*	radeonsi: add code for combining and uploading shaders from 3 shader parts	Marek Olšák	2016-02-21	2	-8/+36
\| \| \| \|	Reviewed-by: Nicolai Hähnle <[email protected]>
*	radeonsi: fail compilation if non-GS non-CS shaders have rodata	Marek Olšák	2016-02-21	1	-0/+13
\| \| \| \|	Reviewed-by: Nicolai Hähnle <[email protected]>
*	radeonsi: separate 2 pieces of code from create_function	Marek Olšák	2016-02-21	1	-31/+51
\| \| \| \|	Reviewed-by: Nicolai Hähnle <[email protected]>
*	radeonsi: add samplemask parameter to si_export_mrt_color	Marek Olšák	2016-02-21	1	-3/+7
\| \| \| \|	Reviewed-by: Nicolai Hähnle <[email protected]>
*	radeonsi: add start_instance parameter to get_instance_index_for_fetch	Marek Olšák	2016-02-21	1	-4/+6
\| \| \| \|	Reviewed-by: Nicolai Hähnle <[email protected]>
*	radeonsi: separate out shader key bits for prologs & epilogs	Marek Olšák	2016-02-21	4	-100/+140
\| \| \| \|	Reviewed-by: Nicolai Hähnle <[email protected]>
*	radeonsi: compute how many input VGPRs fragment shaders have	Marek Olšák	2016-02-21	2	-0/+43
\| \| \| \|	Reviewed-by: Nicolai Hähnle <[email protected]>
*	radeonsi: compute how many input SGPRs and VGPRs shaders have	Marek Olšák	2016-02-21	2	-0/+34
\| \| \| \| \| \| \|	Prologs (shader binaries inserted before the API shader binary) need to know this, so that they won't change the input registers unintentionally. Reviewed-by: Nicolai Hähnle <[email protected]>
*	gallium/radeon: add basic code for setting shader return values	Marek Olšák	2016-02-21	4	-8/+21
\| \| \| \| \| \|	LLVMBuildInsertValue will be used on return_value. Reviewed-by: Nicolai Hähnle <[email protected]>
*	nvc0: enable compute shaders on Fermi	Samuel Pitoiset	2016-02-21	1	-1/+3
\| \| \| \| \| \| \| \|	Kepler compute support is really different than Fermi and it's not ready yet. Signed-off-by: Samuel Pitoiset <[email protected]> Reviewed-by: Ilia Mirkin <[email protected]>
*	nv50/ir: add atomics support on shared memory for Fermi	Samuel Pitoiset	2016-02-21	2	-2/+102
\| \| \| \| \| \| \| \| \| \| \| \| \| \|	Changes from v3: - move the previous OP_SELP change to the previous commit Changes from v2: - make sure the op is OP_SELP when emitting the predicate and add one assert - use bld.getSSA() for mkOp2() - add cross edge between tryLockAndSetBB and joinBB Signed-off-by: Samuel Pitoiset <[email protected]> Acked-by: Ilia Mirkin <[email protected]>
*	nv50/ir: make OP_SELP a compare instruction	Samuel Pitoiset	2016-02-21	3	-10/+19
\| \| \| \| \| \| \| \| \| \|	This OP_SELP insn will be used to handle compare and swap subops. Changes from v2: - fix logic for GK110+ Signed-off-by: Samuel Pitoiset <[email protected]> Reviewed-by: Ilia Mirkin <[email protected]>
*	nv50/ir: add lock/unlock subops for load/store	Samuel Pitoiset	2016-02-21	3	-2/+26
\| \| \| \| \|	Signed-off-by: Samuel Pitoiset <[email protected]> Reviewed-by: Ilia Mirkin <[email protected]>
*	nv50/ir: use s[] addr space for shared buffers	Samuel Pitoiset	2016-02-21	1	-11/+30
\| \| \| \| \| \| \| \| \| \| \| \|	Shared memory address space (FILE_MEMORY_SHARED) must be used instead of global memory when a shared memory area is declared. Changes from v2: - oops, do not remove TGSI_FILE_BUFFER in a switch in nv50_ir_from_tgsi.cpp Signed-off-by: Samuel Pitoiset <[email protected]> Reviewed-by: Ilia Mirkin <[email protected]>
*	nvc0: reduce likelihood of collision for real buffers on Fermi	Samuel Pitoiset	2016-02-21	1	-2/+2
\| \| \| \| \| \| \| \| \| \| \|	Reduce likelihood of collision with real buffers by placing the hole at the top of the 4G area. This fixes some indirect draw+compute tests with large buffers. Suggested by Ilia Mirkin. Signed-off-by: Samuel Pitoiset <[email protected]> Reviewed-by: Ilia Mirkin <[email protected]>
*	nvc0: invalidate compute state when switching pipe contexts	Samuel Pitoiset	2016-02-21	1	-1/+2
\| \| \| \| \|	Signed-off-by: Samuel Pitoiset <[email protected]> Reviewed-by: Ilia Mirkin <[email protected]>
*	nvc0: add support for indirect compute on Fermi	Samuel Pitoiset	2016-02-21	6	-20/+81
\| \| \| \| \| \| \| \| \| \| \| \|	When indirect compute is used, the size of the grid (in blocks) is stored as three integers inside a buffer. This requires a macro to set up GRIDDIM_YX and GRIDDIM_Z. Changes from v2: - do not launch the grid if the number of groups for a dimension is 0 Signed-off-by: Samuel Pitoiset <[email protected]> Reviewed-by: Ilia Mirkin <[email protected]>
*	nvc0: bind textures/samplers for compute on Fermi	Samuel Pitoiset	2016-02-21	3	-8/+57
\| \| \| \| \| \| \| \| \| \| \|	Textures and samplers don't seem to be aliased between COMPUTE and 3D. Changes from v2: - refactor the code to share (almost) the same logic between 3d and compute Signed-off-by: Samuel Pitoiset <[email protected]> Reviewed-by: Ilia Mirkin <[email protected]>
*	nvc0: bind shader buffers for compute on Fermi	Samuel Pitoiset	2016-02-21	5	-7/+56
\| \| \| \| \| \| \| \|	This is loosely based on 3D. Shader buffers are bound on c15 (the driver constbuf) at offset 0x200. Signed-off-by: Samuel Pitoiset <[email protected]> Reviewed-by: Ilia Mirkin <[email protected]>
*	nvc0: bind driver constbuf for compute on Fermi	Samuel Pitoiset	2016-02-21	5	-0/+29
\| \| \| \| \| \| \| \| \| \| \|	Changes from v3: - add new validation state for COMPUTE driver constbuf Changes from v2: - always bind the driver consts even if user params come in via clover Signed-off-by: Samuel Pitoiset <[email protected]> Reviewed-by: Ilia Mirkin <[email protected]>
*	nvc0: add a new validation state for 3D driver constbuf	Samuel Pitoiset	2016-02-21	2	-0/+19
\| \| \| \| \| \| \| \| \|	This will be used to invalidate 3D driver constbuf when using COMPUTE and vice-versa. This is needed because this CB contains a bunch of useful information like the addrs of shader buffers. Signed-off-by: Samuel Pitoiset <[email protected]> Reviewed-by: Ilia Mirkin <[email protected]>
*	nvc0: bind constant buffers for compute on Fermi	Samuel Pitoiset	2016-02-21	5	-13/+81
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	Loosely based on 3D. Changs from v3: - invalidate COMPUTE CBs after validating 3D CBs because they are aliased Changes from v2: - get rid of the 's' param to nvc0_cb_bo_push() because it doesn't matter to upload constbufs for compute using the 3d chan Signed-off-by: Samuel Pitoiset <[email protected]> Reviewed-by: Ilia Mirkin <[email protected]>
*	nvc0: allocate an area for compute user constbufs	Samuel Pitoiset	2016-02-21	4	-18/+20
\| \| \| \| \| \| \|	For compute shaders, we might need to upload uniforms. Signed-off-by: Samuel Pitoiset <[email protected]> Reviewed-by: Ilia Mirkin <[email protected]>
*	nv50: do not advertise about compute shaders	Samuel Pitoiset	2016-02-20	1	-1/+1
\| \| \| \| \| \| \| \| \|	Compute shaders are totally unsupported. This avoids Clover to report that OpenCL is supported on Tesla because it's a lie. Signed-off-by: Samuel Pitoiset <[email protected]> Reviewed-by: Pierre Moreau <[email protected]> Reviewed-by: Ilia Mirkin <[email protected]>
*	st/mesa: force depth mode to GL_RED for sized depth/stencil formats	Ilia Mirkin	2016-02-19	1	-9/+25
\| \| \| \| \| \| \| \| \| \| \|	See commit 9db2098d for the i965 version of this. This fixes depth in a bunch of dEQP EXT_texture_border_clamp tests. And probably other ones as well. Signed-off-by: Ilia Mirkin <[email protected]> Reviewed-by: Kenneth Graunke <[email protected]> Cc: [email protected]
*	egl_dri2: set correct error code if swapbuffers fails	Daniel Czarnowski	2016-02-19	1	-1/+6
\| \| \| \| \| \| \| \| \| \| \|	A return value of '-1' means that there was error during swap with a window drawable, in this case we set error as EGL_BAD_NATIVE_WINDOW. v2: coding style cleanup, better commit message Signed-off-by: Matt Roper <[email protected]> Cc: "11.0 11.1" <[email protected] Reviewed-by: Emil Velikov <[email protected]>
*	egl: move Null check to eglGetSyncAttribKHR to prevent Segfault	Dongwon Kim	2016-02-19	2	-5/+8
\| \| \| \| \| \| \| \| \| \| \| \| \| \|	Null-check on "*value" is currently done in _eglGetSyncAttrib, which is after eglGetSyncAttribKHR dereferences it. Move the check a layer up (in the beginning of eglGetSyncAttribKHR) to avoid segfaults. Cc: "11.0 11.1" <[email protected] Signed-off-by: Dongwon Kim <[email protected]> Reviewed-by: Marek Olšák <[email protected]> [Emil Velikov: tweak commit message, add stable tag] Reviewed-by: Emil Velikov <[email protected]>
*	meta/copy_image: use precomputed dst_internal_format to avoid segfault	Ilia Mirkin	2016-02-19	1	-1/+1
\| \| \| \| \| \| \| \| \|	If the destination is a renderbuffer, dst_tex_image will be NULL. This fixes the *to_renderbuffer dEQP copy image tests. Signed-off-by: Ilia Mirkin <[email protected]> Reviewed-by: Jason Ekstrand <[email protected]> Cc: [email protected]
*	mesa: add GL_OES_texture_stencil8 support	Ilia Mirkin	2016-02-19	3	-0/+11
\| \| \| \| \| \| \| \| \|	It's basically the same thing as GL_ARB_texture_stencil8 except that glCopyTexImage isn't supported, so add STENCIL_INDEX to the list of invalid GLES formats for glCopyTexImage. Signed-off-by: Ilia Mirkin <[email protected]> Reviewed-by: Eduardo Lima Mitev <[email protected]>
*	st/mesa: fix pbo uploads	Ilia Mirkin	2016-02-19	1	-10/+18
\| \| \| \| \| \| \| \| \| \| \|	- LOD must be provided in .w for TXF (even for buffer textures) - User buffer must be valid at draw time - Must have a sampler associated with the sampler view This makes PBO uploads work again on nouveau. Signed-off-by: Ilia Mirkin <[email protected]> Reviewed-by: Nicolai Hähnle <[email protected]>
*	mesa: check fbo completeness based on internal format, not driver format	Ilia Mirkin	2016-02-19	1	-3/+2
\| \| \| \| \| \| \| \| \| \| \| \|	The base format is a function of the user-requested format, while the driver format is not. So we should use the base format instead. The driver format can be anything. Specifically in the stencil-only case, it might be a depth/stencil format. However we still want to refuse such an attachment when bound to GL_DEPTH. Signed-off-by: Ilia Mirkin <[email protected]> Reviewed-by: Brian Paul <[email protected]>
*	mesa: small optimization of _mesa_expand_bitmap()	Brian Paul	2016-02-19	1	-7/+4
\| \| \| \| \| \|	Avoid a per-pixel multiply. Reviewed-by: Kenneth Graunke <[email protected]>
*	mesa: add special case ubyte[4] / BGRA conversion function	Brian Paul	2016-02-19	1	-5/+69
\| \| \| \| \| \| \| \|	This reduces a glTexImage(GL_RGBA, GL_UNSIGNED_BYTE) hot spot in when storing the texture as BGRA. Reviewed-by: Jose Fonseca <[email protected]> Reviewed-by: Roland Scheidegger <[email protected]>
*	st/mesa: implement a simple cache for glDrawPixels	Brian Paul	2016-02-19	3	-0/+97
\| \| \| \| \| \| \| \| \|	Instead of discarding the texture we created, keep it around in case the next glDrawPixels draws the same image again. This is intended to help application which draw the same image several times in a row, either within a frame or subsequent frames. Reviewed-by: Charmaine Lee <[email protected]>
*	llvmpipe: add a few const qualifiers	Brian Paul	2016-02-19	3	-4/+4
\| \| \| \|	Reviewed-by: Roland Scheidegger <[email protected]>
*	trace: assorted whitespace and formatting fixes	Brian Paul	2016-02-19	1	-29/+31
\| \| \| \|	Reviewed-by: Eduardo Lima Mitev <[email protected]>
*	trace: remove unneeded inline qualifiers	Brian Paul	2016-02-19	1	-46/+46
\| \| \| \|	Reviewed-by: Eduardo Lima Mitev <[email protected]>
*	glsl: fix emit_inline_matrix_constructor for doubles	Iago Toral Quiroga	2016-02-19	1	-6/+13
\| \| \| \| \| \| \|	Specifically, for the case where we initialize a dmat with a source matrix that has fewer columns/rows. Reviewed-by: Kenneth Graunke <[email protected]>
*	glsl: Mark float constants as such	Iago Toral Quiroga	2016-02-19	1	-5/+5
\| \| \| \| \| \|	So we don't generate double to float conversion code Reviewed-by: Kenneth Graunke <[email protected]>
*	glsl: fix indentation in emit_inline_matrix_constructor	Iago Toral Quiroga	2016-02-19	1	-75/+75
\| \| \| \|	Reviewed-by: Kenneth Graunke <[email protected]>
*	glsl: fix standalone compiler	Rob Clark	2016-02-19	1	-0/+12
\| \| \| \| \| \| \| \| \|	Need to set some non-zero limits for MaxCombinedUniformComponents, otherwise we hit an "Too many <type> shader uniform components" error in the linker. Signed-off-by: Rob Clark <[email protected]> Reviewed-by: Samuel Iglesias Gonsálvez <[email protected]>
*	st/mesa: disable depth/stencil/alpha tests in PBO upload	Nicolai Hähnle	2016-02-18	1	-0/+8
\| \| \| \| \| \|	Noticed by Brian Paul. Reviewed-by: Marek Olšák <[email protected]>
*	svga: allow non-contiguous VS input declarations	Brian Paul	2016-02-18	1	-2/+2
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	This fixes a glDrawPixels regression since b63fe0552b5f. The new quad-drawing utility code uses 3 vertex attributes (xyz, rgba, st). For glDrawPixels path we don't use the rgba attribute so there's a gap in the TGSI VS input declarations (INPUT[0] = pos, INPUT[2] = texcoord). The TGSI->VGPU10 translations code did not handle this correctly. I missed this because my VM was configured for HWv11 while testing. Another way to fix this would be to change the tgsi_scan.c code so that the tgsi_shader_info::num_inputs (and num_outputs) included the unused inputs/outputs. These counts would then actually be "max input register index + 1" rather than "number of used inputs". But that change could impact all drivers so put it off for now. No regressions found with piglit or typical GL apps. v2: also update alloc_system_value_index() to use info.file_max[] Reviewed-by: Charmaine Lee <[email protected]>
*	gallivm: Check whether to stop disassemble only for x86	Oded Gabbay	2016-02-19	1	-0/+2
\| \| \| \| \| \| \| \| \| \|	Because the if statement that checks whether we have a return statement is valid only on x86, surround it with X86 or X86-64 arch defines Signed-off-by: Oded Gabbay <[email protected]> Reviewed-by: Jose Fonseca <[email protected]> Reviewed-by: Roland Scheidegger <[email protected]>
*	gallivm: use sstream for dissasembling	Oded Gabbay	2016-02-19	1	-21/+30
\| \| \| \| \| \| \| \| \| \| \| \| \|	Currently, disassemble() directly prints to stdout. This has broke the profiling support for llvmpipe JIT code. This patch redirects the output to an sstream object, which is then either gets printed to stdout (for assembly debugging) or gets written to a file in /tmp/ (for profiling support). Signed-off-by: Oded Gabbay <[email protected]> Reviewed-by: Jose Fonseca <[email protected]> Reviewed-by: Roland Scheidegger <[email protected]>