mesa.git - Unnamed repository; edit this file 'description' to name the repository.

	Commit message (Collapse)	Author	Age	Files	Lines
*	vc4: Add register allocation support for MUL output rotation.	Eric Anholt	2016-08-25	2	-0/+14
\| \| \| \| \| \| \|	We need the source to be in r0-r3, so make a new register class for it. It will be up to the surrounding passes to make sure that the r0-r3 allocation of its source won't conflict with anything other class requirements on that temp.
*	vc4: Add support for MUL output rotation.	Eric Anholt	2016-08-25	6	-0/+51
\| \| \| \|	Extracted from a patch by jonasarrow on github.
*	vc4: Add support for the 2-bit LOAD_IMM variants.	Eric Anholt	2016-08-25	6	-0/+58
\| \| \| \| \|	Extracted and fixed up from a patch by jonasarrow on github. This ended up not getting used for ddx/ddy, but seems like it might still be useful.
*	vc4: Add QPU scheduling to handle MUL rotate sources.	Eric Anholt	2016-08-25	1	-0/+13
\| \| \| \|	We need MUL rotates to do ddx/ddy support.
*	vc4: Add disassembly for constant MUL rotates	Eric Anholt	2016-08-25	1	-9/+11
\|
*	vc4: Add real validation for MUL rotation.	Eric Anholt	2016-08-25	2	-10/+43
\| \| \| \|	Caught problems in the upcoming DDX/DDY implementation.
*	vc4: Add a QIR value for the QPU element register.	Eric Anholt	2016-08-25	4	-0/+8
\| \| \| \| \|	This will be used in the ddx/ddy support for "Am I the top half?" or "Am I the left half?" checks.
*	radeonsi: don't use allocas for arrays with LLVM 3.8	Marek Olšák	2016-08-25	1	-1/+3
\| \| \| \| \| \|	It crashes. Bugzilla: https://bugs.freedesktop.org/show_bug.cgi?id=97413
*	gallium/radeon: unify and simplify checking for an empty gfx IB	Marek Olšák	2016-08-25	3	-27/+23
\| \| \| \| \| \| \|	We can take advantage of the fact that multi_fence does the obvious thing with NULL fences. This fixes unflushed fences that can get stuck due to empty IBs.
*	radeonsi: disable SDMA texture copying on Carrizo	Marek Olšák	2016-08-25	1	-0/+6
\| \| \| \| \|	Cc: 12.0 <[email protected]> Reviewed-by: Michel Dänzer <[email protected]>
*	gallium/noop: use 3-space indentation	Marek Olšák	2016-08-25	2	-292/+292
\| \| \| \|	Reviewed-by: Brian Paul <[email protected]>
*	gallium: add a pipe_context parameter to resource_get_handle	Marek Olšák	2016-08-25	10	-3/+18
\| \| \| \| \| \| \| \|	radeonsi needs to do some operations (DCC decompression) for OpenGL-OpenCL interop and this is the only way to make it coherent with the current context. It can optionally be set to NULL. Reviewed-by: Brian Paul <[email protected]>
*	nvc0: invalidate textures/samplers on GK104+	Samuel Pitoiset	2016-08-24	2	-12/+22
\| \| \| \| \| \| \| \| \| \| \| \| \| \|	Like Fermi, textures and samplers are aliased between 3D and compute, especially the TIC_FLUSH/TSC_FLUSH methods and we have to re-validate these resources when switching between the two pipelines. This fixes a GPU hang with Elemental (and most likely with other UE4 demos). Tested on GK107 and GM107. Signed-off-by: Samuel Pitoiset <[email protected]> Reviewed-by: Ilia Mirkin <[email protected]> CC: <[email protected]>
*	vc4: Fix GPU hangs with >16 varying values.	Eric Anholt	2016-08-24	2	-19/+68
\| \| \| \|	Fixes glsl-routing in piglit and hangs in glbenchmark 2.0.2.
*	radeonsi: fix VM faults due NULL internal const buffers on CIK	Marek Olšák	2016-08-24	1	-2/+11
\| \| \| \| \| \| \| \|	They are harmless, but the interrupts do decrease performance. Bugzilla: https://bugs.freedesktop.org/show_bug.cgi?id=97039 Cc: 12.0 <[email protected]>
*	gallium: add a cap to expose whether driver supports mixed color/zs bits	Ilia Mirkin	2016-08-23	15	-0/+15
\| \| \| \| \| \| \| \| \| \|	Some hardware can't render to color/depth buffers of mixed bitness. When that happens a fallback has to happen, but this allows the driver to express that this isn't an optimal scenario. The purpose of this is to remove such fbconfigs from the GLX/EGL config list. Signed-off-by: Ilia Mirkin <[email protected]> Reviewed-by: Marek Olšák <[email protected]>
*	nv50/ir: make sure cfg iterator always hits all blocks	Ilia Mirkin	2016-08-23	1	-4/+4
\| \| \| \| \| \| \| \| \| \| \| \|	In some very specially-crafted cases, we could attempt to visit a node that has already been visited, and then run out of bb's to visit, while there were still cross blocks on the list. Make sure that those get moved over in that case. Bugzilla: https://bugs.freedesktop.org/show_bug.cgi?id=96274 Signed-off-by: Ilia Mirkin <[email protected]> Reviewed-by: Samuel Pitoiset <[email protected]> Cc: [email protected]
*	vc4: Tell state_tracker that we would prefer NIR.	Eric Anholt	2016-08-22	3	-8/+31
\| \| \| \| \| \| \| \| \| \|	Before this series, the code generation path was: GLSL IR -> TGSI -> NIR -> NIR clone -> QIR -> QPU Now it's (generally) GLSL IR -> NIR -> NIR clone -> QIR -> QPU
*	vc4: Use proper type sizes for uniforms.	Eric Anholt	2016-08-22	1	-4/+5
\|
*	vc4: Add VARYING_SLOT_PNTC support.	Eric Anholt	2016-08-22	1	-4/+5
\| \| \| \|	We end up with this when doing GLSL-to-NIR.
*	vc4: Fix vc4_nir_lower_io for non-vec4 I/O.	Eric Anholt	2016-08-22	1	-22/+12
\| \| \| \| \|	To support GLSL-to-NIR, we need to be able to support actual float/vec2/vec3 varyings.
*	nir: Define system values for vc4's blending-lowering arguments.	Eric Anholt	2016-08-22	4	-46/+54
\| \| \| \| \| \| \| \| \| \| \| \| \|	In the GLSL-to-NIR conversion of VC4, I had a bit of trouble with what I was calling the "state uniforms" that I was putting into the NIR fighting with its other lowering passes. Instead of using magic uniform base numbers in the backend, follow the lead of load_user_clip_plane and just define system values for them. v2: Fix unintended change to channel_num, drop unspecified const_index value on blend_const_color_r_float. Reviewed-by: Kenneth Graunke <[email protected]>
*	llvmpipe: fix issues with depth clamp	Roland Scheidegger	2016-08-20	3	-49/+94
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	We only did depth clamp when the value was written from the fs. This is very wrong both for d3d10 and GL, and only passed the corresponding piglit test due to pure luck (it no longer does with the enhanced test). Also, interpolation clamped values to 1.0 always, which can legitimately happen if depth clip is disabled, so fix that as well (untested). There is one unresolved issue left, d3d10 always does depth clamping, whereas GL does not (but does [0,1] clamp instead for fs depth outputs) - this information isn't in any gallium state object, leave it as-is for now (though it looks like llvmpipe misses the [0,1] clamp as well). This (with the previous patch) fixes piglit depth-clamp-range test. Reviewed-by: Jose Fonseca <[email protected]>
*	llvmpipe: fix depth clamping wrt reversed near/far values	Roland Scheidegger	2016-08-20	1	-9/+3
\| \| \| \| \| \| \| \| \| \| \|	This wasn't handled before (the result was that no matter what value got clamped, it always ended up as the near value in this case) (if clamping actually happened). Fix this by using the util helper for that (the math is otherwise "mostly" the same, mostly because there could actually be differences due to float rounding, but I don't even know which one would be more correct). Reviewed-by: Jose Fonseca <[email protected]>
*	a4xx: make sure to actually clamp depth as requested	Ilia Mirkin	2016-08-19	2	-2/+29
\| \| \| \| \| \| \| \| \| \| \|	We were previously ... not clamping. I guess this meant that everything got clamped to 1/0, which was enough to pass the existing tests. Or perhaps the clamping would only happen to the rasterized depth value and not the frag shader's output depth value. Bugzilla: https://bugs.freedesktop.org/show_bug.cgi?id=97231 Signed-off-by: Ilia Mirkin <[email protected]> Cc: [email protected]
*	a4xx: only disable depth clipping, not all clipping, when requested	Ilia Mirkin	2016-08-19	2	-1/+4
\| \| \| \| \| \| \| \| \| \|	The previous bit disables the whole clipper, including the regular viewport-related clipping that would go on. The two new bits disable near and far clipping (separately, as verified with the depth-clamp-range piglit). Signed-off-by: Ilia Mirkin <[email protected]> Cc: [email protected]
*	vc4: Switch store_output to using nir_lower_io_to_scalar / component.	Eric Anholt	2016-08-19	2	-44/+16
\|
*	vc4: Use the intrinsic's first_component for vattr VPM index.	Eric Anholt	2016-08-19	2	-7/+3
\| \| \| \|	Avoids another multiplication by 4 of the base in the NIR.
*	vc4: Convert to using nir_lower_io_scalar for FS inputs.	Eric Anholt	2016-08-19	2	-44/+62
\| \| \| \| \|	The scalarizing of FS inputs can be done in a non-driver-dependent manner, so extract it out of the driver.
*	vc4: Switch to using the intrinsic accessors.	Eric Anholt	2016-08-19	3	-23/+29
\| \| \| \| \|	The const_index[] values have always felt magic, and this documents them a bit better.
*	ttn: Use nir_load_front_face instead of the TGSI-style input.	Eric Anholt	2016-08-19	2	-60/+1
\| \| \| \| \| \| \|	This reduces the diff between GLSL-to-NIR and TGSI-to-NIR, and gives NIR more optimization to work on. Reviewed-by: Kenneth Graunke <[email protected]>
*	ttn: Make FRAG_RESULT_DEPTH be a float variable to match gtn and ptn.	Eric Anholt	2016-08-19	3	-8/+1
\| \| \| \| \| \| \|	This lets TTN-using drivers handle FRAG_RESULT_DEPTH the same between all their source paths. Reviewed-by: Rob Clark <[email protected]>
*	vc4: Dump the TGSI before trying to convert it to NIR.	Eric Anholt	2016-08-19	1	-4/+3
\| \| \| \|	In the case of debugging a crash in TTN, this is nice to have.
*	radeon/vce: set flag based on dual instance enablement	Boyuan Zhang	2016-08-19	1	-2/+4
\| \| \| \| \| \| \|	Set the flag on when dual instance encoding is supported, otherwise set it to off. Signed-off-by: Boyuan Zhang <[email protected]>
*	radeonsi: initialize and finalize the LLVM function pass manager	Marek Olšák	2016-08-18	1	-0/+2
\| \| \| \|	Reviewed-by: Tom Stellard <[email protected]>
*	swr: [rasterizer core] only use Viewport/Scissors during SwrDraw* operations	Tim Rowley	2016-08-17	12	-415/+400
\| \| \| \| \| \| \| \| \| \| \|	Add explicit rects for: - SwrClearRenderTarget - SwrDiscardRect - SwrInvalidateTiles - SwrStoreTiles Signed-off-by: Tim Rowley <[email protected]>
*	swr: [rasterizer common] reorder SWR_FORMAT_INFO	Tim Rowley	2016-08-17	2	-825/+1433
\| \| \| \|	Signed-off-by: Tim Rowley <[email protected]>
*	swr: [rasterizer core] make dirtytile list point directly to macrotilequeues	Tim Rowley	2016-08-17	3	-14/+15
\| \| \| \| \| \|	Speeds up high geometry HPC workloads. Signed-off-by: Tim Rowley <[email protected]>
*	swr: [rasterizer core] portability - remove use of INT64	Tim Rowley	2016-08-17	1	-2/+2
\| \| \| \|	Signed-off-by: Tim Rowley <[email protected]>
*	swr: [rasterizer core] viewport transform disabled fix	Tim Rowley	2016-08-17	1	-4/+11
\| \| \| \| \| \| \|	When viewport transform is disabled (ie. screen space coords are passed in directly), the W component should be interpreted as RHW. Signed-off-by: Tim Rowley <[email protected]>
*	swr: [rasterizer core] clamp scissor rects to current tile rect	Tim Rowley	2016-08-17	1	-0/+18
\| \| \| \|	Signed-off-by: Tim Rowley <[email protected]>
*	swr: [rasterizer core] align stats structures	Tim Rowley	2016-08-17	1	-2/+2
\| \| \| \|	Signed-off-by: Tim Rowley <[email protected]>
*	swr: [rasterizer core] use AVX2 permute to simplify PaTriList	Tim Rowley	2016-08-17	1	-1/+35
\| \| \| \|	Signed-off-by: Tim Rowley <[email protected]>
*	swr: [rasterizer core] move some global variables to SWR_CONTEXT	Tim Rowley	2016-08-17	2	-9/+9
\| \| \| \|	Signed-off-by: Tim Rowley <[email protected]>
*	swr: [rasterizer core] change scale on VP matrix element gathers	Tim Rowley	2016-08-17	1	-6/+6
\| \| \| \| \| \| \|	Was 1, which led to pulling denorms for non-zero indices. Changed to sizeof(float). Signed-off-by: Tim Rowley <[email protected]>
*	swr: [rasterizer] implementing native AVX-512 simd16 intrinsics	Tim Rowley	2016-08-17	2	-84/+265
\| \| \| \|	Signed-off-by: Tim Rowley <[email protected]>
*	svga: fix src/dst typo in can_blit_via_copy_region_vgpu10()	Brian Paul	2016-08-17	1	-1/+1
\| \| \| \| \| \| \| \| \| \|	The function was always returning false because of this typo. Retested with piglit. There's some sRGB-related blit failures, but that seems unrelated. Reviewed-by: Charmaine Lee <[email protected]> Reviewed-by: Neha Bhende <[email protected]>
*	svga: initialize a variable to silence a gcc warning	Brian Paul	2016-08-17	1	-1/+1
\| \| \| \|	Reviewed-by: Charmaine Lee <[email protected]>
*	radeonsi: fix up buffer descriptor upper-bound checking	Marek Olšák	2016-08-17	1	-1/+1
\| \| \| \| \| \|	st/mesa does this too, so we're safe. Reviewed-by: Nicolai Hähnle <[email protected]>
*	gallium: change pipe_image_view::first_element/last_element -> offset/size	Marek Olšák	2016-08-17	5	-36/+18
\| \| \| \| \| \| \| \| \|	This is required by OpenGL. Our hardware supports this. Example: Bind RGBA32F with offset = 4 bytes. Acked-by: Ilia Mirkin <[email protected]> Acked-by: Nicolai Hähnle <[email protected]>