mesa.git - Unnamed repository; edit this file 'description' to name the repository.

	Commit message (Collapse)	Author	Age	Files	Lines
*	freedreno: resync generated headers	Rob Clark	2014-02-01	4	-9/+39
\| \| \| \|	Signed-off-by: Rob Clark <[email protected]>
*	freedreno/a3xx: fix const confusion	Rob Clark	2014-02-01	2	-9/+9
\| \| \| \| \| \| \| \| \| \| \| \|	Gallium can leave const buffers bound above what is used by the current shader. Which can have a couple bad effects: 1) write beyond const space assigned, which can trigger HLSQ lockup 2) double emit of immed consts, first with bound const buffer vals followed by with actual immed vals. This seems to be a sort of undefined condition. Signed-off-by: Rob Clark <[email protected]>
*	freedreno/a3xx/compiler: compiler cleanups	Rob Clark	2014-02-01	7	-145/+198
\| \| \| \| \| \|	Drop color/pos/psize_regid, plus a few compiler and IR cleanups. Signed-off-by: Rob Clark <[email protected]>
*	freedreno/compiler/a3xx: remove lowered instructions	Rob Clark	2014-02-01	1	-354/+0
\| \| \| \|	Signed-off-by: Rob Clark <[email protected]>
*	freedreno: add tgsi lowering pass	Rob Clark	2014-02-01	4	-2/+1229
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	Currently lowers the following instructions: DST, XPD, SCS, LRP, FRC, POW, LIT, EXP, LOG, DP4, DP3, DPH, DP2 translating these into equivalent simpler TGSI instructions. This probably should be moved to util so other drivers can use it, but just adding under freedreno for now so that I can clear out a lot of the lowering code in a3xx compiler before beginning to add new compiler. Signed-off-by: Rob Clark <[email protected]>
*	freedreno/a3xx/compiler: add CLAMP	Rob Clark	2014-02-01	1	-7/+24
\| \| \| \|	Signed-off-by: Rob Clark <[email protected]>
*	freedreno/a3xx/compiler: various fixes	Rob Clark	2014-02-01	1	-14/+34
\| \| \| \|	Signed-off-by: Rob Clark <[email protected]>
*	freedreno: ctx should hold ref to dev	Rob Clark	2014-02-01	6	-2/+8
\| \| \| \| \| \| \| \|	The ctx should hold ref to dev to avoid problems if screen is destroyed before ctx. Doesn't really fix the egl/glx issues, but at least it prevents things from getting much worse. Signed-off-by: Rob Clark <[email protected]>
*	freedreno: add prims-emitted driver query	Rob Clark	2014-02-01	1	-0/+1
\| \| \| \|	Signed-off-by: Rob Clark <[email protected]>
*	freedreno: Set PIPE_CAP_MIN_MAP_BUFFER_ALIGNMENT to 64	Ian Romanick	2014-01-29	1	-0/+3
\| \| \| \| \| \| \| \|	Allocations actually have page alignment, but 64 is still a reasonable value. Signed-off-by: Ian Romanick <[email protected]> Reviewed-by: Rob Clark <[email protected]>
*	gallium: remove PIPE_CAP_SCALED_RESOLVE	Marek Olšák	2014-01-23	1	-1/+0
\| \| \| \| \| \| \|	If any driver doesn't support this, it can use a blit after resolving the samples. Reviewed-by: Brian Paul <[email protected]>
*	freedreno: add basic query support	Rob Clark	2014-01-08	8	-1/+275
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	Add for now some simple/basic query support (ie. things not actually requiring the GPU). Might change around a bit when I actually add GPU queries, but for now this enables some useful performance info in the GALLIUM_HUD. For example: GALLIUM_HUD=fps+batches+batches-sysmem+batches-gmem+restores,draw-calls The driver specific specific queries are: + draw-calls + batches - number of batches per second, sum of batches-sysmem plus batches-gmem + batches-gmem - render a set of tiles in GMEM, for each tile (optionally) system mem -> gmem (restore), plus N draws, plus gmem -> system mem (resolve) per second + batches-sysmem - N draws to system memory (GMEM bypass) per second + restores - number of GMEM batches that required restore per second Ideally for GMEM rendering, you want batches-gmem to equal fps. If the app is doing something that triggers multiple passes (ie. requires extra round trip gmem <-> system memory) then the # of batches per second will go up relative to fps. Signed-off-by: Rob Clark <[email protected]>
*	freedreno/a3xx: use cs patch instead of RFI+RMW	Rob Clark	2014-01-08	8	-52/+46
\| \| \| \| \| \| \| \|	Since we now have the cmdstream patch mechanism needed for hw binning, might as well also use it for RB_RENDER_CONTROL updates. This avoids the need to use RMW (and associated WFI) to update RB_RENDER_CONTROL. Signed-off-by: Rob Clark <[email protected]>
*	freedreno/a3xx: support for hw binning pass	Rob Clark	2014-01-08	15	-158/+706
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	The binning pass sorts vertices into which bins/tiles they apply to. The visibility information generated during the binning pass can be used to speed up the rendering pass by filtering out vertices which do not apply to the current tile. See: https://github.com/freedreno/freedreno/wiki/Adreno-tiling#optimized-approach This brings a significant fps boost. A rough assortment of tests (supertuxkart, etracer, tremulous, glmark2 'build' test, etc) seems to yield a ~35-45% fps improvement. For now, to be conservative, the binning pass is not enabled yet by default. To enable it use: FD_MESA_DEBUG=binning So far I haven't found anything that breaks with binning enabled, but I'd like a bit more testing before I enable it as default. Signed-off-by: Rob Clark <[email protected]>
*	freedreno: be more clever about gmem usage	Rob Clark	2014-01-08	2	-9/+18
\| \| \| \| \| \|	Only need to leave room for depth/stencil if it is actually used, etc. Signed-off-by: Rob Clark <[email protected]>
*	freedreno: resync generated headers	Rob Clark	2014-01-08	5	-24/+214
\| \| \| \|	Signed-off-by: Rob Clark <[email protected]>
*	freedreno/a3xx: fix blend state corruption issue	Rob Clark	2013-12-26	8	-33/+66
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	Using RMW on banked context registers is not safe. The value read could be the wrong one. So if there has been a DRAW_IDX launched, the RMW must be preceded by a WAIT_FOR_IDLE to ensure the read part of RMW sees the correct value. To avoid unnecessary WFI's, keep track if there is a need for WFI, and only emit one if needed. Furthermore, keep track if we even need to update the register in the first place. And to cut down on the amount of RMW to avoid excessive WFI's, at the tiling/GMEM level we can always overwrite RB_RENDER_CONTROL, as the state at beginning of draw/clear cmds (which we IB to) is always undefined. In the draw/clear commands, we always still use RMW (with WFI if needed), but only if the register value actually changes. (At points where the current value cannot be known, the saved value is reset to ~0, which includes bits outside of RBRC_DRAW_STATE, so there never is chance for confusion.) Signed-off-by: Rob Clark <[email protected]>
*	freedreno: prepare for hw binning	Rob Clark	2013-12-26	9	-142/+159
\| \| \| \| \| \| \| \| \| \| \| \|	Actually assign VSC_PIPE's properly, which will be needed for tiling. And introduce fd_tile for per-tile state (including the assignment of tile to VSC_PIPE). This gives us the proper pipe setup that we'll need for hw binning pass, and also cleans things up a bit by not having to pass so many parameters around. And will also make it easier to introduce different tiling patterns (since we may no longer render tiles in a simple left-to-right top-to-bottom pattern). Signed-off-by: Rob Clark <[email protected]>
*	freedreno: resync generated headers	Rob Clark	2013-12-26	6	-76/+131
\| \| \| \|	Signed-off-by: Rob Clark <[email protected]>
*	freedreno: dummy-draw workaround for a320	Rob Clark	2013-12-14	2	-1/+17
\| \| \| \| \| \|	Fixes gpu lockups in supertuxkart. Signed-off-by: Rob Clark <[email protected]>
*	gallium/winsys/drm: Prepare for passing prime fds in winsys_handle	Christopher James Halse Rogers	2013-12-10	1	-0/+5
\| \| \| \| \| \|	Signed-off-by: Christopher James Halse Rogers <[email protected]> Reviewed-by: Thomas Hellstrom <[email protected]> Signed-off-by: Maarten Lankhorst <[email protected]>
*	freedreno/a3xx: add adreno 330 support	Rob Clark	2013-12-07	2	-4/+7
\| \| \| \|	Signed-off-by: Rob Clark <[email protected]>
*	freedreno/a3xx/compiler: add ROUND	Rob Clark	2013-12-07	1	-0/+1
\| \| \| \|	Signed-off-by: Rob Clark <[email protected]>
*	gallium: add support for AMD_vertex_shader_layer	Marek Olšák	2013-12-03	1	-0/+1
\|
*	freedreno: Add a few texture formats	Andreas Heider	2013-12-02	1	-0/+3
\|
*	gallium: new shader cap bit for the amount of sampler views	Roland Scheidegger	2013-11-28	1	-0/+1
\| \| \| \| \| \| \| \| \|	Ever since introducing separate sampler and sampler view max this was really missing. Every driver but llvmpipe reports the same number as number of samplers for now, so nothing should break. Reviewed-by: Jose Fonseca <[email protected]>
*	gallium/drivers: compact compiler flags into Automake.inc	Emil Velikov	2013-11-16	1	-6/+4
\| \| \| \| \| \| \| \| \| \|	* minimise flags duplication * distingush between VISIBILITY C and CXX flags * set only required flags - C and/or CXX v2: add LLVM_CFLAGS back to AM_CFLAGS (add missing backslash) Signed-off-by: Emil Velikov <[email protected]>
*	gallium/drivers: enable automake subdir-objects	Emil Velikov	2013-11-16	1	-0/+2
\| \| \| \|	Signed-off-by: Emil Velikov <[email protected]>
*	freedreno: compact a2xx and a3xx makefiles into parent ones	Johannes Obermayr	2013-11-16	6	-66/+36
\| \| \| \| \| \| \| \| \|	Nearly everything within the three Makefile.am's is identical. Let's simplify things a little. v2: Rebase and rewrite the commit message (Emil Velikov) Signed-off-by: Emil Velikov <[email protected]>
*	freedreno/a3xx/texture: min/max lod	Rob Clark	2013-11-01	1	-5/+3
\| \| \| \|	Signed-off-by: Rob Clark <[email protected]>
*	freedreno/a3xx: update envytools headers	Rob Clark	2013-11-01	4	-8/+22
\| \| \| \|	Signed-off-by: Rob Clark <[email protected]>
*	freedreno/a3xx: fix VS out / FS in linking	Rob Clark	2013-11-01	3	-7/+47
\| \| \| \| \| \| \| \|	Actually link VS out / FS in based on semantic info, keeping in mind that position/pointsize can also be an input to the FS. This fixes a few fragment shaders which were using gl_Position. Signed-off-by: Rob Clark <[email protected]>
*	freedreno/a3xx: allow num_samplers != num_textures	Rob Clark	2013-11-01	2	-56/+55
\| \| \| \|	Signed-off-by: Rob Clark <[email protected]>
*	freedreno/a3xx/compiler: highp frag shader	Rob Clark	2013-11-01	4	-12/+14
\| \| \| \| \| \| \| \| \| \| \| \| \| \|	Fixes use of full-precision in fragment shader (ie. don't clobber r0.x since that can be used by future bary instructions for varying fetch). And makes use of full-precision the default in fragment shader (but can be overriden via FD_MESA_DEBUG=fraghalf). Seems like half precision is often not enough for texture coordinates. The blob compiler is clever enough to keep texture coords in full precision registers while using half precision for everything else. But we aren't quite that clever yet, so better to default to full precision. Signed-off-by: Rob Clark <[email protected]>
*	freedreno/a3xx/compiler: relative addressing fixes.	Rob Clark	2013-11-01	1	-28/+48
\| \| \| \| \| \| \|	Handle some relative addressing constraints: cannot handle const or relative in cat5 and src2 of cat3. Signed-off-by: Rob Clark <[email protected]>
*	freedreno: we do actually support sqrt	Rob Clark	2013-11-01	2	-0/+8
\| \| \| \|	Signed-off-by: Rob Clark <[email protected]>
*	freedreno: emulated unsupported primitive types	Rob Clark	2013-10-29	5	-25/+74
\| \| \| \| \| \| \|	Use u_primconvert to convert unsupported primitives into supported primitive plus index buffer. Signed-off-by: Rob Clark <[email protected]>
*	freedreno: update generated headers	Rob Clark	2013-10-29	6	-125/+238
\| \| \| \| \| \|	pull in some fixes to draw-initiator/prim-type. Signed-off-by: Rob Clark <[email protected]>
*	gallium: add PIPE_CAP_MIXED_FRAMEBUFFER_SIZES	Ilia Mirkin	2013-10-26	1	-0/+1
\| \| \| \| \| \| \| \| \|	This CAP will determine whether ARB_framebuffer_object can be enabled. The nv30 driver does not allow mixing swizzled and linear zsbuf/cbuf textures. Signed-off-by: Ilia Mirkin <[email protected]> Signed-off-by: Marek Olšák <[email protected]>
*	freedreno/a3xx/compiler: relative addressing	Rob Clark	2013-10-24	1	-1/+123
\| \| \| \|	Signed-off-by: Rob Clark <[email protected]>
*	freedreno/a3xx: fix const/rel/const-rel encoding	Rob Clark	2013-10-24	4	-88/+300
\| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \| \|	The encoding of constant, relative, and relative-const src registers is a bit more complex than originally thought, which gives an extra bit to encode const reg # at expense of taking a bit from relative offset. In most cases a3xx seems to actually use a scheme whereby it can encode an extra bit for const register. You have three possible encodings in thirteen bits: register: (11 bits for N.c) 00........... rN.c relative: (10 bits for N) 010.......... r<a0.x + N> 011.......... c<a0.x + N> const: (12 bits for N.c) 1............ cN.c Which means we can deal w/ more consts than previously thought. Signed-off-by: Rob Clark <[email protected]>
*	freedreno/a3xx: add blend state	Rob Clark	2013-10-24	2	-5/+23
\| \| \| \|	Signed-off-by: Rob Clark <[email protected]>
*	freedreno/resource: fail more gracefully	Rob Clark	2013-10-24	1	-1/+13
\| \| \| \| \| \|	Fail more gracefully when buffer allocation/import fails. Signed-off-by: Rob Clark <[email protected]>
*	freedreno: fix compile error	Rob Clark	2013-10-23	1	-1/+1
\| \| \| \| \| \|	Small typo introduced in a3ed98f. Signed-off-by: Rob Clark <[email protected]>
*	gallium: new, unified pipe_context::set_sampler_views() function	Brian Paul	2013-10-23	1	-3/+19
\| \| \| \| \| \| \| \| \| \| \| \|	The new function replaces four old functions: set_fragment/vertex/ geometry/compute_sampler_views(). Note: at this time, it's expected that the 'start' parameter will always be zero. Reviewed-by: Roland Scheidegger <[email protected]> Reviewed-by: Marek Olšák <[email protected]> Tested-by: Emil Velikov <[email protected]>
*	freedreno: use new bind_sampler_states() function	Brian Paul	2013-10-03	1	-21/+19
\|
*	freedreno: consolidate C sources list into Makefile.sources	Emil Velikov	2013-10-01	6	-41/+47
\| \| \| \| \|	Signed-off-by: Emil Velikov <[email protected]> Reviewed-by: Tom Stellard <[email protected]>
*	gallium: add flush_resource context function	Marek Olšák	2013-09-20	1	-0/+6
\| \| \| \| \| \| \| \| \|	r600g needs explicit flushing before DRI2 buffers are presented on the screen. v2: add (stub) implementations for all drivers, fix frontbuffer flushing v3: fix galahad Signed-off-by: Marek Olšák <[email protected]>
*	freedreno/a3xx: fix typo mixup w/ mipfilter	Rob Clark	2013-09-19	1	-1/+1
\| \| \| \|	Signed-off-by: Rob Clark <[email protected]>
*	freedreno: fix glReadPixels	Rob Clark	2013-09-19	1	-2/+2
\| \| \| \| \| \| \|	duh, we still need to flush if there are pending draws and it isn't an unsynchronized case. Signed-off-by: Rob Clark <[email protected]>