third_party_mesa3d

Author	SHA1	Message	Date
Nicolai Hähnle	c9fefa062b	ddebug: rewrite to always use a threaded approach This patch has multiple goals: 1. Off-load the writing of records in 'always' mode to another thread for performance. 2. Allow using ddebug with threaded contexts. This really forces us to move some of the "after_draw" handling into another thread. 3. Simplify the different modes of ddebug, both in the code and in the user interface, i.e. GALLIUM_DDEBUG. In particular, there's no 'pipelined' anymore, since we're always pipelined; and 'noflush' is replaced by 'flush', since we no longer flush by default. 4. Fix the fences in pipelining mode. They previously relied on writes via pipe_context::clear_buffer. However, on radeonsi, those could (quite reasonably) end up in the SDMA buffer. So we use the newly added PIPE_FLUSH_{TOP,BOTTOM}_OF_PIPE fences instead. 5. Improve pipelined mode overall, using the finer grained information provided by the new fences. Overall, the result is that pipelined mode should be more useful, and using ddebug in default mode is much less invasive, in the sense that it changes the overall driver behavior less (which is kind of crucial for a driver debugging tool). An example of the new hang debug output: Gallium debugger active. Hang detection timeout is 1000ms. GPU hang detected, collecting information... Draw # driver prev BOP TOP BOP dump file ------------------------------------------------------------- 2 YES YES YES NO /home/nha/ddebug_dumps/shader_runner_19919_00000000 3 YES NO YES NO /home/nha/ddebug_dumps/shader_runner_19919_00000001 4 YES NO YES NO /home/nha/ddebug_dumps/shader_runner_19919_00000002 5 YES NO YES NO /home/nha/ddebug_dumps/shader_runner_19919_00000003 Done. We can see that there were almost certainly 4 draws in flight when the hang happened: the top-of-pipe fence was signaled for all 4 draws, the bottom-of-pipe fence for none of them. In virtually all cases, we'd expect the first draw in the list to be at fault, but due to the GPU parallelism, it's possible (though highly unlikely) that one of the later draws causes a component to get stuck in a way that prevents the earlier draws from making progress as well. (In the above example, there were actually only 3 draws truly in flight: the last draw is a blit that waits for the earlier draws; however, its top-of-pipe fence is emitted before the cache flush and wait, and so the fact that the draw hasn't truly started yet can only be seen from a closer inspection of GPU state.) Acked-by: Marek Olšák <marek.olsak@amd.com>	2017-11-09 14:01:03 +01:00
Nicolai Hähnle	222a2fb998	util: move os_time.[ch] to src/util Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2017-11-09 11:57:21 +01:00
Nicolai Hähnle	81f398dcb1	ddebug: write out final driver log messages with GALLIUM_DDEBUG=always If the last operation happens to be a non-draw, such as a transfer_map that triggers a decompress blit, there may be interesting messages left in the driver log. Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2017-09-13 18:24:18 +02:00
Nicolai Hähnle	a6e7693882	gallium: remove unused PIPE_DUMP_* defines Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2017-08-22 09:53:35 +02:00
Nicolai Hähnle	635a930ad3	ddebug: remove dd_draw_record::driver_state_log It is no longer used. Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2017-08-22 09:53:35 +02:00
Nicolai Hähnle	81d7577d48	ddebug: add driver log to record dumps Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2017-08-22 09:50:44 +02:00
Nicolai Hähnle	877d800d60	ddebug: handle get_query_result_resource as a GPU call Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2017-08-02 09:46:36 +02:00
Nicolai Hähnle	aff9c54125	gallium: add util_dump_query_type and use it in ddebug Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2017-08-02 09:46:32 +02:00
Nicolai Hähnle	16923e42a4	gallium: rename util_dump_* to util_str_* for enum-to-string conversion This is mostly mechanical search-and-replace, plus touching up the macros in u_dump_defines.c manually a bit. Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2017-08-02 09:46:24 +02:00
Nicolai Hähnle	d91f97f91d	ddebug: handle some cases of non-TGSI shaders NIR shaders are not captured properly in pipelined mode currently. This would require shader cloning, which requires linking all the Gallium drivers against NIR. We can always do that later. v2: avoid immediate crashes in pipelined mode Reviewed-by: Marek Olšák <marek.olsak@amd.com> (v1)	2017-07-05 12:27:11 +02:00
Marek Olšák	330d0607ed	gallium: remove pipe_index_buffer and set_index_buffer pipe_draw_info::indexed is replaced with index_size. index_size == 0 means non-indexed. Instead of pipe_index_buffer::offset, pipe_draw_info::start is used. For indexed indirect draws, pipe_draw_info::start is added to the indirect start. This is the only case when "start" affects indirect draws. pipe_draw_info::index is a union. Use either index::resource or index::user depending on the value of pipe_draw_info::has_user_indices. v2: fixes for nine, svga	2017-05-10 19:00:16 +02:00
Marek Olšák	22f6624ed3	gallium: separate indirect stuff from pipe_draw_info - 80 -> 56 bytes For faster initialization of non-indirect draws.	2017-05-10 19:00:16 +02:00
Marek Olšák	c24c3b94ed	gallium: decrease the size of pipe_vertex_buffer - 24 -> 16 bytes	2017-05-10 19:00:16 +02:00
Nicolai Hähnle	db3559da12	ddebug: implement dd_dump_launch_grid Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2017-05-10 08:58:37 +02:00
Nicolai Hähnle	bf4ecfec4b	ddebug: extract dd_dump_shader Will be re-used for compute shaders. Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2017-05-10 08:58:34 +02:00
Nicolai Hähnle	d15b1f6e2d	gallium/ddebug: dump missing members of pipe_draw_info Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2017-04-14 22:50:54 +02:00
Marek Olšák	681adbc18c	ddebug: implement clear_texture Reviewed-by: Nicolai Hähnle <nicolai.haehnle@amd.com>	2017-03-30 18:53:42 +02:00
Timothy Arceri	da40ac65c7	gallium/util: remove PIPE_THREAD_ROUTINE() This was made unnecessary with `fd33a6bcd7`. This was mostly done with: find ./src -type f -exec sed -i -- \ 's:PIPE_THREAD_ROUTINE($[^,]$, $[^)]$):int\n\1(void \*\2):g' {} \; With some small manual tidy ups. Reviewed-by: Plamena Manolova <plamena.manolova@intel.com> Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2017-03-07 09:12:16 +11:00
Timothy Arceri	628e84a58f	gallium/util: replace pipe_mutex_unlock() with mtx_unlock() pipe_mutex_unlock() was made unnecessary with `fd33a6bcd7`. Replaced using: find ./src -type f -exec sed -i -- \ 's:pipe_mutex_unlock($[^)]*$):mtx_unlock(\&\1):g' {} \; Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2017-03-07 08:53:05 +11:00
Timothy Arceri	ba72554f3e	gallium/util: replace pipe_mutex_lock() with mtx_lock() replace pipe_mutex_lock() was made unnecessary with `fd33a6bcd7`. Replaced using: find ./src -type f -exec sed -i -- \ 's:pipe_mutex_lock($[^)]*$):mtx_lock(\&\1):g' {} \; Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2017-03-07 08:52:38 +11:00
Marek Olšák	9e1dc10432	ddebug: fix hang detection with deferred flushes Reviewed-by: Nicolai Hähnle <nicolai.haehnle@amd.com>	2016-11-29 23:52:31 +01:00
Marek Olšák	10e5f126dd	ddebug: dump most driver information with GALLIUM_DDEBUG=always Reviewed-by: Nicolai Hähnle <nicolai.haehnle@amd.com>	2016-10-05 21:03:23 +02:00
Marek Olšák	c723acc03d	ddebug: dump shader buffers and images this was unimplemented Reviewed-by: Nicolai Hähnle <nicolai.haehnle@amd.com>	2016-09-13 20:38:25 +02:00
Marek Olšák	54272e18a6	gallium: add a pipe_context parameter to fence_finish required by glClientWaitSync (GL 4.5 Core spec) that can optionally flush the context Reviewed-by: Rob Clark <robdclark@gmail.com> Reviewed-by: Nicolai Hähnle <nicolai.haehnle@amd.com>	2016-08-10 01:11:10 +02:00
Marek Olšák	a909210131	gallium: add render_condition_enable param to clear_render_target/depth_stencil Reviewed-by: Roland Scheidegger <sroland@vmware.com> Reviewed-by: Nicolai Hähnle <nicolai.haehnle@amd.com>	2016-08-10 01:10:21 +02:00
Marek Olšák	06b2fd04f6	ddebug: dump driver states and shaders for apitrace calls I think this was an oversight when the PIPE_DUMP flags were added. Reviewed-by: Nicolai Hähnle <nicolai.haehnle@amd.com>	2016-08-09 15:35:42 +02:00
Marek Olšák	6573ad69ef	ddebug: print the command line to all logs (v2) for piglit with the pipelined hang detection mode v2: rebase on top of Brian's commit Reviewed-by: Nicolai Hähnle <nicolai.haehnle@amd.com>	2016-08-03 17:46:46 +02:00
Marek Olšák	840353059a	ddebug: don't use fmemopen on non-Linux OS Bugzilla: https://bugs.freedesktop.org/show_bug.cgi?id=97140 Reviewed-by: Nicolai Hähnle <nicolai.haehnle@amd.com>	2016-08-03 17:46:46 +02:00
Nicolai Hähnle	bade0cd0fb	ddebug: use pclose to close a popen()'d FILE Found by Coverity. Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2016-07-28 10:47:51 +01:00
Marek Olšák	b47727a83a	ddebug: implement pipelined hang detection mode For good performance while being able to generate decent hang reports. The report doesn't contain the parsed IB and the buffer list, but it isolates the draw call and dumps shaders while not having to flush the context. This is for GPU hangs that are harder to reproduce and require interactive playing for minutes or even hours. dd_pipe.h explains some implementation details. Initializing, copying (recording) and clearing states is most of the code. The performance should be at least 50% of the normal performance depending on the circumstances. (i.e. 50% is expected to be the worst case scenario, not the best case) The majority of time is spent in dump_debug_state(PIPE_DUMP_CURRENT_SHADERS) and that's after all the optimizations in later patches. There is no obvious way to optimize that further. Reviewed-by: Nicolai Hähnle <nicolai.haehnle@amd.com>	2016-07-26 23:06:46 +02:00
Marek Olšák	0795a3d54f	ddebug: don't save pointers to call parameters Reviewed-by: Nicolai Hähnle <nicolai.haehnle@amd.com>	2016-07-26 23:06:46 +02:00
Marek Olšák	e4079677a7	ddebug: move dd_call into dd_pipe.h Reviewed-by: Nicolai Hähnle <nicolai.haehnle@amd.com>	2016-07-26 23:06:46 +02:00
Marek Olšák	d50f9e9b04	ddebug: separate draw call dumping logic Reviewed-by: Nicolai Hähnle <nicolai.haehnle@amd.com>	2016-07-26 23:06:46 +02:00
Marek Olšák	95c3025a41	ddebug: move all states into a separate structure Reviewed-by: Nicolai Hähnle <nicolai.haehnle@amd.com>	2016-07-26 23:06:46 +02:00
Marek Olšák	f7720948cc	ddebug: write contents of dmesg into hang reports Reviewed-by: Nicolai Hähnle <nicolai.haehnle@amd.com>	2016-07-26 23:06:46 +02:00
Marek Olšák	6b9924ccb6	ddebug: don't use abort() We don't want a core dump. Reviewed-by: Nicolai Hähnle <nicolai.haehnle@amd.com>	2016-07-26 23:06:46 +02:00
Marek Olšák	26ef8158ac	ddebug: make dd_get_file_stream accept the screen only Reviewed-by: Nicolai Hähnle <nicolai.haehnle@amd.com>	2016-07-26 23:06:46 +02:00
Marek Olšák	6bf81de339	gallium: rework flags for pipe_context::dump_debug_state The pipelined hang detection mode will not want to dump everything. (and it's also time consuming) It will only dump shaders after a draw call and then dump the status registers separately if a hang is detected. Reviewed-by: Nicolai Hähnle <nicolai.haehnle@amd.com>	2016-07-26 23:06:46 +02:00
Marek Olšák	642cf400aa	ddebug: add an option to dump info about a specific apitrace call Reviewed-by: Bas Nieuwenhuizen <bas@basnieuwenhuizen.nl> Reviewed-by: Nicolai Hähnle <nicolai.haehnle@amd.com>	2016-07-05 00:47:12 +02:00
Marek Olšák	1daec2b795	ddebug: implement pipe_context::generate_mipmap Reviewed-by: Bas Nieuwenhuizen <bas@basnieuwenhuizen.nl> Reviewed-by: Nicolai Hähnle <nicolai.haehnle@amd.com>	2016-07-05 00:47:12 +02:00
Marek Olšák	50b2235478	ddebug: record and dump apitrace call numbers Reviewed-by: Bas Nieuwenhuizen <bas@basnieuwenhuizen.nl> Reviewed-by: Nicolai Hähnle <nicolai.haehnle@amd.com>	2016-07-05 00:47:12 +02:00
Bas Nieuwenhuizen	ac77fb74a0	gallium/ddebug: Implement launch_grid. Does not implement dumping info. Signed-off-by: Bas Nieuwenhuizen <bas@basnieuwenhuizen.nl> Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2016-05-13 07:43:46 +02:00
Nicolai Hähnle	41875ac4ed	gallium/ddebug: add 'verbose' option This currently just writes out the name of dump files, which can be useful to easily correlate those files with other log outputs (driver debug output, apitrace calls, etc.) Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2016-01-26 09:58:55 -05:00
Nicolai Hähnle	f4c8fa4e49	gallium/ddebug: make 'noflush' also affect 'always' mode This changes the default behavior of 'always' mode to be consistent with hang detection mode. I have used this to more easily compare dumped command streams using diff. Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2016-01-26 09:58:49 -05:00
Nicolai Hähnle	d640f179d3	gallium/ddebug: regularly log the total number of draw calls This helps in the use of GALLIUM_DDEBUG_SKIP: first run a target application with skip set to a very large number and note how many draw calls happen before the bug. Then re-run, skipping the corresponding number of calls. Despite the additional run, this can still be much faster than not skipping anything. Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2015-12-12 15:23:50 -05:00
Nicolai Hähnle	b86d5ccae2	gallium/ddebug: add GALLIUM_DDEBUG_SKIP option When we know that hangs occur only very late in a reproducible run (e.g. apitrace), we can save a lot of debugging time by skipping the flush and hang detection for earlier draw calls. Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2015-12-12 15:23:34 -05:00
Marek Olšák	89f73827d0	ddebug: separate creation of debug files This will be used by radeonsi for logging. Reviewed-by: Michel Dänzer <michel.daenzer@amd.com>	2015-10-03 22:06:07 +02:00
Marek Olšák	525921ed51	gallium/ddebug: new pipe for hang detection and driver state dumping (v2) v2: lots of improvements This is like identity or trace, but simpler. It doesn't wrap most states. Run with: GALLIUM_DDEBUG=1000 [executable] where "executable" is the app and "1000" is in miliseconds, meaning that the context will be considered hung if a fence fails to signal in 1000 ms. If that happens, all shaders, context states, bound resources, draw parameters, and driver debug information (if any) will be dumped into: /home/$username/dd_dumps/$processname_$pid_$index. Note that the context is flushed after every draw/clear/copy/blit operation and then waited for to find the exact call that hangs. You can also do: GALLIUM_DDEBUG=always to do the dumping after every draw/clear/copy/blit operation without flushing and waiting. Examples of driver states that can be dumped are: - Hardware status registers saying which hw block is busy (hung). - Disassembled shaders in a human-readable form. - The last submitted command buffer in a human-readable form. v2: drop pipe-loader changes, drop SConscript rename dd.h -> dd_pipe.h Acked-by: Christian König <christian.koenig@amd.com> Acked-by: Alex Deucher <alexander.deucher@amd.com>	2015-08-26 19:25:18 +02:00

48 Commits