KonstantinSeurer/mesa

Commit Graph

Author	SHA1	Message	Date
Emil Velikov	1301674c39	st/dri: make swrast_no_present member of dri_screen Just like the dri2 options, this is better suited in the dri_screen struct. Signed-off-by: Emil Velikov <emil.velikov@collabora.com>	2018-10-03 13:38:05 +01:00
Emil Velikov	80b62e2d6d	st/dri: inline dri2_buffer.h within dri2.c The header was used only by dri2.c, containing a two-member struct and cast wrapper. Just inline it where it's used/needed. Signed-off-by: Emil Velikov <emil.velikov@collabora.com> Reviewed-by: Eric Engestrom <eric.engestrom@intel.com>	2018-10-03 13:38:05 +01:00
Emil Velikov	89c2c386c0	st/xa: remove unused xa_screen::d[s]_depth_bits_last Unused since the initial import. Signed-off-by: Emil Velikov <emil.velikov@collabora.com> Reviewed-by: Eric Engestrom <eric.engestrom@intel.com>	2018-10-03 13:38:05 +01:00
Emil Velikov	5ade4b10e2	mesa: use C99 initializer in get_gl_override() The overrides array contains entries indexed on the gl_api enum. Use a C99 initializer to make it a bit more obvious. Signed-off-by: Emil Velikov <emil.velikov@collabora.com> Reviewed-by: Ian Romanick <ian.d.romanick@intel.com> Reviewed-by: Eric Engestrom <eric.engestrom@intel.com>	2018-10-03 13:38:05 +01:00
Gabriel Majeri	f0b987646a	anv: Ensure discreteQueuePriorities is at least 2 This is the minimum value according to the spec. Reviewed-by: Jason Ekstrand <jason@jlekstrand.net> Reviewed-by: Iago Toral Quiroga <itoral@igalia.com>	2018-10-03 07:57:37 +02:00
Timothy Arceri	2b5f42068d	r600: use build-id when available for disk cache Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2018-10-03 09:49:21 +10:00
Timothy Arceri	397f2603eb	nouveau: use build-id when available for disk cache Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2018-10-03 09:49:21 +10:00
Timothy Arceri	2169acbf34	radeonsi: use build-id when available for disk cache Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2018-10-03 09:49:21 +10:00
Timothy Arceri	83ea8dd99b	util: add disk_cache_get_function_identifier() This can be used as a drop in replacement for disk_cache_get_function_timestamp(). Here we use build-id to generate a driver-id rather than build timestamp if available. This should resolve issues such as distros using reproducable builds and flatpak not having real build timestamps. Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2018-10-03 09:49:21 +10:00
Timothy Arceri	6a884014e4	util: rename timestamp param in disk_cache_create() Only some drivers use a timestamp here. Others use things such as build-id, or even a combination of build-ids from Mesa and LLVM. Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2018-10-03 09:49:21 +10:00
Józef Kucia	e24a4e05c7	radeonsi: avoid sending GS_EMIT in shaders without outputs Fixes GPU hangs. Cc: 18.1 18.2 <mesa-stable@lists.freedesktop.org> Bugzilla: https://bugs.freedesktop.org/show_bug.cgi?id=107857 Signed-off-by: Józef Kucia <joseph.kucia@gmail.com> Signed-off-by: Marek Olšák <marek.olsak@amd.com>	2018-10-02 17:13:52 -04:00
Fritz Koenig	08f97407fb	i965: Replace checks for rb->Name with FlipY (v2) In the GL_MESA_framebuffer_flip_y implementation _mesa_is_winsys_fbo checks were replaced with FlipY checks. rb->Name is also used to determine if a buffer is winsys. v2: Fixes annotation [for emil] Fixes: `ab05dd183c` ("i965: implement GL_MESA_framebuffer_flip_y [v3]") Reviewed-by: Jason Ekstrand <jason@jlekstrand.net> Reviewed-by: Chad Versace <chadversary@chromium.org>	2018-10-02 11:28:46 -07:00
Marek Olšák	2fd58d8eb2	radeonsi: initialize ac_gpu_info::name when using SI_FORCE_FAMILY so that it's not NULL when loading radeonsi and a GCN GPU is not present in the system.	2018-10-02 12:21:49 -04:00
Marek Olšák	0b062f0419	radeonsi: don't set the VS prolog key for the blit VS	2018-10-02 12:21:49 -04:00
Jason Ekstrand	58360ca09d	spirv: Move function call handling to vtn_cfg It makes way more sense for it to live there with the rest of function handling. Reviewed-by: Iago Toral Quiroga <itoral@igalia.com>	2018-10-02 10:24:56 -05:00
Jason Ekstrand	00f385e6d4	nir/from_ssa: Don't rewrite derefs destinations to registers We already call nir_rematerialize_derefs_in_use_blocks_impl prior to calling nir_lower_ssa_defs_to_regs_block so the assertion that all deref uses in the block should hold. This fixes the following CTS test when SPIR-V optimization recipe 1: dEQP-VK.glsl.struct.local.loop_nested_struct_array_vertex Fixes: `606eb56ab9` "intel/nir: Only lower load/store derefs" Reviewed-by: Iago Toral Quiroga <itoral@igalia.com>	2018-10-02 10:24:56 -05:00
Jason Ekstrand	bfc89c668e	nir/cf: Remove phi sources if needed in nir_handle_add_jump If the block in which the jump is inserted is the predecessor of a phi then we need to remove phi sources otherwise the phi may end up with things improperly connected. This fixes the following CTS test when dEQP is run with SPIR-V optimization recipe 1: dEQP-VK.glsl.functions.control_flow.return_in_nested_loop_vertex Cc: mesa-stable@lists.freedesktop.org Reviewed-by: Iago Toral Quiroga <itoral@igalia.com>	2018-10-02 10:24:56 -05:00
Eric Engestrom	7b0752fb10	anv: suppress warning about unhandled image layout Let's just be explicit that VK_NV_shading_rate_image is not supported. Suggested-by: Jason Ekstrand <jason.ekstrand@intel.com> Fixes: `6ee1709170` "vulkan: Update the XML and headers to 1.1.86" Signed-off-by: Eric Engestrom <eric.engestrom@intel.com> Reviewed-by: Jason Ekstrand <jason.ekstrand@intel.com>	2018-10-02 15:09:29 +01:00
Rob Clark	ae78489d3e	freedreno/a6xx: hwbinning Signed-off-by: Rob Clark <robdclark@gmail.com>	2018-10-02 10:08:18 -04:00
Rob Clark	8ff349e564	freedreno: update generated headers Signed-off-by: Rob Clark <robdclark@gmail.com>	2018-10-02 10:08:18 -04:00
Jason Ekstrand	7e7959fcb7	intel/fs: Fix a typo in need_matching_subreg_offset This fixes a bunch of Vulkan subgroup tests on little core platforms. Fixes: `4150920b95` "intel/fs: Add a helper for emitting scan operations" Reviewed-by: Caio Marcelo de Oliveira Filho <caio.oliveira@intel.com> Tested-by: Mark Janes <mark.a.janes@intel.com> Reviewed-by: Lionel Landwerlin <lionel.g.landwerlin@intel.com>	2018-10-02 07:44:25 -05:00
Timothy Arceri	ea66bfda88	util: disable cache if we have no build-id and timestamp is zero Timestamp can be zero for example when Flatpak is used. In this case just disable the cache rather then segfaulting when incompatible cache items are loaded. V2: actually return false when mtime is 0. Reviewed-by: Bas Nieuwenhuizen <bas@basnieuwenhuizen.nl>	2018-10-02 22:07:55 +10:00
Timothy Arceri	0e6cdfd561	radeonsi: add a workaround for bitfield_extract when count is 0 This ports the fix from `3d41757788`. Both LLVM 7 & 8 continue to have this problem. It fixes rendering issues in some menu and loading screens of Civ VI which can be seen in the trace from bug 104602. Note: This does not fix the black triangles on Vega for bug 104602. Reviewed-by: Marek Olšák <marek.olsak@amd.com> Bugzilla: https://bugs.freedesktop.org/show_bug.cgi?id=104602 Bugzilla: https://bugs.freedesktop.org/show_bug.cgi?id=107276	2018-10-02 08:39:51 +10:00
Jason Ekstrand	e4538b93f5	anv: Implement VK_KHR_driver_properties Reviewed-by: Bas Nieuwenhuizen <bas@basnieuwenhuizen.nl>	2018-10-01 13:21:12 -05:00
Jason Ekstrand	6ee1709170	vulkan: Update the XML and headers to 1.1.86 Reviewed-by: Bas Nieuwenhuizen <bas@basnieuwenhuizen.nl>	2018-10-01 11:43:20 -05:00
Samuel Pitoiset	c2867e4c2a	radv: do not try to set DCC_CONTROL when image doesn't use DCC Unnecessary. While we are at it, remove the check for pre-VI because it's already checked earlier. Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Bas Nieuwenhuizen <bas@basnieuwenhuizen.nl>	2018-10-01 12:13:12 +02:00
Samuel Pitoiset	f622ab889a	radv: add a sanity check for mutable formats and TC-compat HTILE If apps use the MUTABLE bit and the same formats as the image one in the list, we can still enable TC-compat HTILE. I don't think this happens often but given the fact that TC-compat HTILE allows a nice boost in some situations, it's worth checking. Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Bas Nieuwenhuizen <bas@basnieuwenhuizen.nl>	2018-10-01 12:13:09 +02:00
Samuel Pitoiset	dc91c4d40a	radv: disable HTILE for very small depth surfaces Like we disable DCC/CMASK for small color surfaces as well. Serious Sam 2017 creates a 1x1 depth surface and I think it should be faster to do slow clears on the graphics queue instead of fast clears on compute, and eventually a depth expand if the surface isn't TC-compatible HTILE. Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Bas Nieuwenhuizen <bas@basnieuwenhuizen.nl>	2018-10-01 10:16:33 +02:00
Samuel Pitoiset	6cfa321c39	radv: add potential missing fields for DB_EQAA Other drivers set these two as well, just apply the same rule. Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Bas Nieuwenhuizen <bas@basnieuwenhuizen.nl>	2018-10-01 10:16:30 +02:00
Samuel Pitoiset	bd6df2f923	radv: disable complicated point clipping against user clip planes I don't think this is required by Vulkan too. Ported from RadeonSI (AMDVLK doesn't set it either). Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Bas Nieuwenhuizen <bas@basnieuwenhuizen.nl>	2018-10-01 10:16:25 +02:00
Michel Dänzer	cb863de626	gallium/util: Clarify comment in util_init_thread_pinning As discussed in the review of the patch which added the comment: Nothing happens when a thread is created, because pthread_atfork doesn't affect creating threads. However, spawning a child process will likely crash. Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2018-09-28 17:52:11 +02:00
Samuel Pitoiset	3fb4adae83	radv: do not sync CP DMA when copying buffers We already track if the DMA engine is busy/idle with a flag, and we emit a packet that waits for all CP DMA operations to be complete. This is done at end of command buffer because the kernel doesn't wait for them, and also when emitting barriers, so it should be safe. This improves small copies for both aligned and unaligned sizes. Aligned sizes: BEFORE: 1 KB: 59.840000 ms 2 KB: 71.200000 ms AFTER: 1 KB: 31.200000 ms 2 KB: 31.040000 ms Unaligned sizes: BEFORE: 2 KB: 68.3200 ms 3 KB: 79.3600 ms 5 KB: 76.6400 ms 9 KB: 90.8800 ms 17 KB: 116.0000 ms AFTER: 2 KB: 31.0400 ms 3 KB: 32.0000 ms 5 KB: 30.8800 ms 9 KB: 30.5600 ms 17 KB: 29.6000 ms Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Bas Nieuwenhuizen <bas@basnieuwenhuizen.nl>	2018-09-28 09:08:52 +02:00
Samuel Pitoiset	621e70dd40	radv: adjust the CmdUpdateBuffer threshold for optimal performance According to my benchmark results, it appears that we should reduce the threshold to 1024. BEFORE: 1 KB: 68.656000 ms 2 KB: 118.368000 ms AFTER: 1 KB: 31.760000 ms 2 KB: 29.840000 ms Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Bas Nieuwenhuizen <bas@basnieuwenhuizen.nl>	2018-09-28 09:08:44 +02:00
Samuel Pitoiset	5d6a560a29	radv: do not use the availability bit for timestamp queries It's unnecessary because we can just check if the timestamp is to different to the default value when a pool is created or resetted. Instead of waiting for the availability bit to be 1, we have to emit a not equal WAIT_REG_MEM for checking if the timestamp is ready. Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Dave Airlie <airlied@redhat.com>	2018-09-28 09:08:03 +02:00
Kristian H. Kristensen	3e90505224	freedreno/a6xx: Build up draw dword0 outside visibilty if statement Pulling this logic out means we can share the logic and avoid a couple of temporary variables that helped make things clearer before. Note that in either vismode case, we always program vismode 0. Signed-off-by: Kristian H. Kristensen <hoegsberg@chromium.org>	2018-09-27 16:08:52 -04:00
Kristian H. Kristensen	74a87cdaa6	freedreno/a6xx: Simplify draw_emit() branches a bit Now that we've copied the emit logic into each branch of the if (info->index_size) statement, we can simplify the logic a bit according to which case we're in. Signed-off-by: Kristian H. Kristensen <hoegsberg@chromium.org>	2018-09-27 16:08:52 -04:00
Kristian H. Kristensen	2516073cb6	freedreno/a6xx: Copy OUT_RING() part into each branch of the index if Signed-off-by: Kristian H. Kristensen <hoegsberg@chromium.org>	2018-09-27 16:08:52 -04:00
Kristian H. Kristensen	c3d58d9ffc	freedreno/a6xx: Split fd6_draw_emit into direct and indirect paths This splits the two code paths into separate functions and moves the "if (info->indirect)" test into draw_impl(). Signed-off-by: Kristian H. Kristensen <hoegsberg@chromium.org>	2018-09-27 16:08:52 -04:00
Kristian H. Kristensen	adcd83fb22	freedreno/a6xx: Inline fd6_draw() Simplify the code a bit by inlining this helper. Signed-off-by: Kristian H. Kristensen <hoegsberg@chromium.org>	2018-09-27 16:08:52 -04:00
Kristian H. Kristensen	fb1c6b89a2	freedreno/a6xx: Move emit_marker and wfi to draw_impl() This way the markers clearly bracket the draw call and isn't duplicated for both direct and indirect draw code. Signed-off-by: Kristian H. Kristensen <hoegsberg@chromium.org>	2018-09-27 16:08:52 -04:00
Kristian H. Kristensen	0559050557	freedreno/a6xx: Move inline functions out of fd6_draw.h Only used in fd6_draw.c so put them there. Signed-off-by: Kristian H. Kristensen <hoegsberg@chromium.org>	2018-09-27 16:08:52 -04:00
Hyunjun Ko	1a40faa864	freedreno: fix a typo in launch_grid	2018-09-27 16:06:19 -04:00
Hyunjun Ko	aef410f31e	freedreno/ir3: fix the param order of cmpxchg According to the following definition, int AtomicCompSwap(inout int mem, uint compare, uint data); the preceding one in atomic_comp_swap of NIR is compare and data is followed, while src0 for cmpxchg needs vec2(data, compare) So for ssbo/image deref comp_swap, that should be reversed. Fixes: dEQP-GLES31.functional.image_load_store..atomic.comp_swap	2018-09-27 16:05:49 -04:00
Rob Clark	49d22c2dfc	freedreno/a6xx: fix shaders w/ >= 24 regs Possibly these bits mean something else now. Blob always seems to use FOUR_QUADS, and changing to TWO_QUADS seems to cause different threads to overlap registers. Signed-off-by: Rob Clark <robdclark@gmail.com>	2018-09-27 15:49:14 -04:00
Rob Clark	6530fcc4a7	freedreno/a6xx: fix gl_FragCoord.w Signed-off-by: Rob Clark <robdclark@gmail.com>	2018-09-27 15:45:44 -04:00
Rob Clark	919741b8d5	freedreno: handle invalidated buffers harder Do a better job of skipping mem2gmem/gmem2mem.. Signed-off-by: Rob Clark <robdclark@gmail.com>	2018-09-27 15:41:46 -04:00
Rob Clark	19e9d28646	freedreno/a6xx: fix constlen Fix a few bits of confusion, as with previous gen's constlen is aligned to 4, and value in bitfield is left-shifted by 2 (ie. divided by 4). But this is done by the CONSTLEN() accessor/builder fxn, so don't do it twice. Also HLSQ_FS_CNTL.CONSTLEN is not special. Signed-off-by: Rob Clark <robdclark@gmail.com>	2018-09-27 15:33:10 -04:00
Rob Clark	12de415ad1	freedreno: fix inorder rendering case Signed-off-by: Rob Clark <robdclark@gmail.com>	2018-09-27 15:32:39 -04:00
Rob Clark	b65b6f7606	freedreno/a6xx: backface stencil state Signed-off-by: Rob Clark <robdclark@gmail.com>	2018-09-27 15:31:56 -04:00
Rob Clark	93db15d300	freedreno/a6xx: fix gpu crash with separate-stencil Fixes a crash in (of all things) dEQP-GLES2.info.vendor with --deqp-surface-type=fbo.. Signed-off-by: Rob Clark <robdclark@gmail.com>	2018-09-27 15:31:34 -04:00
Rob Clark	a52ef80d24	freedreno/a6xx: fix MRT config Signed-off-by: Rob Clark <robdclark@gmail.com>	2018-09-27 15:30:36 -04:00
Rob Clark	8930e83642	freedreno: fix potential hang when destroying batch batch_flush_reset_dependencies() expects to be called unlocked, and can call fd_batch_reference() which can try to aquire the screen lock again. Signed-off-by: Rob Clark <robdclark@gmail.com>	2018-09-27 15:29:45 -04:00
Rob Clark	ef6d15f8a8	freedreno: fix corrupted fb state In `c3d9f29b` we allowed ctx->batch to be null, and started tracking the current framebuffer state in fd_context. But the existing logic in fd_blitter_pipe_begin() would, if !ctx->batch, set null fb state to be restored after blit. Which broke the world of deqp (and probably other things) Fixes: `c3d9f29b78` freedreno: allocate ctx's batch on demand Signed-off-by: Rob Clark <robdclark@gmail.com>	2018-09-27 15:27:38 -04:00
Rob Clark	5bb96bf73a	freedreno: simplify pctx->clear() This is defined to always clear the entire surface(s) specified, regardless of scissor state.. mesa/st will turn scissored clears into a draw. So rip about a bunch of unnecessary machinery. Also remove a comment that was obsolete since using u_blitter to turn clear into draw (for the cases where there isn't a hw blitter fast-path). Signed-off-by: Rob Clark <robdclark@gmail.com>	2018-09-27 15:26:32 -04:00
Rob Clark	a7fa44cd33	freedreno: fix FD_MESA_DEBUG=flush The logic to force a flush every draw was short-circuited with newer kernels. Also it should apply to clears as well. Signed-off-by: Rob Clark <robdclark@gmail.com>	2018-09-27 15:25:49 -04:00
Rob Clark	83c5c026ee	freedreno: fix scissor state emit The effective scissor changes based on rasterizer->scissor flag, so we need to re-emit scissor state when rasterizer state changes. Signed-off-by: Rob Clark <robdclark@gmail.com>	2018-09-27 15:25:24 -04:00
Rob Clark	106f18258a	freedreno: update generated headers Signed-off-by: Rob Clark <robdclark@gmail.com>	2018-09-27 15:25:01 -04:00
Erik Faye-Lund	c3486cd8c9	st/mesa: do not call update_framebuffer_size with NULL pointer In st_renderbuffer_alloc_storage, we avoid allocating storage for zero-sized buffers, leading to this pointer being NULL. We already take care to avoid dereferencing these pointers for color-buffers, but not for depth/stencil-buffers. So let's thread a bit more carefully here. This avoids a crash while running Piglit's glx/glx-visuals-stencil test, both on virgl and r600g. Signed-off-by: Erik Faye-Lund <erik.faye-lund@collabora.com> Reviewed-by: Guillaume Charifi <guillaume.charifi@sfr.fr> Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2018-09-27 10:33:44 +02:00
Maxime	dd333c66bd	vulkan: Disable randr lease for libxcb < 1.13 Since the Randr lease code was added, compiling against libxcb 1.12 no longer works. CC: mesa-stable@lists.freedesktop.org Bugzilla: https://bugs.freedesktop.org/show_bug.cgi?id=108024 Fixes: `7ab1fffcd2` Tested-By: Maxime <berillions@gmail.com> Fixes: `7ab1fffcd2` "vulkan: Add EXT_acquire_xlib_display [v5]"	2018-09-27 16:31:42 +10:00
Bas Nieuwenhuizen	40585ddb48	radv: Remove garbage comment. Trivial.	2018-09-27 02:04:06 +02:00
Bas Nieuwenhuizen	0207ebcbf1	radv: Do not use multiple draws for multisample copies. Use sample rate shading instead, should give better locality. Makes Nier with 8x msaa on a Raven go 5 fps -> 7 fps in the menu. Reviewed-by: Samuel Pitoiset <samuel.pitoiset@gmail.com>	2018-09-27 02:04:00 +02:00
Jordan Justen	ca1d3fc538	anv: If softpin is supported, use it with the hiz clear value bo Signed-off-by: Jordan Justen <jordan.l.justen@intel.com> Reviewed-by: Jason Ekstrand <jason@jlekstrand.net> Reviewed-by: Nanley Chery <nanley.g.chery@intel.com>	2018-09-26 10:21:23 -07:00
Jordan Justen	2a97390552	anv: s/batch/value_bo/ on anv_device_init_hiz_clear_batch Signed-off-by: Jordan Justen <jordan.l.justen@intel.com> Reviewed-by: Jason Ekstrand <jason@jlekstrand.net> Reviewed-by: Nanley Chery <nanley.g.chery@intel.com>	2018-09-26 10:21:23 -07:00
Jason Ekstrand	b3f477ef7a	intel/isl: Add a unit suffixes to some struct fields and variables I was about to make the claim to someone that every field in isl_surf is either an enum or has explicit units. Then I looked at isl_surf and discovered this claim was wrong. We should fix that. This commit does a few refactors: * Add _B suffixes to some struct fields * Add _B to some variables and parameters * Rename row_pitch_tiles -> row_pitch_tl Reviewed-by: Nanley Chery <nanley.g.chery@intel.com>	2018-09-26 08:52:26 -05:00
Axel Davy	0d495bec25	radeonsi: NaN should pass kill_if Fixes: https://bugs.freedesktop.org/show_bug.cgi?id=105333 Fixes: https://github.com/iXit/Mesa-3D/issues/314 For this application, NaN is passed to KILL_IF and is expected to pass. v2: Explain in the code why UGE is used. Signed-off-by: Axel Davy <davyaxel0@gmail.com> Reviewed-by: Marek Olšák <marek.olsak@amd.com> CC: <mesa-stable@lists.freedesktop.org>	2018-09-25 22:05:24 +02:00
Axel Davy	46814e771a	st/nine: Do not mark both ff vs and ps updated Previously if only ff vs or only ff ps was used, the constants for both were marked as updated, while only the constants of the used ff shader were updated. Now that NINE_STATE_FF_VS and NINE_STATE_FF_PS do not intersect anymore, we can correctly mark the correct set of constant as updated. Fixes: https://github.com/iXit/Mesa-3D/issues/319 Signed-off-by: Axel Davy <davyaxel0@gmail.com>	2018-09-25 22:05:24 +02:00
Axel Davy	8e0526555d	st/nine: Split NINE_STATE_FF_OTHER NINE_STATE_FF_OTHER was mostly ff vs states. Rename it to NINE_STATE_FF_VS_OTHER and move common states with ps to NINE_STATE_FF_PS_CONSTS (renamed from NINE_STATE_FF_PSSTAGES). Signed-off-by: Axel Davy <davyaxel0@gmail.com>	2018-09-25 22:05:24 +02:00
Axel Davy	5f7a41c33b	st/nine: Add dummy ff shader state Some states only affect the ff shader, not its constants. Currently we don't check anything and always recompute the ff shader key. However we do check for NINE_STATE_FF_OTHER and if set we reupload some constants. Thus for those states which had NINE_STATE_FF_OTHER set but didn't need it, replace by a dummy ff shader state (which is easier to understand for an external reader than just setting 0 and more future proof). Signed-off-by: Axel Davy <davyaxel0@gmail.com>	2018-09-25 22:05:24 +02:00
Axel Davy	f6bf1d2db0	st/nine: Mark pointsize states as ff states The pointsize states were missing the ff NINE_STATE_FF_OTHER flag, and thus might miss state updates when using ff. Fixes some wine tests. Signed-off-by: Axel Davy <davyaxel0@gmail.com>	2018-09-25 22:05:24 +02:00
Axel Davy	89beea100f	st/nine: Minor refactor of a few NINE_STATE_* flags Rename NINE_STATE_FOG_SHADER, NINE_STATE_POINTSIZE_SHADER and NINE_STATE_PS1X_SHADER into NINE_STATE_VS_PARAMS_MISC and NINE_STATE_PS_PARAMS_MISC. The behaviour is unchanged, except one minor change: D3DRS_FOGTABLEMODE doesn't need to affect VS. Signed-off-by: Axel Davy <davyaxel0@gmail.com>	2018-09-25 22:05:24 +02:00
Axel Davy	7ae2509ce0	st/nine: Increase maximum number of temp registers With some test app I hit the limit. As we allocate on demand (up to the maximum), it is free to increase the limit. Signed-off-by: Axel Davy <davyaxel0@gmail.com> CC: <mesa-stable@lists.freedesktop.org>	2018-09-25 22:05:24 +02:00
Axel Davy	dc4b53e129	st/nine: Lock the entire buffer in some cases. Previously we had already found that for MANAGED buffers the buffer started dirty (which meant all writes out of bound before the first draw call using the buffer have to be taken into account). Possibly it is the same for the other types of buffers. For now always lock the entire buffer (starting from the offset) for these (except for DYNAMIC buffers, which might hurt performance too much). Fixes: https://github.com/iXit/Mesa-3D/issues/301 Signed-off-by: Axel Davy <davyaxel0@gmail.com>	2018-09-25 22:05:24 +02:00
Axel Davy	0eeb583650	st/nine: Don't call SetCursor until a cursor is set The previous code was ignoring the input until a cursor is set inside d3d (with SetCursorProperties), as expected by wine tests. However it did still make a call to ID3DPresent_SetCursor, which would result into a SetCursor(NULL) call, thus hidding any cursor set outside d3d, which we shouldn't do. Add comment about not avoiding redundant ID3DPresent_SetCursor calls once a cursor has been set in d3d, as it has been tested to cause regressions. Fixes: https://github.com/iXit/Mesa-3D/issues/197 Signed-off-by: Axel Davy <davyaxel0@gmail.com>	2018-09-25 22:05:24 +02:00
Axel Davy	dcfde02bb0	st/nine: Avoid redundant SetCursorPos calls For some applications SetCursorPosition is called when a cursor event is received. Our SetCursorPosition was always calling wine SetCursorPos which would trigger a cursor event. The infinite loop is avoided by not calling SetCursorPos when the position hasn't changed. Found thanks to wine tests. Fixes irresponsive GUI for some applications. Fixes: https://github.com/iXit/Mesa-3D/issues/173 Signed-off-by: Axel Davy <davyaxel0@gmail.com> CC: <mesa-stable@lists.freedesktop.org>	2018-09-25 22:05:24 +02:00
Axel Davy	112c770597	st/nine: Init cursor position at device creation This is only useful for software cursor, but at least now we won't start it at (0, 0). Signed-off-by: Axel Davy <davyaxel0@gmail.com>	2018-09-25 22:05:24 +02:00
Axel Davy	62ea55ec8b	st/nine: Initialize manually cursor structure Initialize manually the cursor structure fields for more clarity on its content. Signed-off-by: Axel Davy <davyaxel0@gmail.com>	2018-09-25 22:05:24 +02:00
Axel Davy	110950318c	st/nine: Check if format is DS before retrieving flags d3d9_get_pipe_depth_format_bindings assumes the input format is a depth stencil format. Previously the user could hit this function with an invalid format. Protect the last non protected call with a depth_stencil_format check. Another solution is to have d3d9_get_pipe_depth_format_bindings support non depth stencil format, but we don't want the user to create depth buffers with d3d formats that can't be one, it's better to check if the format can be depth buffer with d3d. Signed-off-by: Axel Davy <davyaxel0@gmail.com>	2018-09-25 22:05:24 +02:00
Axel Davy	af60fbc0a4	st/nine: Remove clamping when mul_zero_wins Tests show the clamping can be removed when mul_zero_wins is supported. Signed-off-by: Axel Davy <davyaxel0@gmail.com>	2018-09-25 22:05:24 +02:00
Axel Davy	a0afa80889	st/nine: Implement predicated instructions Most of the work was already there, just not implemented. Fixes: https://github.com/iXit/Mesa-3D/issues/318 Signed-off-by: Axel Davy <davyaxel0@gmail.com>	2018-09-25 22:05:24 +02:00
Axel Davy	e7e82bcdc9	st/nine: Fix aliased read in ff Fix aliasing of colorarg_b4 with colorarg_b5. Fixes: https://github.com/iXit/Mesa-3D/issues/302 Signed-off-by: Axel Davy <davyaxel0@gmail.com>	2018-09-25 22:05:24 +02:00
Axel Davy	9fc6aa1bbe	st/nine: Fix ff assignment with aliasing "tex_stage[s][D3DTSS_COLORARG0] >> 4" could be a two bit number, thus colorarg_b4 was incorrectly set. Signed-off-by: Axel Davy <davyaxel0@gmail.com>	2018-09-25 22:05:24 +02:00
Axel Davy	8c35fb0280	st/nine: Clarify some ff assignments colorarg0, etc are 3 bits wide. Make the code more readable by adding an & 0x7 to further indicate we only remember the first 3 bits only. The 4th bit is always 0, and colorarg_b4, colorarg_b5, etc are used to store the 5th and 6th bits. Signed-off-by: Axel Davy <davyaxel0@gmail.com>	2018-09-25 22:05:24 +02:00
Axel Davy	59aaeeb730	st/nine: Print transform matrices in debug This is useful to see the matrices content in the log to debug. Signed-off-by: Axel Davy <davyaxel0@gmail.com>	2018-09-25 22:05:24 +02:00
Axel Davy	d9da0a1f6d	st/nine: Add ff key hash to help debug This is very useful to find in the log the ff shader shource of a given call. Signed-off-by: Axel Davy <davyaxel0@gmail.com>	2018-09-25 22:05:24 +02:00
Axel Davy	fcbb00a502	st/nine: Avoid RefToBind calls in ff When using csmt, ff shader creation happens on the csmt thread. Creating the shaders, then calling RefToBind causes the device ref to be increased then decreased. However the device dtor assumes than no work pending on the csmt thread could increase the device ref, leading to hang. The issue is avoided by creating the shaders with a bind count directly. Fixes: https://github.com/iXit/Mesa-3D/issues/295 Signed-off-by: Axel Davy <davyaxel0@gmail.com>	2018-09-25 22:05:24 +02:00
Axel Davy	e83b15cba0	st/nine: Add new helper for object creation with bind Add a new helper to create objects starting with a bind count instead of a ref count. Signed-off-by: Axel Davy <davyaxel0@gmail.com>	2018-09-25 22:05:24 +02:00
Axel Davy	fd86ce7c14	st/nine: Add parameter to start with bind Add a parameter to start new object with a bind instead of a refcount. Signed-off-by: Axel Davy <davyaxel0@gmail.com>	2018-09-25 22:05:24 +02:00
Axel Davy	a9bf82ecf4	st/nine: Use perspective correction for ps depth fog Emulate perspective interpolation of depth for programmable ps fog ff ps fog uses position z, or 1/w depending on the ff projection matrix set. This is according to public documents found describing the algorithm and tests we made. In the case of programmable ps, we used position's z, which was sufficient to pass wine tests (which test shaders don't set w). Issue https://github.com/iXit/Mesa-3D/issues/315 showed that this calculation was wrong. Using perspective interpolation on z, that is using z * 1/w seems to satisfy both this application and wine tests. Fixes: https://github.com/iXit/Mesa-3D/issues/315 Signed-off-by: Axel Davy <davyaxel0@gmail.com>	2018-09-25 22:05:24 +02:00
Axel Davy	7ee5e5e239	st/nine: Clamp RCP when 0inf!=0 Tests done on several devices of all 3 vendors and of different generations showed that there are several ways of handling infs and NaN for d3d9. Tests showed Intel on windows does always clamp RCP, RSQ and LOG (thus preventing inf/nan generation), for all shader versions (some vendor behaviours vary with shader versions). Doing this in nine avoids 0inf issues for drivers that can't generate 0*inf=0 (which is controled by TGSI's MUL_ZERO_WINS). For now clamp for all drivers. An ulterior optimization would be to avoid clamping for drivers with MUL_ZERO_WINS for the specific shader versions where NV or AMD don't clamp. LOG and RSQ being already clamped, this patch only clamps RCP. Fixes: https://github.com/iXit/Mesa-3D/issues/316 Signed-off-by: Axel Davy <davyaxel0@gmail.com> CC: <mesa-stable@lists.freedesktop.org>	2018-09-25 22:05:23 +02:00
Caio Marcelo de Oliveira Filho	3cf07361ac	intel/compiler: Export TCS passthrough creation Move create_passthrough_tcs() from i965 so can be used in other contexts. Acked-by: Jason Ekstrand <jason@jlekstrand.net>	2018-09-25 09:16:31 -07:00
Gert Wollny	47a6f98e15	mesa/st: In the precense of integer buffers enable per buffer blending Since blending will be disabled later for integer formats we have to consider that in the case of a mixed set of integer/non-integer format buffers blending must be handled on a per buffer basis. Fixes on r600: dEQP-GLES31.functional.draw_buffers_indexed.random. max_required_draw_buffers.13 Fixes: `8fb966688b` st/mesa: Disable blending for integer formats. Signed-off-by: Gert Wollny <gert.wollny@collabora.com> Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2018-09-25 15:54:38 +02:00
Eric Engestrom	f5b41f9121	gallivm: ensure string is null-terminated instead of assert()ing Signed-off-by: Eric Engestrom <eric.engestrom@intel.com> Reviewed-by: Dylan Baker <dylan@pnwbakers.com>	2018-09-25 11:39:30 +01:00
Topi Pohjolainen	1cc17fb731	intel/compiler/icl: Use barrier id bits 24:30 instead of 24:27,31 Fixes gpu hangs with Carchase and Manhattan. Reviewed-by: Anuj Phogat <anuj.phogat@gmail.com> Signed-off-by: Topi Pohjolainen <topi.pohjolainen@intel.com>	2018-09-25 09:59:59 +03:00
Andres Rodriguez	ec1fcf92ae	radv: only emit ZPASS_DONE for timestamp queries on gfx queues A ZPASS_DONE packet doesn't make sense for the compute queue. It will result in a gpu hang. This change resolves a gpu hang for SteamVR+Vega. Cc: mesa-stable@lists.freedesktop.org Fixes: `1f616a840e` "radv: emit a dummy ..." Signed-off-by: Andres Rodriguez <andresx7@gmail.com> Reviewed-by: Dave Airlie <airlied@redhat.com>	2018-09-25 02:30:34 -04:00
Timothy Arceri	72e4287e8f	radv: make use of nir_lower_load_const_to_scalar() This allows NIR to CSE more operations. LLVM does this also so the impact is limited, however doing this in NIR allows other opts to make progress. For example in radeonsi more loops are unrolled in Civilization Beyond Earth. The actual pipeline-db stats are not overwhelming but even in the negatively affected shaders the NIR is clearly better. It just happens that the code shuffling and in some cases calls to max rather than a flt result in the final output from LLVM not giving as good numbers. However this is an incremental opt that further passes build off so the change should be made IMO. Totals from affected shaders: SGPRS: 20192 -> 20184 (-0.04 %) VGPRS: 19516 -> 19524 (0.04 %) Spilled SGPRs: 437 -> 444 (1.60 %) Spilled VGPRs: 0 -> 0 (0.00 %) Private memory VGPRs: 0 -> 0 (0.00 %) Scratch size: 0 -> 0 (0.00 %) dwords per thread Code Size: 1527444 -> 1522276 (-0.34 %) bytes LDS: 6 -> 6 (0.00 %) blocks Max Waves: 1018 -> 1016 (-0.20 %) Wait states: 0 -> 0 (0.00 %) Reviewed-by: Bas Nieuwenhuizen <bas@basnieuwenhuizen.nl>	2018-09-25 09:31:22 +10:00
Eric Engestrom	f2519e3493	vulkan/wsi/display: wsi_display_select_crtc() doesn' need to modify the connector Signed-off-by: Eric Engestrom <eric.engestrom@intel.com>	2018-09-24 17:38:11 +01:00
Eric Engestrom	bde3102c0d	vulkan/wsi/display: check if wsi_swapchain_init() succeeded Fixes: `da997ebec9` "vulkan: Add KHR_display extension using DRM [v10]" Cc: Keith Packard <keithp@keithp.com> Signed-off-by: Eric Engestrom <eric.engestrom@intel.com> Reviewed-by: Jason Ekstrand <jason@jlekstrand.net>	2018-09-24 17:37:43 +01:00
Leo Liu	3e7b5e5db2	radeon/uvd: use bitstream coded number for symbols of Huffman tables Signed-off-by: Leo Liu <leo.liu@amd.com> Fixes: 130d1f456(radeon/uvd: reconstruct MJPEG bitstream) Cc: "18.2" <mesa-stable@lists.freedesktop.org> Reviewed-by: Boyuan Zhang <boyuan.zhang@amd.com>	2018-09-24 09:12:49 -04:00
Rhys Perry	6ca1402c11	nv50/ir: fix link-time build failure Seems this fixes linking problems that occur in some situations. Signed-off-by: Rhys Perry <pendingchaos02@gmail.com> Reviewed-by: Ilia Mirkin <imirkin@alum.mit.edu>	2018-09-23 18:20:08 +01:00
Rhys Perry	b473fcc9a3	nvc0: fix bindless multisampled images on Maxwell+ NVC0_CB_AUX_BINDLESS_INFO isn't written to on Maxwell+ and it's too small anyway. With these changes, TXQ is used to determine the number of samples and the coordinate adjustment information looked up in a small array in the driver constant buffer. v2: rework to use TXQ and a small array instead of a larger array with an entry for each texture v3: get rid of the small array and calculate the adjustments in the shader Signed-off-by: Rhys Perry <pendingchaos02@gmail.com> Fixes: `c2ae9b4052` ('nvc0: implement multisampled images on Maxwell+') Reviewed-by: Ilia Mirkin <imirkin@alum.mit.edu>	2018-09-22 20:13:17 +01:00
Rhys Perry	f580a895b1	nvc0: warn about changing NVC0_CB_AUX_MP_INFO and NVC0_CB_AUX_DRAW_INFO Signed-off-by: Rhys Perry <pendingchaos02@gmail.com> Reviewed-by: Ilia Mirkin <imirkin@alum.mit.edu>	2018-09-22 16:50:39 +01:00
Rhys Perry	01fa76b707	nvc0: Update counter reading shaders to new NVC0_CB_AUX_MP_INFO Fixes: `66ca7e400b` ('nvc0: add support for programmable sample locations') Signed-off-by: Rhys Perry <pendingchaos02@gmail.com> Reviewed-by: Ilia Mirkin <imirkin@alum.mit.edu>	2018-09-22 16:50:22 +01:00
Eric Anholt	cd667edecc	vc4: Remove dead i == 0 code from the cos() implementation. The loop starts at 1.	2018-09-21 17:16:43 -07:00
Eric Anholt	10d5d2d527	vc4: Fix sin(0.0) and cos(0.0) accuracy to fix SDL rendering rotation. SDL has some shaders that compute sin(angle) and cos(angle) for a rotation matrix in the VS, and angle is usually 0.0. Our previous implementation had quite a bit of error around 0.0, causing single-pixel rotations at typical window sizes. SDL2 has changed as of August 28th (commit 12156:e5a666405750) to not need sin/cos in the VS, but we should still fix this for existing implementations or similar patterns that other programs may have. glsl-cos goes from 32 instructions to 36, but 9 uniforms to 7. glsl-sin goes from 32 instructions to 34, but 8 uniforms to 7. This seems like a fine impact to have for the bugfix. Cc: 18.1 18.2 <mesa-stable@lists.freedesktop.org> Fixes: https://github.com/anholt/mesa/issues/110	2018-09-21 17:16:43 -07:00
Anuj Phogat	a0baedb638	intel/icl: Fix URB size for different SKUs Different ICL SKUs have different URB sizes. Signed-off-by: Anuj Phogat <anuj.phogat@gmail.com> Reviewed-by: Kenneth Graunke <kenneth@whitecape.org>	2018-09-21 14:40:04 -07:00
Anuj Phogat	fa1ff71a0f	i965/icl: Set Enabled Texel Offset Precision Fix bit h/w specification requires this bit to be always set. V2: Fix bit mask (Chris Wilson) Suggested-by: Kenneth Graunke <kenneth@whitecape.org> Signed-off-by: Anuj Phogat <anuj.phogat@gmail.com> Reviewed-by: Kenneth Graunke <kenneth@whitecape.org>	2018-09-21 14:40:04 -07:00
Anuj Phogat	5eb173304b	anv/icl: Set Enabled Texel Offset Precision Fix bit h/w specification requires this bit to be always set. Suggested-by: Kenneth Graunke <kenneth@whitecape.org> Signed-off-by: Anuj Phogat <anuj.phogat@gmail.com> Reviewed-by: Kenneth Graunke <kenneth@whitecape.org>	2018-09-21 14:40:04 -07:00
Marek Olšák	f0cd7dbcd7	glsl_to_tgsi: invert gl_SamplePosition.y for the default framebuffer Fixes dEQP-GLES31.functional.shaders.sample_variables.sample_pos.correctness.default_framebuffer with --deqp-gl-config-name=rgba8888d24s8ms4 Cc: 18.1 18.2 <mesa-stable@lists.freedesktop.org>	2018-09-21 13:39:00 -04:00
Caio Marcelo de Oliveira Filho	b29ec31854	util: Add macro to get number of elements in dynarray Reviewed-by: Christian Gmeiner <christian.gmeiner@gmail.com>	2018-09-21 10:12:51 -07:00
Dylan Baker	5dcb77e491	meson: Don't compile pipe loader with dri support when not using dri Corrects building glx as gallium-xlib without any dri targets. v2: - fix ugly formatting Fixes: `66c94b9313` ("meson: build gallium winsys for dri, null, and wrapper") Reviewed-by: Eric Engestrom <eric.engestrom@intel.com>	2018-09-21 10:03:15 -07:00
Samuel Pitoiset	fe3f13cc5a	radv: use the resolve compute path if dest uses multiple layers The hardware path doesn't support resolving layers, for both source and destination images. This fixes a reflection issue when MSAA is enabled which affects GTA V and probably DIRT3. CC: <mesa-stable@lists.freedesktop.org> Bugzilla: https://bugs.freedesktop.org/show_bug.cgi?id=107786 Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Tested-by: Gregor Münch <gr.muench_at_gmail.com> Reviewed-by: Bas Nieuwenhuizen <bas@basnieuwenhuizen.nl>	2018-09-21 16:35:59 +02:00
Jason Ekstrand	ab80889e92	anv,radv: Implement vkAcquireNextImage2 This was added as part of 1.1 but it's very hard to track exactly what extension added it. In any case, we should implement it. Cc: mesa-stable@lists.freedesktop.org Reviewed-by: Dave Airlie <Airlied@redhat.com>	2018-09-21 07:02:35 -05:00
Samuel Pitoiset	674fcfaecc	radv: only enable shaderInt16 on GFX9+ and LLVM7+ The throughput is similar to 32-bit integers on GFX8 and AMDVLK does not expose 16-bit integers on pre Vega as well. On GFX9+, only LLVM 7+ has support. This fixes a bunch of CTS crashes on GFX9/LLVM 6. Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Bas Nieuwenhuizen <bas@basnieuwenhuizen.nl>	2018-09-21 10:56:17 +02:00
Bas Nieuwenhuizen	0a77e70d10	radv: Fix driver UUID SHA1 init. Was missing the init, found by Emil. Fixes: `d17443a459` "radv: Use build ID if available for cache UUID." CC: <mesa-stable@lists.freedesktop.org> Reviewed-by: Eric Engestrom <eric.engestrom@intel.com> Reviewed-by: Samuel Pitoiset <samuel.pitoiset@gmail.com>	2018-09-20 23:38:38 +02:00
Charmaine Lee	64731e7c5e	svga: fix uninitialized fields in DefineDepthStencilView/DefineStreamOutput This patch fixes uninitialized fields in DefineDepthStencilView and DefineStreamOutput commands that are not relevant in SM4 device. Reviewed-by: Brian Paul <brianp@vmware.com>	2018-09-20 13:20:10 -06:00
Brian Paul	7f4e6f4c97	r300g: add PIPE_SHADER_CAP_SCALAR_ISA switch case to silence warning Reviewed-by: Mathias Fröhlich <Mathias.Froehlich@web.de>	2018-09-20 13:20:10 -06:00
Brian Paul	198c50f487	st/mesa: silenced unhanded enum warning in st_glsl_to_tgsi.cpp Add ir_intrinsic_begin_fragment_shader_ordering switch case to silence warning Reviewed-by: Mathias Fröhlich <Mathias.Froehlich@web.de>	2018-09-20 13:20:10 -06:00
Brian Paul	35ea66a68e	mesa: use GLsizeiptrARB, GLintptrARB in bufferobj.c The function pointer declarations in dd.h for the BufferData() and BufferSubData() use the ARB-suffixed datatypes. This patch changes the buffer_data_fallback() and buffer_sub_data_fallback() functions to use those datatypes too. This fixes a build warning when building 32-bit libraries. Evidently, GLsizeiptrARB and GLsizeiptr are defined differently in that situation. All all implementations of these driver hooks use the ARB-suffixed types. Reviewed-by: Mathias Fröhlich <Mathias.Froehlich@web.de>	2018-09-20 13:20:10 -06:00
Neha Bhende	708d34d41a	svga: Enable Opengl 3.3 compatibility profile With this patch, svga driver will start advertising OpenGL 3.3 compatibility profile. Tested with some mesa demos, piglit and glretrace. Reviewed-by: Brian Paul <brianp@vmware.com> Reviewed-by: Charmaine Lee <charmainel@vmware.com>	2018-09-20 13:20:10 -06:00
Neha Bhende	ede805dd19	svga: Apply texcoord scale factors only if there is sampler view We need to convert unnormalized texcoords to normalized texcoords when we are sampling from texture. We don't need this conversion if there is no sampler view. Tested with piglit, glretrace Fixes vmware bug 2101970 Reviewed-by: Brian Paul <brianp@vmware.com> Reviewed-by: Charmaine Lee <charmainel@vmware.com>	2018-09-20 13:20:10 -06:00
Charmaine Lee	1dcf377a76	svga: fix texture array layer index in transfer map In gallium, the layer index of a texture array to be mapped is specified in the z component, whereas in svga device, the index is specified in a separate argument. Currently in svga_texture_transfer_map(), we explicitly modify the z value in the base transfer map to 0 so the layer offset will not be applied twice, but this causes problem when state tracker later refers to the base transfer map and expects the slice index to be specified in z (commit `463b0ea1f6`). To fix the problem, this patch makes a local copy of the box in svga_transfer and modifies the z value in this copy instead. Fixes spec@khr_texture_compression-astc piglit test crashes. Fixes regression in the dma path with commit 1fdd3dd94a. Tested with mtt glretrace, piglit on Windows VM and Linux VM. Reviewed-by: Brian Paul <brianp@vmware.com>	2018-09-20 13:20:10 -06:00
Dylan Baker	18a6e426f3	Revert "utils/u_math: break dependency on gallium/utils" This reverts commit `0abce6d770`. Which broke the windows build.	2018-09-20 10:36:33 -07:00
Caio Marcelo de Oliveira Filho	2567ad28bb	i965: remove outdated comment about TCS passthrough Since commit `75881bed9e` "i965: Rework the TCS passthrough shader to use NIR." the created nir_shader is not dummy, and it is compiled by the backend like the others. Reviewed-by: Jason Ekstrand <jason@jlekstrand.net>	2018-09-20 09:58:55 -07:00
Dylan Baker	0abce6d770	utils/u_math: break dependency on gallium/utils Currently u_math needs gallium utils for cpu detection. Most of what u_math uses out of u_cpu_detection is duplicated in src/mesa/x86 (surprise!), so I've just reworked it as much as possible to use the x86/common_x86_features.h macros instead of the gallium ones. The mesa implementation is a header only approach, with no external dependencies. There is one small function that was copied over, as promoting u_cpu_detection is itself a fairly hefty undertaking, as it depends on u_debug, and this fixes the bug for now. bugzilla: https://bugs.freedesktop.org/show_bug.cgi?id=107870 Tested-by: Vinson Lee <vlee@freedesktop.org>	2018-09-20 05:52:23 -07:00
Emil Velikov	b8b3517a49	egl/android: rework device probing Unlike the other platforms, here we aim do guess if the device that we somewhat arbitrarily picked, is supported or not. In particular: when a vendor is _not_ requested we loop through all devices, picking the first one which can create a DRI screen. When a vendor is requested - we use that and do _not_ fall-back to any other device. The former seems a bit fiddly, but considering EGL_EXT_explicit_device and EGL_MESA_query_renderer are MIA, this is the best we can do for the moment. With those (proposed) extensions userspace will be able to create a separate EGL display for each device, query device details and make the conscious decision which one to use. v2: - update droid_open_device_drm_gralloc() - set the dri2_dpy->fd before using it - return a EGLBoolean for droid_{probe,open}_device* - do not warn on droid_load_driver failure (Tomasz) - plug mem leak on dri2_create_screen failure (Tomasz) - fixup function name typo (Tomasz, Rob) v3: - add forward declaration for droid_load_driver() Fixes the HAVE_DRM_GRALLOC build (Mauro) - split dup() assignment and check in separate lines (Tomasz, Eric) - make droid_load_driver() static (Tomasz) - drop unused prop_set variable (Tomasz) v4: - rebase - fwd declarationi should be for droid_probe_device() Cc: Robert Foss <robert.foss@collabora.com> Cc: Tomasz Figa <tfiga@chromium.org> Cc: Mauro Rossi <issor.oruam@gmail.com> Signed-off-by: Emil Velikov <emil.velikov@collabora.com> Reviewed-by: Tomasz Figa <tfiga@chromium.org> Tested-by: Tomasz Figa <tfiga@chromium.org> Tested-by: Tapani Pälli <tapani.palli@intel.com>	2018-09-20 10:15:38 +01:00
Danylo Piliaiev	18be7403a1	glsl: Add an assert when cloning ir_dereference_record with invalid field Signed-off-by: Danylo Piliaiev <danylo.piliaiev@globallogic.com> Reviewed-by: Timothy Arceri <tarceri@itsqueeze.com>	2018-09-20 08:30:11 +10:00
Danylo Piliaiev	6f3c7374b1	glsl: Avoid propagating incompatible type of initializer do_assignment validated assigment but when rhs type was not compatible it proceeded without issues and returned error_emitted = false. On the other hand process_initializer expected do_assignment to always return compatible type and never fail. As a result when variable was initialized with incompatible type the type of variable changed to the incompatible one. This manifested in unnecessary error messages and in one case in crash. Example GLSL: vec4 tmp = vec2(0.0); tmp.z -= 1.0; Past error messages: initializer of type vec2 cannot be assigned to variable of type vec4 invalid swizzle / mask `z' type mismatch operands to arithmetic operators must be numeric After this patch: initializer of type vec2 cannot be assigned to variable of type vec4 In the other case when we initialize variable with incompatible struct, accessing variable's field leaded to a crash. Example: uniform struct {float field;} data; ... vec4 tmp = data; tmp.x -= 1.0; After the patch there is only error line without a crash: initializer of type #anon_struct cannot be assigned to variable of type vec4 Signed-off-by: Danylo Piliaiev <danylo.piliaiev@globallogic.com> Reviewed-by: Timothy Arceri <tarceri@itsqueeze.com> Bugzilla: https://bugs.freedesktop.org/show_bug.cgi?id=107547	2018-09-20 08:30:11 +10:00
Michal Srb	194bf0a2e0	st/dri: don't set queryDmaBufFormats/queryDmaBufModifiers if the driver does not implement it This is equivalent to commit `a65db0ad1c`, but for dri_kms_init_screen. Without this gbm_dri_is_format_supported always returns false. Bugzilla: https://bugs.freedesktop.org/show_bug.cgi?id=104926 Fixes: `e14fe41e0b` ("st/dri: implement createImageFromRenderbuffer(2)") Reviewed-by: Emil Velikov <emil.velikov@collabora.com> Reviewed-by: Adam Jackson <ajax@redhat.com> Tested-by: Adam Williamson <adamwill@fedoraproject.org>	2018-09-19 15:20:04 -04:00
Jason Ekstrand	c811af767e	anv/so_memcpy: Don't consider src/dst_offset when computing block size The only thing that matters is the size since we never specify any offsets in terms of blocks. Reviewed-by: Lionel Landwerlin <lionel.g.landwerlin@intel.com>	2018-09-19 09:38:04 -05:00
Jakob Bornecrantz	09171705d5	Revert "mesa: only update framebuffer-state for clears" This reverts commit `fb86365148`.	2018-09-19 15:21:26 +01:00
Samuel Pitoiset	121f226471	radv: use a 64-bit unsigned integer when allocating a descriptor pool pool->size is a 64-bit unsigned integer too. Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Bas Nieuwenhuizen <bas@basnieuwenhuizen.nl>	2018-09-19 13:36:12 +02:00
Samuel Pitoiset	35656823b9	radv: enable VK_SUBGROUP_FEATURE_ARITHMETIC_BIT All CTS pass on Polaris/Vega with LLVM 6, 7 and master, so I think it's safe to enable the feature. Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Bas Nieuwenhuizen <bas@basnieuwenhuizen.nl>	2018-09-19 13:36:10 +02:00
Samuel Pitoiset	febdc13a6c	radv: do not support blitting surfaces with depth and stencil Fixes: dEQP-VK.api.copy_and_blit.core.blit_image.all_formats.depth_stencil.d32_sfloat_s8_uint_d32_sfloat_s8_uint.optimal_optimal_nearest And all friends that try to blit a surface with different depth and stencil formats. Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Bas Nieuwenhuizen <bas@basnieuwenhuizen.nl>	2018-09-19 13:36:07 +02:00
Erik Faye-Lund	fb86365148	mesa: only update framebuffer-state for clears If we update the program-state etc, we risk compiling needless shaders, which can cost quite a bit of performance. Signed-off-by: Erik Faye-Lund <erik.faye-lund@collabora.com> Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2018-09-19 11:52:53 +02:00
Juan A. Suarez Romero	0c82e3603e	nir: add initializer data to fix MSVC compile error CC: Jason Ekstrand <jason@jlekstrand.net> Fixes: 82799a5d1b8 ("nir: Add a small pass to rematerialize derefs per-block") Reviewed-by: Samuel Iglesias Gonsálvez <siglesias@igalia.com>	2018-09-19 11:46:44 +02:00
Jason Ekstrand	976046a8d8	nir: Add some asserts that we don't put derefs in phis The lcssa and phis_to_regs passes are used by various NIR optimizations that modify the CFG. Putting a couple of asserts will help ensure that we don't accidentally put derefs in phis as part of an optimization pass. Reviewed-by: Iago Toral Quiroga <itoral@igalia.com>	2018-09-19 02:00:49 -05:00
Jason Ekstrand	864c780566	nir/opt_if: Re-materialize derefs in use blocks before peeling loops Reviewed-by: Iago Toral Quiroga <itoral@igalia.com> Bugzilla: https://bugs.freedesktop.org/show_bug.cgi?id=107879 Cc: "18.2" <mesa-stable@lists.freedesktop.org>	2018-09-19 02:00:49 -05:00
Jason Ekstrand	0796c3934e	nir/loop_unroll: Re-materialize derefs in use blocks before unrolling When we're about to re-arrange a bunch of blocks, it's a good idea to make sure that we don't have deref uses crossing block boundaries. Otherwise we may end up with a deref going through a phi and that would be bad. Reviewed-by: Iago Toral Quiroga <itoral@igalia.com> Cc: "18.2" <mesa-stable@lists.freedesktop.org>	2018-09-19 01:59:40 -05:00
Jason Ekstrand	7d1d1208c2	nir: Add a small pass to rematerialize derefs per-block This pass re-materializes deref instructions on a per-block basis to ensure that every use of a deref occurs in the same block as the instruction which uses it. Reviewed-by: Iago Toral Quiroga <itoral@igalia.com> Cc: "18.2" <mesa-stable@lists.freedesktop.org>	2018-09-19 01:59:40 -05:00
Bas Nieuwenhuizen	95bb7d82ca	Revert "radv: fix descriptor pool allocation size" This reverts commit `90819abb56`. This logic was wrong, the original code is correct. The direct impact is that we allocate up to approximately a squared amount of memory compared to what we should allocate. Acked-by: Samuel Pitoiset <samuel.pitoiset@gmail.com>	2018-09-18 22:51:42 +02:00
Samuel Pitoiset	c9dbe52f84	radv: implement VK_EXT_conservative_rasterization Only supported by GFX9+. The conservativeraster Sascha demo seems to work as expected. Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Bas Nieuwenhuizen <bas@basnieuwenhuizen.nl>	2018-09-18 13:28:01 +02:00
Samuel Pitoiset	450a325858	radv: do not re-create the sampler for every blits in CmdBlitImage() Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Bas Nieuwenhuizen <bas@basnieuwenhuizen.nl>	2018-09-18 13:27:59 +02:00
Samuel Pitoiset	3871dd7a92	radv: allow to force anisotropy via RADV_TEX_ANISO Ported from RadeonSI. Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Bas Nieuwenhuizen <bas@basnieuwenhuizen.nl>	2018-09-18 13:27:58 +02:00
Timothy Arceri	b54a2311a9	mesa: enable EXT_framebuffer_object in core profile Since user defined names are not allowed in core profile we remove the allow_user_names bool and just check if we have a core profile like all other buffer/texture object handling code does. This extension is required by "Wolfenstein: The Old Blood" and is exposed in core in the Nvidia binary driver. Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2018-09-18 19:58:24 +10:00
Timothy Arceri	02843ed768	mesa: move legacy dri config option texture_depth Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2018-09-18 19:43:05 +10:00
Timothy Arceri	f958ea6eff	mesa: move legacy dri config option fthrottle_mode Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2018-09-18 19:43:05 +10:00
Timothy Arceri	4b1a81ef9d	mesa: move legacy dri config option def_max_anisotropy Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2018-09-18 19:43:05 +10:00
Timothy Arceri	6164d59bcc	mesa: move legacy dri config option no_neg_lod_bias Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2018-09-18 19:43:05 +10:00
Timothy Arceri	6d1890fa07	mesa: move legacy dri config option round_mode Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2018-09-18 19:43:05 +10:00
Timothy Arceri	3a1d09fd55	mesa: remove unused dri option float_depth This seems to have only been used by DRI1 drivers which were removed with `e4344161bd`. Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2018-09-18 19:43:05 +10:00
Timothy Arceri	91e76ce493	mesa: move legacy dri config option dither_mode Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2018-09-18 19:43:05 +10:00
Timothy Arceri	2d7dc9591d	mesa: move legacy dri config option color_reduction Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2018-09-18 19:43:05 +10:00
Timothy Arceri	408d41a413	mesa: move legacy TCL dri config options Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2018-09-18 19:43:05 +10:00
Timothy Arceri	024abd3534	util: use force_compat_profile for Wolfenstein The Old Blood This game is looking for some odd extension after creating a core context such as ARB_vertex_program and EXT_framebuffer_object. Rather then enabling these in core this forces the game to use compat. This allows the game to run and seems to work without issues. All other id tech games/engines use a compat profile. Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2018-09-18 19:34:54 +10:00
Timothy Arceri	64ec50d52f	mesa/st: add force_compat_profile option to driconfig Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2018-09-18 19:34:54 +10:00
Timothy Arceri	7a992fcfa0	Revert "radeonsi: avoid syncing the driver thread in si_fence_finish" This reverts commit `bc65dcab3b`. This was manually reverted. Reverting stops the menu hanging in some id tech games such as RAGE and Wolfenstein The New Order. Reviewed-by: Marek Olšák <marek.olsak@amd.com> Bugzilla: https://bugs.freedesktop.org/show_bug.cgi?id=107891	2018-09-18 19:21:32 +10:00
Eric Anholt	4e1af6808c	v3d: Switch from FLUSH_ALL_STATE to FLUSH for ending our bin CLs. The HW for FLUSH_ALL_STATE isn't validated, since the closed driver only uses FLUSH. Now that we don't have any new state at the end of our bin CLs, follow their lead.	2018-09-17 16:35:45 -07:00
Eric Anholt	0b8007b523	v3d: Stop clearing the OQ state at the end of the job. Ever since we added OQ support, we've been clearing OQ state at the start of the job anyway. We're intentionally breaking old-and-new-driver-mix systems, because we need to stop using the unvalidated FLUSH_ALL_STATE.	2018-09-17 16:35:45 -07:00
Eric Anholt	350cb79045	v3d: Always emit a TF disable at the start of drawing on V3D 4.x. The HW's FLUSH_ALL_STATE is not validated, so we probably shouldn't use it, meaning that we need to reset state at the start. By doing this, we also make ourselves more resilient to another client leaving the TF state enabled at the end of their batch (as we now do, ourselves). However, we still need to emit a single TF disable at the end of the frame, for SWVC5-718.	2018-09-17 16:35:45 -07:00
Dylan Baker	7f08bcb73f	build: Don't overlink gallium xlib target Currently gallium's xlib target will fail to link due to multiple definitions of all the symbols in libmesautil, this only shows up in autotools, and not in meson due to differences in the way that meson and autotools handle linking static archives into static archives. Autotools uses -Wl,--whole-archive implicitly, meson requires this behavior to be opted-into. The solution is just to remove libmesautils from the libgl-xlib target, since it will get all of those symbols form libmesagallium. I've dropped the link from meson as well, it doesn't seem to hurt anything and should make linking just a little faster. Fixes: `8396043f30` ("Replace uses of _mesa_bitcount with util_bitcount") bugzilla: https://bugs.freedesktop.org/show_bug.cgi?id=107923 Tested-by: Brian Paul <brianp@vmware.com> Tested-by: Vinson Lee <vlee@freedesktop.org> Cc: Sergii Romantsov<sergii.romantsov@globallogic.com>	2018-09-17 13:21:01 -07:00
Dylan Baker	3acc18fcf7	move pthread_setaffinity_np check to the build system Rather than trying to encode all of the rules in a header, lets just put them in the build system where they belong. This fixes the build on FreeBSD, which does have pthraed_setaffinity_np, but it's in a pthread_np.h, not behind _GNU_SOURCE. FreeBSD also implements cpu_set slightly differently, so additional changes would be required to get it working right there anyway. v2: - fix #define in autotools Fixes: `9f1bbbdbbd` ("util: try to fix the Android and MacOS build") Cc: Emil Velikov <emil.velikov@collabora.com> Reviewed-by: Marek Olšák <marek.olsak@amd.com> Reviewed-by: Eric Engestrom <eric.engestrom@intel.com>	2018-09-17 13:16:46 -07:00
Fritz Koenig	60d0c0d062	mesa: FramebufferParameteri parameter checking Missing break; causes parameter checking to never pass GL_FRAMEBUFFER_FLIP_Y_MESA parameters. Fixes: `318c265160` ("mesa: GL_MESA_framebuffer_flip_y extension [v4]") Reviewed-by: Eric Engestrom <eric.engestrom@intel.com> Reviewed-by: Brian Paul <brianp@vmware.com>	2018-09-17 11:48:00 -07:00
Fritz Koenig	ba6cc32cf9	mesa: Additional FlipY applications Instances where direction was determined based on winsys or user fbo and should be determined based on FlipY. Key STATE_FB_WPOS_Y_TRANSFORM for of FlipY instead of _mesa_is_user_fbo. This corrects gl_FragCoord usage when applying GL_MESA_framebuffer_flip_y. Fixes: `ab05dd183c` ("i965: implement GL_MESA_framebuffer_flip_y [v3]") Reviewed-by: Brian Paul <brianp@vmware.com>	2018-09-17 11:48:00 -07:00
Bas Nieuwenhuizen	d17443a459	radv: Use build ID if available for cache UUID. To get an useful UUID for systems that have a non-useful mtime for the binaries. I started using SHA1 to ensure we get reasonable mixing in the various possibilities and the various build id lengths. CC: <mesa-stable@lists.freedesktop.org> Reviewed-by: Timothy Arceri <tarceri@itsqueeze.com>	2018-09-17 20:19:52 +02:00
Samuel Pitoiset	08103c5f65	radv: enable shaderInt16 capability Not sure if this is all wired up. CTS does pass and the Tangrams demo works fine on Vega. There are corruption issues on Polaris but not sure if that related to 16-bit support. Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Bas Nieuwenhuizen <bas@basnieuwenhuizen.nl>	2018-09-17 15:18:39 +02:00
Samuel Pitoiset	cd76ce0078	ac: add 16-bit support to ac_build_bitfield_reverse() Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Bas Nieuwenhuizen <bas@basnieuwenhuizen.nl>	2018-09-17 15:18:37 +02:00
Samuel Pitoiset	fc398f4d67	ac: add 16-bit support to ac_build_bit_count() Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Bas Nieuwenhuizen <bas@basnieuwenhuizen.nl>	2018-09-17 15:18:34 +02:00
Samuel Pitoiset	94dd08eb7c	ac: add 16-bit support to ac_find_lsb() Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Bas Nieuwenhuizen <bas@basnieuwenhuizen.nl>	2018-09-17 15:18:32 +02:00
Samuel Pitoiset	5a6c8ca3e8	ac: add 16-bit support to ac_build_umsb() Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Bas Nieuwenhuizen <bas@basnieuwenhuizen.nl>	2018-09-17 15:18:30 +02:00
Samuel Pitoiset	3e7f3e2cd1	ac: add 16-bit support to ac_build_isign() Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Bas Nieuwenhuizen <bas@basnieuwenhuizen.nl>	2018-09-17 15:18:28 +02:00
Samuel Pitoiset	cfd6314cfe	ac: add 16-bit constant values for zero and one Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Bas Nieuwenhuizen <bas@basnieuwenhuizen.nl>	2018-09-17 15:18:26 +02:00
Samuel Pitoiset	074e29183c	ac: add ac_build_bifield_reverse() helper Are we missing 64-bit support? Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Bas Nieuwenhuizen <bas@basnieuwenhuizen.nl>	2018-09-17 15:18:23 +02:00
Samuel Pitoiset	371c35e5bb	ac: add ac_build_bit_count() helper Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Bas Nieuwenhuizen <bas@basnieuwenhuizen.nl>	2018-09-17 15:18:20 +02:00
Samuel Pitoiset	aec9151464	radv: fix use of unreachable() in the meta blit path Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Timothy Arceri <tarceri@itsqueeze.com>	2018-09-17 11:29:25 +02:00
Samuel Pitoiset	6521d4a659	Revert "radv: Optimize rebinding the same descriptor set." This introduces random GPU hangs on Vega, at least. This reverts commit `02a43edf18`.	2018-09-17 11:20:57 +02:00
Samuel Pitoiset	90819abb56	radv: fix descriptor pool allocation size The size has to be multiplied by the number of sets. This gets rid of the OUT_OF_POOL_KHR error and fixes a crash with the Tangrams demo. CC: 18.1 18.2 <mesa-stable@lists.freedesktop.org> Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Bas Nieuwenhuizen <bas@basnieuwenhuizen.nl>	2018-09-17 10:18:01 +02:00
Jason Ekstrand	67094e11e9	anv/query: Add an emit_srm helper Reviewed-by: Lionel Landwerlin <lionel.g.landwerlin@intel.com>	2018-09-17 02:57:21 -05:00
Jason Ekstrand	40149441b8	anv: Add a mi_memset and use it for zeroing queries Reviewed-by: Lionel Landwerlin <lionel.g.landwerlin@intel.com>	2018-09-17 02:57:21 -05:00
Jason Ekstrand	b11e9b5ffe	anv/query: Use anv_address everywhere Instead of passing around BOs and offsets, use addresses which are anv's GPU equivalent of pointers. Reviewed-by: Lionel Landwerlin <lionel.g.landwerlin@intel.com>	2018-09-17 02:57:21 -05:00
Jason Ekstrand	07e214f1ce	anv/query: Write both dwords in emit_zero_queries Each query slot is a uint64_t and we were only zeroing half of it. Fixes: `7ec6e4e689` "anv/query: implement multiview interactions" Reviewed-by: Lionel Landwerlin <lionel.g.landwerlin@intel.com>	2018-09-17 02:57:21 -05:00
Jason Ekstrand	c0420a62c9	anv/query: Increment an index while writing results Instead of computing an index at the end which we hope maps to the number of things written, just count the number of things as we go. Reviewed-by: Lionel Landwerlin <lionel.g.landwerlin@intel.com>	2018-09-17 02:57:21 -05:00
Ian Romanick	df9dbc03d3	i965/fs: Don't propagate conditional modifiers from integer compares to adds No shader-db changes on any Intel platform... which probably explains why no bugs have been bisected to this problem since it landed in Mesa 18.1. :( The commit mentioned below is in 18.2, so 18.1 would need a slightly different fix (due to code refactoring). Signed-off-by: Ian Romanick <ian.d.romanick@intel.com> Fixes: `77f269bb56` "i965/fs: Refactor propagation of conditional modifiers from compares to adds" Reviewed-by: Alejandro Piñeiro <apinheiro@igalia.com> (reviewed the original patch) Cc: Matt Turner <mattst88@gmail.com> (reviewed the original patch)	2018-09-17 00:38:22 -07:00
Bas Nieuwenhuizen	0dd8189f15	radv: Only allow 16 user SGPRs for compute on GFX9+. Apparently for compute there are only 16 instead of the 32 for the graphics path. Fixes dEQP-VK.binding_model.descriptorset_random.sets16.noarray.ubolimitlow.sbolimitlow.imglimitlow.noiub.comp.0 CC: <mesa-stable@lists.freedesktop.org> Reviewed-by: Samuel Pitoiset <samuel.pitoiset@gmail.com>	2018-09-16 12:50:58 +02:00
Bas Nieuwenhuizen	d97c892584	radv: Set the user SGPR MSB for Vega. Otherwise using 32 user SGPRs would be broken. CC: <mesa-stable@lists.freedesktop.org> Reviewed-by: Samuel Pitoiset <samuel.pitoiset@gmail.com>	2018-09-16 12:50:58 +02:00
Bas Nieuwenhuizen	02a43edf18	radv: Optimize rebinding the same descriptor set. This makes it cheaper to just change the dynamic offsets with the same descriptor sets. Suggested-by: Philip Rebohle <philip.rebohle@tu-dortmund.de> Reviewed-by: Samuel Pitoiset <samuel.pitoiset@gmail.com>	2018-09-16 12:50:19 +02:00
Gert Wollny	14976817f4	r600/sb: use safe math optimizations when TGSI contains precise operations Fixes: dEQP-GLES3.functional.shaders.invariance.highp.common_subexpression_3 dEQP-GLES3.functional.shaders.invariance.mediump.common_subexpression_3 dEQP-GLES3.functional.shaders.invariance.lowp.common_subexpression_3 Signed-off-by: Gert Wollny <gw.fossdev@gmail.com> Reviewed-by: Roland Scheidegger <sroland@vmware.com>	2018-09-15 20:44:53 +02:00
Mauro Rossi	cc3b99bb48	android: broadcom/cle: export the broadcom top level path headers Fixes the following building error in vc4 build: In file included from external/mesa/src/gallium/drivers/vc4/kernel/vc4_render_cl.c:34: In file included from external/mesa/src/gallium/drivers/vc4/kernel/vc4_drv.h:27: In file included from external/mesa/src/gallium/drivers/vc4/vc4_simulator_validate.h:34: In file included from external/mesa/src/gallium/drivers/vc4/vc4_context.h:39: In file included from external/mesa/src/gallium/drivers/vc4/vc4_cl.h:56: gen/STATIC_LIBRARIES/libmesa_broadcom_genxml_intermediates/broadcom/cle/v3d_packet_v21_pack.h:12:10: fatal error: 'cle/v3d_packet_helpers.h' file not found ^~~~~~~~~~~~~~~~~~~~~~~~~~ 1 error generated. Fixes: `5b102160ae` ("broadcom/genxml: Introduce a V3D packet/struct decoder.") Cc: "18.2" <mesa-stable@lists.freedesktop.org> Acked-by: Eric Anholt <eric@anholt.net> Reviewed-by: Emil Velikov <emil.velikov@collabora.com> Signed-off-by: Mauro Rossi <issor.oruam@gmail.com>	2018-09-15 09:14:46 +02:00
Mauro Rossi	9158e0bd82	android: broadcom/cle: add gallium include path Fixes the following building error: In file included from external/mesa/src/broadcom/cle/v3d_decoder.c:38: In file included from external/mesa/src/broadcom/cle/v3d_packet_helpers.h:29: external/mesa/src/gallium/auxiliary/util/u_math.h:42:10: fatal error: 'pipe/p_compiler.h' file not found ^~~~~~~~~~~~~~~~~~~ 1 error generated. Fixes: `5b102160ae` ("broadcom/genxml: Introduce a V3D packet/struct decoder.") Cc: "18.2" <mesa-stable@lists.freedesktop.org> Acked-by: Eric Anholt <eric@anholt.net> Reviewed-by: Emil Velikov <emil.velikov@collabora.com> Signed-off-by: Mauro Rossi <issor.oruam@gmail.com>	2018-09-15 09:14:42 +02:00
Mauro Rossi	3341429d74	android: broadcom/genxml: fix collision with intel/genxml header-gen macro Fixes the following building error, happening when building both intel and broadcom: Gen Header: libmesa_broadcom_genxml_32 <= v3d_packet_v21_pack.h FAILED: gen/STATIC_LIBRARIES/libmesa_broadcom_genxml_intermediates/broadcom/cle/v3d_packet_v21_pack.h /bin/bash -c "python external/mesa/src/broadcom/cle/gen_pack_header.py \ external/mesa/src/broadcom/cle/v3d_packet_v21.xml \ > gen/STATIC_LIBRARIES/libmesa_broadcom_genxml_intermediates/broadcom/cle/v3d_packet_v21_pack.h" Traceback (most recent call last): File "external/mesa/src/broadcom/cle/gen_pack_header.py", line 626, in <module> p = Parser(sys.argv[2]) IndexError: list index out of range header-gen macro is already defined by Intel genxml building rules and the existing header-gen does not have the $(PRIVATE_VER) argument, infact the bash command line logged in the building error is missing exactly $(PRIVATE_VER) argument Renaming the macro as pack-header-gen in src/broadcom/Android.genxml.mk solves the building error, another possible way is to keep the gen rules commands expanded and not use the macros. Fixes: `7f80a9ff13` ("vc4: Introduce XML-based packet header generation like Intel's.") Cc: "18.2" <mesa-stable@lists.freedesktop.org> Acked-by: Eric Anholt <eric@anholt.net> Reviewed-by: Emil Velikov <emil.velikov@collabora.com> Signed-off-by: Mauro Rossi <issor.oruam@gmail.com>	2018-09-15 09:14:33 +02:00
Caio Marcelo de Oliveira Filho	f9d25f630c	anv/memcpy: fix build after starting to use addresses The offsets now come from the anv_address, these references were not updated and using the old variable. Fixes: `e1ab834557` "anv/memcpy: Use addresses instead of bo+offset" Tested-by: Clayton Craft <clayton.a.craft@intel.com>	2018-09-14 21:45:50 -07:00
Jason Ekstrand	d6a73824bd	anv/cmd_buffer: Take an address in emit_lrm Reviewed-by: Eric Engestrom <eric.engestrom@intel.com>	2018-09-14 22:12:11 -05:00
Jason Ekstrand	e1ab834557	anv/memcpy: Use addresses instead of bo+offset Reviewed-by: Eric Engestrom <eric.engestrom@intel.com>	2018-09-14 22:12:11 -05:00
Jason Ekstrand	90b46f6c17	anv/so_memcpy: Use the correct SO_BUFFER size on gen8+ This shouldn't matter as we'll never write OOB anyway but we may as well get it right. It's supposed to be in dwords - 1. Reviewed-by: Nanley Chery <nanley.g.chery@intel.com>	2018-09-14 22:12:11 -05:00
Timothy Arceri	e29f0ede75	ac: fix get_image_coords() for radeonsi Because this was setting image to true we would end up calling si_load_image_desc() when we sould be calling si_load_sampler_desc(). This fixes an assert() in Deus Ex: MD Reviewed-by: Marek Olšák <marek.olsak@amd.com>	2018-09-15 12:23:32 +10:00
Marek Olšák	914bd3014f	gallium/util: don't let child processes inherit our thread affinity v2: corrected the comment	2018-09-14 21:15:39 -04:00
Marek Olšák	7d41a7593a	gallium/util: start with a random L3 cache index for AMD Zen	2018-09-14 21:05:37 -04:00
Josh Pieper	936e0dcd61	st/mesa: Validate the result of pipe_transfer_map in make_texture (v2) When using Freecad, I was getting intermittent segfaults inside of mesa. I traced it down to this path in st_cb_drawpixels.c where the result of pipe_transfer_map wasn't being checked. In my case, it was returning NULL because nouveau_bo_new returned ENOENT. I'm by no means a mesa developer, but this patch solves the problem for me and seems reasonable enough. v2: Marek - also unmap the PBO and release the texture, and call the make_texture function sooner for less cleanup Cc: 18.1 18.2 <mesa-stable@lists.freedesktop.org>	2018-09-14 21:05:37 -04:00
Samuel Pitoiset	c79aad30ae	radv: emit the initial config only once in the preambles It shouldn't be needed to emit the initial graphics or compute state when beginning a new command buffer. Emitting them in the preamble should be enough and this will reduce IB sizes. Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Bas Nieuwenhuizen <bas@basnieuwenhuizen.nl>	2018-09-14 10:59:52 +02:00
Samuel Pitoiset	9de062ef20	radv: fix setting global locations for indirect descriptors Indirect descriptors only need one entry, we don't have to emit a location for every descriptors. Fixes GPU hangs with new CTS: dEQP-VK.binding_model.descriptorset_random.* CC: 18.2 <mesa-stable@lists.freedesktop.org> Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Bas Nieuwenhuizen <bas@basnieuwenhuizen.nl>	2018-09-14 10:59:52 +02:00
Samuel Pitoiset	748f4cce18	radv: fix flushing indirect descriptors Let say, we first bind a graphics pipeline that needs indirect descriptors sets. The userdata pointers will be emitted at draw time. Then if we bind a compute pipeline that doesn't need any indirect descriptors, the driver will re-emit them for all grpahics stages. To avoid this to happen, just check the bind point type. CC: 18.2 <mesa-stable@lists.freedesktop.org> Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Bas Nieuwenhuizen <bas@basnieuwenhuizen.nl>	2018-09-14 10:59:52 +02:00

... 2 3 4 5 6 ...

96905 Commits