KonstantinSeurer/mesa

Commit Graph

Author	SHA1	Message	Date
Marek Olšák	212c612a63	tgsi/ureg: allow any register file in address operands Reviewed-by: Brian Paul <brianp@vmware.com> Reviewed-by: Nicolai Hähnle <nicolai.haehnle@amd.com>	2017-10-06 02:56:11 +02:00
Marek Olšák	41b85158ab	gallium: add PIPE_CAP_TGSI_ANY_REG_AS_ADDRESS Reviewed-by: Brian Paul <brianp@vmware.com> Reviewed-by: Nicolai Hähnle <nicolai.haehnle@amd.com>	2017-10-06 02:56:11 +02:00
Marek Olšák	cb686a340f	tgsi/scan: scan address operands (v2) v2: set swizzled usage mask Reviewed-by: Nicolai Hähnle <nicolai.haehnle@amd.com>	2017-10-06 02:56:11 +02:00
Marek Olšák	37714c6df2	tgsi/scan: set correct usage mask for tex offsets in scan_src_operand Reviewed-by: Nicolai Hähnle <nicolai.haehnle@amd.com>	2017-10-06 02:56:11 +02:00
Marek Olšák	5cc779197c	tgsi/scan: take advantage of already swizzled usage mask in scan_src_operand It has always been a usage mask after swizzling. Reviewed-by: Nicolai Hähnle <nicolai.haehnle@amd.com>	2017-10-06 02:56:11 +02:00
Marek Olšák	ea85b76519	tgsi/scan: set non-valid src_index for tex offsets in scan_src_operand tex offsets are not "Src" operands. Reviewed-by: Nicolai Hähnle <nicolai.haehnle@amd.com>	2017-10-06 02:56:11 +02:00
Marek Olšák	be3ab867bd	tgsi: implement tgsi_util_get_inst_usage_mask properly All opcodes are handled. Reviewed-by: Nicolai Hähnle <nicolai.haehnle@amd.com>	2017-10-06 02:56:11 +02:00
Marek Olšák	bb8abc10bf	tgsi: add docs for some existing pack opcodes Reviewed-by: Nicolai Hähnle <nicolai.haehnle@amd.com>	2017-10-06 02:56:11 +02:00
Bas Nieuwenhuizen	4ffb9890ef	radv: Enable VK_KHR_maintenance2 extension. Reviewed-by: Dave Airlie <airlied@redhat.com>	2017-10-06 01:41:29 +02:00
Bas Nieuwenhuizen	0c90ca7d37	radv: Make tess winding order a bit more intuitive. Reviewed-by: Dave Airlie <airlied@redhat.com>	2017-10-06 01:41:29 +02:00
Bas Nieuwenhuizen	c62afd094d	radv: Allow setting the domain origin in tess. Reviewed-by: Dave Airlie <airlied@redhat.com>	2017-10-06 01:41:29 +02:00
Bas Nieuwenhuizen	ca21634632	radv: Disable usage checks in metadata for images with extended usage data. The app can extend the usage, so knowing that the usage is limitied does not help us here. Reviewed-by: Dave Airlie <airlied@redhat.com>	2017-10-06 01:41:29 +02:00
Bas Nieuwenhuizen	f800d91019	radv: Implement querying the point clipping behavior. Reviewed-by: Dave Airlie <airlied@redhat.com>	2017-10-06 01:41:29 +02:00
Daniel Stone	bbe2082e7d	broadcom: Fix out-of-tree build include path Reviewed-by: Eric Anholt <eric@anholt.net> Fixes: `5b102160ae` ("broadcom/genxml: Introduce a V3D packet/struct decoder.")	2017-10-05 15:03:11 -07:00
Bas Nieuwenhuizen	908a25ecb0	meson: generate builddir/src/amd/vulkan/dev_icd.json Reviewed-by: Dylan Baker <dylan@pnwbakers.com>	2017-10-05 23:46:21 +02:00
Kenneth Graunke	18bdf73556	mesa: Use a 565 format for GL_RGB and GL_UNSIGNED_SHORT_5_6_5 textures. Found while trying to optimize an application. Not observed to help performance on i965, but should at least reduce the memory usage of such textures a bit. Reviewed-by: Eric Anholt <eric@anholt.net> Tested-by: Eero Tamminen <eero.t.tamminen@intel.com>	2017-10-05 14:30:47 -07:00
Jason Ekstrand	7463d50580	intel/compiler: Don't propagate cmod into integer multiplies No shader-db change on Sky Lake. Reviewed-by: Matt Turner <mattst88@gmail.com> Cc: mesa-stable@lists.freedesktop.org	2017-10-05 11:54:49 -07:00
Jason Ekstrand	b91ecee04a	intel/compiler: Don't cmod propagate into a saturated operation Shader-db results on Sky Lake: total instructions in shared programs: 12954445 -> 12955125 (0.01%) instructions in affected programs: 141862 -> 142542 (0.48%) helped: 0 HURT: 626 Reviewed-by: Matt Turner <mattst88@gmail.com> Cc: mesa-stable@lists.freedesktop.org	2017-10-05 11:54:49 -07:00
Derek Foreman	17d78ece36	broadcom/vc4: Don't advertise tiled dmabuf modifiers if we can't use them If the DRM_VC4_GET_TILING ioctl isn't present then we can't tell if a dmabuf bo is tiled or linear, so will always assume it's linear. By not advertising tiled formats in this situation we ensure the assumption is correct. This fixes a bug where most attempts to render a gl wayland client under weston will result in a client side abort. Signed-off-by: Derek Foreman <derekf@osg.samsung.com> Reviewed-by: Eric Anholt <eric@anholt.net> Acked-by: Daniel Stone <daniels@collabora.com> (on irc)	2017-10-05 11:26:14 -07:00
Adam Jackson	b174a1ae72	egl: Simplify the "driver" interface "Driver" isn't a great word for what this layer is, it's effectively a build-time choice about what OS you're targeting. Despite that both of the extant backends totally ignore the display argument, the old code would only set up the backend relative to a display. That causes problems! One problem is it means eglGetProcAddress can generate X or Wayland protocol when it tries to connect to a default display so it can call into the backend, which is, you know, completely bonkers. Any other EGL API that doesn't reference a display, like EGL_EXT_device_query, would have the same issue. Fortunately this is a problem that can be solved with the delete key. Reviewed-by: Eric Anholt <eric@anholt.net> Signed-off-by: Adam Jackson <ajax@redhat.com>	2017-10-05 13:43:34 -04:00
Thomas Hellstrom	15e208c4cc	loader/dri3: Don't accidently free buffer holding new back content Avoid freeing buffers holding new back content (with GLX_SWAP_COPY_OML and GLX_SWAP_EXCHANGE_OML) Prevously that would have resulted in back buffer content becoming incorrect after a swap, although I haven't managed to trigger such a situation yet. Signed-off-by: Thomas Hellstrom <thellstrom@vmware.com> Reviewed-by: Sinclair Yeh <syeh@vmware.com>	2017-10-05 09:17:12 +02:00
Thomas Hellstrom	1b8e0bed69	loader/dri3: Avoid resizing existing buffers in dri3_find_back_alloc Resize only in loader_dri3_get_buffers(), where the dri driver has a chance to immediately update the viewport. Signed-off-by: Thomas Hellstrom <thellstrom@vmware.com> Reviewed-by: Sinclair Yeh <syeh@vmware.com>	2017-10-05 09:17:12 +02:00
Thomas Hellstrom	622f5e1d9b	loader/dri3: Use local blits and local buffers when resizing When a drawable is resized, and we fill the resized buffers, with data from the old buffers, use a local blit if there is a local buffer (back or fake front), and we have local blitting capability. Signed-off-by: Thomas Hellstrom <thellstrom@vmware.com> Reviewed-by: Sinclair Yeh <syeh@vmware.com>	2017-10-05 09:17:12 +02:00
Ben Crocker	1359af930e	gallivm/ppc64le: allow environmental control of Altivec code generation In check_os_altivec_support(), allow control of Altivec (first PPC vector instruction set) code generation via a new environmental control, GALLIVM_ALTIVEC, which is expected to take on a value of 1 or 0. The default is to enable Altivec code generation. This environmental control of Altivec code generation is initially available only #ifdef DEBUG. Cc: "17.2" <mesa-stable@lists.freedesktop.org> Signed-off-by: Ben Crocker <bcrocker@redhat.com> Acked-by: Roland Scheidegger <sroland@vmware.com>	2017-10-05 02:14:14 +02:00
Ben Crocker	e93f056a4e	gallivm/ppc64le: adjust VSX code generation control. In lp_build_create_jit_compiler_for_module(), advance the minimum version of LLVM for VSX code generation to 4.0; this is the minimum revision at which several known VSX code generation bugs are fixed: https://llvm.org/bugs/show_bug.cgi?id=25503 (fixed in 3.8.1) https://llvm.org/bugs/show_bug.cgi?id=26775 (fixed in 3.8.1) https://llvm.org/bugs/show_bug.cgi?id=33531 (fixed in 4.0) An llc performance bug introduced in LLVM 4.0, https://llvm.org/bugs/show_bug.cgi?id=34647 is still pending as of LLVM 5.0, but only has a pronounced effect on one of the Piglit tests: ext_transform_feedback-max-varyings. All changes tested via Piglit. Cc: "17.2" <mesa-stable@lists.freedesktop.org> Signed-off-by: Ben Crocker <bcrocker@redhat.com> Acked-by: Roland Scheidegger <sroland@vmware.com>	2017-10-05 02:13:47 +02:00
Ben Crocker	5c75f0c8bb	gallivm: allow additional llc options In init_native_targets, allow the passing of additional options to the LLC compiler via new GALLIVM_LLC_OPTIONS environmental control. This option is available only #ifdef DEBUG, initially. At top, add #include <llvm-c/Support.h> for LLVMParseCommandLineOptions() declaration. v2: Fix compile error with old llvm versions (sroland) Cc: "17.2" <mesa-stable@lists.freedesktop.org> Signed-off-by: Ben Crocker <bcrocker@redhat.com> Acked-by: Nicolai Hähnle <nicolai.haehnle@amd.com> Reviewed-by: Roland Scheidegger <sroland@vmware.com>	2017-10-05 02:06:46 +02:00
Ben Crocker	3a9feb4db8	gallivm: fix typo in debug_printf message In gallivm_compile_module, fix a typo in the debug_printf("Invoke as \"llc ..." message. Cc: "17.2" <mesa-stable@lists.freedesktop.org> Signed-off-by: Ben Crocker <bcrocker@redhat.com> Reviewed-by: Roland Scheidegger <sroland@vmware.com>	2017-10-05 01:48:37 +02:00
Samuel Pitoiset	8196a3c63e	radv: remove useless checks around radv_CmdBindPipeline() Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Bas Nieuwenhuizen <bas@basnieuwenhuizen.nl>	2017-10-04 23:18:51 +02:00
Samuel Pitoiset	b53c207659	radv: check that pipeline is different before binding it We only need to dirty the descriptors when the pipeline is a new one, because user SGPRs can be potentially different. Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Bas Nieuwenhuizen <bas@basnieuwenhuizen.nl>	2017-10-04 23:18:48 +02:00
Matt Turner	2572c2771d	i965: Validate "Special Requirements for Handling Double Precision Data Types" I did not implement: CNL's restriction on 64-bit int + align16, because I don't think we'll ever use this combination regardless of hardware generation. The restriction on immediate DF -> F conversions, because there's no reason to ever generate that, and I don't even know how DF -> F conversions are supposed to work in Align16 since (1) the dst stride must be 1, but (2) the dst stride would have to be 2 for src and dst strides to be aligned.	2017-10-04 14:08:54 -07:00
Matt Turner	98298c7e3d	i965: Fix and enable forgotten validation test I seem to have forgotten I still had work to do.	2017-10-04 14:08:54 -07:00
Matt Turner	122ef3799d	i965: Only insert error message if not already present Some restrictions require something like strides to match between src and dest. For multi-source instructions, I'd rather encapsulate the logic for not inserting already present errors in ERROR_IF than open-coding it multiple places.	2017-10-04 14:08:54 -07:00
Matt Turner	5e76cf153c	i965: Avoid validation error when src1 is not present There can be no violation of the restriction that source offsets are aligned if there is only one source offset.	2017-10-04 14:08:54 -07:00
Matt Turner	cacc229ba0	i965: Remove validate_reg() Replaced by the assembly validator, and in fact gets in the way of writing tests for the assembly validator.	2017-10-04 14:08:54 -07:00
Matt Turner	678d88bcee	i965: Add and use STRIDE and WIDTH macros You'll notice there were bugs in some of the code being replaced. Reviewed-by: Iago Toral Quiroga <itoral@igalia.com>	2017-10-04 14:08:54 -07:00
Matt Turner	4c961a5e79	i965: Add parentheses around usage of macro arguments Otherwise I cannot use this macro in test_eu_validate.cpp Reviewed-by: Iago Toral Quiroga <itoral@igalia.com>	2017-10-04 14:08:54 -07:00
Matt Turner	1fcdb1cbea	i965: Add GLK, CFL, CNL to test_eu_validate.c	2017-10-04 14:08:54 -07:00
Matt Turner	d4c39e9cff	i965: Add Atom graphics names to parse_devid_override()	2017-10-04 14:08:54 -07:00
Matt Turner	6db5ec7deb	i965: Fix support for disassembling 64-bit integer immediates The type suffixes were wrong, and the 16 was missing the 0 prefix. Fixes: `92f787ff86` ("i965: Add support for disassembling 64-bit integer immediates") Reviewed-by: Iago Toral Quiroga <itoral@igalia.com>	2017-10-04 14:08:54 -07:00
Matt Turner	7e88f93469	i965/fs: Rewrite fsign64 to skip the float -> double conversion ... without the float -> double conversion. Low power parts have additional restrictions when it comes to operating on 64-bit types, and the instruction used to do the conversion violates one of them: specifically, the restriction that "Source and Destination horizontal stride must be aligned to the same qword". Previously we generated a float and then converted, but we can avoid the conversion by using the same extract-the-sign-bit + or-in-1.0 algorithm by directly operating on the high four bytes of each double-precision component in the result. In SIMD8 and SIMD16 this cuts one instruction from the implementation, and more importantly that instruction is the one which violated the regioning restriction. Along the way I removed some comments that I did not think helped, and some code about double comparisons which does not seem to be necessary today. This prevents validation failures caught by the new EU validation code added in later patches. Reviewed-by: Iago Toral Quiroga <itoral@igalia.com>	2017-10-04 14:08:54 -07:00
Matt Turner	b541945c20	i965/fs: Unpack count argument to 64-bit shift ops on Atom 64-bit operations on Atom parts have additional restrictions over their big-core counterparts (validated by later patches). Specifically, the restriction that "Source and Destination horizontal stride must be aligned to the same qword" is violated by most shift operations since NIR uses a 32-bit value as the shift count argument, and this causes instructions like shl(8) g19<1>Q g5<4,4,1>Q g23<4,4,1>UD where src1 has a 32-bit stride, but the dest and src0 have a 64-bit stride. This caused ~4 pixels in the ARB_shader_ballot piglit test fs-readInvocation-uint.shader_test to be incorrect. Unfortunately no ARB_gpu_shader_int64 test hit this case because they operate on uniforms, and their scalar regions are an exception to the restriction. We work around this by effectively unpacking the shift count, so that we can read it with a 64-bit stride in the shift instruction. Unfortunately the unpack (a MOV with a dst stride of 2) is a partial write, and cannot be copy-propagated or CSE'd. Bugzilla: https://bugs.freedesktop.org/101984	2017-10-04 14:08:54 -07:00
Matt Turner	2082c32950	i965/fs: Don't apply POW/FDIV workaround on Gen10+ The documentation says it applies only to Gens 8 and 9. Reviewed-by: Scott D Phillips <scott.d.phillips@intel.com>	2017-10-04 14:08:37 -07:00
Matt Turner	d407935327	i965: Fix src0 vs src1 typo A typo caused us to copy src0's reg file to src1 rather than reading src1's as intended. This caused us to fail to compact instructions like mov(8) g4<1>D 0D { align1 1Q }; because src1 was set to immediate rather than architecture file. Fixing this reenables compaction (after the precompact() pass changes the data types): mov(8) g4<1>UD 0x00000000UD { align1 1Q compacted }; Fixes: `1cb0a7941b` ("i965: Switch to using the logical register types") Reviewed-by: Kenneth Graunke <kenneth@whitecape.org>	2017-10-04 14:08:24 -07:00
Dave Airlie	ad3d98da9f	radv: enable tc compatible htile for d32s8 also. This enables tc compatible htile for stencil surfaces as well. This gives a 3-5fps boost on Mad Max on high@4k. It also depends on Bas's tc-compat htile patch. Reviewed-by: Bas Nieuwenhuizen <bas@basnieuwenhuizen.nl> Reviewed-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Signed-off-by: Dave Airlie <airlied@redhat.com>	2017-10-04 21:02:23 +01:00
Samuel Pitoiset	844ae722c4	radv: dump SPIRV when a GPU hang is detected Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Bas Nieuwenhuizen <bas@basnieuwenhuizen.nl>	2017-10-04 19:37:08 +02:00
Samuel Pitoiset	a2a350a3be	radv: dump NIR when a GPU hang is detected This looks a bit ugly to me, but the existing codepath is not terribly elegant as well. Signed-off-by: Samuel Pitoiset <samuel.pitoiset@gmail.com> Reviewed-by: Bas Nieuwenhuizen <bas@basnieuwenhuizen.nl>	2017-10-04 19:37:08 +02:00
Marek Olšák	94d800bfa3	ac: silence a warning	2017-10-04 17:00:05 +02:00
Daniel Stone	b65d6dafd6	egl/wayland: Don't use dmabuf with no modifiers The dmabuf interface requires a valid modifier to be sent. If we don't explicitly get a modifier from the driver, we can't know what to send; it must be inferred from legacy side-channels (or assumed to linear, if none exists). If we have no modifier, then we can only have a single-plane format anyway, so fall back to the old wl_drm buffer import path. Fixes: `a65db0ad1c` ("st/dri: don't expose modifiers in EGL if the driver doesn't implement them") Fixes: `02cc359372` ("egl/wayland: Use linux-dmabuf interface for buffers") Signed-off-by: Daniel Stone <daniels@collabora.com> Reviewed-by: Emil Velikov <emil.velikov@collabora.com> Reported-by: Andy Furniss <adf.lists@gmail.com> Cc: Marek Olšák <marek.olsak@amd.com>	2017-10-04 15:17:46 +01:00
Daniel Stone	6273d2f269	egl/wayland: Check queryImage return for wl_buffer When creating a wl_buffer from a DRIImage, we extract all the DRIImage information via queryImage. Check whether or not it actually succeeds, either bailing out if the query was critical, or providing sensible fallbacks for information which was not available in older DRIImage versions. Fixes: `a65db0ad1c` ("st/dri: don't expose modifiers in EGL if the driver doesn't implement them") Fixes: `02cc359372` ("egl/wayland: Use linux-dmabuf interface for buffers") Signed-off-by: Daniel Stone <daniels@collabora.com> Reviewed-by: Emil Velikov <emil.velikov@collabora.com> Reported-by: Andy Furniss <adf.lists@gmail.com> Cc: Marek Olšák <marek.olsak@amd.com>	2017-10-04 15:17:46 +01:00
Eric Engestrom	d246aa3a0d	travis: move include path from $CC to $CFLAGS Signed-off-by: Eric Engestrom <eric.engestrom@imgtec.com> Reviewed-by: Emil Velikov <emil.velikov@collabora.com>	2017-10-04 15:02:37 +01:00

1 2 3 4 5 ...

96324 Commits All Branches Search

96324 Commits

All Branches