AlexIndustrial/mesa

Author	SHA1	Message	Date
José Roberto de Souza	86c9aa6bfe	intel: Add and use intel_engines_class_to_string() Signed-off-by: José Roberto de Souza <jose.souza@intel.com> Reviewed-by: Kenneth Graunke <kenneth@whitecape.org> Reviewed-by: Lionel Landwerlin <lionel.g.landwerlin@intel.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18975>	2022-10-15 20:04:51 +00:00
José Roberto de Souza	772dfd60ad	intel: Convert i915 engine type to intel in tools/ common/ and ds/ This ones were left to be done after initial conversion. Signed-off-by: José Roberto de Souza <jose.souza@intel.com> Reviewed-by: Kenneth Graunke <kenneth@whitecape.org> Reviewed-by: Lionel Landwerlin <lionel.g.landwerlin@intel.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18975>	2022-10-15 20:04:51 +00:00
José Roberto de Souza	5269d91efc	intel: Convert missing i915 engine types to intel This convertions were missed due to bad rebased in my end, sorry. Fixes: `03b959286e` ("intel: Make engine related functions and types not i915 dependent") Signed-off-by: José Roberto de Souza <jose.souza@intel.com> Reviewed-by: Kenneth Graunke <kenneth@whitecape.org> Reviewed-by: Lionel Landwerlin <lionel.g.landwerlin@intel.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18975>	2022-10-15 20:04:51 +00:00
Alyssa Rosenzweig	ac2964dfbd	nir: Be smarter fusing ffma If there is a single use of fmul, and that single use is fadd, it makes sense to fuse ffma, as we already do. However, if there are multiple uses, fusing may impede code gen. Consider the source fragment: a = fmul(x, y) b = fadd(a, z) c = fmin(a, t) d = fmax(b, c) The fmul has two uses. The current ffma fusing is greedy and will produce the following "optimized" code. a = fmul(x, y) b = ffma(x, y, z) c = fmin(a, t) d = fmax(b, c) Actually, this code is worse! Instead of 1 fmul + 1 fadd, we now have 1 fmul + 1 ffma. In effect, two multiplies (and a fused add) instead of one multiply and an add. Depending on the ISA, that could impede scheduling or increase code size. It can also increase register pressure, extending the live range. It's tempting to gate on is_used_once, but that would hurt in cases where we really do fuse everything, e.g.: a = fmul(x, y) b = fadd(a, z) c = fadd(a, t) For ISAs that fuse ffma, we expect that 2 ffma is faster than 1 fmul + 2 fadd. So what we really want is to fuse ffma iff the fmul will get deleted. That occurs iff all uses of the fmul are fadd and will themselves get fused to ffma, leaving fmul to get dead code eliminated. That's easy to implement with a new NIR search helper, checking that all uses are fadd. shader-db results on Mali-G57 [open shader-db + subset of closed]: total instructions in shared programs: 179491 -> 178991 (-0.28%) instructions in affected programs: 36862 -> 36362 (-1.36%) helped: 190 HURT: 27 total cycles in shared programs: 10573.20 -> 10571.75 (-0.01%) cycles in affected programs: 72.02 -> 70.56 (-2.02%) helped: 28 HURT: 1 total fma in shared programs: 1590.47 -> 1582.61 (-0.49%) fma in affected programs: 319.95 -> 312.09 (-2.46%) helped: 194 HURT: 1 total cvt in shared programs: 812.98 -> 813.03 (<.01%) cvt in affected programs: 118.53 -> 118.58 (0.04%) helped: 65 HURT: 81 total quadwords in shared programs: 98968 -> 98840 (-0.13%) quadwords in affected programs: 2960 -> 2832 (-4.32%) helped: 20 HURT: 4 total threads in shared programs: 4693 -> 4697 (0.09%) threads in affected programs: 4 -> 8 (100.00%) helped: 4 HURT: 0 v2: Update trace checksums for virgl due to numerical differences. Signed-off-by: Alyssa Rosenzweig <alyssa@collabora.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18814>	2022-10-15 17:47:31 +00:00
Mike Blumenkrantz	07c654e08f	glthread: fix buffer allocation size with non-signed buffer offset path this needs to always add the start_offset to avoid creating buffers that are too small Reviewed-by: Pierre-Eric Pelloux-Prayer <pierre-eric.pelloux-prayer@amd.com> Reviewed-by: Marek Olšák <marek.olsak@amd.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18994>	2022-10-15 16:02:38 +00:00
Andri Yngvason	9fc4cb8067	gallium/vl: Add opaque rgb pixel formats Signed-off-by: Andri Yngvason <andri@yngvason.is> Reviewed-by: Simon Ser <contact@emersion.fr> Cc: mesa-stable Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18959>	2022-10-15 08:15:02 +00:00
Erik Faye-Lund	8255739a0a	mesa/main: remove driver-cap for ARB_point_sprite It's always supported, no need for checks here. Reviewed-By: Mike Blumenkrantz <michael.blumenkrantz@gmail.com> Begrudgingly-reviewed-by: Alyssa Rosenzweig <alyssa@collabora.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19049>	2022-10-15 02:57:18 +00:00
Erik Faye-Lund	310959d9fe	mesa/st: rip out point-sprite cap All current drivers reports supporting this cap, let's just assume it's always supported. It seems better to lower this in the drivers, like we already do for etnaviv, panfrost and zink... Reviewed-By: Mike Blumenkrantz <michael.blumenkrantz@gmail.com> Begrudgingly-reviewed-by: Alyssa Rosenzweig <alyssa@collabora.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19049>	2022-10-15 02:57:18 +00:00
Italo Nicola	b0d698c532	rusticl: correctly check global argument size As the spec that is quoted in the comment says, if the argument is a memory object, arg_size should be different than sizeof(cl_mem). The previous verification only worked if the underlying type has the same size as sizeof(cl_mem). Signed-off-by: Italo Nicola <italonicola@collabora.com> Reviewed-by: Karol Herbst <kherbst@redhat.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18985>	2022-10-15 02:23:04 +00:00
Italo Nicola	c935232822	rusticl: use 32-bit address format for 32-bit devices Signed-off-by: Italo Nicola <italonicola@collabora.com> Reviewed-by: Karol Herbst <kherbst@redhat.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18985>	2022-10-15 02:23:03 +00:00
Italo Nicola	66b3df3c15	clc: add 32-bit target Signed-off-by: Italo Nicola <italonicola@collabora.com> Reviewed-by: Karol Herbst <kherbst@redhat.com> Reviewed-by: Jesse Natalie <jenatali@microsoft.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18985>	2022-10-15 02:23:03 +00:00
Alyssa Rosenzweig	b3a69d1c31	panfrost/ci: Disable t720 jobs They're dead, Jim! Signed-off-by: Alyssa Rosenzweig <alyssa@collabora.com> Suggested-by: Daniel Stone <daniels@collabora.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19084>	2022-10-15 00:53:22 +00:00
Erik Faye-Lund	eccc5600c3	zink: use util_dynarray_clear We already have a helper for this, let's use it. Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19068>	2022-10-14 23:29:08 +00:00
Erik Faye-Lund	17d8ff3a39	zink: fixup dynarray-type This doesn't make a functional difference, because the size here ends up being the same; the size of a pointer. But let's use the right type for consistency. Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19068>	2022-10-14 23:29:08 +00:00
Erik Faye-Lund	510a34fbf3	zink: fix broken pool-alloc consolidation When appending the content of a util_dynarray to another util_dynarray, we need to copy the content to the end of the util_dynarray, not the beginning. As we've already resized the dynarray, We also shouldn't add to the size once more at the end, otherwise we'll end up with garbage. Fixes: `43dcdf3365` ("zink: rework/improve descriptor pool overflow handling on batch reset") Closes: https://gitlab.freedesktop.org/mesa/mesa/-/issues/7485 Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19068>	2022-10-14 23:29:08 +00:00
Lionel Landwerlin	b49b18f0b7	anv: reduce BT emissions & surface state writes with push descriptors Zink on Anv running Gfxbench gl_driver2 is significantly slower than Iris. The reason is simple, whereas Iris implements uniform updates using push constants and only has to emit 3DSTATE_CONSTANT_* packets, Zink uses push descriptors with a uniform buffer, which on our implementation use both push constants & binding tables. Anv ends up doing the following for each uniform update : - allocate 2 surface states : - one for the uniform buffer as the offset specify by zink - one for the descriptor set buffer - pack the 2 RENDER_SURFACE_STATE - re-emit binding tables - re-emit push constants Of all of those operations, only the last one ends up being useful in this benchmark because all the uniforms have been promoted to push constants. This change defers the 3 first operations at draw time and executes them only if the pipeline needs them. Vkoverhead before / after : descriptor_template_1ubo_push: 40670 / 85786 descriptor_template_12ubo_push: 4050 / 13820 descriptor_template_1combined_sampler_push, 34410 / 34043 descriptor_template_16combined_sampler_push, 2746 / 2711 descriptor_template_1sampled_image_push, 34765 / 34089 descriptor_template_16sampled_image_push, 2794 / 2649 descriptor_template_1texelbuffer_push, 108537 / 111342 descriptor_template_16texelbuffer_push, 20619 / 20166 descriptor_template_1ssbo_push, 41506 / 85976 descriptor_template_8ssbo_push, 6036 / 18703 descriptor_template_1image_push, 88932 / 89610 descriptor_template_16image_push, 20937 / 20959 descriptor_template_1imagebuffer_push, 108407 / 113240 descriptor_template_16imagebuffer_push, 32661 / 34651 Signed-off-by: Lionel Landwerlin <lionel.g.landwerlin@intel.com> Reviewed-by: Emma Anholt <emma@anholt.net> Reviewed-by: Tapani Pälli <tapani.palli@intel.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19050>	2022-10-14 23:03:16 +00:00
Lionel Landwerlin	ff91c5ca42	anv: add analysis for push descriptor uses and store it in shader cache We'll use this information to avoid : - binding table emission - allocation of surface states v2: Fix anv_nir_push_desc_ubo_fully_promoted() Signed-off-by: Lionel Landwerlin <lionel.g.landwerlin@intel.com> Reviewed-by: Emma Anholt <emma@anholt.net> Reviewed-by: Tapani Pälli <tapani.palli@intel.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19050>	2022-10-14 23:03:16 +00:00
Lionel Landwerlin	01e282f23f	anv: initialization pipeline layout to 0s Signed-off-by: Lionel Landwerlin <lionel.g.landwerlin@intel.com> Cc: mesa-stable Reviewed-by: Emma Anholt <emma@anholt.net> Reviewed-by: Tapani Pälli <tapani.palli@intel.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19050>	2022-10-14 23:03:16 +00:00
Lionel Landwerlin	8616f11a39	anv: track descriptor set layout flags To identify push descriptors. Signed-off-by: Lionel Landwerlin <lionel.g.landwerlin@intel.com> Reviewed-by: Emma Anholt <emma@anholt.net> Reviewed-by: Tapani Pälli <tapani.palli@intel.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19050>	2022-10-14 23:03:16 +00:00
Lionel Landwerlin	d7f1569307	anv: limit push constant reemission Signed-off-by: Lionel Landwerlin <lionel.g.landwerlin@intel.com> Reviewed-by: Emma Anholt <emma@anholt.net> Reviewed-by: Tapani Pälli <tapani.palli@intel.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19050>	2022-10-14 23:03:16 +00:00
Lionel Landwerlin	2db45f713a	isl: avoid gfx version switch cases on the hot path Some of the surface state packing functions are called from the hot path in Anv. We can use function pointers to avoid repeatedly going through switch/case. Signed-off-by: Lionel Landwerlin <lionel.g.landwerlin@intel.com> Reviewed-by: Emma Anholt <emma@anholt.net> Reviewed-by: Tapani Pälli <tapani.palli@intel.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19050>	2022-10-14 23:03:16 +00:00
Lionel Landwerlin	06d955ab21	anv: remove multiple push descriptors VUID-VkPipelineLayoutCreateInfo-pSetLayouts-00293 pSetLayouts must not contain more than one descriptor set layout that was created with VK_DESCRIPTOR_SET_LAYOUT_CREATE_PUSH_DESCRIPTOR_BIT_KHR set There is only one push descriptor set with all the descriptor sets, so no need to have an array. Signed-off-by: Lionel Landwerlin <lionel.g.landwerlin@intel.com> Reviewed-by: Emma Anholt <emma@anholt.net> Reviewed-by: Tapani Pälli <tapani.palli@intel.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19050>	2022-10-14 23:03:16 +00:00
Lionel Landwerlin	803f438d85	anv: optimize 3DSTATE_VF emission We can avoid reemitting this when the index buffer index type doesn't change. Also we don't need to update this when the pipeline changes as we do not pull any value from the pipeline. Instead rely on the dynamic state to tell if dyn->ia.primitive_restart_enable changed. Signed-off-by: Lionel Landwerlin <lionel.g.landwerlin@intel.com> Reviewed-by: Emma Anholt <emma@anholt.net> Reviewed-by: Tapani Pälli <tapani.palli@intel.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19050>	2022-10-14 23:03:16 +00:00
Lionel Landwerlin	126f5bc15a	anv: limit calls into cmd_buffer_flush_dynamic_state Avoids a bunch of checks if we can. Signed-off-by: Lionel Landwerlin <lionel.g.landwerlin@intel.com> Reviewed-by: Emma Anholt <emma@anholt.net> Reviewed-by: Tapani Pälli <tapani.palli@intel.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19050>	2022-10-14 23:03:16 +00:00
Lionel Landwerlin	54bc34f70a	anv: comment out the Gfx8/9 VB cache key workaround for newer Gens This code shows up a little on profiling on Gfx12 and since it's only a gfx8/9 workaround we might as well ifdef it out. Signed-off-by: Lionel Landwerlin <lionel.g.landwerlin@intel.com> Reviewed-by: Emma Anholt <emma@anholt.net> Reviewed-by: Tapani Pälli <tapani.palli@intel.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19050>	2022-10-14 23:03:16 +00:00
Lionel Landwerlin	f8136ea5b6	anv: remove unused code Signed-off-by: Lionel Landwerlin <lionel.g.landwerlin@intel.com> Reviewed-by: Emma Anholt <emma@anholt.net> Reviewed-by: Tapani Pälli <tapani.palli@intel.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19050>	2022-10-14 23:03:16 +00:00
Lionel Landwerlin	cea113c977	vulkan/runtime: don't lookup the pipeline disk cache if disabled When the Anv pipeline got migrated to the runtime, we gain/lost a bit of functionality which is that the disk cache is always read regardless of VK_ENABLE_PIPELINE_CACHE=0. This change brings the old behavior back. Signed-off-by: Lionel Landwerlin <lionel.g.landwerlin@intel.com> Fixes: `591da98779` ("vulkan: Add a common VkPipelineCache implementation") Reviewed-by: Emma Anholt <emma@anholt.net> Reviewed-by: Tapani Pälli <tapani.palli@intel.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19050>	2022-10-14 23:03:16 +00:00
Bas Nieuwenhuizen	6558ecf3eb	radv: Mark dEQP-VK.ray_query.misc.dynamic_indexing as crashing in CI. Gitlab: https://gitlab.freedesktop.org/mesa/mesa/-/issues/7493 Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19073>	2022-10-14 21:34:54 +00:00
Ryan Houdek	b516f59490	vulkan/wsi: Add dep_libudev to idep dependencies Otherwise users of `idep_vulkan_wsi` won't pull in the udev dependency, which will cause the linker to fail later on in compiling. The user of this dependency is lavapipe which would fail to link if this isn't provided. Fixes: `4885e63a6d` (vulkan/wsi: implement missing wsi_register_device_event) Reviewed-by: Dylan Baker <dylan@pnwbakers.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19037>	2022-10-14 21:10:29 +00:00
David Heidelberg	9cb251a0b0	ci/traces: Blender demo (Cube Diorama) flakes on Intel APL Acked-by: Karol Herbst <kherbst@redhat.com> Signed-off-by: David Heidelberg <david.heidelberg@collabora.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19067>	2022-10-14 18:44:30 +00:00
Gert Wollny	2e50bf19cd	nir: move fusing csel and comparisons to opt_late_algebraic With that simple comparisons are cleaned up properly. This helps with some tesselation shaders on r600. Shader-db stats R600/Cayman: -------------------------------------------------------------- total dw in shared programs: 1621806 -> 1620884 (-0.06%) dw in affected programs: 41650 -> 40728 (-2.21%) helped: 211 HURT: 4 helped stats (abs) min: 2 max: 26 x̄: 4.46 x̃: 4 helped stats (rel) min: 0.30% max: 9.68% x̄: 2.87% x̃: 2.52% HURT stats (abs) min: 2 max: 8 x̄: 5.00 x̃: 5 HURT stats (rel) min: 0.23% max: 1.67% x̄: 1.02% x̃: 1.09% 95% mean confidence interval for dw value: -4.81 -3.77 95% mean confidence interval for dw %-change: -3.03% -2.57% Dw are helped. total gprs in shared programs: 41192 -> 41182 (-0.02%) gprs in affected programs: 731 -> 721 (-1.37%) helped: 53 HURT: 45 helped stats (abs) min: 1 max: 3 x̄: 1.23 x̃: 1 helped stats (rel) min: 5.88% max: 40.00% x̄: 16.56% x̃: 14.29% HURT stats (abs) min: 1 max: 2 x̄: 1.22 x̃: 1 HURT stats (rel) min: 7.69% max: 40.00% x̄: 19.42% x̃: 20.00% 95% mean confidence interval for gprs value: -0.37 0.16 95% mean confidence interval for gprs %-change: -3.92% 3.85% Inconclusive result (value mean confidence interval includes 0). total alu_groups in shared programs: 203677 -> 203632 (-0.02%) alu_groups in affected programs: 2876 -> 2831 (-1.56%) helped: 68 HURT: 30 helped stats (abs) min: 1 max: 4 x̄: 1.46 x̃: 1 helped stats (rel) min: 0.84% max: 25.00% x̄: 7.48% x̃: 5.41% HURT stats (abs) min: 1 max: 6 x̄: 1.80 x̃: 1 HURT stats (rel) min: 1.98% max: 33.33% x̄: 10.09% x̃: 5.61% 95% mean confidence interval for alu_groups value: -0.81 -0.11 95% mean confidence interval for alu_groups %-change: -4.20% <.01% Alu_groups are helped. total loops in shared programs: 72 -> 72 (0.00%) loops in affected programs: 0 -> 0 helped: 0 HURT: 0 total cf in shared programs: 88230 -> 88233 (<.01%) cf in affected programs: 71 -> 74 (4.23%) helped: 1 HURT: 4 helped stats (abs) min: 1 max: 1 x̄: 1.00 x̃: 1 helped stats (rel) min: 33.33% max: 33.33% x̄: 33.33% x̃: 33.33% HURT stats (abs) min: 1 max: 1 x̄: 1.00 x̃: 1 HURT stats (rel) min: 1.89% max: 33.33% x̄: 17.14% x̃: 16.67% 95% mean confidence interval for cf value: -0.51 1.71 95% mean confidence interval for cf %-change: -24.20% 38.29% Inconclusive result (value mean confidence interval includes 0). total stack in shared programs: 3827 -> 3827 (0.00%) stack in affected programs: 0 -> 0 helped: 0 HURT: 0 LOST: 0 GAINED: 0 Total CPU time (seconds): 45.32 -> 41.69 (-8.01%) -------------------------------------------------------------- v2: Simplify replacement pattern (Rhys Perry) v3: fix ws (Alexander Orzechowski) v4: move the original lowering to opt_late_algebraic and drop cleanup code (Alyssa) v5: Add shader-sb stats (Alyssa) Signed-off-by: Gert Wollny <gert.wollny@collabora.com> Reviewed-by: Alyssa Rosenzweig <alyssa@collabora.com> Reviewed-by: Emma Anholt <emma@anholt.net> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18970>	2022-10-14 13:08:15 +00:00
Gert Wollny	aea311dbef	r600/sfn: run cleanup passes after late algebraic opt Signed-off-by: Gert Wollny <gert.wollny@collabora.com> Reviewed-by: Emma Anholt <emma@anholt.net> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18970>	2022-10-14 13:08:15 +00:00
Väinö Mäkelä	cfc6bdb760	hasvk: Correctly set NonPerspectiveBarycentricEnable on gfx7 The incorrect #else has existed since commit `bfd9942cdc`, but the issue was already present before that. NonPerspectiveBarycentricEnable must be enabled if non-perspective interpolation is used. Closes: https://gitlab.freedesktop.org/mesa/mesa/-/issues/7449 Reviewed-by: Emma Anholt <emma@anholt.net> Reviewed-by: Lionel Landwerlin <lionel.g.landwerlin@intel.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/19004>	2022-10-14 08:18:40 +00:00
Lionel Landwerlin	eec49374b0	nir: fix NIR_DEBUG=validate_ssa_dominance validate_ssa_def_dominance() asserts : validate_assert(state, !BITSET_TEST(state->ssa_defs_found, def->index)); Because the previous validation lefts bits set when it processed the IR. Signed-off-by: Lionel Landwerlin <lionel.g.landwerlin@intel.com> Reviewed-by: Rhys Perry <pendingchaos02@gmail.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18966>	2022-10-14 10:36:56 +03:00
Yonggang Luo	44ccaca41d	util/mesa/wide: Rename _SIMPLE_MTX_INITIALIZER_NP to SIMPLE_MTX_INITIALIZER Signed-off-by: Yonggang Luo <luoyonggang@gmail.com> Reviewed-by: Jesse Natalie <jenatali@microsoft.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18393>	2022-10-14 03:27:41 +00:00
Guilherme Gallo	be3c46964b	ci/bin: Remove whitespace from token files There was a security problem with some `gitlab_gql.py` scenarios because of `\r` and `\n` in the token file, which interrupted the requests for Gitlab endpoints. Stripping the token file after reading the file content solves the problem. Signed-off-by: Guilherme Gallo <guilherme.gallo@collabora.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18066>	2022-10-14 01:46:52 +00:00
Guilherme Gallo	d52d51b24d	ci/bin: Fix requirements.txt Add missing aiohttp and PyYAML packages Signed-off-by: Guilherme Gallo <guilherme.gallo@collabora.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18066>	2022-10-14 01:46:52 +00:00
Alyssa Rosenzweig	bb6c43027e	agx: Reserve live-in regs at the start of block ...Rather than reserving the union of the registers live-out of the predecessors. This avoids reserving registers that are killed along a control flow edge (where the predecessor has another successor that does use the register). glmark2 subset of shaderdb: total instructions in shared programs: 6442 -> 6440 (-0.03%) instructions in affected programs: 42 -> 40 (-4.76%) helped: 1 HURT: 0 total bytes in shared programs: 42186 -> 42174 (-0.03%) bytes in affected programs: 270 -> 258 (-4.44%) helped: 1 HURT: 0 total halfregs in shared programs: 1769 -> 1757 (-0.68%) halfregs in affected programs: 75 -> 63 (-16.00%) helped: 3 HURT: 0 helped stats (abs) min: 4.0 max: 4.0 x̄: 4.00 x̃: 4 helped stats (rel) min: 16.00% max: 16.00% x̄: 16.00% x̃: 16.00% Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18804>	2022-10-14 01:37:39 +00:00
Alyssa Rosenzweig	de6e11b848	agx: Pass in max regs as a paramter to RA This will allow us to restrict max regs later. Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18804>	2022-10-14 01:37:39 +00:00
Alyssa Rosenzweig	68f89d4cc5	agx: Introduce ra_ctx data structure We have more parameters to pass, this will get unwieldly otherwise. Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18804>	2022-10-14 01:37:39 +00:00
Alyssa Rosenzweig	bcb2cf9688	agx: Write to r0l with a "nesting" instruction This avoids modeling the r0l register explicitly in the IR, which would complicate RA for little benefit at this stage. Do the simplest thing that could possibly work in SSA. glmark2 subset. total instructions in shared programs: 6442 -> 6442 (0.00%) instructions in affected programs: 701 -> 701 (0.00%) helped: 4 HURT: 5 helped stats (abs) min: 1.0 max: 3.0 x̄: 2.00 x̃: 2 helped stats (rel) min: 1.46% max: 7.69% x̄: 4.03% x̃: 3.48% HURT stats (abs) min: 1.0 max: 3.0 x̄: 1.60 x̃: 1 HURT stats (rel) min: 0.81% max: 7.41% x̄: 2.67% x̃: 1.14% 95% mean confidence interval for instructions value: -1.58 1.58 95% mean confidence interval for instructions %-change: -3.70% 3.08% Inconclusive result (value mean confidence interval includes 0). total bytes in shared programs: 42196 -> 42186 (-0.02%) bytes in affected programs: 7768 -> 7758 (-0.13%) helped: 8 HURT: 5 helped stats (abs) min: 2.0 max: 18.0 x̄: 7.25 x̃: 4 helped stats (rel) min: 0.13% max: 7.26% x̄: 2.02% x̃: 0.97% HURT stats (abs) min: 6.0 max: 18.0 x̄: 9.60 x̃: 6 HURT stats (rel) min: 0.82% max: 6.32% x̄: 2.37% x̃: 1.02% 95% mean confidence interval for bytes value: -7.02 5.48 95% mean confidence interval for bytes %-change: -2.30% 1.63% Inconclusive result (value mean confidence interval includes 0). total halfregs in shared programs: 1926 -> 1769 (-8.15%) halfregs in affected programs: 1395 -> 1238 (-11.25%) helped: 71 HURT: 0 helped stats (abs) min: 1.0 max: 10.0 x̄: 2.21 x̃: 2 helped stats (rel) min: 1.92% max: 52.63% x̄: 15.33% x̃: 11.76% 95% mean confidence interval for halfregs value: -2.69 -1.73 95% mean confidence interval for halfregs %-change: -17.98% -12.68% Halfregs are helped. Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18804>	2022-10-14 01:37:39 +00:00
Alyssa Rosenzweig	c9a96d4615	agx: Preload vertex/instance ID only at start This means we don't reserve the registers, which improves RA considerably. Using a special preload psuedo-op instead of a regular move allows us to constrain semantics and gaurantee coalescing. shader-db on glmark2 subset: total instructions in shared programs: 6448 -> 6442 (-0.09%) instructions in affected programs: 230 -> 224 (-2.61%) helped: 4 HURT: 0 total bytes in shared programs: 42232 -> 42196 (-0.09%) bytes in affected programs: 1530 -> 1494 (-2.35%) helped: 4 HURT: 0 total halfregs in shared programs: 2291 -> 1926 (-15.93%) halfregs in affected programs: 2185 -> 1820 (-16.70%) helped: 75 HURT: 0 Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18804>	2022-10-14 01:37:39 +00:00
Alyssa Rosenzweig	f665229d77	agx: Print agx_dim appropriately Easier to read, and gets us closer to proper disasm in Mesa. Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18804>	2022-10-14 01:37:39 +00:00
Alyssa Rosenzweig	6c95572ef0	agx: Print instructions as "dest = src" This makes the dataflow easier to read, especially with splits and collects (which take variable numbers of sources/destinations). Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18804>	2022-10-14 01:37:39 +00:00
Alyssa Rosenzweig	72a1e1f33f	agx: Emit trap at pack-time, not during isel This makes the shaderdb stats make more sense. Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18804>	2022-10-14 01:37:39 +00:00
Alyssa Rosenzweig	1dcaade3e2	agx: Rename "combine" to "collect" For consistency with ir3 and bifrost. Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18804>	2022-10-14 01:37:39 +00:00
Alyssa Rosenzweig	82e8e709cb	agx: Dynamically size split instruction This is more flexible. Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18804>	2022-10-14 01:37:39 +00:00
Alyssa Rosenzweig	7c9fba34bc	agx: Switch to dynamic allocation of srcs/dests So we can handle parallel copies later. Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18804>	2022-10-14 01:37:39 +00:00
Alyssa Rosenzweig	544c60a132	agx: Improve printing of immediate sources For floats, decode the float. Regardless, the size speciifer is redundant. Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18804>	2022-10-14 01:37:39 +00:00
Alyssa Rosenzweig	c2bc8c1384	agx: Don't prefix pseudo-ops It's not really buying us anything and it clutters the IR. Signed-off-by: Alyssa Rosenzweig <alyssa@rosenzweig.io> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/18804>	2022-10-14 01:37:39 +00:00

1 2 3 4 5 ...

161195 Commits