AlexIndustrial/mesa

Author	SHA1	Message	Date
Iago Toral Quiroga	9e90d95508	v3d,v3dv: support up to 8 render targets in v7.1+ Reviewed-by: Alejandro Piñeiro <apinheiro@igalia.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/25450>	2023-10-13 22:37:43 +00:00
Alejandro Piñeiro	452421dfe5	v3dv: no specific separate_segments flag for V3D 7.1 On V3D 7.1 there is not a flag on the Shader State Record to specify if we are using shared or separate segments. This is done by setting the vpm input size to 0 (so we need to ensure that the output would be the max needed for input/output). We were already doing the latter on the prog_data_vs, so we just need to use those values, instead of assigning default values. As we are here, we also add some comments on the compiler part. Reviewed-by: Iago Toral Quiroga <itoral@igalia.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/25450>	2023-10-13 22:37:43 +00:00
Iago Toral Quiroga	8c191d1103	broadcom/compiler: update thread end restrictions validation for v71 Reviewed-by: Alejandro Piñeiro <apinheiro@igalia.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/25450>	2023-10-13 22:37:43 +00:00
Iago Toral Quiroga	1f5a3391bb	broadcom/compiler: only assign rf0 as last resort in V3D 7.x So we can use it for ldunif(a) and avoid generating ldunif(a)rf which can't be paired with conditional instructions. shader-db (pi5): total instructions in shared programs: 11357802 -> 11338883 (-0.17%) instructions in affected programs: 7117889 -> 7098970 (-0.27%) helped: 24264 HURT: 17574 Instructions are helped. total uniforms in shared programs: 3857808 -> 3857815 (<.01%) uniforms in affected programs: 92 -> 99 (7.61%) helped: 0 HURT: 1 total max-temps in shared programs: 2230904 -> 2230199 (-0.03%) max-temps in affected programs: 52309 -> 51604 (-1.35%) helped: 1219 HURT: 725 Max-temps are helped. total sfu-stalls in shared programs: 15021 -> 15236 (1.43%) sfu-stalls in affected programs: 6848 -> 7063 (3.14%) helped: 1866 HURT: 1704 Inconclusive result total inst-and-stalls in shared programs: 11372823 -> 11354119 (-0.16%) inst-and-stalls in affected programs: 7149177 -> 7130473 (-0.26%) helped: 24315 HURT: 17561 Inst-and-stalls are helped. total nops in shared programs: 273624 -> 273711 (0.03%) nops in affected programs: 31562 -> 31649 (0.28%) helped: 1619 HURT: 1854 Inconclusive result (value mean confidence interval includes 0). Reviewed-by: Alejandro Piñeiro <apinheiro@igalia.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/25450>	2023-10-13 22:37:42 +00:00
Iago Toral Quiroga	c8e4ee8ecb	broadcom/compiler: don't assign registers to unused nodes/temps In programs with a lot of unused temps, if we don't do this, we may end up recycling previously used rfs more often, which can be detrimental to instruction pairing. total instructions in shared programs: 11464335 -> 11444136 (-0.18%) instructions in affected programs: 8976743 -> 8956544 (-0.23%) helped: 33196 HURT: 33778 Inconclusive result total max-temps in shared programs: 2230150 -> 2229445 (-0.03%) max-temps in affected programs: 86413 -> 85708 (-0.82%) helped: 2217 HURT: 1523 Max-temps are helped. total sfu-stalls in shared programs: 18077 -> 17104 (-5.38%) sfu-stalls in affected programs: 8669 -> 7696 (-11.22%) helped: 2657 HURT: 2182 Sfu-stalls are helped. total inst-and-stalls in shared programs: 11482412 -> 11461240 (-0.18%) inst-and-stalls in affected programs: 8995697 -> 8974525 (-0.24%) helped: 33319 HURT: 33708 Inconclusive result total nops in shared programs: 298140 -> 296185 (-0.66%) nops in affected programs: 52805 -> 50850 (-3.70%) helped: 3797 HURT: 2662 Inconclusive result Reviewed-by: Alejandro Piñeiro <apinheiro@igalia.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/25450>	2023-10-13 22:37:42 +00:00
Iago Toral Quiroga	ce13aa4ee7	broadcom/compiler: improve allocation for final program instructions The last 3 instructions can't use specific registers so flag all the nodes for temps used in the last program instructions and try to avoid assigning any of these. This may help us avoid injecting nops for the last thread switch instruction. Because regisster allocation needs to happen before QPU scheduling and instruction merging we can't tell exactly what the last 3 instructions will be, so we do this for a few more instructions than just 3. We only do this for fragment shaders because other shader stages always end with VPM store instructions that take an small immediate and therefore will never allow us to merge the final thread switch earlier, so limiting allocation for these shaders will never improve anything and might instead be detrimental. total instructions in shared programs: 11471389 -> 11464335 (-0.06%) instructions in affected programs: 582908 -> 575854 (-1.21%) helped: 4669 HURT: 578 Instructions are helped. total max-temps in shared programs: 2230497 -> 2230150 (-0.02%) max-temps in affected programs: 5662 -> 5315 (-6.13%) helped: 344 HURT: 44 Max-temps are helped. total sfu-stalls in shared programs: 18068 -> 18077 (0.05%) sfu-stalls in affected programs: 264 -> 273 (3.41%) helped: 37 HURT: 48 Inconclusive result (value mean confidence interval includes 0). total inst-and-stalls in shared programs: 11489457 -> 11482412 (-0.06%) inst-and-stalls in affected programs: 585180 -> 578135 (-1.20%) helped: 4659 HURT: 588 Inst-and-stalls are helped. total nops in shared programs: 301738 -> 298140 (-1.19%) nops in affected programs: 14680 -> 11082 (-24.51%) helped: 3252 HURT: 108 Nops are helped. Reviewed-by: Alejandro Piñeiro <apinheiro@igalia.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/25450>	2023-10-13 22:37:42 +00:00
Iago Toral Quiroga	818fc41e7e	broadcom/compiler: don't allocate spill base to rf0 in V3D 7.x Otherwise it can be stomped by instructions doing implicit rf0 writes. Reviewed-by: Alejandro Piñeiro <apinheiro@igalia.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/25450>	2023-10-13 22:37:42 +00:00
Iago Toral Quiroga	84c912c1d4	broadcom/compiler: fix up copy propagation for v71 Update rules for unsafe copy propagations to match v7.x. Reviewed-by: Alejandro Piñeiro <apinheiro@igalia.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/25450>	2023-10-13 22:37:42 +00:00
Iago Toral Quiroga	1e85be415a	broadcom/compiler: lift restriction on vpmwt in last instruction for V3D 7.x Reviewed-by: Alejandro Piñeiro <apinheiro@igalia.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/25450>	2023-10-13 22:37:42 +00:00
Iago Toral Quiroga	2774601780	broadcom/compiler: validate restrictions after TLB Z write Reviewed-by: Alejandro Piñeiro <apinheiro@igalia.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/25450>	2023-10-13 22:37:42 +00:00
Iago Toral Quiroga	d4285d7f2a	broadcom/compiler: start allocating from RF 4 in V7.x In V3D 4.x we start at RF3 so that we allocate RF0-2 only if there aren't any other RFs available. This is useful with small shaders to ensure that our TLB writes don't use these registers because these are the last instructions we emit in fragment shaders and the last instructions in a program can't write to these registers, so if we do, we need to emit NOPs. In V3D 7.x the registers affected by this restriction are RF2-3, so we choose to start at RF4. Reviewed-by: Alejandro Piñeiro <apinheiro@igalia.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/25450>	2023-10-13 22:37:42 +00:00
Iago Toral Quiroga	2b39bb35c5	broadcom/compiler: lift restriction for branch + msfign after setmsf for v7.x Reviewed-by: Alejandro Piñeiro <apinheiro@igalia.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/25450>	2023-10-13 22:37:42 +00:00
Iago Toral Quiroga	5e9b405aa7	broadcom/compiler: update ldvary thread switch delay slot restriction for v7.x In V3D 7.x we don't have accumulators which would not survive a thread switch, so the only restriction is that ldvary can't be placed in the second delay slot of a thread switch. shader-db results for UnrealEngine4 shaders: total instructions in shared programs: 446458 -> 446401 (-0.01%) instructions in affected programs: 13492 -> 13435 (-0.42%) helped: 58 HURT: 3 Instructions are helped. total nops in shared programs: 19571 -> 19541 (-0.15%) nops in affected programs: 161 -> 131 (-18.63%) helped: 30 HURT: 0 Nops are helped. Reviewed-by: Alejandro Piñeiro <apinheiro@igalia.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/25450>	2023-10-13 22:37:42 +00:00
Iago Toral Quiroga	526c1889e5	broadcom/compiler: update thread end restrictions for v7.x In 4.x it is not allowed to write to the register file in the last 3 instructions, but in 7.x we only have this restriction in the thread end instruction itself, and only if the write comes from the ALU ports. Reviewed-by: Alejandro Piñeiro <apinheiro@igalia.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/25450>	2023-10-13 22:37:42 +00:00
Iago Toral Quiroga	ced83e7803	broadcom/compiler: implement small immediates for v71 Reviewed-by: Alejandro Piñeiro <apinheiro@igalia.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/25450>	2023-10-13 22:37:42 +00:00
Iago Toral Quiroga	e4d30600a4	broadcom/compiler: convert mul to add when needed to allow merge V3D 7.x added 'mov' opcodes to the ADD alu, so now it is possible to move these to the ADD alu to facilitate merging them with other MUL instructions. Reviewed-by: Alejandro Piñeiro <apinheiro@igalia.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/25450>	2023-10-13 22:37:42 +00:00
Iago Toral Quiroga	cbedf14687	broadcom/compiler: don't assign rf0 to temps that conflict with ldvary ldvary writes to rf0 implicitly, so we don't want to allocate rf0 to any temps that are live across ldvary's rf0 live ranges. Reviewed-by: Alejandro Piñeiro <apinheiro@igalia.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/25450>	2023-10-13 22:37:42 +00:00
Iago Toral Quiroga	3a36a618d7	broadcom/compiler: try to use ldunif(a) instead of ldunif(a)rf in v71 The rf variants need to encode the destination in the cond bits, which prevents these to be merged with any other instruction that need them. In 4.x, ldunif(a) write to r5 which is a special register that only ldunif(a) and ldvary can write so we have a special register class for it and only allow it for them. Then when we need to choose a register for a node, if this register is available we always use it. In 7.x these instructions write to rf0, which can be used by any instruction, so instead of restricting rf0, we track the temps that are used as ldunif(a) destinations and use that information to favor rf0 for them. Reviewed-by: Alejandro Piñeiro <apinheiro@igalia.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/25450>	2023-10-13 22:37:42 +00:00
Iago Toral Quiroga	d8a25bdb07	broadcom/compiler: enable ldvary pipelining on v71 Reviewed-by: Alejandro Piñeiro <apinheiro@igalia.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/25450>	2023-10-13 22:37:42 +00:00
Iago Toral Quiroga	a8014be2b0	broadcom/compiler: handle rf0 flops storage restriction in v71 Reviewed-by: Alejandro Piñeiro <apinheiro@igalia.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/25450>	2023-10-13 22:37:42 +00:00
Iago Toral Quiroga	d1281d857f	broadcom/compiler: update peripheral access restrictions for v71 In V3D 4.x only a couple of simultaneous accesses where allowed, but V3D 7.x is a bit more flexible, so rather than trying to check for all the allowed combinations it is easier to check if we are one of the disallows. Shader-db (pi5): total instructions in shared programs: 11338883 -> 11307386 (-0.28%) instructions in affected programs: 2727201 -> 2695704 (-1.15%) helped: 12555 HURT: 289 Instructions are helped. total max-temps in shared programs: 2230199 -> 2229260 (-0.04%) max-temps in affected programs: 20508 -> 19569 (-4.58%) helped: 608 HURT: 4 Max-temps are helped. total sfu-stalls in shared programs: 15236 -> 15293 (0.37%) sfu-stalls in affected programs: 148 -> 205 (38.51%) helped: 38 HURT: 64 Inconclusive result (%-change mean confidence interval includes 0). total inst-and-stalls in shared programs: 11354119 -> 11322679 (-0.28%) inst-and-stalls in affected programs: 2732262 -> 2700822 (-1.15%) helped: 12550 HURT: 304 Inst-and-stalls are helped. total nops in shared programs: 273711 -> 274095 (0.14%) nops in affected programs: 9626 -> 10010 (3.99%) helped: 186 HURT: 397 Nops are HURT. Reviewed-by: Alejandro Piñeiro <apinheiro@igalia.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/25450>	2023-10-13 22:37:42 +00:00
Alejandro Piñeiro	ce66c9aead	broadcom/compiler: update payload registers handling when computing live intervals As for v71 the payload registers are not the same. Specifically now rf3 is used as payload register, so this is needed to avoid rf3 being selected as a instruction dst by the register allocator, overwriting the payload value that could be still used. Reviewed-by: Iago Toral Quiroga <itoral@igalia.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/25450>	2023-10-13 22:37:42 +00:00
Alejandro Piñeiro	d72e57fe30	broadcom/compiler: update ldunif/ldvary comment for v71 For v42 and below ldunif/ldvary write both on r5, but with a different delay, so we need to take that into account when scheduling both. For v71 the register used is rf0, but the behaviour is the same. So the scheduling code can be the same, but the comment needs update. Reviewed-by: Iago Toral Quiroga <itoral@igalia.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/25450>	2023-10-13 22:37:42 +00:00
Alejandro Piñeiro	a3aba3f352	broadcom/compiler: update one TMUWT restriction for v71 TMUWT not allowed in the final instruction restriction doesn't apply for v71. Reviewed-by: Iago Toral Quiroga <itoral@igalia.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/25450>	2023-10-13 22:37:42 +00:00
Iago Toral Quiroga	c9fcd5d786	broadcom/compiler: v71 isn't affected by double-rounding of viewport X,Y coords Reviewed-by: Alejandro Piñeiro <apinheiro@igalia.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/25450>	2023-10-13 22:37:42 +00:00
Iago Toral Quiroga	5c7224b81f	broadcom/compiler: generalize check for shaders using pixel center W V3D 4.x has pixel center W in rf0 and V3D 7.x has it in rf3. We already account for this when we setup the c->payload_w, so use that. Reviewed-by: Alejandro Piñeiro <apinheiro@igalia.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/25450>	2023-10-13 22:37:42 +00:00
Iago Toral Quiroga	b4e0c9bac4	broadcom/compiler: allow instruction merges in v71 In v3d 4.x there were restrictions based on the number of raddrs used by the combined instructions, but we don't have these restrictions in v3d 7.x. It should be noted that while there are no restrictions on the number of raddrs addressed, a QPU instruction can only address a single small immediate, so we should be careful about that when we add support for small immediates. Reviewed-by: Alejandro Piñeiro <apinheiro@igalia.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/25450>	2023-10-13 22:37:42 +00:00
Iago Toral Quiroga	28631a5550	broadcom/compiler: don't schedule rf0 writes right after ldvary ldvary writes rf0 implicitly on the next cycle so they would clash. This case is not handled correctly by our normal dependency tracking, which doesn't know anything about delayed writes from instructions and thinks the rf0 write happens on the same cycle ldvary is emitted. Fixes (v71): dEQP-VK.glsl.conversions.matrix_to_matrix.mat2x3_to_mat4x2_fragment Reviewed-by: Alejandro Piñeiro <apinheiro@igalia.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/25450>	2023-10-13 22:37:42 +00:00
Iago Toral Quiroga	42b70f624b	broadcom/compiler: CS payload registers have changed in v71 Reviewed-by: Alejandro Piñeiro <apinheiro@igalia.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/25450>	2023-10-13 22:37:42 +00:00
Iago Toral Quiroga	2b15df963e	broadcom/compiler: don't assign rf0 to temps across implicit rf0 writes In platforms that don't have accumulators and have implicit writes to the register file we need to be careful and avoid assigning a physical register to a temp that lives across an implicit write to that same physical register. For now, we have the case of implicit writes to rf0 from various signals, but it should be easy to extend this to include additional registers if needed. Reviewed-by: Alejandro Piñeiro <apinheiro@igalia.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/25450>	2023-10-13 22:37:42 +00:00
Iago Toral Quiroga	03594b3dca	broadcom/compiler: only handle accumulator classes if present Reviewed-by: Alejandro Piñeiro <apinheiro@igalia.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/25450>	2023-10-13 22:37:41 +00:00
Iago Toral Quiroga	b1548b18d3	broadcom/compiler: rename vir_writes_rX to vir_writes_rX_implicitly Since that represents more accurately what they check.. Reviewed-by: Alejandro Piñeiro <apinheiro@igalia.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/25450>	2023-10-13 22:37:41 +00:00
Iago Toral Quiroga	20b37b273f	broadcom/compiler: make vir_write_rX return false on platforms without accums Reviewed-by: Alejandro Piñeiro <apinheiro@igalia.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/25450>	2023-10-13 22:37:41 +00:00
Iago Toral Quiroga	caf28e5681	broadcom/compiler: prevent rf2-3 usage in thread end delay slots for v71 Reviewed-by: Alejandro Piñeiro <apinheiro@igalia.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/25450>	2023-10-13 22:37:41 +00:00
Iago Toral Quiroga	572fba0bf4	broadcom/compiler: implement read stall check for v71 Reviewed-by: Alejandro Piñeiro <apinheiro@igalia.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/25450>	2023-10-13 22:37:41 +00:00
Iago Toral Quiroga	4e26d2c156	broadcom/compiler: implement "reads/writes too soon" checks for v71 Reviewed-by: Alejandro Piñeiro <apinheiro@igalia.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/25450>	2023-10-13 22:37:41 +00:00
Alejandro Piñeiro	083d082d8e	broadcom/compiler: update register classes to not include accumulators on v71 Reviewed-by: Iago Toral Quiroga <itoral@igalia.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/25450>	2023-10-13 22:37:41 +00:00
Alejandro Piñeiro	d12eb68d3a	broadcom/qpu_schedule: update write deps for v71 We just need to add a write dep if rf0 is written implicitly. Note that we don't need to check if we have accumulators when checking for r3/r4/r5, as v3d_qpu_writes_rX would return false for hw version that doesn't have accumulators. Reviewed-by: Iago Toral Quiroga <itoral@igalia.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/25450>	2023-10-13 22:37:41 +00:00
Alejandro Piñeiro	5a035af931	broadcom/compiler: payload_w is loaded on rf3 for v71 And in general rf0 is now used for other needs. Reviewed-by: Iago Toral Quiroga <itoral@igalia.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/25450>	2023-10-13 22:37:41 +00:00
Alejandro Piñeiro	edfc36817a	broadcom/compiler: add support for varyings on nir to vir generation for v71 Needs update as v71 doesn't have accumulators anymore, and ldvary uses now rf0 to return the value. Reviewed-by: Iago Toral Quiroga <itoral@igalia.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/25450>	2023-10-13 22:37:41 +00:00
Alejandro Piñeiro	a766cc3a5a	broadcom/qpu_schedule: add process_raddr_deps On v71 we don't have muxes, but more raddr. Adding a equivalent add deps function. Reviewed-by: Iago Toral Quiroga <itoral@igalia.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/25450>	2023-10-13 22:37:41 +00:00
Alejandro Piñeiro	dad6917d5e	broadcom/compiler: update vir_to_qpu::set_src for v71 Reviewed-by: Iago Toral Quiroga <itoral@igalia.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/25450>	2023-10-13 22:37:41 +00:00
Alejandro Piñeiro	136d934c80	broadcom/vir: implement is_no_op_mov for v71 Did some refactoring/splitting. Reviewed-by: Iago Toral Quiroga <itoral@igalia.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/25450>	2023-10-13 22:37:41 +00:00
Alejandro Piñeiro	d00f7ef23e	broadcom/compiler: don't favor/select accum registers for hw not supporting it Note that what we do is to just return false on the favor/select accum methods. We could just avoid to call them, but as the select is called more than once, it is just easier this way. Reviewed-by: Iago Toral Quiroga <itoral@igalia.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/25450>	2023-10-13 22:37:41 +00:00
Alejandro Piñeiro	1260b202be	broadcom/compiler: phys index depends on hw version For 7.1 there are not accumulators. So we replace the macro with a function call. Reviewed-by: Iago Toral Quiroga <itoral@igalia.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/25450>	2023-10-13 22:37:41 +00:00
Iago Toral Quiroga	63d633ca7a	broadcom/compiler: update node/temp translation for v71 As the offset applied needs to take into account if we have accumulators or not. Reviewed-by: Alejandro Piñeiro <apinheiro@igalia.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/25450>	2023-10-13 22:37:41 +00:00
Alejandro Piñeiro	347065525f	broadcom/qpu: define v3d_qpu_input, use on v3d_qpu_alu_instr At this point it just tidy up a little the alu_instr structure. But also serves to prepare the structure for new changes, as 7.x uses raddr instead of mux, and it is just easier to add the raddr to the new input structure. Signed-off-by: Alejandro Piñeiro <apinheiro@igalia.com> Signed-off-by: Iago Toral Quiroga <itoral@igalia.com> Reviewed-by: Iago Toral Quiroga <itoral@igalia.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/25450>	2023-10-13 22:37:41 +00:00
Alejandro Piñeiro	3d0c3667dd	broadcom/compiler: add small_imm a/c/d on v3d_qpu_sig small_imm_a, small_imm_c and small_imm_d added on top of the already existing small_imm_b, as V3D 7.1 defines 4 small immediates, tied to the 4 raddr. Note that this is only the definition, and just a inst validation rule to check that are not used before v71. Any real use is still pending. Reviewed-by: Iago Toral Quiroga <itoral@igalia.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/25450>	2023-10-13 22:37:41 +00:00
Alejandro Piñeiro	e5011e19c7	broadcom/compiler: rename small_imm to small_imm_b Current small_imm is associated with the "B" read address. We do this change in advance for v71 support, where we will have 4 different small_imm (a/b/c/d), so we start with a renaming. Reviewed-by: Iago Toral Quiroga <itoral@igalia.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/25450>	2023-10-13 22:37:41 +00:00
Iago Toral Quiroga	4afbf4ad31	v3d: get rid of shader_state pointer in v3d_key Having this pointer in the key is undesirable since it makes copying keys difficult and error prone (as seen in previous patches), also, it is only there for convenience and we don't strictly need it (in fact the vulkan driver doesn't use it at all), so let's just get rid of it so our v3d_key is fully static. Reviewed-by: Juan A. Suarez <jasuarez@igalia.com> Part-of: <https://gitlab.freedesktop.org/mesa/mesa/-/merge_requests/25418>	2023-10-02 06:35:07 +00:00

1 2 3 4 5 ...

809 Commits