petrbalvin
|
4632ac1bb9
|
feat(arch): add the AVX-VNNI-INT16 dot products to the extension layer
Assisted-by: GLM 5.3 Flash
|
2026-10-07 19:49:15 +02:00 |
|
petrbalvin
|
96000dd64d
|
feat(arch): add the VEX encoder to the extension layer
Assisted-by: GLM 5.3 Flash
|
2026-10-07 19:49:15 +02:00 |
|
petrbalvin
|
47d561b229
|
refactor(arch): extract the displacement tail and SIB builders
Assisted-by: GLM 5.3 Flash
|
2026-10-07 19:49:15 +02:00 |
|
petrbalvin
|
27ef71859b
|
feat(arch): add VMINMAXSH to the extension layer
Assisted-by: GLM 5.3 Flash
|
2026-10-07 19:49:15 +02:00 |
|
petrbalvin
|
03f9ef0ac6
|
feat(arch): add the FP16 complex fused multiply-add to the extension layer
Assisted-by: GLM 5.3 Flash
|
2026-10-07 19:49:15 +02:00 |
|
petrbalvin
|
509afbb6c9
|
feat(arch): add the FP16 complex multiply and minimum-maximum to the extension layer
Assisted-by: GLM 5.3 Flash
|
2026-10-07 18:59:09 +02:00 |
|
petrbalvin
|
8080e0acef
|
feat(arch): add the AVX512-FP16 FMA families to the extension layer
Assisted-by: GLM 5.3 Flash
|
2026-10-07 18:58:49 +02:00 |
|
petrbalvin
|
d03de62c07
|
feat(arch): add the amd64 fp16 packed imm8-control group
Assisted-by: GLM 5.3
|
2026-10-07 13:51:48 +02:00 |
|
petrbalvin
|
8de1b371da
|
feat(arch): add the amd64 fp16 packed conversion family
Assisted-by: GLM 5.3
|
2026-10-07 13:51:48 +02:00 |
|
petrbalvin
|
fb6d01a7d0
|
feat(arch): encode the amd64 embedded rounding and SAE decorations
Assisted-by: GLM 5.3
|
2026-10-07 13:51:48 +02:00 |
|
petrbalvin
|
c2adde948f
|
feat(arch): add the write mask to the packed amd64 destinations
Assisted-by: GLM 5.3 Flash
|
2026-10-07 02:21:57 +02:00 |
|
petrbalvin
|
11cac26508
|
feat(arch): add the scaled index to the amd64 memory operands
Assisted-by: GLM 5.3 Flash
|
2026-10-07 02:06:11 +02:00 |
|
petrbalvin
|
ccb155437e
|
feat(arch): add the {1toN} broadcast to the packed amd64 memory sources
Assisted-by: GLM 5.3 Flash
|
2026-10-07 02:06:11 +02:00 |
|
petrbalvin
|
8d611bfdaf
|
feat(arch): add the remaining scalar FP16 memory forms to the extension layer
Assisted-by: GLM 5.3 Flash
|
2026-10-07 01:35:48 +02:00 |
|
petrbalvin
|
454a21f5b7
|
feat(arch): add the packed FP16 and BF16 memory forms to the extension layer
Assisted-by: GLM 5.3 Flash
|
2026-10-07 01:35:48 +02:00 |
|
petrbalvin
|
d275dee3ae
|
feat(arch): add the scalar FP16 memory forms to the amd64 extension layer
Assisted-by: GLM 5.3 Flash
|
2026-10-07 01:35:48 +02:00 |
|
petrbalvin
|
7d69dda874
|
feat(arch): add the amd64 memory-operand mechanism to the extension layer
Assisted-by: GLM 5.3 Flash
|
2026-10-07 01:35:48 +02:00 |
|
petrbalvin
|
103864e8b2
|
feat(arch): add the VL packed FP16 forms to the amd64 extension layer
Assisted-by: GLM 5.3 Flash
|
2026-10-07 00:37:41 +02:00 |
|
petrbalvin
|
aa9c7ca030
|
feat(arch): add the imm8 scalar FP16 controls to the amd64 extension layer
Assisted-by: GLM 5.3 Flash
|
2026-10-07 00:33:30 +02:00 |
|
petrbalvin
|
0354a1f4c1
|
feat(arch): scale and exponent-extract the scalar FP16 in the extension layer
Assisted-by: GLM 5.3 Flash
|
2026-10-07 00:07:58 +02:00 |
|
petrbalvin
|
22055b9bf3
|
feat(arch): add the packed FP16 arithmetic to the amd64 extension layer
Assisted-by: GLM 5.3 Flash
|
2026-10-07 00:07:58 +02:00 |
|
petrbalvin
|
bc37ea5b79
|
feat(arch): add the AVX512-FP16 scalar family to the amd64 extension layer
Assisted-by: GLM 5.3 Flash
|
2026-10-07 00:07:58 +02:00 |
|
petrbalvin
|
28bea95128
|
feat(arch): add the amd64 extended-instruction layer with BF16 and VP2INTERSECT
Assisted-by: GLM 5.3 Flash
|
2026-10-07 00:07:58 +02:00 |
|