mirror of
https://github.com/FEX-Emu/FEX.git
synced 2026-10-07 14:00:17 +02:00
Lets the AVX implementation get all the optimizations that the SSE variant has, reducing the overhead a little. Even with the individual lane handling, this is still leagues better than all of the individual inserts that are pretty beefy with SVE. For example: vpshufd ymm0, ymm1, 0b00000011 drops from 50 instructions to 9