libyuv

mirror of https://chromium.googlesource.com/libyuv/libyuv synced 2026-01-01 03:12:16 +08:00

History

George Steed 8f659daffd [AArch64] Add SVE2 implementations of NV{12,21}ToRGB24Row Now that we have the `_2X` versions of the macros we can use these to implement `ToRGB24` kernels. These cannot use the bottom/top approach previously used by other SVE kernels since there are three rather than two or four elements each. Reduction in runtimes observed compared to the existing Neon implementations: \| NV12ToRGB24Row \| NV21ToRGB24Row Cortex-A510 \| -60.7% \| -60.7% Cortex-A520 \| -46.0% \| -46.0% Cortex-A715 \| -25.2% \| -25.2% Cortex-A720 \| -25.2% \| -25.2% Cortex-X2 \| -28.9% \| -29.0% Cortex-X3 \| -28.2% \| -28.1% Cortex-X4 \| -30.8% \| -30.7% Cortex-X925 \| -28.8% \| -28.9% Change-Id: I39853d124bfdcac38584109870b398b8ecd5b632 Reviewed-on: https://chromium-review.googlesource.com/c/libyuv/libyuv/+/6067149 Reviewed-by: Frank Barchard <fbarchard@chromium.org>	2024-12-04 17:51:08 +00:00
..
libyuv	[AArch64] Add SVE2 implementations of NV{12,21}ToRGB24Row	2024-12-04 17:51:08 +00:00
libyuv.h	NV12 Copy, include scale_uv.h	2020-12-08 18:54:16 +00:00

George Steed 8f659daffd [AArch64] Add SVE2 implementations of NV{12,21}ToRGB24Row

Now that we have the `_2X` versions of the macros we can use these to
implement `ToRGB24` kernels. These cannot use the bottom/top approach
previously used by other SVE kernels since there are three rather than
two or four elements each.

Reduction in runtimes observed compared to the existing Neon
implementations:

            | NV12ToRGB24Row | NV21ToRGB24Row
Cortex-A510 |         -60.7% |         -60.7%
Cortex-A520 |         -46.0% |         -46.0%
Cortex-A715 |         -25.2% |         -25.2%
Cortex-A720 |         -25.2% |         -25.2%
  Cortex-X2 |         -28.9% |         -29.0%
  Cortex-X3 |         -28.2% |         -28.1%
  Cortex-X4 |         -30.8% |         -30.7%
Cortex-X925 |         -28.8% |         -28.9%

Change-Id: I39853d124bfdcac38584109870b398b8ecd5b632
Reviewed-on: https://chromium-review.googlesource.com/c/libyuv/libyuv/+/6067149
Reviewed-by: Frank Barchard <fbarchard@chromium.org>

2024-12-04 17:51:08 +00:00

libyuv

[AArch64] Add SVE2 implementations of NV{12,21}ToRGB24Row

2024-12-04 17:51:08 +00:00

libyuv.h

NV12 Copy, include scale_uv.h

2020-12-08 18:54:16 +00:00