mirror of
https://github.com/FFmpeg/FFmpeg.git
synced 2026-08-12 22:16:21 +00:00
Tests for changes or regressions in the generated micro-ops. This will be
instrumental in my development of the micro-ops optimizer, and my plans to
phase out some of the macro-op optimization passes in favor of doing those
optimizations on the uop level instead.
rgb24 16x16 -> rgb24 16x32:
[ u8 +++X] SWS_OP_READ : 3 elem(s) packed >> 0
min: {0 0 0 _}, max: {255 255 255 _}
[ u8 ...X] SWS_OP_FILTER_V : 16 -> 32 bilinear (2 taps)
min: {0 0 0 _}, max: {255 255 255 _}
[f32 ...X] SWS_OP_DITHER : 16x16 matrix + {0 3 2 -1}
min: {1/512 1/512 1/512 _}, max: {255.998047 255.998047 255.998047 _}
[f32 ...X] SWS_OP_MIN : x <= {255 255 255 _}
min: {1/512 1/512 1/512 _}, max: {255 255 255 _}
[f32 +++X] SWS_OP_CONVERT : f32 -> u8
min: {0 0 0 _}, max: {255 255 255 _}
[ u8 XXXX] SWS_OP_WRITE : 3 elem(s) packed >> 0
(X = unused, z = byteswapped, + = exact, 0 = zero)
Retrying with split passes:
[ u8 +++X] SWS_OP_READ : 3 elem(s) packed >> 0
min: {0 0 0 _}, max: {255 255 255 _}
[ u8 XXXX] SWS_OP_WRITE : 3 elem(s) planar >> 0
(X = unused, z = byteswapped, + = exact, 0 = zero)
+ translated micro-ops:
+ u8_read_packed_xyz
+ u8_write_planar_xyz
Sub-pass #1:
[ u8 ...X] SWS_OP_READ : 3 elem(s) planar >> 0 + 2 tap bilinear filter (V)
min: {0 0 0 _}, max: {255 255 255 _}
[f32 ...X] SWS_OP_DITHER : 16x16 matrix + {0 3 2 -1}
min: {1/512 1/512 1/512 _}, max: {255.998047 255.998047 255.998047 _}
[f32 ...X] SWS_OP_MIN : x <= {255 255 255 _}
min: {1/512 1/512 1/512 _}, max: {255 255 255 _}
[f32 +++X] SWS_OP_CONVERT : f32 -> u8
min: {0 0 0 _}, max: {255 255 255 _}
[ u8 XXXX] SWS_OP_WRITE : 3 elem(s) packed >> 0
(X = unused, z = byteswapped, + = exact, 0 = zero)
+ translated micro-ops:
+ u8_read_planar_fv_xyz
+ f32_dither_xyz_0_3_2_16x16
+ f32_min_xyz
+ f32_to_u8_xyz
+ u8_write_packed_xyz
...
Signed-off-by: Niklas Haas <git@haasn.dev>