• levytamar82's avatar
    AVX2 SubPixel AVG Variance Optimization · ea149096
    levytamar82 authored
    Optimizing 2 functions to process 32 elements in parallel instead of 16:
    1. vp9_sub_pixel_avg_variance64x64
    2. vp9_sub_pixel_avg_variance32x32
    both of those function were calling vp9_sub_pixel_avg_variance16xh_ssse3
    instead of calling that function, it calls vp9_sub_pixel_avg_variance32xh_avx2
    that is written in avx2 and process 32 elements in parallel.
    This Optimization gave 80% function level gain and 2% user level gain
    
    Change-Id: Iea694654e1b7612dc6ed11e2626208c2179502c8
    ea149096
Name
Last commit
Last update
..
common Loading commit data...
decoder Loading commit data...
encoder Loading commit data...
exports_dec Loading commit data...
exports_enc Loading commit data...
vp9_common.mk Loading commit data...
vp9_cx_iface.c Loading commit data...
vp9_dx_iface.c Loading commit data...
vp9_iface_common.h Loading commit data...
vp9cx.mk Loading commit data...
vp9dx.mk Loading commit data...