Function _mm256_maskz_reduce_ph

#[target_feature(enable = "avx512fp16", enable = "avx512vl")]
pub fn _mm256_maskz_reduce_ph(k: __mmask16, a: __m256h, IMM8: i32) -> __m256h

Extract the reduced argument of packed half-precision (16-bit) floating-point elements in a by the number of bits specified by imm8, and store the results in dst using zeromask k (elements are zeroed out when the corresponding mask bit is not set).

Rounding is done according to the imm8 parameter, which can be one of:

Intel's documentation