Function _mm_maskz_sub_round_ss

#[target_feature(enable = "avx512f")]
pub fn _mm_maskz_sub_round_ss(k: __mmask8, a: __m128, b: __m128, ROUNDING: i32) -> __m128

Subtract the lower single-precision (32-bit) floating-point element in b from the lower single-precision (32-bit) floating-point element in a, store the result in the lower element of dst using zeromask k (the element is zeroed out when mask bit 0 is not set), and copy the upper 3 packed elements from a to the upper elements of dst.\

Rounding is done according to the rounding[3:0] parameter, which can be one of:\

Intel's documentation