Function _mm_maskz_fmsub_round_ss

#[target_feature(enable = "avx512f")]
pub fn _mm_maskz_fmsub_round_ss(k: __mmask8, a: __m128, b: __m128, c: __m128, ROUNDING: i32) -> __m128

Multiply the lower single-precision (32-bit) floating-point elements in a and b, and subtract the lower element in c from the intermediate result. Store the result in the lower element of dst using zeromask k (the element is zeroed out when mask bit 0 is not set), and copy the upper 3 packed elements from a to the upper elements of dst.\

Rounding is done according to the rounding[3:0] parameter, which can be one of:\

Intel's documentation