Function _mm_mask_add_round_ss

#[target_feature(enable = "avx512f")]
pub fn _mm_mask_add_round_ss(src: __m128, k: __mmask8, a: __m128, b: __m128, ROUNDING: i32) -> __m128

Add the lower single-precision (32-bit) floating-point element in a and b, store the result in the lower element of dst using writemask k (the element is copied from src when mask bit 0 is not set), and copy the upper 3 packed elements from a to the upper elements of dst.\

Rounding is done according to the rounding[3:0] parameter, which can be one of:\

Intel's documentation