Function _mm_mask_roundscale_round_ss
#[target_feature(enable = "avx512f")]
pub fn _mm_mask_roundscale_round_ss(src: __m128, k: __mmask8, a: __m128, b: __m128, IMM8: i32, SAE: i32) -> __m128
Round the lower single-precision (32-bit) floating-point element in b to the number of fraction bits specified by imm8, store the result in the lower element of dst using writemask k (the element is copied from src when mask bit 0 is not set), and copy the upper 3 packed elements from a to the upper elements of dst.
Rounding is done according to the imm8[2:0] parameter, which can be one of:\
_MM_FROUND_TO_NEAREST_INT: round to nearest_MM_FROUND_TO_NEG_INF: round down_MM_FROUND_TO_POS_INF: round up_MM_FROUND_TO_ZERO: truncate_MM_FROUND_CUR_DIRECTION: useMXCSR.RC- see_MM_SET_ROUNDING_MODE
Exceptions can be suppressed by passing _MM_FROUND_NO_EXC in the sae parameter. Intel's documentation