Function _mm512_mask_fmadd_round_ph

#[target_feature(enable = "avx512fp16")]
pub fn _mm512_mask_fmadd_round_ph(a: __m512h, k: __mmask32, b: __m512h, c: __m512h, ROUNDING: i32) -> __m512h

Multiply packed half-precision (16-bit) floating-point elements in a and b, add the intermediate result to packed elements in c, and store the results in dst using writemask k (the element is copied from a when the corresponding mask bit is not set).

Rounding is done according to the rounding parameter, which can be one of:

Intel's documentation