Function _mm512_maskz_fmsub_round_ph

#[target_feature(enable = "avx512fp16")]
pub fn _mm512_maskz_fmsub_round_ph(k: __mmask32, a: __m512h, b: __m512h, c: __m512h, ROUNDING: i32) -> __m512h

Multiply packed half-precision (16-bit) floating-point elements in a and b, subtract packed elements in c from the intermediate result, and store the results in dst using zeromask k (the element is zeroed out when the corresponding mask bit is not set).

Rounding is done according to the rounding parameter, which can be one of:

Intel's documentation