Function _mm_fnmadd_ss

#[target_feature(enable = "fma")]
pub const fn _mm_fnmadd_ss(a: __m128, b: __m128, c: __m128) -> __m128

Multiplies the lower single-precision (32-bit) floating-point elements in a and b, and add the negated intermediate result to the lower element in c. Store the result in the lower element of the returned value, and copy the 3 upper elements from a to the upper elements of the result.

Intel's documentation