Function _mm_mask_multishift_epi64_epi8

#[target_feature(enable = "avx512vbmi", enable = "avx512vl")]
pub fn _mm_mask_multishift_epi64_epi8(src: __m128i, k: __mmask16, a: __m128i, b: __m128i) -> __m128i

For each 64-bit element in b, select 8 unaligned bytes using a byte-granular shift control within the corresponding 64-bit element of a, and store the 8 assembled bytes to the corresponding 64-bit element of dst using writemask k (elements are copied from src when the corresponding mask bit is not set).

Intel's documentation