How do I convert _m128i to an unsigned int with SSE?

浪尽此生 提交于 2019-12-03 09:14:00

Unfortunately, there's no instruction to do that even in AVX (none that I'm aware of). So you will have to do it manually like are right now.

However, your current method is very sub-optimal and you're relying on .m128i_u8 which is an MSVC extension. Based on my experience with MSVC, it will use an aligned buffer to access the individual elements. This has a very heavy penalty because of partial-word access.

Instead of .m128i_u8, use _mm_extract_epi32(). This is in SSE4.1. But you're already relying with SSE4.1 with _mm_cvtepu8_epi32().

This situation is particularly bad since you're working with 1-byte granularity. If you were working with 2-byte (16-bit integer) granularity instead, there is an efficient solution using shuffle intrinsics.

Marat Dukhan

A combination of _mm_shuffle_epi8 and _mm_cvtsi128_si32 is what you need:

static const __m128i shuffleMask = _mm_setr_epi8(0,  4,  8, 12, -1, -1, -1, -1,
                                               -1, -1, -1, -1, -1, -1, -1, -1);
UINT color = _mm_cvtsi128_si32(_mm_shuffle_epi8(iClr, shuffleMask));
易学教程内所有资源均来自网络或用户发布的内容,如有违反法律规定的内容欢迎反馈
该文章没有解决你所遇到的问题?点击提问,说说你的问题,让更多的人一起探讨吧!