Programming Languages Research Group: Git

[X86] Improved lowering of packed v8i16 vector shifts by non-constant count.

Before this patch, the backend sub-optimally expanded the non-constant shift
count of a v8i16 shift into a sequence of two 'movd' plus 'movzwl'.

With this patch the backend checks if the target features sse4.1. If so, then
it lets the shuffle legalizer deal with the expansion of the shift amount.

Example:
;;
define <8 x i16> @test(<8 x i16> %A, <8 x i16> %B) {
  %shamt = shufflevector <8 x i16> %B, <8 x i16> undef, <8 x i32> zeroinitializer
  %shl = shl <8 x i16> %A, %shamt
  ret <8 x i16> %shl
}
;;

Before (with -mattr=+avx):
  vmovd  %xmm1, %eax
  movzwl  %ax, %eax
  vmovd  %eax, %xmm1
  vpsllw  %xmm1, %xmm0, %xmm0
  retq

Now:
  vpxor  %xmm2, %xmm2, %xmm2
  vpblendw  $1, %xmm1, %xmm2, %xmm1
  vpsllw  %xmm1, %xmm0, %xmm0
  retq

git-svn-id: https://llvm.org/svn/llvm-project/llvm/trunk@223660 91177308-0d34-0410-b5e6-96231b3b80d8

author	Andrea Di Biagio <Andrea_DiBiagio@sn.scee.net>
	Mon, 8 Dec 2014 14:36:51 +0000 (14:36 +0000)
committer	Andrea Di Biagio <Andrea_DiBiagio@sn.scee.net>
	Mon, 8 Dec 2014 14:36:51 +0000 (14:36 +0000)
commit	ae16ff1c4266a17e9a46940453a64c0889eb49f7
tree	e435f2c4bb0899745b28a1d78584c294851f68cb	tree \| snapshot
parent	968f0454b85489b7ee19e73774a50cff383fdafb	commit \| diff

lib/Target/X86/X86ISelLowering.cpp		diff \| blob \| history
test/CodeGen/X86/lower-vec-shift-2.ll		diff \| blob \| history