« Previous | Next » 

Revision 6341838f

Parent ace7f813
Child f7cf0f31

Added by Ronald S. Bultje almost 11 years ago

Use word-writing instead of dword-writing (with two cached but otherwise
unchanged bytes) in the horizontal simple loopfilter. This makes the filter
quite a bit faster in itself (~30 cycles less on Core1), probably mostly
because we don't need a complex 4x4 transpose, but only a simple byte
interleave. Also allows using pextrw on SSE4, which speeds up even more
(e.g. 25% faster on Core i7).

Originally committed as revision 24638 to svn://svn.ffmpeg.org/ffmpeg/trunk


  • added
  • modified
  • copied
  • renamed
  • deleted

View differences