ffmpeg/libavcodec/x86
Ronald S. Bultje 6341838f3c Use word-writing instead of dword-writing (with two cached but otherwise
unchanged bytes) in the horizontal simple loopfilter. This makes the filter
quite a bit faster in itself (~30 cycles less on Core1), probably mostly
because we don't need a complex 4x4 transpose, but only a simple byte
interleave. Also allows using pextrw on SSE4, which speeds up even more
(e.g. 25% faster on Core i7).

Originally committed as revision 24638 to svn://svn.ffmpeg.org/ffmpeg/trunk
2010-07-31 23:13:15 +00:00
..
cavsdsp_mmx.c Make ff_pw_4 128 bits 2010-07-11 22:52:55 +00:00
cpuid.c Remove FF_MM_SSE2/3 flags for CPUs where this is generally not faster than 2010-07-19 22:38:23 +00:00
dct32_sse.c Move SSE optimized 32-point DCT to its own file. Should fix breakage with YASM 2010-07-06 17:48:23 +00:00
deinterlace.asm Convert deinterlacing MMX code to YASM 2010-07-31 14:50:51 +00:00
dnxhd_mmx.c
dsputil_h264_template_mmx.c Remove DECLARE_ALIGNED_{8,16} macros 2010-03-06 14:24:59 +00:00
dsputil_h264_template_ssse3.c Make ff_pw_4 128 bits 2010-07-11 22:52:55 +00:00
dsputil_mmx_avg_template.c Add bitexact versions of put_no_rnd_pixels8 _x2 and _y2 for vp3/theora 2010-06-04 04:46:26 +00:00
dsputil_mmx_qns_template.c
dsputil_mmx_rnd_template.c
dsputil_mmx.c relicense h264 deblock sse2 to lgpl 2010-07-22 00:39:49 +00:00
dsputil_mmx.h Convert deinterlacing MMX code to YASM 2010-07-31 14:50:51 +00:00
dsputil_yasm.asm Update x264asm header files to latest versions. 2010-06-23 19:20:46 +00:00
dsputilenc_mmx.c Remove FF_MM_SSE2/3 flags for CPUs where this is generally not faster than 2010-07-19 22:38:23 +00:00
fdct_mmx.c Replace remaining uses of ATTR_ALIGNED with DECLARE_ALIGNED 2010-03-18 15:00:17 +00:00
fft_3dn2.c Remove DECLARE_ALIGNED_{8,16} macros 2010-03-06 14:24:59 +00:00
fft_3dn.c
fft_mmx.asm more credits to D. J. Bernstein for fft 2010-07-18 20:06:42 +00:00
fft_sse.c Move SSE optimized 32-point DCT to its own file. Should fix breakage with YASM 2010-07-06 17:48:23 +00:00
fft.c Move SSE optimized 32-point DCT to its own file. Should fix breakage with YASM 2010-07-06 17:48:23 +00:00
fft.h SSE optimized 32-point DCT 2010-07-06 16:58:54 +00:00
h264_deblock_sse2.asm relicense h264 deblock sse2 to lgpl 2010-07-22 00:39:49 +00:00
h264_i386.h Remove explicit filename from Doxygen @file commands. 2010-04-20 14:45:34 +00:00
h264_idct_sse2.asm Update x264asm header files to latest versions. 2010-06-23 19:20:46 +00:00
h264_intrapred.asm Fix h264/vp8 intra pred on Athlon XP 2010-07-01 10:29:47 +00:00
h264dsp_mmx.c Fix h264/vp8 intra pred on Athlon XP 2010-07-01 10:29:47 +00:00
idct_mmx_xvid.c Add some missing #includes 2010-03-06 22:36:36 +00:00
idct_mmx.c Fix compilation in x86_64. I broke it with r24580. 2010-07-29 22:45:21 +00:00
idct_sse2_xvid.c Remove explicit filename from Doxygen @file commands. 2010-04-20 14:45:34 +00:00
idct_xvid.h Remove explicit filename from Doxygen @file commands. 2010-04-20 14:45:34 +00:00
lpc_mmx.c
Makefile Convert deinterlacing MMX code to YASM 2010-07-31 14:50:51 +00:00
mathops.h Adding missing () to mathops.h. 2010-05-11 00:22:50 +00:00
mlpdsp.c
motion_est_mmx.c
mpegaudiodec_mmx.c Fix compilation on x64. 2010-06-24 08:53:32 +00:00
mpegvideo_mmx_template.c Remove DECLARE_ALIGNED_{8,16} macros 2010-03-06 14:24:59 +00:00
mpegvideo_mmx.c
rv40dsp_mmx.c Remove DECLARE_ALIGNED_{8,16} macros 2010-03-06 14:24:59 +00:00
simple_idct_mmx.c
snowdsp_mmx.c Separate DWT from snow and dsputil 2010-03-14 17:50:12 +00:00
vc1dsp_mmx.c Move ff_pw_* from vc1dsp_mmx.c to dsputil_mmx.c 2010-07-21 10:02:03 +00:00
vc1dsp_yasm.asm MMX/SSE VC1 loop filter 2010-07-11 22:53:01 +00:00
vp3dsp_mmx.c vp3: The DC-only IDCT is surprisingly not supposed to be bitexact to the 2010-05-28 07:01:34 +00:00
vp3dsp_mmx.h vp3: DC-only IDCT 2010-04-17 02:04:30 +00:00
vp3dsp_sse2.c Remove explicit filename from Doxygen @file commands. 2010-04-20 14:45:34 +00:00
vp3dsp_sse2.h
vp6dsp_mmx.c Remove explicit filename from Doxygen @file commands. 2010-04-20 14:45:34 +00:00
vp6dsp_mmx.h
vp6dsp_sse2.c Remove explicit filename from Doxygen @file commands. 2010-04-20 14:45:34 +00:00
vp6dsp_sse2.h
vp8dsp-init.c Use word-writing instead of dword-writing (with two cached but otherwise 2010-07-31 23:13:15 +00:00
vp8dsp.asm Use word-writing instead of dword-writing (with two cached but otherwise 2010-07-31 23:13:15 +00:00
vp56_arith.h Inline asm for VP56 arith coder 2010-07-23 21:46:30 +00:00
x86inc.asm sync yasm macros from x264 2010-07-21 22:45:16 +00:00
x86util.asm MMX/SSE VC1 loop filter 2010-07-11 22:53:01 +00:00