FFmpeg

mirror of https://git.ffmpeg.org/ffmpeg.git synced 2024-10-10 19:53:26 +00:00

Author	SHA1	Message	Date
Ronald S. Bultje	dc5eec8085	Use pextrw for SSE4 mbedge filter result writing, speedup 5-10cycles on CPUs supporting it. Originally committed as revision 24437 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-07-22 19:59:34 +00:00
Ronald S. Bultje	003243c3c2	Fix and enable horizontal >=SSE2 mbedge loopfilter. Originally committed as revision 24409 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-07-22 01:35:26 +00:00
Loren Merritt	c7b1d9768c	relicense h264 deblock sse2 to lgpl Originally committed as revision 24408 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-07-22 00:39:49 +00:00
Loren Merritt	532e769701	sync yasm macros from x264 Originally committed as revision 24406 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-07-21 22:45:16 +00:00
Jason Garrett-Glaser	8731dbd890	Eliminate one instruction in VP8 dc_add_sse4 Originally committed as revision 24405 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-07-21 22:41:37 +00:00
Jason Garrett-Glaser	7dd224a42d	Various VP8 x86 deblocking speedups SSSE3 versions, improve SSE2 versions a bit. SSE2/SSSE3 mbedge h functions are currently broken, so explicitly disable them. Originally committed as revision 24403 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-07-21 22:11:03 +00:00
Jason Garrett-Glaser	b8b231b5dc	Make mmx VP8 WHT faster Avoid pextrw, since it's slow on many older CPUs. Now it doesn't require mmxext either. Originally committed as revision 24397 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-07-21 20:51:01 +00:00
David Conrad	af521abc28	Add header declarations for mmx/sse constants missing them Originally committed as revision 24381 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-07-21 10:02:07 +00:00
David Conrad	c7eec58170	Move ff_pw_* from vc1dsp_mmx.c to dsputil_mmx.c Should fix compilation with icc and should help prevent any future duplicates Originally committed as revision 24380 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-07-21 10:02:03 +00:00
Ronald S. Bultje	e9e456d850	VP8 MBedge loopfilter MMX/MMX2/SSE2 functions for both luma (width=16) and chroma (width=8). Originally committed as revision 24378 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-07-20 22:58:56 +00:00
Ronald S. Bultje	268821e76e	Chroma (width=8) inner loopfilter MMX/MMX2/SSE2 for VP8 decoder. Originally committed as revision 24377 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-07-20 22:04:18 +00:00
Ronald S. Bultje	c60ed66dbe	Revert r24339 (it causes fate failures on x86-64) - I'll figure out what's wrong with it tomorrow or so, then re-submit. Originally committed as revision 24341 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-07-19 23:57:09 +00:00
Ronald S. Bultje	6526976f0c	Remove FF_MM_SSE2/3 flags for CPUs where this is generally not faster than regular MMX code. Examples of this are the Core1 CPU. Instead, set a new flag, FF_MM_SSE2/3SLOW, which can be checked for particular SSE2/3 functions that have been checked specifically on such CPUs and are actually faster than their MMX counterparts. In addition, use this flag to enable particular VP8 and LPC SSE2 functions that are faster than their MMX counterparts. Based on a patch by Loren Merritt <lorenm AT u washington edu>. Originally committed as revision 24340 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-07-19 22:38:23 +00:00
Ronald S. Bultje	1878f685c0	Implement chroma (width=8) inner loopfilter MMX/MMX2/SSE2 functions. Originally committed as revision 24339 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-07-19 21:53:28 +00:00
Ronald S. Bultje	fb9bdf048c	Be more efficient with registers or stack memory. Saves 8/16 bytes stack for x86-32, or 2 MM registers on x86-64. Originally committed as revision 24338 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-07-19 21:45:36 +00:00
Ronald S. Bultje	3facfc99da	Change function prototypes for width=8 inner and mbedge loopfilter functions so that it does both U and V planes at the same time. This will have speed advantages when using SSE2 (or higher) optimizations, since we can do both the U and V rows together in a single xmm register. This also renames filter16 to filter16y and filter8 to filter8uv so that it's more obvious what each function is used for. Originally committed as revision 24337 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-07-19 21:18:04 +00:00
Loren Merritt	1ee076b1b1	more credits to D. J. Bernstein for fft Originally committed as revision 24308 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-07-18 20:06:42 +00:00
Ronald S. Bultje	819b2dd2b1	Attempt to fix x86-64 testsuite on fate. Originally committed as revision 24275 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-07-16 21:35:30 +00:00
Ronald S. Bultje	6f323f1251	Remove duplicate define. Originally committed as revision 24272 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-07-16 19:54:47 +00:00
Ronald S. Bultje	889b2c26ee	Revert 24270, it contained some stuff that shouldn't have been in there. Originally committed as revision 24271 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-07-16 19:54:25 +00:00
Ronald S. Bultje	2356a7834b	Remove duplicate define. Originally committed as revision 24270 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-07-16 19:42:32 +00:00
Ronald S. Bultje	ede1b9665a	Give x86 r%d registers names, this will simplify implementation of the chroma inner loopfilter, and it also allows us to save one register on x86-64/sse2. Originally committed as revision 24269 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-07-16 19:38:10 +00:00
Ronald S. Bultje	526e831a46	Change return statement, the REP_RET is a mistake since the else case (x86-64, sse2) doesn't actually loop, so REP_RET isn't necessary. Originally committed as revision 24268 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-07-16 18:29:14 +00:00
Ronald S. Bultje	a711eb4829	VP8 H/V inner loopfilter MMX/MMXEXT/SSE2 optimizations. Originally committed as revision 24250 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-07-15 23:02:34 +00:00
David Conrad	faa26db28b	MMX/SSE VC1 loop filter Originally committed as revision 24208 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-07-11 22:53:01 +00:00
David Conrad	7af8fbd348	Make ff_pw_4 128 bits Originally committed as revision 24207 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-07-11 22:52:55 +00:00
Vitor Sessak	881fd7a62f	Move SSE optimized 32-point DCT to its own file. Should fix breakage with YASM disabled. Originally committed as revision 24078 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-07-06 17:48:23 +00:00
Vitor Sessak	4dcc4f8eaa	SSE optimized 32-point DCT Originally committed as revision 24077 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-07-06 16:58:54 +00:00
Ronald S. Bultje	f2a30bd840	Simple H/V loopfilter for VP8 in MMX, MMX2 and SSE2 (yay for yasm macros). Originally committed as revision 24029 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-07-03 19:26:30 +00:00
Jason Garrett-Glaser	b06855f18a	SSSE3 versions of vp8 width4 bilinear MC functions Originally committed as revision 24013 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-07-03 00:48:12 +00:00
Jason Garrett-Glaser	dcc602d802	SSSE3 versions of width4 VP8 6-tap MC functions Also make some small changes to saturation order of 4-tap SSSE3 MC to fix a non-bitexactness bug. Patch mostly by Eli Friedman <eli.friedman AT gmail DOT com>. Originally committed as revision 23965 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-07-02 05:27:41 +00:00
Jason Garrett-Glaser	8434fc26eb	Fix 100L in vp8dsp asm init Originally committed as revision 23946 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-07-01 22:09:22 +00:00
Jason Garrett-Glaser	17dc7c7a60	Fix h264/vp8 intra pred on Athlon XP Whose idea was it to have a CPU that didn't SIGILL on an invalid instruction? Originally committed as revision 23927 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-07-01 10:29:47 +00:00
Måns Rullgård	49bd8e4b84	Fix grammar errors in documentation Originally committed as revision 23904 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-06-30 15:38:06 +00:00
Jason Garrett-Glaser	82a8d0f114	Use add instead of lshift in mmxext vp8 idct Originally committed as revision 23891 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-06-29 17:23:17 +00:00
Ronald S. Bultje	565344e7e4	Remove unused macros (duplicates from the now-LGPL x86util.asm). Originally committed as revision 23890 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-06-29 17:04:29 +00:00
Ronald S. Bultje	2dd2f71692	MMX idct_add for VP8. Originally committed as revision 23886 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-06-29 14:43:11 +00:00
Jason Garrett-Glaser	29e719377f	Add missing mm_support call toff_h264_pred_init_x86. I'm not sure if this is supposed to be here, but it can't hurt. Originally committed as revision 23885 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-06-29 12:28:06 +00:00
Jason Garrett-Glaser	004cda8e79	Add mmxext version of VP8 DC Hadamard transform Originally committed as revision 23878 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-06-29 01:41:59 +00:00
Jason Garrett-Glaser	37355fe823	Make x86util.asm LGPL so we can use it in LGPL asm Strip out most x264-specific stuff (not used anywhere in ffmpeg). Originally committed as revision 23877 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-06-29 00:40:12 +00:00
Jason Garrett-Glaser	bc14f04b2f	MMXEXT version of vp8 4x4 vertical pred Originally committed as revision 23876 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-06-29 00:23:52 +00:00
Jason Garrett-Glaser	fb9927ad7d	Add mmx/mmxext/ssse3 4x4 TM intra pred functions for vp8 Originally committed as revision 23875 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-06-28 23:53:07 +00:00
Jason Garrett-Glaser	8b746bb473	Add missing comment header for predict_4x4_dc_mmxext Originally committed as revision 23874 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-06-28 23:37:24 +00:00
Jason Garrett-Glaser	270a85d259	Fix some intra pred MMX functions that used MMXEXT instructions Also add predict_4x4_dc MMXEXT function for vp8/h264. Originally committed as revision 23873 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-06-28 23:35:17 +00:00
Jason Garrett-Glaser	a912da761d	Fix VP8 bilinear mc on x86_64 Originally committed as revision 23872 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-06-28 22:13:14 +00:00
Baptiste Coudurier	50f70541d3	Change MMXEXT to MMX2, MMXEXT is deprecated Originally committed as revision 23865 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-06-28 21:12:00 +00:00
Jason Garrett-Glaser	0fecad09fe	Add x86 asm functions for VP8 put_pixels Originally committed as revision 23858 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-06-28 19:14:40 +00:00
Jason Garrett-Glaser	a173aa8940	Add MMX, SSE2, SSSE3 asm for VP8 bilinear MC Originally committed as revision 23857 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-06-28 18:56:24 +00:00
Måns Rullgård	1f65b67c46	Fix x86 build with h264dsp disabled Originally committed as revision 23844 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-06-28 10:02:15 +00:00
Eli Friedman	b3858964d6	Add const to some pointer parameters. Patch by Eli Friedman, eli D friedman A gmail Originally committed as revision 23826 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-06-27 15:11:38 +00:00
David Conrad	30bdefd1de	Fix build without yasm Originally committed as revision 23816 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-06-27 02:52:43 +00:00
Jason Garrett-Glaser	0178d14fe5	First shot at VP8 optimizations: - MMXEXT, SSE2 and SSSE3 MC functions - MMX and SSE4 IDCT dc_add functions Patch by Jason Garrett-Glaser <darkshikari gmail com> and myself. Originally committed as revision 23815 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-06-27 02:01:45 +00:00
Måns Rullgård	0912db0206	Make vp8 select h264dsp and use this to pull in mmx intrapred Originally committed as revision 23790 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-06-25 19:10:08 +00:00
Carl Eugen Hoyos	0c59074868	Fix compilation without --enable-gpl. Originally committed as revision 23789 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-06-25 19:06:29 +00:00
Carl Eugen Hoyos	96da2a6967	Cosmetics: Fix indentation. Originally committed as revision 23785 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-06-25 18:34:03 +00:00
Jason Garrett-Glaser	4af8cdfc3f	16x16 and 8x8c x86 SIMD intra pred functions for VP8 and H.264 Originally committed as revision 23783 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-06-25 18:25:49 +00:00
Vitor Sessak	89c7d8058c	Fix compilation on x64. Originally committed as revision 23753 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-06-24 08:53:32 +00:00
Vitor Sessak	57dbd12b6d	Fix asm constraints in apply_window() Originally committed as revision 23752 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-06-24 08:46:47 +00:00
Vitor Sessak	bc2b368215	SSE-optimized MP3 floating point windowing functions Originally committed as revision 23750 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-06-24 07:44:50 +00:00
Jason Garrett-Glaser	2966cc1849	Update x264asm header files to latest versions. Modify the asm accordingly. GLOBAL is now no longoer necessary for PIC-compliant loads. Originally committed as revision 23739 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-06-23 19:20:46 +00:00
David Conrad	413abbe164	Add bitexact versions of put_no_rnd_pixels8 _x2 and _y2 for vp3/theora Originally committed as revision 23463 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-06-04 04:46:26 +00:00
David Conrad	179655b6c6	vp3: The DC-only IDCT is surprisingly not supposed to be bitexact to the full IDCT. Fix this. Originally committed as revision 23358 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-05-28 07:01:34 +00:00
Michael Niedermayer	22cb6fb60f	Adding missing () to mathops.h. Originally committed as revision 23083 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-05-11 00:22:50 +00:00
Reimar Döffinger	1c71b5c89a	Replace more "m" constraints with MANGLE to fix compilation issues with x86_32 gcc 4.4.4 and -fPIC. Originally committed as revision 23082 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-05-10 21:16:08 +00:00
Diego Biurrun	ba87f0801d	Remove explicit filename from Doxygen @file commands. Passing an explicit filename to this command is only necessary if the documentation in the @file block refers to a file different from the one the block resides in. Originally committed as revision 22921 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-04-20 14:45:34 +00:00
David Conrad	eb6a6cd788	vp3: DC-only IDCT 2-4% faster overall decode Originally committed as revision 22896 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-04-17 02:04:30 +00:00
Reimar Döffinger	27eecec359	Convert two "m" constraints to MANGLE to fix compilation with some compilers. Originally committed as revision 22760 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-04-01 16:52:14 +00:00
Måns Rullgård	d343d59837	Replace remaining uses of ATTR_ALIGNED with DECLARE_ALIGNED Originally committed as revision 22593 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-03-18 15:00:17 +00:00
Måns Rullgård	3bd74e9243	Simplify arch-specific object file lists Originally committed as revision 22570 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-03-16 21:23:03 +00:00
Måns Rullgård	43f60eba19	Move arch-specific makefile parts into $arch/Makefile Originally committed as revision 22569 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-03-16 21:22:59 +00:00
Måns Rullgård	4693b031a3	Move H264 dsputil functions into their own struct This moves the H264-specific functions from DSPContext to the new H264DSPContext. The code is made conditional on CONFIG_H264DSP which is set by the codecs requiring it. The qpel and chroma MC functions are not moved as these are used by non-h264 code. Originally committed as revision 22565 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-03-16 01:17:00 +00:00
Måns Rullgård	05aec7bb87	Separate DWT from snow and dsputil This moves the DWT functions from snow.c and dsputil.c to a file of their own. A new struct, DWTContext, holds the function pointers previously part of DSPContext. Originally committed as revision 22522 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-03-14 17:50:12 +00:00
Måns Rullgård	f49747e904	x86: move function prototypes to header files Originally committed as revision 22266 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-03-06 22:37:08 +00:00
Måns Rullgård	c26e58e32c	Add some missing #includes Originally committed as revision 22258 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-03-06 22:36:36 +00:00
Måns Rullgård	1429224b04	Move FFT parts from dsputil.h to fft.h Originally committed as revision 22235 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-03-06 14:34:46 +00:00
Måns Rullgård	84dc2d8afa	Remove DECLARE_ALIGNED_{8,16} macros These macros are redundant. All uses are replaced with the generic DECLARE_ALIGNED macro instead. Originally committed as revision 22233 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-03-06 14:24:59 +00:00
Måns Rullgård	5e46be96f8	Move NEG_[US]SR32 macros to mathops.h Originally committed as revision 21873 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-02-17 23:58:59 +00:00
David Conrad	19530266a5	Enable SSE2 (put\|avg)_pixels_16_sse2 SVQ1 chroma has been special-cased aligned to 16-bytes since at least r15466 Other architectures also assume 16-byte alignment here too but set STRIDE_ALIGN to 16. Originally committed as revision 21736 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-02-10 02:02:06 +00:00
Reimar Döffinger	3d05c1fbec	Make the jump-table section-relative for x86_64 with PIC enabled. This allows to get rid of the macho64 specific hack that moves them to rodata (with worse cache behaviour) and avoids textrels which e.g. Gentoo does not allow for x86_64 libraries. Originally committed as revision 21551 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-01-30 19:26:47 +00:00
Loren Merritt	900479bb74	optimize h264_loop_filter_strength_mmx2 244->160 cycles on core2 Originally committed as revision 21462 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-01-26 17:17:48 +00:00
Alex Converse	3deb53849e	Implement an sse version of scalarproduct_float(). Originally committed as revision 21386 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-01-22 23:07:58 +00:00
Måns Rullgård	c67278098d	Move array specifiers outside DECLARE_ALIGNED() invocations Originally committed as revision 21377 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-01-22 03:25:11 +00:00
David Conrad	1f630b9717	Use two separate memory arguments since 8+() is invalid gas syntax Originally committed as revision 21360 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-01-21 09:46:57 +00:00
Michael Niedermayer	b4c2ada528	Attempt to fix asm compilation failure. Only tested on gcc 4 & x86_64. Originally committed as revision 21355 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-01-20 19:23:19 +00:00
Måns Rullgård	5e7dfb7de1	Move COPY3_IF_LT to lavc/mathops.h This obscure macro is only used in motion_est.c so having it in lavc makes more sense. See discussion here: http://lists.mplayerhq.hu/pipermail/ffmpeg-devel/2008-November/056561.html Originally committed as revision 21346 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-01-20 06:01:54 +00:00
David Conrad	c4f2b6dce3	Use constant offsets for memory operands since gcc is unable to This fixes gcc failing to fit 6 memory locations into 7 registers on x86-32 Originally committed as revision 21337 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-01-20 00:34:10 +00:00
Michael Niedermayer	9ac4548ff7	Fix h264_loop_filter_strength_mmx2() so it works with b frames. Originally committed as revision 21327 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-01-19 16:40:36 +00:00
Michael Niedermayer	ebddd2e253	Remove -2 -> -1 remapping, its not needed anymore as we must remap all references per LUT anyway. Originally committed as revision 21323 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-01-19 14:28:19 +00:00
Gwenole Beauchesne	5716aec3f9	Fix XvMC. XvMCCreateBlocks() may not allocate 16-byte aligned blocks, so we can't use SSE-optimized routines. Originally committed as revision 21011 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-01-04 09:19:32 +00:00
Reimar Döffinger	4a1289450a	Reduce number of ASM constraints for ff_lpc_compute_autocorr_sse2 since it causes no significant speed difference and can avoid compilation issues with --enable-pic. Originally committed as revision 21003 to svn://svn.ffmpeg.org/ffmpeg/trunk	2010-01-02 17:48:08 +00:00
Diego Biurrun	4052cbf161	Get rid of pointless CONFIG_ANY_H263 preprocessor definition. Originally committed as revision 20975 to svn://svn.ffmpeg.org/ffmpeg/trunk	2009-12-30 11:33:59 +00:00
Loren Merritt	758c7455f1	fix a crash in ape decoding on x86_32 sse2 Originally committed as revision 20777 to svn://svn.ffmpeg.org/ffmpeg/trunk	2009-12-08 21:24:01 +00:00
Loren Merritt	a4605efdf5	slightly faster scalarproduct_and_madd_int16_ssse3 on penryn, no change on conroe Originally committed as revision 20743 to svn://svn.ffmpeg.org/ffmpeg/trunk	2009-12-05 17:53:11 +00:00
Loren Merritt	91e644ff77	r20739 broke compilation on systems without yasm Originally committed as revision 20742 to svn://svn.ffmpeg.org/ffmpeg/trunk	2009-12-05 17:51:57 +00:00
Loren Merritt	b1159ad928	refactor and optimize scalarproduct 29-105% faster apply_filter, 6-90% faster ape decoding on core2 (Any x86 other than core2 probably gets much less, since this is mostly due to ssse3 cachesplit avoidance and I haven't written the full gamut of other cachesplit modes.) 9-123% faster ape decoding on G4. Originally committed as revision 20739 to svn://svn.ffmpeg.org/ffmpeg/trunk	2009-12-05 15:09:10 +00:00
Loren Merritt	b10fa1bb8b	port ape dsp functions from sse2 to mmx now requires yasm Originally committed as revision 20722 to svn://svn.ffmpeg.org/ffmpeg/trunk	2009-12-03 18:53:12 +00:00
Loren Merritt	4521308363	s/movdqa/movaps/ in sse1 fft. (regression in r20293) Originally committed as revision 20371 to svn://svn.ffmpeg.org/ffmpeg/trunk	2009-10-25 03:09:53 +00:00
Loren Merritt	b07781b6e4	fix linking on systems with a function name prefix (10l in r20287) Originally committed as revision 20294 to svn://svn.ffmpeg.org/ffmpeg/trunk	2009-10-18 21:44:03 +00:00
Loren Merritt	29e4edbbe7	sync yasm macros to x264 Originally committed as revision 20293 to svn://svn.ffmpeg.org/ffmpeg/trunk	2009-10-18 21:42:28 +00:00
Loren Merritt	e17ccf60fe	huffyuv: add some const qualifiers Originally committed as revision 20290 to svn://svn.ffmpeg.org/ffmpeg/trunk	2009-10-18 20:47:25 +00:00

1 2 3 4 5

213 Commits