FFmpeg

Commit Graph

Author	SHA1	Message	Date
Ronald S. Bultje	b27b54de31	arm/h264pred: add missing argument type.	14 years ago
Oskar Arvidsson	19a0729b4c	Adds 8-, 9- and 10-bit versions of some of the functions used by the h264 decoder. This patch lets e.g. dsputil_init chose dsp functions with respect to the bit depth to decode. The naming scheme of bit depth dependent functions is <base name>_<bit depth>[_<prefix>] (i.e. the old clear_blocks_c is now named clear_blocks_8_c). Note: Some of the functions for high bit depth is not dependent on the bit depth, but only on the pixel size. This leaves some room for optimizing binary size. Preparatory patch for high bit depth h264 decoding support. Signed-off-by: Ronald S. Bultje <rsbultje@gmail.com>	14 years ago
Gavin Kinsey	25347c880f	Fix compilation.for iOS ARMv7.	14 years ago
Bill Pringlemeir	fccff6e83a	Allow h264pred_init_arm.c to compile. SOB: Bill Pringlemeir <bpringlemeir@yahoo.com> Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	14 years ago
Aurelien Jacobs	13d4ec844a	cosmetics: alignment	14 years ago
Oskar Arvidsson	8dbe585641	Adds 8-, 9- and 10-bit versions of some of the functions used by the h264 decoder. This patch lets e.g. dsputil_init chose dsp functions with respect to the bit depth to decode. The naming scheme of bit depth dependent functions is <base name>_<bit depth>[_<prefix>] (i.e. the old clear_blocks_c is now named clear_blocks_8_c). Note: Some of the functions for high bit depth is not dependent on the bit depth, but only on the pixel size. This leaves some room for optimizing binary size. Preparatory patch for high bit depth h264 decoding support. Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	14 years ago
pin xue	05c062e9da	replace movw instruction in ac3dsp_armv6.S AS libavcodec/arm/ac3dsp_armv6.o ffmpeg-src/libavcodec/arm/ac3dsp_armv6.S: Assembler messages: ffmpeg-src/libavcodec/arm/ac3dsp_armv6.S:40: Error: selected processor does not support `movw r8,#0x1fe0' make[1]: *** [libavcodec/arm/ac3dsp_armv6.o] Error 1 MOVW is ARMv7 way to load constant: * movw, or move wide, will move a 16-bit constant into a register, implicitly zeroing the top 16 bits of the target register. * movt, or move top, will move a 16-bit constant into the top half of a given register without altering the bottom 16 bits To load 32 bit constant, movw lower16; movt upper16; is better than ldr if available, because: While this approach takes two instructions, it does not require any extra space to store the constant so both the movw/movt method and the ldr method will end up using the same amount of memory. Memory bandwidth is precious in and the movw/movt approach avoids an extra read on the data side, not to mention the read could have missed the cache. But here it is armv6 optimization, so that we have to use ldr. Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	14 years ago
Mans Rullgard	5f2e6c0fd1	ac3enc: NEON optimised extract_exponents Signed-off-by: Mans Rullgard <mans@mansr.com>	14 years ago
Mans Rullgard	f7653904c8	ARM: NEON fixed-point forward MDCT Signed-off-by: Mans Rullgard <mans@mansr.com>	14 years ago
Mans Rullgard	dba9852935	ARM: NEON fixed-point FFT Signed-off-by: Mans Rullgard <mans@mansr.com>	14 years ago
Mans Rullgard	aa05f2126e	ac3enc: ARM optimised ac3_compute_matissa_size Signed-off-by: Mans Rullgard <mans@mansr.com>	14 years ago
Mans Rullgard	182826c884	ac3: armv6 optimised bit_alloc_calc_bap Signed-off-by: Mans Rullgard <mans@mansr.com>	14 years ago
Mans Rullgard	d782bca415	ac3enc: NEON optimised float_to_fixed24 Signed-off-by: Mans Rullgard <mans@mansr.com>	14 years ago
Michael Niedermayer	34c27ada10	Revert some silly renamings that leaked in from a pull.	14 years ago
Mans Rullgard	d743065e18	ARM: fix ff_apply_window_int16_neon() prototype The length argument should be unsigned. No change in code. Signed-off-by: Mans Rullgard <mans@mansr.com>	14 years ago
Mans Rullgard	2d3b21ffb9	ARM: NEON optimised apply_window_int16() Signed-off-by: Mans Rullgard <mans@mansr.com>	14 years ago
Mans Rullgard	245c78313f	ac3enc: NEON optimised shift functions	14 years ago
Mans Rullgard	f4855a904e	ac3enc: NEON optimised ac3_max_msb_abs_int16 and ac3_exponent_min	14 years ago
Mans Rullgard	0aded9484d	Move dct and rdft definitions to separate files This leaves fft.h with only the core FFT and MDCT definitions thus making it more managable. Signed-off-by: Mans Rullgard <mans@mansr.com>	14 years ago
Mans Rullgard	2912e87a6c	Replace FFmpeg with Libav in licence headers Signed-off-by: Mans Rullgard <mans@mansr.com>	14 years ago
Mans Rullgard	e9634db1dc	ARM: VP8: fix build on systems with global symbol prefix Signed-off-by: Mans Rullgard <mans@mansr.com> (cherry picked from commit `0b32da90f8`)	14 years ago
Mans Rullgard	cf9c227e58	ARM: fix vp8 neon with pic enabled The assembler emits literal pools too far from the load instructions, so we must do it explicitly at a suitable location. Signed-off-by: Mans Rullgard <mans@mansr.com> (cherry picked from commit `8b454c352f`)	14 years ago
Mans Rullgard	0b32da90f8	ARM: VP8: fix build on systems with global symbol prefix Signed-off-by: Mans Rullgard <mans@mansr.com>	14 years ago
Mans Rullgard	8b454c352f	ARM: fix vp8 neon with pic enabled The assembler emits literal pools too far from the load instructions, so we must do it explicitly at a suitable location. Signed-off-by: Mans Rullgard <mans@mansr.com>	14 years ago
Loren Merritt	11ab1e409f	FFT: factor a shuffle out of the inner loop and merge it into fft_permute. 6% faster SSE FFT on Conroe, 2.5% on Penryn. Signed-off-by: Janne Grunau <janne-ffmpeg@jannau.net> (cherry picked from commit `e6b1ed693a`)	14 years ago
Loren Merritt	e6b1ed693a	FFT: factor a shuffle out of the inner loop and merge it into fft_permute. 6% faster SSE FFT on Conroe, 2.5% on Penryn. Signed-off-by: Janne Grunau <janne-ffmpeg@jannau.net>	14 years ago
Mans Rullgard	4ae3ee4ae9	VP8: ARM optimised decode_block_coeffs_internal Approximately 5% faster on Cortex-A8. Signed-off-by: Mans Rullgard <mans@mansr.com> (cherry picked from commit `a7878c9f73`)	14 years ago
Mans Rullgard	5da7494dc5	ARM optimised vp56_rac_get_prob() Approximately 3% faster on Cortex-A8. Signed-off-by: Mans Rullgard <mans@mansr.com> (cherry picked from commit `7da48fd011`)	14 years ago
Mans Rullgard	a7878c9f73	VP8: ARM optimised decode_block_coeffs_internal Approximately 5% faster on Cortex-A8. Signed-off-by: Mans Rullgard <mans@mansr.com>	14 years ago
Mans Rullgard	7da48fd011	ARM optimised vp56_rac_get_prob() Approximately 3% faster on Cortex-A8. Signed-off-by: Mans Rullgard <mans@mansr.com>	14 years ago
Mans Rullgard	ef15d71c1f	VP8: ARM NEON optimisations for dsp functions This adds NEON optimised versions of all functions in VP8DSPContext. Based on initial work by Rob Clark. Signed-off-by: Mans Rullgard <mans@mansr.com> (cherry picked from commit `a1c1d3c003`)	14 years ago
Mans Rullgard	a1c1d3c003	VP8: ARM NEON optimisations for dsp functions This adds NEON optimised versions of all functions in VP8DSPContext. Based on initial work by Rob Clark. Signed-off-by: Mans Rullgard <mans@mansr.com>	14 years ago
Mans Rullgard	01b75fa931	ARM: add helper macro for declaring constant data Signed-off-by: Mans Rullgard <mans@mansr.com> (cherry picked from commit `b9a639ddd6`)	14 years ago
Justin Ruggles	fe2ff6d247	Separate format conversion DSP functions from DSPContext. This will be beneficial for use with the audio conversion API without requiring it to depend on all of dsputil. Signed-off-by: Mans Rullgard <mans@mansr.com> (cherry picked from commit `c73d99e672`)	14 years ago
Mans Rullgard	b9a639ddd6	ARM: add helper macro for declaring constant data Signed-off-by: Mans Rullgard <mans@mansr.com>	14 years ago
Justin Ruggles	c73d99e672	Separate format conversion DSP functions from DSPContext. This will be beneficial for use with the audio conversion API without requiring it to depend on all of dsputil. Signed-off-by: Mans Rullgard <mans@mansr.com>	14 years ago
Justin Ruggles	a8ae4e0e7b	Remove unneeded add bias from 3 functions. DSPContext.vector_fmul_window() DCADSPContext.lfe_fir() SynthFilterContext.synth_filter_float() Signed-off-by: Mans Rullgard <mans@mansr.com> (cherry picked from commit `80ba1ddb58`)	14 years ago
Justin Ruggles	80ba1ddb58	Remove unneeded add bias from 3 functions. DSPContext.vector_fmul_window() DCADSPContext.lfe_fir() SynthFilterContext.synth_filter_float() Signed-off-by: Mans Rullgard <mans@mansr.com>	14 years ago
Mans Rullgard	451b4b8635	Rearrange MpegEncContext to simplify access from asm This moves the fields needed by asm near the top, before any structs or other members which complicate the offset calculation. Modifying other structs will no longer require updating the offsets, and the asm code is slightly simpler due to the smaller offsets. Signed-off-by: Mans Rullgard <mans@mansr.com> (cherry picked from commit `d461a47317`)	14 years ago
Mans Rullgard	8afac88e14	ARM: update MpegEncContext offsets (cherry picked from commit `0745116c10`)	14 years ago
Mans Rullgard	d461a47317	Rearrange MpegEncContext to simplify access from asm This moves the fields needed by asm near the top, before any structs or other members which complicate the offset calculation. Modifying other structs will no longer require updating the offsets, and the asm code is slightly simpler due to the smaller offsets. Signed-off-by: Mans Rullgard <mans@mansr.com>	14 years ago
Mans Rullgard	0745116c10	ARM: update MpegEncContext offsets	14 years ago
Mans Rullgard	0fc1961ecc	ARM: NEON: fix overflow in h264 16x16 planar pred Signed-off-by: Mans Rullgard <mans@mansr.com> (cherry picked from commit `78f318be59`)	14 years ago
Mans Rullgard	78f318be59	ARM: NEON: fix overflow in h264 16x16 planar pred Signed-off-by: Mans Rullgard <mans@mansr.com>	14 years ago
Justin Ruggles	015f9f1ad3	Change DSPContext.vector_fmul() from dst=dstsrc to dest=src0src1. Signed-off-by: Mans Rullgard <mans@mansr.com> (cherry picked from commit `6eabb0d3ad`)	14 years ago
Justin Ruggles	0d8837bdda	Move lpc_compute_autocorr() from DSPContext to a new struct LPCContext. Signed-off-by: Mans Rullgard <mans@mansr.com> (cherry picked from commit `56f8952b25`)	14 years ago
Justin Ruggles	6eabb0d3ad	Change DSPContext.vector_fmul() from dst=dstsrc to dest=src0src1. Signed-off-by: Mans Rullgard <mans@mansr.com>	14 years ago
Justin Ruggles	56f8952b25	Move lpc_compute_autocorr() from DSPContext to a new struct LPCContext. Signed-off-by: Mans Rullgard <mans@mansr.com>	14 years ago
Janne Grunau	2c3589bfda	consolidate .gitignore patters into a single file Signed-off-by: Janne Grunau <janne-ffmpeg@jannau.net>	14 years ago
Janne Grunau	348b8218f7	convert svn:ignore properties to .gitignore files Signed-off-by: Janne Grunau <janne-ffmpeg@jannau.net>	14 years ago

1 2 3 4 5 ...

288 Commits (1125571b736b664a5ef079ec9e6f09640682eeda)