FFmpeg

mirror of https://github.com/FFmpeg/FFmpeg.git synced 2024-12-23 12:43:46 +02:00

Author	SHA1	Message	Date
James Almer	70d685a77f	x86: use the new helper macros where useful Reviewed-by: Michael Niedermayer <michael@niedermayer.cc> Signed-off-by: James Almer <jamrial@gmail.com>	2016-02-14 20:00:21 -03:00
Timothy Gu	180f9a0958	all: Make header guard names consistent	2016-01-31 15:44:11 -08:00
Ganesh Ajjanagadde	26937fb416	swr/resample: use av_clip_int16 instead of av_clip Reviewed-by: Michael Niedermayer <michael@niedermayer.cc> Signed-off-by: Ganesh Ajjanagadde <gajjanagadde@gmail.com>	2015-12-24 11:29:52 -08:00
Clément Bœsch	c1f114a8c4	swresample: use AV_OPT_TYPE_BOOL for linear_interp and cheby options	2015-12-04 15:43:33 +01:00
Ganesh Ajjanagadde	0bd0af6e68	swresample/resample: remove redundant L for floating literal It is inherently double precision, and 1.0 is perfectly represented anyway. Signed-off-by: Ganesh Ajjanagadde <gajjanagadde@gmail.com>	2015-11-15 10:26:43 -05:00
Michael Niedermayer	351e625d60	swresample/resample: increase precision for compensation Signed-off-by: Michael Niedermayer <michael@niedermayer.cc>	2015-11-11 18:17:18 +01:00
Ganesh Ajjanagadde	cf491a925e	swresample/resample: speed up Blackman Nuttall filter This may be a slightly surprising optimization, but is actually based on an understanding of how math libraries compute trigonometric functions. Explanation is given here so that future development uses libm more effectively across the codebase. All libm's essentially compute transcendental functions via some kind of polynomial approximation, be it Taylor-Maclaurin or Chebyshev. Correction terms are added via polynomial correction factors when needed to squeeze out the last bits of accuracy. Lookup tables are also inserted strategically. In the case of trigonometric functions, periodicity is exploited via first doing a range reduction to an interval around zero, and then using some polynomial approximation. This range reduction is the most natural way of doing things - else one would need polynomials for ranges in different periods which makes no sense whatsoever. To avoid the need for the range reduction, it is helpful to feed in arguments as close to the origin as possible for the trigonometric functions. In fact, this also makes sense from an accuracy point of view: IEEE floating point has far more resolution for small numbers than big ones. This patch does this for the Blackman-Nuttall filter, and yields a non-negligible speedup. Sample benchmark (x86-64, Haswell, GNU/Linux) test: fate-swr-resample-dblp-2626-44100 old: 18893514 decicycles in build_filter (loop 1000), 256 runs, 0 skips 18599863 decicycles in build_filter (loop 1000), 512 runs, 0 skips 18445574 decicycles in build_filter (loop 1000), 1000 runs, 24 skips new: 16290697 decicycles in build_filter (loop 1000), 256 runs, 0 skips 16267172 decicycles in build_filter (loop 1000), 512 runs, 0 skips 16251105 decicycles in build_filter (loop 1000), 1000 runs, 24 skips Reviewed-by: Michael Niedermayer <michael@niedermayer.cc> Signed-off-by: Ganesh Ajjanagadde <gajjanagadde@gmail.com>	2015-11-09 18:41:32 -05:00
Ganesh Ajjanagadde	b87ca4bf25	swresample/resample: speed up upsampling by precomputing sines When upsampling, factor is set to 1 and sines need to be evaluated only once for each phase, and the complexity should not depend on the number of filter taps. This does the desired precomputation, yielding significant speedups. Hard guarantees on the gain are not possible, but gains themselves are obvious and are illustrated below. Sample benchmark (x86-64, Haswell, GNU/Linux) test: fate-swr-resample-dblp-2626-44100 old: 29161085 decicycles in build_filter (loop 1000), 256 runs, 0 skips 28821467 decicycles in build_filter (loop 1000), 512 runs, 0 skips 28668201 decicycles in build_filter (loop 1000), 1000 runs, 24 skips new: 14351936 decicycles in build_filter (loop 1000), 256 runs, 0 skips 14306652 decicycles in build_filter (loop 1000), 512 runs, 0 skips 14299923 decicycles in build_filter (loop 1000), 1000 runs, 24 skips Note that this does not statically allocate the sin lookup table. This may be done for the default 1024 phases, yielding a 512*8 = 4kB array which should be small enough. This should yield a small improvement. Nevertheless, this is separate from this patch, is more ambiguous due to the binary increase, and requires a lut to be generated offline. Reviewed-by: Michael Niedermayer <michael@niedermayer.cc> Signed-off-by: Ganesh Ajjanagadde <gajjanagadde@gmail.com>	2015-11-09 18:41:03 -05:00
Ganesh Ajjanagadde	a5202bc968	swresample/resample: improve bessel function accuracy and speed This improves accuracy for the bessel function at large arguments, and this in turn should improve the quality of the Kaiser window. It also improves the performance of the bessel function and hence build_filter by ~ 20%. Details are given below. Algorithm: taken from the Boost project, who have done a detailed investigation of the accuracy of their method, as compared with e.g the GNU Scientific Library (GSL): http://www.boost.org/doc/libs/1_52_0/libs/math/doc/sf_and_dist/html/math_toolkit/special/bessel/mbessel.html. Boost source code (also cited and licensed in the code): https://searchcode.com/codesearch/view/14918379/. Accuracy: sample values may be obtained as follows. i0 denotes the old bessel code, i0_boost the approach here, and i0_real an arbitrary precision result (truncated) from Wolfram Alpha: type "bessel i0(6.0)" to reproduce. These are evaluation points that occur for the default kaiser_beta = 9. Some illustrations: bessel(8.0) i0 (8.000000) = 427.564115721804739678191254 i0_boost(8.000000) = 427.564115721804796521610115 i0_real (8.000000) = 427.564115721804785177396791 bessel(6.0) i0 (6.000000) = 67.234406976477956163762428 i0_boost(6.000000) = 67.234406976477970374617144 i0_real (6.000000) = 67.234406976477975326188025 Reason for accuracy: Main accuracy benefits come at larger bessel arguments, where the Taylor-Maclaurin method is not that good: 23+ iterations (at large arguments, since the series is about 0) can cause significant floating point error accumulation. Benchmarks: Obtained on x86-64, Haswell, GNU/Linux via a loop calling build_filter 1000 times: test: fate-swr-resample-dblp-44100-2626 new: 995894468 decicycles in build_filter(loop 1000), 256 runs, 0 skips 1029719302 decicycles in build_filter(loop 1000), 512 runs, 0 skips 984101131 decicycles in build_filter(loop 1000), 1024 runs, 0 skips old: 1250020763 decicycles in build_filter(loop 1000), 256 runs, 0 skips 1246353282 decicycles in build_filter(loop 1000), 512 runs, 0 skips 1220017565 decicycles in build_filter(loop 1000), 1024 runs, 0 skips A further ~ 5% may be squeezed by enabling -ftree-vectorize. However, this is a separate issue from this patch. Reviewed-by: Michael Niedermayer <michael@niedermayer.cc> Signed-off-by: Ganesh Ajjanagadde <gajjanagadde@gmail.com>	2015-11-08 21:18:16 -05:00
Ganesh Ajjanagadde	1bed09a30e	swresample: allow double precision beta value for the Kaiser window Kaiser windows inherently don't require beta to be an integer. This was an arbitrary restriction. Moreover, soxr does not require it, and in fact often estimates beta to a non-integral value. Thus, this patch allows greater flexibility for swresample clients. Micro version is updated. Reviewed-by: Derek Buitenhuis <derek.buitenhuis@gmail.com> Reviewed-by: Michael Niedermayer <michael@niedermayer.cc> Signed-off-by: Ganesh Ajjanagadde <gajjanagadde@gmail.com>	2015-11-08 21:11:07 -05:00
Ganesh Ajjanagadde	c8780822ba	swresample/resample: speed up build_filter for Blackman-Nuttall filter This uses the trigonometric double and triple angle formulae to avoid repeated (expensive) evaluation of libc's cos(). Sample benchmark (x86-64, Haswell, GNU/Linux) test: fate-swr-resample-dblp-44100-2626 old: 1104466600 decicycles in build_filter(loop 1000), 256 runs, 0 skips 1096765286 decicycles in build_filter(loop 1000), 512 runs, 0 skips 1070479590 decicycles in build_filter(loop 1000), 1024 runs, 0 skips new: 588861423 decicycles in build_filter(loop 1000), 256 runs, 0 skips 591262754 decicycles in build_filter(loop 1000), 512 runs, 0 skips 577355145 decicycles in build_filter(loop 1000), 1024 runs, 0 skips This results in small differences with the old expression: difference (worst case on [0, 2*M_PI]), argmax 0.008: max diff (relative): 0.000000000000157289807188 blackman_old(0.008): 0.000363951585488813192382 blackman_new(0.008): 0.000363951585488755946507 These are judged to be insignificant for the performance gain. PSNR to reference file is unchanged up to second decimal point for instance. Reviewed-by: Michael Niedermayer <michael@niedermayer.cc> Signed-off-by: Ganesh Ajjanagadde <gajjanagadde@gmail.com>	2015-11-05 21:52:40 -05:00
Ganesh Ajjanagadde	9bec6d71a2	swresample/resample: speed up build_filter by 50% This speeds up build_filter by ~ 50%. This gain should be pretty consistent across all architectures and platforms. Essentially, this relies on a observation that the filters have some even/odd symmetry that may be exploited during the construction of the polyphase filter bank. In particular, phases (scaled to [0, 1]) in [0.5, 1] are easily derived from [0, 0.5] and expensive reevaluation of function points are unnecessary. This requires some rather annoying even/odd bookkeeping as can be seen from the patch. I vaguely recall from signal processing theory more general symmetries allowing even greater optimization of the construction. At a high level, "even functions" correspond to 2, and one can imagine variations. Nevertheless, for the sake of some generality and because of existing filters, this is all that is being exploited. Currently, this patch relies on phase_count being even or (trivially) 1, though this is not an inherent limitation to the approach. This assumption is safe as phase_count is 1 << phase_bits, and is hence a power of two. There is no way for user API to set it to a nontrivial odd number. This assumption has been placed as an assert in the code. To repeat, this assumes even symmetry of the filters, which is the most common way to get generalized linear phase anyway and is true of all currently supported filters. As a side note, accuracy should be identical or perhaps slightly better due to this "forcing" filter symmetries leading to a better phase characteristic. As before, I can't test this claim easily, though it may be of interest. Patch tested with FATE. Sample benchmark (x86-64, Haswell, GNU/Linux): test: swr-resample-dblp-44100-2626 new: 527376779 decicycles in build_filter(loop 1000), 256 runs, 0 skips 524361765 decicycles in build_filter(loop 1000), 512 runs, 0 skips 516552574 decicycles in build_filter(loop 1000), 1024 runs, 0 skips old: 974178658 decicycles in build_filter(loop 1000), 256 runs, 0 skips 972794408 decicycles in build_filter(loop 1000), 512 runs, 0 skips 954350046 decicycles in build_filter(loop 1000), 1024 runs, 0 skips Note that lower level optimizations are entirely possible, I focussed on getting the high level semantics correct. In any case, this should provide a good foundation. Reviewed-by: Michael Niedermayer <michael@niedermayer.cc> Signed-off-by: Ganesh Ajjanagadde <gajjanagadde@gmail.com>	2015-11-04 17:05:57 -05:00
wm4	80580bb240	swr: do not reject channel layouts that use channel 63 Channel layouts are essentially uint64_t, and every value is valid.	2015-10-28 19:25:49 +01:00
Ganesh Ajjanagadde	c7131762c0	all: add const-correctness to qsort comparators This adds const-correctness when needed for the comparators. Reviewed-by: Ronald S. Bultje <rsbultje@gmail.com> Signed-off-by: Ganesh Ajjanagadde <gajjanagadde@gmail.com>	2015-10-25 10:07:20 -04:00
Ganesh Ajjanagadde	8507b98c10	avfilter,swresample,swscale: use fabs, fabsf instead of FFABS It is well known that fabs and fabsf are at least as fast and sometimes faster than the FFABS macro, at least on the gcc+glibc combination. For instance, see the reference: http://patchwork.sourceware.org/patch/6735/. This was a patch to glibc in order to remove their usages of a macro. The reason essentially boils down to fabs using the __builtin_fabs of the compiler, while FFABS needs to infer to not use a branch and to simply change the sign bit. Usually the inference works, but sometimes it does not. This may be easily checked by looking at the asm. This also has the added benefit of reducing macro usage, which has problems with side-effects. Note that avcodec is not handled here, as it is huge and most things there are integer arithmetic anyway. Tested with FATE. Reviewed-by: Clément Bœsch <u@pkh.me> Signed-off-by: Ganesh Ajjanagadde <gajjanagadde@gmail.com>	2015-10-22 16:13:26 -04:00
Ganesh Ajjanagadde	ef62f573ca	swresample/swresample_internal: add av_warn_unused_result This will trigger a few warnings that need to be fixed. Reviewed-by: Michael Niedermayer <michael@niedermayer.cc> Signed-off-by: Ganesh Ajjanagadde <gajjanagadde@gmail.com>	2015-10-15 22:27:23 -04:00
wm4	cdf4a13f86	swresample: slightly nicer debug output for auto matrix This is the matrix that will be used for up/downmixing.	2015-10-15 20:16:13 +02:00
Ganesh Ajjanagadde	f3fc103c6a	doc/resampler, swresample/options: use proper capitalization Proper names should be capitalized in all user facing API as far as possible. The option names themselves have not been changed since: 1. We consistently keep option names in lower case. 2. Changing them would break existing scripts. 3. I suspect that we want to be similar to Sox and its relevant options. The converse is also true: improper names should not be capitalized generally. Signed-off-by: Ganesh Ajjanagadde <gajjanagadde@gmail.com> Signed-off-by: Michael Niedermayer <michael@niedermayer.cc>	2015-10-10 20:49:54 +02:00
Michael Niedermayer	1bc873acd6	swresample/resample: manually unroll the main loop in bessel() About 10% faster Signed-off-by: Michael Niedermayer <michael@niedermayer.cc>	2015-10-07 18:00:58 +02:00
Michael Niedermayer	6024c865ef	swresample/resample: merge first iteration into init in bessel() speedup of about 1% Signed-off-by: Michael Niedermayer <michael@niedermayer.cc>	2015-10-07 17:33:00 +02:00
James Almer	acdd672506	x86/audio_convert: fix clobbering of xmm registers Reviewed-by: Michael Niedermayer <michaelni@gmx.at> Signed-off-by: James Almer <jamrial@gmail.com>	2015-10-01 22:40:50 -03:00
Michael Niedermayer	7d636d02b1	swresample/dither_template: Add missing license header Signed-off-by: Michael Niedermayer <michael@niedermayer.cc>	2015-09-27 13:09:10 +02:00
Hendrik Leppkes	160e92c8bf	Merge commit 'e88103a7f92cf27a2868b50acc8a9912f6088249' * commit 'e88103a7f92cf27a2868b50acc8a9912f6088249': Bump major versions of all libraries Merged-by: Hendrik Leppkes <h.leppkes@gmail.com>	2015-09-05 21:35:46 +02:00
Michael Niedermayer	32f53958b8	swresample/swresample: Fix integer overflow in seed calculation Fixes CID1322333 Signed-off-by: Michael Niedermayer <michael@niedermayer.cc>	2015-09-03 09:32:43 +02:00
Michael Niedermayer	fb42e77516	swresample/swresample-test: Make layouts static const Signed-off-by: Michael Niedermayer <michael@niedermayer.cc>	2015-08-30 13:10:11 +02:00
Ganesh Ajjanagadde	24e6729a04	swresample/dither: use integer arithmetic This fixes a -Wabsolute-value reported by clang 3.5+ complaining about misuse of fabs() for integer absolute value. An additional benefit is the removal of floating point calculations. Signed-off-by: Ganesh Ajjanagadde <gajjanagadde@gmail.com> Signed-off-by: Michael Niedermayer <michael@niedermayer.cc>	2015-08-23 23:19:31 +02:00
James Almer	5750d6c5e9	x86: move XOP emulation code back to x86inc Only two functions that use xop multiply-accumulate instructions where the first operand is the same as the fourth actually took advantage of the macros. This further reduces differences with x264's x86inc. Reviewed-by: Ronald S. Bultje <rsbultje@gmail.com> Signed-off-by: James Almer <jamrial@gmail.com>	2015-08-03 17:11:13 -03:00
James Almer	f37a5dcb55	swresample/x86: add missing colon to labels Silences warnings with Nasm Signed-off-by: James Almer <jamrial@gmail.com>	2015-07-26 02:51:13 -03:00
Carl Eugen Hoyos	a77401e1f7	lswr: Allow 64 channels internally.	2015-07-17 00:17:08 +02:00
Michael Niedermayer	d4325b2fea	swr: Remember previously set int_sample_format from user Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2015-06-22 18:34:39 +02:00
Michael Niedermayer	0dd2790df5	swresample/swresample: Clear delayed_samples_fixup in clear_context() This probably makes no difference but its more proper Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2015-06-22 18:34:39 +02:00
Rob Sykes	c70c6be225	swresample: soxr implementation for swr_get_out_samples() Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2015-06-21 23:38:44 +02:00
Michael Niedermayer	5de3a589f1	swresample/swresample: Print used int_sample_fmt Suggested-by: wm4 Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2015-06-21 18:12:56 +02:00
Michael Niedermayer	4977692461	swresample: Choose 16bit internally only if input and output is 16bit or less or if no rematrix and no resampling is performed and the input is 16bit note reampling and rematrix itself always use more than 16bit internally the "internal" sampling format is the format between these steps Its unlikely the difference from this commit is audible in any case unless there is some bug either before or after the change. but multiple people prefer this and it slightly improves the precission of computations. Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2015-06-21 17:33:46 +02:00
Michael Niedermayer	56f0fe6b84	swr: Fix ASSERT_LEVEL warning Found-by: cehoyos Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2015-06-08 20:16:06 +02:00
Clément Bœsch	c5a08956a3	swresample: fix initilaize/initialize typo	2015-06-06 12:16:18 +02:00
Michael Niedermayer	b14361486b	swresample/resample: fix typos Found-by: wm4 <nfxjfg@googlemail.com> Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2015-06-04 13:04:09 +02:00
Michael Niedermayer	c3f87f7545	swresample/swresample: Cleanup on init failure. This avoids leaks if the user doest call swr_close() after a failed init Found-by: James Almer <jamrial@gmail.com> Reviewed-by: James Almer <jamrial@gmail.com> Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2015-06-04 12:35:04 +02:00
Michael Niedermayer	cc17b43d8d	swresample: Add swr_get_out_samples() Previous version reviewed-by: Pavel Koshevoy <pkoshevoy@gmail.com> Previous version reviewed-by: wm4 <nfxjfg@googlemail.com> Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2015-06-04 05:37:32 +02:00
Michael Niedermayer	52acd22a7d	libswresample/rematrix: Check for malloc errors Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2015-06-04 02:36:30 +02:00
Ganesh Ajjanagadde	196b885a5f	swresample/dither: check memory allocation check memory allocation in swri_get_dither() Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2015-06-03 11:46:37 +02:00
Michael Niedermayer	02915602d9	swresample: Check the return value of resampler->init() Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2015-06-03 01:05:19 +02:00
James Almer	c16e99e3b3	x86: check for AV_CPU_FLAG_AVXSLOW where useful Signed-off-by: James Almer <jamrial@gmail.com> Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2015-06-01 00:15:35 +02:00
Rainer Hochecker	adb7372f74	swr: fix alignment issue caused by 8ch sse functions Fix crash when doing 8 ch conversion from apps compiled with MSVS Thanks to Ronald for giving this hint: https://ffmpeg.org/pipermail/ffmpeg-devel/2015-May/173049.html Reviewed-by: "Ronald S. Bultje" <rsbultje@gmail.com> Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2015-05-13 22:44:06 +02:00
Michael Niedermayer	223a859853	swresample/dither_template: Do not define macro functions to nothing This avoids potential warnings Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2015-05-07 01:07:30 +02:00
Michael Niedermayer	ff50b1b13b	swresample/swresample-test: Randomly wipe out channel counts Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2015-04-12 22:21:53 +02:00
Michael Niedermayer	3c77bb5f23	swresample: Check channel layouts and channels against each other and print human readable error messages Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2015-04-12 22:21:34 +02:00
Michael Niedermayer	80a28c7509	swresample: Allow reinitialization without ever setting channel layouts	2015-04-12 22:21:34 +02:00
Michael Niedermayer	d7b9cb2f7a	swresample: Allow reinitialization without ever setting channel counts Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2015-04-12 22:21:34 +02:00
James Almer	43482bd1a5	swr/resample: use av_clip functions Signed-off-by: James Almer <jamrial@gmail.com> Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2015-04-05 22:46:40 +02:00
Michael Niedermayer	1a10134e20	swresample/swresample: Use av_mallocz_array() Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2015-03-30 23:24:33 +02:00
Michael Niedermayer	e16592c42e	swresample/resample: Fix undefined shifts Found-by: Clang -fsanitize=shift Reported-by: Thierry Foucu <tfoucu@google.com> Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2015-03-14 01:15:37 +01:00
Michael Niedermayer	4d00860ac7	swresample: Add prefix to soxr_resampler also move declaration to header Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2015-02-27 19:20:43 +01:00
Michael Niedermayer	c0e3b46118	swresample: add av_cold to init functions Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2015-02-21 00:33:09 +01:00
Michael Niedermayer	0cb95f9082	swresample/resample_template: Add () to protect the arguments of the OUT() macro Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2015-02-17 00:36:35 +01:00
Michael Niedermayer	37013fd018	swresample/swresample-test: Add () to protect uint_rand() argument Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2015-02-16 23:09:09 +01:00
James Almer	f7ed997a6d	x86/swr: make pack_8ch functions work with compilers without aligned stack Signed-off-by: James Almer <jamrial@gmail.com>	2015-02-15 13:57:37 -03:00
Michael Niedermayer	b74ecb82fa	swresample/x86/rematrix_init: Check av_malloc* return codes, forward errors Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2015-02-09 10:15:56 +01:00
Michael Niedermayer	48ffaaaaef	swresample/x86/rematrix_init: Use av_mallocz_array() Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2015-02-09 10:15:56 +01:00
Michael Niedermayer	9d7ae72725	swresample: Use int instead of enum for fields which are accessed through AVOptions as int Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2015-02-02 23:27:26 +01:00
Michael Niedermayer	c77cc2c176	swresample/dither: Cleanup number suffixes The <<31 case needs LL Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2015-02-01 20:00:10 +01:00
Reimar Döffinger	6efd0ba977	swresample_internal.h: Move struct declaration before first use. It seems more logical and works with more restricted C compilers like tinycc. Signed-off-by: Reimar Döffinger <Reimar.Doeffinger@gmx.de>	2015-01-27 21:29:51 +01:00
James Almer	59ac93f6af	x86/swr: add SSE/AVX unpack_6ch functions int32/float only Reviewed-by: Michael Niedermayer <michaelni@gmx.at> Signed-off-by: James Almer <jamrial@gmail.com>	2015-01-12 15:40:03 -03:00
James Almer	6abf00d615	x86/swr: load constants outside the loop in pack_6ch functions Reviewed-by: Michael Niedermayer <michaelni@gmx.at> Signed-off-by: James Almer <jamrial@gmail.com>	2015-01-11 01:11:46 -03:00
James Almer	975ff6a3c6	x86/swr: disable pack_8ch functions on msvc/icl x86_32 Until a proper fix is committed. Signed-off-by: James Almer <jamrial@gmail.com>	2014-12-31 16:38:33 -03:00
James Almer	5f14f9e984	x86/swr: add missing alignment check to pack_6ch functions Reviewed-by: Michael Niedermayer <michaelni@gmx.at> Signed-off-by: James Almer <jamrial@gmail.com>	2014-12-31 13:35:11 -03:00
James Almer	37b35feb64	x86/swr: add SSE2/AVX pack_8ch functions Reviewed-by: Michael Niedermayer <michaelni@gmx.at> Reviewed-by: Ronald S. Bultje <rsbultje@gmail.com> Signed-off-by: James Almer <jamrial@gmail.com>	2014-12-30 23:05:27 -03:00
Michael Niedermayer	649c158e8c	Add FFMPEG_VERSION into the binary libs This simplifies identifying from which revision a binary of a lib came from Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2014-12-19 19:32:40 +01:00
Rob Sykes	4b6f225374	swresample/soxr_resample: fix error handling Fixes CID1257659 Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2014-12-13 21:12:56 +01:00
James Almer	edff061fb0	x86/swr: add ff_float_to_int32_a_avx2 13797 decicycles in ff_float_to_int32_a_sse2, 32768 runs, 0 skips 8603 decicycles in ff_float_to_int32_a_avx2, 32766 runs, 2 skips Reviewed-by: Christophe Gisquet <christophe.gisquet@gmail.com> Signed-off-by: James Almer <jamrial@gmail.com>	2014-11-07 15:01:35 -03:00
James Almer	b385c4c6a3	x86/swr: replace sse4 instructions in pack_6ch with sse ones There's no benefit from using blendps here except on CPUs with AVX, where it's faster than shufps according to Intel's documentation. As such, rename the sse4 functions to sse/sse2 and use shufps instead. Reviewed-by: Michael Niedermayer <michaelni@gmx.at> Signed-off-by: James Almer <jamrial@gmail.com>	2014-11-06 20:54:00 -03:00
Michael Niedermayer	e4f8a973aa	swresample: Fix swr_drop_output so it does not flush the buffers Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2014-11-04 17:17:23 +01:00
Michael Niedermayer	f6bb2cd1b0	swresample/resample: fix invert_initial_buffer() after flush Fixes: asan_heap-uaf_2071250_7_139.ogg Fixes: assertion failure Found-by: Mateusz "j00ru" Jurczyk and Gynvael Coldwind Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2014-10-16 22:35:27 +02:00
Michael Niedermayer	080c846f59	swresample: do not put multiple statements in one line Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2014-10-16 18:18:40 +02:00
Michael Niedermayer	344f8d307a	swresample/audioconvert: Fix undefined behavior (left shift of negative value) Fixes: asan_heap-oob_4da4f3_8_asan_heap-oob_4da4f3_419_scene1a.mm Found-by: Mateusz "j00ru" Jurczyk and Gynvael Coldwind Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2014-10-08 05:35:06 +02:00
Michael Niedermayer	6b347f519d	swresample/swresample: replace always true if() by av_assert0() Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2014-10-06 01:29:15 +02:00
Michael Niedermayer	f9fefa499f	swresample/swresample: fix sample drop loop end condition Fixes Ticket3985 Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2014-10-06 01:29:06 +02:00
Reimar Döffinger	2c5c37ade1	libswresample: move condition to start of loop. This avoids several issue like calculating sum/maxcoef incorrectly due to adding up matrix entries that will be overwritten, as well as out-of-range writes to s->matrix if the maximum allowed number of channels is used. Signed-off-by: Reimar Döffinger <Reimar.Doeffinger@gmx.de>	2014-09-07 11:31:34 +02:00
Reimar Döffinger	284123d7fd	Remove pointless if. A branch to avoid some calculation seems unlikely to have any benefits. Signed-off-by: Reimar Döffinger <Reimar.Doeffinger@gmx.de>	2014-09-07 11:31:33 +02:00
Reimar Döffinger	2231d5b671	libswresample: Avoid needlessly large on-stack array. We only actually need to use a tiny part of it. Unfortunately we seem to have no real test coverage on the code, so this is a bit risky. Signed-off-by: Reimar Döffinger <Reimar.Doeffinger@gmx.de>	2014-09-07 11:31:33 +02:00
Michael Niedermayer	7c51f5bd39	swr: aarch64 audio_convert and neon clobber test Ported from avresample Code by: Mans Rullgard, Janne Grunau, Martin Storsjo Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2014-08-27 20:06:37 +02:00
Michael Niedermayer	b7d5e016a3	swresample: Add AVFrame based API Based on commit `fb1ddcdc8f` by Luca Barbato <lu_zero@gentoo.org> Adapted for libswresample by Michael Niedermayer Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2014-08-16 20:39:37 +02:00
Michael Niedermayer	f4e814f787	swresample: check av_opt_set for failure in swr_alloc_set_opts() Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2014-08-11 01:34:34 +02:00
Luca Barbato	c4ac48c5a1	swresample: document the need to configure the context using AVOptions Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2014-08-11 00:35:56 +02:00
Michael Niedermayer	97f8c7a03e	bump libpostproc and libswresample this is needed / avoids some headaches as one of their dependancies (libavutil) was bumped Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2014-08-10 01:17:02 +02:00
Michael Niedermayer	74be0f82a7	swresample-test: make it independant of the internal SWR_CH_MAX Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2014-08-10 01:15:43 +02:00
Michael Niedermayer	05ff1a2c05	swresample/swresample: Treat mono as planar This might in some cases improve performance. Idea from: `fbc0b86599` Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2014-08-04 03:05:05 +02:00
Andreas Cadhalpun	39a6e02fd4	fix spelling errors Reviewed-by: Timothy Gu <timothygu99@gmail.com> Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2014-07-12 22:33:27 +02:00
Michael Niedermayer	52fafaf474	swresample/libswresample.v: hide ff_* Found-by: Hendrik Leppkes <h.leppkes@gmail.com> Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2014-07-12 18:40:17 +02:00
Timothy Gu	5b58692ed4	swresample: misc. doxy improvements Signed-off-by: Timothy Gu <timothygu99@gmail.com> Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2014-07-08 01:36:56 +02:00
Timothy Gu	064945b3aa	swresample: organize functions into doxy groups Signed-off-by: Timothy Gu <timothygu99@gmail.com> Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2014-07-06 22:07:15 +02:00
Timothy Gu	81f47e272d	swresample: better doxy for configuration-returning functions Signed-off-by: Timothy Gu <timothygu99@gmail.com> Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2014-07-06 22:06:51 +02:00
Timothy Gu	2711b4708a	swresample: improve Doxygen introduction Signed-off-by: Timothy Gu <timothygu99@gmail.com> Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2014-07-06 21:41:31 +02:00
Timothy Gu	77c5f546e7	swresample: add SwrContext doxy Signed-off-by: Timothy Gu <timothygu99@gmail.com> Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2014-07-06 20:58:58 +02:00
Timothy Gu	fc71434e84	swresample: add SwrDitherType doxy Signed-off-by: Timothy Gu <timothygu99@gmail.com> Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2014-07-06 20:58:41 +02:00
Timothy Gu	c0d9b026f9	swresample: group all the option constants in a section in doxy Signed-off-by: Timothy Gu <timothygu99@gmail.com> Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2014-07-06 20:58:34 +02:00
Timothy Gu	0c58388211	swresample: grammar/capitalization fixes Signed-off-by: Timothy Gu <timothygu99@gmail.com> Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2014-07-06 20:46:06 +02:00
Timothy Gu	37715b4594	swresample: split option table to a separate file Signed-off-by: Timothy Gu <timothygu99@gmail.com> Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2014-07-06 03:45:46 +02:00
James Almer	9937362c54	x86/swr: use lavu helper macros to check CPU extensions Signed-off-by: James Almer <jamrial@gmail.com> Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2014-07-04 02:12:16 +02:00
James Almer	8279a15284	x86/swr: split audioconvert and rematrix DSP into separate files Also rename resample_x86_dsp.c to resample_init.c Signed-off-by: James Almer <jamrial@gmail.com> Signed-off-by: Michael Niedermayer <michaelni@gmx.at>	2014-07-04 02:00:11 +02:00

1 2 3 4 5 ...

584 Commits