generic-library/vpx

Author	SHA1	Message	Date
Jim Bankoski	e69b5258fd	fix vp9_vp8 files renamed Change-Id: I20c426e91ee49666db42e20eb074095ab6b8ec5d	2012-11-29 06:53:08 -08:00
Jim Bankoski	13dbf1fb17	more rtcd cleanup Change-Id: Ieefd76e164ca4aa87597da0412977614ddfbacb7	2012-11-28 17:27:15 -08:00
Deb Mukherjee	0de214260b	Merge "Fixing 8x8/4x4 ADST for intra modes with tx select" into experimental	2012-11-28 16:59:17 -08:00
Deb Mukherjee	0742b1e4ae	Fixing 8x8/4x4 ADST for intra modes with tx select This patch allows use of 8x8 and 4x4 ADST correctly for Intra 16x16 modes and Intra 8x8 modes when the block size selected is smaller than the prediction mode. Also includes some cleanups and refactoring. Rebase. Change-Id: Ie3257bdf07bdb9c6e9476915e3a80183c8fa005a	2012-11-28 16:21:12 -08:00
Yaowu Xu	b2f27d909a	Merge "remove the vp9_default_mode_contexts_a" into experimental	2012-11-28 13:56:42 -08:00
Yaowu Xu	1cc5739669	remove the vp9_default_mode_contexts_a Given the way mode_context is updated, the benefit of an additional default is not signficant. Change-Id: I67489453e8781340b18e26a1cc2f04e9221004a2	2012-11-28 11:14:30 -08:00
Jim Bankoski	c67873989f	fixed includes to be fully specified Change-Id: Ia1cce221f8511561b9cbd8edb7726fbc286ff243	2012-11-28 10:53:17 -08:00
Jim Bankoski	926d95cd84	Merge "remove postproc invokes" into experimental	2012-11-28 10:30:42 -08:00
John Koleszar	00e2c6bf7a	Merge "Clamp decoded feature data" into experimental	2012-11-28 10:08:37 -08:00
Jim Bankoski	85cba19e16	remove postproc invokes and some miscellaneous invoke left overs Change-Id: I63191b1bfd3bea4ce30cceaeb686ec850570fc43	2012-11-28 10:00:25 -08:00
Yunqing Wang	d202138621	Further improve macroblock loop filters This change included: 1. Aligned reads in vp9_mbloop_filter_vertical_edge function. Since we actually read 16 bytes, we can align the reads to read starting at (s - 8) instead of (s - 5). 2. Combined u, v loop filters. 3. Added 8x16 transpose. This gave 2% decoder performance gain (tulip clip). Change-Id: Ib14c2f1645c4a3436df17fe2f24789506bf0bb58	2012-11-28 09:27:07 -08:00
Yaowu Xu	12da793d00	removed redundant mode_context data structures This commit removed a couple of redundant data structures in frame coding contextsm, mode_context and mode_context_a, and changed to use vp9_mode_contexts only. The switch of the context for different frame type now relies on the switch of frame coding context between lfc and lfc_a. This commit also removed a number of memcpy among these redundant data structure. Change-Id: I42e8174bd60f466b0860afc44c1263896471b0f3	2012-11-28 09:24:30 -08:00
John Koleszar	a1f15814be	Clamp decoded feature data Not all segment feature data elements are full-range powers of two, so there are values that can be encoded that are invalid. Add a new function to clamp values to the maximum allowed. Change-Id: Ie47cb80ef2d54292e6b8db9f699c57214a915bc4	2012-11-27 16:38:31 -08:00
John Koleszar	fcccbcbb39	Add vp9_ prefix to all vp9 files Support for gyp which doesn't support multiple objects in the same static library having the same basename. Change-Id: Ib947eefbaf68f8b177a796d23f875ccdfa6bc9dc	2012-11-27 14:12:30 -08:00
Yunqing Wang	e7cd80718b	Improve sad3x16 SSE2 function Vp9_sad3x16_sse2() is heavily called in decoder, in which the unaligned reads consume lots of cpu cycles. When CONFIG_SUBPELREFMV is off, the unaligned offset is 1. In this situation, we can adjust the src_ptr to be 4-byte aligned, and then do the aligned reads. This reduced the reading time significantly. Tests on 1080p clip showed over 2% decoder performance gain with CONFIG_SUBPELREFM off. Change-Id: I953afe3ac5406107933ef49d0b695eafba9a6507	2012-11-26 09:53:50 -08:00
Jim Bankoski	510557e2eb	removed the idct rtcd idct calls More cleanup to do after this, but this is a good chunk of removing rtcd. Change-Id: I551db75e341a0a85c3ad650df1e9a60dc305681a	2012-11-24 19:33:58 -08:00
Jim Bankoski	3338af4109	remove subpixel invoke functions Removed the rtcd subpixel invoke functions. Change-Id: I8b7618bd5813333fac66b2817bdf807616e0fb33	2012-11-21 09:16:30 -08:00
Jim Bankoski	4ad2f08c72	Merge "clean out some of the rtcd code." into experimental	2012-11-21 06:41:37 -08:00
John Koleszar	414f68d266	Merge "Pack invisible frames without lengths" into experimental	2012-11-20 17:22:50 -08:00
Yunqing Wang	bbe5e032a4	Fix ref_stride in sad function Used ref_stride. Change-Id: I31f0a3bb935520f54d11a1d87315627f162ae845	2012-11-20 10:01:20 -08:00
Jim Bankoski	f4871b6a3f	clean out some of the rtcd code. This removes functions that are no longer needed and cleans up some warnings. Change-Id: I292a4c3694e9c1d68ce99cea390905b198434719	2012-11-18 12:33:18 -08:00
Jim Bankoski	cb98b83239	removal of temporal invoke Change-Id: I18ca713b02a5241bdb20dddcde0216467b55b596	2012-11-17 06:11:01 -08:00
Yunqing Wang	0eb5590425	Merge "Add const before the dequant(dq)" into experimental	2012-11-16 12:35:17 -08:00
Yunqing Wang	4c7c15ee69	Merge "Optimize 8x8 dequant and idct" into experimental	2012-11-16 12:23:06 -08:00
Yunqing Wang	47d9d48fa4	Add const before the dequant(dq) Modified code to use const before dq. Change-Id: I6fa59c2ed9743ded33ad08df70e15c2fe1ae7b99	2012-11-16 12:13:13 -08:00
Ronald S. Bultje	5b11052ac1	Support 32x32 intra modes in non-keyframe superblocks. Change-Id: Icf8ad313c543462e523bff89690e5daa8d49bcc0	2012-11-16 09:54:43 -08:00
Paul Wilkins	a57dbd957b	Further experimentation with the mode context Experiments with a larger set of contexts and some clean up to replace magic numbers regarding the number of contexts. The starting values and rate of backwards adaption are still suspect and based on a small set of tests. Added forwards adjustment of probabilities. The net result of adding the new context and forward update is small compared to the old context from the legacy find_near function. (down a little on derf but up by a similar amount for HD) HOWEVER.... with the new context and forward update the impact of disabling the reverse update (which may be necessary in some use cases to facilitate parallel decoding) is hugely reduced. For the old context without forward update, the impact of turning off reverse update (Experiment was with SB off) was Derf - 0.9, Yt -1.89, ythd -2.75 and sthd -8.35. The impact was mainly at low data rates. With the new context and forward update enabled the impact for all the test sets was no more than 0.5-1% (again most at the low end). Change-Id: Ic751b414c8ce7f7f3ebc6f19a741d774d2b4b556	2012-11-16 16:58:00 +00:00
Deb Mukherjee	cb2d06ceac	Merge "Compound inter-intra experiment" into experimental	2012-11-16 08:30:34 -08:00
Yaowu Xu	170305dcd3	Merge "changed mv candidate search for superblocks" into experimental	2012-11-16 07:21:55 -08:00
Yaowu Xu	415e6bff4d	changed mv candidate search for superblocks added additional motion vectors at close neighborhood of a superblock to the list of candiate motion vectors, and removed a couple that are further away. The change helped std-hd set about .8% (all metrics) and smaller gain for derf set. Change-Id: Iaa69b98614db43420ed3fd4738d0ca5587b90045	2012-11-16 07:01:13 -08:00
Deb Mukherjee	0c917fc975	Compound inter-intra experiment A patch on compound inter-intra prediction. In compound inter-intra prediction, a new predictor for 16x16 inter coded MBs are obtained by combining a single inter predictor with a 16x16 intra predictor, in a manner that the weight varies with distance from the top/left boundary. The current search strategy is to combine the best inter mode with the best intra mode obtained independently. Results so far: derf +0.31% yt +0.32% std-hd +0.35% hd +0.42% It is conceivable that the results would improve somewhat with a more thorough search strategy where all intra modes are searched given the best mv, or even a joint search for the best mv and the best intra mode. Change-Id: I7951f1ed0d6eb31ca32ac24d120f1585bcd8d79b	2012-11-16 06:56:29 -08:00
Yaowu Xu	1c56946ec1	Merge "subpelrefmv for superblocks" into experimental	2012-11-16 05:49:32 -08:00
John Koleszar	64bcffc1ec	Pack invisible frames without lengths Modify the decoder to return the ending position of the bool decoder and use that as the starting position for the next frame. The constant-space algorithm for parsing the appended frame lengths is O(n^2), which is a potential DoS concern if n is unbounded. Revisit the appended lengths for use as partition lengths when multipartition support is added. In addition, this allows decoding of raw streams outside of a container without additional framing information, though it's insufficient to be able to remux said stream into a container. Change-Id: I71e801a9c3e37abe559a56a597635b0cbae1934b	2012-11-15 15:48:07 -08:00
Yaowu Xu	61416aedc2	subpelrefmv for superblocks duplicate code clean-up and variable name corrections Change-Id: Ibc4703228e652ec425125de5e7bc038fa46595c5	2012-11-15 13:46:52 -08:00
John Koleszar	a9c7597adc	support building vp8 and vp9 into a single lib Change-Id: Ib8f8a66c9fd31e508cdc9caa662192f38433aa3d	2012-11-15 10:46:17 -08:00
John Koleszar	6becad426c	detokenize: use SEG_LVL_EOB feature consistently Update decode_coefs() to break when c >= eob, since it's possible that c starts the loop from 1 and eob is 0. The loop won't terminate in that case. Add new get_eob() function to consistently clamp the eob based on the segment level EOB and the block size. It's possible to code a segment level EOB that's greater than the block size, and that leads to an out of bounds access. Change-Id: I859563b30414615cf1b30dcc2aef8a1de358c42d	2012-11-15 11:44:29 +00:00
pascal massimino	5a955973d9	Merge changes I63348ae3,I658ea409 into experimental * changes: Segment mode coding bug. Silenced a few warnings.	2012-11-15 00:24:57 -08:00
Ronald S. Bultje	127836d11f	Merge "Don't use hybrid transform (ADST) for superblocks." into experimental	2012-11-14 09:18:34 -08:00
Ronald S. Bultje	1e3dd49fe3	Don't use hybrid transform (ADST) for superblocks. This is in line with other cases where we disable ADST if prediction size and transform size don't match. Before this patch, the RD loop will use ADST for superblocks, but frame encoding/decoding won't. Change-Id: I700368c632eb72b5e089c22ef25649d99d7697d0	2012-11-14 08:58:24 -08:00
Paul Wilkins	b527c4dbb7	Segment mode coding bug. There are now more than 16 possible modes so 5 bits required for segment mode feature. Note that it is likely that the mode feature and how it is coded will change but for now the 4 bits was a bug. Change-Id: I63348ae3a9cc31566a656c2dc78f09f5e1a9dcc9	2012-11-14 14:38:03 +00:00
Paul Wilkins	19a1ba1e91	Silenced a few warnings. Silenced a few VS compiler warnings. Change-Id: I658ea409c36c05cd11042675e2e42ccde0ef2420	2012-11-14 14:27:37 +00:00
Yaowu Xu	3fa1348d5f	fix a few typos Change-Id: I7b6f27826052eb706fc6080d4e3a940dff7d3a58	2012-11-13 14:45:53 -08:00
Ronald S. Bultje	1761a6b55a	Merge "Use full 32-pixel edge for superblock bestrefmv motion vector ordering." into experimental	2012-11-13 14:12:58 -08:00
Ronald S. Bultje	b147c64c16	Merge "Fix edge MV handling in SBs." into experimental	2012-11-13 14:12:48 -08:00
Deb Mukherjee	7de64f35d3	A fix in MV_REF experiment This fix ensures that the forward prob update is not turned off for motion vectors. Change-Id: I0b63c9401155926763c6294df6cca68b32bac340	2012-11-13 08:27:04 -08:00
Yunqing Wang	e60478d46d	Optimize 8x8 dequant and idct Similar to 16x16 dequant and idct, based on the value of eobs, the 8x8 dequant and idct calculation was simplified to improve decorder performance. Combined vp9_dequant_idct_add_8x8 and vp9_dequant_dc_idct_add_8x8 to eliminate duplicate code. Change-Id: Ia58e50ab27f7012b7379c495837c9c0b5ba9cf7f	2012-11-12 17:41:53 -08:00
Ronald S. Bultje	c79ae1713c	Use full 32-pixel edge for superblock bestrefmv motion vector ordering. Change-Id: I417e39867c020a17d85370972446a8ce2bbe9a6d	2012-11-12 17:06:56 -08:00
Ronald S. Bultje	722972454c	Fix edge MV handling in SBs. Change-Id: Ia1eddb108ec463835e9de8769572d698e21bca49	2012-11-12 17:06:52 -08:00
Paul Wilkins	2669f42b0d	New inter mode context This change is a fix / extension of the newbestrefmv experiment. As such it is presented without IFDEF. The change creates a new context for coding inter modes in vp9_find_mv_refs(). This replaces the context that was previously calculated in vp9_find_near_mvs(). The new context is unoptimized and not necessarily any better at this stage (results pending), but eliminates the need for a legacy call to vp9_find_near_mvs(). Based on numbers from Scott, this could help decode speed by several %. In a later patch I will add support for forward update of context (assuming this helps) and refine the context as necessary. Change-Id: I1cd991b82c8df86cc02237a34185e6d67510698a	2012-11-12 15:50:02 +00:00
Paul Wilkins	6fb8953c19	Restrict ref mv search range. Experiment to test speed trade off of reducing the extent of the ref mv search. Reducing the maximum number of tested candidates to 9 had minimal net effect on quality in any of the tests sets. Reduction to 7 has a small negative impact (worst was STD-HD at about -0.2%). This change is in response to the apparently high number of decode cycles reported in regard to mv-ref selection. Change-Id: I0e92e92e324337689358495a1ec9ccdeb23dc774	2012-11-12 11:31:12 +00:00

... 11 12 13 14 15

723 Commits