generic-library/vpx

Author	SHA1	Message	Date
John Koleszar	f0d9f10d24	Remove all asm offset files from VP9 The files are empty and unused. Change-Id: Ieb4242d14273efdf24149bda33f9591540bba06a	2013-07-09 14:26:53 -07:00
Frank Galligan	198fa6d0a0	Add Neon horizontal and vertical vp9_mbloop_filter - The vp9 mbfilter C code will branch on flat and mask. This CL will perform both branches and combine the data. A later CL will perform a check to see if all patch will take one branch. - These functions are about 1.75 times faster than the C code on Nexus 7. PS #3 - Changed all functions to dub limit, blimit, and thresh from vld {dx[]}, freeing up r4-r6. - Changed code to use vbif to reduce one instruction and free up a d register. Change-Id: I028dae0e434dc9891c3677bdb182e201ffb04777	2013-07-09 12:40:05 -07:00
Dmitry Kovalev	ec68d25521	Merge "Adding update_tx_ct function, removing duplicated code."	2013-07-09 12:26:11 -07:00
Dmitry Kovalev	aeed28f143	Removing vp9_maskingmv.c and corresponding assembly file. Change-Id: I9842d02d61d78d17dc3449bae8ffbe60f4b3ecb3	2013-07-09 11:22:56 -07:00
Dmitry Kovalev	92a9eaef50	Loop filter code cleanup. Using MAX_LOOP_FILTER constant instead of number 63. Change-Id: If91e0c198331b3041e7cd0707a5948479e9209d8	2013-07-09 11:18:09 -07:00
Ronald S. Bultje	d8fa5d45cc	Merge "Make intra prediction pointers RTCD-based."	2013-07-09 09:54:43 -07:00
Yaowu Xu	df5731273f	Merge "Fix loopfilter bug"	2013-07-09 01:34:25 -07:00
Dmitry Kovalev	c6c279aff0	Merge "Using mi_cols instead of mb_cols."	2013-07-08 20:09:19 -07:00
Dmitry Kovalev	1c65c580d6	Merge "Refactoring setup_pre_planes function."	2013-07-08 20:08:05 -07:00
Dmitry Kovalev	6254c8d780	Merge "Calling set_partition_seg_context() instead of code duplication."	2013-07-08 20:07:06 -07:00
Ronald S. Bultje	8350e7fe38	Make intra prediction pointers RTCD-based. This probably has a mildly negative impact on performance, but will (in future commits - or possibly merged with this one) allow SIMD implementations of individual intra prediction functions. We may perhaps want to consider having separate functions per txfm-size also (i.e. 4x4, 8x8, 16x16 and 32x32 intra prediction functions for each intra prediction mode), but I haven't played much with that yet. Change-Id: Ie739985eee0a3fcbb7aed29ee6910fdb653ea269	2013-07-08 17:25:51 -07:00
John Koleszar	527fc5caf6	Fix loopfilter bug In the rare case were 4x4 interior filtering was called for but no 8x8 or larger filtering takes place, the previous code was skipping the filtering. This patch fixes the issue by including the interior mask in the overall mask for the filter application loops. Change-Id: I4a0b65056c64f97478827c2ff41e0914fc7779d0	2013-07-08 16:49:57 -07:00
Ronald S. Bultje	bd867f1619	Inline vp9_get_mv_joint(). Encode time for first 50 frames of bus (speed 0) @ 1500kbps goes from 2min10.9 to 2min10.5, i.e. 0.3% faster overall, basically because we prevent the call overhead. Change-Id: I1eab1a95dd3eae282f9b866f1f0b3dcadff073d5	2013-07-08 16:22:39 -07:00
Dmitry Kovalev	b7559258a4	Using mi_cols instead of mb_cols. Eliminating usage of mb-units, switching to mi-units. Adding ALIGN_POWER_OF_TWO macro. Change-Id: I2491c969f713207c062011878b57e4e531818607	2013-07-08 14:54:04 -07:00
Tero Rintaluoma	18303b1263	Fix intermediate height in convolve intermediate_height for horizontal filtering must be at least 8 pixels to be able to do vertical filtering correctly. Currently it can be less for small block and y_step_q4 sizes. Change-Id: I2ee28b0591b2041c2fa9844d0ae2ff8a1a59cc21	2013-07-05 14:58:25 +03:00
Dmitry Kovalev	bfcef95c45	Adding update_tx_ct function, removing duplicated code. Change-Id: I8882fe3cd247a5a8304ab8ab2ee9abdb92830133	2013-07-03 18:24:13 -07:00
Dmitry Kovalev	f72e072555	Refactoring setup_pre_planes function. Removing set_refs, adding set_ref function. Change-Id: I5635c478b106ae4e57d317f1c83d929644307e63	2013-07-03 17:42:01 -07:00
Dmitry Kovalev	430bd0c94a	Merge "Replacing 64 / MI_SIZE with MI_BLOCK_SIZE."	2013-07-03 14:16:02 -07:00
Dmitry Kovalev	2ad62c9312	Calling set_partition_seg_context() instead of code duplication. Change-Id: I65be6acc54c99688fd1f0c946cec3511514b8555	2013-07-03 11:15:58 -07:00
Dmitry Kovalev	5a21de8418	Replacing 64 / MI_SIZE with MI_BLOCK_SIZE. Change-Id: I32276552b3ea6dc1dce8e298be114cfe1019b31c	2013-07-03 10:54:50 -07:00
Yaowu Xu	0f02dc2709	Inline a few intra predictors Change-Id: Ib41f0643fdcc088500e7420708f4e72f1f64c710	2013-07-03 10:20:41 -07:00
Ronald S. Bultje	98c493a1c0	Merge "Remove unused function vp9_build_inter4x4_predictors_mbuv()."	2013-07-03 09:05:20 -07:00
Dmitry Kovalev	be77f6bbbf	Removing redundant struct from union b_mode_info. Change-Id: I08fc6e474ff2c12cfa065bae4989c724276e2c83	2013-07-02 16:51:57 -07:00
Ronald S. Bultje	5b87240230	Remove unused function vp9_build_inter4x4_predictors_mbuv(). Change-Id: Ibfd2def2c088f4bc541a1de25990d73480b53d4b	2013-07-02 16:34:24 -07:00
Deb Mukherjee	8d3d2b76f3	Tx size selection enhancements (1) Refines the modeling function and uses that to add some speed features. Specifically, intead of using a flag use_largest_txfm as a speed feature, an enum tx_size_search_method is used, of which two of the types are USE_FULL_RD and USE_LARGESTALL. Two other new types are added: USE_LARGESTINTRA (use largest only for intra) USE_LARGESTINTRA_MODELINTER (use largest for intra, and model for inter) (2) Another change is that the framework for deciding transform type is simplified to use a heuristic count based method rather than an rd based method using txfm_cache. In practice the new method is found to work just as well - with derf only -0.01 down. The new method is more compatible with the new framework where certain rd costs are based on full rd and certain others are based on modeled rd or are not computed. In this patch the existing rd based method is still kept for use in the USE_FULL_RD mode. In the other modes, the count based method is used. However the recommendation is to remove it eventually since the benefit is limited, and will remove a lot of complications in the code (3) Finally a bug is fixed with the existing use_largest_txfm speed feature that causes mismatches when the lossless mode and 4x4 WH transform is forced. Results on derf: USE_FULL_RD: +0.03% (due to change in the tables), 0% encode time reduction USE_LARGESTINTRA: -0.21%, 15% encode time reduction (this one is a pretty good compromise) USE_LARGESTINTRA_MODELINTER: -0.98%, 22% encode time reduction (currently the benefit of modeling is limited for txfm size selection, but keeping this enum as a placeholder) . USE_LARGESTALL: -1.05%, 27% encode-time reduction (same as existing use_largest_txfm speed feature). Change-Id: I4d60a5f9ce78fbc90cddf2f97ed91d8bc0d4f936	2013-07-02 13:54:00 -07:00
Dmitry Kovalev	904070ca64	Merge "Removing unused implicit segmentation code."	2013-07-02 11:58:48 -07:00
Ronald S. Bultje	3cc6eb7c00	Merge "Make get_coef_context() branchless."	2013-07-02 11:48:15 -07:00
Dmitry Kovalev	3140c443e4	Merge "Removing vp9_mbpitch.c, moving vp9_setup_block_dptrs to vp9_block.h."	2013-07-02 11:31:35 -07:00
Dmitry Kovalev	a3d2e6c98b	Removing unused implicit segmentation code. Change-Id: I8a2983fb14274a6ac53681fa4cd5d4209cbd2905	2013-07-02 11:16:42 -07:00
Ronald S. Bultje	9df24b41ca	Merge "Update quantize SSSE3 SIMD to cover 32x32 transform case also."	2013-07-02 09:38:08 -07:00
Dmitry Kovalev	1ac0540296	Removing vp9_mbpitch.c, moving vp9_setup_block_dptrs to vp9_block.h. Change-Id: Ia547a5dd7650b771fd00edd673ab9f920270731c	2013-07-01 17:28:08 -07:00
Ronald S. Bultje	26b6318de8	Make get_coef_context() branchless. This should significantly speedup cost_coeffs(). Basically what the patch does is to make the neighbour arrays padded by one item to prevent an eob check in get_coef_context(), then it populates each col/row scan and left/top edge coefficient with two times the same neighbour - this prevents a single/double context branch in get_coef_context(). Lastly, it populates neighbour arrays in pixel order (rather than scan order), so we don't have to dereference the scantable to get the correct neighbours. Total encoding time of first 50 frames of bus (speed 0) at 1500kbps goes from 2min10.1 to 2min5.3, i.e. a 2.6% overall speed increase. Change-Id: I42bcd2210fd7bec03767ef0e2945a665b851df56	2013-07-01 16:34:10 -07:00
Yaowu Xu	ba3b2604f0	Merge "Quantize (64-bit only, for now) SSSE3 SIMD."	2013-07-01 15:58:57 -07:00
Ronald S. Bultje	c8defcfdee	Update quantize SSSE3 SIMD to cover 32x32 transform case also. Encode time of bus (speed 0) 50 frames @ 1500kbps goes from 2min14.4 to 2min10.1, i.e. a 2.3% overall speed increase. Change-Id: I3699580e74ec26c7d24e03681bc47ba25ee1ee87	2013-07-01 11:36:33 -07:00
Ronald S. Bultje	7353ceab9d	Quantize (64-bit only, for now) SSSE3 SIMD. Total encoding time for first 50 frames of bus (speed 0) @ 1500kbps goes 2min34.8 to 2min14.4, i.e. a 10.4% overall speedup. The code is x86-64 only, it needs some minor modifications to be 32bit compatible, because it uses 15 xmm registers, whereas 32bit only has 8. Change-Id: I2df53770c2e850813ffa713e1a91b45b0082b904	2013-07-01 11:36:07 -07:00
Dmitry Kovalev	2ab3bc8871	Removing vp9_modecont.{h, c}. Moving vp9_default_inter_mode_probs array to vp9_entropymode.c. Change-Id: I88ebda86ccc07f2a43c6c01d4b37898214cfb6de	2013-07-01 10:17:15 -07:00
Jingning Han	993942ce0c	Merge "Enable SSE2 4x4 ADST/DCT transform"	2013-06-29 15:57:04 -07:00
Christian Duvivier	466e0cf303	SSE2 version of vp9_short_fdct32x32_rd. 43,000 -> 5,750 cycles, about 7.5x faster. Change-Id: Ibfd92821b9603f4ed9c256e0ececec14fa4565d0	2013-06-29 13:53:00 -07:00
Johann	6098e359f4	Merge "add Neon optimized add constant residual functions"	2013-06-28 19:50:38 -07:00
James Zern	84d08fa9c4	Merge "fix test compile error"	2013-06-28 19:48:05 -07:00
Ronald S. Bultje	a487af8d35	Merge "Inline vp9_get_coef_context() (and remove vp9_ prefix)."	2013-06-28 19:37:11 -07:00
chm	a83cfd4da1	add Neon optimized add constant residual functions - Add add_constant_residual_8x8 16x16 32x32 functions - Tested under RealView debugger enviroment Change-Id: I5c3a432f651b49bf375de6496353706a33e3e68e	2013-06-28 19:06:51 -07:00
James Zern	a63e31e81e	fix test compile error since: `92479d9` Make update_partition_context faster fixes: vp9/common/vp9_blockd.h:408:22: error: non-constant-expression cannot be narrowed from type 'int' to 'char' in initializer list [-Wc++11-narrowing] char pcvalue[2] = {~(0xe << boffset), ~(0xf <<boffset)}; ^~~~~~~~~~~~~~~~~ Change-Id: Id5b00b9a72d00a2b314081a23879bd1fa3ce983b	2013-06-28 18:07:37 -07:00
Jingning Han	1109b6b888	Enable SSE2 4x4 ADST/DCT transform This commit enables SSE2 4x4 foward hybrid transform. The runtime goes from 249 cycles down to 74 cycles. Overall around 2% speed-up at no compression performance change. Change-Id: Iad4d526346e05c7be896466c05500711bb763660	2013-06-28 17:24:43 -07:00
Dmitry Kovalev	228b8232d3	Cosmetic reordering of FRAME_CONTEXT members. Change-Id: Id641e5188adf55e53e606e5813ae45feaf7abbd2	2013-06-28 16:16:03 -07:00
Dmitry Kovalev	59070f6e3c	Merge "Removing CONFIG_DEBUG checks on assertions."	2013-06-28 14:03:28 -07:00
Ronald S. Bultje	ec5d09b950	Merge "Make coefficient skip condition an explicit RD choice."	2013-06-28 11:54:28 -07:00
Ronald S. Bultje	d00b8e5f82	Inline vp9_get_coef_context() (and remove vp9_ prefix). Makes cost_coeffs() a lot faster: 4x4: 236 -> 181 cycles 8x8: 888 -> 588 cycles 16x16: 3550 -> 2483 cycles 32x32: 17392 -> 12010 cycles Total encode time of first 50 frames of bus (speed 0) @ 1500kbps goes from 2min51.6 to 2min43.9, i.e. 4.7% overall speedup. Change-Id: I16b8d595946393c8dc661599550b3f37f5718896	2013-06-28 10:40:21 -07:00
Dmitry Kovalev	0345fc3ad9	Merge "Decoder's code cleanup."	2013-06-28 10:38:54 -07:00
Dmitry Kovalev	8e6ce6bb9e	Removing CONFIG_DEBUG checks on assertions. Adding CHECK_MEM_ERROR macro to vp9_common.h and removing two duplicated ones from vp9_onyx_int.h and vp9_onyxd_int.h. Change-Id: I916afec61b3019f18193135dac7c35ed0f89b8b6	2013-06-28 10:36:20 -07:00
Ronald S. Bultje	af660715c0	Make coefficient skip condition an explicit RD choice. This commit replaces zrun_zbin_boost, a method of biasing non-zero coefficients following runs of zero-coefficients to be rounded towards zero, with an explicit skip-block choice in the RD loop. The logic is basically that if individual coefficients should be rounded towards zero (from a RD point of view), the trellis/optimize loop should take care of it. If whole blocks should be zero (from a RD point of view), a single RD check is much more efficient than a complete serialization of the quantization loop. Quality change: derf +0.5% psnr, +1.6% ssim; yt +0.6% psnr, +1.1% ssim. SIMD for quantize will follow in a separate patch. Results for other test sets pending. Change-Id: Ife5fa641163ac5150ac428011e87188f1937c1f4	2013-06-28 10:28:49 -07:00
Dmitry Kovalev	3231da0a9e	Decoder's code cleanup. Using vp9_set_pred_flag function instead of custom code, adding decode_tokens function which is now called from decode_atom, decode_sb_intra, and decode_sb. Change-Id: Ie163a7106c0241099da9c5fe03069bd71f9d9ff8	2013-06-27 16:15:43 -07:00
Frank Galligan	1d6dc1b702	Add Neon optimized loop filter functions. - Added vp9_loop_filter_horizontal_edge_neon and vp9_loop_filter_vertical_edge_neon. - The functions are based off the vp8 loopfilter functions. - Matches x86 md5 checksum. Change-Id: Id1c4dddb03584227e5ecd29f574a6ac27738fdd0	2013-06-27 16:14:45 -07:00
Dmitry Kovalev	a3664258c5	Merge "General cleanup in segmentation-related code."	2013-06-27 14:57:07 -07:00
Dmitry Kovalev	be83ef3104	Merge "Moving subexp encoding functions in separate vp9_dsubexp.c file."	2013-06-27 14:55:18 -07:00
Jingning Han	fc1cfd8e32	Merge "Make intra predictor reference buffer configurable"	2013-06-26 19:02:02 -07:00
Jingning Han	4c10515f89	Merge "Make update_partition_context faster"	2013-06-26 19:01:45 -07:00
Yaowu Xu	896dc47cac	Merge "Change to use LUT for mode-to-txfm conversion"	2013-06-26 17:19:47 -07:00
Jingning Han	861cb06c67	Make intra predictor reference buffer configurable This commit enables configurable reference buffer pointer for intra predictor. This allows later removal of spatial dependency between blocks inside a 64x64 superblock in the rate-distortion optimization loop. Change-Id: I02418c2077efe19adc86e046a6b49364a980f5b1	2013-06-26 17:17:21 -07:00
Jingning Han	92479d9526	Make update_partition_context faster Use vpx_memset for updating the partition contexts. Thanks to Noah for pointing out the need of refactoring in this part. Change-Id: I67fb78429d632298f1cd8a0be346cc76f79392a6	2013-06-26 17:05:51 -07:00
Yaowu Xu	25fe05fd92	Change to use LUT for mode-to-txfm conversion Change-Id: Ieb989830f49e6708ee7728eddebf7a2144c37c6f	2013-06-26 14:10:43 -07:00
Dmitry Kovalev	be07485e9a	General cleanup in segmentation-related code. Using consistent function and variable names. Change-Id: I2deb3fded8797453a2081836c9ce2e79ade06eb7	2013-06-26 10:27:28 -07:00
John Koleszar	8137e24f3d	Merge "Move vp9_counts_to_nmv_context to encoder"	2013-06-25 22:44:21 -07:00
John Koleszar	7bbb0633cd	Merge "Move vp9_full_to_model_counts to encoder"	2013-06-25 22:44:16 -07:00
Jingning Han	3cc8c8c3a0	Merge "Refactor intra predictor block"	2013-06-25 19:46:55 -07:00
Jingning Han	d19ea3861d	Refactor intra predictor block Remove vp9_intra4x4_predict(). Use the common intra prediction function for all block sizes. Change-Id: Ibd19d51dfa3da8bbdfb79ddeb81530b2e2089560	2013-06-25 16:33:13 -07:00
Dmitry Kovalev	6fb10f2de4	Renaming "nmv" to "mv". Change-Id: I8299f55c3b930221e52c2237f2ddea65b94fd33b	2013-06-25 15:19:18 -07:00
Ronald S. Bultje	c24d922396	Add averaging-SAD functions for 8-point comp-inter motion search. Makes first 50 frames of bus @ 1500kbps encode from 3min22.7 to 3min18.2, i.e. 2.3% faster. In addition, use the sub_pixel_avg functions to calc the variance of the averaging predictor. This is slightly suboptimal because the function is subpixel-position-aware, but it will (at least for the SSE2 version) not actually use a bilinear filter for a full-pixel position, thus leading to approximately the same performance compared to if we implemented an actual average-aware full-pixel variance function. That gains another 0.3 seconds (i.e. encode time goes to 3min17.4), thus leading to a total gain of 2.7%. Change-Id: I3f059d2b04243921868cfed2568d4fa65d7b5acd	2013-06-25 12:57:28 -07:00
Dmitry Kovalev	9467571777	Moving subexp encoding functions in separate vp9_dsubexp.c file. Change-Id: Idbb2ea80f764fa830fe2ddcfc54ef7fe232f05a8	2013-06-25 11:53:17 -07:00
Dmitry Kovalev	5ae096778e	Merge "Removing unused code."	2013-06-25 11:50:55 -07:00
Yaowu Xu	c2e3ee13e7	Merge "Changed size of mb_mode_context to 8 bits"	2013-06-25 10:44:47 -07:00
Scott LaVarnway	855e23ce8c	Merge "Small mode_info_context cleanup in filter_block_plane"	2013-06-25 10:34:19 -07:00
Dmitry Kovalev	87ee34aacb	Removing unused code. Removing block index (ib) parameter from get_tx_type_{8x8, 16x16} functions. Change-Id: Ia213335aae7a7cb027f97b9cc9b04519840250f1	2013-06-25 10:17:19 -07:00
Dmitry Kovalev	70e9622185	Merge "Removing find_seg_id and using vp9_get_pred_mi_segid instead."	2013-06-25 10:16:06 -07:00
Dmitry Kovalev	529679bd52	Merge "Transforming scale_mv_component_q4 into scale_mv_q4 function."	2013-06-25 10:15:33 -07:00
Scott LaVarnway	c787f40bc4	Small mode_info_context cleanup in filter_block_plane Unnecessary updates to xd->mode_info_context. Change-Id: I36d2d68ca48366f727548526726b1b5437f62968	2013-06-25 12:28:50 -04:00
Yaowu Xu	b9c934df8e	Merge "Enable sse2 implmentation of 8x8 ADST/DCT"	2013-06-25 09:13:22 -07:00
Jingning Han	a32a086d23	Enable sse2 implmentation of 8x8 ADST/DCT This commit makes use of the butterfly structure to enable the sse2 version implementation of 8x8 ADST/DCT hybrid transform coding. The runtime of hybrid transform module goes down from 1170 cycles to 245 cycles. Overall speed-up around 1.5%. Change-Id: Ic808ffd21ece8a9d0410d8c0243d7b6c28ac3b3f	2013-06-24 18:41:33 -07:00
John Koleszar	4ecd6dbead	Move vp9_counts_to_nmv_context to encoder This function only used from within vp9_encodemv.c. Change-Id: Ib3fc7c30b1e2d27321397ac474cbc8976bc1f4b1	2013-06-24 15:58:18 -07:00
John Koleszar	08b1798ae7	Move vp9_full_to_model_counts to encoder This function is not called from the decoder, so it doesn't need to be in common/. Change-Id: I6977dd462a25b4ff39c9c7e1b0b5b16aa58ee733	2013-06-24 15:46:15 -07:00
John Koleszar	ece724ae16	Merge "Remove unused vp9_build_intra_predictors_sb{y,uv}_s"	2013-06-24 15:08:58 -07:00
John Koleszar	ee4a7e4e46	Merge "Remove unused vp9_model_to_full_probs_sb()"	2013-06-24 15:08:54 -07:00
Scott LaVarnway	dfa2ecc3f1	Changed size of mb_mode_context to 8 bits This reduced the size of the MODE_INFO array (mip and prev_mip) by 425,568 bytes each for 1080p resolutions. Change-Id: Ifa513ec2d0a49e8ec0867ec90620762fb7f1261d	2013-06-24 17:11:16 -04:00
John Koleszar	858475a03a	Fix loopfilter of leftmost 4x4 edges in SB For cases where there's no transform set in bit 0 (the left edge of the SB) but bit 0 of mask_4x4_int is set (the edge 4 pixels from the left edge needs filtering), it was incorrectly being skipped before. This situation only happens on the leftmost edge of the image, as the edge at column 0 is intentionally skipped since there aren't pixels to the left to read. Change-Id: Ib2fbbcb40166e90af31b1a0e13b85b68c226cbd3	2013-06-24 08:26:00 -07:00
John Koleszar	9e7019f7df	Remove unused vp9_build_intra_predictors_sb{y,uv}_s The functions no longer referenced. Change-Id: If2705dfbc607f79ec8ec2242d5e03bec27a35aaf	2013-06-21 16:10:05 -07:00
John Koleszar	5c32215e27	Remove unused vp9_model_to_full_probs_sb() This function never referenced. Change-Id: I1c42cd355bfa88e17d169f7335a44be682af58cc	2013-06-21 15:38:55 -07:00
Dmitry Kovalev	f27f76dfb3	Transforming scale_mv_component_q4 into scale_mv_q4 function. Using MV instead of int_mv for function arguments. Change-Id: Ic25e13dccbc98fac1fa1b3255127e00cca2a57f6	2013-06-21 15:34:29 -07:00
Dmitry Kovalev	40141681c0	Removing find_seg_id and using vp9_get_pred_mi_segid instead. Change-Id: Ia40229903c08f14020e90e94cfdf494aba1be827	2013-06-21 13:05:10 -07:00
Ronald S. Bultje	54b2a59623	Implement SSE2 block_error. Change vp9_block_error() to return a 64bit error variable, change all callers to expect a 64bit return value (this will prevent overflows, which we basically don't check for at all right now). Remove duplicate block_error() function, which fixed that through truncation. Remove old (incompatible) mmx/sse2 block_error SIMD versions and replace with a new one that returns a 64bit value. Encoding time of first 50 frames of bus @ 1500kbps goes from 3min29 to 3min23, i.e. a 3% overall speedup. Change-Id: Ib71ac5508b5ee8a80f1753cd85d72df1629abe68	2013-06-21 12:54:52 -07:00
Ronald S. Bultje	7756e9892b	Merge "Add subtract_block SSE2 version and unit test."	2013-06-21 12:49:50 -07:00
Ronald S. Bultje	9a480482cb	Merge "SSE2/SSSE3 optimizations and unit test for sub_pixel_avg_variance()."	2013-06-21 12:49:43 -07:00
Ronald S. Bultje	25c588b1e4	Add subtract_block SSE2 version and unit test. 3% faster overall (3min35.0 to 3min28.5). Change-Id: I5ff8a5c2c91586b6632ca5009ad1ea51ce94af5e	2013-06-21 09:35:37 -07:00
Yaowu Xu	e6cd5ed307	Merge "Implement sse2 and ssse3 versions for all sub_pixel_variance sizes."	2013-06-20 17:42:50 -07:00
Ronald S. Bultje	1e6a32f1af	SSE2/SSSE3 optimizations and unit test for sub_pixel_avg_variance(). Encoding of bus @ 1500kbps (first 50 frames) goes from 3min57 to 3min35, i.e. approximately a 10.5% speedup. Note that the SIMD versions which use a bilinear filter (x_offset & 7 \|\| y_offset & 7) aren't perfectly interleaved, and can probably be improved further in the future. I've marked this with a few TODOs/FIXMEs in the code. Change-Id: I5c9e900c0f0d32e431a50fecae213b510b2549f9	2013-06-20 15:59:48 -07:00
Frank Galligan	c259af4f73	Fix win64 warning. - size_t vs int. Change-Id: Ib47ebd932a4b69db9f52a43000bb69d0a96b9134	2013-06-20 14:07:11 -07:00
Dmitry Kovalev	8283d893eb	Merge "Renaming 'nmv' to 'mv' for several functions."	2013-06-20 10:17:12 -07:00
Ronald S. Bultje	8fb6c58191	Implement sse2 and ssse3 versions for all sub_pixel_variance sizes. Overall speedup around 5% (bus @ 1500kbps first 50 frames 4min10 -> 3min58). Specific changes to timings for each function compared to original assembly-optimized versions (or just new version timings if no previous assembly-optimized version was available): sse2 4x4: 99 -> 82 cycles sse2 4x8: 128 cycles sse2 8x4: 121 cycles sse2 8x8: 149 -> 129 cycles sse2 8x16: 235 -> 245 cycles (?) sse2 16x8: 269 -> 203 cycles sse2 16x16: 441 -> 349 cycles sse2 16x32: 641 cycles sse2 32x16: 643 cycles sse2 32x32: 1733 -> 1154 cycles sse2 32x64: 2247 cycles sse2 64x32: 2323 cycles sse2 64x64: 6984 -> 4442 cycles ssse3 4x4: 100 cycles (?) ssse3 4x8: 103 cycles ssse3 8x4: 71 cycles ssse3 8x8: 147 cycles ssse3 8x16: 158 cycles ssse3 16x8: 188 -> 162 cycles ssse3 16x16: 316 -> 273 cycles ssse3 16x32: 535 cycles ssse3 32x16: 564 cycles ssse3 32x32: 973 cycles ssse3 32x64: 1930 cycles ssse3 64x32: 1922 cycles ssse3 64x64: 3760 cycles Change-Id: I81ff6fe51daf35a40d19785167004664d7e0c59d	2013-06-20 09:34:25 -07:00
Jim Bankoski	2c6bdbbc78	new debug modes code The new print out includes skips and has prefixed sections so you can grep to find things like transforms chosen on each frame. Change-Id: I195043424647d9514cfc3ff6720a5b20d010fa1b	2013-06-20 09:33:11 -07:00
Yaowu Xu	12180c8329	Remove unnecessary copying of probs. Change-Id: Ic924f07c6ab0c929c6cdf11880d3c625806e272c	2013-06-18 23:02:27 -07:00
Dmitry Kovalev	87e1fa7627	Renaming 'nmv' to 'mv' for several functions. Change-Id: I183a38997a9d01e4a1b869e92509f6915216fa09	2013-06-18 18:28:10 -07:00
Jingning Han	7088426976	Merge "Make fdct32 computation flow within 16bit range"	2013-06-18 11:40:14 -07:00
Dmitry Kovalev	dfc0385291	Merge "Removing vp9_invtrans.{c, h} files."	2013-06-18 10:16:25 -07:00
Jingning Han	a41a4860c0	Make fdct32 computation flow within 16bit range This commit makes use of dual fdct32x32 versions for rate-distortion optimization loop and encoding process, respectively. The one for rd loop requires only 16 bits precision for intermediate steps. The original fdct32x32 that allows higher intermediate precision (18 bits) was retained for the encoding process only. This allows speed-up for fdct32x32 in the rd loop. No performance loss observed. Change-Id: I3237770e39a8f87ed17ae5513c87228533397cc3	2013-06-18 09:46:24 -07:00
Ronald S. Bultje	d9fc451666	Move subpixel variance function from common/ to encoder/. This seems to only be used in the encoder. Also remove an empty wrapper file that contained forward declarations for this function, but didn't actually define any actual functions. Change-Id: Ifc561eef7ebe374a7d03698055e51e105f6d614b	2013-06-17 16:54:09 -07:00
Dmitry Kovalev	686b99741c	Removing vp9_invtrans.{c, h} files. Moving single function from vp9_invtrans.c to vp9_encodemb.c. Change-Id: I26bf6bb90de342a3036c0dbfba78a7dd75a61fe7	2013-06-17 16:09:03 -07:00
John Koleszar	61ecc282b5	Merge "Remove unused need_to_clamp_mvs"	2013-06-17 10:31:58 -07:00
John Koleszar	141ab2d5d0	Merge "Fix type mismatch in array definition"	2013-06-14 17:07:22 -07:00
John Koleszar	c2da365484	Merge "Remove constant vp9_coef_update_prob table"	2013-06-14 17:07:19 -07:00
John Koleszar	a9415d2e4c	Fix type mismatch in array definition vp9_default_inter_mode_probs was being accessed with a different type than it was defined with. Ensure that its declaration is included prior to its definition. Change-Id: I2f963f513ab2f4e339f8a3c17e3d0f03749eba16	2013-06-14 16:38:42 -07:00
John Koleszar	0f7a66e962	Remove constant vp9_coef_update_prob table All elements of this table are equal to 252, so replace it with a single constant VP9_COEF_UPDATE_PROB. Change-Id: I1e2d1d284326ce6df9899a740c2fc344b3ec81c9	2013-06-14 15:12:31 -07:00
Jingning Han	0b7910b9ff	Merge "Enable sse2 version of sad8x4/4x8"	2013-06-14 13:15:49 -07:00
Jingning Han	c43af9a8a3	Enable sse2 version of sad8x4/4x8 The encoding time for bus at CIF goes from 661s to 625s. This commit also enabled unit test of sad8x4/4x8 in sad_test.cc. Change-Id: If3d10ebb56bda584bdb69bcf056599d580b12cb1	2013-06-14 09:19:28 -07:00
Jingning Han	15f50e7b42	Enable sse2 version of sad8x4/4x8 The encoding time for bus at CIF goes from 661s to 625s. This commit also enabled unit test of sad8x4/4x8 in sad_test.cc. Change-Id: If3d10ebb56bda584bdb69bcf056599d580b12cb1	2013-06-13 16:18:18 -07:00
John Koleszar	8e47093c9e	Remove unused need_to_clamp_mvs This flag no longer needed. Change-Id: If13482015ddb92d225792ea5c0ee455d2285d1f6	2013-06-12 16:50:14 -07:00
Scott LaVarnway	a81bd12a2e	Quick modifications to mb loopfilter intrinsic functions Modified to work with 8x8 blocks of memory. Will revisit later for further optimizations. For the HD clip used, the decoder improved by almost 20%. Change-Id: Iaa4785be293a32a42e8db07141bd699f504b8c67	2013-06-12 19:23:03 -04:00
Yaowu Xu	d682243012	Merge "Quick modifications to wide loopfilter intrinsic functions"	2013-06-12 15:16:11 -07:00
Ronald S. Bultje	fa96eeb835	Implement SSE version for sad4x8x4d and SSE2 version for sad8x4x4d. Encoding time of crew (CIF, first 50 frames) @ 1500kbps goes from 4min56 to 4min42. Change-Id: I92c0c8b32980d2ae7c6dafc8b883a2c7fcd14a9f	2013-06-12 17:40:01 -04:00
Scott LaVarnway	26496c52bf	Quick modifications to wide loopfilter intrinsic functions Modified to work with 8x8 blocks of memory. Will revisit later for further optimizations. For the HD clip used, the decoder improved my 20%. Change-Id: Ia0057f55d66d1445882351ea6c43b595a5a980e5	2013-06-12 16:49:08 -04:00
John Koleszar	1fa04e1a03	Merge changes I86fe51b0,I4c9a9e0f * changes: Remove unused vp9_idct_add_{y,uv}_block Remove some unused loopfilter code	2013-06-12 13:43:30 -07:00
Johann	bbd5cb2bd4	Merge "Fix compile warnings on windows."	2013-06-12 13:36:50 -07:00
John Koleszar	495ff8e0c7	Merge "Enable mmx loop filter routines"	2013-06-12 12:52:04 -07:00
John Koleszar	ceee4563d6	Remove unused vp9_idct_add_{y,uv}_block These functions are not used, and appear to have been superceded. Change-Id: I86fe51b088264f6b1b8d4d232bba97b371b98120	2013-06-12 12:24:22 -07:00
Jingning Han	1a5bb3cc76	Fix the comments in boundary block partition check Change-Id: Ic6b2881d8d495269edbc514b33376ca963798b45	2013-06-12 12:05:06 -07:00
John Koleszar	8933a652fc	Remove some unused loopfilter code This code is unreachable, and not useful for later reference. Change-Id: I4c9a9e0fbf859c1081bbcfbcda9710afb4b4741f	2013-06-12 11:36:00 -07:00
Frank Galligan	4524548f80	Fix compile warnings on windows. Change-Id: If74bc6110016bc75ea3883ab136fbbac88f6a913	2013-06-12 11:34:15 -07:00
John Koleszar	0e1e16db90	Enable mmx loop filter routines The mmx routines work as expected for the loop filter, so enable them. Change-Id: I2bbd9b99a4445fcba17bb95002f1fb6e01fe8f85	2013-06-12 11:28:21 -07:00
Yaowu Xu	efe05b7437	fix a mis use of ref_frame Change-Id: I9aac140d775b7b4a8727494d15b185b75501a546	2013-06-12 10:32:38 -07:00
Frank Galligan	15f9077ee2	Fix duplicate const. Change-Id: I86be1f7421ed49d577cacf405f6e4b0daa85cfdc	2013-06-12 08:52:34 -07:00
John Koleszar	9831f20594	Disallow wide loopfilter on some chroma borders Don't do the 15 tap filter if there aren't 8 pixels below/right of the edge. Change-Id: I62f16437c1d9ba59b6901a5fe71ddb2f472da344	2013-06-11 11:28:38 -07:00
Jingning Han	551f37d63d	Fix partition coding of corner block This commit fixed the allowable partition types for bottom-right corner blocks. When a block has over half of its pixels as valid content in both vertical and horizontal directions, allow all the four partition types in the bit-stream. Otherwise, apply partition type constraints. Change-Id: I2252e2de7125a8bfb1c824bf34299a13c81102e3	2013-06-10 21:43:17 -07:00
Deb Mukherjee	51a7c7631d	Merge "New probs for filters/tx_size and a few others" into experimental	2013-06-10 16:39:43 -07:00
Deb Mukherjee	a43ff15399	New probs for filters/tx_size and a few others * New probs for subpel filters/tx_count * Makes a change to not reset to defaults for the tx_size probs if an intermediate frame reverts to using a fixed tx_size. * A few updates to the parameters for backward adaptation for mode/mv * some cosmetic cleanups derf300: +0.06% Change-Id: I22994d659bc31ca7a4fc8820fde24001e64a2920	2013-06-10 16:38:47 -07:00
John Koleszar	091e23c3e6	Merge "Remove remnants of VP8 profiles/versions" into experimental	2013-06-10 16:16:17 -07:00
John Koleszar	0fcb625e35	Remove remnants of VP8 profiles/versions Remove the bilinear filter mode, and the no-loopfilter mode, and the related vp9_setup_version() function. Change-Id: I32311367812faf37863131df3af37d63d03973d7	2013-06-10 15:55:03 -07:00
Jim Bankoski	ba2af976cb	print debugging info from mode info struct This commit has no impact but to help us debug issues. To Use call like this: vp9_print_modes_and_motion_vectors(cpi->common.mi, cpi->common.mi_rows, cpi->common.mi_cols, cpi->common.current_video_frame, "decode_mi.stt"); Change-Id: I89e27725dae351370eb7f311a20a145ed4f1d041	2013-06-10 14:03:17 -07:00
John Koleszar	44db42c114	Merge the new loopfilter experiment Change-Id: I524ba98841f2e1850e3276ac365c501cea31546d	2013-06-10 12:30:12 -07:00
John Koleszar	c37a1e5ef2	Merge "Loopfilter: Fix chroma edge selection" into experimental	2013-06-10 12:17:24 -07:00
John Koleszar	2f3cbfdde1	Merge "Fix use of get_uv_tx_size in loopfilter" into experimental	2013-06-10 12:17:11 -07:00
Adrian Grange	c4e5b77d74	Merge "Implement intra-coded frames" into experimental	2013-06-10 12:08:09 -07:00
Deb Mukherjee	995ce523eb	Cosmetic cleanups of filters No bitstream change. Removes unused filters and the code for the case of 2 switchable filters; also changes the 8tap-smooth filter coefficients for integer shifts to be interpolating to be consistent with the way it is implemented currently. Change-Id: I96c542fd8c06f4e0df507a645976f58e6de92aae	2013-06-10 12:06:36 -07:00
Adrian Grange	eac344ef10	Implement intra-coded frames Implements ability to signal and decode frames that are encoded using only intra coding modes. Only the decode side has been implemented here. Change-Id: I53ac6a8d90422cd08ba389e5236e15b45f9e93de	2013-06-10 11:43:16 -07:00
John Koleszar	48b7cbcac5	Loopfilter: Fix chroma edge selection A 32x32 transform should have no internal filtering (check c==4) Change-Id: I7414cf4748ed053208217692ef00cd8b20d49a91	2013-06-10 11:40:57 -07:00
John Koleszar	717d744a01	Fix use of get_uv_tx_size in loopfilter Change the argument of get_uv_tx_size() to be an MBMI pointer, so that the correct column's MBMI can be passed to the function. Change-Id: Ied6b8ec33b77cdd353119e8fd2d157811815fc98	2013-06-10 11:40:57 -07:00
John Koleszar	ec38b6150d	Merge "Fixed point reference picture scaling" into experimental	2013-06-10 09:45:34 -07:00
Ronald S. Bultje	549258b1c2	Merge "border mvref issue" into experimental	2013-06-10 09:22:49 -07:00
Jim Bankoski	75459d65df	border mvref issue Fixes mvref issue. Change-Id: I07dc1b0682845bc18fe0efa6af5e4f4da3abfa3a	2013-06-10 09:21:11 -07:00
Yaowu Xu	7f99844e91	Merge "Loopfilter: bug fix in sb_type usage" into experimental	2013-06-10 08:56:38 -07:00
Tero Rintaluoma	86bb6df005	Fixed point reference picture scaling Fixed point scaling factors are calculated once for each reference frame by using integer division. Otherwise fixed point scaling routines are used in all scaling calculations. This makes it possible to calculate fixed point scaling factors on device driver software and pass them to hardware and thus avoid division on hardware. TODO: - Missing check for maximum frame dimensions (currently scaling uses 14 bits) - Missing check for maximum scaling ratio (upscaling 16:1, downscaling 2:1) Problems: - Straightforward fixed point implementation can cause error +-1 compared to integer division (i.e. in x_step_q4). Should only be an issue for frames larger than 16k. Change-Id: I3cf4dabd610a4dc18da3bdb31ae244ebaf5d579c	2013-06-10 08:07:55 -07:00
Janne Salonen	548f90d2ce	Loopfilter: bug fix in sb_type usage Was always using sb_type of first column in a row of 8x8 units when determining decoded block edges as a subcondition for loop filter skipping. Change-Id: Ib17554633a63a90b70cdaa7bed65db035a8ad9d8	2013-06-10 06:40:05 -07:00
Yaowu Xu	4852a8023d	Merge "Loopfilter: Always filter intra edges" into experimental	2013-06-09 21:18:00 -07:00
Yaowu Xu	9c44ce9f4b	Merge "Loopfilter: use the current block only for skip" into experimental	2013-06-09 21:17:54 -07:00
Yaowu Xu	2e1fd0a497	Merge "Modified loop filter edge skipping" into experimental	2013-06-09 21:17:47 -07:00
John Koleszar	140ac34e57	Loopfilter: Always filter intra edges Change-Id: Ifb1ce2bd52147981ca1aec9ec6cfea8738a23e45	2013-06-09 09:02:47 -07:00
Ronald S. Bultje	c3f9b070ca	Merge "New comp_inter defaults." into experimental	2013-06-09 06:40:02 -07:00
Ronald S. Bultje	3993d30922	Merge "Fix firstpass if framesize is not a multiple of 16." into experimental	2013-06-08 17:40:17 -07:00
Ronald S. Bultje	d30968c32a	Merge "New default tables" into experimental	2013-06-08 17:39:50 -07:00
Ronald S. Bultje	20760254f6	Merge "Align frame size to 8 instead of 16." into experimental	2013-06-08 17:39:41 -07:00
Ronald S. Bultje	99e10253b0	New comp_inter defaults. It seems like I inverted the meaning of the contexts by accident? Change-Id: Iafb2346d9933930949578342b84519b719dd5dd3	2013-06-08 15:13:57 -07:00
Ronald S. Bultje	073c7d5eec	Fix firstpass if framesize is not a multiple of 16. Change-Id: Iec41736c2b6140715f90f40de5ae6cf52497a9b8	2013-06-08 13:32:05 -07:00
Ronald S. Bultje	b64be43998	New default tables Change-Id: Ice8c73a2a843113877b8f8ed78737a1442c25ced	2013-06-08 13:29:14 -07:00
Deb Mukherjee	17da2cab78	TX_SIZE contexts simplification. Reduces TX_SIZE contexts to 2 for each kind. The code is cleaner and there is hardly any performance difference with more than two contexts. Results: almost neutral Change-Id: I17656bd6db76224ae2856adf882504560e7dbaa4	2013-06-08 12:32:26 -07:00
Deb Mukherjee	67cb1f093c	Minor fix in TX_SIZE contexts Change-Id: I9e81f84877e18ba7e55d66389ed60e64a5b7abcc	2013-06-08 07:14:58 -07:00
Yaowu Xu	b7da6d0c5a	Merge "Handle partition type coding of boundary blocks" into experimental	2013-06-07 18:16:16 -07:00
John Koleszar	f7e4b72df8	Loopfilter: use the current block only for skip Use the current block's skip flag to determine edge skipping. Change-Id: I4ba81f899286afbc3f6bb83eba2ef146a01b6fa4	2013-06-07 17:48:57 -07:00
Ronald S. Bultje	71701f3d40	Align frame size to 8 instead of 16. Change-Id: Ic606ef1b31e49963a779455a1e010a9ebb0f3f1f	2013-06-07 17:20:50 -07:00
Adrian Grange	07a5777bde	Frame header changes to support intra_only frames Made changes to the frame header to write the sync code in the frame header for a non-displayable, intra-only frame. Extended reset_frame_context to 2-bits. (Submitting on behalf of Dmitri) Change-Id: Ie836ae0df9ed572fb4f08aabe9351a555c4f3b96	2013-06-07 16:19:34 -07:00
Deb Mukherjee	21401942b0	Coding tx-size selection by use of spatial context Adds coding of transform size within a frame by use of context of transform sizes selected in left and above blocks. Also incorporates code for generating stats. TODO: generate and incorporate new default stats Change-Id: I6a7af099f6ad61d448521d9a51167aedaf638ed6	2013-06-07 16:07:58 -07:00
Deb Mukherjee	869a39ba60	Cleans up mbskip encoding Refactors mbskip coding to be compatible with coding of the rest of the symbols. Adds forward/backward adaptation and removes a lot of the legacy code. Results: fast50: +1.6% derfraw300: +0.317% Change-Id: I395a2976d15af044d3b8ded5acfa45f6f065f980	2013-06-07 16:00:26 -07:00
Jingning Han	78b8190cc7	Handle partition type coding of boundary blocks The partition types of blocks sitting on the frame boundary are constrained by the block size and the position of each sub-block relative to the frame. Hence we use truncated probability models to handle the coding of such information. 100 frames run: yt 0.138% Change-Id: I85d9b45665c15280069c0234ea6f778af586d87d	2013-06-07 14:19:40 -07:00
Ronald S. Bultje	28164eb962	Fix segment feature data size. Change-Id: I4331cfd99a717938f4f970cad81c468cbf287b00	2013-06-07 13:57:28 -07:00
Ronald S. Bultje	fb1f6f1db4	Fix segment feature data type. It has a range of -255,255, so should be int16_t, not int8_t. Change-Id: I5ef4b6aefb6212b0f35f4754f3c4d73fddbc52a0	2013-06-07 13:57:27 -07:00
Ronald S. Bultje	363dc6ceda	Don't crash if motion vector ref points to out-of-bounds area. This can only happen if partition is partly out-of-frame, in which case the referenced mv is either out-of-frame also (and thus has the same value as an already-read one), or it is actually uninitialized, in which case we don't want to use it. Change-Id: Icf39fa4d987c7abcbebb9bbdcdd6311e8fb9d3c9	2013-06-07 13:57:27 -07:00
Paul Wilkins	340c7a48e6	Change to segment ref frame feature. Simplify feature to only support a single reference frame instead of a mask. Change-Id: I5dd3a98c7a224aafb35708850ab82e2f220e68fb	2013-06-07 21:42:22 +01:00
Yaowu Xu	0bb6da3668	Merge "Remove two un-used entries in mode_lf_delta[]" into experimental	2013-06-07 10:10:45 -07:00
Yaowu Xu	254f46bc5b	Merge "Specify mv neighborhood for block larger than 8x8" into experimental	2013-06-07 10:09:35 -07:00
Yaowu Xu	b097a3ba82	Remove two un-used entries in mode_lf_delta[] With the removal of i4X4 and SPLIT_MV modes, the two entries for the modes are no longer used. This patch remove the coding of the deltas. Change-Id: Iea4eb500404ebe9706159380a03b8eca542fb4c3	2013-06-07 09:24:09 -07:00
Deb Mukherjee	78fbaf4d84	Merge "Coding updates for tx-size selection" into experimental	2013-06-07 09:19:36 -07:00
Ronald S. Bultje	def6bc765c	Merge "Revert "Align frame size to 8 instead of 16."" into experimental	2013-06-07 09:01:33 -07:00
Yaowu Xu	8b3ad75266	Specify mv neighborhood for block larger than 8x8 The new neighorbhood adapts to the shape and size of the block type cif +.16% stdhd +.13% Change-Id: I978db58278e9ae3fbd6726ef831bdfc5f5f37d02	2013-06-07 08:59:48 -07:00
Ronald S. Bultje	e7d306aae6	Revert "Align frame size to 8 instead of 16." This reverts commit `c2574414d4` Change-Id: Ie9013cb0bb43e639e01b4588f630b1da59295d38	2013-06-07 08:59:27 -07:00
Deb Mukherjee	3ee1a21a42	Coding updates for tx-size selection Changes to the coding of transform sizes, along with forward and backward probability updates. Results: derf300: +0.241% Context based coding of transform sizes will be in a separate patch. Change-Id: I97241d60a926f014fee2de21fa4446ca56495756	2013-06-07 08:54:00 -07:00
Janne Salonen	5c5223860a	Modified loop filter edge skipping Added condition to not to skip filtering of transform block edges when the edge is also a decoding block edge. Change-Id: Iaccb6206c4202b78e5dca3b89379556e0f4aba0c	2013-06-07 06:36:22 -07:00
Paul Wilkins	576c2bb021	Fix bug in segment skip. Wrong max data size (skip has no data) and use of vp9_get_segdata() when it should be vp9_segfeature_active(). Change-Id: I1eb97d33df6e2a42cc589049f704266fe3639902	2013-06-07 13:27:08 +01:00
Yaowu Xu	4df9e7883c	Merge "Removed rectangular intra prediction code" into experimental	2013-06-06 22:58:07 -07:00
Yaowu Xu	472669befb	Fix a merge conflict ref_frame in MB_Mode_Info was changed in the ref frame coding patch to be an array to handle first and second reference frame, this patch fix the loop filter code that use the pointer directly as reference frame. Change-Id: I71afa5a49deb50c1bc38029fd07470b984c6dfe9	2013-06-06 22:10:07 -07:00
Yaowu Xu	9470c1a2a1	Removed rectangular intra prediction code As all intra predictions happen on squared transform block now. Change-Id: I7ec91e3f0ad01383a03d2bd3099bbf32e87e3466	2013-06-06 21:35:10 -07:00
Jim Bankoski	fa9db8da15	Merge "Fix FIXME." into experimental	2013-06-06 20:50:51 -07:00
Jim Bankoski	686f437264	Merge "Align frame size to 8 instead of 16." into experimental	2013-06-06 20:49:59 -07:00
John Koleszar	736c7b804a	Merge "Reimplementation of loop filter" into experimental	2013-06-06 17:34:26 -07:00
Ronald S. Bultje	c2574414d4	Align frame size to 8 instead of 16. Change-Id: Ic22f416a33de558519d5c30a929f6a954546ade9	2013-06-06 17:28:11 -07:00
Ronald S. Bultje	bc41af00cf	Fix FIXME. Change-Id: I47a9857d35da1bff6153f8090c6b98b689b31a61	2013-06-06 17:28:11 -07:00
Ronald S. Bultje	6ef805eb9d	Change ref frame coding. Code intra/inter, then comp/single, then the ref frame selection. Use contextualization for all steps. Don't code two past frames in comp pred mode. Change-Id: I4639a78cd5cccb283023265dbcc07898c3e7cf95	2013-06-06 17:28:09 -07:00
Ronald S. Bultje	ad34368786	New intra mode and partitioning probabilities. Split partition probabilities between keyframes and non-keyframes, since they are fairly different. Also have per-blocksize interframe y intramode probabilities, since these vary heavily between different blocksizes. Lastly, replace default probabilities for partitioning and intra modes with new ones generated from current codec. Replace counts with actual probabilities also. Change-Id: I77ca996e25e4a28e03bdbc542f27a3e64ca1234f	2013-06-06 10:45:30 -07:00
John Koleszar	043d348aae	Reimplementation of loop filter This version of the loop filter supports non-4:2:0 subsampling and a fourth plane, as well as changing the filtering order to be more friendly to hardware implementations. The filters are applied first to all vertical edges within the 64x64 SB, followed by the top horizontal edge and any internal horizontal edges. Since filtering is applied on each 4x4 edge serially, a dependency is created from filtering one block edge to the next. It would be possible to remove this depencnecy by building all filtering decisions from the unfiltered reconstruction data. Change-Id: I08f3e9683eb7bded8a76651cbc50fc0dfdd05fa7	2013-06-06 08:45:45 -07:00
Jim Bankoski	5a88271b09	don't tokenize & encode tokens for blocks in UMV This avoids encoding tokens for blocks that are entirely in the UMV border. This changes the bitstream. Change-Id: I32b4df46ac8a990d0c37cee92fd34f8ddd4fb6c9	2013-06-06 06:10:25 -07:00
Dmitry Kovalev	28d31aed7f	Merge "Moving bits from compressed header to uncompressed one." into experimental	2013-06-06 01:15:44 -07:00
Jingning Han	61e6586230	Merge "Fix UV intra coding rd loop" into experimental	2013-06-05 21:47:00 -07:00
Jingning Han	f04b15486a	Fix UV intra coding rd loop This commit makes the coding/reconstruction operations of intra coding rate-distortion loop for UV components consistent with those of the encoding process. key frame coding gains: derf: 0.11% stdhd: 0.42% Change-Id: I8d49f83924a320e3689ef2d60096c49d7f0c7a40	2013-06-05 21:18:02 -07:00
Dmitry Kovalev	12345cb391	Moving bits from compressed header to uncompressed one. Bits moved: refresh_frame_flags, active_ref_idx[], ref_frame_sign_bias[], allow_high_precision_mv, mcomp_filter_type, ref_pred_probs[]. Derf results: +0.040% Change-Id: I011f43c7eac0371d533b255fd99aee5ed75b85a5	2013-06-05 20:56:37 -07:00
Deb Mukherjee	30226a658f	Cosmetic renaming VP9_MVREFS to VP9_INTER_MODES NO bitstream change Change-Id: I79f6146dac5fdd157051b6f8dc611c0b7b5e5f7f	2013-06-05 11:24:01 -07:00
Deb Mukherjee	83885235a7	Clean-ups on switchable interpolation and mv_ref Adds backward adaptation and differential forward updates of switchable interpolation filter probabilities. Also adds some cosmetic cleanups and minor fixes on mv_ref probabilities. derfraw300: +0.353% (with most coming from switchable interp changes) Change-Id: Ie2718be73528c945fd0d80cfd63ca2d9cb3032de	2013-06-05 10:11:52 -07:00
Yaowu Xu	0449ee0fec	Fix a off-by-one bug in the calculation of maximum number of tiles in log2 scale. Change-Id: Id283d6e51a8b926015fd3fc631cdbfb4b8268d4a	2013-06-03 14:25:28 -07:00
Paul Wilkins	6dd3a6320e	Merge "Replace scatter scan 32x32 with HW friendly scan." into experimental	2013-06-03 02:42:37 -07:00
Paul Wilkins	3f380d5252	Merge "vp9_find_mv_refs_idx change for last frame." into experimental	2013-06-03 02:34:46 -07:00
Dmitry Kovalev	317d832d38	Merge "Adding plane_block_width and plane_block_height functions." into experimental	2013-05-31 15:28:45 -07:00
Dmitry Kovalev	d771bba27e	Renaming 'motion_vector' to 'mv' for consistency. Change-Id: Ie869ea4992e26867caec46cb878fc86a646aeb9f	2013-05-31 12:32:53 -07:00
Dmitry Kovalev	120a878199	Adding plane_block_width and plane_block_height functions. Change-Id: I02c17fb733c0f3c22dc3167c3d3182797415f1ae	2013-05-31 12:31:49 -07:00
Ronald S. Bultje	a288cb3b10	Merge "Merge all various transform size data trackers into single variables." into experimental	2013-05-31 09:59:24 -07:00
Scott LaVarnway	1e025dbfd1	Merge "Moved use_prev_in_find_mv_refs check to frame level" into experimental	2013-05-31 09:35:51 -07:00
Ronald S. Bultje	e9d68a5e36	Merge all various transform size data trackers into single variables. Change-Id: I2dfc569106b29fbe4da20585a0e85e5e9ea6a4db	2013-05-31 09:18:59 -07:00
Paul Wilkins	cf61fae8ee	vp9_find_mv_refs_idx change for last frame. Restrict get_matching_candidate() to considering mvs at 8x8 and larger sizes for last frame case. This is to reduce the HW load of using vectors down to the 4x4 level from the previous frame. Change-Id: I6505e610fd63a4e22d67f136aec7905a01b893ba	2013-05-31 15:37:27 +01:00
Sami Pietila	0835a35347	Fix inter mode context adaptation. Change-Id: Ibaa47be878c1cd84d88d7518418d2d8d38224e70	2013-05-31 12:58:31 +03:00
Paul Wilkins	aaf61dfbca	Merge "Patch to remove implicit segmentation." into experimental	2013-05-31 02:56:20 -07:00
Yaowu Xu	7ca651a383	Merge "Changed to use a new variant of WHT" into experimental	2013-05-30 21:53:12 -07:00
Ronald S. Bultje	a4e7c6bd4d	Merge "Remove unused define." into experimental	2013-05-30 20:58:22 -07:00
Ronald S. Bultje	310bc1030a	Merge "Merge VP9_YMODES, VP9_UV_MODES, INTRA_MODE_COUNT and cousins." into experimental	2013-05-30 20:58:19 -07:00
Ronald S. Bultje	7d549870f7	Merge "Remove TX_SIZE_MAX_MB." into experimental	2013-05-30 20:58:16 -07:00
Ronald S. Bultje	6ea6f4d253	Merge "Remove one (unused) entry from mvref tables." into experimental	2013-05-30 20:58:13 -07:00
Jim Bankoski	21595f8e38	Merge "Creates a new speed 1:" into experimental	2013-05-30 20:36:05 -07:00
Jim Bankoski	ced21bd6a6	Creates a new speed 1: This speed 1 - uses variance threshold stolen from static-thresh to determine split. Any superblock with greater than the variance set by static thresh * quantizer index squared is split. In addition transform size is set to largest size less than or equal to partition size, sub pixel filter is set to normal, and only 12 modes are used at all. Change-Id: If7a2858ee70f96d1eb989c04fd87a332b147abef	2013-05-30 19:53:00 -07:00
Ronald S. Bultje	16482bddf7	Merge "Remove splitmv." into experimental	2013-05-30 19:07:12 -07:00
Ronald S. Bultje	d2205f92c3	Merge changes I98c18fe5,I80c37cff into experimental * changes: Remove i4x4_pred. Remove unused table.	2013-05-30 19:06:44 -07:00
Ronald S. Bultje	117282a690	Remove unused define. Change-Id: Ic6555128206d61f47a46c550cb3dcaf3b4ec6374	2013-05-30 17:21:06 -07:00
Ronald S. Bultje	a433abbcad	Merge VP9_YMODES, VP9_UV_MODES, INTRA_MODE_COUNT and cousins. These are now merged in a new define called VP9_INTRA_MODES. Change-Id: I0890f895756a7395d84c92f98f43e43f4cf9050d	2013-05-30 17:21:06 -07:00
Ronald S. Bultje	4d3d00b195	Remove TX_SIZE_MAX_MB. Change-Id: I715870513d1fef8471bfd0f5218a79360a1ef126	2013-05-30 17:21:06 -07:00
Ronald S. Bultje	580d29bdbb	Remove one (unused) entry from mvref tables. Change-Id: Ieb4669ae564bec9f3051485ecdf186cb4e00decb	2013-05-30 17:21:06 -07:00
Ronald S. Bultje	e6485581fe	Remove splitmv. We leave it in rdopt.c as a local define for now - this can be removed later. In all other places, we remove it, thereby slightly decreasing the size of some arrays in the bitstream. Change-Id: Ic2a9beb97a4eda0b086f62c039d994b192f99ca5	2013-05-30 17:21:01 -07:00
Ronald S. Bultje	1efa79d32f	Remove i4x4_pred. It remains as a local define in rdopt.c so we can distinguish between split and non-split modes in the RD loop, but disappears outside that scope in the codec. Change-Id: I98c18fe5ab7e4fbd1d6620ec5695e2ea20513ce9	2013-05-30 16:44:58 -07:00
Ronald S. Bultje	9175082c4e	Remove unused table. Change-Id: I80c37cffa176bac942ab3051abdfd585ed5555e1	2013-05-30 16:44:56 -07:00
Yaowu Xu	042e70e45e	Changed to use a new variant of WHT The commit changed to use a new variant of Walsh-Hadamard Transform by Tim Terriberry. This new variant has the best compression among a number of variants that developed by Tim. Change-Id: Icb3a88515463cfc644b17ca046fcd139db2557e9	2013-05-30 15:37:52 -07:00
Ronald S. Bultje	f5827699bf	Merge "Merge all intra mode coding trees into a single one." into experimental	2013-05-30 11:27:51 -07:00
Adrian Grange	6f361f5841	Merge "Add intra_only and reset_frame_context flags" into experimental	2013-05-30 10:56:25 -07:00
Ronald S. Bultje	98c192ae83	Merge all intra mode coding trees into a single one. Also merge all counters. This removes a few unused probability updates from the bitstream. Change-Id: I20f58853e9dac84d8c0d9703ae012c55917516eb	2013-05-30 09:58:53 -07:00
Deb Mukherjee	c98bfcfbbb	Merge "Balancing coef-tree to reduce bool decodes" into experimental	2013-05-30 08:10:47 -07:00
Sami Pietila	5700b4ea42	Replace scatter scan 32x32 with HW friendly scan. The first 240 coeff positions (15 top-left blocks) are scanned in the same order as in scatter scan, after that the coeffs are scanned in "block bands", each band at a time, all coeffs in one band before moving on to the next band. This brings down the amount of 4x4 coeff blocks that need to be buffered while scanning, from 15 blocks to 8 blocks. Change-Id: I478a991d63c48bd5e64d36e59fed7a00c9a651ba	2013-05-30 15:32:46 +03:00
Paul Wilkins	1b103f250f	Patch to remove implicit segmentation. This patch removes the implicit segmentation experiment from the code base as the benefits were still unproven as of the bitstream deadline. Change-Id: I273b99d8d621d1853eac4182f97982cb5957247e	2013-05-30 11:06:29 +01:00
Ronald S. Bultje	17544d1478	Merge "Remove some unused code related to macroblock/splitmv coding." into experimental	2013-05-29 17:35:05 -07:00
Ronald S. Bultje	7873de1481	Merge "Remove unused and outdated debug code." into experimental	2013-05-29 17:33:32 -07:00
Adrian Grange	9e5bb9598c	Add intra_only and reset_frame_context flags Added two flags to the frame header: intra_only: Signals that the frame is encoded using only INTRA coding modes. reset_frame_context: Indicates that the coding context specified in the frame header should be reset to default values before the frame is encoded/decoded. Change-Id: I182d46f1f84fb67a13c46ad767f246a38d7861a2	2013-05-29 17:16:00 -07:00
Deb Mukherjee	b8b3f1a46d	Balancing coef-tree to reduce bool decodes This patch changes the coefficient tree to move the EOB to below the ZERO node in order to save number of bool decodes. The advantages of moving EOB one step down as opposed to two steps down in the other parallel patch are: 1. The coef modeling based on the One-node becomes independent of the tree structure above it, and 2. Fewer conext/counter increases are needed. The drawback is that the potential savings in bool decodes will be less, but assuming that 0s are much more predominant than 1's the potential savings is still likely to be substantial. Results on derf300: -0.237% Change-Id: Ie784be13dc98291306b338e8228703a4c2ea2242	2013-05-29 16:25:52 -07:00
Dmitry Kovalev	38cb616fbf	Merge "Compressed/uncompressed frame header changes." into experimental	2013-05-29 15:29:44 -07:00
Scott LaVarnway	353642bc53	Moved use_prev_in_find_mv_refs check to frame level This patch checks at the frame level to see if the previous mode info context can be used. This patch eliminates the flag check that was done for every mode and removes another check that was done prior to every vp9_find_mv_refs(). Change-Id: I9da5e18b7e7e28f8b1f90d527cad087073df2d73	2013-05-29 16:42:23 -04:00
Jingning Han	6c97bba403	Merge "further clean-ups on intra4x4 coding" into experimental	2013-05-29 10:55:14 -07:00
Sami Pietila	88a4d4c510	Residual coding to cache energy class of tokens. Proposal for tuning the residual coding by changing how the context from previous tokens is calculated. Storing the energy class of previous tokens instead of the token itself eases the critical path of HW implementations. Change-Id: I6d71d856b84518f6c88de771ddd818436f794bab	2013-05-29 15:21:01 +01:00
Ronald S. Bultje	4487f5a690	Remove some unused code related to macroblock/splitmv coding. Change-Id: Ic40d56fb162f4e201547dfae33e62ccd9e865889	2013-05-29 06:29:56 -07:00
Ronald S. Bultje	2afc3422c6	Remove unused and outdated debug code. Change-Id: I0e789bdeaed60f920f7a470e56a8d4ea374233fc	2013-05-28 19:15:57 -07:00
Dmitry Kovalev	18c83b3714	Compressed/uncompressed frame header changes. Adding API to read/write uncompressed frame header bits (it is not final yet). Separate functions to read/write uncompressed header. Moving clr_type, error_resilient_mode, refresh_frame_context, frame_parallel_decoding_mode, frame_context_idx from compressed partition to uncompressed frame header. Change-Id: Id3ed8a387980c652ae147549412f4ec24a0a5bd0	2013-05-28 18:07:54 -07:00
Deb Mukherjee	3d4e032e16	Merge "Clean up related to coefficient modeling" into experimental	2013-05-28 16:55:02 -07:00
Deb Mukherjee	d8c0989d56	Clean up related to coefficient modeling Uses reduced arrays for probabilities and branch counts in the encoder. No change in bitstream. Change-Id: Iec605446f44db4cd325eb45fa12a3003a6ee29db	2013-05-28 16:32:03 -07:00
Jingning Han	4729a6f389	further clean-ups on intra4x4 coding Removed one 4x4 prediction step that was unnessary in the rd loop. Removed a unused modecosts estimate from encoder side. Change-Id: I65221a52719d6876492996955ef04142d2752d86	2013-05-28 11:19:05 -07:00
Paul Wilkins	245a11553a	Merge "Remove loop dering experiment." into experimental	2013-05-28 05:34:14 -07:00
Dmitry Kovalev	1a24011469	Revert "Adding API to read/write uncompressed frame header bits." because of bitstream mismatches. This reverts commit `df037b615f` Change-Id: I1a529f2590df7bc912f5035d22311268933e3dd6	2013-05-28 02:24:52 -07:00
Yaowu Xu	2b96ffe025	a few clean-ups 1. remove prediction mode conversion 2. unified bmode, same for key and non-key frame 3. set I4X4_PRED count for pdf to 0, as I4X4_PRED is no longer coded ever. It is determined by ref_frame and block partition Change-Id: If5b282957c24339b241acdb9f2afef85658fe47d	2013-05-27 13:53:56 -07:00
Timothy B. Terriberry	95339d6825	Reduce WHT complexity. Saves 1 add, 3 shifts (and a shift bias) per 1-D transform. Change-Id: I1104bb1679fe342b2f9677df8a9cdc0cb9699e7d	2013-05-27 13:23:52 -07:00
Jingning Han	de735929cf	Reduce bmi buffer length from 16 to 4 This commit removes the use of bmi_ in the first-pass encoding by forcing encode_intra4x4block_ to use DC_PRED, followed by DCT_DCT only, as John suggested. This makes the need for bmi buffer only up to 4 entries, instead of 16. Change-Id: I3410007dfae789ee46a09ae20c39d3ce3c7954aa	2013-05-27 08:59:15 -07:00
Ronald S. Bultje	5cac66078e	Remove splitmv. Also do per-partition motion vector referencing in <sb8x8 partitions, and adjust mvref finding for sub8x8 partitions. Change-Id: Id3ed1ed4d2a8910d11d327db6cc63b8eb79f941f	2013-05-26 14:40:49 -07:00
Paul Wilkins	845bc13ba9	Remove loop dering experiment. Change-Id: I1a979bf74c286b157c31bab6bdcba0494acb4918	2013-05-25 10:09:23 +01:00
Dmitry Kovalev	0b2b81249b	Merge "Adding API to read/write uncompressed frame header bits." into experimental	2013-05-24 13:43:19 -07:00
Paul Wilkins	e41fd6e3e2	Fix bug in 4x4 band definition. Also some unused data structures/references removed. Change-Id: I295809e887173543e794250cb60ddaf1475ffd24	2013-05-23 17:51:02 -07:00
Yaowu Xu	22694ca1ad	Change txfm_type decision The changing in intra coding to base on transform block, i.e. pred-> txfm->quant->dequant-itxfm->recon, made all blocks within a prediction unit behave consistently, there is no longer a need to handle blocks differently based on the position within a predicitn block. So this commit simplifies the decision of transform type to be based on prediction mode only. Change-Id: If96cb72386f2e9186126ace88afa35ef085b6c96	2013-05-23 17:46:28 -07:00
Paul Wilkins	33ecd6ad54	Merge Scatter Scan experiment. Removal from under configure flag. A bit renaming Change-Id: I2213229dfe852001dfec16b149f47c52ce88f3aa	2013-05-23 13:09:27 +01:00
Jingning Han	7ac5ac52f9	Merge 4x4 block level partition into codebase Move 4x4/4x8/8x4 partition coding out of experimental list. This commit fixed the unit test failure issues. It also resolved the merge conflicts between 4x4 block level partition and iterative motion search for comp_inter_inter. Change-Id: I898671f0631f5ddc4f5cc68d4c62ead7de9c5a58	2013-05-23 11:58:50 +01:00
Yunqing Wang	0812c121e7	Merge "Optimize variance functions" into experimental	2013-05-22 13:41:04 -07:00
Deb Mukherjee	ddb2309568	Merge "Using 128 entry look up table for coef models" into experimental	2013-05-22 10:38:35 -07:00
Yunqing Wang	f4fcfe3075	Optimize variance functions Added SSE2 version of variance functions for super blocks. Change-Id: Ibeaae8771ca21c99d41dd74067574a51e97b412d	2013-05-22 10:29:38 -07:00
Jingning Han	d2cacdc530	Merge "Make the intra rd search support 8x4/4x8" into experimental	2013-05-22 10:00:15 -07:00
Deb Mukherjee	de4d682ca4	Using 128 entry look up table for coef models Reverts to using 128 bit LUT for the coef models rather than 48 to ease hardware implementation. Also incorporates some cleanups including removing various hooks to support different lookup tables based on block_type and ref_type. Change-Id: I54100c120cca07a2ebd3a7776bc4630fa6a153f6	2013-05-22 08:44:31 -07:00
Paul Wilkins	4e08fa96f3	Merge "changes intra coding to be based on txfm block" into experimental	2013-05-22 06:53:12 -07:00
Paul Wilkins	22d7f0703a	Merge "Generalized intra 4x4 encoding for all sizes" into experimental	2013-05-22 06:52:32 -07:00
Scott LaVarnway	d679fdf7b0	Merge "Removed unused idct functions" into experimental	2013-05-22 05:36:36 -07:00
Yaowu Xu	8ba92a0bed	changes intra coding to be based on txfm block This commit changed the encoding and decoding of intra blocks to be based on transform block. In each prediction block, the intra coding iterates thorough each transform block based on raster scan order. This commit also fixed a bug in D135 prediction code. TODO next: The RD mode/txfm_size selection should take this into account when computing RD values. Change-Id: I6d1be2faa4c4948a52e830b6a9a84a6b2b6850f6	2013-05-22 11:53:19 +01:00
Yaowu Xu	232d90d8fd	Generalized intra 4x4 encoding for all sizes Change-Id: I1b86744fa247233c8df031b3f4b87b212c8dd094	2013-05-22 11:44:12 +01:00
Jingning Han	f153a5d063	Make the intra rd search support 8x4/4x8 This commit allows the rate-distortion optimization of intra coding capable of supporting 8x4 and 4x8 partition settings. It enables the entropy coding of intra modes in key frame using a unified contextual probability model conditioned on its above/left prediction modes. Coding performance: derf 0.464% Change-Id: Ieed055084e11fcb64d5d5faeb0e706d30268ba18	2013-05-21 21:03:00 -07:00
John Koleszar	ddf13be8ef	Merge "Initial version of alpha channel support" into experimental	2013-05-21 17:29:51 -07:00
Dmitry Kovalev	df037b615f	Adding API to read/write uncompressed frame header bits. The API is not final yet and can be changed. Actual layout of uncompressed frame part will be finalized later. Right now moving clr_type, error_resilient_mode, refresh_frame_context, frame_parallel_decoding_mode from first compressed partition to uncompressed frame part. Change-Id: I3afc5d4ea92c5a114f4c3d88f96858cccc15b76e	2013-05-21 15:31:32 -07:00
Scott LaVarnway	a143152600	Removed unused idct functions No longer used. Change-Id: Id28c9247cebba183c6fa786dff96824ae100132c	2013-05-21 17:59:54 -04:00
Deb Mukherjee	7a645e4e12	Merging the model coef prob experiment Merges the experiment. Change-Id: I4eb19af6de6df6aa3a96a2e82f231d47ed9b3ae9	2013-05-21 14:44:38 -07:00
Deb Mukherjee	90a7723f8c	Merge "Refinements on modelcoef expt to reduce storage" into experimental	2013-05-21 13:41:54 -07:00
Scott LaVarnway	3d0110fd8e	Removed diff from macroblockd_plane No longer used. Change-Id: I171c5fa33a7600ad45b9466af23a46ccbdfe0480	2013-05-21 14:22:09 -04:00
Scott LaVarnway	0c3f3bf1d5	Removed vp9_recon functions No longer used. Change-Id: Ica5166f7117f4693dffdf7633dcfc1b263103d0d	2013-05-21 13:57:50 -04:00
Scott LaVarnway	1db6373267	Merge "WIP: 4x4 idct/recon merge" into experimental	2013-05-21 10:45:53 -07:00
Dmitry Kovalev	060d93d704	Merge "Removing clamp_type from the bitstream." into experimental	2013-05-21 10:35:27 -07:00
Deb Mukherjee	07443f1589	Refinements on modelcoef expt to reduce storage Uses more aggrerssive interpolation to reduce storage for the model tables by almost more than half. Only 48 lists of probs are stored (as opposed to 128 before), corresponding to ONE_NODE probabilities of: 1, 3, 7, 11, ..., 115, 119, 127, 135, ..., 247, 255. Besides, only 1 table is used as opposed to 2 before. So the overall memory needed for the tables is just 48 * 8 = 384 bytes. The table currently used is based on a new Pareto distribution with heavier tail than a generalized Gaussian - which improves results on derf by about 0.1% over a single table Generaized Gaussian. Results overall on derfraw300 is -0.14%. Change-Id: I19bd03559cbf5894a9f8594b8023dcc3e546f6bd	2013-05-21 10:06:56 -07:00
Deb Mukherjee	39a90bc8e8	Updating the model coef experiment Cleans up the experiment. Actually uses reduced counts for backward updates, and reduced number of probabilities in the context. No change in bitstream when the experiment is on. Between expt on and off: derfraw300 is down only -0.062% (which is better than when expts were run previously). Change-Id: I55285a049a0c22810bdb42914212ab5a4f8521b5	2013-05-20 12:46:36 -07:00
Scott LaVarnway	ba48a11130	WIP: 4x4 idct/recon merge This patch eliminates the intermediate diff buffer usage by combining the short idct and the add residual into one function. The encoder can use the same code as well. Change-Id: I296604bf73579c45105de0dd1adbcc91bcc53c22	2013-05-20 13:03:17 -04:00
Jingning Han	810b612c23	Enable bit-stream support to 8x4 and 4x8 partition The recursive partition type search is enabled down to 4x4, 4x8 and 8x4, followed by the corresponding rate-distortion optimization for the per-partition encoding mode decisions. The bit-stream writing/reading synchronized in supporting the rectangular partition of 8x8 block. This provides above 1% coding performance gains on derf. To do next: 1. re-design the rate-distortion loop for inter prediction below 8x8. 2. re-design the rate-distortion loop for intra prediction below 4x4. 3. make the loop-filter aware of rectangular partition of 8x8 block. 4. clean the unused probability models. 5. update default probability values. Change-Id: Idd41a315b16879db08f045a322241f46f1d53f20	2013-05-19 14:59:04 -07:00
Dmitry Kovalev	498b6460a1	Removing clamp_type from the bitstream. Change-Id: Ica75bdd4905c4a04b7f92795d0b8ce6836a99ef4	2013-05-17 15:50:26 -07:00
Jim Bankoski	d9705b200f	Merge "holds utility debugging functions" into experimental	2013-05-17 09:23:36 -07:00
Paul Wilkins	6f0c8e82c0	Merge "Replace default counts with default probs." into experimental	2013-05-17 08:58:53 -07:00
Paul Wilkins	61e2eac61d	Merge "New inter mode context." into experimental	2013-05-17 06:59:09 -07:00
Paul Wilkins	99c4b1eea1	Replace default counts with default probs. Replace vp9_kf_default_bmode_counts structure with direct default probabilities. The probability structure is smaller and it removes the need to specify in the bitstream how to convert the counts to probabilities. Note that I have concerns still about the size and value of the large intra mode context. This may cause problems for HW but it also means we rely heavily on reverse update as forwards update of a structure this size is problematic. I intend to review this more generally in the next few days to see if we can come up with a competitive solution that does not rely on such a large context. Change-Id: I0a36071079d5d26a57ab0e9fbf91af4199aa7984	2013-05-17 14:54:52 +01:00
Jim Bankoski	b67e46b33c	holds utility debugging functions This one prints out a visual version of the partitioning for human eyes to follow... Change-Id: Iba434589a2f55eb069484686d99a382db93b9548	2013-05-17 06:29:28 -07:00
Paul Wilkins	51bc4bf4a0	Remove MODE_STATS flag and code Change-Id: I6c70a8a8a4633399842ac74792003ae5f7859ffa	2013-05-17 12:34:10 +01:00
John Koleszar	679e4abdd5	Initial version of alpha channel support This is a mostly-working implementation of an extra channel in the bitstream. Configure with --enable-alpha to test. Notable TODOs: - Add extra channel to all mismatch tests, PSNR, SSIM, etc - Configurable subsampling - Variable number of planes (currently always uses all 4) - Loop filtering - Per-plane lossless quantizer - ARNR support This implementation just uses the same contents as the Y channel for the A channel, due to lack of content and general pain in playing back 4 channel content. A later patch will use the actual alpha channel passed in from outside the codec. Change-Id: Ibf81f023b1c570bd84b3064e9b4b8ae52e087592	2013-05-16 22:21:09 -07:00
John Koleszar	f07602e403	Merge "Remove vp9_extend_mb_row()" into experimental	2013-05-16 15:22:38 -07:00
Scott LaVarnway	9aa37a51b2	Merge "WIP: 8x8 idct/recon merge" into experimental	2013-05-16 14:28:30 -07:00
Yaowu Xu	e3869e9cfc	Removed Q threshold in the usage of ADST Test on cif set showed small but consistent compression gain for almost all encodings with overall impact of .08%. The gains average aournd .12% combined with D63 adst change. Test encoding on std-hd set is ongoing.. Change-Id: If4d94799cf0486fb9c770b193e5c386d13d99d59	2013-05-16 13:33:07 -07:00
John Koleszar	16ac5a5cde	Remove vp9_extend_mb_row() This code is no longer needed for correct intra prediction. Change-Id: I822d1a8b0ad0a00e7c4c6e7b2931790c39d1267d	2013-05-16 11:56:00 -07:00
Scott LaVarnway	794a7bedbd	WIP: 8x8 idct/recon merge This patch eliminates the intermediate diff buffer usage by combining the short idct and the add residual into one function. The encoder can use the same code as well. Change-Id: Iacfd57324fbe2b7beca5d7f3dcae25c976e67f45	2013-05-16 13:52:15 -04:00
Jingning Han	8e3d0e4d7d	Add building blocks for 4x8/8x4 rd search These building blocks enable rate-distortion optimization search over block sizes of 8x4 and 4x8. Need to convert them into mmx/sse forms. Change-Id: I570ea2d22d14ceec3fe3575128d7dfa172a577de	2013-05-16 10:41:29 -07:00
Jingning Han	c0f70cca40	Merge "Fix the transform type selection in 4x4 partition" into experimental	2013-05-16 10:14:49 -07:00
Paul Wilkins	6ff3eb1647	New inter mode context. This patch creates a new inter mode contest that avoids a dependence on the reconstructed motion vectors from neighboring blocks. This was a change requested by a hardware vendor to improve decode performance. As part of this change I have also made some modifications to stats output code (under a flag) to allow accumulation of inter mode context flags over multiple clips Some further changes will be required to accommodate the deprecation of the split mv mode over the next few days. Performance as stands is around -0.25% on derf and std-hd but up on the YT and YT-HD sets. With further tuning or some adjustment to the context criteria it should be possible to make this change broadly neutral. Change-Id: Ia15cb4470969b9e87332a59c546ae0bd40676f6c	2013-05-16 12:09:19 +01:00
Paul Wilkins	18e07420a2	Merge "Further Implicit Segmentation Changes" into experimental	2013-05-16 03:03:44 -07:00
John Koleszar	7addafb5b1	Merge "Fix vp9_build_intra_predictors_sbuv_s for non-4:2:0" into experimental	2013-05-15 20:58:15 -07:00
John Koleszar	501ae3484c	Fix vp9_build_intra_predictors_sbuv_s for non-4:2:0 Remove an assumption about chroma size, and the number of planes. Change-Id: I286a7fac296ec334c6a8ad847f663f3adbb9f43e	2013-05-15 17:57:08 -07:00
Dmitry Kovalev	5c5582242d	Moving the same code to new function vp9_setup_scale_factors. Change-Id: I2408ad22717784a40e23701ccb9d978265440e4f	2013-05-15 16:33:36 -07:00
Jingning Han	8468a5c1a0	Fix the transform type selection in 4x4 partition This commit allows proper transform type (DCT/ADST) selection in the settings of partition 4x4 level. Change-Id: Iec6f922a46480d777e7ca9142a99e8c131f0077b	2013-05-15 16:09:58 -07:00
Dmitry Kovalev	cd16fe9160	Merge "Preparing vp9_deblock and vp9_denoise to alpha support." into experimental	2013-05-15 15:40:52 -07:00
Dmitry Kovalev	80c0715375	Merge "Moving several static functions from vp9_reconinter.h to vp9_reconinter.c." into experimental	2013-05-15 15:39:43 -07:00
Scott LaVarnway	a272ff25cd	WIP: 16x16 idct/recon merge This patch eliminates the intermediate diff buffer usage by combining the short idct and the add residual into one function. The encoder can use the same code as well. Change-Id: Iea7976b22b1927d24b8004d2a3fddae7ecca3ba1	2013-05-15 13:16:02 -04:00
Paul Wilkins	7d7e5b5131	Further Implicit Segmentation Changes Trial use of a combination of reference frame, prediction block size and mv to define segmentation. Change-Id: Ie8946a0446dbad777fdcf7626f89e5af0994db50	2013-05-15 16:00:06 +01:00
Dmitry Kovalev	6706e674b1	Moving several static functions from vp9_reconinter.h to vp9_reconinter.c. Change-Id: I5da9c16bab26f6ff0c9d3a2a29ef6c84f5093161	2013-05-14 17:49:41 -07:00
Scott LaVarnway	2cf0d4be12	WIP: 32x32 idct/recon merge This patch eliminates the intermediate diff buffer usage by combining the short idct and the add residual into one function. The encoder can use the same code as well. Change-Id: I4ea09df0e162591e420d869b7431c2e7f89a8c1a	2013-05-14 15:54:17 -07:00
Jingning Han	1f26840fbf	Enable recursive partition down to 4x4 This commit allows the rate-distortion optimization recursion at encoder to go down to 4x4 block size. It deprecates the use of I4X4_PRED and SPLITMV syntax elements from bit-stream writing/reading. Will remove the unused probability models in the next patch. The partition type search and bit-stream are now capable of supporting the rectangular partition of 8x8 block, i.e., 8x4 and 4x8. Need to revise the rate-distortion parts to get these two partition tested in the rd loop. Change-Id: I0dfe3b90a1507ad6138db10cc58e6e237a06a9d6	2013-05-14 12:39:56 -07:00
Dmitry Kovalev	7bbf716f04	Preparing vp9_deblock and vp9_denoise to alpha support. Change-Id: I299feefa64b93bd62263aea1ff1e41e85faeb6ca	2013-05-14 11:01:57 -07:00
Yaowu Xu	8be35347a8	Merge "changed to use adst for D63_PRED" into experimental	2013-05-14 09:30:17 -07:00
John Koleszar	d31a632dcd	Merge "Revert "Preparing vp9_deblock and vp9_denoise to alpha support."" into experimental	2013-05-14 06:47:44 -07:00
John Koleszar	56efb73be3	Revert "Preparing vp9_deblock and vp9_denoise to alpha support." This reverts commit `a933311131` Change-Id: I2321f88011178381adbcffeda1bcc6a430ab8f1d	2013-05-14 06:46:11 -07:00
Yaowu Xu	da3aec6c8a	changed to use adst for D63_PRED To be consistent with other prediciton modes Change-Id: If9e1464e5c807f0b36047a046c4ac59d91b1b868	2013-05-13 22:09:38 -07:00
Dmitry Kovalev	571aa44606	Merge "Preparing vp9_deblock and vp9_denoise to alpha support." into experimental	2013-05-13 21:48:37 -07:00
Dmitry Kovalev	ce70ad5967	Merge "Code cleanup inside vp9_firstpass.c." into experimental	2013-05-13 21:46:14 -07:00
Dmitry Kovalev	2149852935	Merge "Removing simple loopfilter and code duplication from loopfilter code." into experimental	2013-05-13 18:25:15 -07:00
Dmitry Kovalev	1c43e643b7	Removing simple loopfilter and code duplication from loopfilter code. Change-Id: Ib19352e391408507f2237985501406900a355964	2013-05-13 18:09:11 -07:00
Dmitry Kovalev	a933311131	Preparing vp9_deblock and vp9_denoise to alpha support. Change-Id: Id1cc1c2663b9c2219cb830ffb4b0c6ab3468dc04	2013-05-13 14:03:29 -07:00
Jingning Han	089ed30d07	Merge "Use consistent partition context setup in enc/dec" into experimental	2013-05-13 10:51:51 -07:00
Jingning Han	fc31ae479d	Merge "Move get_sb_index to vp9_blockd.h" into experimental	2013-05-13 10:51:29 -07:00
Paul Wilkins	e5f715201a	Change to band calculation. Change band calculation back to simpler model based on the order in which coefficients are coded in scan order not the absolute coefficient positions. With the scatter scan experiment enabled the results were appear broadly neutral on derf (-0.028) but up a little on std-hd +0.134). Without the scatterscan experiment on the results were up derf as well. Change-Id: Ie9ef03ce42a6b24b849a4bebe950d4a5dffa6791	2013-05-13 17:21:49 +01:00
Jingning Han	6910f178f1	Use consistent partition context setup in enc/dec Move set_partition_seg_context_ to common file. Use consistent context setup conditions for partition probability model update at encoder and decoder. Change-Id: I24b7ed3b1c48e3d2568191a46b70136b99b67b1a	2013-05-11 15:22:13 -07:00
Jingning Han	2117d4ee96	Move get_sb_index to vp9_blockd.h Use common function to fetch/assign sb_index in rd loop, bit-stream writing and reading. Change-Id: I1d8a214a57ed9cbcd026040436ef33e5e39d65b7	2013-05-11 13:26:59 -07:00
John Koleszar	5f221e7ac4	Merge "Fix token allocation for non-4:2:0" into experimental	2013-05-10 16:57:09 -07:00
John Koleszar	998b540fe2	Merge "Fix non-4:2:0 chroma MV calculation for SPLITMV" into experimental	2013-05-10 16:57:03 -07:00
John Koleszar	64667d5af7	Merge "Subsampling aware allocs and bitstream" into experimental	2013-05-10 16:55:00 -07:00
Dmitry Kovalev	4a559d3448	Merge "Removing unused simple loopfilter code." into experimental	2013-05-10 12:14:34 -07:00
Dmitry Kovalev	effaa3263d	Removing unused simple loopfilter code. Change-Id: Ic11dc052fb641687c015e1bbc37181b9babcd43e	2013-05-10 11:04:43 -07:00
Yunqing Wang	9f5811c2da	Add joint motion search in comp_inter_inter mode(experiment) In current code, motion vectors got from single prediction mode are used in compound prediction mode directly. These motion vectors may not give accurate prediction since they are searched independently. In this patch, we took Pascal's suggestion, and did joint motion search in compound prediction mode to find better motion vectors in this situation. Test results: Overall PSNR: 0.570%(derf), 0.918%(stdhd); SSIM: 0.572%(derf), 1.009%(stdhd); The encoder is a little slower. This can be improved since some c code is used in motion search. Change-Id: Ib30c9240f6c56c9b070867b4ca89412a76d9f3c6	2013-05-10 10:15:43 -07:00
John Koleszar	da5054c5af	Fix token allocation for non-4:2:0 Increase the allocated size of the token array to support 4:4:4. Change-Id: I7766a7bedc74b819dcc1f3622d634f340fd3186d	2013-05-09 20:22:59 -07:00
John Koleszar	a37ee9d2e8	Fix non-4:2:0 chroma MV calculation for SPLITMV The previous code was somewhat vestigial for 16x16 MI units, but was incorrect when called with chroma blocks larger than 4x4 because the block index caused a reference to a non-existent BMI. This patch uses the same MV for all chroma subblocks in SPLITMV mode, which is suboptimal for non-4:2:0 subsamplings, but as SPLITMV may be removed in the near future, will use this as a stop gap. Change-Id: I3211cee5ccf1cfb426e5eef5353b0ce5bb92b4cd	2013-05-09 20:14:39 -07:00
John Koleszar	da58436f43	Subsampling aware allocs and bitstream Make framebuffer allocations according to the chroma subsamping factors in use. A bit is placed in the raw part of the frame header for each of the two subsampling factors. This will be moved in a future commit to make them part of the TBD feature set bits, probably only set on keyframes, etc. Change-Id: I59ed38d3a3c0d4af3c7c277617de28d04a001853	2013-05-09 17:50:12 -07:00
John Koleszar	eab6a421ea	Merge "Use common get_uv_tx_size()" into experimental	2013-05-09 12:18:10 -07:00
Dmitry Kovalev	eb93893bee	Updating comments for prediction modes. Change-Id: If4063184f7b37dc011ec6a7a3e75260f4251e984	2013-05-09 11:37:51 -07:00
John Koleszar	1fec23bef6	Use common get_uv_tx_size() Use a single method for calculating the transform size of non-luma planes. Change-Id: I16ebd10e7944d7b9075ab79d15e6a5b5f9bab775	2013-05-08 20:48:32 -07:00
Dmitry Kovalev	dc5418050a	Code cleanup inside vp9_firstpass.c. Change-Id: Ia2814402e3c2ec97c24c536c05f0f526fe1a431c	2013-05-08 18:13:46 -07:00
Dmitry Kovalev	f66320abff	Removing LOOPFILTER_TYPE and corresponding bit in bitstream. We don't have two loopfilter types anymore. Change-Id: I53c0137361342c7d00887ad03be3490f0dfa3532	2013-05-08 16:44:08 -07:00
Dmitry Kovalev	267e9331e2	Merge "Using 4-iteration loop for extra_mb_col inside loopfilter function." into experimental	2013-05-08 16:37:34 -07:00
Dmitry Kovalev	7b602cba65	Merge "Eliminating several YV12_BUFFER_CONFIG usages." into experimental	2013-05-08 16:36:24 -07:00
Jingning Han	944ad130b6	Merge "Extend left/above partition context to per mi(8x8)" into experimental	2013-05-08 16:33:25 -07:00
Dmitry Kovalev	81f33bc091	Eliminating several YV12_BUFFER_CONFIG usages. Change-Id: Ia85b987c935d545920dcae5a6f44136b1a08a008	2013-05-08 14:11:47 -07:00
Dmitry Kovalev	673cc21dfc	Using loop to iterate through YV12_BUFFER_CONFIG planes. Change-Id: I22f1066eb0022c8d75f65a78435ee4ffecdfe0c9	2013-05-08 13:39:16 -07:00
Dmitry Kovalev	a0b6b8a7d4	Merge "Removing unused code + little cleanup." into experimental	2013-05-08 11:23:14 -07:00
Jingning Han	4a88ad89fd	Extend left/above partition context to per mi(8x8) Update and buffer left/above partition information context per 8x8 block. This allows to further enable recursive partition down to 4x4 block size, and hence deprecating I4X4_PRED and SPLITMV. This commit also fixes a context buffer swap/restore issue in 32x32 partition type search. This gives 0.1% performance gain for derf/yt. Will refactor the superblock partition type search into recursion form. Change-Id: Ib61975aca5f12b78d8018481d7fa1393d085689b	2013-05-08 10:20:34 -07:00
John Koleszar	7465f52f81	Merge "Make setup_pred_block subsampling-aware." into experimental	2013-05-07 21:53:31 -07:00
Dmitry Kovalev	8e39295934	Using 4-iteration loop for extra_mb_col inside loopfilter function. Change-Id: I3a4f456035628a9397bdc57c19cdb03439ab1ed3	2013-05-07 17:18:57 -07:00
John Koleszar	d6c490cb15	Merge "Deprecate code_zerogroup experiment." into experimental	2013-05-07 17:09:38 -07:00
Dmitry Kovalev	9cd5406c32	Merge "Removing vp9_swap_yv12_buffer function and corresponding files." into experimental	2013-05-07 17:02:38 -07:00
Dmitry Kovalev	cba0a5db2b	Removing unused code + little cleanup. Change-Id: I81c19a8f19cfb5c7183609656ade833d72feb500	2013-05-07 16:56:22 -07:00
Paul Wilkins	a14ae84749	Deprecate code_zerogroup experiment. Delete code under the CONFIG_CODE_ZEROGROUP flag. Change-Id: I5fe6c7b42a5da9b73118e33594301da4129f320a	2013-05-07 16:52:55 -07:00
Dmitry Kovalev	b05247df95	Removing vp9_swap_yv12_buffer function and corresponding files. Adding static swap_yv12 function to vp9_firstpass.c. Change-Id: I7da9caab9720498db4a74c627901bf37816ed06c	2013-05-07 16:49:22 -07:00
Paul Wilkins	1ed57a6a62	Deprecate comp_interintra_pred experiment. Delete code under the CONFIG_COMP_INTERINTRA_PRED flag. Change-Id: I3d1079cf46305c08f7e11d738596ea112e7b547f	2013-05-07 16:24:08 -07:00
Paul Wilkins	0bfcd30768	Remove enable_6tap filter experiment. Clean out code under CONFIG_ENABLE_6TAP flag. Change-Id: Ic45b624081181027d6ba24d55dd644c3197f9830	2013-05-07 16:13:02 -07:00
Paul Wilkins	8c1b516d10	Deprecate the newbintramode experiment. Clean out code relating to newbintramode. Change-Id: Ie91f4f156cdf60ce0da8ca407c1c9cb00c7d0705	2013-05-07 16:00:59 -07:00
Paul Wilkins	9afb6700c2	Adjust q range Skip Q values between the q.0 mode and a real q of 2.0 as these are not valuable from an RD perspective. Change-Id: I110c4858c57f97315953f4d88a2596d4764360df	2013-05-07 15:34:17 -07:00
Jingning Han	b0cd64f189	Merge "Add building blocks for partition down to 4x4" into experimental	2013-05-07 15:33:20 -07:00
Dmitry Kovalev	847e184011	Merge "General code cleanup inside treewriter-related files." into experimental	2013-05-07 15:04:28 -07:00
Jingning Han	cf8b5a09ed	Add building blocks for partition down to 4x4 Macro ab4x4 contains experiments for recursive partition down to 4x4 block size. Change-Id: Ic727842fa98a4df9fd51e0025a545dc76a5c76c1	2013-05-07 12:11:51 -07:00
John Koleszar	e559e14fa6	Make setup_pred_block subsampling-aware. Code previously set up the pointers by scaling by MI_UV_SIZE, which is 4:2:0 only. Change-Id: Ic13a92895cff018ec1345736746ed84cb31e6e31	2013-05-07 11:47:45 -07:00
Jingning Han	c0504a9b24	Merge "Merge SB8X8 into the codebase" into experimental	2013-05-07 09:23:47 -07:00
Jingning Han	776c1482a3	Merge SB8X8 into the codebase Pull sb8x8 out of experimental list. verified via borg run tests. Fixed unit test failures. Change-Id: I12a4bbd17395930580c048ab68becad1ffe46e76	2013-05-07 09:08:25 -07:00
Scott LaVarnway	cb7955d83e	Removed vp9_setup_intra_recon() This setup is now handled by vp9_build_intra_predictors() when left_available and/or up_available is zero. Change-Id: I59cec0ab95f8be69ce885fd20727510e4deef8a0	2013-05-06 16:13:06 -04:00
Ronald S. Bultje	f7fa367094	Fix first-pass intra4x4 for sb8x8 experiment. Change-Id: I1df17f45721c690d157800daa6a0b377e3d32bc2	2013-05-04 15:49:41 -07:00
John Koleszar	acc9c125dd	Remove old_block_idx_4x4 Removes several instances where the old block numbering was still in use. Change-Id: Id35130591455a4abe6844613e45c0b70c1220c08	2013-05-03 17:19:13 -07:00
John Koleszar	6c622e2783	Merge "Separate transform and quant from vp9_encode_sb" into experimental	2013-05-03 17:19:01 -07:00
John Koleszar	4529c68b3b	Separate transform and quant from vp9_encode_sb This allows removing a large number of transform size specific functions, as well as supporting 444/alpha by routing all code through the subsampling-aware path. Change-Id: Ieb085cebe9f37f24fc24de179898b22abfda08a4	2013-05-03 12:14:50 -07:00
Adrian Grange	7aae782c37	Merge "Extend number of reference buffers to 8." into experimental	2013-05-03 09:59:54 -07:00
Adrian Grange	d7eea782f2	Extend number of reference buffers to 8. The number of reference buffers is extended to 8 and a reference sign-bias added for the LAST_FRAME. Whilst the number of reference buffers used by an individual frame remains unchanged at 3, these may now be selected from 8 possible buffers. Change-Id: I2d247b9c1c2b3a339d6c9fac125e81ba373f75a7	2013-05-03 09:17:18 -07:00
Scott LaVarnway	3041cf8c8b	Merge "Reduced y_dequant, uv_dequant size" into experimental	2013-05-03 07:30:31 -07:00
Dmitry Kovalev	519d9f3e16	Merge "Using treed_read/treed_write functions for segment ids." into experimental	2013-05-02 10:40:58 -07:00
Ronald S. Bultje	704fb4866e	Fix right-edge availability for intra prediction in sb8x8. Fixes valgrind uninitialized value use warnings. Change-Id: Ie9314d684e2ad194f8aca5bde1729fb9b7c0221d	2013-05-02 10:16:48 -07:00
Ronald S. Bultje	ec6cf519d1	Merge "Fix some more offset errors in sb8x8." into experimental	2013-05-02 09:08:33 -07:00
Jingning Han	73a4824c34	Merge "Fix bug in sb8x8 partition context" into experimental	2013-05-02 09:04:28 -07:00
Ronald S. Bultje	3e345cd4d8	Fix some more offset errors in sb8x8. Change-Id: I83677227f7610fdf2db9f15f87fecd4d8e072427	2013-05-02 07:54:18 -07:00
Ronald S. Bultje	dd1e6b8e6f	Merge "Fix block reconstruction with sb8x8 enabled." into experimental	2013-05-02 07:11:36 -07:00
Jingning Han	ba24a28f69	Fix bug in sb8x8 partition context Fix the issue that causes array bound excess in getting partition context. Change-Id: I66166f047f0bcaefebb0bcf441c5b1f777d8da44	2013-05-01 22:34:27 -07:00
Ronald S. Bultje	ff37688a91	Fix block reconstruction with sb8x8 enabled. The encoder reconstruction is now correct. Decoder to follow shortly. Change-Id: Iedf98cdaebb4ca1256c7714cad7024a75853ad6a	2013-05-01 19:28:17 -07:00
Jingning Han	b8decb0313	Fix bugs in sb8x8 experiment/context prob update Fix bugs occur in contextual partition probability update, when sb8x8 is enabled. Change-Id: I19e2cec8a54c2dafd2be2803bbfde7337a2ae45f	2013-05-01 15:16:50 -07:00
Ronald S. Bultje	b6c2d872f0	Fix some crashes in sb8x8 experiment. Change-Id: I390bb1cedc835f439fd5dd6cda6572b29cbb139c	2013-05-01 14:45:27 -07:00
Jingning Han	650e632400	Merge "Enable bit-stream support to SB8X8" into experimental	2013-05-01 13:48:14 -07:00
Scott LaVarnway	94ed11d89d	Reduced y_dequant, uv_dequant size Currently, only two values are used. Removed the unused values. Change-Id: Idc5b8be354d84ffc68df39ea3e45f9f50d977b35	2013-05-01 16:25:10 -04:00
Jingning Han	2bf1dc2e23	Enable bit-stream support to SB8X8 This commit enables bit-stream writing and reading for recursive partition down to block 8x8. Change-Id: I163cd48d191cc94ead49cbb7fc91374f6bf204e2	2013-05-01 13:03:54 -07:00
Dmitry Kovalev	6e4ed2f0fe	Merge "Adding vp9_get_qindex function." into experimental	2013-05-01 12:04:21 -07:00
John Koleszar	d139655b14	Merge "Make vp9_optimize_sb* common" into experimental	2013-04-30 21:43:26 -07:00
John Koleszar	1f80a568d2	Make vp9_optimize_sb* common Unify the various vp9_optimize_sb functions into one that handles all transform sizes. Change-Id: I48b642fbfb3e72cc2e0bcf1d0317a80a80547882	2013-04-30 21:34:58 -07:00
Dmitry Kovalev	79590f186c	Merge "Cleaning up encoder segmentation code." into experimental	2013-04-30 17:49:55 -07:00
Ronald S. Bultje	ff2d69573e	Use more restrictive block radius for 8x8 block MV references. Change-Id: If02e006aa8a89da9de23da92362bd2e7718ea07c	2013-04-30 17:34:02 -07:00
Dmitry Kovalev	aea29cd278	General code cleanup inside treewriter-related files. Change-Id: Ifaa40612a9c054d96112ba350c6f4adb46b1bd5b	2013-04-30 16:39:07 -07:00
Ronald S. Bultje	d068d869b9	sb8x8 integration in rd loop. Work-in-progress, not yet ready for review. TODO items: - bitstream writing (encoder) and reading (decoder) - decoder reconstruction Change-Id: I5afb7284e7e0480847b47cd0097cb469433c9081	2013-04-30 16:13:20 -07:00
Dmitry Kovalev	b5364d4f3b	Using treed_read/treed_write functions for segment ids. Changing the order of probabilities inside mb_segment_tree_probs in order to use treed_read/treed_write function instead of custom code. Change-Id: I843487d5057913b9358db73da270893eefecc6c8	2013-04-30 14:06:49 -07:00
Dmitry Kovalev	3f6c6ffc86	Adding vp9_get_qindex function. Moving common code from encoder and decoder to vp9_get_qindex function. Also moving quant-related constants from vp9_onyxc_int.h to vp9_quant_common.h. Change-Id: I70c5bfbaa1c8bf00fde0bfc459d077f88b6d46c8	2013-04-30 13:39:50 -07:00
Yaowu Xu	ad6890316c	Merge "Removed code no longer being used." into experimental	2013-04-30 12:26:09 -07:00
Yaowu Xu	df6f82c3e8	Removed code no longer being used. Change-Id: Iab9a88f250614a790b6ad96bf3150a74210910df	2013-04-30 12:09:27 -07:00
Dmitry Kovalev	15b5e465f2	Adding vp9_update_frame_size function. Moving common code from encoder and decoder to vp9_update_frame_size. Change-Id: I6ca758b7d05ffd52821bd3f7ad68089da11e4165	2013-04-30 11:14:27 -07:00
Dmitry Kovalev	70d12c3a75	Merge "Renaming refresh_entropy_probs to refresh_frame_context." into experimental	2013-04-30 10:21:24 -07:00
Dmitry Kovalev	51a73fbba2	Merge "Consistent names for quant-related functions and variables." into experimental	2013-04-30 10:19:48 -07:00
Jingning Han	7492edac93	Merge "Separate I4X4_PRED coding from macroblock modules" into experimental	2013-04-29 21:51:59 -07:00
Jingning Han	94191b5c82	Separate I4X4_PRED coding from macroblock modules Separate the functionality of I4X4_PRED from decode_mb. Use decode_atom_intra instead, to enable recursive partition of superblock down to 8x8. Change-Id: Ifc89a3be82225398954169d0a839abdbbfd8ca3b	2013-04-29 18:59:36 -07:00
Yaowu Xu	0b48548eeb	Merge "fixed new intra code for rectanglar blocks" into experimental	2013-04-29 16:16:27 -07:00
Dmitry Kovalev	ee97da2c03	Cleaning up encoder segmentation code. Moving code from vp9_pack_bitstream to new function encode_segmentation. Change-Id: I1f1e59a1f038618ad95162b7db4b6f8164850ea8	2013-04-29 16:07:17 -07:00
Yaowu Xu	caea860a0f	Merge "Enabled i4x4 to use right above pixels" into experimental	2013-04-29 16:05:19 -07:00
Yaowu Xu	7ea12f2c5f	Merge "Use same intra prediction for all block size" into experimental	2013-04-29 16:02:35 -07:00
Yaowu Xu	4747c6ed90	fixed new intra code for rectanglar blocks Also fixed two minor subtle boundary conditions in intra prediction code, and replaced memcpy/memset with vpx_ prefixed version. Change-Id: I9cddff3be831228b628f1f2f065a61feacbcbee6	2013-04-29 16:02:00 -07:00
Yaowu Xu	e388251d5d	Enabled i4x4 to use right above pixels Change-Id: I7442b4600b6812bed13e655ccf68f9ea56cc83a2	2013-04-29 15:16:59 -07:00
Yaowu Xu	3d655805f2	Use same intra prediction for all block size The commmit changed to use same intra prediction function for all block sizes. Some details on the changes: 1. All directional modes except DC/TM/V/H now have built-in filtering for all pixels with filter taps either (1, 2, 1)/4 or (1, 1)/2. 2. Above edge get automatic extended to double width (bw*2), which makes a lot of the prediciton mode computation simpler. 3. Same intra prediction function is called with different size for i4x4_pred and all other larger size. Overall, the change helped keyframe only coding for both cif size and std-hd size test sets by .5% consistently on all encodings. For normal coding with single/auto key frame, the change now also is consistently net positive for all encodings. The overall gains is about .15% on std-hd set. Change-Id: I01ceb31fbc73d49776262e6bdc06853b03bbd1d1	2013-04-29 15:15:30 -07:00
Ronald S. Bultje	f5ad774814	Merge "Change above/left_context to use an 8x8 basis." into experimental	2013-04-29 11:28:25 -07:00
Ronald S. Bultje	2dbaa4f4f4	Change above/left_context to use an 8x8 basis. Output changes slightly because of a minor bug in (at least) the sb32x16 block2above tx16x16 tables that previously existed in vp9_blockd.c. Change-Id: I624af28ac200a8322d64454cf05c79e9502968cc	2013-04-29 10:37:25 -07:00
Deb Mukherjee	040eeed9d0	Turning model based reverse update on for coefs Turns model based reverse updates on for coefficients in an effort to reduce the memory requirement for counters. With this patch the counters needed will be reduced by about 75% since only 3 counts are needed instead of 12. The impact in performance is: derf300: -0.252% stdhd250: -0.046% However retraining should alleviate some of the drop in performance. Change-Id: I6f2b3e13f6d5520aa3400b0b228fb5e8b4a43caa	2013-04-29 10:09:57 -07:00
Paul Wilkins	fb3e4ed9eb	Merge "Minor tweak to implicit segmentation experiment." into experimental	2013-04-27 11:58:13 -07:00
Ronald S. Bultje	7f8cbda333	Merge "Grow MODE_INFO array to use an 8x8 basis." into experimental	2013-04-26 14:46:50 -07:00
Dmitry Kovalev	9713a68719	Renaming refresh_entropy_probs to refresh_frame_context. Change-Id: I5429c02246d198eb1b6aadbc3313b26bf3436062	2013-04-26 14:39:58 -07:00
Johann	9e23bd5df5	Merge "Merge branch 'master' into experimental" into experimental	2013-04-26 13:35:28 -07:00
Johann	32a5c52856	Merge branch 'master' into experimental Conflicts: vp9/common/vp9_findnearmv.c vp9/common/vp9_rtcd_defs.sh vp9/decoder/vp9_decodframe.c vp9/decoder/x86/vp9_dequantize_sse2.c vp9/encoder/vp9_rdopt.c vp9/vp9_common.mk Resolve file name changes in favor of master. Resolve rdopt changes in favor of experimental, preserving the newer experiments. Change-Id: If51ed8f457470281c7b20a5c1a2f4ce2cf76c20f	2013-04-26 12:57:10 -07:00
Dmitry Kovalev	5a5a1f25a8	Consistent names for quant-related functions and variables. Change-Id: I3a6d601e90e8740b9c26dd0afbfe9d467b75d367	2013-04-26 12:30:20 -07:00
Ronald S. Bultje	1a46b30ebe	Grow MODE_INFO array to use an 8x8 basis. Change-Id: I087e08e7909a406b71715b8525c104208daa6889	2013-04-26 11:57:17 -07:00
John Koleszar	bb41ab4a0c	Remove BLOCKD structure All members can be referenced from their per-plane counterparts, and removes assumptions about 24 blocks per macroblock. Change-Id: I7ff2fa72d22c29163eb558981c8193765a8113d9	2013-04-26 10:35:54 -07:00
John Koleszar	4f55c5618a	Remove destination pointers from BLOCKD Access these members from MACROBLOCKD instead. Change-Id: I7907230dd473ff12ebe182b9280d8b7f12a888c4	2013-04-26 10:14:07 -07:00
Scott LaVarnway	57f180b388	Removed bmi from blockd This originally was "Removed update_blockd_bmi()". Now, this patch removed bmi from blockd and uses the bmi found in mode_info_context. Eliminates unnecessary bmi copies between blockd and mode_info_context. Change-Id: I287a4972974bb363f49e528daa9b2a2293f4bc76	2013-04-26 10:19:43 -04:00
Ronald S. Bultje	8d028402d7	Remove implicit assumption that mode_info_stride == mb_cols + 1. Change-Id: I3030d7adac73109aeaa1ecc0f78ac968c092d9aa	2013-04-25 14:21:01 -07:00
Ronald S. Bultje	c849eaca59	Use b_width/height_log2 instead of mb_ where appropriate. Basic assumption: when talking about transform units, use b_; when talking about macroblock indices, use mb_. Change-Id: Ifd163f595d4924ff892de4eb0401ccd56dc81884	2013-04-25 14:20:59 -07:00
John Koleszar	a99e1aa8ca	Remove predictor pointers from BLOCKD Access these members from MACROBLOCKD instead. Change-Id: I2574622e577bb9feede47f6b7ccbb11f3e928ca8	2013-04-25 12:04:07 -07:00
John Koleszar	6c0c6b86c1	Remove diff from BLOCKD The underlying storage for these buffers is in the per-plane MACROBLOCKD area, so read it from there directly. Change-Id: Id6bd835117fdd9dea07db95ad06eff9f12afaaf7	2013-04-25 11:57:22 -07:00
John Koleszar	15255eef82	Move dequant from BLOCKD to per-plane MACROBLOCKD This data can vary per-plane, but not per-block. Change-Id: I1971b0b2c2e697d2118e38b54ef446e52f63c65a	2013-04-25 11:57:20 -07:00
John Koleszar	4bd0f4f646	Remove BLOCK structure All members can be referenced from their per-plane counterparts, and removes assumptions about 24 blocks per macroblock. Change-Id: I593fb0715e74cd84b48facd1c9b18c3ae1185d4b	2013-04-25 11:33:17 -07:00
Johann	c5b127afea	Rename vp9_idct_x86.c Remove similarly named header file. It is obsolete. Move file to match naming style. Adjust make file to include the file correctly and remove extra unnecessary #if guard. Change-Id: Ifba07ba9938a5df08a9f4eda54a3ac4d6983f7bf	2013-04-25 11:13:02 -07:00
Dmitry Kovalev	61a47da869	Adding is_inter_mode function. Change-Id: I2d32d46002cb92c63050c2b8328865c406103621	2013-04-25 10:23:00 -07:00
Dmitry Kovalev	2cf0675a52	Merge "Removing unused mi_mv_pred_row and mi_mv_pred_col functions." into experimental	2013-04-25 10:18:07 -07:00
Dmitry Kovalev	9b081e6807	Merge "Using ROUND_POWER_OF_TWO macro inside vp9_loopfilter_filters.c." into experimental	2013-04-25 10:17:54 -07:00
Dmitry Kovalev	22c6ce03fa	Merge "Handling frame references and scale factors in one for loop." into experimental	2013-04-25 10:17:34 -07:00
Jingning Han	b42b41c856	Merge "Move sbsegment out of experimental list" into experimental	2013-04-25 09:18:01 -07:00
Scott LaVarnway	a426c7f343	Merge "Moved dequantization into the token decoder" into experimental	2013-04-25 08:53:42 -07:00
Dmitry Kovalev	994c79cccf	Handling frame references and scale factors in one for loop. Using ALLOWED_REFS_PER_FRAME constants instead of hard coded 3, replacing memcpy with plain struct assignment. Change-Id: Ibc86f5d175fcb3f3a3eddacf593525370f1f854c	2013-04-24 17:20:53 -07:00
Yaowu Xu	bcf82cf503	Merge two similar functions into one Function set_mb_row() and set_mb_col() do similar work and are always called together, this commit merged them into a single function for clarity and easy maintainence. This was a TODO item. Change-Id: I956bd9ed6afb8b2b0469b20fd8bc893b26f8a0f3	2013-04-24 15:58:03 -07:00
Dmitry Kovalev	a2d46434b5	Merge "Fixing PRED_SWITCHABLE_INTERP case in vp9_get_pred_context function." into experimental	2013-04-24 15:44:39 -07:00
Jingning Han	b0e3b3df18	Move sbsegment out of experimental list Move rectangular superblock coding out of experimental list. Change-Id: I96c37547d122330d666a67b4bf577ae54547857f	2013-04-24 15:19:17 -07:00
Jingning Han	ff2b8aa2c9	Contextual entropy coding of partition syntax This commit enables selecting probability models for recursive block partition information syntax, depending on its above/left partition information, as well as the current block size. These conditional probability models are reasonably stationary and consistent across frames, hence the backward adaptive approach is used to maintain and update the contextual models. It achieves coding performance gains (on top of enabling rectangular block sizes): derf: 0.242% yt: 0.391% hd: 0.376% stdhd: 0.645% Change-Id: Ie513d9673337f0d27abd65fb566b711d0844ec2e	2013-04-24 14:23:14 -07:00
Dmitry Kovalev	2e3f3e4fbb	Using ROUND_POWER_OF_TWO macro inside vp9_loopfilter_filters.c. Change-Id: Icb671cd011f645a3361684207840d14330ca7488	2013-04-24 11:50:49 -07:00
Ronald S. Bultje	41a8a95bd1	Merge "Change chroma loopfilter to skip inner SB edges for tx16x16 also." into experimental	2013-04-24 11:45:26 -07:00
Ronald S. Bultje	4149ba1cbf	Merge "Minor indent changes in loopfilter code." into experimental	2013-04-24 11:45:19 -07:00
Ronald S. Bultje	cc7ce53140	Merge "Add basic building blocks for 8x8 superblocks experiment." into experimental	2013-04-24 11:45:13 -07:00
Dmitry Kovalev	bd994ed42d	Fixing PRED_SWITCHABLE_INTERP case in vp9_get_pred_context function. Adding xd->up_available as additional check for above context. Change-Id: If5654e4cae184b9c369b7b2e08076cb2951d00ed	2013-04-24 10:45:32 -07:00
Ronald S. Bultje	5b57580cd9	Change chroma loopfilter to skip inner SB edges for tx16x16 also. Change-Id: I6ea9e110b5c5b07ab7d092886dbd51a6eccc0217	2013-04-24 09:45:43 -07:00
Paul Wilkins	6579720e6a	Merge "Extension of segmentation to 8 segments." into experimental	2013-04-24 09:44:48 -07:00
Paul Wilkins	e307e924b5	Merge "Simplify Segment Coding" into experimental	2013-04-24 09:43:01 -07:00
Deb Mukherjee	0e905e03f2	Merge "Fix in token allocation with zerogroup expt" into experimental	2013-04-24 08:57:25 -07:00
Paul Wilkins	da04312f79	Minor tweak to implicit segmentation experiment. This minor tweak makes segment 0 neutral and used by key frames and also extends beyond 4 segments. Change-Id: Ife4744602aba66ac9432746db3113cc5cd88a482	2013-04-24 16:43:01 +01:00
Paul Wilkins	31ee193a9c	Extension of segmentation to 8 segments. Also some further simplification following removal of top node code. There is an issue in regards to the shared file vp8cx.h in regard to the roi_map as this interface assumes that there are only 4 segments. I have left the value here as 4 for now meaning that the roi_map interface is broken for VP9. Note that this change would have been easier if I hadn't had to search for hard wire instances of the number 4 and <= 3. Change-Id: Ia8b6deea4be4dbd20deb1656e689dd43a5f190e8	2013-04-24 16:36:47 +01:00
Paul Wilkins	c77aff1286	Simplify Segment Coding Remove top node optimization. The improvement this gives is not sufficient to justify the extra complexity. Change-Id: I2bb4a12a50ffd52cacfa4a3e8acbb2e522066905	2013-04-24 10:47:12 +01:00
Paul Wilkins	27bb4777cd	Simple implicit segmentation experiment. Change-Id: Iaef16122732c2a81e0927f9862b51b68dc788712	2013-04-24 10:04:27 +01:00
Dmitry Kovalev	156d912025	Merge "Code cleanup inside vp9_get_pred_context function." into experimental	2013-04-23 18:03:29 -07:00
Dmitry Kovalev	97ac785e65	Merge "Simple cleanup inside vp9_decodframe.c and vp9_entropymode.c." into experimental	2013-04-23 18:02:46 -07:00
Dmitry Kovalev	afeff1acd1	Merge "Removing redundant code in vp9_entropymode.c." into experimental	2013-04-23 18:01:37 -07:00
Jingning Han	adbbd26517	Merge "Enable rectangular support for comp inter-intra" into experimental	2013-04-23 17:16:16 -07:00
Dmitry Kovalev	9828a33ebb	Removing unused mi_mv_pred_row and mi_mv_pred_col functions. Change-Id: If8ba37bf0b86e8dea88c27d911e8ddb0f6d5a3c5	2013-04-23 16:34:22 -07:00
John Koleszar	4f35e3e1c1	Merge "Move src_diff to per-plane MACROBLOCK data" into experimental	2013-04-23 16:24:08 -07:00
Dmitry Kovalev	de7c25c9f0	Code cleanup inside vp9_get_pred_context function. Change-Id: Id06b7a299a26ed944a401faae51907537f722a7e	2013-04-23 16:18:09 -07:00
Dmitry Kovalev	d811558f63	Removing redundant code in vp9_entropymode.c. Change-Id: Ia7266b8d3aa3d5cff2db0c3b2f014def045759af	2013-04-23 15:56:27 -07:00
Dmitry Kovalev	144f49c6aa	Simple cleanup inside vp9_decodframe.c and vp9_entropymode.c. Change-Id: I62dde981f5201c5fbc22001609ee4b5fd0a9bdf5	2013-04-23 15:50:56 -07:00
Jingning Han	a26c1edbb4	Enable rectangular support for comp inter-intra This commit enables rectangular block prediction of compound inter-intra mode. It combines the mb/sb32/sb64 prediction functions into a unified version with configurable block width and height. This fixes the enc/dec mismatch of the codebase when comp-interintra-pred is enabled. Change-Id: I1d0db2f1f184007802df04fcd12b9dadb3189ff0	2013-04-23 15:39:19 -07:00
Ronald S. Bultje	276a1106e6	Merge changes I54acef34,I72d42971 into experimental * changes: Make some sb_type comparisons independent of literal enum values. Make loopfilter aware of rectangular blocks.	2013-04-23 15:29:19 -07:00
Dmitry Kovalev	d0d1094a05	Merge "Adding get_scan_{4x4, 8x8, 16x16} functions." into experimental	2013-04-23 12:44:51 -07:00
Johann	7af58d4338	Resolve declaration and implementation. Clean Windows build warnings: warning C4028: formal parameter <N> different from declaration This was fixed independently in master and experimental but the fixes were in opposite directions. One added const to the declaration and the other removed it from the implementation. Also update the variable names. This doesn't modify the data so call it ref, matching the functions in the vicinity, rather than dst. Change-Id: I2ffc6b4a874cb98c26487b909d20a5e099b5582c	2013-04-23 12:42:31 -07:00
Ronald S. Bultje	e0fc9201fe	Minor indent changes in loopfilter code. Change-Id: I0cdc951558e4d7748f63df8c03b1c9dce086acb0	2013-04-23 12:37:05 -07:00
Ronald S. Bultje	94297863bf	Add basic building blocks for 8x8 superblocks experiment. Change-Id: I274a1d2e461e6ffdb106bac4ad6951692ace314e	2013-04-23 12:34:37 -07:00
Ronald S. Bultje	5ba98ebcf1	Make some sb_type comparisons independent of literal enum values. Change-Id: I54acef342b8e787e05af0febd7cf0d7d10288383	2013-04-23 12:34:32 -07:00
Ronald S. Bultje	0db636619f	Make loopfilter aware of rectangular blocks. Also use explicitely named enum values in sb_type comparisons, rather than relying on absolute integer values, because enum values may change in the future. Change-Id: I72d42971a98157af93413a25ac2c7e6f9b369cec	2013-04-23 12:34:32 -07:00
John Koleszar	cbd1315ac4	Move src_diff to per-plane MACROBLOCK data First in a series of commits making certain MACROBLOCK members addressable per-plane. This commit also refactors the block subtraction functions vp9_subtract_b, vp9_subtract_sby_c, etc to be loops-over-planes and variable subsampling aware. Change-Id: I371d092b914ae0a495dfd852ea1a3d2467be6ec3	2013-04-23 12:18:51 -07:00
Ronald S. Bultje	c4cae4cd5d	Remove unused corner_predictor and log2_minus_1 functions. Change-Id: Ic659544ca12b1bc191b93ddfa378964bd133bfc9	2013-04-23 11:19:39 -07:00
Scott LaVarnway	e3167b8c23	Merge "Eliminated prev_mip memsets/memcpys" into experimental	2013-04-23 10:22:13 -07:00
Deb Mukherjee	64e7f017ce	Fix in token allocation with zerogroup expt Fix to increase allocation of tokens at very high rates. Change-Id: Ia27aa0316b0fab664230800f9c9947b5c68ecd58	2013-04-23 08:51:31 -07:00
Deb Mukherjee	735febf1ce	Removing the implicit compound inter experiment Removing this experiment for now, since it has been broken with the latest code changes. Change-Id: I1be2181b56de490fcb577f5905b5e147a8ed82d8	2013-04-22 16:46:54 -07:00
Scott LaVarnway	e732bc298c	Moved dequantization into the token decoder Mostly for cleanup purposes. Now we should be able to rework the encoder/decoder to use a common idct/add function. Change-Id: I1597cc59812f362ecec0a3493b6101a6cc6fa7ff	2013-04-22 17:53:07 -04:00
Dmitry Kovalev	5de7e16ca2	Adding get_scan_{4x4, 8x8, 16x16} functions. Change-Id: Id4306ef6d65d4a3984aed50b775bdf48d4f6c438	2013-04-22 14:08:41 -07:00
John Koleszar	01e41a531b	Remove vp9_recon_intra_mbuv Use common vp9_recon_sbuv instead. Change-Id: I146f79adfdfda2b52257a52fa783727f12afa246	2013-04-22 12:05:24 -07:00
John Koleszar	c2c15e8eb3	Rewrite vp9_recon_sb* Rewrite vp9_recon_sb{,y,uv} to be a loop over planes. Change-Id: Ica2bbbb3105a1d29b2ff2ead07b76cde9683154c	2013-04-22 12:05:24 -07:00
John Koleszar	a443447b8b	Move pre, second_pre to per-plane MACROBLOCKD data Continue moving framebuffers to per-plane data. Change-Id: I237e5a998b364c4ec20316e7249206c0bff8631a	2013-04-22 12:05:24 -07:00
Deb Mukherjee	f12509f640	Merge "Removes the code_nonzerocount experiment" into experimental	2013-04-22 11:53:14 -07:00
Deb Mukherjee	0aa79be7d5	Removes the code_nonzerocount experiment This patch does not seem to give any benefits. Change-Id: I9d2b4091d6af3dfc0875f24db86c01e2de57f8db	2013-04-22 10:58:49 -07:00
Deb Mukherjee	6ce718eb18	Merge "End of orientation zero group experiment" into experimental	2013-04-22 10:33:12 -07:00
Deb Mukherjee	70d9f116fd	End of orientation zero group experiment Adds an experiment that codes an end-of-orientation symbol for every eligible zero encountered in scan order. This cleans out various other sub-experiments that were part of the origiinal patch, which will be later included if found useful. Results are slightly positive on all sets (0.1 - 0.2% range). Change-Id: I57765c605fefc7fb9d1b57f1b356843602abefaf	2013-04-22 09:27:59 -07:00
John Koleszar	6d5ac8f2e1	reconinter: remove unnecessary functions, params Removes the redundant dst pointers from vp9_build_inter_predictors_sb{y,uv} and the remaining mb specific functions. Change-Id: I7b6bf439d9394b85ea79b4fe61a3ffc1025720da	2013-04-22 08:20:54 -07:00
Paul Wilkins	f82c61b886	Merge "make DC_PRED for i4x4 to use real pixels only" into experimental	2013-04-22 05:10:36 -07:00
Dmitry Kovalev	c7a38f77ef	Merge "Removing get_segment_id function and using existing vp9_get_pred_mb_segid." into experimental	2013-04-20 11:05:50 -07:00
Dmitry Kovalev	5c632dbb19	Merge "Renaming vp9_extra_bit_struct to vp9_extra_bit." into experimental	2013-04-20 11:03:25 -07:00
John Koleszar	fa8ddbd2a6	Merge "Move dst to per-plane MACROBLOCKD data" into experimental	2013-04-19 16:33:45 -07:00
John Koleszar	588c3cb02e	Merge "Remove vp9_recon_mb{,y}" into experimental	2013-04-19 16:33:09 -07:00
Yaowu Xu	e3465a63d7	make DC_PRED for i4x4 to use real pixels only Wherever there are real pixels available before falling back to use assumed values 127 and 129. This also make DC_PRED for i4x4 consistent with DC_PRED for larger blocks. Change-Id: I54372924826118da023f402c802ac6ce0caa70c3	2013-04-19 16:22:07 -07:00
John Koleszar	95c6c13ce6	Merge "Remove redundant pointers from void vp9_recon_sb{y,uv}" into experimental	2013-04-19 16:17:42 -07:00
John Koleszar	d12376aa2c	Move dst to per-plane MACROBLOCKD data First in a series of commits moving the framebuffers pointers to per-plane data, so that they can be indexed numerically rather than by name. Change-Id: I6e0d60fd4d51e6375c384eb7321776564df21775	2013-04-19 16:16:10 -07:00
Scott LaVarnway	9662531d77	Eliminated prev_mip memsets/memcpys For 1080 material, this buffer is currently 2,270,928 bytes. This patch swaps ptrs instead of copying and uses the last show_frame flag instead of setting the entire buffer to zero. For the test clip used, the decoder improved by up to 1%. Change-Id: I686825712ad56043e09ada9808dc489f875a6ce0	2013-04-19 18:38:10 -04:00
Dmitry Kovalev	c09f652590	Removing get_segment_id function and using existing vp9_get_pred_mb_segid. Change-Id: Iff35d4b2f8f65511f80c594958c01fb4673fa033	2013-04-19 14:25:32 -07:00
Paul Wilkins	fb754fd37e	Merge "Mv ref candidates cut to 2." into experimental	2013-04-19 14:09:44 -07:00
John Koleszar	9ec0f658a1	Remove vp9_recon_mb{,y} Use the common sb functions instead. Change-Id: I4fa0a8ee3c6ada56271dd09bf895b97642f55858	2013-04-19 12:12:00 -07:00
John Koleszar	d747986d29	Remove redundant pointers from void vp9_recon_sb{y,uv} Remove the unnecessary _s_ from their names, and add a new vp9_recon_sb() that calls the y and uv variants. Change-Id: I7ffaa5ff5605a8472cac2a53de8cf889353039a6	2013-04-19 12:06:07 -07:00

... 8 9 10 11 12 ...

1481 Commits