generic-library/vpx

Author	SHA1	Message	Date
Jingning Han	10da059b52	Remove unused is_background function Change-Id: Ia540eac5f066ae95280c2f898370eddf0110c279	2014-11-05 21:19:23 -08:00
Jingning Han	caaf63b2c4	Rework cut-off decisions in cyclic refresh aq mode This commit removes the cyclic aq mode dependency on in_static_area and reworks the corresponding cut-off thresholds. It improves the compression performance of speed -5 by 1.47% in PSNR and 2.07% in SSIM, and the compression performance of speed -6 by 3.10% in PSNR and 5.25% in SSIM. Speed wise, about 1% faster in both settings at high bit-rates. Change-Id: I1ffc775afdc047964448d9dff5751491ba4ff4a9	2014-11-05 21:17:09 -08:00
hkuang	e8860693ea	Merge "Totally remove prev_mi in VP9 decoder."	2014-11-05 17:48:47 -08:00
hkuang	4cc7c5a17f	Totally remove prev_mi in VP9 decoder. This will save the memory and improve the decode speed due to removing unnecessary memset of big prev_mi array for all the key frames. Decoding a all key frames 1080p video shows speed improve around 2%. Change-Id: I6284a445c1291056e3c15135c3c20d502f791c10	2014-11-05 16:14:30 -08:00
Yaowu Xu	2c4fee17bc	Fix visual studio 2013 compiler warnings For configured with --enable-vp9-highbitdepth Change-Id: I2b181519d7192f8d7a241ad5760c3578255f24e6	2014-11-05 13:47:28 -08:00
Hui Su	2c95a3f374	Merge "Simplify interface of write_selected_tx_size and read_tx_size"	2014-11-05 13:33:09 -08:00
Jingning Han	a7889cac9a	Merge "Skip ref frame mode search conditioned on predicted mv residuals"	2014-11-05 12:04:10 -08:00
Hui Su	709c634b84	Simplify interface of write_selected_tx_size and read_tx_size Change-Id: Ia2b2a895deefaaf7b34bf26df86add56dbab082c	2014-11-04 16:11:50 -08:00
Minghai Shang	9f9e30d7bf	Merge "[spatial svc] Make spatial svc working for one pass rate control"	2014-11-04 15:57:16 -08:00
Minghai Shang	86c36a504d	[spatial svc] Make spatial svc working for one pass rate control Change-Id: Ibd9114485c3d747f9d148f64f706bf873ea473ac	2014-11-04 11:46:48 -08:00
Jingning Han	1e753387c8	Merge "Refactor sub-pixel motion search unit"	2014-11-04 09:11:15 -08:00
Jingning Han	1434f7695b	Skip ref frame mode search conditioned on predicted mv residuals This commit makes the RTC coding mode to conditionally skip the reference frame mode search, when the predicted motion vector of the current reference frame gives more than two times sum of absolute difference compared to that of other reference frames. It reduces the runtim by 1% - 4% for speed -5 and -6. The average compression performance is improved by about 0.1% in both settings. It is of particular benefit to light change scenarios. The compression performance of test clip mmmovingvga.y4m is improved by 6.39% and 15.69% at high bit rates for speed -5 and -6, respectively. Speed -5 vidyo1 16555 b/f, 40.818 dB, 12422 ms -> 16552 b/f, 40.804 dB, 12100 ms nik 33211 b/f, 39.138 dB, 11341 ms -> 33228 b/f, 39.139 dB, 11023 ms mmmoving 33263 b/f, 40.935 dB, 13508 ms -> 33256 b/f, 41.068 dB, 12861 ms Speed -6 vidyo1 16541 b/f, 40.227 dB, 8437 ms -> 16540 b/f, 40.220 dB, 8216 ms nik 33272 b/f, 38.399 dB, 7610 ms -> 33267 b/f, 38.414 dB, 7490 ms mmmoving 33255 b/f, 40.555 dB, 7523 ms -> 33257 b/f, 40.975 dB, 7493 ms Change-Id: Id2aef76ef74a3cba5e9a82a83b792144948c6a91	2014-11-04 09:10:19 -08:00
Marco	343acaa8f2	Merge "Allow disable of refresh golden for more than 1 layer encoding."	2014-11-03 14:38:05 -08:00
Jingning Han	e083f6bd08	Refactor sub-pixel motion search unit This commit unfolds the legacy macro definitions used in the sub-pixel motion search and refactors the operational flow for later optimizations. Change-Id: I3e3f770cad961d03d1a6eb0b2a0186cc77eaf2b8	2014-11-03 09:02:57 -08:00
Jingning Han	0ca5908ff6	Merge "Fix the THR_MODES array used in vp9_pick_inter_mode"	2014-11-03 08:46:42 -08:00
Yaowu Xu	2fe893c94f	Merge "Fix speed 7 and speed 12 for rt"	2014-11-03 08:02:58 -08:00
Marco	d6b688375f	Allow disable of refresh golden for more than 1 layer encoding. The current logic was allowing for disabling golden refresh only for two pass svc encoding. This change disables it as long as more than 1 layer encoding is used (for example temporal layers under 1pass CBR). Change-Id: I4dc5204a7ad365c821ec7963e93b59da82e1826b	2014-11-02 22:24:00 -08:00
Jingning Han	7e119e2946	Fix the THR_MODES array used in vp9_pick_inter_mode Fix the alignment of entries fo intra prediction modes. Change-Id: Ie32ad87cf90694efd591a4b1cc29c916c4cd56f7	2014-11-02 12:25:57 -08:00
Yaowu Xu	0271ff7775	Fix speed 7 and speed 12 for rt A recent change has introduced big quality drops for speed 7 and 12 for --rt mode. The change reverted the big drop and improved quality by 9.5% for speed 7 and 13.4% for speed 12. Change-Id: I07b82e3bb6002a73af486a083458c88877bdad01	2014-10-31 17:29:02 -07:00
hkuang	55577431ae	Bind motion vectors with frame buffer structure. This will save a lot of memory for decoder due to removing of prev_mi, but prev_mi is still needed in encoder. So this will increase a little bit memory for encoder. Change-Id: I24b2f1a423ebffa55a9bd2fcee1077dac995b2ed	2014-10-31 17:01:08 -07:00
Jingning Han	1c84e73ebd	Merge "Fix mode index use case in vp9_pick_inter_mode"	2014-10-31 08:55:40 -07:00
Jingning Han	61966b1d10	Merge "Refactor vp9_update_rd_thresh_fact"	2014-10-31 08:55:28 -07:00
Jingning Han	1cffea9fb7	Merge "Rework pred pixel buffer system in non-RD coding mode"	2014-10-31 08:55:24 -07:00
Jingning Han	64348d9f8d	Fix mode index use case in vp9_pick_inter_mode This improves coding performance of speed -5 and -6 by 0.6%, respectively. Change-Id: Ic5a7746a88c73285f0b14333d35dc16b02152c25	2014-10-30 11:10:06 -07:00
Jingning Han	f7b46d8c5e	Refactor vp9_update_rd_thresh_fact Reduce the scope of function parameters. Change-Id: Ifef2cfb559908a97498ffdbd6ea53da1cd45a73c	2014-10-30 11:09:40 -07:00
Jingning Han	7bea8c59f9	Rework pred pixel buffer system in non-RD coding mode This commit makes the inter prediction buffer system to support hybrid partition search. It reduces the runtime of speed -5 by about 3%. No compression performance change. vidyo1 720p 1000 kbps 11831 ms -> 11497 ms nik 720p 1000 kbps 10919 ms -> 10645 ms Change-Id: I5b2da747c6395c253cd074d3907f5402e1840c36	2014-10-30 11:08:35 -07:00
Hui Su	d478d2df37	Merge "Move the definition of switchable filter numbers into enum INTERP_FILTER; Modify the macro ADD_MV_REF_LIST and IF_DIFF_REF_FRAME_ADD_MV."	2014-10-30 11:05:04 -07:00
Hui Su	66906da066	Merge "Combine vp9_encode_block_intra and encode_block_intra"	2014-10-30 11:02:31 -07:00
Yunqing Wang	aed48c786a	Remove unused speed feature Partition_check was unused and removed. Change-Id: I15ec9162d86dc61f04c09229c498629878ed7155	2014-10-29 17:05:04 -07:00
Jingning Han	afa31ab9b8	Merge "Enable mode search threshold update in non-RD coding mode"	2014-10-29 12:42:22 -07:00
Jingning Han	9349a28e80	Enable mode search threshold update in non-RD coding mode Adaptively adjust the mode thresholds after each mode search round to skip checking less likely selected modes. Local tests indicate 5% - 10% speed-up in speed -5 and -6. Average coding performance loss is -1.055%. speed -5 vidyo1 720p 1000 kbps 16533 b/f, 40.851 dB, 12607 ms -> 16556 b/f, 40.796 dB, 11831 ms nik 720p 1000 kbps 33229 b/f, 39.127 dB, 11468 ms -> 33235 b/f, 39.131 dB, 10919 ms speed -6 vidyo1 720p 1000 kbps 16549 b/f, 40.268 dB, 10138 ms -> 16538 b/f, 40.212 dB, 8456 ms nik 720p 1000 kbps 33271 b/f, 38.433 dB, 7886 ms -> 33279 b/f, 38.416 dB, 7843 ms Change-Id: I2c2963f1ce4ed9c1cf233b5b2c880b682e1c1e8b	2014-10-29 10:55:34 -07:00
Adrian Grange	4074099ed8	Simplify vp9_set_rd_speed_thresholds_sub8x8 Change-Id: I4bf0f9a38697f5aea564a47afd7f02bb8b2888b6	2014-10-29 09:09:46 -07:00
Hui Su	0928da3b6e	Combine vp9_encode_block_intra and encode_block_intra Change-Id: I79091fb677b64892ecca2fb466fde14602d8cdfc	2014-10-28 18:57:01 -07:00
Jingning Han	982dab6050	Merge "Use zero motion vector in choose_partitioning"	2014-10-28 12:00:13 -07:00
JackyChen	50e5c30536	Merge "vp9_denoiser_sse2: refactor the code."	2014-10-28 11:06:05 -07:00
Yaowu Xu	7d7b43b9af	Merge "Allow update of golden refernce buffer in CBR mode"	2014-10-28 10:48:02 -07:00
JackyChen	99a8dac4de	vp9_denoiser_sse2: refactor the code. Combined vp9_denoiser_8xM_sse2 and vp9_denoiser_4xM_sse2 into one function vp9_denoiser_NxM_sse2_small and passed the bitexact testing. Changed the name of the function vp9_denoiser_64_32_16xM_sse2 to vp9_denoiser_NxM_sse2_big. Change-Id: Ib22478df585994dd347ebae04202c0b701e7f451	2014-10-28 09:36:58 -07:00
Yaowu Xu	2a506e33b4	Merge "Add a new control of golden frame boost in CBR mode"	2014-10-28 09:32:58 -07:00
Yaowu Xu	e5cd51880e	Allow update of golden refernce buffer in CBR mode This commit changes to allow the usage of golden reference frame in VP9 CBR mode to improve quality. VP9 supports potentially up to 8 reference buffers, it has reference buffers available for this purpose. This was not possible in VP8 as golden and alt-ref buffers were used for temporal scalability purpose in CBR mode in WebRTC. For frames that update golden frame, there can be a quality boost. The amount of allowed bitrate boost can be controlled via parameter rc_max_inter_bitrate_pct. The inital value of the boost ratior is currently based on over_shoot_pct. Further experiments will work out the adaption of this boost value. Change-Id: I0c5f010c8fd8b7b598f69779c1b30e5b2ac30a4d	2014-10-28 09:31:10 -07:00
Paul Wilkins	422d7bc918	Relax maximum Q for extreme overshoot. Added code to relax the active maximum Q in response to extreme local overshoot to reduce bandwidth peaks. The impact is small in metrics terms, but it this helps reduce bandwidth spikes and overall overshoot in a number of clips in our tests sets (especially the YT test set). In particular this should help prevent very big spikes where a clip is mainly easy but has a short hard section. In such a case a choice of maximum Q for the clip as a whole may allow us to hit the overall target rate but give some extreme spikes. The chunked encoding in YT mitigates this problem but it can show up where a longer clip is coded as a single chunk. Change-Id: I213d09950ccb8489d10adf00fda1e53235b39203	2014-10-28 13:03:06 +00:00
Jingning Han	07436abb86	Use zero motion vector in choose_partitioning The zero motion vector was effectively used in the subsampled pixel based variance calculation. This commit makes it directly use zero mv to generate prediction. Change-Id: Ica83dc843e9f8da2f89c3ef451e50f16214c0def	2014-10-27 19:38:43 -07:00
Jingning Han	d56b3eb0cf	Refactor encoder tile data structure Make the common tile info as one element in the encoder tile data struct. Change-Id: I8c474b4ba67ee3e2c86ab164f353ff71ea9992be	2014-10-27 19:37:13 -07:00
Yaowu Xu	03a60b78db	Add a new control of golden frame boost in CBR mode 0 means that golden boost is off, and uses average frame target rate, a non-zero number means the percentage of boost over average frame bitrate is given initially to golden frames in CBR mode. Change-Id: If4334fe2cc424b65ae0cce27f71b5561bf1e577d	2014-10-27 13:55:18 -07:00
Jingning Han	192010d218	Refactor rtc coding mode to support tile encoding Use per tile threshold in the prediction mode search process. Change-Id: I6c74ee5a3b069bb4281002dfe51310911a0756c0	2014-10-27 09:53:46 -07:00
Yaowu Xu	aa2af3ff6e	Merge "Add a new control of max bitrate for inter frame"	2014-10-27 08:11:54 -07:00
Jingning Han	ac53c41e64	Merge "Tile based adaptive mode search in RD loop"	2014-10-24 18:44:52 -07:00
Yaowu Xu	636099f7b6	Add a new control of max bitrate for inter frame Change-Id: I205de3611622cff7f751ea8baf9f82784581730a	2014-10-24 10:19:28 -07:00
Jingning Han	eee201c221	Tile based adaptive mode search in RD loop Make the spatially adaptive mode search in rate-distortion optimization loop inter tile independent. Experiments suggest that this does not significantly change the coding staticstics. Single tile, speed 3: pedestrian_area 1080p 1500 kbps 59192 b/f, 40.611 dB, 101689 ms blue_sky 1080p 1500 kbps 58505 b/f, 36.347 dB, 62458 ms mobile_cal 720p 1000 kbps 13335 b/f, 35.646 dB, 45655 ms as compared to 4 column tiles, speed 3: pedestrian_area 1080p 1500 kbps 59329 b/f, 40.597 dB, 101917 ms blue_sky 1080p 1500 kbps 58712 b/f, 36.320 dB, 62693 ms mobile_cal 720p 1000 kbps 13191 b/f, 35.485 dB, 45319 ms Change-Id: I35c6e1e0a859fece8f4145dec28623cbc6a12325	2014-10-24 10:00:27 -07:00
Paul Wilkins	60d192db04	Merge "Enable dual arf with constant q."	2014-10-24 05:51:25 -07:00
Paul Wilkins	3758650c98	Merge "Move frame re-sizing into the recode loop"	2014-10-24 05:50:39 -07:00
Adrian Grange	65753eeb8a	Move frame re-sizing into the recode loop The point at which frames are scaled to their coded dimensions is moved into the re-code loop. This is in preparation for a further patch that will add logic into the re-code loop to reduce the coded frame size if the encoder is struggling to hit the target data rate at the native frame size. Change-Id: Ie4131f5ec6fb93148879f6ce96123296442bf2d1	2014-10-23 16:20:57 -07:00
Yaowu Xu	86777f2e1e	Merge "Move filter_ref initialization"	2014-10-23 11:20:22 -07:00
Yaowu Xu	065809d286	Move filter_ref initialization To outside the loop to avoid repeating the operations. Change-Id: I66c1986e98ce0d7594caad3d3b45de655b299bff	2014-10-23 08:27:25 -07:00
Paul Wilkins	8fc3ab774f	Enable dual arf with constant q. Add second level arf Q adjustment when using dual arfs in constant Q mode. Previously in constant Q mode enabling dual arf hurt by ~5% but with this change the average benefit is ~1-1.5% with some mid range data points up ~10%. Note however that it still hurts on some clips including some very low motion show content. Change-Id: I5b7789a2f42a6127d9e801cc010c20a7113bdd9b	2014-10-23 13:19:31 +01:00
Paul Wilkins	9363425daa	Merge "Initialization bug for multi arf."	2014-10-23 02:02:48 -07:00
Jingning Han	41a17f4457	Merge "Allow checking zeromv mode in vp9_pick_inter_mode"	2014-10-22 18:46:20 -07:00
Yunqing Wang	330a6b2756	Merge "vp9_ethread: allocate frame contexts outside VP9_COMMON struct"	2014-10-22 17:10:39 -07:00
Yunqing Wang	7c7e4d4eb8	vp9_ethread: allocate frame contexts outside VP9_COMMON struct This patch allocated frame contexts outside VP9_COMMON. This allows multiple threads to share the same copy of frame contexts, and reduces the overhead. It also guarantees the correct update of these contexts during bitstream packing. This patch doesn't change encoding result. Change-Id: Ic181a2460b891d1d587278a6d02d8057b9dbd353	2014-10-22 15:03:12 -07:00
Yaowu Xu	7c48a295ae	Merge "Fix a subtle issue in re-use inter_pred"	2014-10-22 14:53:06 -07:00
Jingning Han	08cdd006e1	Allow checking zeromv mode in vp9_pick_inter_mode This improves the compression performance of speed -5 by 0.6%. The speed impact is less than 1%. Change-Id: Ie77daa561976dfc8b479061e1221bdf428eb0c3b	2014-10-22 14:47:15 -07:00
JackyChen	897500b9ba	Merge "vp9_denoiser_sse2.c: improve code style."	2014-10-22 13:52:03 -07:00
Yaowu Xu	3f79359e0a	Fix a subtle issue in re-use inter_pred The initialization of this_mode_pred does not work when the ref_frame loop ever goes beyond LAST_FRAME. This commit fixes the subtle issue and allows potentially expanding the loop to test GOLDEN_FRAME. Change-Id: Ibbd427a22160d1d9eacb8ed0c87f88d6cef9c0f3	2014-10-22 12:06:27 -07:00
JackyChen	5cba6516aa	vp9_denoiser_sse2.c: improve code style. denoiser_sse2.c: fix typos in comment. Change-Id: Ic0fb102331b0e533c058da3cab1fbc30de9a0070	2014-10-22 10:55:54 -07:00
Paul Wilkins	7cd6330ef3	Initialization bug for multi arf. Moved erroneous reset of cpi->multi_arf_last_grp_enabled. Change-Id: Ibb0b96f6ed1d5eeb575a3b1c798e0fe2ee651d06	2014-10-22 18:51:07 +01:00
Hangyu Kuang	9ce3a7d76c	Implement frame parallel decode for VP9. Using 4 threads, frame parallel decode is ~3x faster than single thread decode and around 30% faster than tile parallel decode for frame parallel encoded video on both Android and desktop with 4 threads. Decode speed is scalable to threads too which means decode could be even faster with more threads. Change-Id: Ia0a549aaa3e83b5a17b31d8299aa496ea4f21e3e	2014-10-22 10:50:58 -07:00
Jingning Han	0e64aa5073	Merge "Refactor rate distortion cost structure in non-RD coding mode"	2014-10-22 08:41:36 -07:00
Yaowu Xu	87665f16f4	Merge "Change speed features for good quality(cpu-used=5)"	2014-10-22 08:40:15 -07:00
Jingning Han	be212d4db3	Refactor rate distortion cost structure in non-RD coding mode This commit refactors the rate distortion structure used in the non-RD coding mode and saves a few RDCOST calculations. Change-Id: I62c3416c300d2c5372f21b96d93a6b633a34ab3a	2014-10-21 17:17:11 -07:00
Hui Su	8947b18fa3	Move the definition of switchable filter numbers into enum INTERP_FILTER; Modify the macro ADD_MV_REF_LIST and IF_DIFF_REF_FRAME_ADD_MV. Change-Id: Ic36c9eb6ccb8ec324d991f7241e42b40b60b1dcb	2014-10-21 15:41:37 -07:00
Yaowu Xu	c30f7e6cc5	Change speed features for good quality(cpu-used=5) The existing speed features produce horrible encoding results, almost 30% worse than cpu-used=4, this commit adjust the speed features to produce relatively resonable results to be within 3%-5% of cpu-used=4. Change-Id: I0ca6ebafb33024d4a0cbcf04c78a4a00b8dd1ecf	2014-10-21 11:59:12 -07:00
Jingning Han	1ed1dde06d	Remove unused copy_partitioning Change-Id: I75a2a3772ed17e73180eb4f263cc838cae4927b0	2014-10-21 09:47:58 -07:00
Jingning Han	f2c21cfa1e	Merge "Remove deprecated constrain_copy_partitioning function"	2014-10-21 09:44:11 -07:00
Jingning Han	55ec7ebca7	Merge "Remove unused sb_has_motion function in vp9_encodeframe.c"	2014-10-21 09:43:55 -07:00
Jingning Han	072d844aff	Merge "Remove deprecated use_lastframe_partitioning feature"	2014-10-21 09:43:45 -07:00
Jingning Han	61ff08ef61	Merge "Hybrid partition search for rtc coding mode"	2014-10-21 09:43:35 -07:00
Paul Wilkins	4163da33c2	Merge "Extend --auto-alt-ref so it can enable multi-alt ref."	2014-10-21 07:02:24 -07:00
Paul Wilkins	889c53a507	Merge "Resolve compiler warning."	2014-10-21 07:02:13 -07:00
Jingning Han	abb2fbb10e	Remove deprecated constrain_copy_partitioning function Its functionality has been replaced with choose_partitioning and threshold based control on split mode check. Change-Id: Ic9bb321df06b524f5c38ea5874dc6f6a8f93c5e3	2014-10-20 17:08:21 -07:00
Jingning Han	ef53898c48	Remove unused sb_has_motion function in vp9_encodeframe.c Change-Id: I035fb6aa5c10741b065e27befb097d8087e3c62f	2014-10-20 17:08:11 -07:00
Jingning Han	e62ce79e1a	Remove deprecated use_lastframe_partitioning feature This speed feature has been deprecated in both yt and rtc coding modes. This commit removes the related operations. Change-Id: I079c79c6adafe45581af2ebf8b98faebcface1ce	2014-10-20 17:03:38 -07:00
Jingning Han	9f128b3ed9	Hybrid partition search for rtc coding mode This commit re-designs the recursive partition search scheme in rtc speed -5. It first checks if the current block is under cyclic refresh mode. If so, apply recursive partition search. Otherwise, perform sub-sampled pixel based partition selection. When the pre-selection finds the partition size should be 32x32 or above, use the partition size directly. Otherwise, apply partition search at nearby levels around the preset partition size. It is enabled in speed -5. The compression performance of rtc speed -5 is improved by 9.4%. Speed wise, the run-time goes slower from 1% to 10%. nik_720p, 1000 kbps 33220 b/f, 38.977 dB, 10109 ms -> 33200 b/f, 39.119 dB, 10210 ms vidyo1_720p, 1000 kbps 16536 b/f, 40.495 dB, 10119 ms -> 16536 b/f, 40.827 dB, 11287 ms Change-Id: I65adba352e3adc03bae50854ddaea1b421653c6c	2014-10-20 13:02:12 -07:00
Yunqing Wang	687c56e802	Merge "SAD32xh and SAD64xh for AVX2"	2014-10-20 12:37:55 -07:00
Yunqing Wang	67c866750c	Merge "Remove the dependency in token storing locations"	2014-10-20 08:26:46 -07:00
Paul Wilkins	6f0ae3a2d1	Extend --auto-alt-ref so it can enable multi-alt ref. Extend --auto-alt-ref from parameter so we can use it to turn multi-arf on and off from the command line. For now the range is 0-off, 1-on, 2-multi-arf on. Rename play_alternate to enable_auto_arf Change-Id: Id7b64407cfbe76ba0090a83b588a03e22a240386	2014-10-20 16:09:37 +01:00
Paul Wilkins	9626a0cb62	Resolve compiler warning. conversion from 'const int64_t' to 'int', possible loss of data. Change-Id: I471a73bba5d448d9be0ef9cbf1590fa73aa74be1	2014-10-20 12:08:33 +01:00
Paul Wilkins	9c98fb2bab	Merge "Alter adjustment of two pass GF/ARF boost with Q."	2014-10-20 03:12:06 -07:00
levytamar82	7045aec00a	SAD32xh and SAD64xh for AVX2 All sad function that process above 32 consecutive elements are optimized for AVX2: vp9_sad64x64 vp9_sad64x32 vp9_sad32x64 vp9_sad32x32 vp9_sad32x16 vp9_sad64x64_avg vp9_sad64x32_avg vp9_sad32x64_avg vp9_sad32x32_avg vp9_sad32x16_avg The functions that appeared as a hotspot is vp9_sad32x32 and vp9_sad64x64 vp9_sad32x32 was optimized by 68% and vp9_sad64x64 was optimized by 90% both of them gave and overall ~2.3% user level gain Change-Id: Iccf86b375a2b54c5fbbe685902ead0c9a561b9fd	2014-10-19 13:59:10 -07:00
Debargha Mukherjee	6202c75f84	Merge "Add highbitdepth function for vp9_avg_8x8"	2014-10-18 14:37:10 -07:00
Yaowu Xu	06e65269c7	Merge "Remove unused VAR_BASED_FIXED_PARTITION flag"	2014-10-18 13:31:47 -07:00
Yaowu Xu	7bf475926b	Merge "Use rate/distortion thresholds to control non-RD partition search"	2014-10-18 13:31:41 -07:00
Peter de Rivaz	73ae6e495c	Add highbitdepth function for vp9_avg_8x8 Cherry-picked from https://gerrit.chromium.org/gerrit/#/c/71914/ (`a92f987a6b`) on highbitdepth branch. Change-Id: I6903e4e4cb57d90590725c8a1c64c23da7ae65e8	2014-10-17 17:04:37 -07:00
Yunqing Wang	7c4992c466	Remove the dependency in token storing locations Currently, the tokens for a tile are stored immediately after its preceding tile, which causes a dependency. This is unnecessary since we always allocate enough memory for tokens. Removing the dependency allows token writing done in parallel. This patch doesn't change encoding result. Change-Id: I7365a6e5e2c2833eb14377c37e1503c9d0f26543	2014-10-17 14:25:33 -07:00
JackyChen	0615a196e2	Merge "vp9_denoiser_sse2.c: solve windows build error."	2014-10-17 11:16:22 -07:00
Jingning Han	3bc94cd2eb	Merge "Add init and reset functions for RD_COST struct"	2014-10-17 11:15:19 -07:00
Jingning Han	4a6c99e5fd	Merge "Reset rate cost value in rd mode search"	2014-10-17 11:15:03 -07:00
Jingning Han	94ecfa323f	Reset rate cost value in rd mode search When early termination is triggered, properly reset the rate cost to invalid value to avoid potential ioc issue. Change-Id: I3444390be2e49a34bb02cf8a74c33d5dbd96d88d	2014-10-17 09:33:59 -07:00
JackyChen	6356d21a47	vp9_denoiser_sse2.c: solve windows build error. Change-Id: Ib5df91c8580d5dbeb0b3554edc9c2ca906ba4c4d	2014-10-17 09:28:22 -07:00
Jingning Han	e1111fba7e	Remove unused VAR_BASED_FIXED_PARTITION flag Change-Id: I4ce19b7cb1c45fed86e81ee785e787630020fb4f	2014-10-17 09:02:25 -07:00
Paul Wilkins	f0c3da930a	Alter adjustment of two pass GF/ARF boost with Q. Delete gfboost_qadjust() and move Q based adjustment into calc_frame_boost(). Also remove clamping. Making the adjustment here means that it influences not just the boost level but also the selection of the GF/ARF interval. This change gives a small average gain in PSNR but larger gains in SSIM, especially for harder std-hd set (1.5%) Change-Id: I3aa81b8feccaeff93d915e19fb9cf5cd64c86327	2014-10-17 16:37:17 +01:00
James Zern	00f1cf40ed	Merge "vp9_denoiser_sse2.c: eliminate gcc warnings"	2014-10-17 03:26:06 -07:00
JackyChen	8514d03402	vp9_denoiser_sse2.c: eliminate gcc warnings Change-Id: I5f63f48e11e31ea9951223c5b18f42a2471e4560	2014-10-17 11:00:57 +02:00
Jingning Han	4b5f01e12e	Merge "Fix an ioc issue in super_block_uvrd"	2014-10-16 12:35:11 -07:00
Jingning Han	ed100c0b00	Fix an ioc issue in super_block_uvrd This commit fixes an ioc issue that will happen when the cumulative variables are not in effective use. The fix discards these redundant additions. Change-Id: Idbac5bfb989c0cedc5f8a323effce938519b2457	2014-10-16 11:07:39 -07:00
Paul Wilkins	716ae78ce4	Change initialization of static_scene_max_gf_interval. This removes an unnecessary restriction that causes a problem (noticed by AWG) when the forced key frame interval is set to a very small value, such as 10. In this case we were being forced to code minimal length GF groups. Change-Id: I76ef5861a09638ff51f61fea02359554184ada53	2014-10-16 16:18:29 +01:00
Minghai Shang	68b550f551	[spatial svc]Another workaround to avoid using prev_mi We encode a empty invisible frame in front of the base layer frame to avoid using prev_mi. Since there's a restriction for reference frame scaling factor, we have to make it smaller and smaller gradually until its size is 16x16. Change remerged. Change-Id: I9efab38bba7da86e056fbe8f663e711c5df38449	2014-10-16 16:09:40 +01:00
Paul Wilkins	d5130af568	Revert "Move input frame scaling into the recode loop" This reverts commit `452dc21500`. This change has introduced a significant quality regression on content with forced key frames. (e.g. the YT and yt-hd set). It is most noticeable in static content where the kf bits dominate. Here, despite key frames being apparently coded at the same Q, there is a drop in all metrics of ~20% (e.g clXR and BFa0). Change-Id: Iba14cc61778c0846fa0a59c33c55a9fc49512cb4	2014-10-16 15:54:40 +01:00
Paul Wilkins	468032961d	Revert "[spatial svc]Another workaround to avoid using prev_mi" This reverts commit `c113457af9`. Temporary revert to allow clean revert of another commit. Change-Id: Ia9b7b755e6c48e1b6e383329f121fef175a24b27	2014-10-16 15:52:08 +01:00
Marco	48ea5b7190	Merge "Some updates for Speed 6/VAR_BASED_PARTITION."	2014-10-15 15:57:21 -07:00
Jingning Han	e2612fbd70	Add init and reset functions for RD_COST struct Change-Id: I2902de7051a883fd22e27a655209233733969cfd	2014-10-15 15:02:06 -07:00
Jingning Han	801b77d48c	Merge "Replace copy_partitioning use case with choose_partitioning"	2014-10-15 14:54:52 -07:00
Jingning Han	5e766ccee0	Use rate/distortion thresholds to control non-RD partition search Compare the estimated rate and distortion to the thresholds scaled according to the operating block size and determine if further split partition search will be run. The compression performance of speed -5 is changed by -0.074%. The encoding speed is 10% - 15% faster. vidyo1 720p 16545 b/f, 40.492 dB, 11475 ms -> 16535 b/f, 40.486 dB, 10100 ms nik720p 16624 b/f, 36.310 dB, 10071 ms -> 16617 b/f, 36.313 dB, 8346 ms Change-Id: Ic9197ab5761279ae55d2fb7813b2af0e0db497b8	2014-10-15 13:40:33 -07:00
Marco	09ea74f194	Some updates for Speed 6/VAR_BASED_PARTITION. Reduce the intra_cost_penalty for non-rd mode, and some updates to VAR_BASED_PARTITION. Visual tests show some improvement at Speed 6, for RTC clips. Change-Id: If9090daf7aed14906a32d931a538ab544bbca606	2014-10-15 12:06:48 -07:00
Jingning Han	89b8c7a513	Replace copy_partitioning use case with choose_partitioning This commit replaces the use of copy_partitioning with choose_partitioning based on the sse of subsamped pixels, which provides significantly better coding performance and runs at similar speed, as compared to copy_partitioning. It improves rtc speed 5 coding performance by 3%. Change-Id: I52d3682a12dce0147f5e52383a594fc242ca3228	2014-10-15 11:37:20 -07:00
Minghai Shang	c113457af9	[spatial svc]Another workaround to avoid using prev_mi We encode a empty invisible frame in front of the base layer frame to avoid using prev_mi. Since there's a restriction for reference frame scaling factor, we have to make it smaller and smaller gradually until its size is 16x16. Change-Id: I60b680314e33a60b4093cafc296465ee18169c19	2014-10-14 16:26:39 -07:00
Adrian Grange	2040bb58fb	Merge "Move input frame scaling into the recode loop"	2014-10-14 15:30:42 -07:00
Alex Converse	00a9671bbd	Merge "Add a 32-bit friendly sse2 quantizer."	2014-10-14 14:35:02 -07:00
Yunqing Wang	a614f2288c	Remove an unneeded function call set_tile_limits() is called in vp9_change_config() already. Change-Id: I91c3a0df2c1c7fd7e71546d8f51fd5b65838a7da	2014-10-14 11:41:37 -07:00
Alex Converse	7497d2fb23	Add a 32-bit friendly sse2 quantizer. This is based on the 64-bit ssse3 quantizer. 1.1x speedup for screen content at speed 7. Change-Id: I57d15415ef97c49165954bbe3daaaf9318e37448	2014-10-14 11:37:41 -07:00
Jingning Han	f67e75a6f4	Merge "Refactor super_block_uvrd function to remove goto statement"	2014-10-14 11:33:00 -07:00
Jingning Han	f3a5de816d	Refactor super_block_uvrd function to remove goto statement Use return value 0/1 as indicator of the validity of the rate- distortion cost. Change-Id: I6244126fbf03472cebcba4f177a6cd329fae4743	2014-10-14 09:58:11 -07:00
Adrian Grange	452dc21500	Move input frame scaling into the recode loop Move the point at which input frames are scaled into the recode loop. This will allow us to change the coded frame size dynamically in response to previous attempts to encode the frame at a higher resolution. A following patch will implement a scheme for resizing the frame in the recode loop. Change-Id: I6a59c02d6ac1626512edad6de8b60063b79433e6	2014-10-14 09:27:55 -07:00
Jingning Han	d0369d6fd4	Merge "Use speed feature variable in vp9_rd_pick_inter/intra_mode"	2014-10-14 09:10:24 -07:00
Jingning Han	fdf2205558	Merge "Fix vp9_rd_pick_inter/intra function types"	2014-10-14 09:10:11 -07:00
Jingning Han	790a96c94f	Merge "Refactor rate distortion cost structure"	2014-10-14 08:58:55 -07:00
Jingning Han	69a09a70e9	Use speed feature variable in vp9_rd_pick_inter/intra_mode Replace repeated fetch cpi->sf with a local sf pointer. Change-Id: I5a55bba3e1c41fbdbc6ad5f078d2fa49dd95ee67	2014-10-13 16:15:00 -07:00
Jingning Han	3bdb6bfcee	Fix vp9_rd_pick_inter/intra function types The returned value is not used anywhere, hence changing the function type into void. Change-Id: I0ece49ed61e7aab6df01140135503ad41d4ef4a4	2014-10-13 16:00:46 -07:00
Jingning Han	811cef97c9	Refactor rate distortion cost structure This commit makes a struct that contains rate value, distortion value, and the rate-distortion cost. The goal is to provide a better interface for rate-distortion related operation. It is first used in rd_pick_partition and saves a few RDCOST calculations. Change-Id: I1a6ab7b35282d3c80195af59b6810e577544691f	2014-10-13 14:27:16 -07:00
Paul Wilkins	6dbb9e4d44	Clamp rate error estimate. Add back clamp which ensures that the Q adaptation is turned off when the over_shoot_pct and under_shoot_pct parameters are set to 100. Change-Id: Id0161b114d39a3029cd3eb28020caab0c3914922	2014-10-13 18:07:58 +01:00
Paul Wilkins	f7f0eaa581	Add adaptation option for VBR. Allow min and maxQ to creep when the undershoot or overshoot exceeds thresholds controlled by the command line under_shoot_pct and over_shoot_pct values. Default is 100%,100% which ~disables adaptation. Derf results for example undershoot% / overshoot%:- Head:- Mean abs (%rate error) = 14.4% This check in:- 25%/25% - Mean abs (%rate error) = 6.7% PSNR hit -1% SSIM -0.1% 5% / 5% - Mean abs (%rate error) = 2.2% PSNR hit -3.3% SSIM - 1.1% Most of the remaining error and most of the quality hit is at extreme data rates. The adaptation code still has an exception for material that is in effect static so that we don't over adjust and over spend on YT slide show type content. (Rebase of If25a2449a415449c150acff23df713e9598d64c9 to resolve a auto-merge error) Change-Id: Iec4e1613ef0d067454751d8220edb7058dfbd816	2014-10-13 10:16:44 +01:00
Jingning Han	a62acf3c0a	Fix ActiveMapTest valgrind warning This fixes a valgrind warning in the ActiveMapTest unit test reported in issue 870. Change-Id: Idf172ab0244ebefe630c3577e649bc9ba7c43d10	2014-10-11 22:36:58 -07:00
Alex Converse	a90255c366	Revert "Add adaptation option for VBR." This reverts commit `869d4ca519`. This breaks the build via conflict with `e18edd5eb6`. Change-Id: If544b99e367a449452834eb8cce600f58c34ec0d	2014-10-10 11:34:00 -07:00
Paul Wilkins	169949dd74	Merge "Add adaptation option for VBR."	2014-10-10 09:22:58 -07:00
Yaowu Xu	bdea0055b2	Merge "vp9/choose_partitioning: add missing clear_system_state"	2014-10-10 09:16:19 -07:00
James Zern	a3e1a9291a	vp9/choose_partitioning: add missing clear_system_state set_vt_partitioning does double math Change-Id: I8e9d73d5c89b937a5326abf04164d24d9d88c5ef	2014-10-10 08:14:46 -07:00
Paul Wilkins	869d4ca519	Add adaptation option for VBR. Allow min and maxQ to creep when the undershoot or overshoot exceeds thresholds controlled by the command line under_shoot_pct and over_shoot_pct values. Default is 100%,100% which ~disables adaptation. Derf results for example undershoot% / overshoot%:- Head:- Mean abs (%rate error) = 14.4% This check in:- 25%/25% - Mean abs (%rate error) = 6.7% PSNR hit -1% SSIM -0.1% 5% / 5% - Mean abs (%rate error) = 2.2% PSNR hit -3.3% SSIM - 1.1% Most of the remaining error and most of the quality hit is at extreme data rates. The adaptation code still has an exception for material that is in effect static so that we don't over adjust and over spend on YT slide show type content. Change-Id: If25a2449a415449c150acff23df713e9598d64c9	2014-10-10 12:54:16 +01:00
James Zern	7c6fec672f	vp9_avg_intrin_sse2: correct intrinsics include immintrin.h -> emmintrin.h fixes build where newer intrinsics are unavailable Change-Id: I79311b39bfa782fc2abeb45884ecb417050cb9f8	2014-10-10 10:05:47 +02:00
Deb Mukherjee	9a29fdbae7	Merge "Rename highbitdepth functions to use highbd prefix"	2014-10-09 15:39:56 -07:00
Deb Mukherjee	1929c9b391	Rename highbitdepth functions to use highbd prefix Uses highbd_ prefix convention consistently. Change-Id: I58f7f799a7ff8e32701bcd71c955bcf1cdd4581e	2014-10-09 14:40:40 -07:00
Jingning Han	112789d4f2	Merge "Remove sub8x8 block index from rd_pick_partition argument"	2014-10-09 11:16:11 -07:00
Deb Mukherjee	3117830af3	Merge "Subpel search cleanups and enhancements"	2014-10-09 11:14:51 -07:00
Alex Converse	9ffbc31367	Merge "Move the high freq coeff check outside store_coding_context"	2014-10-09 11:12:02 -07:00
Jingning Han	6a0d291fce	Remove sub8x8 block index from rd_pick_partition argument This parameter is deprecated. Its function is replaced with other explicit condition check. Change-Id: I61337e350ba8ca9eb50382db8b4d4acbf45cb7eb	2014-10-09 09:20:16 -07:00
Yaowu Xu	3af39b9997	Merge "Fix src frame buffer copy and extend"	2014-10-09 07:08:48 -07:00
James Zern	cec763bd97	set_vt_partitioning: fix type conversion warning double -> int64 + make threshold_multiplier an int Change-Id: I6d3607fdf13d670f57c9d9b04a80acb2be1346a0	2014-10-09 11:41:36 +02:00
Deb Mukherjee	d78dbff09a	Subpel search cleanups and enhancements - Some fixes to surface fit. - Returns variance function as cost rather than sad in the pattern search and diamond search functions. Only vp9_pattern_search_sad function used in bigdia search uses sad as integer 1-away costs. - Deploys SUBPEL_TREE_PRUNED_MORE for speed 4+. Results: derf [Speed 3]: About +0.036% in coding efficiency without any discernible speed loss. derf [Speed 4]: About 2-3% faster at -0.199% loss in coding efficiency. derf [Speed 5]: About 3-4% faster at -0.149% loss in coding efficiency. Change-Id: I8462f94f6adb46966ca964f2bd0400977357fd63	2014-10-08 23:59:43 -07:00
Yunqing Wang	189566db58	Merge "Allow mode search breakout at very low prediction errors"	2014-10-08 19:58:18 -07:00
Yunqing Wang	e18edd5eb6	Allow mode search breakout at very low prediction errors In model_rd_for_sb function, the spatial domain SSE and variance are checked to see if transform coefficients are quantized to 0. Besides that, this patch adds another set of thresholds that are much more strict. These thresholds are used to conduct a partition block level check to measure if all its TX blocks are skippable for YUV planes. If it is true, x->skip is set for this partition block, and thus its mode search is terminated. This speeds up the encoding at very low prediction error case, such as screen sharing application. This patch covers what rd_encode_breakout_test() does, so that function is removed. Borg test at speed 3 shows: For stdhd set, psnr: +0.008%, ssim: +0.014%; For derf set, psnr: +0.018%, ssim: +0.025%. No noticeable speed change. Change-Id: I4e5f15cf10016a282a68e35175ff854b28195944	2014-10-08 17:46:22 -07:00
Jingning Han	5fcbcf1b22	Move the high freq coeff check outside store_coding_context This fixes valgrind message issue 870. Change-Id: Ibbc2481923a2995029ab05de30c9e8a6e9f0f9a8	2014-10-08 16:10:32 -07:00
Jingning Han	41cea46154	Use local variable in vp9_rd_pick_inter_mode_sb Change-Id: Ie35a965a6b8de536ccaf61ff61498620d22db205	2014-10-08 16:09:47 -07:00
Yaowu Xu	d602500d4a	Fix src frame buffer copy and extend For input source with size that is not multiple of 8, the size is rounded to 8 and saved in width or height, the original source sizes are saved in crop_width and crop_height. This commit corrects the computation of bottom and right extension amounts to use the orignal sizes, hence crop_width and crop_height. In addition, this commit also adds the missed initialization for uv_crop_width and uv_crop_height. This addresses issue #834 Change-Id: I084543ca7645a4964b88f7cf8ff668f517d3a39b	2014-10-08 11:07:04 -07:00
Jim Bankoski	20254d1daa	Merge "experimental : partition using 1/8 x 1/8 image"	2014-10-08 09:04:26 -07:00
Jim Bankoski	4130e691bb	Merge "Force better lower quantizer keyframe in case of high quantizer."	2014-10-08 09:04:01 -07:00
Paul Wilkins	2faff64866	Merge "Improve two pass VBR accuracy."	2014-10-08 04:23:30 -07:00
Paul Wilkins	679a1c2f5a	Merge "Two pass rc changes."	2014-10-08 04:23:20 -07:00
Jim Bankoski	0ce51d823f	experimental : partition using 1/8 x 1/8 image The concept: There's too much noise in source pixels for variance and at low bitrate the reconstructed looks nothing like the source so we have problems getting good partitionings with either. This skirts the issue by using a box blur scaled down version for variance calculations. To compare against source_var_ moved keyframe to be rd based like source_var. Change-Id: Ie3babdbfadae324b7b5a76bea192893af27f0624	2014-10-07 16:36:14 -07:00
Jim Bankoski	dae97868da	Force better lower quantizer keyframe in case of high quantizer. Change-Id: Ie69a164bc166b6a8819777038d65a7d9f9c3361f	2014-10-07 15:40:02 -07:00
Jingning Han	3bbec7b422	Merge "Replace mi_width_log2() with mi_width_log2_lookup table"	2014-10-07 15:33:52 -07:00
Jingning Han	27c9577f8e	Merge "Take out repeated block width/height lookup functions"	2014-10-07 15:33:45 -07:00
Yunqing Wang	0cac69f594	Merge "Fix skip_txfm issue in rdopt code"	2014-10-07 15:13:44 -07:00
Jingning Han	bd9706506f	Merge "Move inter filter defs to vp9_filter.h"	2014-10-07 13:42:26 -07:00
Jingning Han	91442e1626	Merge "Reduce the scope of the header file used in vp9_context_tree.h"	2014-10-07 13:42:17 -07:00
Jingning Han	92b45e5d4e	Merge "Remove redundant header file from vp9_encoder.h"	2014-10-07 13:42:13 -07:00
Yunqing Wang	a4aa14020a	Fix skip_txfm issue in rdopt code Fixed an encoder crash. Set skip_txfm to 0 for cases that skip_txfm isn't calculated. Put memcpy of skip_txfm at right place. Change-Id: Ib3b6afc1b251a85b2a853c8138fb3393f48cfef6	2014-10-07 12:47:43 -07:00
Jingning Han	7ee58985bd	Replace mi_width_log2() with mi_width_log2_lookup table Change-Id: If0ea98aa139d14d40cd924114e18396aff36b5a5	2014-10-07 12:45:25 -07:00
Jingning Han	b66f7016c1	Take out repeated block width/height lookup functions The functions b_width_log2 and b_height_log2 only do direct table fetch. This commit unifies such use cases by using the table directly and removes these functions. Change-Id: I3103fc6ba959c1182886a2799d21b8b77c8a7b6b	2014-10-07 12:33:07 -07:00
Jingning Han	5d9cdac087	Move inter filter defs to vp9_filter.h Add comments on the use case of these definitions. Further reduce the scope of header file in vp9_context_tree.h. Change-Id: Ic4a7638e838d0ac441b64abfc56e57354c059d75	2014-10-07 12:16:37 -07:00
Deb Mukherjee	cfc337aae8	Merge "Resolves some static analysis / undefined warnings"	2014-10-07 12:15:26 -07:00
Deb Mukherjee	fced63ed30	Resolves some static analysis / undefined warnings Also fixes a case of distortion becoming negative and messing up the RDCOST computation. Change-Id: Id345af9e8dfff31ade622be5756e51f2cdface53	2014-10-07 11:20:56 -07:00
Jingning Han	5e32036b97	Reduce the scope of the header file used in vp9_context_tree.h Change-Id: I264ee35044a5973c7725daba7af870968353a3c1	2014-10-07 11:13:35 -07:00
JackyChen	a9f479682a	Merge "Add SSE2 code and unit test for VP9 denoiser."	2014-10-07 10:51:55 -07:00
Jingning Han	3f93f23120	Remove redundant header file from vp9_encoder.h Change-Id: Ia212390cf8d36db5436bb0f0e1b696f70066341a	2014-10-07 10:49:58 -07:00
Jingning Han	a75551585b	Fix eobs buffer pointer mis-use This commit fixes a buffer pointer mis-use in store_coding_context. The compression performance for stdhd set of speed 3 is improved by 0.097%. It fixes issue 869. Change-Id: Idc59e22035eaf39f7133ca04174894374d647ff7	2014-10-06 15:57:13 -07:00
JackyChen	80465dae88	Add SSE2 code and unit test for VP9 denoiser. This SSE2 is based on VP8 denoiser's SSE2 code. In VP8, there are only 16x16 blocks in denoiser, while in VP9, there are 13 different block sizes. By adding this SSE2 code, the improvement of encoder speed is around 20%(using C code vs using SSE2 code), vary for different clips. The unit test for VP9 denoiser is to confirm that the SSE2 code is bit-exact with the C code. The unit test covers all block size. Change-Id: Ic8d8ac26db4ea40a5f146b5678a065af07eaaa3d	2014-10-06 15:27:40 -07:00
Jingning Han	1b8c57e915	Merge "Fix an IOC issue in vp9_rd_pick_inter_mode_sb"	2014-10-06 09:29:29 -07:00
Yaowu Xu	5966acc1be	Merge "Properly initialize segmentID in nonrd coding path"	2014-10-06 07:57:36 -07:00
Paul Wilkins	0e1068a4bd	Improve two pass VBR accuracy. Adjustments to the GF interval choice and minimum boost. Adjustment to the calculation of 2 pass worst q. Compared to 09/29 head there is metrics hit on derf of (-0.123%,-0.191%) Compared to the September 29 head and a baseline on September 18 baseline the accuracy of the VBR rate control measured on the derf set is as follows:- Mean error % / Mean abs(error %) Sept 18 baseline (-7.0% / 14.76%) Sept 29 head (-15.7%, 19.8%) This check in (-1.5% / 14.4%) The mean undershoot is reduced slightly but the worst case overshoot on e.g. harbour/highway is increased. This will be addressed in a later patch. Change-Id: Iffd9b0ab7432a131c98fbaaa82d1e5b40be72b58	2014-10-06 14:20:09 +01:00
Jingning Han	085b97aa5c	Fix an IOC issue in vp9_rd_pick_inter_mode_sb It is possible that the GOLDEN reference frame is not avaiable, in which setting the predicted mv will be associated with a residual value of INT_MAX. This commit checks this condition before left shift and comparison with that of ALTREF frame, to avoid overflow issue. Change-Id: Ib98c3149dbdd016f2fe5beaafb13f67d469dd07c	2014-10-05 12:05:14 -07:00
Jingning Han	a021924092	Merge "Fix indent in encode_rd_sb_row"	2014-10-03 15:24:02 -07:00
Jingning Han	a1088e0b5f	Merge "Rework partition search skip scheme"	2014-10-03 15:23:54 -07:00
Yaowu Xu	0065b73481	Properly initialize segmentID in nonrd coding path This commit adds proper initialization of segment id for variance AQ mode in non-rd coding path. It fixes the enc/dec mismatch issue of rt=7 with --aq-mode=1, as reported in issue #816 Change-Id: I02fa41b96345bf2e66077d5ea553f85ba800f7bb	2014-10-03 15:01:53 -07:00
Deb Mukherjee	8a01074d04	Merge "Incorporate WRAPLOW macro into non-highbitdepth tx"	2014-10-03 12:45:39 -07:00
Jingning Han	ef62233396	Fix indent in encode_rd_sb_row Change-Id: Icbcfe7b56d88474f4398b4c5b52f6719d551ab4a	2014-10-03 11:57:36 -07:00
Jingning Han	bb260d9076	Rework partition search skip scheme This commit enables the encoder to skip split partition search if the bigger block size has all non-zero quantized coefficients in low frequency area and the total rate cost is below a certain threshold. It logarithmatically scales the rate threshold according to the current block size. For speed 3, the compression performance loss: derf -0.093% stdhd -0.066% Local experiments show 4% - 20% encoding speed-up for speed 3. blue_sky_1080p, 1500 kbps 51051 b/f, 35.891 dB, 67236 ms -> 50554 b/f, 35.857 dB, 59270 ms (12% speed-up) old_town_cross_720p, 1500 kbps 14431 b/f, 36.249 dB, 57687 ms -> 14108 b/f, 36.172 dB, 46586 ms (19% speed-up) pedestrian_area_1080p, 1500 kbps 50812 b/f, 40.124 dB, 100439 ms -> 50755 b/f, 40.118 dB, 96549 ms (4% speed-up) mobile_calendar_720p, 1000 kbps 10352 b/f, 35.055 dB, 51837 ms -> 10172 b/f, 35.003 dB, 44076 ms (15% speed-up) Change-Id: I412e34db49060775b3b89ba1738522317c3239c8	2014-10-03 11:54:30 -07:00
Deb Mukherjee	d50716face	Incorporate WRAPLOW macro into non-highbitdepth tx Incorporates the WRAPLOW macro into the non-highbitdepth transforms to aid hardware verification between a software C model and an intended hardware implementation though the use of the configure options: --enable-experimental --enable-emulate-hardware. Note that to avoid further discrepancies between the sse/sse2 implementations of the transforms and the C implementation, when the emulate hardware option is invoked, we also disable sse/sse2/etc. Also incudes some minor cleanups/renaming etc. Change-Id: Ib864d8493313927d429cce402982f1c8e45b3287	2014-10-03 11:38:05 -07:00
Deb Mukherjee	2c7b94f6ec	Merge "Prevent negative cost for highbitdepth"	2014-10-03 11:37:47 -07:00
Deb Mukherjee	431cdc33ee	Prevent negative cost for highbitdepth Adds proper scaling for highbitdepth in a rdopt cost. Change-Id: I066694799a7f491b830945ef1c66eb202071c355	2014-10-03 10:22:21 -07:00
Deb Mukherjee	00a4b20fbe	rdmult data type change To fix a VS warning. Change-Id: I4c530c0afe8d06acdb8cc78b7995aba57a25373d	2014-10-03 00:09:41 -07:00
Yaowu Xu	f809475c73	Merge "Make iscan and scan neighbor arrays static const."	2014-10-02 15:15:58 -07:00
Yaowu Xu	9712bc691d	Make iscan and scan neighbor arrays static const. This commit changes the tables to be read only, which fixes issue #866 Change-Id: I85bbe03f9d344f50570f8c1c61699bdc5cee248f	2014-10-02 14:08:14 -07:00
Alex Converse	a0befb93e7	Fix subsampling check for images 1 pixel wide/tall Change-Id: I0e262ede7eb4a4ae0c86181922d744e542e93350	2014-10-02 11:02:57 -07:00
Deb Mukherjee	35e8fa1458	rdmult data type change to fix high bit-depth Fixes an intermittent assert failure for highbitdepth. Change-Id: If8cad0209a94f1184b69c7b3f1d587934f857d9b	2014-10-02 07:37:26 -07:00
Jingning Han	9641b1b9ac	Merge "Remove unused header files from vp9_encodemb.h"	2014-10-01 17:05:25 -07:00
Deb Mukherjee	30fbf23fda	Merge "High-bitdepth bugfixes"	2014-10-01 16:47:43 -07:00
Yunqing Wang	e350e3fe68	Merge "Modify block transform skipping check"	2014-10-01 16:19:56 -07:00
Jingning Han	72a78a0c40	Remove unused header files from vp9_encodemb.h Change-Id: Icfc3fb62cc0b05e435814035bfe1f2e2870442b4	2014-10-01 14:50:24 -07:00
Deb Mukherjee	a160d72522	High-bitdepth bugfixes Miscellaneous bug-fixes for high bitdepth functionality. With this patch, high bit-depth profiles become mostly functional, except for an intermittent assert failure issue that is being tracked. Change-Id: I6a7fcbdcf1e5b09842e88535f8442d2e1230748c	2014-10-01 14:18:11 -07:00
Jingning Han	0a9f5fa146	Remove repeated header files from vp9_block.h This commit removes unused header file vp9_onyxc_int.h and repeatedly included file vpx_ports/mem.h from vp9_block.h Change-Id: I400b210bd1da48f1880bd50a8f4a6e2c690e15a1	2014-10-01 13:01:43 -07:00
Yunqing Wang	e4aac6bb61	Modify block transform skipping check Block transform skipping was implemented based on DCT's energy conservation property. Modified the thresholds using zero bin parameters. AC and DC coefficients were checked separately to allow better identifying of skippable blocks. Borg test at speed 3 showed: stdhd set: psnr gain: 0.153%, ssim gain: 0.051%; derf set: psnr gain: 0.023%, ssim gain: 0.036% For most test clips, the encoding speedup is 1% - 2%. parkrun(720p): 7.5% speedup, park_joy(1080p): 3.5% speedup. Change-Id: If28eb81113a077414f5ca7b021c14f9069b373bb	2014-10-01 12:58:09 -07:00
Jingning Han	20a37391d9	Merge "Conditionally skip reference frame check"	2014-10-01 11:19:10 -07:00
Jingning Han	891793a540	Conditionally skip reference frame check For regular inter frames, if the distance from GOLDEN_FRAME is larger than 2 and if the predicted motion vector of LAST_FRAME gives lower sse than that of GOLDEN_FRAME, skip the GOLDE_FRAME mode checking in the rate-distortion optimization. It provides about 5% speed-up at expense of -0.137% and -0.230% performance down for speed 3. Local experiment results: pedestrian 1080p 2000 kbps 66712 b/f, 40.908 dB, 113688 ms -> 66768 b/f, 40.911 dB, 108752 ms blue_sky 1080p 2000 kbps 51054 b/f, 35.894 dB, 70406 ms -> 51051 b/f, 35.891 dB, 67236 ms old_town_cross 720p 1500 kbps 14412 b/f, 36.252 dB, 60690 ms -> 14431 b/f, 36.249 dB, 57346 ms Change-Id: Idfcafe7f63da7a4896602fc60bd7093f0f0d82ca	2014-10-01 08:32:15 -07:00
Yunqing Wang	b1b6fd85db	Merge "Skip the partition search for still frames"	2014-09-30 11:59:05 -07:00
Yunqing Wang	c8d01b1eaf	Merge "Refactor encode_rd_sb_row function"	2014-09-30 11:58:39 -07:00
Deb Mukherjee	40479dfe92	Misc. high-bit-depth fixes Change-Id: Ie9fb6a4078eb6a3fb7c4ff1453831ab9afe23121	2014-09-30 10:37:53 -07:00
Deb Mukherjee	63e49be340	Merge "Adds two new subpel search methods"	2014-09-29 20:11:04 -07:00
JackyChen	7ba646f7e6	Fix a bug in calculating delta in VP9 denoiser. When calculating delta in VP8 denoiser, since the block size is fixed to 16x16, the divisor is 256, which is the number of the pixel. But in VP9, the block size varies, the divisor should correspond to the block size. Change-Id: Ibdc1e5d23ba8c788b0d0dc6d406bcdfc34c1b142	2014-09-29 13:09:18 -07:00
Deb Mukherjee	4e9c0d2ad4	Adds two new subpel search methods One is a more aggressive version of the pruned subpel tree search where only a single halfpel candidate is searched. The search candidate is based on a surface fit result. The other is a method to obtain the subpel position at one shot based on the same surface fit. The methods have not been deployed in any speed setting yet. Change-Id: I34fef3f2e34f11396c9d1ba97f4be8c4ffca62d3	2014-09-29 12:51:20 -07:00
Jingning Han	8b4dd536a5	Merge "Skip certain ALTREF inter modes in ARF coding"	2014-09-29 10:43:45 -07:00
Deb Mukherjee	d4713f1d50	Fix a bug introduced in a previous patch on highbd Change-Id: Ice692334f75157446a44a6e81503cada977934f4	2014-09-26 15:43:55 -07:00
Jingning Han	ccdb518ff8	Skip certain ALTREF inter modes in ARF coding This commit enables the encoder to skip checking ALTREF inter modes in ARF coding, if the predicted motion vectors suggest that the GOLDEN_FRAME provides higher prediction accuracy than ALTREF_FRAME. It improves the speed 3 encoding speed by about 5%, at the expense of compression performance loss -0.041% and -0.225% for derf and stdhd, respectively. pedestrian_area 1080p 2000 kbps 66705 b/f, 40.909 dB, 118738 ms -> 66732 b/f, 40.908 dB, 113688 ms old_town_cross 720p 1500 kbps 14427 b/f, 36.256 dB, 62746 ms -> 14412 b/f, 36.252 dB, 60690 ms blue_sky 1080p 1500 kbps 51026 b/f, 35.897 dB, 73310 ms -> 50921 b/f, 35.893 dB, 70406 ms bus CIF 1000 kbps 21301 b/f, 34.841 dB, 7326 ms -> 21248 b/f, 34.837 dB, 7196 ms Change-Id: I76cf88b4d655e1ee3c0cb03c8a5745493040e8d2	2014-09-26 12:53:43 -07:00
Paul Wilkins	d3bbd87d5e	Two pass rc changes. Adjustments to the GF interval choice and minimum boost. Change-Id: I29951621484e1ee339adfb73ab430aa65f310ad8	2014-09-26 17:13:02 +01:00
Yunqing Wang	1fcbf6ed56	Skip the partition search for still frames This patch re-enabled the feature in Pengchong's patch (commit `1286126073`). Originally, it was turned on while use_lastframe_partitioning > 0(not used anymore). Now it was added as a feature, and turned on while speed >= 2. As described in the original patch, this feature helps speed up the slideshows in YouTube. Change-Id: I1b0f18d65da1ee1c8d1e117dabba910c5207c471	2014-09-26 09:03:52 -07:00
Deb Mukherjee	993d10a217	Adds various high bit-depth encode functions Change-Id: I6f67b171022bbc8199c6d674190b57f6bab1b62f	2014-09-25 01:50:36 -07:00
Jingning Han	6989e81d61	Remove unused variable in handle_inter_mode Change-Id: Id757d2c940756ce1b0ead2ea24af9ac0a493de05	2014-09-24 18:27:44 -07:00
Paul Wilkins	76035d16d9	Merge "Fix build issue with stats enabled."	2014-09-24 10:32:37 -07:00
Yunqing Wang	14ee2805a3	Refactor encode_rd_sb_row function Simplified the code and removed some code that was not used anymore. This patch didn't change encoding result. Change-Id: I7e54a74c8f35a6726dfc8a1c55b337448b7ea124	2014-09-24 10:24:18 -07:00
Paul Wilkins	5b724fc78e	Fix build issue with stats enabled. Compiler build issue when output stats enabled. Change-Id: I7b5409108f3f27ba61b0241b9340b412683eff45	2014-09-24 11:48:58 +01:00
Deb Mukherjee	e1d3c36525	Adds high bit-depth frame resize functions Change-Id: I35b015a759325d72d0da427c61a09f19f8e69697	2014-09-23 22:55:33 -07:00
Yaowu Xu	8751e49a6f	Merge "Adapt mode based rd_threshold for similar block size"	2014-09-23 22:28:08 -07:00
Yaowu Xu	60737c9fc8	Merge "Fix an IOC"	2014-09-23 20:44:35 -07:00
Deb Mukherjee	4109372af3	Adds high bit-depth psnr/sse functions Also adds some miscellaneous high bit-depth setup functions. Change-Id: I66488b08a5a2a8cb9518ca10497cf1c1501ceded	2014-09-23 17:28:05 -07:00
Deb Mukherjee	e2a90c0b21	Merge "High bit-depth loop/arf/postproc filter functions"	2014-09-23 17:26:32 -07:00
Deb Mukherjee	6c6213d960	Merge "Pruned subpel search for speed 3."	2014-09-23 17:12:03 -07:00
Deb Mukherjee	931ed516ba	High bit-depth loop/arf/postproc filter functions Adds high-bitdepth loopfilter, temporal filter and postproc functions Change-Id: I81c8a9176890784686bc4f2af0d550d243b3b2d3	2014-09-23 16:20:43 -07:00
Yaowu Xu	4a101310e8	Adapt mode based rd_threshold for similar block size The rd_thresholds are adaptively changed based on best mode tested. It was only changed for the same block size, this commit makes the adaptation for similar block sizes too. The commit also made minor adjustment and code cleanups. The impact on encoding time for _ped: 118089 ms -> 111927 ms The impact on compression: derf: -0.339% stdhd: -0.303% Change-Id: I8817fed1102350497f2ec631849e43f753878e5d	2014-09-23 16:10:59 -07:00
Yaowu Xu	56032b471d	Fix an IOC Change-Id: I0ca6746696d81657c035b0f6523c9af370da3c95	2014-09-23 16:07:22 -07:00
Deb Mukherjee	c94b17f4b2	Pruned subpel search for speed 3. Adds code to return an integer cost list for NSTEP search. Then uses it for pruned subpel search in speed 3. derf: -0.06% Speed on mobcal 720p increaes from 10.28 fps to 10.65 fps. [Subject to further testing]. Change-Id: Ib591382d25b2c11bcaba9d3a27a93a9d1ab27a96	2014-09-23 11:27:58 -07:00
Yaowu Xu	7feede9869	Merge "Remove code duplication"	2014-09-22 17:13:59 -07:00
Yaowu Xu	052bc8ea6a	Merge "Simplify rd_pick_intra_sby_mode()"	2014-09-22 17:13:55 -07:00
Yaowu Xu	c7ab18fe56	Remove code duplication Change-Id: I453b3e0d946951665d5919248445fc4f3222d2ad	2014-09-22 15:22:51 -07:00
Yaowu Xu	f46326c7a2	Simplify rd_pick_intra_sby_mode() Change-Id: Ifb0915c94c2db48827ddbd446314cb6e3155b99c	2014-09-22 14:58:51 -07:00
Minghai Shang	38b6aed8fd	Merge "[spatial svc] Remove vpx_svc_parameters_t and the loop that sets it for each layer"	2014-09-22 14:01:24 -07:00
Jingning Han	f7023ea014	Remove unnecessary local variable declaration This commit removes a repetitive local variable declaration in vp9_rd_pick_inter_mode_sb. Change-Id: I1b0afa98ff1ecbfb46e17d3d1cee95d32c4309db	2014-09-22 09:29:28 -07:00
Jingning Han	eee904c9b9	Adaptive mode search scheduling This commit enables an adaptive mode search order scheduling scheme in the rate-distortion optimization. It changes the compression performance by -0.433% and -0.420% for derf and stdhd respectively. It provides speed improvement for speed 3: bus CIF 1000 kbps 24590 b/f, 35.513 dB, 7864 ms -> 24696 b/f, 35.491 dB, 7408 ms (6% speed-up) stockholm 720p 1000 kbps 8983 b/f, 35.078 dB, 65698 ms -> 8962 b/f, 35.054 dB, 60298 ms (8%) old_town_cross 720p 1000 kbps 11804 b/f, 35.666 dB, 62492 ms -> 11778 b/f, 35.609 dB, 56040 ms (10%) blue_sky 1080p 1500 kbps 57173 b/f, 36.179 dB, 77879 ms -> 57199 b/f, 36.131 dB, 69821 ms (10%) pedestrian_area 1080p 2000 kbps 74241 b/f, 41.105 dB, 144031 ms -> 74271 b/f, 41.091 dB, 133614 ms (8%) Change-Id: Iaad28cbc99399030fc5f9951eb5aa7fa633f320e	2014-09-22 09:28:16 -07:00
hkuang	c70cea97ac	Remove mi_grid_* structures. mi_grid_* are arrays of pointer to pointer. They save the pointers that point to the MIs in cm->mi. But they are unnecessary and complicated. The original goal was to remove MODE_INFO_t copy. But with an extra MODE_INFO_t pointer inside MODE_INFO_t, same goal could be achieved. This commit totally removes the mi_grid_* structures. But there are still many dummy MODE_INFO_t inside cm->mi which are a waste of memory. Next commit will do on-demand MODE_INFO_t allocation in order to save these memories. Change-Id: I3a05cf1610679fed26e0b2eadd315a9ae91afdd6	2014-09-19 21:27:11 -07:00
Deb Mukherjee	822b51609b	High bit-depth coefficient coding functions Tokenization and Detokenization enhancements for 10/12 bit Change-Id: I3c269ec30f8eb160ee024905638a193975237559	2014-09-19 15:21:24 -07:00
Minghai Shang	209ee12110	[spatial svc] Remove vpx_svc_parameters_t and the loop that sets it for each layer vpx_svc_parameters_t contains id, resolution and min/max qp for each spatial layer. In this change we will use extra config to send min/max qp and scaling factors, then calculate layer resolution inside encoder. Change-Id: Ib673303266605fe803c3b067284aae5f7a25514a	2014-09-18 18:05:07 -07:00
Minghai Shang	f66be91f61	Merge "[spatial svc] Use same golden frame for all temporal layers"	2014-09-18 12:29:40 -07:00
Minghai Shang	f780b16bb8	[spatial svc] Use same golden frame for all temporal layers Overhead goes down from 8% to 3% for 1080 60p Change-Id: Idf3e5ca8712402a914a8cb79df17d3cdab63b163	2014-09-18 11:16:29 -07:00
Deb Mukherjee	6d0ee9860e	Merge "Adds high bitdepth convolve, interpred & scaling"	2014-09-18 10:52:23 -07:00
Deb Mukherjee	0d3c3d3ce7	Adds high bitdepth convolve, interpred & scaling Change-Id: Ie51c352a6b250547207cbc1ebba833a01ed053e3	2014-09-18 07:26:17 -07:00
Paul Wilkins	c389b37bb4	Substantial reworking of code for arf and kf groups. Substantial restructuring of the way we estimate the rate of decay in prediction quality and determine the arf interval and amount of boost used. Also other changes to support moving to a lower first pass Q which exposes some new features and allows us to better distinguish genuinely static blocks from low motion or noisy blocks. Net gains now visible on all the test sets with std-hd PSNR up 1.87%. There are still some bad outlier cases but most of these are low motion or slide show type content where the metrics are already high at any given rate. The best + case is up by more than 10%. Change-Id: I18e25170053bdf3188f493ff8062f48a74515815	2014-09-18 12:53:48 +01:00
Deb Mukherjee	5cd0aab81a	Adds high bitdepth quantization functions Adds various high bitdepth quantization functions. Change-Id: I36fc0bf75a1bd15128ed271df8723de0ac134b0c	2014-09-16 14:55:37 -07:00
Jingning Han	66f812fb56	Merge "Use non-zero mode threshold for NEARESTMV modes"	2014-09-16 13:39:54 -07:00
Adrian Grange	2b3b63f422	Merge "Fix ARF construction when scaling"	2014-09-16 12:35:23 -07:00
Adrian Grange	99df7ded95	Merge "Move call to vp9_rc_get_second_pass_params()."	2014-09-16 11:37:33 -07:00
Adrian Grange	1def634f1a	Fix ARF construction when scaling The ARF frame should always be the same size as the native resolution of the input frames. It will be scaled to the required resolution at encode time. Change-Id: I0afe858129aa6ef65b1648f43476331715346896	2014-09-16 11:12:49 -07:00
Jingning Han	56fa3ab886	Use non-zero mode threshold for NEARESTMV modes This commit makes the encoder to use non-zero mode threshold for NEARESTMV modes. The runtime for test clips of speed 3 is reduced by about 1%. pedestrian 1080p 2000 kbps, 143239 ms -> 141989 ms bus CIF 1000 kbps, 7835 ms -> 7749 ms The compression performance change is about -0.02% for both derf and stdhd. Change-Id: Ib71808922c41ae2997100cb7c561f68dcebfa08e	2014-09-16 09:56:10 -07:00
Jingning Han	ffaebfc7b4	Merge "Add ARF validation for compound inter mode check"	2014-09-15 21:26:37 -07:00
Jingning Han	c50256c157	Merge "Remove redundant reference frame check in sub8x8 RD search"	2014-09-15 21:26:11 -07:00
Jingning Han	fe96932c69	Merge "Replace best_ref_index table fetch with best_mbmode"	2014-09-15 21:25:48 -07:00
Yunqing Wang	57eb2a4e83	Merge "Simplify the skip flag cost code"	2014-09-15 18:50:30 -07:00
Yunqing Wang	c60ef810a1	Merge "Set the skip flag to 1 for skippable blocks"	2014-09-15 18:50:19 -07:00
Yunqing Wang	200ec69abb	Simplify the skip flag cost code Code refactoring. Change-Id: Idad53cb80497d13551a142a642f7529fc305b0bc	2014-09-15 17:11:16 -07:00
Yunqing Wang	46aed7b8d0	Set the skip flag to 1 for skippable blocks If the partition block is skippable, which means no coefficients for Y, U, and V planes, its skip flag is set to 1. No quality change (verified by borg tests), and no noticeable speed change. Change-Id: I9231f720f8dd6364384cf05aa148ca24d75450f1	2014-09-15 16:50:19 -07:00
Jingning Han	f897dd5f09	Merge "Fix format in vp9_rd_pick_inter_mode_sub8x8"	2014-09-15 15:34:22 -07:00
Jingning Han	f1581b3b2e	Add ARF validation for compound inter mode check This commit enforces ARF validation check for compound inter modes. It avoids potential access to ARF in the encoding process if it is not allowed. Change-Id: I055fec946b5d19d97937dc9001e1e564923e2439	2014-09-15 12:20:57 -07:00
Jingning Han	252822e81c	Remove redundant reference frame check in sub8x8 RD search The valid reference frame check in sub8x8 rate-distortion optimization search has been included in the ref_frame_skip_mask scheme. This commit removes the later further validation checks that are not in effect. Change-Id: I853b477c44037d3dc0afec6cbfce08a96c597a75	2014-09-15 12:20:04 -07:00
Jingning Han	cc00eea676	Replace best_ref_index table fetch with best_mbmode This commit replaces the best_ref_index table fetch with the use of best_mbmode in vp9_rd_pick_inter_mode_sub8x8. Change-Id: I882ee9ee6a8c0e61befcca1f4dba6d2ea8de8f13	2014-09-15 09:59:20 -07:00
Jingning Han	73805bfa70	Fix format in vp9_rd_pick_inter_mode_sub8x8 Change-Id: I9b6a74bdf003b39235f14f8b5b7f3b861f6bf131	2014-09-15 09:44:09 -07:00
Yunqing Wang	10a9456ade	Merge "Refactor encode_superblock function"	2014-09-15 09:28:31 -07:00
Paul Wilkins	cd95543ee4	Move call to vp9_rc_get_second_pass_params(). Call to vp9_rc_get_second_pass_params() moved from Pass2Encode() to earlier in vp9_get_compressed_data(), to ensure that two pass stats and parameters are available before decisions such as frame scaling. Change-Id: If21537f0073919b04696a7d5e9aac78e23d76f39	2014-09-15 12:45:42 +01:00
Jingning Han	95f67f09ac	Merge "Remove redundant reference frame threshold settings"	2014-09-13 10:44:00 -07:00
Jingning Han	59dd83a3ea	Merge "Refactor reference frame control in sub8x8 block RD search"	2014-09-13 10:43:36 -07:00
Jingning Han	e6d927343e	Merge "Format fixes in vp9_rd_pick_inter_mode_sb"	2014-09-13 10:43:24 -07:00
Jingning Han	ad3c92b9b7	Merge "Remove unused best_inter_rd variable"	2014-09-13 10:43:14 -07:00
Jingning Han	f02e0b6cf6	Merge "Remove unused speed feature"	2014-09-13 10:43:03 -07:00
Deb Mukherjee	c0dfecfb89	Merge "Use bigdia search with pruned subpel search"	2014-09-12 16:42:18 -07:00
Yunqing Wang	1bf0beb5fc	Refactor encode_superblock function The code covers both x->skip=0 & x->skip=1 cases. Change-Id: I09745c10e5994dc700ae4c01b4b62979cdaf3306	2014-09-12 15:58:17 -07:00
Jingning Han	888a848453	Remove redundant reference frame threshold settings When a reference frame type is not in the frame buffer, the mode search threshold will be set to INT_MAX, so as to effectively turn off the mode entries in the rate-distortion optimization loop that involves this reference frame type. This operation is now integrated in the ref_frame_skip_mask scheme. This commit hence removes the redundant mode search threshold setting. Change-Id: Ib18f45da611afda2af275201efd367df7f5101ab	2014-09-12 14:36:51 -07:00
Jingning Han	adb20849b6	Refactor reference frame control in sub8x8 block RD search This commit unifies the reference frame control in the rate- distortion optimization search loop of sub8x8 block size to remove the control dependency on mode search order. Change-Id: I3a174099f71a7cc176ede9fd60e2374243ae9232	2014-09-12 11:03:03 -07:00
Minghai Shang	3e7b04af54	Merge "[spatial svc] Output psnr for all layers in one packet."	2014-09-12 10:52:42 -07:00
Deb Mukherjee	83c76118eb	Use bigdia search with pruned subpel search Improves function to return sad of integer pels by reusing integer pels already visited in the smallest scale. Turns on BIGDIA search for speed 4. Also, turns on the first version of the pruned subpel search at this speed. derf: -0.32% (speed 4) Speed seems to improve by at least 5% but subject to verification. Change-Id: Iaec8eaffd61d6237ac029e6a2a1b0a88b2a35271	2014-09-12 10:25:12 -07:00
Jingning Han	7f77a1c3c9	Merge "Unify intra mode mask into mode_skip_mask scheme"	2014-09-12 09:06:35 -07:00
Deb Mukherjee	10783d4f3a	Adds high bitdepth transform functions and tests Adds various high bitdepth transform functions and tests. Much of the changes are related to using typedefs tran_low_t and tran_high_t for the final transform cofficients and intermediate stages of the transform computation respectively rather than fixed types int16_t/int. When vp9_highbitdepth configure flag is off, these map tp int16_t/int32_t, but when the flag is on, they map to int32_t/int64_t to make space for needed extra precision. Change-Id: I3c56de79e15b904d6f655b62ffae170729befdd8	2014-09-11 19:56:33 -07:00
Deb Mukherjee	1e4136d35d	Adds high bit depth sad and variance functions Moves high bit depth sad/var functions from highbitdepth branch to master. Change-Id: If03845d8ef9c9c494e13350e7a587c289306b94d	2014-09-11 17:30:44 -07:00
Jingning Han	74ddde01c0	Format fixes in vp9_rd_pick_inter_mode_sb Change-Id: Ie45687405dcaa34ba465dce2aa14f76017d3a794	2014-09-11 17:15:15 -07:00
Minghai Shang	e3fff31aff	[spatial svc] Output psnr for all layers in one packet. Change-Id: I97d0cf095e9cfefdfa0f65eb5e96d6848cc9ffca	2014-09-11 16:21:35 -07:00
Jingning Han	8e3f7a52a1	Remove unused best_inter_rd variable The variable best_inter_rd is effectively not in use in the rate- distortion mode search loops of both regular block sizes and sub8x8 block sizes. Change-Id: I178f909f8c9629772e13adc6257908653b2adf31	2014-09-11 16:16:26 -07:00
Jingning Han	00fe92c22f	Remove unused speed feature The speed feature that skips compound inter prediction modes was subsumed by other speed features and effectively was not in use. This commit removes it. Change-Id: I22b0c71a8ddd15d93b25d86fa63a1dce2ba6a1a9	2014-09-11 15:54:53 -07:00
James Zern	d555dfdc69	Merge "vp9_picklpf: search_filter_level: remove filt_err"	2014-09-11 15:39:19 -07:00
Jingning Han	71b4bee33f	Merge "Remove inter_mode_mask from rate-distortion search loop"	2014-09-11 12:08:13 -07:00
James Zern	49d7abc0ec	vp9_picklpf: search_filter_level: remove filt_err inspect ss_err[] directly, removes an unnecessary assignment Change-Id: I14db5e8e567e7e541a57fce73389ffe7651d5614	2014-09-11 11:37:56 -07:00
Jingning Han	0cf599b573	Merge "Move intra block size skip outside mode search loop"	2014-09-11 11:15:35 -07:00
Jingning Han	387ec881d3	Merge "Fix format in vp9_rd_pick_inter_mode_sub8x8"	2014-09-11 11:15:25 -07:00
Jingning Han	3556ab56f6	Merge "Move overlay frame speed feature setting out of mode search loop"	2014-09-11 11:14:37 -07:00
Jingning Han	82757250d6	Merge "Refactor to remove speed feature dependency on mode search order"	2014-09-11 11:14:26 -07:00
Jingning Han	bdd8eb6fcc	Unify intra mode mask into mode_skip_mask scheme Integrate intra mode mask speed feature with the mode_skip_mask scheme. Move it outside the mode search loop in the vp9_rd_pick_inter_mode_sb function. Change-Id: I7738fea749bfdc08ad05d7f2524feb8ff67568d9	2014-09-11 10:36:48 -07:00
Minghai Shang	fb754540e9	Merge "[spatial svc]Add golden frame to first pass rate control"	2014-09-11 10:24:28 -07:00
Jingning Han	8cefed1568	Remove inter_mode_mask from rate-distortion search loop This speed feature is used in real-time setting only. Remove the related condition check in the rate-distortion optimization search loop. Change-Id: Iaacc1e268214634e6f95c5048c28a60cec6c42fc	2014-09-11 10:18:55 -07:00
Jingning Han	238b2ace86	Move intra block size skip outside mode search loop Unify this speed feature in the ref_frame_skip_mask scheme. Change-Id: I7ea5646da02d3ea643680c22d50dabd448d55a27	2014-09-11 09:54:19 -07:00
Jingning Han	8b06a24ce7	Fix format in vp9_rd_pick_inter_mode_sub8x8 Change-Id: I0da29c858c6c1eb5ef07cee8f599329f5a002da9	2014-09-11 09:28:47 -07:00
Jingning Han	8d42fad9c1	Move overlay frame speed feature setting out of mode search loop Refactor overlay frame speed-up related function. Make it unified with the ref_frame_skip_mask system and Move it out of the mode search loop. Change-Id: I0dde9baf44354f6ba00b4679cba02fa6a30c7316	2014-09-10 19:44:58 -07:00
JackyChen	d9050af683	Merge "Fix the bug which made VP8 denoiser not bit-exact between C code and SSE code."	2014-09-10 18:08:59 -07:00
Minghai Shang	0a0ccf669b	[spatial svc]Add golden frame to first pass rate control Change-Id: If3035f0e7dfcfe88c4bbf4eec66761e070476df0	2014-09-10 17:35:02 -07:00
Jingning Han	f9f0879756	Refactor to remove speed feature dependency on mode search order This commit refactor the rate-distortion optimization search for regular block sizes to remove the speed feature dependency on mode search order. Change-Id: Ied033ee484c2957e17baa7b6450b720fe7dd0e7d	2014-09-10 17:09:14 -07:00
JackyChen	47380c3350	Fix the bug which made VP8 denoiser not bit-exact between C code and SSE code. This issue is found when the denoising mode is set to kDenoiserOnYUVAggressive. Updated the C code to make it the same with SSE version. I also changed several lines in VP9 denoiser for the code style. Change-Id: I640d48cf946fe8c6a400e6e252107501d1e226d3	2014-09-10 16:18:43 -07:00
Jingning Han	6facdfdd7d	Merge "Fix a bug in vp9_rd_pick_inter_mode_sb"	2014-09-10 14:46:19 -07:00
Jingning Han	5ac97d101d	Merge "Remove redundant ref frame pointer assignment"	2014-09-10 14:46:11 -07:00
Jingning Han	68d79146ea	Fix a bug in vp9_rd_pick_inter_mode_sb This commit fixes a bug related to skipping intra mode checking, by using a separate variable to store the best prediction error from inter mode. It avoids unintentionally overwriting intra mode rate-distortion cost, and hence affecting other speed features. Change-Id: I99e12993339c84c8b4f597996b372012e5858fae	2014-09-09 15:39:54 -07:00
Jingning Han	9a9e2aef09	Remove redundant ref frame pointer assignment Assigning selected reference frame pointer is done in the encode_superblock function. No need to do this at the end of rate-distortion optimization search. Change-Id: I33fcede0fd304b4a4c4deef2d126d79546a9c070	2014-09-09 15:15:11 -07:00
Jingning Han	89ffda0ddf	Merge "Remove dependency of intra mode search skip check on mode order"	2014-09-09 14:10:28 -07:00
Jingning Han	1614306913	Merge "Replace best_mode_index table retrieve with fetching best_mbmode"	2014-09-09 14:10:19 -07:00
Jingning Han	33593d1f03	Remove dependency of intra mode search skip check on mode order This commit refactors the vp9_rd_pick_inter_mode_sb function to remove the intra mode early termination dependency on the mode search order. Change-Id: If6ac49aa7c530c7b9a5bd31b0ab84db83e192bec	2014-09-09 12:30:47 -07:00
Jingning Han	d96228a07c	Replace best_mode_index table retrieve with fetching best_mbmode This commit allows the encoder to find current best prediction mode state using best_mbmode, instead of fetching from the static mode search table via best_mode_index. Change-Id: Ibefeab83aed33a49c2be03e83f09153856ca4271	2014-09-09 11:58:10 -07:00
Yunqing Wang	f10d7eeda2	Remove the use of use_lastframe_partitioning at speed 4 The use of use_lastframe_partitioning is totally removed in good- quality encoding. Its usage in real-time encoding needs to be evaluated to see if it can be removed too. The Borg tests at speed 4 showed: stdhd set: 0.220% psnr gain, 0.166% ssim gain; derf set: 0.329% psnr gain, 0.476% ssim gain. Speed test on selected clips showed 1.54% speedup.(Worst case: pedestrian_area_1080p25.y4m, speed loss: 1.5%) Change-Id: I1c844d329b0b5678558439b887297c1be7ddab00	2014-09-09 10:54:07 -07:00
James Zern	3610d0a3e0	Merge changes I4c74dcab,Ifbfc1422,I2450b485,Ibdb07f6d,I3737772f,Ic3be55ed * changes: vp9_pick_inter_mode: normalize some types vp9_pick_inter_mode: cosmetics: localize var. defs vp9_pick_inter_mode: cosmetics: add const vp9_pick_inter_mode: cosmetics: fix indent vp9_pickmode: move PRED_BUFFER definition to .c vp9_pickmode: make vp9_pick_inter_mode() void	2014-09-08 19:19:31 -07:00
Yaowu Xu	b73c9df1a4	Merge "No longer use use_lastframe_partitioning speed feature"	2014-09-08 18:10:20 -07:00
Paul Wilkins	f24054574d	Fix VS build issue. Compile fails when CONFIG_INTERNAL_STATS flag is set. Change-Id: Iba7701c058169ca3fc0b9008619ac55a1fe1a8b6	2014-09-08 15:29:33 -07:00
Jingning Han	a61973bf29	Merge "Enable adaptive motion search for ARF coding"	2014-09-08 08:51:05 -07:00
Dmitry Kovalev	1f19ebbab6	Replacing vp9_get_mb_ss_sse2 asm implementation with intrinsics. Change-Id: Ib4f5dd733eb2939b108070a01e83da5d9990bac0	2014-09-06 00:10:25 -07:00
James Zern	49bb8fbaca	vp9_pick_inter_mode: normalize some types Change-Id: I4c74dcab6358817f03d3bc4d526006d241f0c10e	2014-09-05 19:22:54 -07:00
James Zern	7fe86bba2e	vp9_pick_inter_mode: cosmetics: localize var. defs Change-Id: Ifbfc142291697a1847ef85ced0b0eb4d6dab161e	2014-09-05 19:22:54 -07:00
James Zern	6f094e2a71	vp9_pick_inter_mode: cosmetics: add const Change-Id: I2450b4856e48dbc4d5b938b2edcea0704f756c8e	2014-09-05 19:22:53 -07:00
James Zern	0adfacad75	vp9_pick_inter_mode: cosmetics: fix indent + delete a dead comment Change-Id: Ibdb07f6dbdb30fc7888f6115ddc326fcec1157a7	2014-09-05 19:22:53 -07:00
James Zern	5ed806a608	vp9_pickmode: move PRED_BUFFER definition to .c Change-Id: I3737772fe53f9885c82e2ac4c1af478ab951c16c	2014-09-05 19:22:53 -07:00
James Zern	94968c6d14	vp9_pickmode: make vp9_pick_inter_mode() void the previous return value was constant and unused. Change-Id: Ic3be55edb4a884448c7bb07977a80dfb58b7b940	2014-09-05 19:22:53 -07:00
Yunqing Wang	1092140379	No longer use use_lastframe_partitioning speed feature The speedup in rd_pick_partition() function makes it possible to drop use_lastframe_partitioning feature. By doing that, we achieve good PSNR gain with small speed loss. Also, this makes encoding loop less complicated. The code cleanup patch will follow. Borg tests showed: 1. At speed 2, stdhd set: 0.201% PSNR gain, 0.133% SSIM gain; derf set: 0.262% PSNR gain, 0.276% SSIM gain. 2. At speed 3, stdhd set: 0.139% PSNR gain, 0.109% SSIM gain; derf set: 0.447% PSNR gain, 0.442% SSIM gain. The average speed loss over selected test clips is within 1% with the worst case of 4%. Change-Id: Icfd2ded7869372b585a6972855d933b3d0280d90	2014-09-05 16:24:41 -07:00
Yunqing Wang	ebac8f3487	Merge "Correct the mode decisions in special cases"	2014-09-05 13:45:41 -07:00
Dmitry Kovalev	54bec0971f	Merge "Initializing intra modes without vpx_once()."	2014-09-05 12:03:36 -07:00
Yunqing Wang	1dd9a63929	Correct the mode decisions in special cases The rate costs calculated for inter modes are not precise in some cases, which causes NEWMV is chosen instead of NEARESTMV, NEARMV, and ZEROMV. This patch added checks for these cases, and corrected the mode decisions. Borg tests at speed 3 showed: 1. stdhd set: 0.102% PSNR gain and 0.088% SSIM gain. 2. derf set: 0.147% PSNR gain and 0.132% SSIM gain. No speed change. Change-Id: I35d17684b89ad4734fb610942d707899146426db	2014-09-05 12:01:07 -07:00
JackyChen	f8e5105b47	Merge "Map motion magnitude in VP9 denoiser."	2014-09-04 16:59:53 -07:00
Jingning Han	d435148fe6	Enable adaptive motion search for ARF coding This commit turns on adaptive motion search for ARF coding, in addition to other normal inter frame coding. It improves the average compression efficiency: stdhd 0.1% derf 0.04% For the test sequences, the speed 3 runtime is reduced: pedestrian 1080p 2000 kbps, 149932 ms -> 144580 ms, (3.3% speed-up) bus CIF 1000 kbps, 8050 ms -> 7895 ms, (1.9%) highway CIF 100 bkps, 45033 ms -> 44078 ms, (2.2%) Change-Id: I5228565b609f99e8ae04f6140a2bf2b64a831d21	2014-09-04 16:26:40 -07:00
Jingning Han	3de038f396	Merge "Speed up compound inter prediction mode check"	2014-09-04 16:09:07 -07:00
JackyChen	b869b970c1	Merge "Update the condition when COPY_BLOCK is chosen."	2014-09-04 15:48:22 -07:00
Dmitry Kovalev	46b83391e2	Merge "Removing local set_speed_features() function."	2014-09-04 15:36:52 -07:00
JackyChen	b1153f34d4	Map motion magnitude in VP9 denoiser. This is to keep the same with VP8 denoiser. If motion magnitude is small, make denoiser more aggressive. Change-Id: I942a6e2f2ed9aec6f0c4c1f9e5fa47066cadcc0c	2014-09-04 14:53:33 -07:00
JackyChen	d75266f141	Update the condition when COPY_BLOCK is chosen. The change is just to keep the condition the same with VP8. Change-Id: I9662b40996126605945dd853c0cbe8916c1ce578	2014-09-04 14:28:12 -07:00
JackyChen	7ba600dc89	Merge "Fix a bug in VP9 denoiser."	2014-09-04 14:16:26 -07:00
JackyChen	e30f7698f5	Fix a bug in VP9 denoiser. When the first try of denoising turns out to be too much, we will use a softer filter by adopting an adjustment to make the result closer to original pixel (as in VP8 denoiser). The old code made the adjustment in the wrong direction. Change-Id: I84e28fa9e01eef47c5a37d5a2e6d3d378a06786b	2014-09-04 11:46:36 -07:00
Dmitry Kovalev	48197f0a70	Adding sse2 variant for vp9_mse{8x8, 8x16, 16x8}. Change-Id: I6786d25ce4f32b8d8912f2d239a45ca15b310c4b	2014-09-03 19:02:14 -07:00
Dmitry Kovalev	ab73dba65f	Merge "Replacing asm 16x16 variance calculation with intrinsics."	2014-09-03 18:57:33 -07:00
Dmitry Kovalev	406404af63	Merge "Small cleanup: reusing existing code."	2014-09-03 18:57:25 -07:00
Jingning Han	d62d804e64	Speed up compound inter prediction mode check This commit allows the encoder to store outcomes of single reference frame modes and compares them to decide if the inter prediction filter, forward transform, and quantization can be skipped. The compression performance of speed 3 is down derf -0.364% stdhd -0.198% For test sequences, the speed 3 runtime is reduced highway CIF 100 kbps, 51976 ms -> 45033 ms, 13% speed-up stockholm 720p 1000 kbps, 71826 ms -> 67838 ms, 5.5% speed-up pedestrian 1080p 2000 kbps, 154924 ms -> 150702 ms, 2.6% speed-up Change-Id: I5aa26f918d2b4b5197a2c0afa2779319f1c88e44	2014-09-03 15:28:01 -07:00
Yaowu Xu	7ab5de04fd	Merge "Change last_partition_redo_frequency for speed 3"	2014-09-03 14:57:02 -07:00
Yaowu Xu	44879ceea7	Merge "Remove redundant code"	2014-09-03 14:55:28 -07:00
Dmitry Kovalev	7f4c3b8d93	Merge "Cleaning up vp9_variance_avx2.c."	2014-09-03 13:21:38 -07:00
Yaowu Xu	ad3616a1fb	Merge "Merge two similar functions into one"	2014-09-03 13:00:02 -07:00
Dmitry Kovalev	a7ccc12973	Small cleanup: reusing existing code. Change-Id: Iac4775ad98e988f2b9cf5bd0dc91ab994d0262ce	2014-09-03 12:20:29 -07:00
Dmitry Kovalev	4eab7c28b8	Merge "Removing duplicated code."	2014-09-03 12:11:37 -07:00
Yaowu Xu	9a15835812	Merge "select_tx_mode(): remove special case for key frame"	2014-09-03 11:54:44 -07:00
Dmitry Kovalev	bf778e7d8e	Initializing intra modes without vpx_once(). Change-Id: I0a9d52432f2500f1bd8f43f229e70e38bb9a0343	2014-09-03 11:39:02 -07:00
Yaowu Xu	e759d95743	Merge two similar functions into one intra_super_block_yrd() and inter_super_block_yrd() are largely same, this commit merges them into one to reduce code duplication. Change-Id: I64d7042a5b099345627cf55663010c185b25ec37	2014-09-03 11:21:06 -07:00
Dmitry Kovalev	095d48a419	Merge "Removing clear_system_state() call from update_coef_probs()."	2014-09-03 11:05:45 -07:00
Minghai Shang	759afe525c	Merge "[svc] Temporal svc with two pass rate control"	2014-09-03 10:51:19 -07:00
Yaowu Xu	7a33712475	Change last_partition_redo_frequency for speed 3 From 3 to 2, which seems to be slightly positive on compression for all test sets, also reduces encoding time by 2%-5%, varying on the test clips. Change-Id: If045417bd27311700c919b4a335eff0dc1130ae0	2014-09-03 09:34:10 -07:00
Yaowu Xu	cdda17ed77	Remove redundant code Change-Id: I453b167f03811a3cd3592089593b3f2823f62ab3	2014-09-03 09:34:10 -07:00
Yaowu Xu	c1058e5bbe	select_tx_mode(): remove special case for key frame This commit removes the special case for key frame, as transform size decision is controlled by the appropriate speed feature for all lossy coding modes: tx_size_search_method. Change-Id: I9677171e3f2432ec23705f7c5ea8170dd4562fae	2014-09-03 09:34:10 -07:00
Paul Wilkins	819e231b93	Merge "Skip comp inter mode test in RD loop with same frame bias signs"	2014-09-03 02:26:47 -07:00
Jingning Han	801fef26ec	Skip comp inter mode test in RD loop with same frame bias signs This commit allows the encoder to skip check on compound inter modes in the rate-distortion optimization loop, if the reference frame bias signs are the same. Change-Id: Ib753e6bb11cbdd338aee69dbe2b649671f75a6b0	2014-09-02 18:17:33 -07:00
Dmitry Kovalev	070210e20b	Removing duplicated code. Change-Id: I7b5c776d5e6f5ca428b87fa9411ae4012a9538ba	2014-09-02 17:57:35 -07:00
Dmitry Kovalev	0ecc75c819	Merge "Removing MMX SAD calculation code."	2014-09-02 17:35:59 -07:00
Deb Mukherjee	a4ef1a0819	Merge "Adds config opt for highbitdepth + misc. vpx"	2014-09-02 15:41:27 -07:00
Dmitry Kovalev	318fc0c34f	Removing MMX SAD calculation code. Removed functions: * vp9_sad_16x16_mmx * vp9_sad_8x16_mmx * vp9_sad_16x8_mmx * vp9_sad_8x8_mmx * vp9_sad_4x4_mmx Change-Id: Ic5174b93b64d65d846f0c11e72cab149e9472bc3	2014-09-02 14:41:36 -07:00
Deb Mukherjee	5acfafb18e	Adds config opt for highbitdepth + misc. vpx Adds config parameter vp9_highbitdepth, to support highbitdepth profiles. Also includes most vpx level high bit-depth functions. However encode/decode in the highbitdepth profiles will not work until the rest of the code is in place. Change-Id: I34c53b253c38873611057a6cbc89a1361b8985a6	2014-09-02 14:37:10 -07:00
Dmitry Kovalev	6f6bd282c9	Replacing asm 16x16 variance calculation with intrinsics. New code is 20% faster for 64-bit and 15% faster for 32-bit. Compiled using clang. Change-Id: Icfea461238411001fd093561293dbfedfbf8d0bb	2014-09-02 13:54:34 -07:00
Minghai Shang	be3b08da3e	[svc] Temporal svc with two pass rate control It's built based on current spatial svc code. We only support one spatial two temporal layers at this time. Change-Id: I1fdc8584354b910331e626bfae60473b3b701ba1	2014-09-02 12:05:14 -07:00
Jingning Han	33176fef87	Skip comp inter mode tests for arf coding This commit skips the compound inter mode prediction check in the rate-distortion optimization loop for ARF coding. It reduces the runtime for certain test clips at speed 3, at no compression performance change: bus CIF 1000 kbps, 8260 ms -> 8090 ms, 1.8% speed-up stockholm 720p 1000 kbps, 74453 ms -> 71826 ms, 2.9% speed-up No visible speed-up for pedestrian area 1080p at 2000 kbps. Change-Id: Ic68aa56837159b726563b784e2e3729e846465ad	2014-09-02 11:23:47 -07:00
Dmitry Kovalev	5c937db029	Cleaning up vp9_variance_avx2.c. Change-Id: I75eb47dd21f87015efd673dbd2aa71f4386afdf5	2014-09-02 11:01:29 -07:00
Dmitry Kovalev	0a4403992a	Merge "Removing 'frames' field from VP9_COMP."	2014-09-02 10:01:20 -07:00
Dmitry Kovalev	7c24d21f2e	Merge "Removing lookup_next_frame_stats()."	2014-09-02 09:25:16 -07:00
Jingning Han	bac0268716	Merge "Skip intra mode tests depending on inter residuals"	2014-09-02 08:32:52 -07:00
Dmitry Kovalev	dbe2170595	Merge "Replacing asm 8x8 variance calculation with intrinsics."	2014-08-31 18:39:46 -07:00
Dmitry Kovalev	4ab2241f5b	Removing dummy_packing member from VP9_COMP. Change-Id: I571ce84c97087f8a1a36a10058393bfdcefbf72a	2014-08-29 17:33:20 -07:00
Dmitry Kovalev	0b721db543	Replacing asm 8x8 variance calculation with intrinsics. New code is 10% faster for 64-bit and 25% faster for 32-bit. Compiled using clang. Change-Id: I8ba1544c30dd6f3ca479db806384317549650dfc	2014-08-29 17:28:31 -07:00
Jingning Han	deb8882cca	Merge "Fix int64_t to unsigned int conversion warnings"	2014-08-29 17:15:46 -07:00
Jingning Han	dc3327c9dc	Merge "Extend block level sse to support multiple txfm blocks"	2014-08-29 17:15:30 -07:00
Jingning Han	6ddf1e152a	Fix int64_t to unsigned int conversion warnings Use unsigned int type to store the sse in the pixel domain. The precision is sufficient to handle sse of block size up to 64x64. The transform domain version however needs int64_t, since there is a transfer gain applied in the forward transformation that might cause unsigned int overflow. Change-Id: Ifef97c38597e426262290f35341fbb093cf0a079	2014-08-29 14:29:31 -07:00
Dmitry Kovalev	72037944df	Merge "Removing variance MMX code."	2014-08-29 14:08:02 -07:00
Yunqing Wang	a4a1ca109c	Merge "Minor fix in vp9_encoder.h"	2014-08-29 13:44:10 -07:00
Yunqing Wang	96c43e8aa9	Minor fix in vp9_encoder.h Added the missing "int". Change-Id: I7c8af3dee700837b40f010d53e1431a59370ae3a	2014-08-29 11:27:24 -07:00
Dmitry Kovalev	12cd6f421d	Removing variance MMX code. Removed functions: * vp9_mse16x16_mmx * vp9_get_mb_ss_mmx * vp9_get4x4var_mmx * vp9_get8x8var_mmx * vp9_variance4x4_mmx * vp9_variance8x8_mmx * vp9_variance16x16_mmx * vp9_variance16x8_mmx * vp9_variance8x16_mmx They all have SSE2 equivalent. Change-Id: I3796f2477c4f59b35b4828f46a300c16e62a2615	2014-08-29 10:26:42 -07:00
Jingning Han	4282955ee1	Skip intra mode tests depending on inter residuals This commit allows encoder to skip intra coding mode test, when the known inter residual is less than the source variance. It reduces the runtime of speed 3 for test clips: bus cif 1000 kbps: 8587 ms -> 8260 ms, 3.8% speed-up pedestrian 1080p 2000 kbps: 161381 ms -> 155241 ms, 3.7% speed-up. The compression performance is down by derf -0.36% stdhd -0.25% Change-Id: I75ce1e035b4da2153cb1ac14111d1a07c05a735d	2014-08-29 08:37:35 -07:00
Jingning Han	02e6ecdc4c	Extend block level sse to support multiple txfm blocks This commit extends the sse and forward transform computation flag to support the case 64x64 blocks where there are 4 32x32 2D-DCT blocks. Change-Id: I86a3e805dfaa0f3abd812f590520c71aa0e40473	2014-08-29 08:29:34 -07:00
Dmitry Kovalev	dcac083cf3	Implementing 4x4 variance calculation with SSE2. New SSE2 function is three times faster than MMX one. Change-Id: I4f387ce9f75b88379176ec7bdc62d86eb5f70fbe	2014-08-28 15:01:16 -07:00
Dmitry Kovalev	e9d106bd45	Merge "Removing unused arnr_type from VP9EncoderConfig and vp9_extracfg."	2014-08-28 13:50:05 -07:00
Yunqing Wang	5ac75188cb	Merge "Early termination in encoding partition search"	2014-08-28 13:49:39 -07:00
Dmitry Kovalev	c0383912df	Merge "Removing unused debug code under WRITE_RECON_BUFFER."	2014-08-28 11:46:45 -07:00
Dmitry Kovalev	57e0b2baf3	Merge "Converting configure_skippable_frame() to is_skippable_frame()."	2014-08-28 11:45:32 -07:00
Yunqing Wang	4d2c376923	Early termination in encoding partition search In the partition search, the encoder checks all possible partitionings in the superblock's partition search tree. This patch proposed a set of criteria for partition search early termination, which effectively decided whether or not to terminate the search in current branch based on the "skippable" result of the quantized transform coefficients. The "skippable" information was gathered during the partition mode search, and no overhead calculations were introduced. This patch gives significant encoding speed gains without sacrificing the quality. Borg test results: 1. At speed 1, stdhd set: psnr: +0.074%, ssim: +0.093%; derf set: psnr: -0.024%, ssim: +0.011%; 2. At speed 2, stdhd set: psnr: +0.033%, ssim: +0.100%; derf set: psnr: -0.062%, ssim: +0.003%; 3. At speed 3, stdhd set: psnr: +0.060%, ssim: +0.190%; derf set: psnr: -0.064%, ssim: -0.002%; 4. At speed 4, stdhd set: psnr: +0.070%, ssim: +0.143%; derf set: psnr: -0.104%, ssim: +0.039%; The speedup ranges from several percent to 60+%. speed1 speed2 speed3 speed4 (1080p, 100f): old_town_cross: 48.2% 23.9% 20.8% 16.5% park_joy: 11.4% 17.8% 29.4% 18.2% pedestrian_area: 10.7% 4.0% 4.2% 2.4% (720p, 200f): mobcal: 68.1% 36.3% 34.4% 17.7% parkrun: 15.8% 24.2% 37.1% 16.8% shields: 45.1% 32.8% 30.1% 9.6% (cif, 300f) bus: 3.7% 10.4% 14.0% 7.9% deadline: 13.6% 14.8% 12.6% 10.9% mobile: 5.3% 11.5% 14.7% 10.7% Change-Id: I246c38fb952ad762ce5e365711235b605f470a66	2014-08-28 11:27:28 -07:00
Deb Mukherjee	bb2a9abb1e	Merge "Updates vp9_pattern search to return integer sads"	2014-08-28 09:38:56 -07:00
Dmitry Kovalev	c4c0b2e765	Merge "Replacing int_mv with MV."	2014-08-28 09:18:11 -07:00
Deb Mukherjee	04b100b23e	Updates vp9_pattern search to return integer sads Updates the vp9_pattern_search function to return integer one-away neighbors' sad values, for subsequent use in speeding up the sub-pel search. Also, removes code for the do_refine option which is not being used currently. Updates the integer and subpel functions to pass in a 5-element sad list for output or input. A new pruned sub-pel search algorithm is implemented that uses the sad returned from the integer pel search. But it is not deployed yet. Change-Id: Ifa9f5ad024b5b660570366d2bd900343e1891520	2014-08-28 06:49:58 -07:00
Jingning Han	143be253b6	Merge "Re-work RD modeling based on inter frame prediction residual"	2014-08-27 18:48:49 -07:00
Jingning Han	34675e6631	Merge "Re-use switchable rate value in handle_inter_mode"	2014-08-27 18:48:41 -07:00
Jingning Han	4e4f4ba868	Merge "Add an early termination check in handle_inter_mode"	2014-08-27 18:48:32 -07:00
Jingning Han	6924fddb08	Merge "Use max txfm size unit in rate-distortion cost modeling"	2014-08-27 18:48:24 -07:00
Jingning Han	993ef8bd4c	Re-work RD modeling based on inter frame prediction residual This commit re-work the operation flow related to prediction residual generation and the rate-distortion modeling. It saves one call for model_rd_for_sb. Change-Id: Icaf96c0ff09c903637ed5283448afe01d798195f	2014-08-27 15:03:32 -07:00
Jingning Han	4db022c368	Re-use switchable rate value in handle_inter_mode The value of switchable rate has been stored in a local variable. This change skips the second call to vp9_get_switchable_rate() by reusing the local variable. Change-Id: Ib7d3fef7621cc4bde94c6d6e6b3a71f1fd4559f2	2014-08-27 15:03:16 -07:00
Jingning Han	cd228fcdb8	Add an early termination check in handle_inter_mode Check the mode and motion vector cost. If it is already above the existing best rate-distortion cost, skip the rest check process on this mode. Change-Id: Ie065cebdfda2a3be3be18b8e8b43dc29aaa8c179	2014-08-27 14:59:52 -07:00
Jingning Han	ec7ce316d2	Use max txfm size unit in rate-distortion cost modeling This commit makes the rate distortion modeling run in the unit of maximum transform block size. No compression/speed change observed. It is for the use of later fast forward transform purpose. Change-Id: Ibaaedb69c765e8d0c5d5012f0ec07f36fd9f68fd	2014-08-27 14:59:02 -07:00
Yaowu Xu	bcfb1ffb9d	Merge "add a new interp filter search strategy."	2014-08-26 17:30:42 -07:00
Dmitry Kovalev	668d3cf402	Replacing int_mv with MV. Change-Id: I483a2fefc5f9ea4533dfd64448f3b6b426dd9eed	2014-08-26 10:53:05 -07:00
Yaowu Xu	1144fee3d5	add a new interp filter search strategy. This commit addes a new strategy to reduce the search for optimal interpolation filter type. The encoder counts and store how many each filter type is selected and used for each of the reference frames. A filter type that is rarely used for all three reference frames is masked out to avoid computation. The impact on compression is neglectible: -0.02% on derf +0.02% on stdhd Encoding time is seen to reduce by 2~3%. Change-Id: Ibafa92291b51185de40da513716222db4b230383	2014-08-26 09:05:04 -07:00
Dmitry Kovalev	33f4e5707c	Removing unused arnr_type from VP9EncoderConfig and vp9_extracfg. Change-Id: Icab9a4399c5687453f4bec14b8cb5000464335e5	2014-08-25 23:48:52 -07:00
Dmitry Kovalev	a00278c6dc	Removing 'frames' field from VP9_COMP. Using local variable instead. Change-Id: If592d73ba2b04972cdae938751155c183a6db25a	2014-08-25 23:27:08 -07:00
Dmitry Kovalev	0586975912	Merge "Removing tx_stepdown_count from VP9_COMP."	2014-08-25 18:37:40 -07:00
Dmitry Kovalev	48edc8df31	Merge "Adding oxcf temp variable."	2014-08-25 18:37:33 -07:00
Dmitry Kovalev	0082727cb7	Merge "Adding is_keyframe temp var."	2014-08-25 18:36:59 -07:00
Dmitry Kovalev	4478553efc	Removing tx_stepdown_count from VP9_COMP. The variable is never read. Change-Id: I94141c1667fa5d10604cd6f83c5f64df107dee94	2014-08-25 14:42:05 -07:00
Minghai Shang	42ad07a138	Merge "[spatial svc]Multiple frame context feature"	2014-08-25 14:29:49 -07:00

... 6 7 8 9 10 ...

4920 Commits