Commits · d415d287175bff26f94e75055cdd9a6e7d2a4a3b · BC / public / external / libvpx

11 Apr, 2013 - 4 commits
- Merge loop over all macroblock modes into encode_sb_row(). · d415d287
  Ronald S. Bultje authored 11 years ago
```
Rename pick_mb_modes to pick_mb_mode, since it now handles only a
single macroblock. This is consistent with pick_sb_mode handling a
single non-macroblock.

Change-Id: I896fdfa06436b2d8c24d6474718cc74420df6b3b
```
  d415d287
- Remove "tplist" from VP9_COMP. · 2e2b8a53
  Ronald S. Bultje authored 11 years ago
```
It is write-only.

Change-Id: I2412344688d96593cc01c038e7f51410d0f85ed0
```
  2e2b8a53
- Merge pick_sb_modes and pick_sb64_modes. · 605ff051
  Ronald S. Bultje authored 11 years ago
```
Change-Id: Iad69e7a3b7e470acf6094f6a52e7da69066fd552
```
  605ff051
- Slight simplification of SB RD loop recursion conditions. · 38d79453
  Ronald S. Bultje authored 11 years ago
```
Change-Id: I87a406fcd18ab043253ca0c009d1182fdc5c3046
```
  38d79453
10 Apr, 2013 - 2 commits

Make RD superblock mode search size-agnostic. · b4f6098e

Ronald S. Bultje authored 11 years ago

Merge various super_block_yrd and super_block_uvrd versions into one
common function that works for all sizes. Make transform size selection
size-agnostic also. This fixes a slight bug in the intra UV superblock
code where it used the wrong transform size for txsz > 8x8, and stores
the txsz selection for superblocks properly (instead of forgetting it).
Lastly, it removes the trellis search that was done for 16x16 intra
predictors, since trellis is relatively expensive and should thus only
be done after RD mode selection.

Gives basically identical results on derf (+0.009%).

Change-Id: If4485c6f0a0fe4038b3172f7a238477c35a6f8d3

b4f6098e

Make SB coding size-independent. · a3874850

Ronald S. Bultje authored 11 years ago

Merge sb32x32 and sb64x64 functions; allow for rectangular sizes. Code
gives identical encoder results before and after. There are a few
macros for rectangular block sizes under the sbsegment experiment; this
experiment is not yet functional and should not yet be used.

Change-Id: I71f93b5d2a1596e99a6f01f29c3f0a456694d728

a3874850

05 Apr, 2013 - 1 commit

Remove full-pixel-related code. · 36c3a67c

Ronald S. Bultje authored 11 years ago

This is a VP8-only feature (part of profile 3) that is unsupported in
VP9.

Change-Id: I78016eede8d9c834d44d4c517f3e8b8fc2a378b1

36c3a67c

03 Apr, 2013 - 1 commit

Renaming sb32_coded and sb64_coded fields. · dca8ad17

Dmitry Kovalev authored 11 years ago

Renaming sb32_coded to prob_sb32_coded and sb64_coded to prob_sb64_coded.

Change-Id: I6de5cad00a57c3e066d53467f8c38cb6073dce11

dca8ad17

28 Mar, 2013 - 2 commits

Framework changes in nzc to allow more flexibility · fe9b5143

Deb Mukherjee authored 11 years ago

The patch adds the flexibility to use standard EOB based coding
on smaller block sizes and nzc based coding on larger blocksizes.
The tx-sizes that use nzc based coding and those that use EOB based
coding are controlled by a function get_nzc_used().
By default, this function uses nzc based coding for 16x16 and 32x32
transform blocks, which seem to bridge the performance gap
substantially.

All sets are now lower by 0.5% to 0.7%, as opposed to ~1.8% before.

Change-Id: I06abed3df57b52d241ea1f51b0d571c71e38fd0b

fe9b5143

Fix crash when --tune=ssim is selected. · befb0393

Paul Wilkins authored 11 years ago

Crash fix only. No functional change or testing.

Change-Id: I0c6d114d024c29fc11ae61666f5938f11b01dd6a

befb0393

27 Mar, 2013 - 1 commit

Removing redundant function arguments. · 17cddb4e

Dmitry Kovalev authored 11 years ago

Almost all arguments for vp9_build_inter32x32_predictors_sb and
vp9_build_inter64x64_predictors_sb can be deduced from the first macroblock
argument.

Change-Id: I5d477a607586d05698d5b3b9b9bc03891dd3fe83

17cddb4e

26 Mar, 2013 - 3 commits

Implicit weighted prediction experiment · 23144d23

Deb Mukherjee authored 12 years ago

Adds an experiment to use a weighted prediction of two INTER
predictors, where the weight is one of (1/4, 3/4), (3/8, 5/8),
(1/2, 1/2), (5/8, 3/8) or (3/4, 1/4), and is chosen implicitly
based on consistency of the predictors to the already
reconstructed pixels to the top and left of the current macroblock
or superblock.

Currently the weighting is not applied to SPLITMV modes, which
default to the usual (1/2, 1/2) weighting. However the code is in
place controlled by a macro. The same weighting is used for Y and
UV components, where the weight is derived from analyzing the Y
component only.

Results (over compound inter-intra experiment)
derf: +0.18%
yt: +0.34%
hd: +0.49%
stdhd: +0.23%

The experiment suggests bigger benefit for explicitly signaled weights.

Change-Id: I5438539ff4485c5752874cd1eb078ff14bf5235a

23144d23

Use above/left (instead of previous in scan-order) as token context. · 790fb132

Ronald S. Bultje authored 11 years ago

Pearson correlation for above or left is significantly higher than for
previous-in-scan-order (absolute values depend on position in scan, but
in general, we gain about 0.1-0.2 by using either above or left; using
both basically just makes this even better). For eob branch skipping,
we continue to use the previous token in scan order.

This helps about 0.9% on derf after re-training on a limited data set.
Full re-training and results on larger-resolution clips are pending.

Note that this commit breaks trellis, so we can probably get further
gains out of it by fixing trellis at some later point.

Change-Id: Iead68e296fc3a105cca746b5e3da9555d6010cfe

790fb132

Modeling default coef probs with distribution · fd18d5df

Deb Mukherjee authored 12 years ago

Replaces the default tables for single coefficient magnitudes with
those obtained from an appropriate distribution. The EOB node
is left unchanged. The model is represeted as a 256-size codebook
where the index corresponds to the probability of the Zero or the
One node. Two variations are implemented corresponding to whether
the Zero node or the One-node is used as the peg. The main advantage
is that the default prob tables will become considerably smaller and
manageable. Besides there is substantially less risk of over-fitting
for a training set.

Various distributions are tried and the one that gives the best
results is the family of Generalized Gaussian distributions with
shape parameter 0.75. The results are within about 0.2% of fully
trained tables for the Zero peg variant, and within 0.1% of the
One peg variant.

The forward updates are optionally (controlled by a macro)
model-based, i.e. restricted to only convey probabilities from the
codebook. Backward updates can also be optionally (controlled by
another macro) model-based, but is turned off by default. Currently
model-based forward updates work about the same as unconstrained
updates, but there is a drop in performance with backward-updates
being model based.

The model based approach also allows the probabilities for the key
frames to be adjusted from the defaults based on the base_qindex of
the frame. Currently the adjustment function is a placeholder that
adjusts the prob of EOB and Zero node from the nominal one at higher
quality (lower qindex) or lower quality (higher qindex) ends of the
range. The rest of the probabilities are then derived based on the
model from the adjusted prob of zero.

Change-Id: Iae050f3cbcc6d8b3f204e8dc395ae47b3b2192c9

fd18d5df

22 Mar, 2013 - 1 commit
- Minor code clean up · 815734e5
  Paul Wilkins authored 12 years ago
```
Change-Id: Ifa864e0acb253b238b03cdeed0fe5d6ee30a45d8
```
  815734e5
16 Mar, 2013 - 1 commit

Remove some unused rate control variables · b8ac9f2f

John Koleszar authored 12 years ago

These variables are unused, and are subject to overflowing, causing
assertions when built with -ftrapv.

Change-Id: Ia00a3201af309906c05bcd4b23a643925ed6ea86

b8ac9f2f

14 Mar, 2013 - 2 commits
- force lossless coding at very high quality end · 374a1736
  Yaowu Xu authored 12 years ago
```
Change-Id: I75fc4eee10bee9efd419d248827290cce8e6d637
```
  374a1736
- Remove leftover reference to 2nd order dc/ac quant · f4d2ad69
  Yaowu Xu authored 12 years ago
```
Change-Id: Ib8dacf1d2797743569771b8f699e40e1aeb085cb
```
  f4d2ad69
13 Mar, 2013 - 1 commit

removed reference to "LLM" and "x8" · 00555263

Yaowu Xu authored 12 years ago

The commit changed the name of files and function to remove obselete
reference to LLM and x8.

Change-Id: I973b20fc1a55149ed68b5408b3874768e6f88516

00555263

09 Mar, 2013 - 1 commit

Continued experiment with nonzero count · a28139c8

Deb Mukherjee authored 12 years ago

Adds probability updates for extra bits for the nzcs, code for
getting nzc stats, plus some minor cleanups and fixes.

Change-Id: If2814e7f04fb52f5025ad9f400f3e6c50a00b543

a28139c8

08 Mar, 2013 - 2 commits

Extend diff MV limit from +/-256 to +/-1024 · 2a5278bd

Jingning Han authored 12 years ago

Increase the motion search range by 4x. Change MV_CLASS tree of the
entropy coding to allow two additional mv classes to cover the
extended motion vector limit. The codec determines the effective
motion search range conditioned on the actual frame dimension.

It provides coding gains:

stdhd 0.39%
yt    0.56%
hd    0.47%

Major coding performance gains are packed in several sequences with
intense motion activities, e.g., ped_1080p gains 7% at high bit-rates,
and on average 3%.

TODO: Need to further tune the rate control and motion search units.

Change-Id: Ib842540a6796fbee5a797809433ef6a477c6d78d

2a5278bd

Add support for tx_select in i8x8 encoding in keyframes. · b41dee84
Ronald S. Bultje authored 12 years ago
```
Also enable tx_select for keyframes.

Change-Id: Iadb1231d9fa7af0c8dce3d9b41830b93a302479e
```
b41dee84

07 Mar, 2013 - 1 commit

Coding con-zero count rather than EOB for coeffs · eb6ef241

Deb Mukherjee authored 12 years ago

This patch revamps the entropy coding of coefficients to code first
a non-zero count per coded block and correspondingly remove the EOB
token from the token set.

STATUS:
Main encode/decode code achieving encode/decode sync - done.
Forward and backward probability updates to the nzcs - done.
Rd costing updates for nzcs - done.
Note: The dynamic progrmaming apporach used in trellis quantization
is not exactly compatible with nzcs. A suboptimal approach has been
used instead where branch costs are updated to account for changes
in the nzcs.

TODO:
Training the default probs/counts for nzcs

Change-Id: I951bc1e22f47885077a7453a09b0493daa77883d

eb6ef241

05 Mar, 2013 - 1 commit

Make superblocks independent of macroblock code and data. · 111ca421

Ronald S. Bultje authored 12 years ago

Split macroblock and superblock tokenization and detokenization
functions and coefficient-related data structs so that the bitstream
layout and related code of superblock coefficients looks less like it's
a hack to fit macroblocks in superblocks.

In addition, unify chroma transform size selection from luma transform
size (i.e. always use the same size, as long as it fits the predictor);
in practice, this means 32x32 and 64x64 superblocks using the 16x16 luma
transform will now use the 16x16 (instead of the 8x8) chroma transform,
and 64x64 superblocks using the 32x32 luma transform will now use the
32x32 (instead of the 16x16) chroma transform.

Lastly, add a trellis optimize function for 32x32 transform blocks.

HD gains about 0.3%, STDHD about 0.15% and derf about 0.1%. There's
a few negative points here and there that I might want to analyze
a little closer.

Change-Id: Ibad7c3ddfe1acfc52771dfc27c03e9783e054430

111ca421

04 Mar, 2013 - 1 commit

Support 16K sequence coding · 5957b2b5

Jingning Han authored 12 years ago

Fixed a couple of variable/function definitions, as well as header
handling to support 16K sequence coding at high bit-rates.

The width and height are each specified by two bytes in the header.
Use an extra byte to explicitly indicate the scaling factors in
both directions, each ranging from 0 to 15.

Tested coding up to 16400x16400 dimension.

Change-Id: Ibc2225c6036620270f2c0cf5172d1760aaec10ec

5957b2b5

28 Feb, 2013 - 1 commit

Code cleanup. · 0d9cc0a9

Dmitry Kovalev authored 12 years ago

Removing redundant 'extern' keyword, better formatting, code
simplification.

Change-Id: I132fea14f08c706ee9ea147d19464d03f833f25b

0d9cc0a9

27 Feb, 2013 - 2 commits

Use ref_frame_map vice active_ref_idx on the encoder · 800ad0b8

John Koleszar authored 12 years ago

This patch makes the encoder's use of ref_frame_map and active_ref_idx
consistent with the decoder. ref_frame_map[] maps a reference buffer
index to its actual location in the yv12_fb array, since many
references may share an underlying buffer. active_ref_idx[] mirrors
cpi->{lst,gld,alt}_fb_idx, holding the active references in each
slot.

This also fixes a bug in setup_buffer_inter() where the incorrect
reference was used to populate the scaling factors.

Change-Id: Id3728f6d77cffcd27c248903bf51f9c3e594287e

800ad0b8

Spatial resamping of ZEROMV predictors · eb939f45

John Koleszar authored 12 years ago

This patch allows coding frames using references of different
resolution, in ZEROMV mode. For compound prediction, either
reference may be scaled.

To test, I use the resize_test and enable WRITE_RECON_BUFFER
in vp9_onyxd_if.c. It's also useful to apply this patch to
test/i420_video_source.h:

  --- a/test/i420_video_source.h
  +++ b/test/i420_video_source.h
  @@ -93,6 +93,7 @@ class I420VideoSource : public VideoSource {

     virtual void FillFrame() {
       // Read a frame from input_file.
  +    if (frame_ != 3)
       if (fread(img_->img_data, raw_sz_, 1, input_file_) == 0) {
         limit_ = frame_;
       }

This forces the frame that the resolution changes on to be coded
with no motion, only scaling, and improves the quality of the
result.

Change-Id: I1ee75d19a437ff801192f767fd02a36bcbd1d496

eb939f45

26 Feb, 2013 - 1 commit

Refactor inter recon functions to support scaling · 6a4f708c

John Koleszar authored 12 years ago

Ensure that all inter prediction goes through a common code path
that takes scaling into account. Removes a bunch of duplicate
1st/2nd predictor code. Also introduces a 16x8 mode for 8x8
MVs, similar to the 8x4 trick we were doing before. This has an
unexpected effect with EIGHTTAP_SMOOTH, so it's disabled in that
case for now.

Change-Id: Ia053e823a8bc616a988a0af30452e1e75a739cba

6a4f708c

23 Feb, 2013 - 1 commit
- Split coefficient token tables intra vs. inter. · 0c9e2e9a
  Ronald S. Bultje authored 12 years ago
```
Change-Id: I5416455f8f129ca0f450d00e48358d2012605072
```
  0c9e2e9a
20 Feb, 2013 - 1 commit
- Merge lossless experiment · d262e26c
  Yaowu Xu authored 12 years ago
```
Change-Id: I7b7b8d4fda3a23699e0c920d727f8c15d37d43aa
```
  d262e26c
19 Feb, 2013 - 1 commit

Use lossless for Q0 · 93d6b86c

Yaowu Xu authored 12 years ago

The commit changes the coding mode to lossless whenever the lowest
quantizer is choosen.

As expected, test results showed no difference for cif and std-hd
set where Q0 is rarely used. For yt and yt-hd set, Q0 is used for
a number of clips, where this commit helped a lot in the high end.

Average over all clips in the sets:
yt: 2.391% 1.017% 1.066%
hd: 1.937%  .764%  .787%

Change-Id: I9fa9df8646fd70cb09ffe9e4202b86b67da16765

93d6b86c

15 Feb, 2013 - 1 commit
- Remove some Y2-related code. · 46dff5d2
  Ronald S. Bultje authored 12 years ago
```
Change-Id: I4f46d142c2a8d1e8a880cfac63702dcbfb999b78
```
  46dff5d2
13 Feb, 2013 - 2 commits

Add support for tile rows. · 89a206ef

Ronald S. Bultje authored 12 years ago

These allow sending partial bitstream packets over the network before
encoding a complete frame is completed, thus lowering end-to-end
latency. The tile-rows are not independent.

Change-Id: I99986595cbcbff9153e2a14f49b4aa7dee4768e2

89a206ef

fix the lossless experiment · 16f25f9d
Yaowu Xu authored 12 years ago
```
Change-Id: I95acfc1417634b52d344586ab97f0abaa9a4b256
```
16f25f9d

12 Feb, 2013 - 1 commit

Add tile column size limits (256 pixels min, 4096 pixels max). · f496f601

Ronald S. Bultje authored 12 years ago

This is after discussion with the hardware team. Update the unit test
to take these sizes into account. Split out some duplicate code into
a separate file so it can be shared.

Change-Id: I8311d11b0191d8bb37e8eb4ac962beb217e1bff5

f496f601

08 Feb, 2013 - 1 commit

Pass macroblock index to pick inter functions · 6125a1ed

John Koleszar authored 12 years ago

Pass the current mb row and column around rather than the
recon_yoffset and recon_uvoffset, since those offsets will
change from predictor to predictor, based on the reference
frame selection.

Change-Id: If3f9df059e00f5048ca729d3d083ff428e1859c1

6125a1ed

07 Feb, 2013 - 1 commit

Added skip switches for SB32 and SB64 · 29731308

Paul Wilkins authored 12 years ago

Added switches and code to skip/breakout from
doing SB32 and SB64 tests based on whether
the 16x16 MB tests used split modes. Also to
optionally skip 64x64 if 16x16 was chosen over
32x32.

Impact varies depending on clip from a few %
up to almost 50% on encode speed. Only the
split mode breakout is currently enabled.

Change-Id: Ib5836140b064b350ffa3057778ed2cadcc495cf8

29731308

05 Feb, 2013 - 1 commit

[WIP] Add column-based tiling. · 1407bdc2

Ronald S. Bultje authored 12 years ago

This patch adds column-based tiling. The idea is to make each tile
independently decodable (after reading the common frame header) and
also independendly encodable (minus within-frame cost adjustments in
the RD loop) to speed-up hardware & software en/decoders if they used
multi-threading. Column-based tiling has the added advantage (over
other tiling methods) that it minimizes realtime use-case latency,
since all threads can start encoding data as soon as the first SB-row
worth of data is available to the encoder.

There is some test code that does random tile ordering in the decoder,
to confirm that each tile is indeed independently decodable from other
tiles in the same frame. At tile edges, all contexts assume default
values (i.e. 0, 0 motion vector, no coefficients, DC intra4x4 mode),
and motion vector search and ordering do not cross tiles in the same
frame.
t log

Tile independence is not maintained between frames ATM, i.e. tile 0 of
frame 1 is free to use motion vect...

1407bdc2

28 Jan, 2013 - 1 commit

Segment Skip Flag · 0ff9b033

Paul Wilkins authored 12 years ago

First step in simplifying the segment mode and
segment EOB flags into a simpler segment skip
flag that implies 0,0 mv and EOB at position 0.

Change-Id: Ib750cac31a7a02dc21082580498efd9f7d8d72a5

0ff9b033