Highly Parallel Rate-Distortion Optimized Intra-Mode Decision on Multicore Graphics Processors

被引:25
|
作者
Cheung, Ngai-Man [1 ]
Au, Oscar C. [1 ]
Kung, Man-Cheung [2 ]
Wong, Peter H. W. [2 ]
Liu, Chun Hung [1 ]
机构
[1] Hong Kong Univ Sci & Technol, Dept Elect & Comp Engn, Kowloon, Hong Kong, Peoples R China
[2] Visual Percept Dynam Labs Mobile Ltd, Shatin, Hong Kong, Peoples R China
关键词
Graphics processing unit; greedy approach; multicore; parallel processing; rate-distortion optimized intra-prediction;
D O I
10.1109/TCSVT.2009.2031515
中图分类号
TM [电工技术]; TN [电子技术、通信技术];
学科分类号
0808 ; 0809 ;
摘要
Rate-distortion (RD)-based mode selections are important techniques in video coding. In these methods, an encoder may compute the RD costs for all the possible coding modes, and select the one which achieves the best trade-off between encoding rate and compression distortion. Previous papers have demonstrated that RD-based mode selections can lead to significant improvements in coding efficiency. RD-based mode selections, however, would incur considerable increases in encoding complexity, since these methods require computing the RD costs for numerous candidate coding modes. In this paper, we consider the scenario where software-based video encoding is performed on personal computers or game consoles, and investigate how multi-core graphics processing units (GPUs) may be efficiently utilized to undertake the task of RD optimized intra-prediction mode selections in audio and video coding standards and H. 264 video encoding. Achieving efficient GPU-based intra-mode decisions, however, could be nontrivial for two reasons. First, intra-mode decision tends to be sequential. Specifically, the mode decision of the current block would depend on the reconstructed data of the neighboring blocks. Therefore, the coding modes of neighboring blocks would need to be computed first before that of the current block can be determined. This dependency poses challenges to GPU-based computation, which relies heavily on parallel data processing to achieve superior speedups. Second, RD-based intra-mode decision may require conditional branchings to determine the encoding bit-rate, and these branching operations may incur substantial performance penalties when being executed on GPUs due to pipeline architectural designs. To address these issues, we analyze the data dependency in intra-mode decision, and propose novel greedy-based encoding orders to achieve highly parallel processing of data blocks. We also prove that the proposed greedy-based orders are optimal in our problem, i.e., they require the minimum number of iterations to process a video frame given the dependency constraints. In addition, we propose a method to estimate the coding rate suitable for GPU implementation. Experimental results suggest our proposed solution can be more than 50 times faster than the previously proposed parallel intra-prediction, since our work can efficiently exploit the massive parallel opportunity in GPUs.
引用
收藏
页码:1692 / 1703
页数:12
相关论文
共 48 条
  • [11] Performance Exploration of Jointly Rate-Distortion Optimized HEVC Intra Encoder
    Zhang, Yingwen
    Wang, Meng
    Li, Junru
    Wang, Shiqi
    2024 DATA COMPRESSION CONFERENCE, DCC, 2024, : 603 - 603
  • [12] Sum of Absolute Difference-based Rate-Distortion Optimization Cost Function for H.265/HEVC Intra-Mode Prediction
    Alariao, Naomi Z.
    Caguia, Kimberly P.
    Olaguer, Patricia Anne J.
    Sison, Giljohn Andrew, V
    Valenton, Jian Mae C.
    Serrano, Kanny Krizzy D.
    Dela Cruz, Angelo R.
    2019 IEEE 11TH INTERNATIONAL CONFERENCE ON HUMANOID, NANOTECHNOLOGY, INFORMATION TECHNOLOGY, COMMUNICATION AND CONTROL, ENVIRONMENT, AND MANAGEMENT (HNICEM), 2019,
  • [13] FAST PREDICTION MODE DECISION WITH HADAMARD TRANSFORM BASED RATE-DISTORTION COST ESTIMATION FOR HEVC INTRA CODING
    Zhu, Jia
    Liu, Zhenyu
    Wang, Dongsheng
    Han, Qingrui
    Song, Yang
    2013 20TH IEEE INTERNATIONAL CONFERENCE ON IMAGE PROCESSING (ICIP 2013), 2013, : 1977 - 1981
  • [14] Rate-distortion optimized mode selection method for multiple description video coding
    Yu-Chen Sun
    Wen-Jiin Tsai
    Multimedia Tools and Applications, 2014, 72 : 1411 - 1439
  • [15] Rate-distortion optimized mode selection method for multiple description video coding
    Sun, Yu-Chen
    Tsai, Wen-Jiin
    MULTIMEDIA TOOLS AND APPLICATIONS, 2014, 72 (02) : 1411 - 1439
  • [16] Parallel H.264/AVC Fast Rate-Distortion Optimized Motion Estimation by Using a Graphics Processing Unit and Dedicated Hardware
    Shahid, Muhammad Usman
    Ahmed, Ashfaq
    Martina, Maurizio
    Masera, Guido
    Magli, Enrico
    IEEE TRANSACTIONS ON CIRCUITS AND SYSTEMS FOR VIDEO TECHNOLOGY, 2015, 25 (04) : 701 - 715
  • [17] Rate distortion optimized mode decision in the scalable video coding
    Yang, ZJ
    Wu, F
    Li, SP
    2003 INTERNATIONAL CONFERENCE ON IMAGE PROCESSING, VOL 3, PROCEEDINGS, 2003, : 781 - 784
  • [18] Fast Inter-Mode Decision Based on Rate-Distortion Cost Characteristics
    Hu, Sudeng
    Zhao, Tiesong
    Wang, Hanli
    Kwong, Sam
    ADVANCES IN MULTIMEDIA INFORMATION PROCESSING-PCM 2010, PT II, 2010, 6298 : 145 - +
  • [19] Rate-Distortion Optimized Transforms Based on the Lloyd-Type Algorithm for Intra Block Coding
    Zou, Feng
    Au, Oscar C.
    Pang, Chao
    Dai, Jingjing
    Zhang, Xingyu
    Fang, Lu
    IEEE JOURNAL OF SELECTED TOPICS IN SIGNAL PROCESSING, 2013, 7 (06) : 1072 - 1083
  • [20] A rate-distortion optimized real time intra-update method for packet video transmission
    Deng, Y
    Peng, Q
    Yang, TW
    2003 IEEE INTELLIGENT TRANSPORTATION SYSTEMS PROCEEDINGS, VOLS. 1 & 2, 2003, : 915 - 918