GPU Accelerated H.264 Video Compression for Broadcast
Metadata
- Publisher
- SMPTE — White Plains, NY, USA
- Doc Type
- Journal Article
- Content Type
- Original Research
- Abbreviated Title
- SMPTE Mot. Imag. J
- Volume
- 120, No. 3, pp. 28–35
- Abstract
- In recent years, the ability to perform massive parallel computations has become readily available on every desktop due to improved graphic processors and new programming tools such as CUDA. The standard central processing units (CPUs) are also evolving rapidly in the same direction, making this technology even more accessible. Video compression is an illustrative example of an area where the speed/computation power trade-off is especially steep. To render all the features of the modern H.264 compression standard, a supercomputing level of power might be necessary. The modern graphics processing unit (GPU) and the next generation of the CPU could deliver the required computation power. We investigate what can be done to exploit the features of the modern GPUs for H.264 compression and how the new video compression standards such as H.265 might be adopted in the future of massive parallel processing.
- Publication Date
- 2011-04-01
- DOI
10.5594/j18028- ISSN
- Print:
1545-0279 - Link
- https://doi.org/10.5594/j18028
- Author(s)
- Dmitry KlimovJan Weigner
- Copyright
- © 2011 Society of Motion Picture and Television Engineers, Inc.
Bibliographic Reference(s)
- Intel , “Intel Advanced Vector Extensions Programming Reference,” www.software.intel.com/file/10069 . EXTERNAL
- H265.net , http://www.h265.net . EXTERNAL
- Nvidia , “Nvidia Speed Results,” http://www.mainconcept.com/fileadmin/user_upload/download/product_sheets/CUDA-Sheets_06-2010.pdf , 2010 . EXTERNAL
- Cinegy Cinecoder , http://www.cinecoder.com . EXTERNAL
- Elemental Technologies, Inc. , “Harness the Power of Massively Parallel Video Processing,” http://www.elementaltechnologies.com/products/product-overview . EXTERNAL
- Bailey David H. , “A High-Performance FFT Algorithm for Vector Supercomputers,” Proc. Third SIAM Conference on Parallel Processing for Scientific Computing , Abstract, p. 114 , December 1987 . EXTERNAL
- Franchetti F. Püschel M. Voronenko Y. Chellappa S. Moura J. M. F. , “Discrete Fourier Transform on Multicore,” IEEE Signal Processing Magazine, special issue on Signal Processing on Platforms with Multiple Cores , 26 ( 6 ): 90 – 102 , 2009 . EXTERNAL
- Frigo M. Johnson S. G. , “The Design and Implementation of FFTW3,” Proc. IEEE , 93 : 216 – 231 , 2005 . EXTERNAL
- Genovese L. , “Graphic Processing Units: A Possible Answer to HPC,” Fourth ABINIT Developer Workshop , 2009 . EXTERNAL
- Govindaraju Naga K. Lloyd Brandon Dotsenko Yuri Smith Burton Manferdelli John , “High Performance Discrete Fourier Transforms on Graphics Processors,” Proc. 2008 ACM/IEEE Conf. on Supercomp ., Nov. 15–21, Austin, TX, 2008 . EXTERNAL
- Johnson J. R. Johnson R. W. Rodriquez D. Tolimieri R. , “A Methodology for Designing, Modifying, and Implementing Fourier Transform Algorithms on Various Architectures,” CSSP , 9 ( 4 ): 449 – 500 , 1990 . EXTERNAL
- Intel , “Intel SSE4 Programming Reference,” 2007 , www.developers.net/intelisdshowcase/view/2550 . EXTERNAL
- Kumar Sanjeev Kim Daehyun Smelyanskiy Mikhail Chen Yen-Kuang Chhugani Jatin Hughes Christopher J. Kim Changkyu Lee Victor W. Nguyen Anthony D. , “Atomic Vector Operations on Chip Multiprocessors,” Proc. 35th Annual Int. Symp. on Comp. Arch ., pp. 441 – 452 , June 21–25, 2008 . EXTERNAL
- Ramanathan R. , “Extending the World's Most Popular Processor Architecture,” Intel Architecture White Paper , 2006 . EXTERNAL
- Satish Nadathur Kim Changkyu Chhugani Jatin Nguyen Anthony D. Lee Victor W. Kim Daehyun Dubey Pradeep , “Fast Sort on CPUs and GPUs: A Case for Bandwidth Oblivious SIMD Sort,” Proc. 2010 Int. Conf. on Management of Data , Indianapolis, IN , June 2010 . EXTERNAL
- Seiler Larry Carmean Doug Sprangle Eric Forsyth Tom Abrash Michael Dubey Pradeep Junkins Stephen Lake Adam Sugerman Jeremy Cavin Robert Espasa Roger Grochowski Ed Juan Toni Hanrahan Pat , “Larrabee: A Many-Core x86 Architecture for Visual Computing,” ACM TOG , 27 ( 3 ): 18 , Aug. 2008 . EXTERNAL
- Silberstein Mark Schuster Assaf Geiger Dan Patney Anjul Owens John D. , “Efficient Computation of Sum-Products on GPUs Through Software-Managed Cache,” Proc. 22nd Annual Int. Conf. on Supercomputing , Island of Kos , Greece , June 2008 . EXTERNAL
- Volkov V. Demmel J. , “LU, QR and Cholesky Factorizations Using Vector Capabilities of GPUs,” Tech. Report UCB/EECS-2008-49 , EECS Department, University of California , Berkeley , May 2008 . EXTERNAL
- Volkov Vasily Demmel James W. , “Benchmarking GPUs to Tune Dense Linear Algebra,” Proc. 2008 ACM/IEEE Conf. on Supercomputing . Austin, TX , Nov. 2008 . EXTERNAL
- Williams Samuel Waterman Andrew Patterson David , “Roofline: An Insightful Visual Performance Model for Multicore Architectures,” Commun. ACM , 52 ( 4 ): 65 – 76 , April 2009 . EXTERNAL
- Xu W. Mueller K. , “A Performance-Driven Study of Regularization Methods for GPU-Accelerated Iterative CT,” 2nd High Performance Reconstruction Workshop (HPIR ), Beijing, China , Sept. 2009 . EXTERNAL
- Yang Zhiyi Zhu Yating Pu Yong , “Parallel Image Processing Based on CUDA,” Proc. 2008 Int. Conf. on Comp. Science and Software Eng ., pp. 198 – 201 , Dec. 2008 . EXTERNAL
- Leischner N. Osipov V. Sanders P. , Fermi Architecture White Paper, 2009 . EXTERNAL
- Nvidia CUDA Zone , http://www.nvidia.com/object/cuda_home.html , 2010 . EXTERNAL
- General-Purpose Computation on Graphics Hardware , http://gpgpu.org , 2009 . EXTERNAL
- CUDA BLAS Library , http://developer.download.nvidia.com/compute/cuda/2_1/toolkit/docs/CUBLAS_Library_2.1.pdf , 2008 . EXTERNAL
- CUDA CUFFT Library , http://developer.download.nvidia.com/compute/cuda/2_1/toolkit/docs/CUFFT_Library_2.1.pdf , 2008 . EXTERNAL
- ISO/IEC 14496–10, International Organization for Standardization , www.iso.org . EXTERNAL
- Video Coding Experts Group , Meeting Report for 31st VCEG Meeting, Marrakech, Morocco, 15–16 Jan. 2007. EXTERNAL
Source Data (JSON)
Full registry record with provenance metadata. Open directly: /api/doc/10.5594-j18028.json
Reference Tree
Explore all references and references to this document, as a navigable tree.
Open Reference TreeReference this Doc
Plain text (ISO 690 compliant)
Preview:
Dmitry Klimov and Jan Weigner; GPU Accelerated H.264 Video Compression for Broadcast, SMPTE Motion Imaging Journal ( Volume: 120, Issue: 3, April 2011); SMPTE, 2011. Available at https://doi.org/10.5594/j18028
Snippet:
Dmitry Klimov and Jan Weigner; GPU Accelerated H.264 Video Compression for Broadcast, SMPTE Motion Imaging Journal ( Volume: 120, Issue: 3, April 2011); SMPTE, 2011. Available at https://doi.org/10.5594/j18028
HTML (ISO 690 compliant)
Preview:
Dmitry Klimov and Jan Weigner; GPU Accelerated H.264 Video Compression for Broadcast, SMPTE Motion Imaging Journal ( Volume: 120, Issue: 3, April 2011); SMPTE, 2011. Available at https://doi.org/10.5594/j18028
Snippet:
<span class="citation">Dmitry Klimov and Jan Weigner; <cite>GPU Accelerated H.264 Video Compression for Broadcast</cite>, SMPTE Motion Imaging Journal ( Volume: 120, Issue: 3, April 2011); SMPTE, 2011. Available at <a href="https://doi.org/10.5594/j18028" target="_blank" rel="noopener">https://doi.org/10.5594/j18028</a></span>
SMPTE's HTML Pub
Preview:
Dmitry Klimov and Jan Weigner; GPU Accelerated H.264 Video Compression for Broadcast, SMPTE Motion Imaging Journal ( Volume: 120, Issue: 3, April 2011); SMPTE, 2011
doi: 10.5594/j18028
url: https://doi.org/10.5594/j18028
doi: 10.5594/j18028
url: https://doi.org/10.5594/j18028
Snippet:
<li> Dmitry Klimov and Jan Weigner; <cite id="bib-10-5594-j18028">GPU Accelerated H.264 Video Compression for Broadcast</cite>, SMPTE Motion Imaging Journal ( Volume: 120, Issue: 3, April 2011); SMPTE, 2011 <span class="doi">10.5594/j18028</span> </li>