{"id":{"repo_id":"mit","oai_identifier":"oai:dspace.mit.edu:1721.1/66707"},"canonical_url":"https://search.dev.ndltd.org/etd/mit/oai:dspace.mit.edu:1721.1/66707","repository":{"repo_id":"mit","name":"MIT","base_url":"https://dspace.mit.edu/oai/request"},"display":{"title":"Analysis of one-dimensional transforms in coding motion compensation prediction residuals for video applications","abstract":"In video coding, motion compensation prediction provides significant increases in overall compression efficiency. The prediction residuals are typically treated as images and compressed by applying two-dimensional transforms such as the two-dimensional discrete cosine transform (2D-DCT). Previous work has found that the use of direction-adaptive one-dimensional discrete cosine transforms (1D-DCTs) in coding motion compensation residuals can provide significant additional bitrate savings. However, this requires optimization over all of the available transforms to minimize the overall bitrate, which can be expensive in terms of time and computation. In this thesis, we examine the use of only the horizontal and vertical 1D-DCTs in addition to the 2D-DCT for coding motion compensation residuals. By reducing the number of available transforms, the amount of required computation decreases significantly, with a potential cost in performance. We perform experiments using a modified H.264/AVC codec to compare the performance of using different sets of available transforms. The results indicate that for typical applications of video coding, most of the performance benefit from using directional 1D-DCTs can be retained by keeping only the horizontal and vertical 1D-DCTs.","abstract_html":"In video coding, motion compensation prediction provides significant increases in overall compression efficiency. The prediction residuals are typically treated as images and compressed by applying two-dimensional transforms such as the two-dimensional discrete cosine transform (2D-DCT). Previous work has found that the use of direction-adaptive one-dimensional discrete cosine transforms (1D-DCTs) in coding motion compensation residuals can provide significant additional bitrate savings. However, this requires optimization over all of the available transforms to minimize the overall bitrate, which can be expensive in terms of time and computation. In this thesis, we examine the use of only the horizontal and vertical 1D-DCTs in addition to the 2D-DCT for coding motion compensation residuals. By reducing the number of available transforms, the amount of required computation decreases significantly, with a potential cost in performance. We perform experiments using a modified H.264/AVC codec to compare the performance of using different sets of available transforms. The results indicate that for typical applications of video coding, most of the performance benefit from using directional 1D-DCTs can be retained by keeping only the horizontal and vertical 1D-DCTs.","abstract_has_math":false,"creators":["Zhang, Harley (Harley H.)"],"institution":"Massachusetts Institute of Technology","degree_name":null,"degree_level":null,"degree_discipline":null,"degree_department":"Massachusetts Institute of Technology. Dept. of Electrical Engineering and Computer Science.","school":null,"contributors":[],"advisors":["Jae S. Lim."],"committee_chairs":[],"committee_members":[],"year":2011,"date_issued":"2011","date_published":"2011","updated_at":"2026-07-22T22:22:26Z","subjects":["Electrical Engineering and Computer Science."],"languages":["eng"],"rights":["M.I.T. theses are protected by copyright. They may be viewed from this source for any purpose, but reproduction or distribution in any format is prohibited without written permission. See provided URL for inquiries about permission."],"rights_urls":["http://dspace.mit.edu/handle/1721.1/7582"],"identifier_entries":[]},"links":{"outbound_url":"http://hdl.handle.net/1721.1/66707","outbound_label":"Handle","outbound_source":"dc:identifier.uri"},"metadata_groups":[{"id":"people","label":"People","entries":[{"key":"dc:contributor.advisor","label":"Advisor","values":["Jae S. Lim."]},{"key":"dc:contributor.department","label":"Department","values":["Massachusetts Institute of Technology. Dept. of Electrical Engineering and Computer Science."]},{"key":"dc:contributor.other","label":"Dc Contributor Other","values":["Massachusetts Institute of Technology. Dept. of Electrical Engineering and Computer Science."]},{"key":"dc:creator","label":"Author","values":["Zhang, Harley (Harley H.)"]}]},{"id":"academic_context","label":"Academic Context","entries":[{"key":"dc:date.accessioned","label":"Dc Date Accessioned","values":["2011-11-01T18:06:02Z"]},{"key":"dc:date.available","label":"Dc Date Available","values":["2011-11-01T18:06:02Z"]},{"key":"dc:date.issued","label":"Date","values":["2011"]},{"key":"dc:publisher","label":"Institution","values":["Massachusetts Institute of Technology"]},{"key":"dc:type","label":"Dc Type","values":["Thesis"]}]},{"id":"subjects_keywords","label":"Subjects and Keywords","entries":[{"key":"dc:subject","label":"Dc Subject","values":["Electrical Engineering and Computer Science."]}]},{"id":"language_rights","label":"Language and Rights","entries":[{"key":"dc:language.iso","label":"Language (ISO)","values":["eng"]},{"key":"dc:rights","label":"Dc Rights","values":["M.I.T. theses are protected by copyright. They may be viewed from this source for any purpose, but reproduction or distribution in any format is prohibited without written permission. See provided URL for inquiries about permission."]},{"key":"dc:rights.uri","label":"Rights URI","values":["http://dspace.mit.edu/handle/1721.1/7582"]}]},{"id":"identifiers","label":"Identifiers","entries":[{"key":"dc:identifier.uri","label":"Identifier URI","values":["http://hdl.handle.net/1721.1/66707"]}]},{"id":"additional","label":"Additional Metadata","entries":[{"key":"dc:description","label":"Description","values":["Thesis (M. Eng.)--Massachusetts Institute of Technology, Dept. of Electrical Engineering and Computer Science, 2011.","This electronic version was submitted by the student author. The certified thesis is available in the Institute Archives and Special Collections.","Cataloged from student submitted PDF version of thesis.","Includes bibliographical references (p. 49)."]},{"key":"dc:description.abstract","label":"Abstract","values":["In video coding, motion compensation prediction provides significant increases in overall compression efficiency. The prediction residuals are typically treated as images and compressed by applying two-dimensional transforms such as the two-dimensional discrete cosine transform (2D-DCT). Previous work has found that the use of direction-adaptive one-dimensional discrete cosine transforms (1D-DCTs) in coding motion compensation residuals can provide significant additional bitrate savings. However, this requires optimization over all of the available transforms to minimize the overall bitrate, which can be expensive in terms of time and computation. In this thesis, we examine the use of only the horizontal and vertical 1D-DCTs in addition to the 2D-DCT for coding motion compensation residuals. By reducing the number of available transforms, the amount of required computation decreases significantly, with a potential cost in performance. We perform experiments using a modified H.264/AVC codec to compare the performance of using different sets of available transforms. The results indicate that for typical applications of video coding, most of the performance benefit from using directional 1D-DCTs can be retained by keeping only the horizontal and vertical 1D-DCTs."]},{"key":"dc:description.degree","label":"Dc Description Degree","values":["M.Eng."]},{"key":"dc:title","label":"Title","values":["Analysis of one-dimensional transforms in coding motion compensation prediction residuals for video applications"]}]}],"canonical_facts":{"dc:contributor.advisor":["Jae S. Lim."],"dc:contributor.department":["Massachusetts Institute of Technology. Dept. of Electrical Engineering and Computer Science."],"dc:contributor.other":["Massachusetts Institute of Technology. Dept. of Electrical Engineering and Computer Science."],"dc:creator":["Zhang, Harley (Harley H.)"],"dc:date.accessioned":["2011-11-01T18:06:02Z"],"dc:date.available":["2011-11-01T18:06:02Z"],"dc:date.issued":["2011"],"dc:description":["Thesis (M. Eng.)--Massachusetts Institute of Technology, Dept. of Electrical Engineering and Computer Science, 2011.","This electronic version was submitted by the student author. The certified thesis is available in the Institute Archives and Special Collections.","Cataloged from student submitted PDF version of thesis.","Includes bibliographical references (p. 49)."],"dc:description.abstract":["In video coding, motion compensation prediction provides significant increases in overall compression efficiency. The prediction residuals are typically treated as images and compressed by applying two-dimensional transforms such as the two-dimensional discrete cosine transform (2D-DCT). Previous work has found that the use of direction-adaptive one-dimensional discrete cosine transforms (1D-DCTs) in coding motion compensation residuals can provide significant additional bitrate savings. However, this requires optimization over all of the available transforms to minimize the overall bitrate, which can be expensive in terms of time and computation. In this thesis, we examine the use of only the horizontal and vertical 1D-DCTs in addition to the 2D-DCT for coding motion compensation residuals. By reducing the number of available transforms, the amount of required computation decreases significantly, with a potential cost in performance. We perform experiments using a modified H.264/AVC codec to compare the performance of using different sets of available transforms. The results indicate that for typical applications of video coding, most of the performance benefit from using directional 1D-DCTs can be retained by keeping only the horizontal and vertical 1D-DCTs."],"dc:description.degree":["M.Eng."],"dc:identifier.uri":["http://hdl.handle.net/1721.1/66707"],"dc:language.iso":["eng"],"dc:publisher":["Massachusetts Institute of Technology"],"dc:rights":["M.I.T. theses are protected by copyright. They may be viewed from this source for any purpose, but reproduction or distribution in any format is prohibited without written permission. See provided URL for inquiries about permission."],"dc:rights.uri":["http://dspace.mit.edu/handle/1721.1/7582"],"dc:subject":["Electrical Engineering and Computer Science."],"dc:title":["Analysis of one-dimensional transforms in coding motion compensation prediction residuals for video applications"],"dc:type":["Thesis"]},"updated_at":"2026-07-22T22:22:26Z"}