{"id":{"repo_id":"nus","oai_identifier":"oai:scholarbank.nus.edu.sg:10635/14885"},"canonical_url":"https://search.dev.ndltd.org/etd/nus/oai:scholarbank.nus.edu.sg:10635/14885","repository":{"repo_id":"nus","name":"National University of Singapore","base_url":"https://scholarbank.nus.edu.sg/oai/request"},"display":{"title":"Semantic concept detection from visual content with statistical learning","abstract":"This thesis addresses semantic concept detection from visual content with statistical learning methods. The highest level semantic concept is genre, the lowest one is object. Accordingly, two parts of research work have been conducted, namely sports news video genre identification and automatic image annotation. For the former, two novel features were proposed to classify sports news video shots; due to the variation of content and shot length, this problem is high challenging. For the latter, we proposed a novel automatic image annotation framework and achieved promising results which outperform the state of art works in two famous dataset: CorelCD and TRECVID2003. Our contributions can be summarized from two aspects: first, proposed a novel image representation scheme with which an image can be treated as a text document, so many text document techniques can be employed; second, proposed two flexible information fusion methods for fusing diverse visual features and multiple modalities.","abstract_html":"This thesis addresses semantic concept detection from visual content with statistical learning methods. The highest level semantic concept is genre, the lowest one is object. Accordingly, two parts of research work have been conducted, namely sports news video genre identification and automatic image annotation. For the former, two novel features were proposed to classify sports news video shots; due to the variation of content and shot length, this problem is high challenging. For the latter, we proposed a novel automatic image annotation framework and achieved promising results which outperform the state of art works in two famous dataset: CorelCD and TRECVID2003. Our contributions can be summarized from two aspects: first, proposed a novel image representation scheme with which an image can be treated as a text document, so many text document techniques can be employed; second, proposed two flexible information fusion methods for fusing diverse visual features and multiple modalities.","abstract_has_math":false,"creators":["WANG DEHONG"],"institution":null,"degree_name":null,"degree_level":null,"degree_discipline":null,"degree_department":null,"school":null,"contributors":[],"advisors":[],"committee_chairs":[],"committee_members":[],"year":2005,"date_issued":"2005-12-06","date_published":"2005-12-06","updated_at":"2026-07-24T03:30:47Z","subjects":["semantic concept, genre identification, image annotation, visual features, information fusion, statistical learning"],"languages":[],"rights":[],"rights_urls":[],"identifier_entries":[]},"links":{"outbound_url":null,"outbound_label":null,"outbound_source":null},"metadata_groups":[{"id":"people","label":"People","entries":[{"key":"dc:creator","label":"Author","values":["WANG DEHONG"]}]},{"id":"academic_context","label":"Academic Context","entries":[{"key":"dc:date.issued","label":"Date","values":["2005-12-06"]},{"key":"dc:relation.isreferencedby","label":"Dc Relation Isreferencedby","values":["https://scholarbank.nus.edu.sg/handle/10635/14885"]},{"key":"dc:type","label":"Dc Type","values":["Thesis"]}]},{"id":"subjects_keywords","label":"Subjects and Keywords","entries":[{"key":"dc:subject","label":"Dc Subject","values":["semantic concept, genre identification, image annotation, visual features, information fusion, statistical learning"]}]},{"id":"identifiers","label":"Identifiers","entries":[{"key":"dc:identifier.uri","label":"Identifier URI","values":["https://scholarbank.nus.edu.sg/bitstreams/7b9ee493-8fbc-450f-9194-df8dfa4a543f/download"]}]},{"id":"additional","label":"Additional Metadata","entries":[{"key":"dc:description.abstract","label":"Abstract","values":["This thesis addresses semantic concept detection from visual content with statistical learning methods. The highest level semantic concept is genre, the lowest one is object. Accordingly, two parts of research work have been conducted, namely sports news video genre identification and automatic image annotation. For the former, two novel features were proposed to classify sports news video shots; due to the variation of content and shot length, this problem is high challenging. For the latter, we proposed a novel automatic image annotation framework and achieved promising results which outperform the state of art works in two famous dataset: CorelCD and TRECVID2003. Our contributions can be summarized from two aspects: first, proposed a novel image representation scheme with which an image can be treated as a text document, so many text document techniques can be employed; second, proposed two flexible information fusion methods for fusing diverse visual features and multiple modalities."]},{"key":"dc:format.checksum.md5","label":"Dc Format Checksum Md5","values":["8085f641bebeab4432bf6f0cd2fe7b36","8d400e17d581849a42fefa536e37c851"]},{"key":"dc:title","label":"Title","values":["Semantic concept detection from visual content with statistical learning"]}]}],"canonical_facts":{"dc:creator":["WANG DEHONG"],"dc:date.issued":["2005-12-06"],"dc:description.abstract":["This thesis addresses semantic concept detection from visual content with statistical learning methods. The highest level semantic concept is genre, the lowest one is object. Accordingly, two parts of research work have been conducted, namely sports news video genre identification and automatic image annotation. For the former, two novel features were proposed to classify sports news video shots; due to the variation of content and shot length, this problem is high challenging. For the latter, we proposed a novel automatic image annotation framework and achieved promising results which outperform the state of art works in two famous dataset: CorelCD and TRECVID2003. Our contributions can be summarized from two aspects: first, proposed a novel image representation scheme with which an image can be treated as a text document, so many text document techniques can be employed; second, proposed two flexible information fusion methods for fusing diverse visual features and multiple modalities."],"dc:format.checksum.md5":["8085f641bebeab4432bf6f0cd2fe7b36","8d400e17d581849a42fefa536e37c851"],"dc:identifier.uri":["https://scholarbank.nus.edu.sg/bitstreams/7b9ee493-8fbc-450f-9194-df8dfa4a543f/download"],"dc:relation.isreferencedby":["https://scholarbank.nus.edu.sg/handle/10635/14885"],"dc:subject":["semantic concept, genre identification, image annotation, visual features, information fusion, statistical learning"],"dc:title":["Semantic concept detection from visual content with statistical learning"],"dc:type":["Thesis"]},"updated_at":"2026-07-24T03:30:47Z"}