{"id":{"repo_id":"uiuc","oai_identifier":"oai:www.ideals.illinois.edu:2142/16720"},"canonical_url":"https://search.dev.ndltd.org/etd/uiuc/oai:www.ideals.illinois.edu:2142/16720","repository":{"repo_id":"uiuc","name":"University of Illinois - Urbana-Champaign","base_url":"https://www.ideals.illinois.edu/oai-pmh"},"display":{"title":"Database support for top-down proteomics","abstract":"Top-down proteomics is a revolutionary application for the identification and characterization of protein, known to be one of the most complicated and challenging issues in biology. In top-down proteomics, the quality and speed of the data warehouse is very important, as high accuracy results are returned by a database search. ProSight Warehouse fills the critical role as the data warehouse for ProSight PTM, the first publicly available top-down proteomics software suite. MySQL, a free relational database, was the base of this warehouse. Many annotated and predicted protein forms have been successfully incorporated into the organism-specific database and in the integrated database for human strains. To achieve high quality and efficiency, a database schema (Absolute Mass Search), data annotation methods (Shotgun and Extended Shotgun Annotation), data population strategies (on-the-fly population, bulk-loading method), and a database integration methodology for human protein were developed. With the successful implementation of ProSight Warehouse, ProSight PTM achieved its aspiration, highly accurate protein identification and characterization.","abstract_html":"Top-down proteomics is a revolutionary application for the identification and characterization of protein, known to be one of the most complicated and challenging issues in biology. In top-down proteomics, the quality and speed of the data warehouse is very important, as high accuracy results are returned by a database search. ProSight Warehouse fills the critical role as the data warehouse for ProSight PTM, the first publicly available top-down proteomics software suite. MySQL, a free relational database, was the base of this warehouse. Many annotated and predicted protein forms have been successfully incorporated into the organism-specific database and in the integrated database for human strains. To achieve high quality and efficiency, a database schema (Absolute Mass Search), data annotation methods (Shotgun and Extended Shotgun Annotation), data population strategies (on-the-fly population, bulk-loading method), and a database integration methodology for human protein were developed. With the successful implementation of ProSight Warehouse, ProSight PTM achieved its aspiration, highly accurate protein identification and characterization.","abstract_has_math":false,"creators":["Kim, Yong-Bin"],"institution":"University of Illinois at Urbana-Champaign","degree_name":"Ph.D.","degree_level":"Dissertation","degree_discipline":"Computer Science","degree_department":null,"school":null,"contributors":["Belford, Geneva","Belford, Geneva G.","Kelleher, Neil L.","Han, Jiawei","Zhai, ChengXiang"],"advisors":[],"committee_chairs":[],"committee_members":[],"year":2010,"date_issued":"2010-08-20T17:55:52Z","date_published":"2010-08-20T17:55:52Z","updated_at":"2026-07-22T22:25:09Z","subjects":["Top-down proteomics","Database support for proteomics","Prosight PTM","Prosight Warehouse","Biological warehouse integration","Data warehouse"],"languages":["en"],"rights":["Copyright 2010 Yong-Bin Kim"],"rights_urls":[],"identifier_entries":[]},"links":{"outbound_url":"http://hdl.handle.net/2142/16720","outbound_label":"Handle","outbound_source":"dc:identifier"},"metadata_groups":[{"id":"people","label":"People","entries":[{"key":"dc:contributor","label":"Contributor","values":["Belford, Geneva","Belford, Geneva G.","Kelleher, Neil L.","Han, Jiawei","Zhai, ChengXiang"]},{"key":"dc:creator","label":"Author","values":["Kim, Yong-Bin"]}]},{"id":"academic_context","label":"Academic Context","entries":[{"key":"dc:date","label":"Dc Date","values":["2010-08-20T17:55:52Z","2010-08"]},{"key":"thesis:degree_discipline","label":"Discipline","values":["Computer Science"]},{"key":"thesis:degree_level","label":"Degree Level","values":["Dissertation"]},{"key":"thesis:degree_name","label":"Degree Name","values":["Ph.D."]},{"key":"thesis:institution_name","label":"Thesis Institution Name","values":["University of Illinois at Urbana-Champaign"]}]},{"id":"subjects_keywords","label":"Subjects and Keywords","entries":[{"key":"dc:subject","label":"Dc Subject","values":["Top-down proteomics","Database support for proteomics","Prosight PTM","Prosight Warehouse","Biological warehouse integration","Data warehouse"]}]},{"id":"language_rights","label":"Language and Rights","entries":[{"key":"dc:language","label":"Dc Language","values":["en"]},{"key":"dc:rights","label":"Dc Rights","values":["Copyright 2010 Yong-Bin Kim"]}]},{"id":"identifiers","label":"Identifiers","entries":[{"key":"dc:identifier","label":"Identifier","values":["http://hdl.handle.net/2142/16720"]}]},{"id":"additional","label":"Additional Metadata","entries":[{"key":"dc:description","label":"Description","values":["Top-down proteomics is a revolutionary application for the identification and characterization of protein, known to be one of the most complicated and challenging issues in biology. In top-down proteomics, the quality and speed of the data warehouse is very important, as high accuracy results are returned by a database search. ProSight Warehouse fills the critical role as the data warehouse for ProSight PTM, the first publicly available top-down proteomics software suite. MySQL, a free relational database, was the base of this warehouse. Many annotated and predicted protein forms have been successfully incorporated into the organism-specific database and in the integrated database for human strains. To achieve high quality and efficiency, a database schema (Absolute Mass Search), data annotation methods (Shotgun and Extended Shotgun Annotation), data population strategies (on-the-fly population, bulk-loading method), and a database integration methodology for human protein were developed. With the successful implementation of ProSight Warehouse, ProSight PTM achieved its aspiration, highly accurate protein identification and characterization.","Item withdrawn by Mark Zulauf (zulauf@illinois.edu) on 2010-07-12T20:45:20Z Item was in collections: University of Illinois Theses & Dissertations (ID: 1) No. of bitstreams: 1 Kim_Yong-Bin.pdf: 3028764 bytes, checksum: 357c9d473a8aee82ad6792d262611222 (MD5)","Made available in DSpace on 2010-08-20T17:55:52Z (GMT). No. of bitstreams: 3 Kim_Yong-Bin.pdf: 3028764 bytes, checksum: 357c9d473a8aee82ad6792d262611222 (MD5) 1_Kim_Yong-Bin.pdf: 3028771 bytes, checksum: 2c456fde326686e4e86b12a5c74d2cd8 (MD5) license.txt: 4055 bytes, checksum: f0093863896a06144826951bcbf44042 (MD5)"]},{"key":"dc:title","label":"Title","values":["Database support for top-down proteomics"]}]}],"canonical_facts":{"dc:contributor":["Belford, Geneva","Belford, Geneva G.","Kelleher, Neil L.","Han, Jiawei","Zhai, ChengXiang"],"dc:creator":["Kim, Yong-Bin"],"dc:date":["2010-08-20T17:55:52Z","2010-08"],"dc:description":["Top-down proteomics is a revolutionary application for the identification and characterization of protein, known to be one of the most complicated and challenging issues in biology. In top-down proteomics, the quality and speed of the data warehouse is very important, as high accuracy results are returned by a database search. ProSight Warehouse fills the critical role as the data warehouse for ProSight PTM, the first publicly available top-down proteomics software suite. MySQL, a free relational database, was the base of this warehouse. Many annotated and predicted protein forms have been successfully incorporated into the organism-specific database and in the integrated database for human strains. To achieve high quality and efficiency, a database schema (Absolute Mass Search), data annotation methods (Shotgun and Extended Shotgun Annotation), data population strategies (on-the-fly population, bulk-loading method), and a database integration methodology for human protein were developed. With the successful implementation of ProSight Warehouse, ProSight PTM achieved its aspiration, highly accurate protein identification and characterization.","Item withdrawn by Mark Zulauf (zulauf@illinois.edu) on 2010-07-12T20:45:20Z Item was in collections: University of Illinois Theses & Dissertations (ID: 1) No. of bitstreams: 1 Kim_Yong-Bin.pdf: 3028764 bytes, checksum: 357c9d473a8aee82ad6792d262611222 (MD5)","Made available in DSpace on 2010-08-20T17:55:52Z (GMT). No. of bitstreams: 3 Kim_Yong-Bin.pdf: 3028764 bytes, checksum: 357c9d473a8aee82ad6792d262611222 (MD5) 1_Kim_Yong-Bin.pdf: 3028771 bytes, checksum: 2c456fde326686e4e86b12a5c74d2cd8 (MD5) license.txt: 4055 bytes, checksum: f0093863896a06144826951bcbf44042 (MD5)"],"dc:identifier":["http://hdl.handle.net/2142/16720"],"dc:language":["en"],"dc:rights":["Copyright 2010 Yong-Bin Kim"],"dc:subject":["Top-down proteomics","Database support for proteomics","Prosight PTM","Prosight Warehouse","Biological warehouse integration","Data warehouse"],"dc:title":["Database support for top-down proteomics"],"thesis:degree_discipline":["Computer Science"],"thesis:degree_level":["Dissertation"],"thesis:degree_name":["Ph.D."],"thesis:institution_name":["University of Illinois at Urbana-Champaign"]},"updated_at":"2026-07-22T22:25:09Z"}