{"id":{"repo_id":"uiuc","oai_identifier":"oai:www.ideals.illinois.edu:2142/46741"},"canonical_url":"https://search.dev.ndltd.org/etd/uiuc/oai:www.ideals.illinois.edu:2142/46741","repository":{"repo_id":"uiuc","name":"University of Illinois - Urbana-Champaign","base_url":"https://www.ideals.illinois.edu/oai-pmh"},"display":{"title":"A study of automatic email routing for an information technology help desk","abstract":"Document classification has been a classic problem in both machine learning and information retrieval. One domain for document classification is automatic email routing. Given an email (a document), the system attempts to guess the location that the email should be routed to. An automatic system would in theory be able to replace a person doing the job of sorting emails, which can save time and money. However, incorrectly sorted emails would then need to be re-sorted manually, so it is important for the system to be accurate. The Engineering IT department at the University of Illinois at Urbana-Champaign has a helpdesk that users can email with technical problems. The IT department services the entire College of Engineering, encompassing many departments, so they need a way to sort emails based on which IT professional would be best suited to fixing a user’s problem. Thus, this problem would be well suited for some type of document classification solution. However, the helpdesk stands out from similar problems; the IT department already employs an email routing technique that is already quite accurate. It becomes very obvious that a stand-alone document classification solution would be subpar, but perhaps combining it with the existing routing method would provide higher accuracy. In this thesis, we explore ways to combine the classic document classification techniques with the existing routing strategy used by the helpdesk. We test out different text-based features, but we find that since the existing method is already very accurate, it is quite difficult to improve on.","abstract_html":"Document classification has been a classic problem in both machine learning and information retrieval. One domain for document classification is automatic email routing. Given an email (a document), the system attempts to guess the location that the email should be routed to. An automatic system would in theory be able to replace a person doing the job of sorting emails, which can save time and money. However, incorrectly sorted emails would then need to be re-sorted manually, so it is important for the system to be accurate. The Engineering IT department at the University of Illinois at Urbana-Champaign has a helpdesk that users can email with technical problems. The IT department services the entire College of Engineering, encompassing many departments, so they need a way to sort emails based on which IT professional would be best suited to fixing a user’s problem. Thus, this problem would be well suited for some type of document classification solution. However, the helpdesk stands out from similar problems; the IT department already employs an email routing technique that is already quite accurate. It becomes very obvious that a stand-alone document classification solution would be subpar, but perhaps combining it with the existing routing method would provide higher accuracy. In this thesis, we explore ways to combine the classic document classification techniques with the existing routing strategy used by the helpdesk. We test out different text-based features, but we find that since the existing method is already very accurate, it is quite difficult to improve on.","abstract_has_math":false,"creators":["Di Febo, Joseph"],"institution":"University of Illinois at Urbana-Champaign","degree_name":"M.S.","degree_level":"Thesis","degree_discipline":"Computer Science","degree_department":null,"school":null,"contributors":["Zhai, ChengXiang"],"advisors":[],"committee_chairs":[],"committee_members":[],"year":2014,"date_issued":"2014-01-16T18:00:52Z","date_published":"2014-01-16T18:00:52Z","updated_at":"2026-07-22T22:25:36Z","subjects":["document classification","email routing","information retrieval"],"languages":["en"],"rights":["Copyright 2013 Joseph Di Febo"],"rights_urls":[],"identifier_entries":[]},"links":{"outbound_url":"http://hdl.handle.net/2142/46741","outbound_label":"Handle","outbound_source":"dc:identifier"},"metadata_groups":[{"id":"people","label":"People","entries":[{"key":"dc:contributor","label":"Contributor","values":["Zhai, ChengXiang"]},{"key":"dc:creator","label":"Author","values":["Di Febo, Joseph"]}]},{"id":"academic_context","label":"Academic Context","entries":[{"key":"dc:date","label":"Dc Date","values":["2014-01-16T18:00:52Z","2013-12"]},{"key":"dc:type","label":"Dc Type","values":["text"]},{"key":"thesis:degree_discipline","label":"Discipline","values":["Computer Science"]},{"key":"thesis:degree_level","label":"Degree Level","values":["Thesis"]},{"key":"thesis:degree_name","label":"Degree Name","values":["M.S."]},{"key":"thesis:institution_name","label":"Thesis Institution Name","values":["University of Illinois at Urbana-Champaign"]}]},{"id":"subjects_keywords","label":"Subjects and Keywords","entries":[{"key":"dc:subject","label":"Dc Subject","values":["document classification","email routing","information retrieval"]}]},{"id":"language_rights","label":"Language and Rights","entries":[{"key":"dc:language","label":"Dc Language","values":["en"]},{"key":"dc:rights","label":"Dc Rights","values":["Copyright 2013 Joseph Di Febo"]}]},{"id":"identifiers","label":"Identifiers","entries":[{"key":"dc:identifier","label":"Identifier","values":["http://hdl.handle.net/2142/46741"]}]},{"id":"additional","label":"Additional Metadata","entries":[{"key":"dc:description","label":"Description","values":["Document classification has been a classic problem in both machine learning and information retrieval. One domain for document classification is automatic email routing. Given an email (a document), the system attempts to guess the location that the email should be routed to. An automatic system would in theory be able to replace a person doing the job of sorting emails, which can save time and money. However, incorrectly sorted emails would then need to be re-sorted manually, so it is important for the system to be accurate. The Engineering IT department at the University of Illinois at Urbana-Champaign has a helpdesk that users can email with technical problems. The IT department services the entire College of Engineering, encompassing many departments, so they need a way to sort emails based on which IT professional would be best suited to fixing a user’s problem. Thus, this problem would be well suited for some type of document classification solution. However, the helpdesk stands out from similar problems; the IT department already employs an email routing technique that is already quite accurate. It becomes very obvious that a stand-alone document classification solution would be subpar, but perhaps combining it with the existing routing method would provide higher accuracy. In this thesis, we explore ways to combine the classic document classification techniques with the existing routing strategy used by the helpdesk. We test out different text-based features, but we find that since the existing method is already very accurate, it is quite difficult to improve on.","Item withdrawn by Laura Spradlin (lspradl2@illinois.edu) on 2013-12-11T17:58:20Z Item was in collections: University of Illinois Theses & Dissertations (ID: 1) No. of bitstreams: 2 difebo_joseph.docx: 39263 bytes, checksum: 051b28ea3e4dc93d559ae8b55e55831e (MD5) difebo_joseph.pdf: 333453 bytes, checksum: b74f9f15c7c2d213e5645613d1f46476 (MD5)","Made available in DSpace on 2014-01-16T18:00:52Z (GMT). No. of bitstreams: 3 Joseph_Di Febo.pdf: 333536 bytes, checksum: 9e88cd600da79417b9627debb3359b4b (MD5) difebo_joseph.docx: 39292 bytes, checksum: 8f28f508ec535ff369dfb7fa62aa2dab (MD5) license.txt: 4063 bytes, checksum: 9bb72955e81e37536fa899493bda911d (MD5)"]},{"key":"dc:title","label":"Title","values":["A study of automatic email routing for an information technology help desk"]}]}],"canonical_facts":{"dc:contributor":["Zhai, ChengXiang"],"dc:creator":["Di Febo, Joseph"],"dc:date":["2014-01-16T18:00:52Z","2013-12"],"dc:description":["Document classification has been a classic problem in both machine learning and information retrieval. One domain for document classification is automatic email routing. Given an email (a document), the system attempts to guess the location that the email should be routed to. An automatic system would in theory be able to replace a person doing the job of sorting emails, which can save time and money. However, incorrectly sorted emails would then need to be re-sorted manually, so it is important for the system to be accurate. The Engineering IT department at the University of Illinois at Urbana-Champaign has a helpdesk that users can email with technical problems. The IT department services the entire College of Engineering, encompassing many departments, so they need a way to sort emails based on which IT professional would be best suited to fixing a user’s problem. Thus, this problem would be well suited for some type of document classification solution. However, the helpdesk stands out from similar problems; the IT department already employs an email routing technique that is already quite accurate. It becomes very obvious that a stand-alone document classification solution would be subpar, but perhaps combining it with the existing routing method would provide higher accuracy. In this thesis, we explore ways to combine the classic document classification techniques with the existing routing strategy used by the helpdesk. We test out different text-based features, but we find that since the existing method is already very accurate, it is quite difficult to improve on.","Item withdrawn by Laura Spradlin (lspradl2@illinois.edu) on 2013-12-11T17:58:20Z Item was in collections: University of Illinois Theses & Dissertations (ID: 1) No. of bitstreams: 2 difebo_joseph.docx: 39263 bytes, checksum: 051b28ea3e4dc93d559ae8b55e55831e (MD5) difebo_joseph.pdf: 333453 bytes, checksum: b74f9f15c7c2d213e5645613d1f46476 (MD5)","Made available in DSpace on 2014-01-16T18:00:52Z (GMT). No. of bitstreams: 3 Joseph_Di Febo.pdf: 333536 bytes, checksum: 9e88cd600da79417b9627debb3359b4b (MD5) difebo_joseph.docx: 39292 bytes, checksum: 8f28f508ec535ff369dfb7fa62aa2dab (MD5) license.txt: 4063 bytes, checksum: 9bb72955e81e37536fa899493bda911d (MD5)"],"dc:identifier":["http://hdl.handle.net/2142/46741"],"dc:language":["en"],"dc:rights":["Copyright 2013 Joseph Di Febo"],"dc:subject":["document classification","email routing","information retrieval"],"dc:title":["A study of automatic email routing for an information technology help desk"],"dc:type":["text"],"thesis:degree_discipline":["Computer Science"],"thesis:degree_level":["Thesis"],"thesis:degree_name":["M.S."],"thesis:institution_name":["University of Illinois at Urbana-Champaign"]},"updated_at":"2026-07-22T22:25:36Z"}