{"id":{"repo_id":"regina","oai_identifier":"oai:uregina.scholaris.ca:10294/9255"},"canonical_url":"https://search.dev.ndltd.org/etd/regina/oai:uregina.scholaris.ca:10294/9255","repository":{"repo_id":"regina","name":"University of Regina","base_url":"https://uregina.scholaris.ca/server/oai/request"},"display":{"title":"An Approach for Building and Querying Linked Datasets","abstract":"The World Wide Web (WWW) was first envisioned as a global web of linked documents and it is now becoming a global web of linked data. The World Wide Web Consortium (W3C) uses the term semantic web1 to refer to this new vision. This thesis is focused on building linked datasets from data which may not be fully accessible on the web. Once transformed into linked datasets, the public is more able to use and query the data. Various semantic web technologies for handling data are used in this thesis. These technologies includeWeb Ontology Language (OWL) for describing the semantics of the information, Resource Description Framework (RDF) for expressing the relationships in the information, and Simple Protocol and RDF Query Language (SPARQL) for querying the information used in web documents. In 2005, Sterling [55] described an object whose “informational support is so overwhelmingly extensive and rich that it is regarded as a material instantiation of an immaterial system”. He called this object a spime, a term that he coined by combining the words space and time. Spimes about particular items can collect linked information, which is shared on the web by citizen-consumers. Though Sterling used examples from the physical world, the semantic web may be the perfect environment for spimes to thrive. Hepting et al. [33] described the design of an informational spime, using an example of chicken eggs. This thesis helps to realize that design with a proof of concept for a spime with chicken egg information. The proof of concept was done using existing, 1https://www.w3.org/standards/semanticweb/ primarily open source, tools that allow for: extraction and remediation of data on the web, addition of data not already on the web, and querying of all that data. As a starting point for this thesis, information about eggs was found on a webpage marked up as a table in Hypertext Markup Language (HTML), which provided structure but no semantics. The approach including the following steps: extracting egg data from an existing web source2; developing an egg ontology; publishing the egg information as linked data; connecting this linked data to external sources such as Wikidata3; and querying the published linked data with SPARQL. The egg spime illustrated in the thesis contains information about eggs available at a specific place and time. Citizen-consumers in Regina, Canada and Patiala, India wanting to buy eggs could query their own local egg spimes, which may differ substantially in structure and content from one other. This thesis contributes to the realization of Berners-Lee’s vision for the semantic web. There are several opportunities for future work, including the addition of geospatial information and improved interface functionalities.","abstract_html":"The World Wide Web (WWW) was first envisioned as a global web of linked documents and it is now becoming a global web of linked data. The World Wide Web Consortium (W3C) uses the term semantic web1 to refer to this new vision. This thesis is focused on building linked datasets from data which may not be fully accessible on the web. Once transformed into linked datasets, the public is more able to use and query the data. Various semantic web technologies for handling data are used in this thesis. These technologies includeWeb Ontology Language (OWL) for describing the semantics of the information, Resource Description Framework (RDF) for expressing the relationships in the information, and Simple Protocol and RDF Query Language (SPARQL) for querying the information used in web documents. In 2005, Sterling [55] described an object whose “informational support is so overwhelmingly extensive and rich that it is regarded as a material instantiation of an immaterial system”. He called this object a spime, a term that he coined by combining the words space and time. Spimes about particular items can collect linked information, which is shared on the web by citizen-consumers. Though Sterling used examples from the physical world, the semantic web may be the perfect environment for spimes to thrive. Hepting et al. [33] described the design of an informational spime, using an example of chicken eggs. This thesis helps to realize that design with a proof of concept for a spime with chicken egg information. The proof of concept was done using existing, 1https://www.w3.org/standards/semanticweb/ primarily open source, tools that allow for: extraction and remediation of data on the web, addition of data not already on the web, and querying of all that data. As a starting point for this thesis, information about eggs was found on a webpage marked up as a table in Hypertext Markup Language (HTML), which provided structure but no semantics. The approach including the following steps: extracting egg data from an existing web source2; developing an egg ontology; publishing the egg information as linked data; connecting this linked data to external sources such as Wikidata3; and querying the published linked data with SPARQL. The egg spime illustrated in the thesis contains information about eggs available at a specific place and time. Citizen-consumers in Regina, Canada and Patiala, India wanting to buy eggs could query their own local egg spimes, which may differ substantially in structure and content from one other. This thesis contributes to the realization of Berners-Lee’s vision for the semantic web. There are several opportunities for future work, including the addition of geospatial information and improved interface functionalities.","abstract_has_math":false,"creators":["Kaur, Harmanjeet"],"institution":"Faculty of Graduate Studies and Research, University of Regina","degree_name":"Master of Science (MSc)","degree_level":"Master&apos;s","degree_discipline":"Computer Science","degree_department":null,"school":null,"contributors":[],"advisors":["Hepting, Daryl"],"committee_chairs":[],"committee_members":["Yao, Yiyu","Yang, Xue-Dong"],"year":2019,"date_issued":"2019-12","date_published":"2019-12","updated_at":"2026-07-24T04:03:38Z","subjects":[],"languages":["en"],"rights":[],"rights_urls":[],"identifier_entries":[{"key":"dc:identifier.doi","label":"DOI","values":["https://doi.org/10.82465/4394"],"render_values":[{"text":"https://doi.org/10.82465/4394","href":"https://doi.org/10.82465/4394","code":true}]}]},"links":{"outbound_url":"https://hdl.handle.net/10294/9255","outbound_label":"Handle","outbound_source":"dc:identifier.uri"},"metadata_groups":[{"id":"people","label":"People","entries":[{"key":"dc:contributor.advisor","label":"Advisor","values":["Hepting, Daryl"]},{"key":"dc:contributor.committeemember","label":"Committee Member","values":["Yao, Yiyu","Yang, Xue-Dong"]},{"key":"dc:creator","label":"Author","values":["Kaur, Harmanjeet"]}]},{"id":"academic_context","label":"Academic Context","entries":[{"key":"dc:date.accessioned","label":"Dc Date Accessioned","values":["2020-08-29T19:04:44Z"]},{"key":"dc:date.available","label":"Dc Date Available","values":["2020-08-29T19:04:44Z"]},{"key":"dc:date.issued","label":"Date","values":["2019-12"]},{"key":"dc:publisher","label":"Institution","values":["Faculty of Graduate Studies and Research, University of Regina"]},{"key":"dc:type","label":"Dc Type","values":["master thesis"]},{"key":"thesis:degree_discipline","label":"Discipline","values":["Computer Science"]},{"key":"thesis:degree_level","label":"Degree Level","values":["Master&apos;s"]},{"key":"thesis:degree_name","label":"Degree Name","values":["Master of Science (MSc)"]},{"key":"thesis:institution_name","label":"Thesis Institution Name","values":["Faculty of Graduate Studies and Research, University of Regina"]}]},{"id":"language_rights","label":"Language and Rights","entries":[{"key":"dc:language.iso","label":"Language (ISO)","values":["en"]}]},{"id":"identifiers","label":"Identifiers","entries":[{"key":"dc:identifier.doi","label":"DOI","values":["https://doi.org/10.82465/4394"]},{"key":"dc:identifier.uri","label":"Identifier URI","values":["https://hdl.handle.net/10294/9255"]}]},{"id":"additional","label":"Additional Metadata","entries":[{"key":"dc:description","label":"Description","values":["A Thesis Submitted to the Faculty of Graduate Studies and Research In Partial Fulfillment of the Requirements for the Degree of Master of Science in Computer Science, University of Regina. ix, 83 p."]},{"key":"dc:description.abstract","label":"Abstract","values":["The World Wide Web (WWW) was first envisioned as a global web of linked documents and it is now becoming a global web of linked data. The World Wide Web Consortium (W3C) uses the term semantic web1 to refer to this new vision. This thesis is focused on building linked datasets from data which may not be fully accessible on the web. Once transformed into linked datasets, the public is more able to use and query the data. Various semantic web technologies for handling data are used in this thesis. These technologies includeWeb Ontology Language (OWL) for describing the semantics of the information, Resource Description Framework (RDF) for expressing the relationships in the information, and Simple Protocol and RDF Query Language (SPARQL) for querying the information used in web documents. In 2005, Sterling [55] described an object whose “informational support is so overwhelmingly extensive and rich that it is regarded as a material instantiation of an immaterial system”. He called this object a spime, a term that he coined by combining the words space and time. Spimes about particular items can collect linked information, which is shared on the web by citizen-consumers. Though Sterling used examples from the physical world, the semantic web may be the perfect environment for spimes to thrive. Hepting et al. [33] described the design of an informational spime, using an example of chicken eggs. This thesis helps to realize that design with a proof of concept for a spime with chicken egg information. The proof of concept was done using existing, 1https://www.w3.org/standards/semanticweb/ primarily open source, tools that allow for: extraction and remediation of data on the web, addition of data not already on the web, and querying of all that data. As a starting point for this thesis, information about eggs was found on a webpage marked up as a table in Hypertext Markup Language (HTML), which provided structure but no semantics. The approach including the following steps: extracting egg data from an existing web source2; developing an egg ontology; publishing the egg information as linked data; connecting this linked data to external sources such as Wikidata3; and querying the published linked data with SPARQL. The egg spime illustrated in the thesis contains information about eggs available at a specific place and time. Citizen-consumers in Regina, Canada and Patiala, India wanting to buy eggs could query their own local egg spimes, which may differ substantially in structure and content from one other. This thesis contributes to the realization of Berners-Lee’s vision for the semantic web. There are several opportunities for future work, including the addition of geospatial information and improved interface functionalities."]},{"key":"dc:title","label":"Title","values":["An Approach for Building and Querying Linked Datasets"]}]}],"canonical_facts":{"dc:contributor.advisor":["Hepting, Daryl"],"dc:contributor.committeemember":["Yao, Yiyu","Yang, Xue-Dong"],"dc:creator":["Kaur, Harmanjeet"],"dc:date.accessioned":["2020-08-29T19:04:44Z"],"dc:date.available":["2020-08-29T19:04:44Z"],"dc:date.issued":["2019-12"],"dc:description":["A Thesis Submitted to the Faculty of Graduate Studies and Research In Partial Fulfillment of the Requirements for the Degree of Master of Science in Computer Science, University of Regina. ix, 83 p."],"dc:description.abstract":["The World Wide Web (WWW) was first envisioned as a global web of linked documents and it is now becoming a global web of linked data. The World Wide Web Consortium (W3C) uses the term semantic web1 to refer to this new vision. This thesis is focused on building linked datasets from data which may not be fully accessible on the web. Once transformed into linked datasets, the public is more able to use and query the data. Various semantic web technologies for handling data are used in this thesis. These technologies includeWeb Ontology Language (OWL) for describing the semantics of the information, Resource Description Framework (RDF) for expressing the relationships in the information, and Simple Protocol and RDF Query Language (SPARQL) for querying the information used in web documents. In 2005, Sterling [55] described an object whose “informational support is so overwhelmingly extensive and rich that it is regarded as a material instantiation of an immaterial system”. He called this object a spime, a term that he coined by combining the words space and time. Spimes about particular items can collect linked information, which is shared on the web by citizen-consumers. Though Sterling used examples from the physical world, the semantic web may be the perfect environment for spimes to thrive. Hepting et al. [33] described the design of an informational spime, using an example of chicken eggs. This thesis helps to realize that design with a proof of concept for a spime with chicken egg information. The proof of concept was done using existing, 1https://www.w3.org/standards/semanticweb/ primarily open source, tools that allow for: extraction and remediation of data on the web, addition of data not already on the web, and querying of all that data. As a starting point for this thesis, information about eggs was found on a webpage marked up as a table in Hypertext Markup Language (HTML), which provided structure but no semantics. The approach including the following steps: extracting egg data from an existing web source2; developing an egg ontology; publishing the egg information as linked data; connecting this linked data to external sources such as Wikidata3; and querying the published linked data with SPARQL. The egg spime illustrated in the thesis contains information about eggs available at a specific place and time. Citizen-consumers in Regina, Canada and Patiala, India wanting to buy eggs could query their own local egg spimes, which may differ substantially in structure and content from one other. This thesis contributes to the realization of Berners-Lee’s vision for the semantic web. There are several opportunities for future work, including the addition of geospatial information and improved interface functionalities."],"dc:identifier.doi":["https://doi.org/10.82465/4394"],"dc:identifier.uri":["https://hdl.handle.net/10294/9255"],"dc:language.iso":["en"],"dc:publisher":["Faculty of Graduate Studies and Research, University of Regina"],"dc:title":["An Approach for Building and Querying Linked Datasets"],"dc:type":["master thesis"],"thesis:degree_discipline":["Computer Science"],"thesis:degree_level":["Master&apos;s"],"thesis:degree_name":["Master of Science (MSc)"],"thesis:institution_name":["Faculty of Graduate Studies and Research, University of Regina"]},"updated_at":"2026-07-24T04:03:38Z"}