Global ETD Search
Search theses and dissertations gathered from participating repositories worldwide. Every result links back to the library that holds it. No account is needed.
Results
Showing 1 to 19 of 19 for “"Web archives"”.
-
Aggregating Private and Public Web Archives Using the Mementity Framework
<p>Web archives preserve the live Web for posterity, but the content on the Web one cares about may not be preserved. The ability to access this content in the future requires the assurance that those sites will continue to exist on the Web until the content is requested and that the content will …
-
Using Web Archives to Enrich the Live Web Experience Through Storytelling
… our cultural discourse occurs primarily on the Web. Thus, Web preservation is a fundamental precondition for multiple disciplines. Archiving Web pages into themed collections is a method for ensuring these resources are available for posterity. Services such as Archive-It exists to allow …
-
Expanding the Usage of Web Archives by Recommending Archived Webpages Using Only the URI
<p>Web archives are a window to view past versions of webpages. When a user requests a webpage on the live Web, such as http://tripadvisor.com/where_to_t ravel/, the webpage may not be found, which results in an HyperText Transfer Protocol (HTTP) 404 response. The user then may search for the …
-
Improving Collection Understanding for Web Archives with Storytelling: Shining Light Into Dark and Stormy Archives
… sense of an ever-increasing number of archived web pages. As collections themselves grow, we need tools to make sense of them. Tools that work on the general web, like search engines, are not a good fit for these collections because search engines do not currently represent multiple document …
-
A Framework for Verifying the Fixity of Archived Web Resources
<p>The number of public and private web archives has increased, and we implicitly trust content delivered by these archives. Fixity is checked to ensure that an archived resource has remained unaltered (i.e., fixed) since the time it was captured. Currently, end users do not have the ability to …
-
To Relive the Web: A Framework for the Transformation and Archival Replay of Web Pages
<p>When replaying an archived web page (known as a memento), the fundamental expectation is that the page should be viewable and function exactly as it did at archival time. However, this expectation requires web archives to modify the page and its embedded resources, so that they no longer …
-
Web Archive Services Framework for Tighter Integration Between the Past and Present Web
<p>Web archives have contained the cultural history of the web for many years, but they still have a limited capability for access. Most of the web archiving research has focused on crawling and preservation activities, with little focus on the delivery methods. The current access methods are …
-
MementoMap: A Web Archive Profiling Framework for Efficient Memento Routing
<p>With the proliferation of public web archives, it is becoming more important to better profile their contents, both to understand their immense holdings as well as to support routing of requests in Memento aggregators. A memento is a past version of a web page and a Memento aggregator is a tool …
-
Opal: In Vivo Based Preservation Framework for Locating Lost Web Pages
… a framework for interactively locating missing web pages (http status code 404). Opal is an example of "in vivo" preservation: harnessing the collective behavior of web archives, commercial search engines, and research projects for the purpose of preservation. Opal servers learn from their …
-
Lazy Preservation: Reconstructing Websites from the Web Infrastructure
<p>Backup or preservation of websites is often not considered until after a catastrophic event has occurred. In the face of complete website loss, webmasters or concerned third parties have attempted to recover some of their websites from the Internet Archive. Still others have sought to retrieve …
-
An Extensible Framework for Creating Personal Archives of Web Resources Requiring Authentication
… key factors for the success of the World Wide Web are its large size and the lack of a centralized control over its contents. In recent years, many advances have been made in preserving web content but much of this content (namely, social media content) was not archived, or still to this day is …
-
Avoiding Spoilers on Mediawiki Fan Sites Using Memento
… shows, novels, movies) exist on the World Wide Web. These wikis provide a wealth of information about complex stories, but if readers are behind in their viewing they run the risk of encountering spoilers" -- information that gives away key plot points before the intended time of the show's …
-
Advancing Internet Viewpoint Diversity: A Novel Algorithm and a Corpus Creation Tool
… creating indexed corpora from the Common Crawl web archives. The architecture is instantiated into an automated tool that generates an intelligible topical corpus through a series of steps involving processing, filtering, cleaning, and removing duplicate content. Utilizing this tool, we …
-
Large Web Archive Collection Infrastructure and Services
The web has evolved to be the primary carrier of human knowledge during the information age. The ephemeral nature of much web content makes web knowledge preservation vital in preserving human knowledge and memories. Web archives are created to preserve the current web and make it available for …
-
Scripts in a Frame: A Framework for Archiving Deferred Representations
<p>Web archives provide a view of the Web as seen by Web crawlers. Because of rapid advancements and adoption of client-side technologies like JavaScript and Ajax, coupled with the inability of crawlers to execute these technologies effectively, Web resources become harder to archive as they become …
-
Detecting, Modeling, and Predicting User Temporal Intention
… we start by analyzing the content on the live web and its persistence. We noticed that a portion of the resources shared in social media disappear, and with further analysis we unraveled a relationship between this disappearance and time. We lose around 11% of the resources after one year of …
-
Visualizing Digital Collections at Archive-It
… create,maintain, and view digital collections of web resources. The current interface of Archive-It is largely text-based, supporting drill-down navigation using lists of URIs.While this interface provides good searching capabilities, it is not efficient for browsing. In the absence of keywords, a …
-
Web-based library for student projects/theses and faculty research papers
… purpose of this project is to make available a Web-based Library, a web application developed for the Department of Computer Science at CSUSB to manage student projects/theses and faculty papers. The project is designed in accordance with Model-View-Controller (MVC) design pattern using the …
-
Web-based library for student projects/theses and faculty research papers
… purpose of this project is to make available a Web-based Library, a web application developed for the Department of Computer Science at CSUSB to manage student projects/theses and faculty papers. The project is designed in accordance with Model-View-Controller (MVC) design pattern using the …