Global ETD Search

Search theses and dissertations gathered from participating repositories worldwide. Every result links back to the library that holds it. No account is needed.

Results

Showing 1 to 17 of 17 for “"Web archiving"”.

  1. Performance Measurement and Analysis of Transactional Web Archiving

    Web archiving is necessary to retain the history of the World Wide Web and to study its evolution. It is important for the cultural heritage community. Some organizations are legally obligated to capture and archive Web content. The advent of transactional Web archiving makes the archiving process …

    vt Repository record for Performance Measurement and Analysis of Transactional Web Archiving (opens in a new tab)

  2. Performance Evaluation of Web Archiving Through In-Memory Page Cache

    … study proposes and evaluates a new method for Web archiving. We leverage the caching infrastructure in Web servers for archiving. Redis is used as the page cache and its persistence mechanism is exploited for archiving. We experimentally evaluate the performance of our archival technique using …

    vt Repository record for Performance Evaluation of Web Archiving Through In-Memory Page Cache (opens in a new tab)

  3. Improving Web Search Ranking Using the Internet Archive

    Current web search engines retrieve relevant results only based on the latest content of web pages stored in their indices despite the fact that many web resources update frequently. We explore possible techniques and data sources for improving web search result ranking using web page historical …

    vt Repository record for Improving Web Search Ranking Using the Internet Archive (opens in a new tab)

  4. Web Archive Services Framework for Tighter Integration Between the Past and Present Web

    <p>Web archives have contained the cultural history of the web for many years, but they still have a limited capability for access. Most of the web archiving research has focused on crawling and preservation activities, with little focus on the delivery methods. The current access methods are …

    odu Repository record for Web Archive Services Framework for Tighter Integration Between the Past and Present Web (opens in a new tab)

  5. Large Web Archive Collection Infrastructure and Services

    The web has evolved to be the primary carrier of human knowledge during the information age. The ephemeral nature of much web content makes web knowledge preservation vital in preserving human knowledge and memories. Web archives are created to preserve the current web and make it available for …

    vt Repository record for Large Web Archive Collection Infrastructure and Services (opens in a new tab)

  6. A Framework for Verifying the Fixity of Archived Web Resources

    <p>The number of public and private web archives has increased, and we implicitly trust content delivered by these archives. Fixity is checked to ensure that an archived resource has remained unaltered (i.e., fixed) since the time it was captured. Currently, end users do not have the ability to …

    odu Repository record for A Framework for Verifying the Fixity of Archived Web Resources (opens in a new tab)

  7. An Extensible Framework for Creating Personal Archives of Web Resources Requiring Authentication

    … key factors for the success of the World Wide Web are its large size and the lack of a centralized control over its contents. In recent years, many advances have been made in preserving web content but much of this content (namely, social media content) was not archived, or still to this day is …

    odu Repository record for An Extensible Framework for Creating Personal Archives of Web Resources Requiring Authentication (opens in a new tab)

  8. Scripts in a Frame: A Framework for Archiving Deferred Representations

    <p>Web archives provide a view of the Web as seen by Web crawlers. Because of rapid advancements and adoption of client-side technologies like JavaScript and Ajax, coupled with the inability of crawlers to execute these technologies effectively, Web resources become harder to archive as they become …

    odu Repository record for Scripts in a Frame: A Framework for Archiving Deferred Representations (opens in a new tab)

  9. The French in London on-land and on-line: an ethnosemiotic analysis

    … has curated a Special Collection of community Web resources in the UK Web Archive, laying the foundations for a theory of selective thematic Web archiving. An innovative “ethnosemiotic” paradigm, combining Bourdieu’s ethnographic principles and the multimodal social semiotic approach advocated …

    westminster Repository record for The French in London on-land and on-line: an ethnosemiotic analysis (opens in a new tab)

  10. Expanding the Usage of Web Archives by Recommending Archived Webpages Using Only the URI

    <p>Web archives are a window to view past versions of webpages. When a user requests a webpage on the live Web, such as http://tripadvisor.com/where_to_t ravel/, the webpage may not be found, which results in an HyperText Transfer Protocol (HTTP) 404 response. The user then may search for the …

    odu Repository record for Expanding the Usage of Web Archives by Recommending Archived Webpages Using Only the URI (opens in a new tab)

  11. Aggregating Private and Public Web Archives Using the Mementity Framework

    <p>Web archives preserve the live Web for posterity, but the content on the Web one cares about may not be preserved. The ability to access this content in the future requires the assurance that those sites will continue to exist on the Web until the content is requested and that the content will …

    odu Repository record for Aggregating Private and Public Web Archives Using the Mementity Framework (opens in a new tab)

  12. MementoMap: A Web Archive Profiling Framework for Efficient Memento Routing

    <p>With the proliferation of public web archives, it is becoming more important to better profile their contents, both to understand their immense holdings as well as to support routing of requests in Memento aggregators. A memento is a past version of a web page and a Memento aggregator is a tool …

    odu Repository record for MementoMap: A Web Archive Profiling Framework for Efficient Memento Routing (opens in a new tab)

  13. Lazy Preservation: Reconstructing Websites from the Web Infrastructure

    <p>Backup or preservation of websites is often not considered until after a catastrophic event has occurred. In the face of complete website loss, webmasters or concerned third parties have attempted to recover some of their websites from the Internet Archive. Still others have sought to retrieve …

    odu Repository record for Lazy Preservation: Reconstructing Websites from the Web Infrastructure (opens in a new tab)

  14. Using the Web Infrastructure for Real Time Recovery of Missing Web Pages

    <p>Given the dynamic nature of the World Wide Web, missing web pages, or "404 Page not Found" responses, are part of our web browsing experience. It is our intuition that information on the web is rarely completely lost, it is just missing. In whole or in part, content often moves from one URI to …

    odu Repository record for Using the Web Infrastructure for Real Time Recovery of Missing Web Pages (opens in a new tab)

  15. Intelligent Event Focused Crawling

    … event focused crawling system to collect Web data about key events. When an event occurs, many users try to locate the most up-to-date information about that event. Yet, there is little systematic collecting and archiving anywhere of information about events. We propose intelligent event …

    vt Repository record for Intelligent Event Focused Crawling (opens in a new tab)

  16. Bootstrapping Web Archive Collections From Micro-Collections in Social Media

    <p>In a Web plagued by disappearing resources, Web archive collections provide a valuable means of preserving Web resources important to the study of past events. These archived collections start with seed URIs (Uniform Resource Identifiers) hand-selected by curators. Curators produce high quality …

    odu Repository record for Bootstrapping Web Archive Collections From Micro-Collections in Social Media (opens in a new tab)

  17. To Relive the Web: A Framework for the Transformation and Archival Replay of Web Pages

    <p>When replaying an archived web page (known as a memento), the fundamental expectation is that the page should be viewable and function exactly as it did at archival time. However, this expectation requires web archives to modify the page and its embedded resources, so that they no longer …

    odu Repository record for To Relive the Web: A Framework for the Transformation and Archival Replay of Web Pages (opens in a new tab)