Home » Posts tagged 'link rot' (Page 4)
Tag Archives: link rot
From Link Rot to Web Sanctuary
Here is an interesting story about preserving British government information.
- From Link Rot to Web Sanctuary: Creating the Digital Educational Resource Archive (DERA), by Bernard M. Scaife, Ariadne, Issue 67 (4 July-2011)
It occurred to us that this software could enable us to eradicate our link rot problem, whilst building in a core level of digital preservation and increasing the discoverability of these documents. We were convinced that a citation which linked to a record in a Web archive was far more likely to survive than one which did not.They knew that government budget cuts were increasing the risk of losing content from government departments. The article describes their experiences and summarizes what they learned:
- Placing files in a repository gives digital preservation to key documents in the subject field and eradicates the link rot problem.
- Adding high-quality metadata enhances the resource and allows it to hold its head high and become an integral part of a library's collection.
- A specialist library can play an important role in preserving domain-specific government content as part of its long-term strategy and ensure high-quality resources remain available.
- Provided you are prepared to get to grips with its complexity, the EPrints software is well suited to the task and provides good interoperability with other legacy systems for importing metadata
- The added value of being able to search the full text provides a potentially very rich resource for data mining whether by current or future researchers of educational history. Continue reading
Sometimes even the live links are dead (or languishing)
Readers of FGI are well acquainted with link rot, where internet links break over time. Today I'd like to talk about something more subtle with no obvious way to detect the problem. On the Alaska page of the State Agency Databases Across the Fifty States project, I had a link to APOC InfoQuick, a database of disclosure information for public officials and lobbyists from the Alaska Public Offices Commission. Today I visited the link at https://webapp.state.ak.us/apoc/index.jsp and chose the "lobbyist reporting" menu item because I thought it would be fun to list BP lobbyists in a personal blog entry I was drafting. The lobbyist reporting section had a Search Lobbyist Registrations link. I clicked on it, searched for BP and got some listings. But only from 2007, the first year that Sarah Palin was Governor. Searches in other parts of the lobbyist reporting system confirmed that NO information was available after 2007. I started to wonder if I'd missed the session law that repealed lobbying reporting requirements. Then I noticed that the URL started with "webapp" and thought that it might be good to see if this database was still linked from the APOC home page. It wasn't. Now they had a link called "search reports" at http://doa.alaska.gov/apoc/SearchReports/index.html. The page features two reporting systems for public officials - An "interim reporting system" for reports filed 2010 and later and "searchable campaign reporting" which is the public official/candidate portion of APOC InfoQuick. This explains why APOC InfoQuick wasn't taken off the live web. Current information on lobbyists in Alaska is still available, just not database searchable. You can access various PDF lobbyist reports from 2005 forward at http://doa.alaska.gov/apoc/TrainingReports/lobbyist.html. I have no information on why lobbyist information is no longer database searchable and speculating why would take me out of my comfort zone of not discussing policy choices made by the level of government I work for. The main point I'm making is that most librarians and other information specialists are pretty comfortable with link checking and fixing broken links when we find them. But what can we do when a site remains on the web but has stopped being updated? Especially when there's no note on the old site about the change? Continue reading
Pinboard report on link rot
A new report on link rot on the blog of the social bookmarking service Pinboard:
- Remembrance of Links Past, by Maciej Ceglowski, Pinboard Blog (May 26, 2011).
New Link Rot Report
For libraries that rely on pointing to URLs rather than preserving information in their own digital libraries, the new report from the Chesapeake Project provides sobering, factual data on the reliability of that strategy. In an examination of "link rot" the project found that 30.4% of URLs examined no longer provide access to their original information. This study is particularly relevant to government information specialists because more than 90% of their sample URLs were from state governments (state.[state code].us), organizations (.org), and government (.gov) the top-level domains. The Chesapeake Project Legal Information Archive, which harvests and preserves relevant digital legal information from the web, has been producing reports on "link rot" for several years. They define link rot as "a URL that no longer provides direct access to files matching the content originally harvested from the URL and currently preserved in the Chesapeake Group's digital archive." Their new report is now available:
- "Link Rot" and Legal Resources on the Web: A 2011 Analysis, by the Chesapeake Digital Preservation Group, [April 2011].
2011 Report on Link Rot
How reliable are those URLs in your OPAC? The Chesapeake Project Legal Information Archive which harvests and preserves relevant digital information from the web, has been producing reports on "link rot" for several years. They define link rot as "a URL that no longer provides direct access to files matching the content originally harvested from the URL and currently preserved in the Chesapeake Project's digital archive." Their new report is now available:
- Breaking Down Link Rot: The Chesapeake Project Legal Information Archive's Examination of URL Stability, By Sarah Rhodes, LLRX (March 1, 2011).
Latest Comments