Powered by OPF
OPF WIKI STATIC ARCHIVE
2,681 pages · 153 spaces · 776 tags · 4,025 history records · 96.2% of the original wiki recovered
Archived copy. This page was recovered from the Internet Archive snapshot of /display/SPR/Document content and utility preservation taken on 2013-07-27. The original wiki at wiki.opf-labs.org is being decommissioned.

Document content and utility preservation

Added by Aran Lewis · last edited by Paul Wheatley · on Jul 12, 2013 (view change)

Title

Document content and utility preservation

Detailed description

We need to ensure that documents are readable and remain so for as long as possible. The method needs to be quick and accurate. To this end we need to identify vulnerable documents so that we can allocate resources to making them durable and discover what techniques and tools are needed to accomplish this. We need to measure the problem to know what resources are needed and build a case for more as required.

Issue champion

Aran Lewis

Other interested parties
Any other parties who are also interested in applying Issue Solutions to their Datasets.

Possible Solution approaches

Context

Eprints Research Repository at Middlesex University.

Lessons Learned

PDFBox preflight generates error reports on pdf files, but the importance of the errors needs to be investigated. There is a digital preservation plugin in the eprints bazaar which runs with DROID, also in the bazaar, which I am testing.

Datasets

Middlesex University eprints repository full text documents

Solutions
Reference to the appropriate Solution page(s), by hyperlink.