Archived copy.
This page was recovered from the Internet Archive snapshot of
/display/REQ/Identifying web content taken on 2013-07-27.
The original wiki at wiki.opf-labs.org is being decommissioned.
Identifying web content
| Title |
Identifying web content |
| Detailed description | The Web archives team at BnF has long suspected that the MIME types declared by the server for their pages' content were not accurate, which is a problem for preservation and could impede future emulation, for instance.Thus, when expanding BnF's preservation system (SPAR) to ingest our Web archives collection, we had an ARC module developed for JHOVE 2 (http://bitbucket.org/lbihanic/jhove2-bnf |
| Issue champion | |
| Other interested parties |
Any other parties who are also interested in applying Issue Solutions to their Datasets |
| Possible Solution approaches | Brief brainstorm of possible approaches to solving the Issue. Each approach should be described in a single sentence as part of a bulleted list |
| Context | Bibliothèque nationale de France (National Library of France, BnF) |
| Lessons Learned | Notes on Lessons Learned from tackling this Issue that might be useful to inform digital preservation best practice |
| Datasets | French Web Archives |
| Solutions | Server MIME Type Correction |