Website Classification Dataset - UK Selective Web Archive

& Andrew Jackson
The dataset comprises a manually curated selective archive produced by UKWA which includes the classification of sites into a two-tiered subject hierarchy. In partnership with the Internet Archive and JISC, UKWA had obtained access to the subset of the Internet Archive’s web collection that relates to the UK. The JISC UK Web Domain Dataset (1996 - 2013) contains all of the resources from the Internet Archive that were hosted on domains ending in ‘.uk’, or...
This data repository is not currently reporting usage information. For information on how your repository can submit usage information, please see our documentation.