Learning to WALK:
Building a National Web Archiving Collaborative Platform
Ian Milligan
Assistant Professor
@ianmilligan1
Nick Ruest
Digital Assets Librarian
@ruebot
Plan for the Talk
The Current Team
Historian
CS
Librarian/Archivist
The PI/Co-PIs
PhD RA
MA RA
MA RA
Postdoctoral
Fellow (Governance)
The Support
So what are we all about?
Let’s start with an opportunity not a problem.
We have fantastic web archival collections in Canada.
But not many people use them…
(At least until the Archive-It API comes online and transforms everything)
Canadian archival data is silo’d by institution and collection.
The Canadian Web Archival Landscape
Right now, to use Canadian web archives – you have to really want to use them.
i.e. you need to be an expert.
We want web archives to be used on page 150 of a random book.
Enter the Web Archives for Longitudinal Knowledge Project
WALKin’
We want to break down silos, bring Canadian web archives into a centralized portal with access to derivative datasets.
WALKin’
So what can we do to make sense of all this data?
Plan for the Talk
Workflow
In the back end we generate derivative datasets...
Workflow
So in the front end they…
… and we’ll ultimately.
Landing page
Collection lists
Collection detail
… but … how can they do those searches?
Plan for the Talk
Enter our impending community contribution
SCALING
💗 gifcities.org 💗
SolrCloud
Enter our impending community contribution
Hello Blacklight
WARCLight
Plan for the Talk
Derivative Datasets
Suddenly web archives aren’t boutique – they speak to a broader audience and you can imagine how to use them.
Data can then play with other people and platforms
The goal:
One central hub for web archiving search, research, derivatives...
So researchers can cite web archives on page 150 of a book without needing to be an expert!
Thanks!
Ian Milligan
Assistant Professor
@ianmilligan1
Nick Ruest
Digital Assets Librarian
@ruebot