Sign in
Enter the admin token set on this deployment. It is kept for this browser tab only.
What is in the archive
Live counts from D1 and R2.
Copy attached files into your own storage
Documents that are scanned PDFs still live on the site they came from. Until they are copied, readers are sent there. Copying them into R2 makes the archive self-contained and keeps it working if the source site is reorganised.
Remove stored files
Deletes the cleaned text and copied attachments held in R2 for one section. The database rows are left alone, so the site will show documents with their text missing until you import that section again. Use it when you want to reclaim space or clear out a bad import.
- Source
- Section
- File
- Fields
- Import
Which website did this come from?
A source is a website you take documents from. Pick one, or add a new one.
Add a new source
The short code appears in web addresses and in storage keys. Letters, digits and hyphens.
Which kind of document?
Add a new section
Choose the JSON file
Match the fields
Each row is something the archive stores. Pick which part of your JSON holds it. The guesses below come from reading the file; change any that look wrong. This gets saved, so the next export of the same kind will skip this step.
How the first few records come out
Import
Sources and sections
Everything the archive knows about. Counts are live.
Import history
The last twenty runs.