A useful starting point

The quick answer

  1. Keep “Include linked documents” selected in Zipasite.
  2. Archive the pages that link to the required PDFs.
  3. Extract the ZIP, find the documents and open each important file.

List the documents you need

Write down the public pages that contain the PDF links and the document titles you expect. Include the edition, language or year where those details matter. A catalogue may offer several files with very similar names.

A crawl follows discoverable references and access rules. A document stored on the server becomes reachable through the links exposed to the crawler. Use your inventory to assess coverage, especially when documents appear after a search or a form submission.

Use the linked-document setting

Open Zipasite, enter the site address and choose a small first page limit. “Include linked documents” is selected by default. Keep it selected for PDFs and other supported document links.

The page limit counts pages rather than the number of PDF files. Several documents on one page can add substantial download size. Start with a manageable sample and check that the links you care about are represented in the saved result.

Find and verify the PDFs

Extract the ZIP and open the domain’s assets folder. Search for .pdf files in your file manager, then use manifest.json to match local filenames with the original document URLs. Some download endpoints use a generic filename or extension; the manifest helps trace them.

Open the important documents in a PDF reader. Check their titles, editions and page counts against your inventory. A PDF-style link can lead to a login screen or an intermediate page, so the file itself is the useful evidence.

Keep a clear reference collection

Copy selected PDFs into a separate working folder if that makes reading easier. Retain the complete site archive so the original page context remains available. Add a source list with document titles, URLs and dates.

  • Check required documents individually.
  • Review failed-file notes and access restrictions.
  • Keep language and edition details with each document.
  • Share files under the permissions that apply to them.

Common questions

Does Zipasite produce a PDF-only archive?

The ZIP contains the saved website and its supporting files. Select the PDFs you need from the extracted copy.

What about private documents?

Use the owner’s approved export or download workflow for documents behind sign-in. Zipasite’s crawler works with public responses.

Keep a copy you can check.

Start with a small set of public pages, then review your saved files.

Open Zipasite