A useful starting point
The quick answer
- Keep “Include linked documents” selected in Zipasite.
- Archive the pages that link to the required PDFs.
- Extract the ZIP, find the documents and open each important file.
List the documents you need
Write down the public pages that contain the PDF links and the document titles you expect. Include the edition, language or year where those details matter. A catalogue may offer several files with very similar names.
A crawl follows discoverable references and access rules. A document stored on the server becomes reachable through the links exposed to the crawler. Use your inventory to assess coverage, especially when documents appear after a search or a form submission.
Use the linked-document setting
Open Zipasite, enter the site address and choose a small first page limit. “Include linked documents” is selected by default. Keep it selected for PDFs and other supported document links.
The page limit counts pages rather than the number of PDF files. Several documents on one page can add substantial download size. Start with a manageable sample and check that the links you care about are represented in the saved result.
Find and verify the PDFs
Extract the ZIP and open the domain’s assets folder. Search for .pdf files in your file manager, then use manifest.json to match local filenames with the original document URLs. Some download endpoints use a generic filename or extension; the manifest helps trace them.
Open the important documents in a PDF reader. Check their titles, editions and page counts against your inventory. A PDF-style link can lead to a login screen or an intermediate page, so the file itself is the useful evidence.
Common questions
Does Zipasite produce a PDF-only archive?
The ZIP contains the saved website and its supporting files. Select the PDFs you need from the extracted copy.
What about private documents?
Use the owner’s approved export or download workflow for documents behind sign-in. Zipasite’s crawler works with public responses.