How to get Public Document Finder in n8n
Find public PDFs and files for a company, domain, or keyword - no logins, no API keys. A run saves one row per result.
Set up the Apify actor
Apify is where this run is collected. An actor is the program that collects the rows. A task is that program saved with your settings, so the next section can ask for the latest rows. You need a free Apify account.
- 01Open Public Document Finder and click Try for free. Create an Apify account if you do not have one. The Free plan does not ask for a card.
- 02You are on the input form. Keep the sample for this first run.
{ "queries": [ "who.int" ], "includeSubdomains": true, "maxSubdomains": 25, "maxItems": 5, "maxResults": 5 } - 03Leave the proxy setting as it is. Click Start and wait until the run says Succeeded.
- 04Open Storage, then Dataset. You should see
interestScore,title,url,fileType. - 05When you want the real list, raise maxItems. The sample cap is only there so the first run stays small.
- 06Go back to the input and choose Save as a new task. Name it after this list.
- 07Open the task. Copy the task ID from the address bar. Excel, Sheets, Power BI, and the other guides ask for this ID.
- 08On the task, open Schedules and add a weekly run if you want fresh rows. Each run pays the start charge and the charge for the rows. The amounts are in the price list on this page.
- 09Open Settings, then API & Integrations, and copy an API token. Excel, Sheets, and the other guides put this token in a download link. Anyone with that link can download the rows, so treat it like a password.
How to set up in n8n
Use the task ID and the API token from Set up the Apify actor, above.
- 01Add an Apify node. Choose Run Actor. Pick Public Document Finder. Turn on Wait for finish. Set the input to:
{ "queries": [ "who.int" ], "includeSubdomains": true, "maxSubdomains": 25, "maxItems": 5, "maxResults": 5 } - 02Add Get Dataset Items and map
interestScore,title,url,fileType,hostClass. - 03The weekly schedule from Set up the Apify actor, above, keeps the weekly run. You can also schedule the n8n workflow.
Schema
Examples come from the actor sample, not a live result.
| Name | Description | Example |
|---|---|---|
interestScore | Interest Score | 0-100 |
title | Title | Best-effort |
url | Document URL | Direct |
fileType | File type | Extension |
hostClass | Host class | website, |
channel | Discovery channel | Which |
discoveryFlags | Discovery flags | item-a, |
interestScoreInterest Score0-100heuristic OSINT triage score (cloud shares, budgets, decks, etc. rank higher; robots.txt-style noise is skipped) titleTitleBest-efforttitle urlDocument URLDirector share URL fileTypeFile typeExtensionhint (pdf, docx, …) hostClassHost classwebsite,google_drive, amazon_s3, wayback, github, … channelDiscovery channelWhichdiscovery path found it discoveryFlagsDiscovery flagsitem-a,item-b