Process documents for ocr & extraction
Runs on a schedule across Google Drive, Slack and Google Sheets, looks the answer up in a vector store, then parses the incoming file into rows. Built for ocr & extraction, around a "List Changed Docs (Google Drive)" step, in 21 documented nodes.
Process documents with Google Drive and Google Sheets for your team. An original, ready-to-import n8n workflow built around the 'File Processing' pattern for legal & compliance. It runs on a schedule and orchestrates 21 nodes across Google Drive, Slack, Google Sheets. It uses n8n's LangChain AI nodes (Agent/Chain + chat model). Retry-on-fail is enabled on external calls. Import the JSON, add your credentials, and adapt the parameters. The result is a hands-off process that just works.
Key benefits
- ✓Keeps Google Drive and Slack in step for ocr & extraction, without anyone re-typing a row
- ✓Safe to re-run: records already in Google Drive or Slack are skipped, across all 21 nodes of this ocr-&-extraction build
- ✓A failed Google Drive or Slack call is retried rather than dropped, at any of the 21 nodes in this ocr-&-extraction build
- ✓Every ocr-&-extraction run leaves an audit row behind, alongside Google Drive
- ✓Fans the same result out to Google Sheets in the same run, for ocr & extraction
Use cases
- →Teams keeping Google Drive and Slack aligned by hand for ocr & extraction
- →Answering ocr-&-extraction questions from your own indexed documents, then replying in Google Drive
- →Ocr & extraction work where "List Changed Docs (Google Drive)" is still a manual step before Google Drive
- →Ocr & extraction work that only reaches Google Drive on a schedule, when someone remembers to start it
Integrations
Workflow preview (21 nodes)
The template's actual node layout and connections. Node parameters and credential slots unlock with the download.
What's inside (21 nodes)
A look at the node types this workflow uses. No purchase required — full parameters and credentials are yours after you buy.
From download to running in 3 steps
- 1Import. In n8n, open Workflows → menu → “Import from File” and pick the downloaded JSON.
- 2Connect. Add your own credentials on the app nodes - n8n highlights exactly which ones need them.
- 3Activate. Run it once to test, then toggle Active. The automation is live.
What you'll learn
Production levelProduction-grade: idempotency, dead-letter handling, and reconciliation included.
- •Downloading and parsing CSV, XLSX, or PDF into JSON items
- •Serializing processed items back out to a file
- •Archiving the processed file
Every node carries its own documentation on the canvas, plus notes explaining why the architecture was built this way, a credential setup guide, troubleshooting, and three practice exercises. Sample data comes pinned to the trigger, so you can hit Execute and watch data flow before connecting a single account.
Common questions
What exactly do I get?
The complete, ready-to-import n8n workflow as a JSON file, delivered instantly after payment. Import it into your own n8n (cloud or self-hosted), add your credentials, and it runs.
How do I import it into n8n?
In n8n, open Workflows, click the three-dot menu, choose “Import from File”, and select the downloaded JSON. Then connect your own app credentials on the highlighted nodes.
Do I need anything else for it to work?
You need your own n8n instance and accounts/credentials for the apps this workflow connects to (Google Drive, Slack, Google Sheets). No coding is required.
What if it doesn't work for me?
If the file is faulty, won't import, or isn't as described, contact us within 7 days and we'll fix it or refund you - see our refund policy.
Can I modify or resell it?
You can freely adapt and use it in your own or your clients' projects. Reselling or redistributing the template file itself is not permitted.
Related templates
Extract from receipts: Slack, scheduled
Ocr & extraction on n8n: runs on a schedule across Slack, Google Sheets and Google Drive, starting at "Export All Rows (Google Sheets)". 16 documented nodes.
Distill documents for ocr & extraction
Ocr & extraction on n8n: runs when a webhook arrives across Google Sheets, Slack, Google Drive and Outlook, starting at "Collect Raw Text". 22 documented nodes.
Extract from documents: Slack, scheduled
Ocr & extraction on n8n: runs on a schedule across Slack, Google Sheets, Notion and Google Drive, starting at "Export All Rows (Notion)". 23 documented nodes.
Issue contracts: Google Sheets & Slack, webhook
Ocr & extraction on n8n: runs when a webhook arrives across Google Sheets, Slack, Google Drive and Twilio, starting at "Gather API Numbers". Drops anything it has already handled. 21 documented nodes.