Bulk PDF & Photo Data Extraction
Automated pipeline that extracts text from hundreds of PDFs and photos, analyzes content with AI, and writes structured data to a spreadsheet.

The Problem
Processing hundreds of PDF documents and photos manually to build a structured database was taking days of repetitive work — opening each file, reading content, and typing data into a spreadsheet.
The Solution
Created an n8n workflow that reads a project list from a sheet, loops through project folders in Google Drive, downloads PDFs, extracts text, analyzes document quality, and feeds everything into a DeepSeek LLM for structured analysis. The workflow also handles photos by listing image files from designated folders. All extracted and analyzed data is combined and written to a Google Sheet as structured records.
Key Outcomes
- Saved days of manual data entry from hundreds of PDFs
- AI-powered extraction and structuring of document content
- Handles both PDF text and photo image sources
- All data written to Sheets in a consistent, queryable format
Want something similar for your business?
Book a Free Strategy Call