Extract multi-page website data and save to Google Drive automatically
Workflow Description
Automation that fetches content from multiple web pages using HTTP requests, processes and structures data in batches, filters results, and automatically saves extracted information to Google Drive with advanced data handling capabilities.
How it works
- 1.Trigger automation manually or on schedule
- 2.Send HTTP requests to scrape content from multiple target pages
- 3.Parse and extract data from XML/HTML responses
- 4.Organize results into manageable batches for processing
- 5.Apply filters and transformations to clean data
- 6.Save final results as files in Google Drive
Use cases
- Collect product prices and details from multi-page e-commerce sites
- Gather contact lists or public data from websites at scale
- Monitor and archive website content updates automatically
Requirements
- Active Google Drive account for storage
- Understanding of target website structure and data fields
- Definition of page URLs or pagination patterns
Service Value
Ready-made workflow template for automation delivery and service execution.
Apps Used
Details
How to Use
- 1.Click "Download Template"
- 2.Open your n8n dashboard
- 3.Go to Workflows > Import from File
- 4.Select downloaded file and configure credentials
Nodes Used (16)
Sticky Note
Sticky Note
When clicking ‘Test workflow’
Manual Trigger
Loop Over Items
Split In Batches
Wait
Wait
Limit
Limit
Get List of Website URLs
HTTP Request
Convert to JSON
Xml
Create List of Website URLs
Split Out
Filter By Topics or Pages
Filter
Set Website URL
Set
Jina.ai Web Scraper
HTTP Request
Save Webpage Contents to Google Drive
Google Drive
Extract Title & Markdown Content
Code
Sticky Note1
Sticky Note
Sticky Note2
Sticky Note
Sticky Note3
Sticky Note