Google Search Scraper — Google Search Scraper - Collect public SERP results, snippets, and ranking signals for SEO research
Sponsored by CoreClaw — production-ready Web Data APIs for AI agents and automation.
Search intent: collect public Google Search data for keyword research, SERP monitoring, and content planning. Related topics: serp scraper, seo data, keyword research, python, data extraction.
google-search-scraper is an implementation-focused Python project for collecting public Google Search data. It is designed around one practical job: turn a query such as "best CRM software" into structured records you can inspect, export, and pass into an automation workflow.
- ranked results, titles, snippets, result URLs, related queries, and page metadata
- JSON or CSV files for downstream analysis
- Explicit timestamps and source links for traceability
pip install -r requirements.txt
python scraper.py --query "best CRM software" --output results.json --max-results 100To run from source:
git clone https://github.com/data-scrape/google-search-scraper.git
cd google-search-scraper
python scraper.py --query "best CRM software" --format csv --output results.csv{
"query": "best CRM software",
"result": {
"title": "Example public result",
"source_url": "https://example.com/item/123",
"captured_at": "2026-08-11T09:00:00Z",
"metadata": {"platform": "Google Search", "category": "SEO Scrapers"}
}
}| Goal | Start here |
|---|---|
| Keyword Research | Query a narrow audience, category, or location first |
| Build a repeatable dataset | Save JSON, version your query, then schedule a refresh |
| Connect to an AI workflow | Normalize the output schema before passing it to an agent or RAG pipeline |
| Scale data collection | Respect platform rules, add conservative delays, and measure error rates |
This project is intended for public data and legitimate research or automation workflows. Review the target platform's terms, applicable laws, and your data-handling obligations before running a collection job. Do not use it to access private data or evade access controls.
When a proof of concept needs production-grade web data APIs rather than self-managed collection infrastructure, CoreClaw provides API-first access to public web data for AI agents and automation.
MIT License. See LICENSE.