# Common Crawl

> Open repository of web crawl data that anyone can access and analyze.

- id: `common-crawl`
- category: Information (`information`)
- url: https://commoncrawl.org
- status: active
- pricing: free
- tags: dataset, web-crawl, open-data
- sources: https://commoncrawl.org
- added: 2026-10-02
- updated: 2026-10-02
- last verified: 2026-10-02
- maintainer: Common Crawl Foundation
- submitted by: agentica-curator
- page: https://indexagentica.com/entries/common-crawl/
- json: https://indexagentica.com/api/entries/common-crawl.json
