Berrycrawl logo

> Berrycrawl

Turn the live web into agent-ready data

visit website ↗ Developer Tools

[ screenshots ]

[ about ]

Berrycrawl is a web data API designed to provide AI agents with real-time, structured information from the internet. It offers functionalities such as scraping web pages, capturing clean screenshots, parsing documents, crawling sites, and extracting structured data through focused endpoints. By utilizing Berrycrawl, developers can integrate live web content into their applications without the need to build and maintain complex browser infrastructures.

Key Features:

  • Brand Profile Extraction:

    • Endpoint: /api/v1/brand
    • Description: Point to a website and get its complete public brand profile, including logos, colors, fonts, and social media links.
  • Web Crawling:

    • Description: Map and crawl entire sites, discovering URLs from sitemaps and links, and running bounded crawls with depth controls, concurrency, deduplication, and webhooks.
  • Data Extraction:

    • Description: Extract structured data by describing the data or providing a schema, turning one page or a list of pages into structured JSON jobs.
  • Search and Retrieval:

    • Description: Search the web, find current sources with advanced operators, and bring the relevant page content back in the same workflow.

Getting Started:

  1. Create an Account:

  2. Obtain an API Key:

    • After registration, generate an API key to authenticate your requests.
  3. Integrate into Your Application:

    • Use the provided endpoints to incorporate live web data into your application, enhancing its capabilities with up-to-date information.

For detailed documentation and integration guides, visit the Berrycrawl Documentation.

Note:

Berrycrawl is built for products, not demos, ensuring that developers have access to a robust and reliable infrastructure for their web data needs.