







Firecrawl is the web data API to search, scrape, and interact with the web at scale. Turn any source into clean Markdown or structured data your agents can ship with.
firecrawl/firecrawl
The context API to search, scrape, and interact with the web at scale. 🔥

Interact — Scrape a page, then take actions on it with AI prompts or code | Firecrawl
Scrape any page, then click, type, fill forms, and extract data with natural language or Playwright code. Live view, persistent profiles, session chaining.

Interact | Firecrawl
Interact with a page you fetched by prompting or running code.

Keynote: The Death of the Browser - Rachel-Lee Nabors, AgentQL
Scraping Framework for Golang
Scraping framework for extracting the data you need from websites, used for a wide range of applications, like data mining, data processing or archiving

Welcome to Tabstack
Tabstack API is a powerful web content extraction and transformation toolkit designed specifically for AI agent builders. It provides intelligent web scraping capabilities, content processing, and structured data extraction through a simple REST API.

Welcome to Tabstack
Tabstack API is a powerful web content extraction and transformation toolkit designed specifically for AI agent builders. It provides intelligent web scraping capabilities, content processing, and structured data extraction through a simple REST API.

How to Scrape Web Content for RAG with Readability.js
Tabstack - Web Data and Browser Automation APIs
Get structured output from one API call. Run extraction, web research, and browser automation without managing LLM orchestration, browser infra, or pipelines.

Common Crawl - Open Repository of Web Crawl Data
We build and maintain an open repository of web crawl data that can be accessed and analyzed by anyone.
WDC - RDFa, Microdata, and Microformat Data Sets
More and more websites have started to embed structured data describing products, people, organizations, places, and events into their HTML pages using markup standards such as Microdata, JSON-LD, RDFa, and Microformats. The Web Data Commons project extracts this data from several billion web pages. So far the project provides 12 different data set releases extracted from the Common Crawls 2010 to 2023. The project provides the extracted data for download and publishes statistics about the deployment of the different formats.
steel-dev/steel-browser
🔥 Open Source Browser API for AI Agents & Apps. Steel Browser is a batteries-included browser sandbox that lets you automate the web without worrying about infrastructure.
Web search· Ollama Blog
A new web search API is now available in Ollama. Ollama provides a generous free tier of web searches for individuals to use, and higher rate limits are available via Ollama’s cloud.

KiloClaw — Your AI assistant that actually does things
KiloClaw reads your email, manages your calendar, monitors your projects, and talks to you wherever you already are. Start free.

OWL - Semantic Web Standards
The W3C Web Ontology Language (OWL) is a Semantic Web language designed to represent rich and complex knowledge about things, groups of things, and relations between things. OWL is a computational logic-based language such that knowledge expressed in OWL can be exploited by computer programs, e.g., to verify the consistency of that knowledge or to make implicit knowledge explicit. OWL documents, known as ontologies, can be published in the World Wide Web and may refer to or be referred from other OWL ontologies. OWL is part of the W3C’s Semantic Web technology stack, which includes RDF, RDFS, SPARQL, etc.