







"I found it very weird that there essentially is no way to browse the web in an open manner. So that's what I am trying to build," the founder of Stract said.
A Former Google Engineer Built a Search Engine for Finding Every Privacy Violation You Face Online
Former Google engineer Tim Libert is releasing a search engine, webXray, that aims to find illicit online data collection and tracking—with the goal of becoming “the Henry Ford of tech lawsuits.”

Welcome - OpenWebSearch.eu – Promoting Europe‘s Independence in Web Search
Michael Granitzer University Passau and Open Search Foundation, Coordinator of OpenWebSearch.EU

Why Google’s New AI-Saturated Search Page Will Be A Disaster
Google didn’t invent full-text search of the Internet – that honor belongs to early pioneers such as WebCrawler, Lycos and AltaVista. But for the last 25 years or so, Google has…

The Web Can Thrive Without Google’s Search Monopoly
Viable browsers and meaningful contributions to web standards can be sustained with more modest revenue streams, writes Alissa Cooper.

Google’s broken link to the web
With AI search results coming to the masses, the human-powered web recedes further into the background

Search privately and without ads — Uruky
Search privately and without ads using Uruky, the private search engine.

Is Misinformation More Open? A Study of robots.txt Gatekeeping on the Web
Web scraping, the automated process of extracting information from websites, has long played a foundational role in the Internet ecosystem (Gray, 1995). It supports services such as search engine indexing, price comparison tools, and competitive intelligence. More recently, it has become a core component in the development of large-scale generative AI models. These Large Language Models (LLMs) require enormous volumes of training data, often in the terabyte range (Kaplan et al., 2020; Lehane, 2025), and the public web remains a low-cost, attractive source. Major model developers, including those behind OpenAI’s Chat-GPT (OpenAI, 2025), Google’s Bard (now known as Gemini) & Vertex AI (Romain, Danielle, 2023), and Anthropic’s Claude (Romain, Danielle, 2025), openly acknowledge the use of web scraping to construct their training corpora (Abdin et al., 2024; Brown et al., 2020; Chowdhery et al., 2023; Grattafiori et al., 2024; Team et al., 2024; Touvron et al., 2023).
Hister | Your Own Search Engine
Hister is a private, self hosted search engine for the pages you visit and the files you keep. Search from the web, terminal, command line, or MCP.

OpenAI Might Be Making a Web Browser
Link to: https://www.theinformation.com/articles/openai-considers-taking-on-google-with-browser


Web search· Ollama Blog
A new web search API is now available in Ollama. Ollama provides a generous free tier of web searches for individuals to use, and higher rate limits are available via Ollama’s cloud.

Common Crawl - Open Repository of Web Crawl Data
We build and maintain an open repository of web crawl data that can be accessed and analyzed by anyone.
Search Engine for Source Code - PublicWWW.com
Search engine for source code - ultimate solution for digital marketing and affiliate marketing research.