







Sure. AI companies have ALWAYS been training their models on Wikipedia content, which under the free and open access model is available to anyone — including AI companies. Agreements like these require AI companies to limit and offset the strain they place on Wikimedia infrastructure.
Kulusevski's Patella Spurs 🤦♂️
Hoping that @molly.wiki can help explain.
Jan 15, 2026 at 6:47 PM
Wikidata Launches Free Vector Database as Open Alternative to Closed AI Systems
Wikidata has launched something big.

OKA – Disseminating free content on Wikipedia and open platforms through targeted funding
We are a non-profit organization dedicated to improving Wikipedia and other open platforms. We do so by providing monthly stipends to full-time contributors and translators. We leverage AI (Large Language Models) to automate most of the work.
‘Impossible’ to create AI tools like ChatGPT without copyrighted material, OpenAI says
Pressure grows on artificial intelligence firms over the content used to train their products

Using AI in open source
If you want to use AI to help you contribute to one of the projects I maintain, I would be delighted. But I have rules.

Cloudflare's Matthew Prince has a plan to get Google and the AI oliigarchs to pay for your content even though many are used to getting it for free. He might have enough leverage to make them.
AI chatbots are blowing up the 30-year economic relationship publishers have had with search engines. Many think this will kill the web if not addressed. Google needs to change first. It's resisting. Cloudflare's Matthew Prince is going to try and make them.
Wikipedia Says AI Is Causing a Dangerous Decline in Human Visitors
“With fewer visits to Wikipedia, fewer volunteers may grow and enrich the content, and fewer individual donors may support this work.”
For Most of the World, Open-Source AI Is the Only Way Forward
Proprietary AI is both too expensive and too centralized in control for most countries and companies to rely upon.

“Wait, not like that”: Free and open access in the age of generative AI
The real threat isn’t AI using open knowledge — it’s AI companies killing the projects that make knowledge free

Nobody needs AI to search the Internet, court says in ruling against Google
Google AI Overview court loss in Germany could spell doom for AI search industry.

AI News Audit: AI, Canadian Journalism, and Paths for Policy Action — Centre for Media, Technology and Democracy
March 16, 2026 - AI companies built their products using Canadian journalism without permission and without compensation, and are now delivering that journalism to consumers as their own product. Existing copyright and media policy frameworks were not designed to address this. In February and Marc

Is Misinformation More Open? A Study of robots.txt Gatekeeping on the Web
Web scraping, the automated process of extracting information from websites, has long played a foundational role in the Internet ecosystem (Gray, 1995). It supports services such as search engine indexing, price comparison tools, and competitive intelligence. More recently, it has become a core component in the development of large-scale generative AI models. These Large Language Models (LLMs) require enormous volumes of training data, often in the terabyte range (Kaplan et al., 2020; Lehane, 2025), and the public web remains a low-cost, attractive source. Major model developers, including those behind OpenAI’s Chat-GPT (OpenAI, 2025), Google’s Bard (now known as Gemini) & Vertex AI (Romain, Danielle, 2023), and Anthropic’s Claude (Romain, Danielle, 2025), openly acknowledge the use of web scraping to construct their training corpora (Abdin et al., 2024; Brown et al., 2020; Chowdhery et al., 2023; Grattafiori et al., 2024; Team et al., 2024; Touvron et al., 2023).
Digital Science on Twitter / X
An AI-native workspace that already knows your project?The new Papers AI from Digital Science keeps your drafts, data & references together, so the AI assistant has full context of your work - instead of starting fresh each time. 👉 Available now: https://t.co/DNcG173GDk… pic.twitter.com/FQWxdfKfS4— Digital Science (@digitalsci) August 4, 2026

Unpacking Open Source Artificial Intelligence: Toward a Framework for Openness in Foundation Models
Openness has long driven innovation in software,9 and AI is no exception.12 While some see openness in foundation models (FMs) as a security threat,18 others argue that restricting access will not meaningfully reduce risk and will limit the benefits of transparency, research, and global participation.3 As the EU AI Act reporting requirements on FMs—also referred to as general-purpose AI models (GPAIMs)—move toward implementation, there is an urgent need for a more nuanced and informed understanding of openness in AI systems.

Apple almost open-sourced its AI models, here’s why it didn’t: report
A new report outlines the internal AI drama Apple has experienced recently, including the decision not to open-source its AI models.

Public AI Inference Utility
A nonprofit, open-source service to make public and sovereign AI models more accessible.
OpenAI strikes Reddit deal to train its AI on your posts
Reddit’s signed AI licensing deals with Google and OpenAI.
