







The scientific community must consider the longevity of open research infrastructure—why it might fail and how to prevent it.
Who Will Keep Research Data Infrastructure Open and Running?
The scientific community must consider the longevity of open research infrastructure—why it might fail and how to prevent it.

Doing Data Science on the Shoulders of Giants: The Value of Open Source Software for the Data Science Community
Open source software is ubiquitous throughout data science, and enables the work of nearly every data scientist in some way or another. Open source projects, however, are disproportionately maintained by a small number of individuals, some of whom are institutionally supported, but many of whom do this maintenance on a purely volunteer basis. The health of the data science ecosystem depends on the support of open source projects, on an individual and institutional level.

Governing by dismantling: tech oligarchy and the stifling of public data infrastructure
Published in Science as Culture (Ahead of Print, 2026)

COS goes FOSS The sorry state of scientific publishing and how we could move to an open and resilient infrastructure
“The data files remains our property and are not deposited for free access.”
The Politics of Open Infrastructures: Power, Governance, and Justice in Digital Knowledge Practices
This volume examines how openness is designed, governed, contested and lived in contemporary digital knowledge infrastructures. From open source software and internet standards, to citizen science platforms, public sector data systems and alternative computing practices, the book shows that infrastructures are never neutral technical backbones.

Position paper: persistent identifiers in research infrastructure policy - Crossref
PIDs have become central to national and international open research strategies, but identifiers alone cannot deliver the connected, open record that researchers, institutions, funders, publishers, and policymakers depend on. Effective research infrastructure rests on three interdependent elements: open, persistent identifiers; rich, open, and linked metadata; and the sustainable governance and resilient operation of the organisations involved. Crossref urges policymakers to evaluate all three together.

A Polycentric Governance Lens on Data Infrastructures
Funding policies for data infrastructure promote open data sharing to drive positive social impact. However, concerns regarding the long-term management of data within and across distributed infrastructures can hinder data sharing. We draw upon the concept of polycentric governance to demonstrate how collaborative practices of data curation, in preparing and maintaining data for (future) sharing, provide a solid foundation for understanding data governance within data infrastructures. Based on a qualitative case study of a distributed ecological network, we investigate the conditions under which data are managed as a shared resource by local actors to ensure the long-term (re)usability of data. We contribute to CSCW by conceptualising data curation as a complex form of governance practice with multiple centres of decision-making, each of which operates with some degree of autonomy in data infrastructures. A polycentric governance lens on data infrastructures advances the CSCW conception of data curation as a collective governance practice that can cultivate a data democracy culture within and across organisations, empower individuals to be accountable for their data, and foster a mindset shift toward decentralised data governance.

Investing in resilient infrastructure to safeguard scientific knowledge
Political threats to scientific data demand coordinated, resilient open infrastructure — IOI is helping build it.

A Data Utopia for Science-of-Science
Here I want to briefly sketch out a vision for how to solve a key set of problems facing science-of-science researchers, using the relatively new idea of a ‘data trust.’ In my ideal wor…

Broadening Access to Data Science Education in High School and Higher Education through Open Source Tools, Infrastructure, and Training
Equitable data science education requires a multifaceted approach, involving high school and higher education, community involvement, and accessible tools. A renewed investment in public digital infrastructure is needed to support these efforts. Nonprofits play a crucial role in supporting these efforts, and increased representation in leadership can enhance their impact. By addressing these disparities, we can ensure a more inclusive future in data science.
DataColada: No Comments - Replicability-Index
Science is like an iceberg. The published record is only a fraction of the things that university -paid academics do. Some time ago, Brian Nosek dreamed about a scientific utopia of open science that would make the workings of academia more transparent, but all we got was preprints and some badges - that are apparently DataColada: Open Science, but Closed Comments?
From Albums to Streams: How Modularity Changes Systems
Scientific publishing is breaking under document-centric formats designed for a physical world. Borrowing from music’s shift from albums to streaming, we make the case that open access alone cannot deliver reuse, trust, or scale. The future of science depends on modular, interoperable research components that move across tools, enabling new workflows, tools, and ecosystems.
Center for Open Science
At the Center for Open Science, our mission is to increase openness, integrity, and reproducibility of scholarly research. Promoting these practices within the research funding and publishing communities accelerates scientific progress

A Polycentric Governance Lens on Data Infrastructures
Computer Supported Cooperative Work (CSCW) - Funding policies for data infrastructure promote open data sharing to drive positive social impact. However, concerns regarding the long-term management...

Empowering science communities with open, democratic, researcher-owned infrastructure.
We’re building communities and tech for publishing, curating, sharing, and discussing research online using ATProto and other decentralized protocols.

Empowering science communities with open, democratic, researcher-owned infrastructure.
We’re building communities and tech for publishing, curating, sharing, and discussing research online using ATProto and other decentralized protocols.
