







been thinking a lot about options for maintaining a sort of partial/sparse mst index over hubble’s mst-less record storage. and i have some ideas. (hubble doesn’t need this to serve sync.getRepo, bc star-lite’s mst reconstruction algorithm is fast enough to rebuild from scratch on demand)
/star/star-lite at main · microcosm.blue/star
tangled.orgJun 7, 2026 at 11:28 PM
Folders Are Broken, So I Built Heaper
Folders Are Broken, So I Built Heaper
turbopuffer: fast search on object storage
Inaugural blog post about the development of turbopuffer, a search engine that uses object storage and SSD caching for cost-effective, low latency search. This post describes into the motivation behind its creation, its unique architecture, and how it significantly reduces costs for large-scale vector searches. Discover how turbopuffer is transforming search infrastructure for companies like Cursor and Suno, offering a scalable and reliable solution.

Graft
Graft is an open-source transactional storage engine designed for efficient data synchronization at the edge.
Spacedrive — A local-first data engine for everything you own
Index any data source. Search everything from one place. Keep it on your machine.

9001/copyparty
Portable file server with accelerated resumable uploads, dedup, WebDAV, SFTP, FTP, TFTP, zeroconf, media indexer, thumbnails++ all in one file
WinFS, Integrated/Unified Storage, and Microsoft – Part 1
People have been bugging me to write about Integrated Storage for some time, and with Bill Gates having just disclosed that failure to ship WinFS was his biggest product regret now seemed like a g…

Backing up Spotify
We backed up Spotify (metadata and music files). It’s distributed in bulk torrents (~300TB). It’s the world’s first “preservation archive” for music which is fully open (meaning it can easily be mirrored by anyone with enough disk space), with 86 million music files, representing around 99.6% of listens.

outl — local-first markdown outliner & LLM second brain
Open-source bullet-point outliner. Plain markdown on disk (no UUIDs), peer-to-peer tree-CRDT sync that doesn't lose data when devices edit offline — your devices talk straight to each other, end-to-end encrypted, no cloud or third-party server holding your notes — and an MCP server so Claude, Cursor and ChatGPT use your notes as a second brain. A modern Roam, Logseq and Obsidian alternative built in Rust.

2TB SanDisk Extreme PRO with USB4 | Sandisk
The SanDisk Extreme PRO with USB4 portable SSD is engineered for on-the-move professionals and modern adventurers, with turbocharged read speeds up to 3800MB/s and capacities up to 4TB.

a neat thing about hubble is that you can get a point-in-time* snapshot of the whole network. vs doing your own backfill crawl, which will give you samples spread over a >24h period any researchers who want this, feel free to reach out
fig (aka:[phil])
24,347,088,288 records in 41,172,356 repos in the reachable network fits in 1.86TiB in hubble
I was looking at an older version of the software I'm yet again working on from what, 15 years ago, and I can't believe we're still mostly stuck with the same private blob storage options that can serve as user-data storage. S3-compatible storage + Dropbox was the list: github.com/icidasset/ongaku-ryoho-v1?tab…
Hubble is so cool, I highly recommend using it. I published a post yesterday that you can check out as an example.
Building the Social Graph with Hubble and Jetstream
50-shades.pckt.blogmaxine s. fritz
hubble is live #atprotonyc