







A lightweight TUI application to view and query tabular data files, such as CSV, TSV, and parquet. - shshemi/tabiew
Parquet
Apache Parquet Documentation Releases Apache Parquet is an open source, column-oriented data file format designed for efficient data storage and retrieval. It provides high performance compression and encoding schemes to handle complex data in bulk and is supported in many programming languages and analytics tools.
jdefrancesco/dskDitto
Ultra fast and easy duplicate file finder. Awesome TUI/GUI to manage results.
Welcome to Tabstack
Tabstack API is a powerful web content extraction and transformation toolkit designed specifically for AI agent builders. It provides intelligent web scraping capabilities, content processing, and structured data extraction through a simple REST API.

Welcome to Tabstack
Tabstack API is a powerful web content extraction and transformation toolkit designed specifically for AI agent builders. It provides intelligent web scraping capabilities, content processing, and structured data extraction through a simple REST API.

Tabstack - Web Data and Browser Automation APIs
Get structured output from one API call. Run extraction, web research, and browser automation without managing LLM orchestration, browser infra, or pipelines.

Dremel: interactive analysis of web-scale datasets: Proceedings of the VLDB Endowment: Vol 3, No 1-2
Dremel is a scalable, interactive ad-hoc query system for analysis of read-only nested data. By combining multi-level execution trees and columnar data layout, it is capable of running aggregation queries over trillion-row tables in seconds. The system ...

Tabstack Research: Verified Answers from the Open Web
Offload browser orchestration and agentic loops to Tabstack. Bridge the synthesis gap with a research primitive that delivers verified, cited web data.

BookStack
BookStack is a simple, open-source, self-hosted, easy-to-use platform for organising and storing information.
Spacedrive — A local-first data engine for everything you own
Index any data source. Search everything from one place. Keep it on your machine.

Column Storage for the AI Era
In the past few years, we’ve seen a cambrian explosion of new columnar formats, challenging the hegemony of Parquet: Lance, Fastlanes, Nimble, Vortex, AnyBlox, F3 (File Format for the Future). The thinking is that the context has changed so much that the design of yore (the previous decade) is not going to cut it moving forward. This seemed a bit intriguing to me, especially since the main contribution of Parquet has been to provide a standard for columnar storage. Parquet is not simply a file format. As an open source project hosted by the ASF, it acts as a consensus building machine for the industry. Creating six new formats is not going to help with interroperability. I spent some time to understand a bit better how things actually changed and how Parquet needs to adapt to meet the demands of this new era. In this post I’ll discuss my findings.
fx – a terminal JSON viewer & processor
A terminal viewer & processor for JSON, YAML, & TOML (TUI and CLI)

Apache Calcite | Proceedings of the 2018 International Conference on Management of Data
As part of the Digital Library's transition to Open Access, new features for researchers are available in the Premium Edition. Click here to learn more.

About the App

The Modern Electronic Lab Notebook (ELN) - LabArchives
Tally Counter & Tracker Daily - Apps on Google Play
Mobile Data Collection and Collaboration App
ODK Collect
Field Book Documentation