Warp is an agentic development environment, born out of the terminal.
parquet
Counted from the Cargo.toml manifests of the 136 indexed repositories that declare parquet as a dependency — not download counts. Dependency data last verified 2026-08-13.
Crates that show up unusually often in parquet projects. The most distinctive pairings rank first — crates these projects use far more than the average indexed Rust project does, not just crates that are popular everywhere. Each percentage is the share of parquet projects that also use it.
Rust full node implementation of the Fuel v2 protocol.
Scalable datastore for metrics, events, and real-time analytics
Search infrastructure for AI
Production-grade Rust-native trading engine with deterministic event-driven architecture
Neon: Serverless Postgres. We separated storage and compute to offer autoscaling, code-like database branching, and scale to zero.
A high-performance observability data pipeline.
Minimalist ML framework for Rust
Generate any location from the real world in Minecraft with a high level of detail.
dbt enables data analysts and engineers to transform their data using the same practices that software engineers use to build applications.
TensorZero is an open-source LLMOps platform that unifies an LLM gateway, observability, evaluation, optimization, and experimentation.
Cloud-native search engine for observability. An open-source alternative to Datadog, Elasticsearch, Loki, and Tempo.
Visualize, query, and stream to train on multimodal robotics data.
Data Agent Ready Warehouse : One for Analytics, Search, AI, Python Sandbox. — rebuilt from scratch. Unified architecture on your S3.
Event streaming platform for agentic AI. Continuously ingest, transform, and serve event streams in real time, at scale.
Apache DataFusion SQL Query Engine
Sui, a next-generation smart contract platform with high throughput, low latency, and an asset-oriented programming model powered by the Move programming language
Open Lakehouse Format for Multimodal AI. Convert from Parquet in 2 lines of code for 100x faster random access, vector index, and data versioning. Compatible with Pandas, DuckDB, Polars, Pyarrow, and PyTorch with more integrations coming..
The open-source Observability 2.0 database. One engine for metrics, logs, and traces — replacing Prometheus, Loki & ES.
The live data layer for apps and AI agents. Create up-to-the-second views into your business, just using SQL
High-performance data engine for AI and multimodal workloads. Process images, audio, video, and structured data at any scale
One SQL interface over APIs, files, and live sources — built for agents.
Distributed stream processing engine in Rust
Apache Iggy: Hyper-Efficient Message Streaming at Laser Speed
The CSV magician
Language model tokenization at GB/s
DORA (Dataflow-Oriented Robotic Architecture) is middleware designed to streamline and simplify the creation of AI-based robotic applications. It offers low latency, composable, and distributed dataflow capabilities. Applications are modeled as directed graphs, also referred to as pipelines.
Official Rust implementation of Apache Arrow
Drop-in Apache Spark replacement written in Rust, unifying batch processing, stream processing, and compute-intensive AI workloads.
A native Rust library for Delta Lake, with bindings into Python
Graph Node indexes data from blockchains such as Ethereum and serves it over GraphQL
An extensible, state-of-the-art framework for columnar compression, and the fastest FOSS columnar file format. Formerly at @spiraldb, now an Incubation Stage project at LFAI&Data, part of the Linux Foundation.
GlueSQL is quite sticky. It sticks to anything.
A portable accelerated SQL query, search, and LLM-inference engine, written in Rust, for data-grounded AI apps and agents.
Parseable is an observability datalake built from first principles.
Stream your Postgres data anywhere in real-time. Simple Rust building blocks for change data capture (CDC) pipelines.
Apache Mahout - an environment for quickly creating scalable, performant machine learning applications.
📺(tv) Tidy Viewer is a cross-platform CLI csv pretty printer that uses column styling to maximize viewer enjoyment.
The Feldera Incremental Computation Engine
Apache Kafka® compatible broker with S3, PostgreSQL, SQLite, Apache Iceberg and Delta Lake
The Auron accelerator for distributed computing framework (e.g., Spark) leverages native vectorized execution to accelerate query processing
Tonbo is an embedded database for serverless and edge runtimes.
Apache Iceberg
Open-source Pricing and Billing Infrastructure 🚀 Subscription management, Invoicing, Pricing, Usage-based billing, Cost limiting, Grandfathering, Experiments, Revenue analytics & Actionable insights
Local-first ETL/ELT studio: a drag-and-drop visual pipeline designer that compiles to SQL and runs on DuckDB. Tiny desktop app, no servers, git-friendly workspaces.
Local-first ETL/ELT studio: a drag-and-drop visual pipeline designer that compiles to SQL and runs on DuckDB. Tiny desktop app, no servers, git-friendly workspaces.
🌊 Continuously synchronize the systems where your data lives, to the systems where you _want_ it to live, by managing your data flows with Estuary. 🌊
Postgres Foreign Data Wrapper development framework in Rust.
Fast, streaming indexing, query, and agentic LLM applications in Rust
Peek inside Parquet files right from your terminal